Self-training improves few-shot learning in legal artificial intelligence tasks

Artificial Intelligence and Law 33 (3):809-825 (2025)
  Copy   BIBTEX

Abstract

As the labeling costs in legal artificial intelligence tasks are expensive. Therefore, it becomes a challenge to utilize low cost to train a robust model. In this paper, we propose a LAIAugment approach, which aims to enhance the few-shot learning capability in legal artificial intelligence tasks. Specifically, we first use the self-training approach to label the amount of unlabelled data to enhance the feature learning capability of the model. Moreover, we also search for datasets that are similar to the training set by improving the text similarity function. We conducted experimental analyses for three legal artificial intelligence tasks, including evidence extraction, legal element extraction, and case multi-label prediction, which composed of 3500 judgement documents. The experimental results show that the proposed LAIAugment method has an average F1-score of 72.3% on the three legal AI tasks, which is 1.93% higher than the baseline model. At the same time, it shows a huge improvement in few-shot learning.

Other Versions

No versions found

Links

PhilArchive

External links

Setup an account with your affiliations in order to access resources via your University's proxy server

Through your library

Similar books and articles

A task-based interface to legal databases.Luuk Matthijssen - 1998 - Artificial Intelligence and Law 6 (1):81-103.

Analytics

Added to PP
2024-05-18

Downloads
68 (#876,064)

6 months
23 (#456,593)

Historical graph of downloads
How can I increase my downloads?