LEVEN: A Large-Scale Chinese Legal Event Detection Dataset

Yao, Feng; Xiao, Chaojun; Wang, Xiaozhi; Liu, Zhiyuan; Hou, Lei; Tu, Cunchao; Li, Juanzi; Liu, Yun; Shen, Weixing; Sun, Maosong

Computer Science > Computation and Language

arXiv:2203.08556 (cs)

[Submitted on 16 Mar 2022]

Title:LEVEN: A Large-Scale Chinese Legal Event Detection Dataset

Authors:Feng Yao, Chaojun Xiao, Xiaozhi Wang, Zhiyuan Liu, Lei Hou, Cunchao Tu, Juanzi Li, Yun Liu, Weixing Shen, Maosong Sun

View PDF

Abstract:Recognizing facts is the most fundamental step in making judgments, hence detecting events in the legal documents is important to legal case analysis tasks. However, existing Legal Event Detection (LED) datasets only concern incomprehensive event types and have limited annotated data, which restricts the development of LED methods and their downstream applications. To alleviate these issues, we present LEVEN a large-scale Chinese LEgal eVENt detection dataset, with 8,116 legal documents and 150,977 human-annotated event mentions in 108 event types. Not only charge-related events, LEVEN also covers general events, which are critical for legal case understanding but neglected in existing LED datasets. To our knowledge, LEVEN is the largest LED dataset and has dozens of times the data scale of others, which shall significantly promote the training and evaluation of LED methods. The results of extensive experiments indicate that LED is challenging and needs further effort. Moreover, we simply utilize legal events as side information to promote downstream applications. The method achieves improvements of average 2.2 points precision in low-resource judgment prediction, and 1.5 points mean average precision in unsupervised case retrieval, which suggests the fundamentality of LED. The source code and dataset can be obtained from this https URL.

Comments:	Accepted to ACL2022 Findings
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2203.08556 [cs.CL]
	(or arXiv:2203.08556v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2203.08556

Submission history

From: Feng Yao [view email]
[v1] Wed, 16 Mar 2022 11:40:02 UTC (1,332 KB)

Computer Science > Computation and Language

Title:LEVEN: A Large-Scale Chinese Legal Event Detection Dataset

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:LEVEN: A Large-Scale Chinese Legal Event Detection Dataset

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators