paper-with-me

Papers

Breaking the Manual Annotation Bottleneck: Creating a Comprehensive Legal Case Criticality Dataset through Semi-Automated Labeling

2024-10-17 · Ronja Stern, Ken Kawamura, Matthias Stürmer, Ilias Chalkidis, Joel Niklaus

Predicting case criticality helps legal professionals in the court system manage large volumes of case law. This paper introduces the Criticality Prediction dataset, a new resource for evaluating the potential influence of Swiss Federal Supreme Court decisions on future jurisprudence. Unlike existing approaches that rely on resource-intensive manual annotations, we semi-automatically derive labels leading to a much larger dataset than otherwise possible. Our dataset features a two-tier labeling system: (1) the LD-Label, which identifies cases published as Leading Decisions (LD), and (2) the Citation-Label, which ranks cases by their citation frequency and recency. This allows for a more nuanced evaluation of case importance. We evaluate several multilingual models, including fine-tuned variants and large language models, and find that fine-tuned models consistently outperform zero-shot baselines, demonstrating the need for task-specific adaptation. Our contributions include the introduction of this task and the release of a multilingual dataset to the research community.

📄 PDF Abstract BibTeX arXiv:2410.13460

Code (0)

등록된 구현이 없습니다.

Tasks

Jurisprudence

Similar Papers 제목 키워드 기반

A Guide for Manual Annotation of Scientific Imagery: How to Prepare for Large Projects

2025-08-20 · Azim Ahmadzadeh, Rohan Adhyapak, Armin Iraji, Kartik Chaurasiya 외 arxiv

Despite the high demand for manually annotated image data, managing complex and costly annotation projects remains under-discussed. This is partly due to the fact that leading such projects requires dealing with a set of…

SAM2Auto: Auto Annotation Using FLASH

2025-06-09 · Arash Rocky, Q. M. Jonathan Wu

Vision-Language Models (VLMs) lag behind Large Language Models due to the scarcity of annotated datasets, as creating paired visual-textual annotations is labor-intensive and expensive. To address this bottleneck, we int…

Instance SegmentationObjectobject-detectionObject Detection+5

Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

2023-12-14 · Peiyi Wang, Lei LI, Zhihong Shao, R. X. Xu 외

In this paper, we present an innovative process-oriented math process reward model called \textbf{Math-Shepherd}, which assigns a reward score to each step of math problem solutions. The training of Math-Shepherd is achi…

Arithmetic ReasoningGSM8KMathMathematical Reasoning+2

Breaking Language Barriers with MMTweets: Advancing Cross-Lingual Debunked Narrative Retrieval for Fact-Checking

2023-08-10 · Iknoor Singh, Carolina Scarton, Xingyi Song, Kalina Bontcheva

Finding previously debunked narratives involves identifying claims that have already undergone fact-checking. The issue intensifies when similar false claims persist in multiple languages, despite the availability of deb…

Fact CheckingMisinformationRe-RankingRetrieval+1

BREAKING! Presenting Fake News Corpus for Automated Fact Checking

2019-07-01 · ACL 2019 7 · Archita Pathak, Rohini Srihari

Popular fake news articles spread faster than mainstream articles on the same topic which renders manual fact checking inefficient. At the same time, creating tools for automatic detection is as challenging due to lack o…

ArticlesFact CheckingFake News Detection