paper-with-me

Papers

Large Language Models Are Effective Human Annotation Assistants, But Not Good Independent Annotators

2025-03-09 · Feng Gu, Zongxia Li, Carlos Rafael Colon, Benjamin Evans, Ishani Mondal, Jordan Lee Boyd-Graber

Event annotation is important for identifying market changes, monitoring breaking news, and understanding sociological trends. Although expert annotators set the gold standards, human coding is expensive and inefficient. Unlike information extraction experiments that focus on single contexts, we evaluate a holistic workflow that removes irrelevant documents, merges documents about the same event, and annotates the events. Although LLM-based automated annotations are better than traditional TF-IDF-based methods or Event Set Curation, they are still not reliable annotators compared to human experts. However, adding LLMs to assist experts for Event Set Curation can reduce the time and mental effort required for Variable Annotation. When using LLMs to extract event variables to assist expert annotators, they agree more with the extracted variables than fully automated LLMs for annotation.

📄 PDF Abstract BibTeX arXiv:2503.06778

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
Focus 설명 없음

Similar Papers 제목 키워드 기반

Model-in-the-Loop (MILO): Accelerating Multimodal AI Data Annotation with LLMs

2024-09-16 · Yifan Wang, David Stevens, Pranay Shah, WenWen Jiang 외

The growing demand for AI training data has transformed data annotation into a global industry, but traditional approaches relying on human annotators are often time-consuming, labor-intensive, and prone to inconsistent …

Linear Alignment: A Closed-form Solution for Aligning Human Preferences without Tuning and Feedback

2024-01-21 · Songyang Gao, Qiming Ge, Wei Shen, Shihan Dou 외

The success of AI assistants based on Language Models (LLMs) hinges on Reinforcement Learning from Human Feedback (RLHF) to comprehend and align with user intentions. However, traditional alignment algorithms, such as PP…

Form

Generating Natural Questions from Images for Multimodal Assistants

2020-11-17 · Alkesh Patel, Akanksha Bindal, Hadas Kotek, Christopher Klein 외

Generating natural, diverse, and meaningful questions from images is an essential task for multimodal assistants as it confirms whether they have understood the object and scene in the images properly. The research in vi…

Common Sense ReasoningNatural QuestionsQuestion AnsweringQuestion Generation+3

dafny-annotator: AI-Assisted Verification of Dafny Programs

2024-11-05 · Gabriel Poesia, Chloe Loughridge, Nada Amin

Formal verification has the potential to drastically reduce software bugs, but its high additional cost has hindered large-scale adoption. While Dafny presents a promise to significantly reduce the effort to write verifi…

Friction

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning

2026-05-17 · Zeyu Wang, Chang Liu, Eduardus Tjitrahardja, Yuntao Wang 외 arxiv

Despite extensive efforts on egocentric video datasets and benchmarks, understanding users' internal states, which is crucial for enabling seamless AI assistant experiences, remains largely overlooked. In this work, we i…