paper-with-me

홈 › Papers

Annif at the GermEval-2025 LLMs4Subjects Task: Traditional XMTC Augmented by Efficient LLMs

2025-08-21 · Osma Suominen, Juho Inkinen, Mona Lehtinen arxiv

This paper presents the Annif system in the LLMs4Subjects shared task (Subtask 2) at GermEval-2025. The task required creating subject predictions for bibliographic records using large language models, with a special focus on computational efficiency. Our system, based on the Annif automated subject indexing toolkit, refines our previous system from the first LLMs4Subjects shared task, which produced excellent results. We further improved the system by using many small and efficient language models for translation and synthetic data generation and by using LLMs for ranking candidate subjects. Our system ranked 1st in the overall quantitative evaluation of and 1st in the qualitative evaluation of Subtask 2.

📄 PDF Abstract BibTeX arXiv:2508.15877

Code (0)

등록된 구현이 없습니다.

Tasks

Synthetic Data GenerationComputational Efficiency

Similar Papers 제목 키워드 기반

Annif at SemEval-2025 Task 5: Traditional XMTC augmented by LLMs

2025-04-28 · Osma Suominen, Juho Inkinen, Mona Lehtinen

This paper presents the Annif system in SemEval-2025 Task 5 (LLMs4Subjects), which focussed on subject indexing using large language models (LLMs). The task required creating subject predictions for bibliographic records…

Synthetic Data Generation

UPAppliedCL at GermEval 2021: Identifying Fact-Claiming and Engaging Facebook Comments Using Transformers

2021-09-01 · GermEval 2021 9 · Robin Schaefer, Manfred Stede

In this paper we present UPAppliedCL’s contribution to the GermEval 2021 Shared Task. In particular, we participated in Subtasks 2 (Engaging Comment Classification) and 3 (Fact-Claiming Comment Classification). While acc…

ClassificationEngaging Comment ClassificationFact-Claiming Comment Classification

AIT_FHSTP at GermEval 2021: Automatic Fact Claiming Detection with Multilingual Transformer Models

2021-09-01 · GermEval 2021 9 · Jaqueline Böck, Daria Liakhovets, Mina Schütz, Armin Kirchknopf 외

Spreading ones opinion on the internet is becoming more and more important. A problem is that in many discussions people often argue with supposed facts. This year’s GermEval 2021 focuses on this topic by incorporating a…

WLV-RIT at GermEval 2021: Multitask Learning with Transformers to Detect Toxic, Engaging, and Fact-Claiming Comments

2021-07-30 · GermEval 2021 9 · Skye Morgan, Tharindu Ranasinghe, Marcos Zampieri

This paper addresses the identification of toxic, engaging, and fact-claiming comments on social media. We used the dataset made available by the organizers of the GermEval-2021 shared task containing over 3,000 manually…

IRCologne at GermEval 2021: Toxicity Classification

2021-09-01 · GermEval 2021 9 · Fabian Haak, Björn Engelmann

In this paper, we describe the TH Köln’s submission for the Shared Task on the Identification of Toxic Comments at GermEval 2021. Toxicity is a severe and latent problem in comments in online discussions. Complex languag…

ClassificationLanguage ModelingLanguage ModellingToxic Comment Classification