paper-with-me

홈 › Papers

SemEval-2025 Task 5: LLMs4Subjects -- LLM-based Automated Subject Tagging for a National Technical Library's Open-Access Catalog

2025-04-09 · Jennifer D'Souza, Sameer Sadruddin, Holger Israel, Mathias Begoin, Diana Slawig

We present SemEval-2025 Task 5: LLMs4Subjects, a shared task on automated subject tagging for scientific and technical records in English and German using the GND taxonomy. Participants developed LLM-based systems to recommend top-k subjects, evaluated through quantitative metrics (precision, recall, F1-score) and qualitative assessments by subject specialists. Results highlight the effectiveness of LLM ensembles, synthetic data generation, and multilingual processing, offering insights into applying LLMs for digital library classification.

📄 PDF Abstract BibTeX arXiv:2504.07199

Code (1)

natlibfi/annif-llms4subjects 공식 구현

Tasks

Synthetic Data Generation

Methods 이 논문이 사용한 방법론

Library 설명 없음

Similar Papers 제목 키워드 기반

DNB-AI-Project at SemEval-2025 Task 5: An LLM-Ensemble Approach for Automated Subject Indexing

2025-04-30 · Lisa Kluge, Maximilian Kähler

This paper presents our system developed for the SemEval-2025 Task 5: LLMs4Subjects: LLM-based Automated Subject Tagging for a National Technical Library's Open-Access Catalog. Our system relies on prompting a selection …

Annif at SemEval-2025 Task 5: Traditional XMTC augmented by LLMs

2025-04-28 · Osma Suominen, Juho Inkinen, Mona Lehtinen

This paper presents the Annif system in SemEval-2025 Task 5 (LLMs4Subjects), which focussed on subject indexing using large language models (LLMs). The task required creating subject predictions for bibliographic records…

Synthetic Data Generation

Annif at the GermEval-2025 LLMs4Subjects Task: Traditional XMTC Augmented by Efficient LLMs

2025-08-21 · Osma Suominen, Juho Inkinen, Mona Lehtinen arxiv

This paper presents the Annif system in the LLMs4Subjects shared task (Subtask 2) at GermEval-2025. The task required creating subject predictions for bibliographic records using large language models, with a special foc…

Synthetic Data GenerationComputational Efficiency

NBF at SemEval-2025 Task 5: Light-Burst Attention Enhanced System for Multilingual Subject Recommendation

2025-05-06 · Baharul Islam, Nasim Ahmad, Ferdous Ahmed Barbhuiya, Kuntal Dey

We present our system submission for SemEval 2025 Task 5, which focuses on cross-lingual subject classification in the English and German academic domains. Our approach leverages bilingual data during training, employing…

GPURetrievalSentenceSentence Embeddings

Semantic Relation Classification via Convolutional Neural Networks with Simple Negative Sampling

2015-06-25 · EMNLP 2015 9 · Kun Xu, Yansong Feng, Songfang Huang, Dongyan Zhao

Syntactic features play an essential role in identifying relationship in a sentence. Previous neural network models often suffer from irrelevant information introduced when subjects and objects are in a long distance. In…

General ClassificationRelationRelation ClassificationSentence