paper-with-me

홈 › Papers

Enriching Social Science Research via Survey Item Linking

2024-12-20 · Tornike Tsereteli, Daniel Ruffinelli, Simone Paolo Ponzetto

Questions within surveys, called survey items, are used in the social sciences to study latent concepts, such as the factors influencing life satisfaction. Instead of using explicit citations, researchers paraphrase the content of the survey items they use in-text. However, this makes it challenging to find survey items of interest when comparing related work. Automatically parsing and linking these implicit mentions to survey items in a knowledge base can provide more fine-grained references. We model this task, called Survey Item Linking (SIL), in two stages: mention detection and entity disambiguation. Due to an imprecise definition of the task, existing datasets used for evaluating the performance for SIL are too small and of low-quality. We argue that latent concepts and survey item mentions should be differentiated. To this end, we create a high-quality and richly annotated dataset consisting of 20,454 English and German sentences. By benchmarking deep learning systems for each of the two stages independently and sequentially, we demonstrate that the task is feasible, but observe that errors propagate from the first stage, leading to a lower overall task performance. Moreover, mentions that require the context of multiple sentences are more challenging to identify for models in the first stage. Modeling the entire context of a document and combining the two stages into an end-to-end system could mitigate these problems in future work, and errors could additionally be reduced by collecting more diverse data and by improving the quality of the knowledge base. The data and code are available at https://github.com/e-tornike/SIL .

📄 PDF Abstract BibTeX arXiv:2412.15831

Code (1)

e-tornike/sil 공식 구현 pytorch

Tasks

BenchmarkingEntity DisambiguationSurvey

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

A methodology for co-constructing an interdisciplinary model: from model to survey, from survey to model

2020-11-27 · Elise Beck, Julie Dugdale, Carole Adam, Christelle Gaïdatzis 외

How should computer science and social science collaborate to build a common model? How should they proceed to gather data that is really useful to the modelling? How can they design a survey that is tailored to the targ…

modelSurvey

Word Embedding for Social Sciences: An Interdisciplinary Survey

2022-07-07 · Akira Matsui, Emilio Ferrara

To extract essential information from complex data, computer scientists have been developing machine learning models that learn low-dimensional representation mode. From such advances in machine learning research, not on…

BIG-bench Machine LearningSurvey

AI for social science and social science of AI: A Survey

2024-01-22 · Ruoxi Xu, Yingfei Sun, Mengjie Ren, Shiguang Guo 외

Recent advancements in artificial intelligence, particularly with the emergence of large language models (LLMs), have sparked a rethinking of artificial general intelligence possibilities. The increasing human-like capab…

Towards Automated Survey Variable Search and Summarization in Social Science Publications

2022-09-14 · Yavuz Selim Kartal, Sotaro Takeshita, Tornike Tsereteli, Kai Eckert 외

Nowadays there is a growing trend in many scientific disciplines to support researchers by providing enhanced information access through linking of publications and underlying datasets, so as to support research with inf…

Variable Detection

Query Expansion for Survey Question Retrieval in the Social Sciences

2015-06-18 · Dulisch Nadine, Kempf Andreas Oskar, Schaer Philipp

In recent years, the importance of research data and the need to archive and to share it in the scientific community have increased enormously. This introduces a whole new set of challenges for digital libraries. In the …

Retrieval