paper-with-me

Papers

Wikipedia as a Resource for Text Analysis and Retrieval

2019-07-01 · ACL 2019 7 · Marius Pasca

This tutorial examines the role of Wikipedia in tasks related to text analysis and retrieval. Text analysis tasks, which take advantage of Wikipedia, include coreference resolution, word sense and entity disambiguation and information extraction. In information retrieval, a better understanding of the structure and meaning of queries helps in matching queries against documents, clustering search results, answer and entity retrieval and retrieving knowledge panels for queries asking about popular entities.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Clusteringcoreference-resolutionCoreference ResolutionEntity DisambiguationEntity RetrievalInformation RetrievalRetrieval

Similar Papers 제목 키워드 기반

The Role of Wikipedia in Text Analysis and Retrieval

2016-12-01 · COLING 2016 12 · Marius Pa{\c{s}}ca

This tutorial examines the characteristics, advantages and limitations of Wikipedia relative to other existing, human-curated resources of knowledge; derivative resources, created by converting semi-structured content in…

Coreference ResolutionInformation RetrievalRetrieval

Wikipedia Text Reuse: Within and Without

2018-12-21 · Alshomary Milad, Völske Michael, Licht Tristan, Wachsmuth Henning 외

We study text reuse related to Wikipedia at scale by compiling the first corpus of text reuse cases within Wikipedia as well as without (i.e., reuse of Wikipedia text in a sample of the Common Crawl). To discover reuse b…

Retrieval

Detecting Corpus-Level Knowledge Inconsistencies in Wikipedia with Large Language Models

2025-09-27 · Sina J. Semnani, Jirayu Burapacheep, Arpandeep Khatua, Thanawan Atchariyachanvanit 외 arxiv

Wikipedia is the largest open knowledge corpus, widely used worldwide and serving as a key resource for training large language models (LLMs) and retrieval-augmented generation (RAG) systems. Ensuring its accuracy is the…

Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR

2025-11-07 · Grigory Kovalev, Natalia Loukachevitch, Mikhail Tikhomirov, Olga Babina 외 arxiv

In this paper, we present a novel series of Russian information retrieval datasets constructed from the "Did you know..." section of Russian Wikipedia. Our datasets support a range of retrieval tasks, including fact-chec…

Information Retrieval

Design Challenges in Low-resource Cross-lingual Entity Linking

2020-05-02 · EMNLP 2020 11 · Xingyu Fu, Weijia Shi, Xiaodong Yu, Zian Zhao 외

Cross-lingual Entity Linking (XEL), the problem of grounding mentions of entities in a foreign language text into an English knowledge base such as Wikipedia, has seen a lot of research in recent years, with a range of p…

Cross-Lingual Entity LinkingEntity Linking