Who mentions whom? Recognizing political actors in proceedings
We show that it is straightforward to train a state of the art named entity tagger (spaCy) to recognize political actors in Dutch parliamentary proceedings with high accuracy. The tagger was trained on 3.4K manually labeled examples, which were created in a modest 2.5 days work. This resource is made available on github. Besides proper nouns of persons and political parties, the tagger can recognize quite complex definite descriptions referring to cabinet ministers, ministries, and parliamentary committees. We also provide a demo search engine which employs the tagged entities in its SERP and result summaries.
Code (0)
등록된 구현이 없습니다.
Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Who Sides with Whom? Towards Computational Construction of Discourse Networks for Political Debates
Understanding the structures of political debates (which actors make what claims) is essential for understanding democratic political decision making. The vision of computational construction of such discourse networks f…
Decision MakingKnowledge Base PopulationConcept Identification of Directly and Indirectly Related Mentions Referring to Groups of Persons
Unsupervised concept identification through clustering, i.e., identification of semantically related words and phrases, is a common approach to identify contextual primitives employed in various use cases, e.g., text dim…
ArticlesClusteringDimensionality ReductionEntity Resolution``Who Mentions Whom?''- Understanding the Psycho-Sociological Aspects of Twitter Mention Network
Labeling Gaps Between Words: Recognizing Overlapping Mentions with Mention Separators
In this paper, we propose a new model that is capable of recognizing overlapping mentions. We introduce a novel notion of mention separators that can be effectively used to capture how mentions overlap with one another. …
A Greek Parliament Proceedings Dataset for Computational Linguistics and Political Analysis
Large, diachronic datasets of political discourse are hard to come across, especially for resource-lean languages such as Greek. In this paper, we introduce a curated dataset of the Greek Parliament Proceedings that exte…