Recognizing Descriptive Wikipedia Categories for Historical Figures
Wikipedia is a useful knowledge source that benefits many applications in language processing and knowledge representation. An important feature of Wikipedia is that of categories. Wikipedia pages are assigned different categories according to their contents as human-annotated labels which can be used in information retrieval, ad hoc search improvements, entity ranking and tag recommendations. However, important pages are usually assigned too many categories, which makes it difficult to recognize the most important ones that give the best descriptions. In this paper, we propose an approach to recognize the most descriptive Wikipedia categories. We observe that historical figures in a precise category presumably are mutually similar and such categorical coherence could be evaluated via texts or Wikipedia links of corresponding members in the category. We rank descriptive level of Wikipedia categories according to their coherence and our ranking yield an overall agreement of 88.27% compared with human wisdom.
Code (0)
등록된 구현이 없습니다.
Tasks
DescriptiveInformation RetrievalRetrievalTAGSimilar Papers 제목 키워드 기반
Studying the Wikipedia Hyperlink Graph for Relatedness and Disambiguation
Hyperlinks and other relations in Wikipedia are a extraordinary resource which is still not fully understood. In this paper we study the different types of links in Wikipedia, and contrast the use of the full graph with …
Entity DisambiguationTowards Recognizing Unseen Categories in Unseen Domains
Current deep visual recognition systems suffer from severe performance degradation when they encounter new images from classes and scenarios unseen during training. Hence, the core challenge of Zero-Shot Learning (ZSL) i…
Domain GeneralizationZero-Shot LearningZero-Shot Learning + Domain GeneralizationIncorporating Figure Captions and Descriptive Text in MeSH Term Indexing
The goal of text classification is to automatically assign categories to documents. Deep learning automatically learns effective features from data instead of adopting human-designed features. In this paper, we focus spe…
ArticlesDeep LearningDescriptiveDocument Classification+3Figure Descriptive Text Extraction using Ontological Representation
Experimental research publications provide figure form resources including graphs, charts, and any type of images to effectively support and convey methods and results. To describe figures, authors add captions, which ar…
ArticlesDescriptiveSentenceSentence Classification