A Text-based Approach For Link Prediction on Wikipedia Articles
This paper present our work in the DSAA 2023 Challenge about Link Prediction for Wikipedia Articles. We use traditional machine learning models with POS tags (part-of-speech tags) features extracted from text to train the classification model for predicting whether two nodes has the link. Then, we use these tags to test on various machine learning models. We obtained the results by F1 score at 0.99999 and got 7th place in the competition. Our source code is publicly available at this link: https://github.com/Tam1032/DSAA2023-Challenge-Link-prediction-DS-UIT_SAT
Code (1)
Tasks
ArticlesLink PredictionPOSPredictionSimilar Papers 제목 키워드 기반
Link Prediction for Wikipedia Articles as a Natural Language Inference Task
Link prediction task is vital to automatically understanding the structure of large knowledge bases. In this paper, we present our system to solve this task at the Data Science and Advanced Analytics 2023 Competition "Ef…
ArticlesLink PredictionNatural Language InferencePrediction+2Matching Cultural Heritage items to Wikipedia
Digitised Cultural Heritage (CH) items usually have short descriptions and lack rich contextual information. Wikipedia articles, on the contrary, include in-depth descriptions and links to related articles, which motivat…
ArticlesEntity LinkingPredicting Links on Wikipedia with Anchor Text Information
Wikipedia, the largest open-collaborative online encyclopedia, is a corpus of documents bound together by internal hyperlinks. These links form the building blocks of a large network whose structure contains important in…
ArticlesLink PredictionWEXEA: Wikipedia EXhaustive Entity Annotation
Building predictive models for information extraction from text, such as named entity recognition or the extraction of semantic relationships between named entities in text, requires a large corpus of annotated text. Wik…
Articlesnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+1DBpedia NIF: Open, Large-Scale and Multilingual Knowledge Extraction Corpus
In the past decade, the DBpedia community has put significant amount of effort on developing technical infrastructure and methods for efficient extraction of structured information from Wikipedia. These efforts have been…
Articles