paper-with-me

홈 › Papers

Topic Extraction and Bundling of Related Scientific Articles

2012-12-21 · Shameem A Puthiya Parambath

Automatic classification of scientific articles based on common characteristics is an interesting problem with many applications in digital library and information retrieval systems. Properly organized articles can be useful for automatic generation of taxonomies in scientific writings, textual summarization, efficient information retrieval etc. Generating article bundles from a large number of input articles, based on the associated features of the articles is tedious and computationally expensive task. In this report we propose an automatic two-step approach for topic extraction and bundling of related articles from a set of scientific articles in real-time. For topic extraction, we make use of Latent Dirichlet Allocation (LDA) topic modeling techniques and for bundling, we make use of hierarchical agglomerative clustering techniques. We run experiments to validate our bundling semantics and compare it with existing models in use. We make use of an online crowdsourcing marketplace provided by Amazon called Amazon Mechanical Turk to carry out experiments. We explain our experimental setup and empirical results in detail and show that our method is advantageous over existing ones.

📄 PDF Abstract BibTeX arXiv:1212.5423

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesClusteringInformation RetrievalRetrieval

Similar Papers 제목 키워드 기반

Exploring the evolution of research topics during the COVID-19 pandemic

2023-10-05 · Francesco Invernici, Anna Bernasconi, Stefano Ceri

The COVID-19 pandemic has changed the research agendas of most scientific communities, resulting in an overwhelming production of research articles in a variety of domains, including medicine, virology, epidemiology, eco…

ArticlesEpidemiologyVirology

A Joint Learning Approach based on Self-Distillation for Keyphrase Extraction from Scientific Documents

2020-10-22 · COLING 2020 8 · Tuan Manh Lai, Trung Bui, Doo Soon Kim, Quan Hung Tran

Keyphrase extraction is the task of extracting a small set of phrases that best describe a document. Most existing benchmark datasets for the task typically have limited numbers of annotated documents, making it challeng…

ArticlesKeyphrase Extraction

Natural Language Processing for Intelligent Access to Scientific Information

2016-12-01 · COLING 2016 12 · Horacio Saggion, Francesco Ronzano

During the last decade the amount of scientific information available on-line increased at an unprecedented rate. As a consequence, nowadays researchers are overwhelmed by an enormous and continuously growing number of a…

ArticlesNatural Language InferenceQuestion Answering

On Representation Learning for Scientific News Articles Using Heterogeneous Knowledge Graphs

2021-04-12 · Angelika Romanou, Panayiotis Smeros, Karl Aberer

In the era of misinformation and information inflation, the credibility assessment of the produced news is of the essence. However, fact-checking can be challenging considering the limited references presented in the new…

ArticlesFact CheckingGraph Neural NetworkKnowledge Graphs+4

Context Selection for Hypothesis and Statistical Evidence Extraction from Full-Text Scientific Articles

2026-03-22 · Sai Koneru, Jian Wu, Sarah Rajtmajer arxiv

Extracting hypotheses and their supporting statistical evidence from full-text scientific articles is central to the synthesis of empirical findings, but remains difficult due to document length and the distribution of s…