paper-with-me

Papers

CompText: Visualizing, Comparing & Understanding Text Corpus

2022-07-27 · Suvi Varshney, Divjeet Singh Jas

A common practice in Natural Language Processing (NLP) is to visualize the text corpus without reading through the entire literature, still grasping the central idea and key points described. For a long time, researchers focused on extracting topics from the text and visualizing them based on their relative significance in the corpus. However, recently, researchers started coming up with more complex systems that not only expose the topics of the corpus but also word closely related to the topic to give users a holistic view. These detailed visualizations spawned research on comparing text corpora based on their visualization. Topics are often compared to idealize the difference between corpora. However, to capture greater semantics from different corpora, researchers have started to compare texts based on the sentiment of the topics related to the text. Comparing the words carrying the most weightage, we can get an idea about the important topics for corpus. There are multiple existing texts comparing methods present that compare topics rather than sentiments but we feel that focusing on sentiment-carrying words would better compare the two corpora. Since only sentiments can explain the real feeling of the text and not just the topic, topics without sentiments are just nouns. We aim to differentiate the corpus with a focus on sentiment, as opposed to comparing all the words appearing in the two corpora. The rationale behind this is, that the two corpora do not many have identical words for side-by-side comparison, so comparing the sentiment words gives us an idea of how the corpora are appealing to the emotions of the reader. We can argue that the entropy or the unexpectedness and divergence of topics should also be of importance and help us to identify key pivot points and the importance of certain topics in the corpus alongside relative sentiment.

📄 PDF Abstract BibTeX arXiv:2207.13771

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Visualizing and Comparing Convolutional Neural Networks

2014-12-20 · Wei Yu, Kuiyuan Yang, Yalong Bai, Hongxun Yao 외

Convolutional Neural Networks (CNNs) have achieved comparable error rates to well-trained human on ILSVRC2014 image classification task. To achieve better performance, the complexity of CNNs is continually increasing wit…

ClassificationGeneral Classificationimage-classificationImage Classification

Visualizing Facets of Text Complexity across Registers

2020-05-01 · LREC 2020 5 · Marina Santini, Arne Jonsson, Evelina Rennes

In this paper, we propose visualizing results of a corpus-based study on text complexity using radar charts. We argue that the added value of this type of visualisation is the polygonal shape that provides an intuitive g…

Visualizing textual models with in-text and word-as-pixel highlighting

2016-06-20 · Abram Handler, Su Lin Blodgett, Brendan O'Connor

We explore two techniques which use color to make sense of statistical text models. One method uses in-text annotations to illustrate a model's view of particular tokens in particular documents. Another uses a high-level…

Scattertext: a Browser-Based Tool for Visualizing how Corpora Differ

2017-03-02 · Jason S. Kessler

Scattertext is an open source tool for visualizing linguistic variation between document categories in a language-independent way. The tool presents a scatterplot, where each axis corresponds to the rank-frequency a term…

Visualizing and Understanding GANs

2019-03-27 · ICLR Workshop DeepGenStruct 2019 · David Bau, Jun-Yan Zhu, Hendrik Strobelt, Bolei Zhou 외

We present an analytic framework to visualize and understand GANs at the unit-, object-, and scene-level. We first identify a group of interpretable units that are closely related to object concepts with a segmentation-b…

Object