paper-with-me

Papers

Statistically Significant Detection of Linguistic Change

2014-11-12 · Vivek Kulkarni, Rami Al-Rfou, Bryan Perozzi, Steven Skiena

We propose a new computational approach for tracking and detecting statistically significant linguistic shifts in the meaning and usage of words. Such linguistic shifts are especially prevalent on the Internet, where the rapid exchange of ideas can quickly change a word's meaning. Our meta-analysis approach constructs property time series of word usage, and then uses statistically sound change point detection algorithms to identify significant linguistic shifts. We consider and analyze three approaches of increasing complexity to generate such linguistic property time series, the culmination of which uses distributional characteristics inferred from word co-occurrences. Using recently proposed deep neural language models, we first train vector representations of words for each time period. Second, we warp the vector spaces into one unified coordinate system. Finally, we construct a distance-based distributional time series for each word to track it's linguistic displacement over time. We demonstrate that our approach is scalable by tracking linguistic change across years of micro-blogging using Twitter, a decade of product reviews using a corpus of movie reviews from Amazon, and a century of written books using the Google Book-ngrams. Our analysis reveals interesting patterns of language usage change commensurate with each medium.

📄 PDF Abstract BibTeX arXiv:1411.3315

Code (0)

등록된 구현이 없습니다.

Tasks

Change Point DetectionTime SeriesTime Series Analysis

Similar Papers 제목 키워드 기반

What if Deception Cannot be Detected? A Cross-Linguistic Study on the Limits of Deception Detection from Text

2025-05-19 · Aswathy Velutharambath, Kai Sassenberg, Roman Klinger

Can deception be detected solely from written text? Cues of deceptive communication are inherently subtle, even more so in text-only communication. Yet, prior studies have reported considerable success in automatic decep…

Deception Detection

Assert, don't describe: Linguistic features that shift LLM reasoning about animal welfare

2026-04-30 · Jasmine Brazilek, Harper Dunn arxiv

Animal-welfare advocates produce a lot of writing, and increasingly that writing trains the language models that millions of people then ask about animal welfare. Using vocabulary-matched stance-contrast probes on a held…

Statistically significant detection of semantic shifts using contextual word embeddings

2021-04-08 · EMNLP (Eval4NLP) 2021 11 · Yang Liu, Alan Medlar, Dorota Glowacka

Detecting lexical semantic change in smaller data sets, e.g. in historical linguistics and digital humanities, is challenging due to a lack of statistical power. This issue is exacerbated by non-contextual embedding mode…

Word Embeddings

To Write or to Automate Linguistic Prompts, That Is the Question

2026-03-26 · Marina Sánchez-Torrón, Daria Akselrod, Jason Rauchwerk arxiv

LLM performance is highly sensitive to prompt design, yet whether automatic prompt optimization can replace expert prompt engineering in linguistic tasks remains unexplored. We present the first systematic comparison of …

Prompt Engineering

Online Graph-Based Change-Point Detection for High Dimensional Data

2019-06-07 · Yang-Wen Sun, Katerina Papagiannouli, Vladmir Spokoiny

Online change-point detection (OCPD) is important for application in various areas such as finance, biology, and the Internet of Things (IoT). However, OCPD faces major challenges due to high-dimensionality, and it is st…

Change Point DetectionVocal Bursts Intensity Prediction