paper-with-me

홈 › Papers

The Dynamics of Human and AI-Generated Language: How Semantics Fluctuates across Different Timescales

2026-06-09 · Han-Jen Chang, Yasir Çatal, Angelika Wolman, Agustín Ibáñez, David Smith, I-Wen Su, Kai-Yuan Cheng, Georg Northoff arxiv

Spoken language, whether produced by humans or large language models (LLM), unfolds over time with varying semantic content. However, we still lack simple, interpretable time-series features that capture how generic versus specific content is distributed over time, and that can be used to compare human and AI-generated speech. We introduce a semantic-timescale analysis pipeline that turns word-level transcripts with timestamps into semantic time-series. For each spoken narrative, we compute (i) semantic specificity using WordNet-based word depth and (ii) contextual similarity using SBERT embeddings and quantify their temporal dependence using autocorrelation-window measures (ACW-0 and related metrics). We then compare original speech to multiple shuffled controls that selectively disrupt lexical identity, temporal order, and word duration. Across human-read autobiographical narratives, TTS readings, and LLM-generated texts rendered with TTS, we find that segments with longer ACW-0 in the semantic time-series tend to contain more generic vocabulary, whereas segments with shorter ACW-0 are enriched in more specific words. These associations are strongly attenuated or abolished when word order and timing are randomized, indicating that ACW-based measures capture non-trivial temporal organization of semantic content beyond static lexical distributions. Our results suggest that ACW-based semantic timescales are a useful family of features for analyzing and comparing the temporal structure of human and AI-generated speech.

📄 PDF Abstract BibTeX arXiv:2606.11371

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Enhancing Systematic Reviews with Large Language Models: Using GPT-4 and Kimi

2025-04-28 · Dandan Chen Kaptur, Yue Huang, Xuejun Ryan Ji, Yanhui Guo 외

This research delved into GPT-4 and Kimi, two Large Language Models (LLMs), for systematic reviews. We evaluated their performance by comparing LLM-generated codes with human-generated codes from a peer-reviewed systemat…

LLM Enhanced Action Recognition via Hierarchical Global-Local Skeleton-Language Model

2026-03-28 · Ruosi Wang, Fangwei Zuo, Lei Li, Zhaoqiang Xia arxiv

Skeleton-based human action recognition has achieved remarkable progress in recent years. However, most existing GCN-based methods rely on short-range motion topologies, which not only struggle to capture long-range join…

Action Recognition

On the Semantics of Large Language Models

2025-07-07 · Martin Schuele arxiv

Large Language Models (LLMs) such as ChatGPT demonstrated the potential to replicate human language abilities through technology, ranging from text generation to engaging in conversations. However, it remains controversi…

Text Generation

Eco-evolutionary dynamics of social dilemmas

2016-05-24

Social dilemmas are an integral part of social interactions. Cooperative actions, ranging from secreting extra-cellular products in microbial populations to donating blood in humans, are costly to the actor and hence cre…

Contrasting Human- and Machine-Generated Word-Level Adversarial Examples for Text Classification

2021-09-09 · EMNLP 2021 11 · Maximilian Mozes, Max Bartolo, Pontus Stenetorp, Bennett Kleinberg 외

Research shows that natural language processing models are generally considered to be vulnerable to adversarial attacks; but recent work has drawn attention to the issue of validating these adversarial inputs against cer…

Sentiment AnalysisSentiment Classificationtext-classificationText Classification+1