paper-with-me

Papers

Towards Stable and Personalised Profiles for Lexical Alignment in Spoken Human-Agent Dialogue

2025-09-04 · Keara Schaaij, Roel Boumans, Tibor Bosse, Iris Hendrickx arxiv

Lexical alignment, where speakers start to use similar words across conversation, is known to contribute to successful communication. However, its implementation in conversational agents remains underexplored, particularly considering the recent advancements in large language models (LLMs). As a first step towards enabling lexical alignment in human-agent dialogue, this study draws on strategies for personalising conversational agents and investigates the construction of stable, personalised lexical profiles as a basis for lexical alignment. Specifically, we varied the amounts of transcribed spoken data used for construction as well as the number of items included in the profiles per part-of-speech (POS) category and evaluated profile performance across time using recall, coverage, and cosine similarity metrics. It was shown that smaller and more compact profiles, created after 10 min of transcribed speech containing 5 items for adjectives, 5 items for conjunctions, and 10 items for adverbs, nouns, pronouns, and verbs each, offered the best balance in both performance and data efficiency. In conclusion, this study offers practical insights into constructing stable, personalised lexical profiles, taking into account minimal data requirements, serving as a foundational step toward lexical alignment strategies in conversational agents.

📄 PDF Abstract BibTeX arXiv:2509.04104

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Improving the generation of personalised descriptions

2017-09-01 · WS 2017 9 · Thiago Castro Ferreira, Iv Paraboni, r{\'e}

Referring expression generation (REG) models that use speaker-dependent information require a considerable amount of training data produced by every individual speaker, or may otherwise perform poorly. In this work we pr…

Referring ExpressionReferring expression generationText Generation

Who Decides What Is Harmful? Content Moderation Policy Through A Multi-Agent Personalised Inference Framework

2026-05-02 · Ewelina Gajewska, Michal Wawer, Katarzyna Budzynska, Jaroslaw A. Chudziak arxiv

The increasing scale and complexity of online platforms raises critical policy questions around harmful content, digital well-being, and user autonomy. Traditional content moderation systems rely on centralised, top-down…

Estimating the number of household TV profiles based in customer behaviour using Gaussian mixture model averaging

2025-05-15 · Gabriel R. Palma, Sally McClean, Brahim Allan, Zeeshan Tariq 외

TV customers today face many choices from many live channels and on-demand services. Providing a personalised experience that saves customers time when discovering content is essential for TV providers. However, a reliab…

Do we read what we hear? Modeling orthographic influences on spoken word recognition

2021-04-01 · EACL 2021 2 · Nicole Macher, Badr M. Abdullah, Harm Brouwer, Dietrich Klakow

Theories and models of spoken word recognition aim to explain the process of accessing lexical knowledge given an acoustic realization of a word form. There is consensus that phonological and semantic information is cruc…

Multilingual Lexical Feature Analysis of Spoken Language for Predicting Major Depression Symptom Severity

2025-11-10 · Anastasiia Tokareva, Judith Dineley, Zoe Firth, Pauline Conde 외 arxiv

Background: Remotely captured spoken language could provide objective, regular indicators of depression symptom severity. However, research to date has largely used non-clinical, cross-sectional written language and comp…