paper-with-me

홈 › Papers

Modeling Pathology-Like Behavioral Patterns in Language Models Through Behavioral Fine-Tuning

2026-05-21 · Nicola Milano, Davide Marocco arxiv

Large language models are increasingly used as computational tools for modeling human-like behavior. We introduce a behavioral induction framework that modifies model policies through fine-tuning on structured decision-making tasks: using synthetic datasets inspired by maladaptive behavioral patterns, including depression and paranoia, we train transformer-based language models to consistently select specific classes of actions across diverse contexts. We then test whether this behavioral optimization produces systematic changes in generative distributions. Across two architectures, fine-tuned models show stable, context-general shifts in next-token probability distributions, including increased probability assigned to negative and threat-related interpretations in open-ended language tasks. These effects generalize beyond training contexts and are detectable in qualitative completions, psychometric-style evaluations, and quantitative distributional metrics such as Jensen-Shannon divergence. Induced behavioral profiles also show partial specificity. Models optimized for different behavioral patterns exhibit dissociable response tendencies across evaluation probes, suggesting that structured behavioral training produces differentiated policy-level biases rather than generic distributional skew. We interpret these findings as evidence that consistent behavioral optimization in LLMs can generate stable behavioral and distributional patterns consistent with altered latent priors, linking action selection and language generation. More broadly, the results support a view of LLMs as policy-based systems in which behavioral constraints shape emergent representational structure, highlighting their potential as controlled testbeds for studying the relationship between behavior, interpretation, and generative language in computational models of cognition.

📄 PDF Abstract BibTeX arXiv:2605.22356

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Alignment Backfire: Language-Dependent Reversal of Safety Interventions Across 16 Languages in LLM Multi-Agent Systems

2026-03-05 · Hiroki Fukui arxiv

In perpetrator treatment, a recurring observation is the dissociation between insight and action: offenders articulate remorse yet behavioral change does not follow. We report four preregistered studies (1,584 multi-agen…

Intervening on psychopathology networks: Evaluating intervention targets through simulations

2022-08-01 · Methods 2022 8 · Gabriela Lunansky

Identifying the different influences of symptoms in dynamic psychopathology models may hold promise for increasing treatment efficacy in clinical applications. Dynamic psychopathology models study the behavioral patterns…

Harmonizing Large Language Models with Collaborative Behavioral Signals for Conversational Recommendation

2025-03-12 · Guanrong Li, Kuo Tian, Jinnan Qi, Qinghan Fu 외

Conversational recommendation frameworks have gained prominence as a dynamic paradigm for delivering personalized suggestions via interactive dialogues. The incorporation of advanced language understanding techniques has…

Collaborative FilteringConversational Recommendation

DevBench: A multimodal developmental benchmark for language learning

2024-06-14 · Alvin Wei Ming Tan, Sunny Yu, Bria Long, Wanjing Anya Ma 외

How (dis)similar are the learning trajectories of vision-language models and children? Recent modeling work has attempted to understand the gap between models' and humans' data efficiency by constructing models trained o…

TRACE-Bot: Detecting Emerging LLM-Driven Social Bots via Implicit Semantic Representations and AIGC-Enhanced Behavioral Patterns

2026-04-02 · Zhongbo Wang, Zhiyu Lin, Zhu Wang, Haizhou Wang arxiv

Large Language Model-driven (LLM-driven) social bots pose a growing threat to online discourse by generating human-like content that evades conventional detection. Existing methods suffer from limited detection accuracy …