paper-with-me

홈 › Papers

Not Too Short, Not Too Long: How LLM Response Length Shapes People's Critical Thinking in Error Detection

2026-03-06 · Natalie Friedman, Adelaide Nyanyo, Kevin Weatherwax, Lifei Wang, Chengchao Zhu, Zeshu Zhu, S. Joy Mountford arxiv

Large language models (LLMs) have become common decision-support tools across educational and professional contexts, raising questions about how their outputs shape human critical thinking. Prior work suggests that the amount of AI assistance can influence cognitive engagement, yet little is known about how specific properties of LLM outputs (e.g., response length) impacts users' critical evaluation of information. In this study, we examine whether the length of LLM responses shapes users' accuracy in evaluating LLM-generated reasoning on critical thinking tasks, particularly in interaction with the correctness of the LLM's reasoning. To begin evaluating this, we conducted a within-subjects experiment with 24 participants who completed 15 modified Watson--Glaser critical thinking items, each accompanied by an LLM-generated explanation that varied in length and correctness. Mixed-effects logistic regression revealed a strong and statistically reliable effect of LLM output correctness on participant accuracy, with participants more likely to answer correctly when the LLM's explanation was correct. Response length appeared to moderated this effect: when the LLM output was incorrect, medium-length explanations were associated with higher participant accuracy than either shorter or longer explanations, whereas accuracy remained high across lengths when the LLM output was correct. Together, these findings suggest that response length alone may be insufficient to support critical thinking, and that how reasoning is presented-including a potential advantage of mid-length explanations under some conditions-points to design opportunities for LLM-based decision-support systems that emphasize transparent reasoning and calibrated expressions of certainty.

📄 PDF Abstract BibTeX arXiv:2603.06878

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Comparison of window shapes and lengths in short-time feature extraction for classification of heart sound signals

2026-04-15 · Mahmoud Fakhry, Abeer FathAllah Brery arxiv

Heart sound signals, phonocardiography (PCG) signals, allow for the automatic diagnosis of potential cardiovascular pathology. Such classification task can be tackled using the bidirectional long short-term memory (biLST…

Efficient RL Training for Reasoning Models via Length-Aware Optimization

2025-05-18 · Danlong Yuan, Tian Xie, Shaohan Huang, Zhuocheng Gong 외

Large reasoning models, such as OpenAI o1 or DeepSeek R1, have demonstrated remarkable performance on reasoning tasks but often incur a long reasoning path with significant memory and time costs. Existing methods primari…

Math

Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training

2026-05-08 · Chen Wang, Hexuan Deng, Yining Zhang, Yuchen Zhang 외 arxiv

Reinforcement learning with verifiable rewards improves LLM reasoning but often induces overthinking, where models generate unnecessarily long reasoning traces. Existing methods mainly rely on length penalties or early-e…

Reinforcement Learning

Statistical reconstruction of pulse shapes from pulse streams

2023-02-06 · Marek W. Rupniewski

A short sample sequence of a finite-length pulse signal allows for its reconstruction only if the signal has a sparse representation in some basis. The recurrence of the pulse allows for a statistical approach to its rec…

Women worry about family, men about the economy: Gender differences in emotional responses to COVID-19

2020-04-17 · Isabelle van der Vegt, Bennett Kleinberg

Among the critical challenges around the COVID-19 pandemic is dealing with the potentially detrimental effects on people's mental health. Designing appropriate interventions and identifying the concerns of those most at …