paper-with-me

Papers

Label Over Logic? How Source Cues Bias Human Fallacy Judgments More Than LLMs

2026-05-28 · Mahjabin Nahar, Nafis Irtiza Tripto, Aiping Xiong, Ting-Hao 'Kenneth' Huang, Dongwon Lee arxiv

As AI-generated and AI-assisted content floods online spaces, source labels attached to such content can distort human reasoning judgments, with downstream consequences for moderation, evaluation, and decision-making. Whether LLMs share this vulnerability, or offer more source-agnostic evaluation, remains an open question with direct implications for human-AI collaboration. We examine this issue using logical fallacies as a controlled setting to isolate source-label effects on reasoning quality, independent of domain knowledge. We conduct an online study (N=505) where participants are assigned to a source condition (human, AI, human with AI assistance, AI with human assistance, or no disclosure) and evaluate comments containing logical fallacies, comparing their judgments with those of LLMs (GPT-5.2, Gemini 2.5 Flash, Claude Sonnet 4.5), who were evaluated across the same source conditions. Human evaluators were significantly more susceptible to fallacies labeled as written by human or human with AI assistance and assigned higher trust and evaluation ratings in these conditions. LLM evaluations remained comparatively stable across source labels, though performance varied across models. Confidence levels were similarly high across conditions for both humans and LLMs, regardless of fallacy presence. Our findings indicate that source-label bias in reasoning evaluation is primarily a human vulnerability and highlight the potential of human-LLM collaboration in increasingly AI-mediated environments.

📄 PDF Abstract BibTeX arXiv:2605.29928

Code (0)

등록된 구현이 없습니다.

Tasks

Logical Fallacies

Similar Papers 제목 키워드 기반

Beyond Surface Cues: Disentangling Sociocultural Signals in Multilingual LLMs

2026-08-24 · Yuanjun Feng, Tanzhou Liu, Stefan Feuerriegel, Yash Raj Shrestha arxiv

Multilingual LLM outputs can vary across sociocultural contexts. However, evidence of cultural grounding can be misleading: identity labels may be inferred from explicit or indirect textual cues, while names and wording …

SleepBand: Single-Source Domain Generalization for Sleep Staging via Physiologically Structured Spectral Modeling

2026-07-06 · Zhi Lu, Yang Hu, Yan Chen arxiv

Generalizing sleep staging models to unseen datasets is challenging, and typical domain generalization (DG) methods often rely on multiple source domains or domain labels that are rarely available in practice. We tackle …

Single-Source Domain Generalization

PEACE: Cross-Platform Hate Speech Detection- A Causality-guided Framework

2023-06-15 · Paras Sheth, Tharindu Kumarage, Raha Moraffah, Aman Chadha 외

Hate speech detection refers to the task of detecting hateful content that aims at denigrating an individual or a group based on their religion, gender, sexual orientation, or other characteristics. Due to the different …

Hate Speech Detection

Cross Pseudo-Labeling for Semi-Supervised Audio-Visual Source Localization

2024-03-05 · Yuxin Guo, Shijie Ma, Yuhao Zhao, Hu Su 외

Audio-Visual Source Localization (AVSL) is the task of identifying specific sounding objects in the scene given audio cues. In our work, we focus on semi-supervised AVSL with pseudo-labeling. To address the issues with v…

Pseudo Label

Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judge

2026-04-07 · Xin Sun, Di Wu, Sijing Qin, Isao Echizen 외 arxiv

Large language models (LLMs) are increasingly used as automated evaluators (LLM-as-a-Judge). This work challenges its reliability by showing that trust judgments by LLMs are biased by disclosed source labels. Using a cou…