paper-with-me

홈 › Papers

Shape vs. Context: Examining Human--AI Gaps in Ambiguous Japanese Character Recognition

2026-02-27 · Daichi Haraguchi arxiv

High text recognition performance does not guarantee that Vision-Language Models (VLMs) share human-like decision patterns when resolving ambiguity. We investigate this behavioral gap by directly comparing humans and VLMs using continuously interpolated Japanese character shapes generated via a $β$-VAE. We estimate decision boundaries in a single-character recognition (shape-only task) and evaluate whether VLM responses align with human judgments under shape in context (i.e., embedding an ambiguous character near the human decision boundary in word-level context). We find that human and VLM decision boundaries differ in the shape-only task, and that shape in context can improve human alignment in some conditions. These results highlight qualitative behavioral differences, offering foundational insights toward human--VLM alignment benchmarking.

📄 PDF Abstract BibTeX arXiv:2602.23746

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Can LLMs Detect Ambiguous Plural Reference? An Analysis of Split-Antecedent and Mereological Reference

2025-10-06 · Dang Anh, Rick Nouwen, Massimo Poesio arxiv

Our goal is to study how LLMs represent and interpret plural reference in ambiguous and unambiguous contexts. We ask the following research questions: (1) Do LLMs exhibit human-like preferences in representing plural ref…

Do Context-Aware Translation Models Pay the Right Attention?

2021-05-14 · ACL 2021 5 · Kayo Yin, Patrick Fernandes, Danish Pruthi, Aditi Chaudhary 외

Context-aware machine translation models are designed to leverage contextual information, but often fail to do so. As a result, they inaccurately disambiguate pronouns and polysemous words that require context for resolu…

Machine TranslationTranslation

Co-Writing with AI, on Human Terms: Aligning Research with User Demands Across the Writing Process

2025-04-16 · Mohi Reza, Jeb Thomas-Mitchell, Peter Dushniku, Nathan Laundry 외

As generative AI tools like ChatGPT become integral to everyday writing, critical questions arise about how to preserve writers' sense of agency and ownership when using these tools. Yet, a systematic understanding of ho…

Recovering Sub-threshold S-wave Arrivals in Deep Learning Phase Pickers via Shape-Aware Loss

2025-11-10 · Chun-Ming Huang, Li-Heng Chang, I-Hsin Chang, An-Sheng Lee 외 arxiv

Deep learning has transformed seismic phase picking, but a systematic failure mode persists: for some S-wave arrivals that appear unambiguous to human analysts, the model produces only a distorted peak trapped below the …

Word forms - not just their lengths- are optimized for efficient communication

2017-03-06 · Stephan C. Meylan, Thomas L. Griffiths

The inverse relationship between the length of a word and the frequency of its use, first identified by G.K. Zipf in 1935, is a classic empirical law that holds across a wide range of human languages. We demonstrate that…