paper-with-me

홈 › Papers

Human-Like Coarse Object Representations in Vision Models

2026-02-12 · Andrey Gizdov, Andrea Procopio, Yichen Li, Daniel Harari, Tomer Ullman arxiv

Humans appear to represent objects for intuitive physics with coarse, volumetric bodies'' that smooth concavities - trading fine visual details for efficient physical predictions - yet their internal structure is largely unknown. Segmentation models, in contrast, optimize pixel-accurate masks that may misalign with such bodies. We ask whether and when these models nonetheless acquire human-like bodies. Using a time-to-collision (TTC) behavioral paradigm, we introduce a comparison pipeline and alignment metric, then vary model training time, size, and effective capacity via pruning. Across all manipulations, alignment with human behavior follows an inverse U-shaped curve: small/briefly trained/pruned models under-segment into blobs; large/fully trained models over-segment with boundary wiggles; and an intermediate ideal body granularity'' best matches humans. This suggests human-like coarse bodies emerge from resource constraints rather than bespoke biases, and points to simple knobs - early checkpoints, modest architectures, light pruning - for eliciting physics-efficient representations. We situate these results within resource-rational accounts balancing recognition detail against physical affordances.

📄 PDF Abstract BibTeX arXiv:2602.12486

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards aligned body representations in vision models

2025-11-29 · Andrey Gizdov, Andrea Procopio, Yichen Li, Daniel Harari 외 arxiv

Human physical reasoning relies on internal "body" representations - coarse, volumetric approximations that capture an object's extent and support intuitive predictions about motion and physics. While psychophysical evid…

Semantic Segmentation

Extremely coarse learning objectives induce human-aligned representations in AI vision models

2026-05-07 · Yash Mehta, Michael F. Bonner arxiv

Artificial neural networks trained on visual tasks develop internal representations resembling those of the primate visual system, a discovery that has guided a decade of computational neuroscience. Research on building …

Investigating Fine- and Coarse-grained Structural Correspondences Between Deep Neural Networks and Human Object Image Similarity Judgments Using Unsupervised Alignment

2025-05-22 · Soh Takahashi, Masaru Sasaki, Ken Takeda, Masafumi Oizumi

The learning mechanisms by which humans acquire internal representations of objects are not fully understood. Deep neural networks (DNNs) have emerged as a useful tool for investigating this question, as they have intern…

ObjectSelf-Supervised Learning

Relational Knowledge Distillation Brings DNN Representations Close Enough to Humans to Be Aligned Without Supervision

2026-08-28 · Yuria Shimizu, Soh Takahashi, Takato Horii, Masafumi Oizumi arxiv

Linking the internal representations of deep neural networks (DNNs) to human mental representations is important for using DNNs as computational models of human vision. Existing DNN representations remain insufficiently …

Knowledge Distillation

Improving Span-based Question Answering Systems with Coarsely Labeled Data

2018-11-05 · Hao Cheng, Ming-Wei Chang, Kenton Lee, Ankur Parikh 외

We study approaches to improve fine-grained short answer Question Answering models by integrating coarse-grained data annotated for paragraph-level relevance and show that coarsely annotated data can bring significant pe…

Multi-Task LearningQuestion Answering