paper-with-me

홈 › Papers

Looking for Confirmations: An Effective and Human-Like Visual Dialogue Strategy

2021-09-11 · EMNLP 2021 11 · Alberto Testoni, Raffaella Bernardi

Generating goal-oriented questions in Visual Dialogue tasks is a challenging and long-standing problem. State-Of-The-Art systems are shown to generate questions that, although grammatically correct, often lack an effective strategy and sound unnatural to humans. Inspired by the cognitive literature on information search and cross-situational word learning, we design Confirm-it, a model based on a beam search re-ranking algorithm that guides an effective goal-oriented strategy by asking questions that confirm the model's conjecture about the referent. We take the GuessWhat?! game as a case-study. We show that dialogues generated by Confirm-it are more natural and effective than beam search decoding without re-ranking.

📄 PDF Abstract BibTeX arXiv:2109.05312

Code (1)

albertotestoni/confirm_it 공식 구현 pytorch

Tasks

Re-Ranking

Similar Papers 제목 키워드 기반

Vision: looking and seeing through our brain's information bottleneck

2025-03-24 · Li Zhaoping

Our brain recognizes only a tiny fraction of sensory input, due to an information processing bottleneck. This blinds us to most visual inputs. Since we are blind to this blindness, only a recent framework highlights this…

Guiding Interaction Behaviors for Multi-modal Grounded Language Learning

2017-08-01 · WS 2017 8 · Jesse Thomason, Jivko Sinapov, Raymond Mooney

Multi-modal grounded language learning connects language predicates to physical properties of objects in the world. Sensing with multiple modalities, such as audio, haptics, and visual colors and shapes while performing …

Grounded language learningRetrieval

Lexical Acquisition through Implicit Confirmations over Multiple Dialogues

2017-08-01 · WS 2017 8 · Kohei Ono, Ryu Takeda, Eric Nichols, Mikio Nakano 외

We address the problem of acquiring the ontological categories of unknown terms through implicit confirmation in dialogues. We develop an approach that makes implicit confirmation requests with an unknown term{'}s predic…

ChatbotTask-Oriented Dialogue Systems

Moving by Looking: Towards Vision-Driven Avatar Motion Generation

2025-09-23 · Markos Diomataris, Berat Mert Albaba, Giorgio Becherini, Partha Ghosh 외 arxiv

The way we perceive the world fundamentally shapes how we move, whether it is how we navigate in a room or how we interact with other humans. Current human motion generation methods, neglect this interdependency and use …

DualFete: Revisiting Teacher-Student Interactions from a Feedback Perspective for Semi-supervised Medical Image Segmentation

2025-11-12 · Le Yi, Wei Huang, Lei Zhang, Kefu Zhao 외 arxiv

The teacher-student paradigm has emerged as a canonical framework in semi-supervised learning. When applied to medical image segmentation, the paradigm faces challenges due to inherent image ambiguities, making it partic…

Semi-supervised Medical Image Segmentation