paper-with-me

홈 › Papers

Iconary: A Pictionary-Based Game for Testing Multimodal Communication with Drawings and Text

2021-12-01 · EMNLP 2021 11 · Christopher Clark, Jordi Salvador, Dustin Schwenk, Derrick Bonafilia, Mark Yatskar, Eric Kolve, Alvaro Herrasti, Jonghyun Choi, Sachin Mehta, Sam Skjonsberg, Carissa Schoenick, Aaron Sarnat, Hannaneh Hajishirzi, Aniruddha Kembhavi, Oren Etzioni, Ali Farhadi

Communicating with humans is challenging for AIs because it requires a shared understanding of the world, complex semantics (e.g., metaphors or analogies), and at times multi-modal gestures (e.g., pointing with a finger, or an arrow in a diagram). We investigate these challenges in the context of Iconary, a collaborative game of drawing and guessing based on Pictionary, that poses a novel challenge for the research community. In Iconary, a Guesser tries to identify a phrase that a Drawer is drawing by composing icons, and the Drawer iteratively revises the drawing to help the Guesser in response. This back-and-forth often uses canonical scenes, visual metaphor, or icon compositions to express challenging words, making it an ideal test for mixing language and visual/symbolic communication in AI. We propose models to play Iconary and train them on over 55,000 games between human players. Our models are skillful players and are able to employ world knowledge in language models to play with words unseen during training. Elite human players outperform our models, particularly at the drawing task, leaving an important gap for future research to address. We release our dataset, code, and evaluation setup as a challenge to the community at http://www.github.com/allenai/iconary.

📄 PDF Abstract BibTeX arXiv:2112.00800

Code (1)

allenai/iconary 공식 구현 pytorch

Tasks

World Knowledge

Similar Papers 제목 키워드 기반

DrawMon: A Distributed System for Detection of Atypical Sketch Content in Concurrent Pictionary Games

2022-11-10 · Nikhil Bansal, Kartik Gupta, Kiruthika Kannan, Sivani Pentapati 외

Pictionary, the popular sketch-based guessing game, provides an opportunity to analyze shared goal cooperative game play in restricted communication settings. However, some players occasionally draw atypical sketch conte…

Sketchtopia: A Dataset and Foundational Agents for Benchmarking Asynchronous Multimodal Communication with Iconic Feedback

2025-01-01 · CVPR 2025 1 · Mohd Hozaifa Khan, Ravi Kiran Sarvadevabhatla

We introduce Sketchtopia, a large-scale dataset and AI framework designed to explore goal-driven, multimodal communication through asynchronous interactions in a Pictionary-inspired setup. Sketchtopia captures natura…

Benchmarking

Game of Sketches: Deep Recurrent Models of Pictionary-style Word Guessing

2018-01-29 · Ravi Kiran Sarvadevabhatla, Shiv Surya, Trisha Mittal, Venkatesh Babu Radhakrishnan

The ability of intelligent agents to play games in human-like fashion is popularly considered a benchmark of progress in Artificial Intelligence. Similarly, performance on multi-disciplinary tasks such as Visual Question…

Question AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Enabling My Robot To Play Pictionary : Recurrent Neural Networks For Sketch Recognition

2016-08-11 · Ravi Kiran Sarvadevabhatla, Jogendra Kundu, Babu R. Venkatesh

Freehand sketching is an inherently sequential process. Yet, most approaches for hand-drawn sketch recognition either ignore this sequential aspect or exploit it in an ad-hoc manner. In our work, we propose a recurrent n…

ObjectObject RecognitionSketch Recognition

MOTIF: Contextualized Images for Complex Words to Improve Human Reading

2022-06-01 · LREC 2022 6 · Xintong Wang, Florian Schneider, Özge Alacam, Prateek Chaudhury 외

MOTIF (MultimOdal ConTextualized Images For Language Learners) is a multimodal dataset that consists of 1125 comprehension texts retrieved from Wikipedia Simple Corpus. Allowing multimodal processing or enriching the con…

Reading Comprehension