paper-with-me

Papers

The REX corpora: A collection of multimodal corpora of referring expressions in collaborative problem solving dialogues

2012-05-01 · LREC 2012 5 · Takenobu Tokunaga, Ryu Iida, Asuka Terai, Naoko Kuriyama

This paper describes a collection of multimodal corpora of referring expressions, the REX corpora. The corpora have two notable features, namely (1) they include time-aligned extra-linguistic information such as participant actions and eye-gaze on top of linguistic information, (2) dialogues were collected with various configurations in terms of the puzzle type, hinting and language. After describing how the corpora were constructed and sketching out each, we present an analysis of various statistics for the corpora with respect to the various configurations mentioned above. The analysis showed that the corpora have different characteristics in the number of utterances and referring expressions in a dialogue, the task completion time and the attributes used in the referring expressions. In this respect, we succeeded in constructing a collection of corpora that included a variety of characteristics by changing the configurations for each set of dialogues, as originally planned. The corpora are now under preparation for publication, to be used for research on human reference behaviour.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

G-TUNA: a corpus of referring expressions in German, including duration information

2017-09-01 · WS 2017 9 · David Howcroft, Jorrig Vogels, Vera Demberg

Corpora of referring expressions elicited from human participants in a controlled environment are an important resource for research on automatic referring expression generation. We here present G-TUNA, a new corpus of r…

Referring ExpressionReferring expression generationText Generation

Resolving Referring Expressions in Images With Labeled Elements

2018-10-24 · Nevan Wichers, Dilek Hakkani-Tur, Jindong Chen

Images may have elements containing text and a bounding box associated with them, for example, text identified via optical character recognition on a computer screen image, or a natural image with labeled objects. We pre…

Optical Character RecognitionOptical Character Recognition (OCR)Referring Expression

Cross-Modal Relationship Inference for Grounding Referring Expressions

2019-06-01 · CVPR 2019 6 · Sibei Yang, Guanbin Li, Yizhou Yu

Grounding referring expressions is a fundamental yet challenging task facilitating human-machine communication in the physical world. It locates the target object in an image on the basis of the comprehension of the rela…

Modeling Context in Referring Expressions

2016-07-31 · Licheng Yu, Patrick Poirson, Shan Yang, Alexander C. Berg 외

Humans refer to objects in their environments all the time, especially in dialogue with other people. We explore generating and comprehending natural language referring expressions for objects in images. In particular, w…

Referring ExpressionReferring expression generationText Generation

Kosmos-2: Grounding Multimodal Large Language Models to the World

2023-06-26 · Zhiliang Peng, Wenhui Wang, Li Dong, Yaru Hao 외

We introduce Kosmos-2, a Multimodal Large Language Model (MLLM), enabling new capabilities of perceiving object descriptions (e.g., bounding boxes) and grounding text to the visual world. Specifically, we represent refer…

Image CaptioningIn-Context LearningLanguage ModelingLanguage Modelling+9