paper-with-me

Papers

Exploring the Functional and Geometric Bias of Spatial Relations Using Neural Language Models

2018-06-01 · WS 2018 6 · Simon Dobnik, Mehdi Ghanimifard, John Kelleher

The challenge for computational models of spatial descriptions for situated dialogue systems is the integration of information from different modalities. The semantics of spatial descriptions are grounded in at least two sources of information: (i) a geometric representation of space and (ii) the functional interaction of related objects that. We train several neural language models on descriptions of scenes from a dataset of image captions and examine whether the functional or geometric bias of spatial descriptions reported in the literature is reflected in the estimated perplexity of these models. The results of these experiments have implications for the creation of models of spatial lexical semantics for human-robot dialogue systems. Furthermore, they also provide an insight into the kinds of the semantic knowledge captured by neural language models trained on spatial descriptions, which has implications for image captioning systems.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Image Captioning

Similar Papers 제목 키워드 기반

What a neural language model tells us about spatial relations

2019-06-01 · WS 2019 6 · Mehdi Ghanimifard, Simon Dobnik

Understanding and generating spatial descriptions requires knowledge about what objects are related, their functional interactions, and where the objects are geometrically located. Different spatial relations have differ…

Image DescriptionLanguage ModelingLanguage Modelling

Clique topology reveals intrinsic geometric structure in neural correlations

2015-02-22

Detecting meaningful structure in neural activity and connectivity data is challenging in the presence of hidden nonlinearities, where traditional eigenvalue-based methods may be misleading. We introduce a novel approach…

Hippocampus

Exploring Geometric Deep Learning For Precipitation Nowcasting

2023-09-11 · Shan Zhao, Sudipan Saha, Zhitong Xiong, Niklas Boers 외

Precipitation nowcasting (up to a few hours) remains a challenge due to the highly complex local interactions that need to be captured accurately. Convolutional Neural Networks rely on convolutional kernels convolving wi…

Deep Learning

FunFact: Building Probabilistic Functional 3D Scene Graphs via Factor-Graph Reasoning

2026-04-04 · Zhengyu Fu, René Zurbrügg, Kaixian Qu, Marc Pollefeys 외 arxiv

Recent work in 3D scene understanding is moving beyond purely spatial analysis toward functional scene understanding. However, existing methods often consider functional relationships between object pairs in isolation, f…

Scene Understanding

G^3-LQ: Marrying Hyperbolic Alignment with Explicit Semantic-Geometric Modeling for 3D Visual Grounding

2024-01-01 · CVPR 2024 1 · YuAn Wang, YaLi Li, Shengjin Wang

Grounding referred objects in 3D scenes is a burgeoning vision-language task pivotal for propelling Embodied AI as it endeavors to connect the 3D physical world with free-form descriptions. Compared to the 2D counter…

3D visual groundingVisual Grounding