paper-with-me

홈 › Papers

Text to 3D Scene Generation with Rich Lexical Grounding

2015-05-23 · IJCNLP 2015 7 · Angel Chang, Will Monroe, Manolis Savva, Christopher Potts, Christopher D. Manning

The ability to map descriptions of scenes to 3D geometric representations has many applications in areas such as art, education, and robotics. However, prior work on the text to 3D scene generation task has used manually specified object categories and language that identifies them. We introduce a dataset of 3D scenes annotated with natural language descriptions and learn from this data how to ground textual descriptions to physical objects. Our method successfully grounds a variety of lexical terms to concrete referents, and we show quantitatively that our method improves 3D scene generation over previous work using purely rule-based methods. We evaluate the fidelity and plausibility of 3D scenes generated with our grounding approach through human judgments. To ease evaluation on this task, we also introduce an automated metric that strongly correlates with human judgments.

📄 PDF Abstract BibTeX arXiv:1505.06289

Code (0)

등록된 구현이 없습니다.

Tasks

Scene GenerationText to 3D

Similar Papers 제목 키워드 기반

SDesc3D: Towards Layout-Aware 3D Indoor Scene Generation from Short Descriptions

2026-04-02 · Jie Feng, Jiawei Shen, Junjia Huang, Junpeng Zhang 외 arxiv

3D indoor scene generation conditioned on short textual descriptions provides a promising avenue for interactive 3D environment construction without the need for labor-intensive layout specification. Despite recent progr…

Scene Generation

Structured Interfaces for Automated Reasoning with 3D Scene Graphs

2025-10-18 · Aaron Ray, Jacob Arkin, Harel Biggie, Chuchu Fan 외 arxiv

In order to provide a robot with the ability to understand and react to a user's natural language inputs, the natural language must be connected to the robot's underlying representations of the world. Recently, large lan…

Instruction FollowingCode Generation

Are Current Decoding Strategies Capable of Facing the Challenges of Visual Dialogue?

2022-10-24 · Amit Kumar Chaudhary, Alex J. Lucassen, Ioanna Tsani, Alberto Testoni

Decoding strategies play a crucial role in natural language generation systems. They are usually designed and evaluated in open-ended text-only tasks, and it is not clear how different strategies handle the numerous chal…

InformativenessText GenerationVisual Grounding

Enhancing Legal LLMs through Metadata-Enriched RAG Pipelines and Direct Preference Optimization

2026-02-25 · Suyash Maniyar, Deepali Singh, Rohith Reddy arxiv

Large Language Models (LLMs) perform well in short contexts but degrade on long legal documents, often producing hallucinations such as incorrect clauses or precedents. In the legal domain, where precision is critical, s…

ZING-3D: Zero-shot Incremental 3D Scene Graphs via Vision-Language Models

2025-10-24 · Pranav Saxena, Jimmy Chiun arxiv

Understanding and reasoning about complex 3D environments requires structured scene representations that capture not only objects but also their semantic and spatial relationships. While recent works on 3D scene graph ge…

Scene Graph Generation