paper-with-me

Scene Generation

6개 벤치마크 · 논문 524편 · 이 태스크의 논문 보기 →

Benchmarks

GoogleEarth

결과 10개

AVD

결과 6개

Replica

결과 6개

VizDoom

결과 6개

OSM

결과 4개

KITTI

결과 2개

Most implemented

Funnel Activation for Visual Recognition

2020-07-23 · 구현 7개

GPD-1: Generative Pre-training for Driving

2024-12-11 · 구현 2개

Papers

StreetDiff: Multi-view Street Scenes Generation via Cross-view Consistent Multi-view Stable Diffusion with Structure Prompts

2026-09-09 · Qi Zhang, Yanyifan Wang, Weiyuan Zhang, Hui Huang arxiv

Multi-view diffusion models have shown strong performance in scenes with strong geometric priors and sparse semantics, such as indoor rooms or simple outdoor environments (e.g., fields, courtyards). However, they often f…

Scene Generation

SceneMosaic: Efficient and Diverse Simulation-Ready Scene Generation via Hybrid Agentic Layout Evolution

2026-09-04 · Xingjian Ran, Xiaoye Mo, Sihao Liu, Jianyu Zhang 외 hf

Diverse and simulation-ready indoor scenes are essential for interactive entertainment and embodied AI, yet their scalable generation remains challenging. Recent agentic text-to-3D scene pipelines that rely on vision-lan…

Scene Generation

ScenePilot: Grow-and-Repair Policy for Text-Driven 3D Indoor Scene Generation

2026-08-31 · Jiawei Zhang, Hongsong Wang, Pan Zhou arxiv

Text-driven 3D indoor scene generation has advanced from dataset-bound layout modeling to open-vocabulary synthesis with large language and vision-language models. Yet existing methods remain limited: one-pass generators…

Scene Generation

SpatialCrafter: Single Image World Modeling with Generative 3D Proxies

2026-08-27 · Chuan Fang, Lingteng Qiu, Yixun Liang, Rui Chen 외 arxiv

Explorable image-to-scene generation is essential for applications in gaming, robotics, and virtual reality. Existing methods based on video diffusion model (VDM) commonly rely on incomplete conditioning signals such as …

Scene GenerationPoint Clouds

Towards Surgical World-Action Modeling: A Preliminary Joint Visual-Trajectory Forecasting for Surgical Motion Planning

2026-08-20 · Weiliang Huang, Huanrong Liu, Bob Zhang, Qi Dou 외 arxiv

Reliable surgical planning requires models to anticipate not only how instruments will move, but also how the operative visual state will evolve together with such motion. Existing approaches typically treat future scene…

Trajectory ForecastingTrajectory PredictionMotion ForecastingScene Generation

Beyond Placement and Articulation: Usage-Driven Code Scenes for Embodied Interaction

2026-08-19 · Zijian Xiao, Zipeng Ye, Jinkun Hao, Xiong Yang 외 arxiv

Indoor scene synthesis provides essential environments for embodied AI, robotic manipulation, and simulation-based policy learning. Recent code-based scene generation methods produce editable and extensible environments,…

Indoor Scene SynthesisScene Generation

전체 524편 보기 →