paper-with-me

홈 › Papers

WorldScore: A Unified Evaluation Benchmark for World Generation

2025-04-01 · Haoyi Duan, Hong-Xing Yu, Sirui Chen, Li Fei-Fei, Jiajun Wu

We introduce the WorldScore benchmark, the first unified benchmark for world generation. We decompose world generation into a sequence of next-scene generation tasks with explicit camera trajectory-based layout specifications, enabling unified evaluation of diverse approaches from 3D and 4D scene generation to video generation models. The WorldScore benchmark encompasses a curated dataset of 3,000 test examples that span diverse worlds: static and dynamic, indoor and outdoor, photorealistic and stylized. The WorldScore metrics evaluate generated worlds through three key aspects: controllability, quality, and dynamics. Through extensive evaluation of 19 representative models, including both open-source and closed-source ones, we reveal key insights and challenges for each category of models. Our dataset, evaluation code, and leaderboard can be found at https://haoyi-duan.github.io/WorldScore/

📄 PDF Abstract BibTeX arXiv:2504.00983

Code (0)

등록된 구현이 없습니다.

Tasks

Scene GenerationVideo Generation

Similar Papers 제목 키워드 기반

Reference-Free Assessment of Physical Consistency in World Model-based Video Generation

2026-06-21 · Yun Oh, Sukmin Yun arxiv

We introduce reference-free measures for evaluating the physical consistency of generated videos, combining relative and absolute approaches to assess fidelity. Although tools like WorldGym or WorldEval enable robotic si…

Video Generation

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models

2026-08-05 · Haiyang Zhou, Wangbo Yu, Chaoran Feng, Xunyu Zhou 외 hf

The abundance of casually captured monocular videos and images on social media provides a valuable source for immersive content creation, where generating novel views from such sparse observations can greatly enhance use…

Novel View Synthesis

Latent Spatial Memory for Video World Models

2026-06-08 · Weijie Wang, Haoyu Zhao, Yifan Yang, Feng Chen 외 arxiv

Video world models that maintain 3D spatial consistency across generated frames typically rely on explicit point cloud memory constructed in RGB space. This design is both computationally expensive, requiring repeated re…

Video Generation

NeoWorld: Neural Simulation of Explorable Virtual Worlds via Progressive 3D Unfolding

2025-09-29 · Yanpeng Zhao, Shanyan Guan, Yunbo Wang, Yanhao Ge 외 arxiv

We introduce NeoWorld, a deep learning framework for generating interactive 3D virtual worlds from a single input image. Inspired by the on-demand worldbuilding concept in the science fiction novel Simulacron-3 (1964), o…

Representation Learning

4DWorldBench: A Comprehensive Evaluation Framework for 3D/4D World Generation Models

2025-11-25 · Yiting Lu, Wei Luo, Peiyan Tu, Haoran Li 외 arxiv

World Generation Models are emerging as a cornerstone of next-generation multimodal intelligence systems. Unlike traditional 2D visual generation, World Models aim to construct realistic, dynamic, and physically consiste…

Autonomous Driving