paper-with-me

Papers

SceneSplat++: A Large Dataset and Comprehensive Benchmark for Language Gaussian Splatting

2025-06-10 · Mengjiao Ma, Qi Ma, Yue Li, Jiahuan Cheng, Runyi Yang, Bin Ren, Nikola Popovic, Mingqiang Wei, Nicu Sebe, Luc van Gool, Theo Gevers, Martin R. Oswald, Danda Pani Paudel

3D Gaussian Splatting (3DGS) serves as a highly performant and efficient encoding of scene geometry, appearance, and semantics. Moreover, grounding language in 3D scenes has proven to be an effective strategy for 3D scene understanding. Current Language Gaussian Splatting line of work fall into three main groups: (i) per-scene optimization-based, (ii) per-scene optimization-free, and (iii) generalizable approach. However, most of them are evaluated only on rendered 2D views of a handful of scenes and viewpoints close to the training views, limiting ability and insight into holistic 3D understanding. To address this gap, we propose the first large-scale benchmark that systematically assesses these three groups of methods directly in 3D space, evaluating on 1060 scenes across three indoor datasets and one outdoor dataset. Benchmark results demonstrate a clear advantage of the generalizable paradigm, particularly in relaxing the scene-specific limitation, enabling fast feed-forward inference on novel scenes, and achieving superior segmentation performance. We further introduce GaussianWorld-49K a carefully curated 3DGS dataset comprising around 49K diverse indoor and outdoor scenes obtained from multiple sources, with which we demonstrate the generalizable approach could harness strong data priors. Our codes, benchmark, and datasets will be made public to accelerate research in generalizable 3DGS scene understanding.

📄 PDF Abstract BibTeX arXiv:2506.08710

Code (0)

등록된 구현이 없습니다.

Tasks

3DGSScene Understanding

Similar Papers 제목 키워드 기반

SceneSplat: Gaussian Splatting-based Scene Understanding with Vision-Language Pretraining

2025-03-23 · Yue Li, Qi Ma, Runyi Yang, Huapeng Li 외

Recognizing arbitrary or previously unseen categories is essential for comprehensive real-world 3D scene understanding. Currently, all existing methods rely on 2D or textual modalities during training, or together at inf…

3DGSBenchmarkingGPUScene Understanding+1

This is the way: designing and compiling LEPISZCZE, a comprehensive NLP benchmark for Polish

2022-11-23 · Łukasz Augustyniak, Kamil Tagowski, Albert Sawczyn, Denis Janiak 외

The availability of compute and data to train larger and larger language models increases the demand for robust methods of benchmarking the true progress of LM training. Recent years witnessed significant progress in sta…

Benchmarking

Advancing the Evaluation of Traditional Chinese Language Models: Towards a Comprehensive Benchmark Suite

2023-09-15 · Chan-Jan Hsu, Chang-Le Liu, Feng-Ting Liao, Po-chun Hsu 외

The evaluation of large language models is an essential task in the field of language understanding and generation. As language models continue to advance, the need for effective benchmarks to assess their performance ha…

Question Answering

The Music Maestro or The Musically Challenged, A Massive Music Evaluation Benchmark for Large Language Models

2024-06-22 · Jiajia Li, Lu Yang, Mingni Tang, Cong Chen 외

Benchmark plays a pivotal role in assessing the advancements of large language models (LLMs). While numerous benchmarks have been proposed to evaluate LLMs' capabilities, there is a notable absence of a dedicated benchma…

EMMA-500: Enhancing Massively Multilingual Adaptation of Large Language Models

2024-09-26 · Shaoxiong Ji, Zihao Li, Indraneil Paul, Jaakko Paavola 외

In this work, we introduce EMMA-500, a large-scale multilingual language model continue-trained on texts across 546 languages designed for enhanced multilingual performance, focusing on improving language coverage for lo…

Cross-Lingual TransferLanguage ModelingLanguage Modelling