paper-with-me

홈 › Papers

Compact Feed-Forward 3D Gaussians via Saliency-Guided Primitive Merging

2026-08-11 · Tim-Felix Fassch, Jochen Kall, Cyrill Stachniss arxiv

3D scene reconstruction, modeling, and rendering are highly relevant for numerous tasks, and 3D Gaussian splatting has become a standard choice in this context. Its feed-forward variants provide fast reconstruction from sparse input views but often produce per-pixel primitives, leading to highly redundant and thus inefficient representations. We present a structure-aware merging pipeline that takes per-pixel primitives from any feed-forward method and consolidates them into a compact, content-adaptive Gaussian set while largely retaining visual quality at just $\frac{1}{20}^\text{th}$ of the Gaussians of a per-pixel method. We group spatially coherent Gaussians of similar appearance into variable-size clusters via adaptive superpixel segmentation guided by a saliency map, which allocates fine segments to textured regions and coarse segments to homogeneous areas. We compress each cluster into a compact latent representation through a learned encoder, then match and consolidate representations across views based on geometric overlap and feature similarity via a learned merger. A level-of-detail decoder then produces the final Gaussians at a controllable resolution, enabling a flexible quality-efficiency trade-off at inference. As a post-processing module, the pipeline is backbone-agnostic, leveraging the strengths of existing feed-forward methods. This leads to better and more robust quality than achieved by previous approaches that target a reduction in primitive count, while providing a highly compact representation, that can be rendered efficiently.

📄 PDF Abstract BibTeX arXiv:2608.10712

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

F4Splat: Feed-Forward Predictive Densification for Feed-Forward 3D Gaussian Splatting

2026-03-22 · Injae Kim, Chaehyeon Kim, Minseong Bae, Minseok Joo 외 arxiv

Feed-forward 3D Gaussian Splatting methods enable single-pass reconstruction and real-time rendering. However, they typically adopt rigid pixel-to-Gaussian or voxel-to-Gaussian pipelines that uniformly allocate Gaussians…

SaLon3R: Structure-aware Long-term Generalizable 3D Reconstruction from Unposed Images

2025-10-16 · Jiaxin Guo, Tongfan Guan, Wenzhen Dong, Wenzhao Zheng 외 arxiv

Recent advances in 3D Gaussian Splatting (3DGS) have enabled generalizable, on-the-fly reconstruction of sequential input views. However, existing methods often predict per-pixel Gaussians and combine Gaussians from all …

Novel View Synthesis3D ReconstructionDepth Estimation

Learning Global Motion with Compact Gaussians for Feed-Forward 4D Reconstruction

2026-05-29 · Mungyeom Kim, Minkyeong Jeon, Honggyu An, Jaewoo Jung 외 arxiv

Dynamic scene reconstruction from monocular video remains a fundamental challenge in computer vision. Existing feed-forward methods predict 3D Gaussians pixel-wise for each frame, suffering from duplicated Gaussians and …

Scene UnderstandingPoint Tracking

Deep Level Sets for Salient Object Detection

2017-07-01 · CVPR 2017 7 · Ping Hu, Bing Shuai, Jun Liu, Gang Wang

Deep learning has been applied to saliency detection in recent years. The superior performance has proved that deep networks can model the semantic properties of salient objects. Yet it is difficult for a deep network…

Objectobject-detectionObject DetectionRGB Salient Object Detection+2

SparseSplat: Towards Applicable Feed-Forward 3D Gaussian Splatting with Pixel-Unaligned Prediction

2026-04-03 · Zicheng Zhang, Xiangting Meng, Ke Wu, Wenchao Ding arxiv

Recent progress in feed-forward 3D Gaussian Splatting (3DGS) has notably improved rendering quality. However, the spatially uniform and highly redundant 3DGS map generated by previous feed-forward 3DGS methods limits the…