paper-with-me

홈 › Papers

Struct-GStream: Towards Efficient Free-Viewpoint Video Streaming at Low-Bitrates with Structured 3D Gaussians

2026-08-02 · Han Jiao, Jiakai Sun, Lei Zhao, Wei Xing, Huaizhong Lin, Zhanjie Zhang, Ao Ma arxiv

Constructing photorealistic Free-Viewpoint Videos (FVVs) of dynamic scenes from a set of posed 2D images has been an intriguing yet challenging task in computer vision. Methods based on neural rendering achieve high-fidelity image quality in FVV construction. However, most of these methods are unable to achieve real-time rendering and often require complete video sequences to train. Despite the existence of some online training methods capable of rendering FVVs in real time, they struggle to meet the requirements for storage and training time for downstream applications. To overcome this problem, we propose Struct-GStream, which can achieve efficient FVV streaming using structured 3D Gaussians (3DGs). Specifically, we introduce dynamic anchor points to generate structured 3DGs to construct basic scenes and model approximate scene movements based on the assumption of local rigidity in object motion. Besides, we introduce a global free 3DGs patching strategy involving free 3DGs' generation, pruning, and optimization to patch and model deficient areas and emerging objects. Our method achieves fast training at low bitrates while maintaining high rendering quality. Extensive experiments demonstrate that Struct-GStream significantly outperforms existing online training methods for FVV construction in terms of training time, storage, and rendering quality while maintaining competitive rendering speed.

📄 PDF Abstract BibTeX arXiv:2608.01053

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

3DGStream: On-the-Fly Training of 3D Gaussians for Efficient Streaming of Photo-Realistic Free-Viewpoint Videos

2024-03-03 · CVPR 2024 1 · Jiakai Sun, Han Jiao, Guangyuan Li, Zhanjie Zhang 외

Constructing photo-realistic Free-Viewpoint Videos (FVVs) of dynamic scenes from multi-view videos remains a challenging endeavor. Despite the remarkable advancements achieved by current neural rendering techniques, thes…

3DGSNeural Rendering

DLGStream: Dynamic Language-embedded Guassian Splatting for Open-vocabulary Enabled Free-viewpoint Video Streaming

2026-06-27 · Zhihui Ke, Yuyang Liu, Xiaobo Zhou, Tie Qiu arxiv

3D Gaussian Splatting~(3DGS) has emerged as a promising paradigm for reconstructing streamable free-viewpoint video~(FVV) from multi-view videos. However, 3DGS-based FVVs typically lack user interaction and editing capab…

Motion Matters: Compact Gaussian Streaming for Free-Viewpoint Video Reconstruction

2025-05-22 · Jiacong Chen, Qingyu Mao, Youneng Bao, Xiandong Meng 외

3D Gaussian Splatting (3DGS) has emerged as a high-fidelity and efficient paradigm for online free-viewpoint video (FVV) reconstruction, offering viewers rapid responsiveness and immersive experiences. However, existing …

3DGSVideo Reconstruction

CogStream: Context-guided Streaming Video Question Answering

2025-06-12 · Zicheng Zhao, Kangyu Wang, Shijie Li, Rui Qian 외

Despite advancements in Video Large Language Models (Vid-LLMs) improving multimodal understanding, challenges persist in streaming video reasoning due to its reliance on contextual information. Existing paradigms feed al…

Question AnsweringVideo Question Answering

Streaming Drag-Oriented Interactive Video Manipulation: Drag Anything, Anytime!

2025-10-03 · Junbao Zhou, Yuan Zhou, Kesen Zhao, Qingshan Xu 외 arxiv

Achieving streaming, fine-grained control over the outputs of autoregressive video diffusion models remains challenging, making it difficult to ensure that they consistently align with user expectations. To bridge this g…