paper-with-me

Papers

Next Best Sense: Guiding Vision and Touch with FisherRF for 3D Gaussian Splatting

2024-10-07 · Matthew Strong, Boshu Lei, Aiden Swann, Wen Jiang, Kostas Daniilidis, Monroe Kennedy III

We propose a framework for active next best view and touch selection for robotic manipulators using 3D Gaussian Splatting (3DGS). 3DGS is emerging as a useful explicit 3D scene representation for robotics, as it has the ability to represent scenes in a both photorealistic and geometrically accurate manner. However, in real-world, online robotic scenes where the number of views is limited given efficiency requirements, random view selection for 3DGS becomes impractical as views are often overlapping and redundant. We address this issue by proposing an end-to-end online training and active view selection pipeline, which enhances the performance of 3DGS in few-view robotics settings. We first elevate the performance of few-shot 3DGS with a novel semantic depth alignment method using Segment Anything Model 2 (SAM2) that we supplement with Pearson depth and surface normal loss to improve color and depth reconstruction of real-world scenes. We then extend FisherRF, a next-best-view selection method for 3DGS, to select views and touch poses based on depth uncertainty. We perform online view selection on a real robot system during live 3DGS training. We motivate our improvements to few-shot GS scenes, and extend depth-based FisherRF to them, where we demonstrate both qualitative and quantitative improvements on challenging robot scenes. For more information, please see our project page at https://arm.stanford.edu/next-best-sense.

📄 PDF Abstract BibTeX arXiv:2410.04680

Code (1)

armlabstanford/NextBestSense 공식 구현 pytorch

Tasks

3DGS

Similar Papers 제목 키워드 기반

The Next Evolution of Artificial Sense of Touch

2023-10-25 · Sonja Groß, Amartya Ganguly, Hendrik Dietz, Sami Haddadin

We propose the next evolution of the artificial sense of touch, including an in-depth examination of the latest advancements in tactile sensing technology and the challenges that remain. We delve into the forefront of DN…

DexTouch: Learning to Seek and Manipulate Objects with Tactile Dexterity

2024-01-23 · Kang-Won Lee, Yuzhe Qin, Xiaolong Wang, Soo-Chul Lim

The sense of touch is an essential ability for skillfully performing a variety of tasks, providing the capacity to search and manipulate objects without relying on visual information. In this paper, we introduce a multi-…

Multimodal perception for dexterous manipulation

2021-12-28 · Guanqun Cao, Shan Luo

Humans usually perceive the world in a multimodal way that vision, touch, sound are utilised to understand surroundings from various dimensions. These senses are combined together to achieve a synergistic effect where th…

3D ReconstructionFrictionTranslation

The Power of the Senses: Generalizable Manipulation from Vision and Touch through Masked Multimodal Learning

2023-11-02 · Carmelo Sferrazza, Younggyo Seo, Hao liu, Youngwoon Lee 외

Humans rely on the synergy of their senses for most essential tasks. For tasks requiring object manipulation, we seamlessly and effectively exploit the complementarity of our senses of vision and touch. This paper draws …

FusionSense: Bridging Common Sense, Vision, and Touch for Robust Sparse-View Reconstruction

2024-10-10 · Irving Fang, Kairui Shi, Xujin He, Siqi Tan 외

Humans effortlessly integrate common-sense knowledge with sensory input from vision and touch to understand their surroundings. Emulating this capability, we introduce FusionSense, a novel 3D reconstruction framework tha…

3D ReconstructionCommon Sense ReasoningObject