paper-with-me

Papers

DiffuDepGrasp: Diffusion-based Depth Noise Modeling Empowers Sim2Real Robotic Grasping

2025-11-17 · Yingting Zhou, Wenbo Cui, Weiheng Liu, Guixing Chen, Haoran Li, Dongbin Zhao arxiv

Transferring the depth-based end-to-end policy trained in simulation to physical robots can yield an efficient and robust grasping policy, yet sensor artifacts in real depth maps like voids and noise establish a significant sim2real gap that critically impedes policy transfer. Training-time strategies like procedural noise injection or learned mappings suffer from data inefficiency due to unrealistic noise simulation, which is often ineffective for grasping tasks that require fine manipulation or dependency on paired datasets heavily. Furthermore, leveraging foundation models to reduce the sim2real gap via intermediate representations fails to mitigate the domain shift fully and adds computational overhead during deployment. This work confronts dual challenges of data inefficiency and deployment complexity. We propose DiffuDepGrasp, a deploy-efficient sim2real framework enabling zero-shot transfer through simulation-exclusive policy training. Its core innovation, the Diffusion Depth Generator, synthesizes geometrically pristine simulation depth with learned sensor-realistic noise via two synergistic modules. The first Diffusion Depth Module leverages temporal geometric priors to enable sample-efficient training of a conditional diffusion model that captures complex sensor noise distributions, while the second Noise Grafting Module preserves metric accuracy during perceptual artifact injection. With only raw depth inputs during deployment, DiffuDepGrasp eliminates computational overhead and achieves a 95.7% average success rate on 12-object grasping with zero-shot transfer and strong generalization to unseen objects.Project website: https://diffudepgrasp.github.io/.

📄 PDF Abstract BibTeX arXiv:2511.12912

Code (0)

등록된 구현이 없습니다.

Tasks

Robotic Grasping

Similar Papers 제목 키워드 기반

JointDiT: Enhancing RGB-Depth Joint Modeling with Diffusion Transformers

2025-05-01 · Kwon Byung-Ki, Qi Dai, Lee Hyoseok, Chong Luo 외

We present JointDiT, a diffusion transformer that models the joint distribution of RGB and depth. By leveraging the architectural benefit and outstanding image prior of the state-of-the-art diffusion transformer, JointDi…

Depth EstimationImage GenerationScheduling

CaDM: Codec-aware Diffusion Modeling for Neural-enhanced Video Streaming

2022-11-15 · Qihua Zhou, Ruibin Li, Song Guo, Peiran Dong 외

Recent years have witnessed the dramatic growth of Internet video traffic, where the video bitstreams are often compressed and delivered in low quality to fit the streamer's uplink bandwidth. To alleviate the quality deg…

DecoderDenoisingSuper-Resolution

RealD$^2$iff: Bridging Real-World Gap in Robot Manipulation via Depth Diffusion

2025-11-27 · Xiujian Liang, Jiacheng Liu, Mingyang Sun, Qichen He 외 arxiv

Robot manipulation in the real world is fundamentally constrained by the visual sim2real gap, where depth observations collected in simulation fail to reflect the complex noise patterns inherent to real sensors. In this …

Robot Manipulation

Towards Robust Time-of-Flight Depth Denoising with Confidence-Aware Diffusion Model

2025-03-25 · Changyong He, Jin Zeng, Jiawei Zhang, Jiajie Guo

Time-of-Flight (ToF) sensors efficiently capture scene depth, but the nonlinear depth construction procedure often results in extremely large noise variance or even invalid areas. Recent methods based on deep neural netw…

Denoising

Edit as You See: Image-guided Video Editing via Masked Motion Modeling

2025-01-08 · Zhi-Lin Huang, Yixuan Liu, Chujun Qin, Zhongdao Wang 외

Recent advancements in diffusion models have significantly facilitated text-guided video editing. However, there is a relative scarcity of research on image-guided video editing, a method that empowers users to edit vide…

Optical Flow EstimationSelf-Supervised LearningVideo Editing