paper-with-me

홈 › Papers

PerchMobi^3: A Multi-Modal Robot with Power-Reuse Quad-Fan Mechanism for Air-Ground-Wall Locomotion

2025-09-16 · Yikai Chen, Zhi Zheng, Jin Wang, Bingye He, Xiangyu Xu, Jialu Zhang, Huan Yu, Guodong Lu arxiv

Achieving seamless integration of aerial flight, ground driving, and wall climbing within a single robotic platform remains a major challenge, as existing designs often rely on additional adhesion actuators that increase complexity, reduce efficiency, and compromise reliability. To address these limitations, we present PerchMobi^3, a quad-fan, negative-pressure, air-ground-wall robot that implements a propulsion-adhesion power-reuse mechanism. By repurposing four ducted fans to simultaneously provide aerial thrust and negative-pressure adhesion, and integrating them with four actively driven wheels, PerchMobi^3 eliminates dedicated pumps while maintaining a lightweight and compact design. To the best of our knowledge, this is the first quad-fan prototype to demonstrate functional power reuse for multi-modal locomotion. A modeling and control framework enables coordinated operation across ground, wall, and aerial domains with fan-assisted transitions. The feasibility of the design is validated through a comprehensive set of experiments covering ground driving, payload-assisted wall climbing, aerial flight, and cross-mode transitions, demonstrating robust adaptability across locomotion scenarios. These results highlight the potential of PerchMobi^3 as a novel design paradigm for multi-modal robotic mobility, paving the way for future extensions toward autonomous and application-oriented deployment.

📄 PDF Abstract BibTeX arXiv:2509.12620

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

M2R2: MulitModal Robotic Representation for Temporal Action Segmentation

2025-04-25 · Daniel Sliwowski, Dongheui Lee

Temporal action segmentation (TAS) has long been a key area of research in both robotics and computer vision. In robotics, algorithms have primarily focused on leveraging proprioceptive information to determine skill bou…

Action SegmentationTemporal Action Segmentation

MeCo: Enhancing LLM-Empowered Multi-Robot Collaboration via Similar Task Memoization

2026-01-28 · Baiqing Wang, Helei Cui, Bo Zhang, Xiaolong Zheng 외 arxiv

Multi-robot systems have been widely deployed in real-world applications, providing significant improvements in efficiency and reductions in labor costs. However, most existing multi-robot collaboration methods rely on e…

Sparse ActionGen: Accelerating Diffusion Policy with Real-time Pruning

2026-01-19 · Kangye Ji, Jianbo Zhou, Yuan Meng, Ye Li 외 arxiv

Diffusion Policy has dominated action generation due to its strong capabilities for modeling multi-modal action distributions, but its multi-step denoising processes make it impractical for real-time visuomotor control. …

VLA-Cache: Towards Efficient Vision-Language-Action Model via Adaptive Token Caching in Robotic Manipulation

2025-02-04 · Siyu Xu, Yunke Wang, Chenghao Xia, Dihao Zhu 외

Vision-Language-Action (VLA) model can process instructions and visual perception to directly generate actions as output in an end-to-end fashion due to its strong multi-modal reasoning capabilities. While the performanc…

Decision MakingSequential Decision MakingvalidVision-Language-Action

rrSDS: Towards a Robot-ready Spoken Dialogue System

2020-07-01 · SIGDIAL (ACL) 2020 7 · Casey Kennington, Daniele Moro, Lucas Marchand, Jake Carns 외

Spoken interaction with a physical robot requires a dialogue system that is modular, multimodal, distributive, incremental and temporally aligned. In this demo paper, we make significant contributions towards fulfilling …