paper-with-me

Papers

Co-policy: Responsive Human-Robot Co-Creation for Musical Performances

2026-06-18 · Xuetao Li, Wenke Huang, Mang Ye, Zijian Liu, Jinhua Xie, Jifeng Xuan, Miao Li arxiv

Art has long stood as a pivotal expression of human creativity. Embodied artificial intelligence offers a route for generative models to participate in that creativity through physical action rather than disembodied digital content. In robotic music co-creation, it is challenging to connect semantic musical understanding with real-time and physically executable performance. We present Co-policy, a framework for human-robot musical co-creation that separates semantic intent grounding, constrained musical variation, and visuomotor execution. To ground musical semantics, Co-policy uses pre-inference semantic anchors and a fine-tuned Qwen-vl planner (F-Qwen) to transform speech, live musical seeds, and visual observations into structured co-creation plans. To support low-latency execution, Co-policy introduces a Gaussian-Mixture Visuomotor Policy (GMP), implemented as a conditional mixture-density policy that maps target notes and visual context to multimodal robot actions in a single forward pass. Unlike robotic playback systems that merely reproduce user-specified notes, Co-policy generates complementary musical responses under both musical and physical constraints. Real-robot chime experiments, ablations, and expert evaluation show improved intent alignment, execution accuracy, and response frequency over diffusion-policy and ablated baselines, supporting physically grounded action generation as a key requirement for embodied human-AI co-creation.

📄 PDF Abstract BibTeX arXiv:2606.19914

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Ghost in the Keys: A Disklavier Demo for Human-AI Musical Co-Creativity

2025-11-03 · Louis Bradshaw, Alexander Spangher, Stella Biderman, Simon Colton arxiv

While generative models for music composition are increasingly capable, their adoption by musicians is hindered by text-prompting, an asynchronous workflow disconnected from the embodied, responsive nature of instrumenta…

Human-Human & Human-Robot Interaction Transformer (H2INT) for Robot Navigation in Dense and Uncertain Crowds

2026-09-04 · Ao Shen, Kaixi Chen, Shiwei Liu, Fang Deng 외 arxiv

Safe robot navigation in dense crowds requires reasoning about pedestrian motion and how it may change in response to a robot. However, many learning-based approaches generate pedestrian motion independently of the robot…

Reinforcement LearningRobot Navigation

Trust in Autonomous Human--Robot Collaboration: Effects of Responsive Interaction Policies

2026-02-25 · Shauna Heron, Meng Cheng Lau arxiv

Trust plays a central role in human--robot collaboration, yet its formation is rarely examined under the constraints of fully autonomous interaction. This pilot study investigated how interaction policy influences trust …

Robot Drummer: Learning Rhythmic Skills for Humanoid Drumming

2025-07-15 · Asad Ali Shahid, Francesco Braghin, Loris Roveda arxiv

Humanoid robots have seen remarkable advances in dexterity, balance, and locomotion, yet their role in expressive domains such as music performance remains largely unexplored. Musical tasks, like drumming, present unique…

Reinforcement Learning

Interactive Multi-Robot Flocking with Gesture Responsiveness and Musical Accompaniment

2024-03-30 · Catie Cuan, Kyle Jeffrey, Kim Kleiven, Adrian Li 외

For decades, robotics researchers have pursued various tasks for multi-robot systems, from cooperative manipulation to search and rescue. These tasks are multi-robot extensions of classical robotic tasks and often optimi…