paper-with-me

홈 › Papers

Program Generation from Diverse Video Demonstrations

2023-02-01 · Anthony Manchin, Jamie Sherrah, Qi Wu, Anton Van Den Hengel

The ability to use inductive reasoning to extract general rules from multiple observations is a vital indicator of intelligence. As humans, we use this ability to not only interpret the world around us, but also to predict the outcomes of the various interactions we experience. Generalising over multiple observations is a task that has historically presented difficulties for machines to grasp, especially when requiring computer vision. In this paper, we propose a model that can extract general rules from video demonstrations by simultaneously performing summarisation and translation. Our approach differs from prior works by framing the problem as a multi-sequence-to-sequence task, wherein summarisation is learnt by the model. This allows our model to utilise edge cases that would otherwise be suppressed or discarded by traditional summarisation techniques. Additionally, we show that our approach can handle noisy specifications without the need for additional filtering methods. We evaluate our model by synthesising programs from video demonstrations in the Vizdoom environment achieving state-of-the-art results with a relative increase of 11.75% program accuracy on prior works

📄 PDF Abstract BibTeX arXiv:2302.00178

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Neural Program Synthesis from Diverse Demonstration Videos

2018-07-01 · ICML 2018 7 · Shao-Hua Sun, Hyeonwoo Noh, Sriram Somasundaram, Joseph Lim

Interpreting decision making logic in demonstration videos is key to collaborating with and mimicking humans. To empower machines with this ability, we propose a neural program synthesizer that is able to explicitly…

Decision MakingProgram Synthesis

CRAFT: Video Diffusion for Bimanual Robot Data Generation

2026-04-04 · Jason Chen, I-Chun Arthur Liu, Gaurav Sukhatme, Daniel Seita arxiv

Bimanual robot learning from demonstrations is fundamentally limited by the cost and narrow visual diversity of real-world data, which constrains policy robustness across viewpoints, object configurations, and embodiment…

Video Generation

Joint Flow Trajectory Optimization For Feasible Robot Motion Generation from Video Demonstrations

2025-09-25 · Xiaoxiang Dong, Matthew Johnson-Roberson, Weiming Zhi arxiv

Learning from human video demonstrations offers a scalable alternative to teleoperation or kinesthetic teaching, but poses challenges for robot manipulators due to embodiment differences and joint feasibility constraints…

Information-Theoretic Detection of Bimanual Interactions for Dual-Arm Robot Plan Generation

2026-01-27 · Elena Merlo, Marta Lagomarsino, Arash Ajoudani arxiv

Programming by demonstration is a strategy to simplify the robot programming process for non-experts via human demonstrations. However, its adoption for bimanual tasks is an underexplored problem due to the complexity of…

BOOST: Bootstrapping Strategy-Driven Reasoning Programs for Program-Guided Fact-Checking

2025-04-03 · Qisheng Hu, Quanyu Long, Wenya Wang

Program-guided reasoning has shown promise in complex claim fact-checking by decomposing claims into function calls and executing reasoning programs. However, prior work primarily relies on few-shot in-context learning (…

Claim VerificationDiversityFact CheckingIn-Context Learning