paper-with-me

홈 › Papers

Learning Manner of Execution from Partial Corrections

2023-02-07 · Mattias Appelgren, Alex Lascarides

Some actions must be executed in different ways depending on the context. For example, wiping away marker requires vigorous force while wiping away almonds requires more gentle force. In this paper we provide a model where an agent learns which manner of action execution to use in which context, drawing on evidence from trial and error and verbal corrections when it makes a mistake (e.g., `no, gently''). The learner starts out with a domain model that lacks the concepts denoted by the words in the teacher's feedback; both the words describing the context (e.g., marker) and the adverbs like `gently''. We show that through the the semantics of coherence, our agent can perform the symbol grounding that's necessary for exploiting the teacher's feedback so as to solve its domain-level planning problem: to perform its actions in the current context in the right way.

📄 PDF Abstract BibTeX arXiv:2302.03338

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Centralized Training with Hybrid Execution in Multi-Agent Reinforcement Learning

2022-10-12 · Pedro P. Santos, Diogo S. Carvalho, Miguel Vasco, Alberto Sardinha 외

We introduce hybrid execution in multi-agent reinforcement learning (MARL), a new paradigm in which agents aim to successfully complete cooperative tasks with arbitrary communication levels at execution time by taking ad…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

CycleVLA: Proactive Self-Correcting Vision-Language-Action Models via Subtask Backtracking and Minimum Bayes Risk Decoding

2026-01-05 · Chenyang Ma, Guangyu Yang, Kai Lu, Shitong Xu 외 arxiv

Current work on robot failure detection and correction typically operates in a post hoc manner, analyzing errors and applying corrections only after failures occur. This work introduces CycleVLA, a system that equips Vis…

Conveyor: Efficient Tool-aware LLM Serving with Tool Partial Execution

2024-05-29 · Yechen Xu, Xinhao Kong, Tingjun Chen, Danyang Zhuo

The complexity of large language model (LLM) serving workloads has substantially increased due to the integration with external tool invocations, such as ChatGPT plugins. In this paper, we identify a new opportunity for …

Language ModelingLanguage ModellingLarge Language Model

An end-to-end deep learning pipeline to derive blood input with partial volume corrections for automated parametric brain PET mapping

2024-02-05 · Rugved Chavan, Gabriel Hyman, Zoraiz Qureshi, Nivetha Jayakumar 외

Dynamic 2-[18F] fluoro-2-deoxy-D-glucose positron emission tomography (dFDG-PET) for human brain imaging has considerable clinical potential, yet its utilization remains limited. A key challenge in the quantitative analy…

Compliant Residual DAgger: Improving Real-World Contact-Rich Manipulation with Human Corrections

2025-06-20 · Xiaomeng Xu, Yifan Hou, Zeyi Liu, Shuran Song

We address key challenges in Dataset Aggregation (DAgger) for real-world contact-rich manipulation: how to collect informative human correction data and how to effectively update policies with this new data. We introduce…

Contact-rich Manipulation