Recognising Affordances in Predicted Futures to Plan with Consideration of Non-canonical Affordance Effects
We propose a novel system for action sequence planning based on a combination of affordance recognition and a neural forward model predicting the effects of affordance execution. By performing affordance recognition on predicted futures, we avoid reliance on explicit affordance effect definitions for multi-step planning. Because the system learns affordance effects from experience data, the system can foresee not just the canonical effects of an affordance, but also situation-specific side-effects. This allows the system to avoid planning failures due to such non-canonical effects, and makes it possible to exploit non-canonical effects for realising a given goal. We evaluate the system in simulation, on a set of test tasks that require consideration of canonical and non-canonical affordance effects.
Code (0)
등록된 구현이 없습니다.
Tasks
Affordance RecognitionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
ActSWM: Action-Sensitive World Models for Long-Horizon Planning in Open-World Games
Latent world models support efficient model-predictive control by optimizing future control sequences in latent space and replanning in a receding-horizon manner. However, existing latent predictors often lack stable lon…
GrASP: Gradient-Based Affordance Selection for Planning
Planning with a learned model is arguably a key component of intelligence. There are several challenges in realizing such a component in large-scale reinforcement learning (RL) problems. One such challenge is dealing eff…
Reinforcement Learning (RL)Is the Future Compatible? Diagnosing Dynamic Consistency in World Action Models
World Action Models (WAMs) enable decision-making through imagined rollouts by predicting future observations and actions. However, the reliability of these imagined futures remains under-examined: is a generated future …
Hierarchical Graph Learning for Calendar Spread Strategies in Commodity Futures Markets
Commodity futures can be represented hierarchically, with underlying assets at the upper level and individual futures contracts at the lower level. Entities at each level can be connected by edges reflecting inherent cor…
Graph LearningContrastive Explanations for Reinforcement Learning via Embedded Self Predictions
We investigate a deep reinforcement learning (RL) architecture that supports explaining why a learned agent prefers one action over another. The key idea is to learn action-values that are directly represented via human-…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)