paper-with-me

홈 › Papers

Graph neural induction of value iteration

2020-09-26 · Andreea Deac, Pierre-Luc Bacon, Jian Tang

Many reinforcement learning tasks can benefit from explicit planning based on an internal model of the environment. Previously, such planning components have been incorporated through a neural network that partially aligns with the computational graph of value iteration. Such network have so far been focused on restrictive environments (e.g. grid-worlds), and modelled the planning procedure only indirectly. We relax these constraints, proposing a graph neural network (GNN) that executes the value iteration (VI) algorithm, across arbitrary environment models, with direct supervision on the intermediate steps of VI. The results indicate that GNNs are able to model value iteration accurately, recovering favourable metrics and policies across a variety of out-of-distribution tests. This suggests that GNN executors with strong supervision are a viable component within deep reinforcement learning systems.

📄 PDF Abstract BibTeX arXiv:2009.12604

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement LearningGraph Neural Networkreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

Graph Neural Network 설명 없음

Similar Papers 제목 키워드 기반

Combining Word Embeddings with Bilingual Orthography Embeddings for Bilingual Dictionary Induction

2020-12-01 · COLING 2020 8 · Silvia Severini, Viktor Hangya, Alexander Fraser, Hinrich Sch{\"u}tze

Bilingual dictionary induction (BDI) is the task of accurately translating words to the target language. It is of great importance in many low-resource scenarios where cross-lingual training data is not available. To per…

TranslationTransliterationWord Embeddings

Provably Faster Gradient Descent via Long Steps

2023-07-12 · Benjamin Grimmer

This work establishes new convergence guarantees for gradient descent in smooth convex optimization via a computer-assisted analysis technique. Our theory allows nonconstant stepsize policies with frequent long steps pot…

Automatic Induction of Bellman-Error Features for Probabilistic Planning

2014-01-16 · Jia-Hong Wu, Robert Givan

Domain-specific features are important in representing problem structure throughout machine learning and decision-theoretic planning. In planning, once state features are provided, domain-independent algorithms such as a…

Planning based on classification by induction graph

2013-11-15 · Sofia Benbelkacem, Baghdad Atmani, Mohamed Benamina

In Artificial Intelligence, planning refers to an area of research that proposes to develop systems that can automatically generate a result set, in the form of an integrated decision-making system through a formal proce…

ClassificationDecision MakingGeneral ClassificationScheduling

Towards Fault Diagnosis in Induction Motor using Fractional Fourier Transform

2024-12-24 · Usman Ali

A method for determining the current signature faults using Fractional Fourier Transform (FrFT) has been developed. The method has been applied to the real-time steady-state current of the inverter-fed high power inducti…

Fault Diagnosis