paper-with-me

Papers

Towards Goal-Oriented Agents for Evolving Problems Observed via Conversation

2024-01-11 · Michael Free, Andrew Langworthy, Mary Dimitropoulaki, Simon Thompson

The objective of this work is to train a chatbot capable of solving evolving problems through conversing with a user about a problem the chatbot cannot directly observe. The system consists of a virtual problem (in this case a simple game), a simulated user capable of answering natural language questions that can observe and perform actions on the problem, and a Deep Q-Network (DQN)-based chatbot architecture. The chatbot is trained with the goal of solving the problem through dialogue with the simulated user using reinforcement learning. The contributions of this paper are as follows: a proposed architecture to apply a conversational DQN-based agent to evolving problems, an exploration of training methods such as curriculum learning on model performance and the effect of modified reward functions in the case of increasing environment complexity.

📄 PDF Abstract BibTeX arXiv:2401.05822

Code (0)

등록된 구현이 없습니다.

Tasks

Chatbot

Similar Papers 제목 키워드 기반

ToM2C: Target-oriented Multi-agent Communication and Cooperation with Theory of Mind

2021-10-15 · NeurIPS 2021 12 · Yuanfei Wang, Fangwei Zhong, Jing Xu, Yizhou Wang

Being able to predict the mental states of others is a key factor to effective social interaction. It is also crucial for distributed multi-agent systems, where agents are required to communicate and cooperate. In this p…

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces

2025-02-11 · Loris Gaven, Thomas Carta, Clément Romac, Cédric Colas 외

Open-ended learning agents must efficiently prioritize goals in vast possibility spaces, focusing on those that maximize learning progress (LP). When such autotelic exploration is achieved by LLM agents trained with onli…

Hierarchical Object-Oriented POMDP Planning for Object Rearrangement

2024-12-02 · Rajesh Mangannavar, Alan Fern, Prasad Tadepalli

We present an online planning framework for solving multi-object rearrangement problems in partially observable, multi-room environments. Current object rearrangement solutions, primarily based on Reinforcement Learning …

ObjectObject Rearrangement

GOAT: A Training Framework for Goal-Oriented Agent with Tools

2025-10-14 · Hyunji Min, Sangwon Jung, Junyoung Sung, Dosung Lee 외 arxiv

Current approaches rely on zero-shot evaluation due to the absence of training data; while proprietary models such as GPT-4 exhibit strong reasoning capabilities, smaller open-source models remain ineffective at complex …

Goal-Oriented Multi-Agent Semantic Networking: Unifying Intents, Semantics, and Intelligence

2025-11-30 · Shutong Chen, Qi Liao, Adnan Aijaz, Yansha Deng arxiv

6G services are evolving toward goal-oriented and AI-native communication, which are expected to deliver transformative societal benefits across various industries and promote energy sustainability. Yet today's networkin…