paper-with-me

홈 › Papers

Adaptable and Verifiable BDI Reasoning

2020-07-23 · Peter Stringer, Rafael C. Cardoso, Xiaowei Huang, Louise A. Dennis

Long-term autonomy requires autonomous systems to adapt as their capabilities no longer perform as expected. To achieve this, a system must first be capable of detecting such changes. In this position paper, we describe a system architecture for BDI autonomous agents capable of adapting to changes in a dynamic environment and outline the required research. Specifically, we describe an agent-maintained self-model with accompanying theories of durative actions and learning new action descriptions in BDI systems.

📄 PDF Abstract BibTeX arXiv:2007.11743

Code (0)

등록된 구현이 없습니다.

Tasks

Position

Similar Papers 제목 키워드 기반

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards

2025-10-01 · Yiran Shen, Yu Xia, Jonathan Chang, Prithviraj Ammanabrolu arxiv

Aligning large language models to human preferences is inherently multidimensional, yet most pipelines collapse heterogeneous signals into a single objective. We seek to answer what it would take to simultaneously align …

Logic-Parametric Neuro-Symbolic NLI: Controlling Logical Formalisms for Verifiable LLM Reasoning

2026-01-09 · Ali Farjami, Luca Redondi, Marco Valentino arxiv

Large language models (LLMs) and theorem provers (TPs) can be effectively combined for verifiable natural language inference (NLI). However, existing approaches rely on a fixed logical formalism, a feature that limits ro…

Natural Language Inference

Zero Reinforcement Learning Towards General Domains

2025-10-29 · Yuyuan Zeng, Yufei Huang, Can Xu, Qingfeng Sun 외 arxiv

Zero Reinforcement Learning (Zero-RL) has proven to be an effective approach for enhancing the reasoning capabilities of large language models (LLMs) by directly applying reinforcement learning with verifiable rewards on…

Reinforcement Learning

Constraints-of-Thought: A Framework for Constrained Reasoning in Language-Model-Guided Search

2025-10-10 · Kamel Alrashedy, Vriksha Srihari, Zulfiqar Zaidi, Ridam Srivastava 외 arxiv

While researchers have made significant progress in enabling large language models (LLMs) to perform multi-step planning, LLMs struggle to ensure that those plans align with high-level user intent and satisfy symbolic co…

Arithmetic ReasoningCode Generation

Wan-R1: Verifiable-Reinforcement Learning for Video Reasoning

2026-03-29 · Ming Liu, Yunbei Zhang, Shilong Liu, Liwen Wang 외 arxiv

Video generation models produce visually coherent content but struggle with tasks requiring spatial reasoning and multi-step planning. Reinforcement learning (RL) offers a path to improve generalization, but its effectiv…

Reinforcement LearningSpatial ReasoningVideo Generation