paper-with-me

홈 › Papers

Unlocking Smarter Device Control: Foresighted Planning with a World Model-Driven Code Execution Approach

2025-05-22 · Xiaoran Yin, Xu Luo, Hao Wu, Lianli Gao, Jingkuan Song

The automatic control of mobile devices is essential for efficiently performing complex tasks that involve multiple sequential steps. However, these tasks pose significant challenges due to the limited environmental information available at each step, primarily through visual observations. As a result, current approaches, which typically rely on reactive policies, focus solely on immediate observations and often lead to suboptimal decision-making. To address this problem, we propose \textbf{Foresighted Planning with World Model-Driven Code Execution (FPWC)},a framework that prioritizes natural language understanding and structured reasoning to enhance the agent's global understanding of the environment by developing a task-oriented, refinable \emph{world model} at the outset of the task. Foresighted actions are subsequently generated through iterative planning within this world model, executed in the form of executable code. Extensive experiments conducted in simulated environments and on real mobile devices demonstrate that our method outperforms previous approaches, particularly achieving a 44.4\% relative improvement in task success rate compared to the state-of-the-art in the simulated environment. Code and demo are provided in the supplementary material.

📄 PDF Abstract BibTeX arXiv:2505.16422

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingNatural Language Understanding

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

"Get ready for a party": Exploring smarter smart spaces with help from large language models

2023-03-24 · Evan King, Haoxiang Yu, Sangsu Lee, Christine Julien

The right response to someone who says "get ready for a party" is deeply influenced by meaning and context. For a smart home assistant (e.g., Google Home), the ideal response might be to survey the available devices in t…

Distill to Think, Foresee to Act: Cognitive-Physical Reinforcement Learning for Autonomous Driving

2026-05-20 · Yang Wu, Qiang Meng, Zhaojiang Liu, Youquan Liu 외 arxiv

Current end-to-end autonomous driving models are fundamentally constrained by the behavioral cloning ceiling of imitation learning. While reinforcement learning offers a path to smarter autonomy, it demands two missing p…

Reinforcement LearningAutonomous Driving

Towards Adaptive Self-Improvement for Smarter Energy Systems

2025-01-31 · Alexander Sommer, Peter Bazan, Jonathan Fellerer, Behnam Babaeian 외

This paper introduces a hierarchical framework for decision-making and optimization, leveraging Large Language Models (LLMs) for adaptive code generation. Instead of direct decision-making, LLMs generate and refine execu…

Code GenerationDecision Making

Low-Cost Compact Theft-Detection System using MPU-6050 and Blynk IoT Platform

2020-12-18 · Atharva Karnik, Diksha Adke, Pushkar Sathe

The system explained in this paper provides a compact smart surveillance system. Recent years have seen the Internet of Things (IoT) dominating in various fields of applications. With devices getting smarter and insurgen…

Coverage Path Planning for Autonomous Sailboats in Inhomogeneous and Time-Varying Oceans: A Spatiotemporal Optimization Approach

2026-02-13 · Yang An, Zhikang Ge, Taiyu Zhang, Jean-Baptiste R. G. Souppez 외 arxiv

Autonomous sailboats are well suited for long-duration ocean observation due to their wind-driven endurance. However, their performance is highly anisotropic and strongly influenced by inhomogeneous and time-varying wind…