paper-with-me

홈 › Papers

ReInAgent: A Context-Aware GUI Agent Enabling Human-in-the-Loop Mobile Task Navigation

2025-10-09 · Haitao Jia, Ming He, Zimo Yin, Likang Wu, Jianping Fan, Jitao Sang arxiv

Mobile GUI agents exhibit substantial potential to facilitate and automate the execution of user tasks on mobile phones. However, exist mobile GUI agents predominantly privilege autonomous operation and neglect the necessity of active user engagement during task execution. This omission undermines their adaptability to information dilemmas including ambiguous, dynamically evolving, and conflicting task scenarios, leading to execution outcomes that deviate from genuine user requirements and preferences. To address these shortcomings, we propose ReInAgent, a context-aware multi-agent framework that leverages dynamic information management to enable human-in-the-loop mobile task navigation. ReInAgent integrates three specialized agents around a shared memory module: an information-managing agent for slot-based information management and proactive interaction with the user, a decision-making agent for conflict-aware planning, and a reflecting agent for task reflection and information consistency validation. Through continuous contextual information analysis and sustained user-agent collaboration, ReInAgent overcomes the limitation of existing approaches that rely on clear and static task assumptions. Consequently, it enables more adaptive and reliable mobile task navigation in complex, real-world scenarios. Experimental results demonstrate that ReInAgent effectively resolves information dilemmas and produces outcomes that are more closely aligned with genuine user preferences. Notably, on complex tasks involving information dilemmas, ReInAgent achieves a 25% higher success rate than Mobile-Agent-v2.

📄 PDF Abstract BibTeX arXiv:2510.07988

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Survey on Context-Aware Multi-Agent Systems: Techniques, Challenges and Future Directions

2024-02-03 · Hung Du, Srikanth Thudumu, Rajesh Vasa, Kon Mouzakis

Research interest in autonomous agents is on the rise as an emerging topic. The notable achievements of Large Language Models (LLMs) have demonstrated the considerable potential to attain human-like intelligence in auton…

Autonomous DrivingCollision AvoidanceManagementNavigate

PECAN: Leveraging Policy Ensemble for Context-Aware Zero-Shot Human-AI Coordination

2023-01-16 · Xingzhou Lou, Jiaxian Guo, Junge Zhang, Jun Wang 외

Zero-shot human-AI coordination holds the promise of collaborating with humans without human data. Prevailing methods try to train the ego agent with a population of partners via self-play. However, these methods suffer …

Diversity

ContextNav: Towards Agentic Multimodal In-Context Learning

2025-10-06 · Honghao Fu, Yuan Ouyang, Kai-Wei Chang, Yiwei Wang 외 arxiv

Recent advances demonstrate that multimodal large language models (MLLMs) exhibit strong multimodal in-context learning (ICL) capabilities, enabling them to adapt to novel vision-language tasks from a few contextual exam…

CP-Agent: Context-Aware Multimodal Reasoning for Cellular Morphological Profiling under Chemical Perturbations

2026-06-02 · Yuxin Zhang, Yiyao Li, Ping Shu Ho, Simon See 외 arxiv

Cell Painting combines multiplexed fluorescent staining, high-content imaging, and quantitative analysis to generate high-dimensional phenotypic readouts to support diverse downstream tasks such as mechanism-of-action (M…

Representation LearningMultimodal ReasoningDrug Discovery

I Was Blind but Now I See: Implementing Vision-Enabled Dialogue in Social Robots

2023-11-15 · Giulio Antonio Abbo, Tony Belpaeme

In the rapidly evolving landscape of human-computer interaction, the integration of vision capabilities into conversational agents stands as a crucial advancement. This paper presents an initial implementation of a dialo…

Computational EfficiencyPrompt Engineering