paper-with-me

홈 › Papers

UI-Oceanus: Scaling GUI Agents with Synthetic Environmental Dynamics

2026-02-11 · Mengzhou Wu, Yuzhe Guo, Yuan Cao, Haochuan Lu, Songhe Zhu, Pingzhe Qu, Xin Chen, Kang Qin, Zhongpu Wang, Xiaode Zhang, Xinyi Wang, Wei Dai, Gang Cao, Yuetang Deng, Zhi Gong, Dezhi Ran, Linyi Li, Wei Yang, Tao Xie arxiv

Scaling generalist GUI agents is hindered by the data scalability bottleneck of expensive human demonstrations and the "distillation ceiling" of synthetic teacher supervision. To transcend these limitations, we propose UI-Oceanus, a framework that shifts the learning focus from mimicking high-level trajectories to mastering interaction physics via ground-truth environmental feedback. Through a systematic investigation of self-supervised objectives, we identify that forward dynamics, defined as the generative prediction of future interface states, acts as the primary driver for scalability and significantly outweighs inverse inference. UI-Oceanus leverages this insight by converting low-cost autonomous exploration, which is verified directly by system execution, into high-density generative supervision to construct a robust internal world model. Experimental evaluations across a series of models demonstrate the decisive superiority of our approach: models utilizing Continual Pre-Training (CPT) on synthetic dynamics outperform non-CPT baselines with an average success rate improvement of 7% on offline benchmarks, which amplifies to a 16.8% gain in real-world online navigation. Furthermore, we observe that navigation performance scales with synthetic data volume. These results confirm that grounding agents in forward predictive modeling offers a superior pathway to scalable GUI automation with robust cross-domain adaptability and compositional generalization.

📄 PDF Abstract BibTeX arXiv:2604.02345

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Budget-Aware Tool Use Enables Effective Agent Scaling

2025-11-21 · Tengxiao Liu, Zifeng Wang, Jin Miao, I-Hung Hsu 외 arxiv

Scaling test-time computation has been extended from language model reasoning to tool-augmented agents, where scaling involves not only thinking in tokens but also acting via tool calls that directly constrain environmen…

Thinking by Doing: Building Efficient World Model Reasoning in LLMs via Multi-turn Interaction

2025-11-28 · Bao Shu, Yan Cai, Jianjian Sun, Chunrui Han 외 arxiv

Developing robust world model reasoning is crucial for large language model (LLM) agents to plan and interact in complex environments. While multi-turn interaction offers a superior understanding of environmental dynamic…

Active Learning

Limits to AI Growth: The Ecological and Social Consequences of Scaling

2025-01-29 · Eshta Bhardwaj, Rohan Alexander, Christoph Becker

The accelerating development and deployment of AI technologies depend on the continued ability to scale their infrastructure. This has implied increasing amounts of monetary investment and natural resources. Frontier AI …

Scaling laws in global corporations as a benchmarking approach to assess environmental performance

2022-06-07 · Rossana Mastrandrea, Rob ter Burg, Yuli Shan, Klaus Hubacek 외

The largest 6,529 international corporations are accountable for almost 30% of global CO2e emissions. A growing awareness of the role of the corporate world in the path toward sustainability has led many shareholders and…

BenchmarkingOpen-Ended Question Answering

Second-order Phase Transition in Phytoplankton Trait Dynamics

2020-04-01 · Jenny Held, Tom Lorimer, Francesco Pomati, Ruedi Stoop 외

Key traits of unicellular species, like cell size, often follow scale-free or self-similar distributions, hinting at the possibility of an underlying critical process. However, linking such empirical scaling laws to the …