paper-with-me

홈 › Papers

Beyond Simply Environment Scaling: Designing Effective Environment Distributions for Multimodal Agent Learning

2026-08-04 · Kejian Zhu, Zhuoran Jin, Dongqi Huang, Hongbang Yuan, Yupu Hao, Kang Liu, Jun Zhao arxiv

Recent works train agents by constructing large-scale multimodal environment pools. However, we find that simply increasing the number of multimodal environments does not always benefit. We further analyze the limitations in current multimodal environment distributions through a series of experiments. Based on these findings, we study how to build more effective training environment distributions from two dimensions: diversity and difficulty structure. For diversity, we propose Ability-aware Environment Selection (AES) to obtain diverse environment sets. For difficulty structure, we propose Hierarchical Difficulty Curriculum (HDC), which organizes curriculum learning through two difficulty levels: harness weakening and state-scale progression. Experiments show that AES and HDC effectively improve multimodal agent training.

📄 PDF Abstract BibTeX arXiv:2608.03571

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Wukong: Towards a Scaling Law for Large-Scale Recommendation

2024-03-04 · Buyun Zhang, Liang Luo, Yuxin Chen, Jade Nie 외

Scaling laws play an instrumental role in the sustainable improvement in model quality. Unfortunately, recommendation models to date do not exhibit such laws similar to those observed in the domain of large language mode…

Language ModellingLarge Language Model

So You Think You Can Scale Up Autonomous Robot Data Collection?

2024-11-04 · Suvir Mirchandani, Suneel Belkhale, Joey Hejna, Evelyn Choi 외

A long-standing goal in robot learning is to develop methods for robots to acquire new skills autonomously. While reinforcement learning (RL) comes with the promise of enabling autonomous data collection, it remains chal…

Imitation LearningReinforcement Learning (RL)

Budget-Aware Tool Use Enables Effective Agent Scaling

2025-11-21 · Tengxiao Liu, Zifeng Wang, Jin Miao, I-Hung Hsu 외 arxiv

Scaling test-time computation has been extended from language model reasoning to tool-augmented agents, where scaling involves not only thinking in tokens but also acting via tool calls that directly constrain environmen…

$k$NN Prompting: Beyond-Context Learning with Calibration-Free Nearest Neighbor Inference

2023-03-24 · Benfeng Xu, Quan Wang, Zhendong Mao, Yajuan Lyu 외

In-Context Learning (ICL), which formulates target tasks as prompt completion conditioned on in-context demonstrations, has become the prevailing utilization of LLMs. In this paper, we first disclose an actual predicamen…

In-Context Learning

Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations

2024-05-28 · Alexander Hägele, Elie Bakouch, Atli Kosson, Loubna Ben allal 외

Scale has become a main ingredient in obtaining strong machine learning models. As a result, understanding a model's scaling properties is key to effectively designing both the right training setup as well as future gene…

GPU