paper-with-me

홈 › Papers

O-Researcher: An Open Ended Deep Research Model via Multi-Agent Distillation and Agentic RL

2026-01-07 · Yi Yao, He Zhu, Piaohong Wang, Jincheng Ren, Xinlong Yang, Qianben Chen, Xiaowan Li, Dingfeng Shi, Jiaxian Li, Qiexiang Wang, Sinuo Wang, Xinpeng Liu, Jiaqi Wu, Minghao Liu, Wangchunshu Zhou arxiv

The performance gap between closed-source and open-source large language models (LLMs) is largely attributed to disparities in access to high-quality training data. To bridge this gap, we introduce a novel framework for the automated synthesis of sophisticated, research-grade instructional data. Our approach centers on a multi-agent workflow where collaborative AI agents simulate complex tool-integrated reasoning to generate diverse and high-fidelity data end-to-end. Leveraging this synthesized data, we develop a two-stage training strategy that integrates supervised fine-tuning with a novel reinforcement learning method, designed to maximize model alignment and capability. Extensive experiments demonstrate that our framework empowers open-source models across multiple scales, enabling them to achieve new state-of-the-art performance on the major deep research benchmark. This work provides a scalable and effective pathway for advancing open-source LLMs without relying on proprietary data or models.

📄 PDF Abstract BibTeX arXiv:2601.03743

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

BioResearcher: Scenario-Guided Multi-Agent for Translational Medicine

2026-05-07 · Remigiusz Kinas, Joanna Krawczyk, Rafał Powalski, Przemysław Pietrzak 외 arxiv

Translational medicine turns underspecified development goals into evidence synthesis that must combine literature, trials, patents, and quantitative multi-omics analysis while preserving identifiers, uncertainty, and re…

Hybrid Open-Ended Tri-Evolution Makes Better Deep Researcher

2026-06-10 · Hongming Piao, Chi Liu, Mengzhuo Chen, Yan Shu 외 arxiv

Deep research and agent evolution serve as de-facto tasks for AI agents in real-world applications toward artificial general intelligence. The former enables autonomous retrieval and integration of information in open-en…

Reinforcement Learning

The Agentic Researcher: A Practical Guide to AI-Assisted Research in Mathematics and Machine Learning

2026-03-16 · Max Zimmer, Nico Pelleriti, Christophe Roux, Sebastian Pokutta arxiv

AI tools and agents are reshaping how researchers work, from proving theorems to training neural networks. Yet for many, it remains unclear how these tools fit into everyday research practice. This paper is a practical g…

Building Open-Ended Embodied Agent via Language-Policy Bidirectional Adaptation

2023-12-12 · Shaopeng Zhai, Jie Wang, Tianyi Zhang, Fuxian Huang 외

Building embodied agents on integrating Large Language Models (LLMs) and Reinforcement Learning (RL) have revolutionized human-AI interaction: researchers can now leverage language instructions to plan decision-making fo…

Decision MakingLanguage ModellingReinforcement Learning (RL)

BLADE: Benchmarking Language Model Agents for Data-Driven Science

2024-08-19 · Ken Gu, Ruoxi Shang, Ruien Jiang, Keying Kuang 외

Data-driven scientific discovery requires the iterative integration of scientific domain knowledge, statistical expertise, and an understanding of data semantics to make nuanced analytical decisions, e.g., about which va…

BenchmarkingDecision MakingLanguage ModelingLanguage Modelling+3