paper-with-me

홈 › Papers

Sci-CoE: Co-evolving Scientific Reasoning LLMs via Geometric Consensus with Sparse Supervision

2026-02-12 · Xiaohan He, Shiyang Feng, Songtao Huang, Lei Bai, Bin Wang, Bo Zhang arxiv

Large language models (LLMs) have demonstrated exceptional reasoning capabilities, and co-evolving paradigms have shown promising results in domains such as code and math. However, in scientific reasoning tasks, these models remain fragile due to unreliable solution evaluation and limited diversity in verification strategies. In this work, we propose Sci-CoE, a two-stage scientific co-evolving framework that enables models to self-evolve as both solver and verifier through a transition from sparse supervision to unsupervised learning. In the first stage, the model uses a small set of annotated data to establish fundamental correctness judgment anchors for the Verifier. In the second stage, we introduce a geometric reward mechanism that jointly considers consensus, reliability, and diversity, driving large-scale self-iteration on unlabeled data. Experiments on several general scientific benchmarks demonstrate that Sci-CoE enhances complex reasoning capabilities and exhibits strong scalability, facilitating the construction of more robust and diverse evaluation systems. Codes are available at https://github.com/InternScience/Sci-CoE.

📄 PDF Abstract BibTeX arXiv:2602.12164

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Can LLMs extract scientific consensus? A case study in high-temperature superconductivity

2026-05-26 · Mouyang Cheng, Wenhao He, Zhuotao Jin, Bowen Yu 외 arxiv

Scientific knowledge is increasingly dispersed across vast and heterogeneous scientific literature, where important claims are often implicit, evolving, and internally debated. While large language models (LLMs) have sho…

Information Extraction

SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments

2026-04-15 · Dinging Li, Yingxiu Zhao, Xinrui Cheng, Kangheng Lin 외 arxiv

Spatial reasoning over three-dimensional scenes is a core capability for embodied intelligence, yet continuous model improvement remains bottlenecked by the cost of geometric annotation. The self-evolving paradigm offers…

Spatial ReasoningPoint Clouds

SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System

2025-11-22 · Zhiyu Xu, Weilong Yan, Yufei Shi, Xin Meng 외 arxiv

Recent advancements in multimodal large language models (MLLMs) and video agent systems have significantly improved general video understanding. However, when applied to scientific video understanding and educating, a do…

TopoAgent: A Self-Evolving Topological Agent for Multimodal Scientific Reasoning

2026-07-16 · Mingze Xu, Yinghui Li, Jiayi Kuang, Zhanhui Kang 외 arxiv

While Multimodal Large Language Models (MLLMs) excel in general tasks, rigorous scientific reasoning remains challenging due to the limitations of monolithic, linear planning. Such sequential designs often suffer from vi…

A Survey of Scientific Large Language Models: From Data Foundations to Agent Frontiers

2025-08-28 · Ming Hu, Chenglong Ma, Wei Li, Wanghan Xu 외 arxiv

Scientific Large Language Models (Sci-LLMs) are transforming how knowledge is represented, integrated, and applied in scientific research, yet their progress is shaped by the complex nature of scientific data. This surve…