paper-with-me

홈 › Papers

INFUSER: Influence-Guided Self-Evolution Improves Reasoning

2026-06-08 · Siyu Chen, Miao Lu, Beining Wu, Heejune Sheen, Fengzhuo Zhang, Shuangning Li, Zhiyuan Li, Jose Blanchet, Tianhao Wang, Zhuoran Yang arxiv

Self-evolution offers a scalable path to stronger reasoning: a pretrained language model improves itself with only minimal external supervision. Yet existing methods either depend on extensively curated or teacher-generated training data, or, when the generator runs unsupervised, reward it by a difficulty heuristic that need not improve the solver. We introduce INFUSER, an iterative co-training framework with two co-evolving roles: a Generator that drafts questions and reference golden answers from a pool of unstructured, automatically collected documents, and a Solver that improves by training on them. The solver is trained with standard correctness rewards against the generator-provided answers, while the generator is rewarded by an optimizer-aware influence score that measures whether each proposed question would actually improve the solver on the target distribution. Because this continuous, noisy influence score is poorly served by standard GRPO, we propose DuGRPO, a dual-normalized variant of GRPO, for generator training. Together, these turn the document pool into an adaptive curriculum that favors questions useful to the current solver, not just hard ones. On Qwen3-8B-Base, INFUSER outperforms strong self-evolution baselines with over 20% relative improvement on Olympiad and SuperGPQA benchmarks, and an 8B INFUSER co-evolving generator outperforms a frozen 32B thinking generator on math and coding. Ablations confirm each design choice is necessary, and two extensions, applying INFUSER to an instruction-finetuned anchor and augmenting it with rule-verifiable RLVR data, further demonstrate the flexibility and generalizability of the framework. Code is available at https://github.com/FFishy-git/INFUSER.

📄 PDF Abstract BibTeX arXiv:2606.09052

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MusicInfuser: Making Video Diffusion Listen and Dance

2025-03-18 · Susung Hong, Ira Kemelmacher-Shlizerman, Brian Curless, Steven M. Seitz

We introduce MusicInfuser, an approach for generating high-quality dance videos that are synchronized to a specified music track. Rather than attempting to design and train a new multimodal audio-video model, we show how…

Video Generation

PromptInfuser: How Tightly Coupling AI and UI Design Impacts Designers' Workflows

2023-10-24 · Savvas Petridis, Michael Terry, Carrie J. Cai

Prototyping AI applications is notoriously difficult. While large language model (LLM) prompting has dramatically lowered the barriers to AI prototyping, designers are still prototyping AI functionality and UI separately…

Language ModelingLanguage ModellingLarge Language Model

StainFuser: Controlling Diffusion for Faster Neural Style Transfer in Multi-Gigapixel Histology Images

2024-03-14 · Robert Jewsbury, Ruoyu Wang, Abhir Bhalerao, Nasir Rajpoot 외

Stain normalization algorithms aim to transform the color and intensity characteristics of a source multi-gigapixel histology image to match those of a target image, mitigating inconsistencies in the appearance of stains…

Computational EfficiencyInstance SegmentationSemantic SegmentationStyle Transfer+1

InfuserKI: Enhancing Large Language Models with Knowledge Graphs via Infuser-Guided Knowledge Integration

2024-02-18 · Fali Wang, Runxue Bao, Suhang Wang, Wenchao Yu 외

Large Language Models (LLMs) have achieved exceptional capabilities in open generation across various domains, yet they encounter difficulties with tasks that require intensive knowledge. To address these challenges, met…

Knowledge Graphs

Disentangled Multimodal Brain MR Image Translation via Transformer-based Modality Infuser

2024-02-01 · Jihoon Cho, Xiaofeng Liu, Fangxu Xing, Jinsong Ouyang 외

Multimodal Magnetic Resonance (MR) Imaging plays a crucial role in disease diagnosis due to its ability to provide complementary information by analyzing a relationship between multimodal images on the same subject. Acqu…

Brain Tumor SegmentationTranslationTumor Segmentation