paper-with-me

홈 › Papers

Gene-R1: Reasoning with Data-Augmented Lightweight LLMs for Gene Set Analysis

2025-09-11 · Zhizheng Wang, Yifan Yang, Qiao Jin, Zhiyong Lu arxiv

The gene set analysis (GSA) is a foundational approach for uncovering the molecular functions associated with a group of genes. Recently, LLM-powered methods have emerged to annotate gene sets with biological functions together with coherent explanatory insights. However, existing studies primarily focus on proprietary models, which have been shown to outperform their open-source counterparts despite concerns over cost and data privacy. Furthermore, no research has investigated the application of advanced reasoning strategies to the GSA task. To address this gap, we introduce Gene-R1, a data-augmented learning framework that equips lightweight and open-source LLMs with step-by-step reasoning capabilities tailored to GSA. Experiments on 1,508 in-distribution gene sets demonstrate that Gene-R1 achieves substantial performance gains, matching commercial LLMs. On 106 out-of-distribution gene sets, Gene-R1 performs comparably to both commercial and large-scale LLMs, exhibiting robust generalizability across diverse gene sources.

📄 PDF Abstract BibTeX arXiv:2509.10575

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Distilling Reasoning Without Knowledge: A Framework for Reliable LLMs

2026-03-15 · Auksarapak Kietkajornrit, Jad Tarifi, Nima Asgharbeygi arxiv

Fact-seeking question answering with large language models (LLMs) remains unreliable when answers depend on up-to-date or conflicting information. Although retrieval-augmented and tool-using LLMs reduce hallucinations, t…

Question Answering

Knowledge Integration Decay in Search-Augmented Reasoning of Large Language Models

2026-02-10 · Sangwon Yu, Ik-hwan Kim, Donghun Kang, Bongkyu Hwang 외 arxiv

Modern Large Language Models (LLMs) have demonstrated remarkable capabilities in complex tasks by employing search-augmented reasoning to incorporate external knowledge into long chains of thought. However, we identify a…

LIR$^3$AG: A Lightweight Rerank Reasoning Strategy Framework for Retrieval-Augmented Generation

2025-12-20 · Guo Chen, Junjie Huang, Huaijin Xie, Fei Sun 외 arxiv

Retrieval-Augmented Generation (RAG) effectively enhances Large Language Models (LLMs) by incorporating retrieved external knowledge into the generation process. Reasoning models improve LLM performance in multi-hop QA t…

Towards Time Series Reasoning with LLMs

2024-09-17 · Winnie Chow, Lauren Gardiner, Haraldur T. Hallgrímsson, Maxwell A. Xu 외

Multi-modal large language models (MLLMs) have enabled numerous advances in understanding and reasoning in domains like vision, but we have not yet seen this broad success for time-series. Although prior works on time-se…

Time SeriesTime Series Forecasting

Double-Calibration: Towards Reliable LLMs via Calibrating Knowledge and Reasoning Confidence

2026-01-17 · Yuyin Lu, Ziran Liang, Yanghui Rao, Wenqi Fan 외 arxiv

Reliable reasoning in Large Language Models (LLMs) is challenged by their propensity for hallucination. While augmenting LLMs with Knowledge Graphs (KGs) improves factual accuracy, existing KG-augmented methods fail to q…

Knowledge Graphs