paper-with-me

Papers

Benchmarking LLM-based agents for single-cell omics analysis

2025-08-16 · Yang Liu, Lu Zhou, Xiawei Du, Ruikun He, Xuguang Zhang, Rongbo Shen, Yixue Li arxiv

Background: The surge in single-cell omics data exposes limitations in traditional, manually defined analysis workflows. AI agents offer a paradigm shift, enabling adaptive planning, executable code generation, traceable decisions, and real-time knowledge fusion. However, the lack of a comprehensive benchmark critically hinders progress. Results: We introduce a novel benchmarking evaluation system to rigorously assess agent capabilities in single-cell omics analysis. This system comprises: a unified platform compatible with diverse agent frameworks and LLMs; multidimensional metrics assessing cognitive program synthesis, collaboration, execution efficiency, bioinformatics knowledge integration, and task completion quality; and 50 diverse real-world single-cell omics analysis tasks spanning multi-omics, species, and sequencing technologies. Our evaluation reveals that Grok3-beta achieves state-of-the-art performance among tested agent frameworks. Multi-agent frameworks significantly enhance collaboration and execution efficiency over single-agent approaches through specialized role division. Attribution analyses of agent capabilities identify that high-quality code generation is crucial for task success, and self-reflection has the most significant overall impact, followed by retrieval-augmented generation (RAG) and planning. Conclusions: This work highlights persistent challenges in code generation, long-context handling, and context-aware knowledge retrieval, providing a critical empirical foundation and best practices for developing robust AI agents in computational biology.

📄 PDF Abstract BibTeX arXiv:2508.13201

Code (0)

등록된 구현이 없습니다.

Tasks

Program SynthesisCode Generation

Similar Papers 제목 키워드 기반

The current state of single-cell proteomics data analysis

2022-10-03 · Christophe Vanderaa, Laurent Gatto

Sound data analysis is essential to retrieve meaningful biological information from single-cell proteomics experiments. This analysis is carried out by computational methods that are assembled into workflows, and their i…

Benchmarking

Initial recommendations for performing, benchmarking, and reporting single-cell proteomics experiments

2022-07-19 · Laurent Gatto, Ruedi Aebersold, Juergen Cox, Vadim Demichev 외

Analyzing proteins from single cells by tandem mass spectrometry (MS) has become technically feasible. While such analysis has the potential to accurately quantify thousands of proteins across thousands of single cells, …

BenchmarkingExperimental Design

scBench-Long: Verifiable Benchmarking of Long-Horizon Single-Cell Biology

2026-06-25 · Ian Diks, Zhen Yang, Arjun Banerjee, Tim Proctor 외 arxiv

Single-cell studies require analysts to convert raw measurements into specific biological claims through multi-step workflows and integration of metadata, assay context, and auxiliary evidence. Existing AI-biology benchm…

Standardised workflow for mass spectrometry-based single-cell proteomics data processing and analysis using the scp package

2023-10-20 · Samuel Grégoire, Christophe Vanderaa, Sébastien Pyr dit Ruys, Gabriel Mazzucchelli 외

Mass spectrometry (MS) based single-cell proteomics (SCP) explores cellular heterogeneity by focusing on the functional effectors of the cells - proteins. However, extracting meaningful biological information from MS dat…

Benchmarking

scMamba: A Scalable Foundation Model for Single-Cell Multi-Omics Integration Beyond Highly Variable Feature Selection

2025-06-25 · Zhen Yuan, Shaoqing Jiao, Yihang Xiao, Jiajie Peng

The advent of single-cell multi-omics technologies has enabled the simultaneous profiling of diverse omics layers within individual cells. Integrating such multimodal data provides unprecedented insights into cellular id…

BenchmarkingContrastive Learningfeature selection