paper-with-me

Papers

GenoTEX: An LLM Agent Benchmark for Automated Gene Expression Data Analysis

2024-06-21 · Haoyang Liu, ShuYu Chen, Ye Zhang, Haohan Wang

Recent advancements in machine learning have significantly improved the identification of disease-associated genes from gene expression datasets. However, these processes often require extensive expertise and manual effort, limiting their scalability. Large Language Model (LLM)-based agents have shown promise in automating these tasks due to their increasing problem-solving abilities. To support the evaluation and development of such methods, we introduce GenoTEX, a benchmark dataset for the automated analysis of gene expression data. GenoTEX provides analysis code and results for solving a wide range of gene-trait association problems, encompassing dataset selection, preprocessing, and statistical analysis, in a pipeline that follows computational genomics standards. The benchmark includes expert-curated annotations from bioinformaticians to ensure accuracy and reliability. To provide baselines for these tasks, we present GenoAgent, a team of LLM-based agents that adopt a multi-step programming workflow with flexible self-correction, to collaboratively analyze gene expression datasets. Our experiments demonstrate the potential of LLM-based methods in analyzing genomic data, while error analysis highlights the challenges and areas for future improvement. We propose GenoTEX as a promising resource for benchmarking and enhancing automated methods for gene expression data analysis. The benchmark is available at https://github.com/Liu-Hy/GenoTEX.

📄 PDF Abstract BibTeX arXiv:2406.15341

Code (1)

liu-hy/genotex 공식 구현

Tasks

AI AgentAutoMLBenchmarkingCode GenerationLarge Language Model

Similar Papers 제목 키워드 기반

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis

2025-07-28 · Haoyang Liu, Yijiang Li, Haohan Wang arxiv

Gene expression analysis holds the key to many biomedical discoveries, yet extracting insights from raw transcriptomic data remains formidable due to the complexity of multiple large, semi-structured files and the need f…

RefDrone: A Challenging Benchmark for Referring Expression Comprehension in Drone Scenes

2025-02-01 · Zhichao Sun, Yepeng Liu, Huachao Zhu, Yuliang Gu 외

Drones have become prevalent robotic platforms with diverse applications, showing significant potential in Embodied Artificial Intelligence (Embodied AI). Referring Expression Comprehension (REC) enables drones to locate…

Referring ExpressionReferring Expression Comprehension

PersonaVlog: Personalized Multimodal Vlog Generation with Multi-Agent Collaboration and Iterative Self-Correction

2025-08-19 · Xiaolu Hou, Bing Ma, Jiaxiang Cheng, Xuhua Ren 외 arxiv

With the growing demand for short videos and personalized content, automated Video Log (Vlog) generation has become a key direction in multimodal content creation. Existing methods mostly rely on predefined scripts, lack…

A Search for Improved Performance in Regular Expressions

2017-04-13 · Brendan Cody-Kenny, Michael Fenton, Adrian Ronayne, Eoghan Considine 외

The primary aim of automated performance improvement is to reduce the running time of programs while maintaining (or improving on) functionality. In this paper, Genetic Programming is used to find performance improvement…

Diversity

Grounding Language in Multi-Perspective Referential Communication

2024-10-04 · Zineng Tang, Lingjun Mao, Alane Suhr

We introduce a task and dataset for referring expression generation and comprehension in multi-agent embodied environments. In this task, two agents in a shared scene must take into account one another's visual perspecti…

Referring ExpressionReferring expression generation