paper-with-me

Papers

GenomeQA: Benchmarking General Large Language Models for Genome Sequence Understanding

2026-04-07 · Weicai Long, Yusen Hou, Junning Feng, Houcheng Su, Shuo Yang, Donglin Xie, Yanlin Zhang arxiv

Large Language Models (LLMs) are increasingly adopted as conversational assistants in genomics, where they are mainly used to reason over biological knowledge, annotations, and analysis outputs through natural language interfaces. However, existing benchmarks either focus on specialized DNA models trained for sequence prediction or evaluate biological knowledge using text-only questions, leaving the behavior of general-purpose LLMs when directly exposed to raw genome sequences underexplored. We introduce GenomeQA, a benchmark designed to provide a controlled evaluation setting for general-purpose LLMs on sequence-based genome inference tasks. GenomeQA comprises 5,200 samples drawn from multiple biological databases, with sequence lengths ranging from 6 to 1,000 base pairs (bp), spanning six task families: Enhancer and Promoter Identification, Splice Site Identification, Taxonomic Classification, Histone Mark Prediction, Transcription Factor Binding Site Prediction, and TF Motif Prediction. Across six frontier LLMs, we find that models consistently outperform random baselines and can exploit local sequence signals such as GC content and short motifs, while performance degrades on tasks that require more indirect or multi-step inference over sequence patterns. GenomeQA establishes a diagnostic benchmark for studying and improving the use of general-purpose LLMs on raw genomic sequences.

📄 PDF Abstract BibTeX arXiv:2604.05774

Code (0)

등록된 구현이 없습니다.

Tasks

Transcription Factor Binding Site Prediction

Similar Papers 제목 키워드 기반

OmniGenBench: Automating Large-scale in-silico Benchmarking for Genomic Foundation Models

2024-10-02 · Heng Yang, Jack Cole, Ke Li

The advancements in artificial intelligence in recent years, such as Large Language Models (LLMs), have fueled expectations for breakthroughs in genomic foundation models (GFMs). The code of nature, hidden in diverse gen…

Benchmarking

Genome-Factory: A Library for Tuning, Deploying, and Interpreting Genomic Foundation Models

2025-09-13 · Weimin Wu, Xuefeng Song, Yibo Wen, Qinjie Lin 외 arxiv

We introduce Genome-Factory, the first integrated Python library for tuning, deploying, and interpreting genomic foundation models. Our core contribution is to simplify and unify the workflow for genomic model developmen…

parameter-efficient fine-tuning

CLMB: deep contrastive learning for robust metagenomic binning

2021-11-18 · Pengfei Zhang, Zhengyuan Jiang, YiXuan Wang, Yu Li

The reconstruction of microbial genomes from large metagenomic datasets is a critical procedure for finding uncultivated microbial populations and defining their microbial functional roles. To achieve that, we need to pe…

BenchmarkingContrastive LearningDenoising

OmniGenBench: A Modular Platform for Reproducible Genomic Foundation Models Benchmarking

2025-05-20 · Heng Yang, Jack Cole, Yuan Li, Renzhi Chen 외

The code of nature, embedded in DNA and RNA genomes since the origin of life, holds immense potential to impact both humans and ecosystems through genome modeling. Genomic Foundation Models (GFMs) have emerged as a trans…

Benchmarking

BEND: Benchmarking DNA Language Models on biologically meaningful tasks

2023-11-21 · Frederikke Isa Marin, Felix Teufel, Marc Horlacher, Dennis Madsen 외

The genome sequence contains the blueprint for governing cellular processes. While the availability of genomes has vastly increased over the last decades, experimental annotation of the various functional, non-coding and…

BenchmarkingLanguage ModelingLanguage Modelling