paper-with-me

홈 › Papers

A data- and compute-efficient chest X-ray foundation model beyond aggressive scaling

2026-02-26 · Chong Wang, Yabin Zhang, Yunhe Gao, Maya Varma, Clemence Mottez, Faidra Patsatzi, Jiaming Liu, Jin Long, Jean-Benoit Delbrouck, Sergios Gatidis, Akshay S. Chaudhari, Curtis P. Langlotz arxiv

Foundation models for medical imaging are typically pretrained on increasingly large datasets, following a "scale-at-all-costs" paradigm. However, this strategy faces two critical challenges: large-scale medical datasets often contain substantial redundancy and severe class imbalance that bias representation learning toward over-represented patterns, and indiscriminate training regardless of heterogeneity in data quality incurs considerable computational inefficiency. Here we demonstrate that active, principled data curation during pretraining can serve as a viable, cost-effective alternative to brute-force dataset enlargement. We introduce CheXficient, a chest X-ray (CXR) foundation model that selectively prioritizes informative training samples. CheXficient is pretrained on only 22.7% of 1,235,004 paired CXR images and reports while consuming under 27.3% of the total compute budget, yet achieving comparable or superior performance to its full-data counterpart and other large-scale pretrained models. We assess CheXficient across 20 individual benchmarks spanning 5 task types, including non-adapted off-the-shelf evaluations (zero-shot findings classification and crossmodal retrieval) and adapted downstream tasks (disease prediction, semantic segmentation, and radiology report generation). Further analyses show that CheXficient systematically prioritizes under-represented training samples, improving generalizability on long-tailed or rare conditions. Overall, our work offers practical insights into the data and computation demands for efficient pretraining and downstream adaptation of medical vision-language foundation models.

📄 PDF Abstract BibTeX arXiv:2602.22843

Code (0)

등록된 구현이 없습니다.

Tasks

Representation LearningSemantic Segmentation

Similar Papers 제목 키워드 기반

Less Could Be Better: Parameter-efficient Fine-tuning Advances Medical Vision Foundation Models

2024-01-22 · Chenyu Lian, Hong-Yu Zhou, Yizhou Yu, Liansheng Wang

Parameter-efficient fine-tuning (PEFT) that was initially developed for exploiting pre-trained large language models has recently emerged as an effective approach to perform transfer learning on computer vision tasks. Ho…

parameter-efficient fine-tuningTransfer Learning

Prompt Compression in Production Task Orchestration: A Pre-Registered Randomized Trial

2026-03-06 · Warren Johnson, Charles Lee arxiv

The economics of prompt compression depend not only on reducing input tokens but on how compression changes output length, which is typically priced several times higher. We evaluate this in a pre-registered six-arm rand…

MedDChest: A Content-Aware Multimodal Foundational Vision Model for Thoracic Imaging

2025-11-06 · Mahmoud Soliman, Islam Osman, Mohamed S. Shehata, Rasika Rajapakshe arxiv

The performance of vision models in medical imaging is often hindered by the prevailing paradigm of fine-tuning backbones pre-trained on out-of-domain natural images. To address this fundamental domain gap, we propose Me…

Data Augmentation

BeyondCT: A deep learning model for predicting pulmonary function from chest CT scans

2024-08-10 · Kaiwen Geng, Zhiyi Shi, Xiaoyan Zhao, Alaa Ali 외

Abstract Background: Pulmonary function tests (PFTs) and computed tomography (CT) imaging are vital in diagnosing, managing, and monitoring lung diseases. A common issue in practice is the lack of access to recorded pulm…

Computed Tomography (CT)

Joint Partitioning and Placement of Foundation Models for Real-Time Edge AI

2025-11-30 · Aladin Djuhera, Fernando Koch, Alecio Binotto arxiv

Inference over large-scale foundation models within heterogeneous edge environments necessitates a fundamentally reconfigurable orchestration substrate. Static partitioning of model layers presumes temporal stability acr…