paper-with-me

Papers

Multi Anatomy X-Ray Foundation Model

2025-09-15 · Nishank Singla, Krisztian Koos, Farzin Haddadpour, Amin Honarmandi Shandiz, Lovish Chum, Xiaojian Xu, Qing Jin, Erhan Bas arxiv

X-ray imaging is a ubiquitous in radiology, yet most existing AI foundation models are limited to chest anatomy and fail to generalize across broader clinical tasks. In this work, we introduce XR-0, the multi-anatomy X-ray foundation model using self-supervised learning on a large, private dataset of 1.15 million images spanning diverse anatomical regions and evaluated across 12 datasets and 20 downstream tasks, including classification, retrieval, segmentation, localization, visual grounding, and report generation. XR-0 achieves state-of-the-art performance on most multi-anatomy tasks and remains competitive on chest-specific benchmarks. Our results demonstrate that anatomical diversity and supervision are critical for building robust, general-purpose medical vision models, paving the way for scalable and adaptable AI systems in radiology.

📄 PDF Abstract BibTeX arXiv:2509.12146

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised LearningVisual Grounding

Similar Papers 제목 키워드 기반

Lamps: Learning Anatomy from Multiple Perspectives via Self-supervision in Chest Radiographs

2025-12-28 · Ziyu Zhou, Haozhe Luo, Mohammad Reza Hosseinzadeh Taher, Jiaxuan Pang 외 arxiv

Foundation models have been successful in natural language processing and computer vision because they are capable of capturing the underlying structures (foundation) of natural languages. However, in medical imaging, th…

Self-Supervised Learning

Anatomy Contextualized Adaptation of CT Foundation Models

2026-07-29 · Roshan Kenia, Stephanie L McNamara, William Lotter arxiv

CT vision-language foundation models have demonstrated promising performance across downstream tasks, but are typically trained with whole-volume representations that dilute fine-grained anatomical signals. Fine-grained …

Towards Foundation Models Learned from Anatomy in Medical Imaging via Self-Supervision

2023-09-27 · Mohammad Reza Hosseinzadeh Taher, Michael B. Gotway, Jianming Liang

Human anatomy is the foundation of medical imaging and boasts one striking characteristic: its hierarchy in nature, exhibiting two intrinsic properties: (1) locality: each anatomical structure is morphologically distinct…

AnatomySelf-Supervised Learning

Surgical Anatomy Recognition with Context Learning using Foundation Representations

2026-06-20 · Ronald L. P. D. de Jong, Tim J. M. Jaspers, Raf A. H. Vervoort, Aron F. H. A. Bakker 외 arxiv

Accurate recognition of anatomical structures is essential for safe and effective minimally invasive surgery (MIS), yet it remains underexplored in surgical computer vision due to limited annotated data and methods tailo…

Video Semantic SegmentationScene UnderstandingObject Tracking

6 Fingers, 1 Kidney: Natural Adversarial Medical Images Reveal Critical Weaknesses of Vision-Language Models

2025-12-03 · Leon Mayer, Piotr Kalinowski, Caroline Ebersbach, Marcel Knopp 외 arxiv

Vision-language models (VLMs) are increasingly integrated into clinical workflows. However, existing benchmarks primarily assess performance on common anatomical presentations and fail to capture the challenges posed by …