paper-with-me

홈 › Papers

Vision-Language Generative Model for View-Specific Chest X-ray Generation

2023-02-23 · Hyungyung Lee, Da Young Lee, Wonjae Kim, Jin-Hwa Kim, Tackeun Kim, Jihang Kim, Leonard Sunwoo, Edward Choi

Synthetic medical data generation has opened up new possibilities in the healthcare domain, offering a powerful tool for simulating clinical scenarios, enhancing diagnostic and treatment quality, gaining granular medical knowledge, and accelerating the development of unbiased algorithms. In this context, we present a novel approach called ViewXGen, designed to overcome the limitations of existing methods that rely on general domain pipelines using only radiology reports to generate frontal-view chest X-rays. Our approach takes into consideration the diverse view positions found in the dataset, enabling the generation of chest X-rays with specific views, which marks a significant advancement in the field. To achieve this, we introduce a set of specially designed tokens for each view position, tailoring the generation process to the user's preferences. Furthermore, we leverage multi-view chest X-rays as input, incorporating valuable information from different views within the same study. This integration rectifies potential errors and contributes to faithfully capturing abnormal findings in chest X-ray generation. To validate the effectiveness of our approach, we conducted statistical analyses, evaluating its performance in a clinical efficacy metric on the MIMIC-CXR dataset. Also, human evaluation demonstrates the remarkable capabilities of ViewXGen, particularly in producing realistic view-specific X-rays that closely resemble the original images.

📄 PDF Abstract BibTeX arXiv:2302.12172

Code (1)

hyn2028/llm-cxr pytorch

Tasks

DiagnosticLanguage ModellingQuantization

Similar Papers 제목 키워드 기반

Any-to-Any Vision-Language Model for Multimodal X-ray Imaging and Radiological Report Generation

2025-05-02 · Daniele Molino, Francesco Di Feola, Linlin Shen, Paolo Soda 외

Generative models have revolutionized Artificial Intelligence (AI), particularly in multimodal applications. However, adapting these models to the medical domain poses unique challenges due to the complexity of medical d…

Language ModelingLanguage Modelling

CheXanatomy: Anatomy-Aware Vision-Language Modeling for Chest Radiographs

2026-06-07 · Sergios Gatidis, Curtis Langlotz, Christian Bluethgen arxiv

Vision-language models (VLMs) pretrained on large-scale image-text pairs demonstrate strong image-level understanding, but are primarily optimized for global alignment and do not explicitly encode fine-grained anatomical…

A Survey of Data Agents: Emerging Paradigm or Overstated Hype?

2025-10-27 · Yizhang Zhu, Liangwei Wang, Chenyu Yang, Xiaotian Lin 외 arxiv

The rapid advancement of large language models (LLMs) has spurred the emergence of data agents, autonomous systems designed to orchestrate Data + AI ecosystems for tackling complex data-related tasks. However, the term "…

ChestGPT: Integrating Large Language Models and Vision Transformers for Disease Detection and Localization in Chest X-Rays

2025-07-04 · Shehroz S. Khan, Petar Przulj, Ahmed Ashraf, Ali Abedi arxiv

The global demand for radiologists is increasing rapidly due to a growing reliance on medical imaging services, while the supply of radiologists is not keeping pace. Advances in computer vision and image processing techn…

Transfer Learning

NetOrchLLM: Mastering Wireless Network Orchestration with Large Language Models

2024-12-13 · Asmaa Abdallah, Abdullatif Albaseer, Abdulkadir Celik, Mohamed Abdallah 외

The transition to 6G networks promises unprecedented advancements in wireless communication, with increased data rates, ultra-low latency, and enhanced capacity. However, the complexity of managing and optimizing these n…

Natural Language Understanding