paper-with-me

Papers

Spatio-Temporal and Clinical Conditioning for Fine-Grained Radiology Report Retrieval

2026-07-02 · P. Sloan, E. Simpson, M. Mirmehdi arxiv

Radiology is vital to modern healthcare, but rising imaging demand and persistent workforce shortages strain reporting capacity and clinical workflows. Automated radiology report generation has the potential to support radiologists and help alleviate this burden; however, existing retrieval-based methods remain rigid, lack explicit anatomical grounding, and do not account for longitudinal disease progression or available clinical context. In this work, we introduce STAR3, a multimodal, spatio-temporal, attentive retrieval framework for radiology report generation that aligns region-level anatomical information with clinical indications and longitudinal changes across chest X-ray studies. Our framework employs an object detector to identify anatomically meaningful regions and retrieves semantically relevant report sentences conditioned on both current clinical context and changes observed between prior and current examinations. This design enables anatomically and temporally grounded report generation that better reflects clinical reporting practice. Experiments on the MIMIC-CXR dataset demonstrate that STAR3 outperforms current retrieval-based approaches on retrieval, NLP and clinical metrics, highlighting the value of conditioning retrieval anatomically, temporally and clinically for advancing automated radiology report generation.

📄 PDF Abstract BibTeX arXiv:2607.02024

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Fine-grained Context and Multi-modal Alignment for Freehand 3D Ultrasound Reconstruction

2024-07-05 · Zhongnuo Yan, Xin Yang, Mingyuan Luo, Jiongquan Chen 외

Fine-grained spatio-temporal learning is crucial for freehand 3D ultrasound reconstruction. Previous works mainly resorted to the coarse-grained spatial features and the separated temporal dependency learning and struggl…

Management

DiffusionVID: Denoising Object Boxes with Spatio-temporal Conditioning for Video Object Detection

2023-10-30 · IEEE Access 2023 10 · Si-Dong Roh, Ki-Seok Chung

Several existing still image object detectors suffer from image deterioration in videos, such as motion blur, camera defocus, and partial occlusion. We present DiffusionVID, a diffusion model-based video object detector,…

DenoisingGPUObjectobject-detection+2

FineMoGen: Fine-Grained Spatio-Temporal Motion Generation and Editing

2023-12-22 · NeurIPS 2023 11 · Mingyuan Zhang, Huirong Li, Zhongang Cai, Jiawei Ren 외

Text-driven motion generation has achieved substantial progress with the emergence of diffusion models. However, existing methods still struggle to generate complex motion sequences that correspond to fine-grained descri…

Mixture-of-ExpertsMotion GenerationMotion Synthesis

Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization

2026-05-22 · Tianyu Wang, Junjie Wu, Jingquan Gao, Shishuo Li arxiv

Temporal Action Localization (TAL) has been extensively studied in generic video understanding, while fine-grained sports scenarios, such as professional badminton, remain underexplored due to their complex and subtle sp…

Temporal Action Localization

Dynamic Spatio-Temporal Specialization Learning for Fine-Grained Action Recognition

2022-09-03 · Tianjiao Li, Lin Geng Foo, Qiuhong Ke, Hossein Rahmani 외

The goal of fine-grained action recognition is to successfully discriminate between action categories with subtle differences. To tackle this, we derive inspiration from the human visual system which contains specialized…

Action RecognitionFine-grained Action Recognition