paper-with-me

홈 › Papers

XAttn-BMD: Multimodal Deep Learning with Cross-Attention for Femoral Neck Bone Mineral Density Estimation

2025-11-18 · Yilin Zhang, Leo D. Westbury, Elaine M. Dennison, Nicholas C. Harvey, Nicholas R. Fuggle, Rahman Attar arxiv

Poor bone health is a significant public health concern, and low bone mineral density (BMD) leads to an increased fracture risk, a key feature of osteoporosis. We present XAttn-BMD (Cross-Attention BMD), a multimodal deep learning framework that predicts femoral neck BMD from hip X-ray images and structured clinical metadata. It utilizes a novel bidirectional cross-attention mechanism to dynamically integrate image and metadata features for cross-modal mutual reinforcement. A Weighted Smooth L1 loss is tailored to address BMD imbalance and prioritize clinically significant cases. Extensive experiments on the data from the Hertfordshire Cohort Study show that our model outperforms the baseline models in regression generalization and robustness. Ablation studies confirm the effectiveness of both cross-attention fusion and the customized loss function. Experimental results show that the integration of multimodal data via cross-attention outperforms naive feature concatenation without cross-attention, reducing MSE by 16.7%, MAE by 6.03%, and increasing the R2 score by 16.4%, highlighting the effectiveness of the approach for femoral neck BMD estimation. Furthermore, screening performance was evaluated using binary classification at clinically relevant femoral neck BMD thresholds, demonstrating the model's potential in real-world scenarios.

📄 PDF Abstract BibTeX arXiv:2511.14604

Code (0)

등록된 구현이 없습니다.

Tasks

Multimodal Deep LearningBinary ClassificationDensity Estimation

Similar Papers 제목 키워드 기반

LV-XAttn: Distributed Cross-Attention for Long Visual Inputs in Multimodal Large Language Models

2025-02-04 · Tzu-Tao Chang, Shivaram Venkataraman

Cross-attention is commonly adopted in multimodal large language models (MLLMs) for integrating visual information into the language backbone. However, in applications with large visual inputs, such as video understandin…

GPUVideo Understanding

XAttnRes: Cross-Stage Attention Residuals for Medical Image Segmentation

2026-03-28 · Xinyu Liu, Qing Xu, Zhen Chen arxiv

In the field of Large Language Models (LLMs), Attention Residuals have recently demonstrated that learned, selective aggregation over all preceding layer outputs can outperform fixed residual connections. We propose Cros…

Medical Image Segmentation

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention

2025-02-06 · Yixin Liu, Lie Lu, Jihui Jin, Lichao Sun 외

The rapid proliferation of generative audio synthesis and editing technologies has raised significant concerns about copyright infringement, data provenance, and the spread of misinformation through deepfake audio. Water…

Audio SynthesisFace SwappingMisinformation

Globular structure of the hypermineralized tissue in human femoral neck

2020-09-17 · Qiong Wang, Tengteng Tang, David Cooper, Felipe Eltit 외

Bone becomes more fragile with ageing. Among many structural changes, a thin layer of highly mineralized and brittle tissue covers part of the external surface of the thin femoral neck cortex in older people and has been…

Segmentation of tibiofemoral joint tissues from knee MRI using MtRA-Unet and incorporating shape information: Data from the Osteoarthritis Initiative

2024-01-23 · Akshay Daydar, Alik Pramanick, Arijit Sur, Subramani Kanagaraj

Knee Osteoarthritis (KOA) is the third most prevalent Musculoskeletal Disorder (MSD) after neck and back pain. To monitor such a severe MSD, a segmentation map of the femur, tibia and tibiofemoral cartilage is usually ac…

Segmentation