paper-with-me

홈 › Papers

Breaking the Modality Barrier: Generative Modeling for Accurate Molecule Retrieval from Mass Spectra

2025-11-09 · Yiwen Zhang, Keyan Ding, Yihang Wu, Xiang Zhuang, Yi Yang, Qiang Zhang, Huajun Chen arxiv

Retrieving molecular structures from tandem mass spectra is a crucial step in rapid compound identification. Existing retrieval methods, such as traditional mass spectral library matching, suffer from limited spectral library coverage, while recent cross-modal representation learning frameworks often encounter modality misalignment, resulting in suboptimal retrieval accuracy and generalization. To address these limitations, we propose GLMR, a Generative Language Model-based Retrieval framework that mitigates the cross-modal misalignment through a two-stage process. In the pre-retrieval stage, a contrastive learning-based model identifies top candidate molecules as contextual priors for the input mass spectrum. In the generative retrieval stage, these candidate molecules are integrated with the input mass spectrum to guide a generative model in producing refined molecular structures, which are then used to re-rank the candidates based on molecular similarity. Experiments on both MassSpecGym and the proposed MassRET-20k dataset demonstrate that GLMR significantly outperforms existing methods, achieving over 40% improvement in top-1 accuracy and exhibiting strong generalizability.

📄 PDF Abstract BibTeX arXiv:2511.06259

Code (0)

등록된 구현이 없습니다.

Tasks

Representation LearningContrastive Learning

Similar Papers 제목 키워드 기반

Breaking the Programming Language Barrier: Multilingual Prompting to Empower Non-Native English Learners

2024-12-17 · James Prather, Brent N. Reeves, Paul Denny, Juho Leinonen 외

Non-native English speakers (NNES) face multiple barriers to learning programming. These barriers can be obvious, such as the fact that programming language syntax and instruction are often in English, or more subtle, su…

Code Generation

Deep generative model-driven multimodal prostate segmentation in radiotherapy

2019-10-23 · Kibrom Berihu Girum, Gilles Créhange, Raabid Hussain, Paul Michael Walker 외

Deep learning has shown unprecedented success in a variety of applications, such as computer vision and medical image analysis. However, there is still potential to improve segmentation in multimodal images by embedding …

Medical Image AnalysisMulti-Task LearningSegmentation

MOOSE-Star: Unlocking Tractable Training for Scientific Discovery by Breaking the Complexity Barrier

2026-03-04 · Zonglin Yang, Lidong Bing arxiv

While large language models (LLMs) show promise in scientific discovery, existing research focuses on inference or feedback-driven training, leaving the direct modeling of the generative reasoning process, $P(\text{hypot…

MedGR$^2$: Breaking the Data Barrier for Medical Reasoning via Generative Reward Learning

2025-08-28 · Weihai Zhi, Jiayan Guo, Shangyang Li arxiv

The application of Vision-Language Models (VLMs) in medicine is critically hampered by the scarcity of high-quality, expert-annotated data. Supervised Fine-Tuning (SFT) on existing datasets often leads to poor generaliza…

Reinforcement Learning

Jointly Modeling Inter- & Intra-Modality Dependencies for Multi-modal Learning

2024-05-27 · Divyam Madaan, Taro Makino, Sumit Chopra, Kyunghyun Cho

Supervised multi-modal learning involves mapping multiple modalities to a target label. Previous studies in this field have concentrated on capturing in isolation either the inter-modality dependencies (the relationships…