paper-with-me

Papers

Devil in the Tail: A Multi-Modal Framework for Drug-Drug Interaction Prediction in Long Tail Distinction

2024-10-16 · Liangwei Nathan Zheng, Chang George Dong, Wei Emma Zhang, Xin Chen, Lin Yue, Weitong Chen

Drug-drug interaction (DDI) identification is a crucial aspect of pharmacology research. There are many DDI types (hundreds), and they are not evenly distributed with equal chance to occur. Some of the rarely occurred DDI types are often high risk and could be life-critical if overlooked, exemplifying the long-tailed distribution problem. Existing models falter against this distribution challenge and overlook the multi-faceted nature of drugs in DDI prediction. In this paper, a novel multi-modal deep learning-based framework, namely TFDM, is introduced to leverage multiple properties of a drug to achieve DDI classification. The proposed framework fuses multimodal features of drugs, including graph-based, molecular structure, Target and Enzyme, for DDI identification. To tackle the challenge posed by the distribution skewness across categories, a novel loss function called Tailed Focal Loss is introduced, aimed at further enhancing the model performance and address gradient vanishing problem of focal loss in extremely long-tailed dataset. Intensive experiments over 4 challenging long-tailed dataset demonstrate that the TFMD outperforms the most recent SOTA methods in long-tailed DDI classification tasks. The source code is released to reproduce our experiment results: https://github.com/IcurasLW/TFMD_Longtailed_DDI.git

📄 PDF Abstract BibTeX arXiv:2410.12249

Code (1)

IcurasLW/TFMD_Longtailed_DDI 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

Focal Loss A Focal Loss function addresses class imbalance during training in tasks like object detection. Focal loss applies a modulating term to the cross entropy loss in order to…

Similar Papers 제목 키워드 기반

The Devil Is in the Details: Tackling Unimodal Spurious Correlations for Generalizable Multimodal Reward Models

2025-03-05 · Zichao Li, Xueru Wen, Jie Lou, Yuqiu Ji 외

Multimodal Reward Models (MM-RMs) are crucial for aligning Large Language Models (LLMs) with human preferences, particularly as LLMs increasingly interact with multimodal data. However, we find that MM-RMs trained on exi…

Learning of Visual Relations: The Devil is in the Tails

2021-08-22 · ICCV 2021 10 · Alakh Desai, Tz-Ying Wu, Subarna Tripathi, Nuno Vasconcelos

Significant effort has been recently devoted to modeling visual relations. This has mostly addressed the design of architectures, typically by adding parameters and increasing model complexity. However, visual relation l…

Graph GenerationScene Graph Generation

The Devil is in the Details: Boosting Guided Depth Super-Resolution via Rethinking Cross-Modal Alignment and Aggregation

2024-01-16 · Xinni Jiang, Zengsheng Kuang, Chunle Guo, Ruixun Zhang 외

Guided depth super-resolution (GDSR) involves restoring missing depth details using the high-resolution RGB image of the same scene. Previous approaches have struggled with the heterogeneity and complementarity of the mu…

cross-modal alignmentfeature selectionSuper-Resolution

MDNN: A Multimodal Deep Neural Network for Predicting Drug-Drug Interaction Events

2021-09-01 · IJCAI 2021 9 · Tengfei Lyu1, Jianliang Gao1∗, Ling Tian1, Zhao Li2 외

The interaction of multiple drugs could lead to serious events, which causes injuries and huge medical costs. Accurate prediction of drug-drug interaction (DDI) events can help clinicians make effective decisions and…

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding

2025-12-07 · Shida Gao, Feng Xue, Xiangfeng Wang, Anlong Ming 외 arxiv

Multimodal large language models (MLLMs) are rapidly expanding from general video understanding to finer-grained understanding such as spatio-temporal video grounding (STVG) and reasoning. In these tasks, an MLLM must lo…

Spatio-Temporal Video Grounding