paper-with-me

Papers

NaNa and MiGu: Semantic Data Augmentation Techniques to Enhance Protein Classification in Graph Neural Networks

2024-03-21 · Yi-Shan Lan, Pin-Yu Chen, Tsung-Yi Ho

Protein classification tasks are essential in drug discovery. Real-world protein structures are dynamic, which will determine the properties of proteins. However, the existing machine learning methods, like ProNet (Wang et al., 2022a), only access limited conformational characteristics and protein side-chain features, leading to impractical protein structure and inaccuracy of protein classes in their predictions. In this paper, we propose novel semantic data augmentation methods, Novel Augmentation of New Node Attributes (NaNa), and Molecular Interactions and Geometric Upgrading (MiGu) to incorporate backbone chemical and side-chain biophysical information into protein classification tasks and a co-embedding residual learning framework. Specifically, we leverage molecular biophysical, secondary structure, chemical bonds, and ionic features of proteins to facilitate protein classification tasks. Furthermore, our semantic augmentation methods and the co-embedding residual learning framework can improve the performance of GIN (Xu et al., 2019) on EC and Fold datasets (Bairoch, 2000; Andreeva et al., 2007) by 16.41% and 11.33% respectively. Our code is available at https://github.com/r08b46009/Code_for_MIGU_NANA/tree/main.

📄 PDF Abstract BibTeX arXiv:2403.14736

Code (1)

r08b46009/code_for_migu_nana 공식 구현 pytorch

Tasks

Data AugmentationDrug Discovery

Methods 이 논문이 사용한 방법론

GIN Per the authors, Graph Isomorphism Network (GIN) generalizes the WL test and hence achieves maximum discriminative power among GNNs.

Similar Papers 제목 키워드 기반

iMiGUE-Speech: A Spontaneous Speech Dataset for Affective Analysis

2026-02-25 · Sofoklis Kakouros, Fang Kang, Haoyu Chen arxiv

This work presents iMiGUE-Speech, an extension of the iMiGUE dataset that provides a spontaneous affective corpus for studying emotional and affective states. The new release focuses on speech and enriches the original d…

Speech Emotion RecognitionSentiment Analysis

iMiGUE: An Identity-free Video Dataset for Micro-Gesture Understanding and Emotion Analysis

2021-07-01 · CVPR 2021 1 · Xin Liu, Henglin Shi, Haoyu Chen, Zitong Yu 외

We introduce a new dataset for the emotional artificial intelligence research: identity-free video dataset for Micro-Gesture Understanding and Emotion analysis (iMiGUE). Different from existing public datasets, iMiGUE fo…

Emotion Recognition

Banana Sub-Family Classification and Quality Prediction using Computer Vision

2022-04-06 · Narayana Darapaneni, Arjun Tanndalam, Mohit Gupta, Neeta Taneja 외

India is the second largest producer of fruits and vegetables in the world, and one of the largest consumers of fruits like Banana, Papaya and Mangoes through retail and ecommerce giants like BigBasket, Grofers and Amazo…

ClassificationData Augmentationimage-classificationImage Classification+2

Unlocking Continual Learning Abilities in Language Models

2024-06-25 · Wenyu Du, Shuang Cheng, Tongxu Luo, Zihan Qiu 외

Language models (LMs) exhibit impressive performance and generalization capabilities. However, LMs struggle with the persistent challenge of catastrophic forgetting, which undermines their long-term sustainability in con…

Continual LearningInductive Bias

Are Female Carpenters like Blue Bananas? A Corpus Investigation of Occupation Gender Typicality

2024-08-06 · Da Ju, Karen Ulrich, Adina Williams

People tend to use language to mention surprising properties of events: for example, when a banana is blue, we are more likely to mention color than when it is yellow. This fact is taken to suggest that yellowness is som…