paper-with-me

Papers

MMPolymer: A Multimodal Multitask Pretraining Framework for Polymer Property Prediction

2024-06-07 · Fanmeng Wang, Wentao Guo, Minjie Cheng, Shen Yuan, Hongteng Xu, Zhifeng Gao

Polymers are high-molecular-weight compounds constructed by the covalent bonding of numerous identical or similar monomers so that their 3D structures are complex yet exhibit unignorable regularity. Typically, the properties of a polymer, such as plasticity, conductivity, bio-compatibility, and so on, are highly correlated with its 3D structure. However, existing polymer property prediction methods heavily rely on the information learned from polymer SMILES sequences (P-SMILES strings) while ignoring crucial 3D structural information, resulting in sub-optimal performance. In this work, we propose MMPolymer, a novel multimodal multitask pretraining framework incorporating polymer 1D sequential and 3D structural information to encourage downstream polymer property prediction tasks. Besides, considering the scarcity of polymer 3D data, we further introduce the "Star Substitution" strategy to extract 3D structural information effectively. During pretraining, in addition to predicting masked tokens and recovering clear 3D coordinates, MMPolymer achieves the cross-modal alignment of latent representations. Then we further fine-tune the pretrained MMPolymer for downstream polymer property prediction tasks in the supervised learning paradigm. Experiments show that MMPolymer achieves state-of-the-art performance in downstream property prediction tasks. Moreover, given the pretrained MMPolymer, utilizing merely a single modality in the fine-tuning phase can also outperform existing methods, showcasing the exceptional capability of MMPolymer in polymer feature extraction and utilization.

📄 PDF Abstract BibTeX arXiv:2406.04727

Code (1)

fanmengwang/mmpolymer 공식 구현 pytorch

Tasks

cross-modal alignmentPredictionProperty Prediction

Similar Papers 제목 키워드 기반

Multimodal machine learning with large language embedding model for polymer property prediction

2025-03-29 · Tianren Zhang, Dai-Bei Yang

Contemporary large language models (LLMs), such as GPT-4 and Llama, have harnessed extensive computational power and diverse text corpora to achieve remarkable proficiency in interpreting and generating domain-specific c…

Property Prediction

Bioplastic Design using Multitask Deep Neural Networks

2022-03-22 · Christopher Kuenneth, Jessica Lalonde, Babetta L. Marrone, Carl N. Iverson 외

Non-degradable plastic waste stays for decades on land and in water, jeopardizing our environment; yet our modern lifestyle and current technologies are impossible to sustain without plastics. Bio-synthesized and biodegr…

Diversity

OmniVec2 - A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning

2024-01-01 · CVPR 2024 1 · Siddharth Srivastava, Gaurav Sharma

We present a novel multimodal multitask network and associated training algorithm. The method is capable of ingesting data from approximately 12 different modalities namely image video audio text depth point cloud ti…

3D Point Cloud ClassificationAction ClassificationAction RecognitionAudio Classification+6

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning

2025-07-06 · Siddharth Srivastava, Gaurav Sharma arxiv

We present a novel multimodal multitask network and associated training algorithm. The method is capable of ingesting data from approximately 12 different modalities namely image, video, audio, text, depth, point cloud, …

FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning

2026-05-10 · Xing Han, Shravan Chaudhari, Tanvi Ranade, Rama Chellappa 외 arxiv

Real-world model deployment across multiple domains requires multimodal models to operate under two complementary regimes: (1) multi-task pretraining, tasks are co-available at design time where related tasks could borro…

Multi-Task LearningContinual Learning