paper-with-me

Papers

InstructMol: Multi-Modal Integration for Building a Versatile and Reliable Molecular Assistant in Drug Discovery

2023-11-27 · He Cao, Zijing Liu, Xingyu Lu, Yuan YAO, Yu Li

The rapid evolution of artificial intelligence in drug discovery encounters challenges with generalization and extensive training, yet Large Language Models (LLMs) offer promise in reshaping interactions with complex molecular data. Our novel contribution, InstructMol, a multi-modal LLM, effectively aligns molecular structures with natural language via an instruction-tuning approach, utilizing a two-stage training strategy that adeptly combines limited domain-specific data with molecular and textual information. InstructMol showcases substantial performance improvements in drug discovery-related molecular tasks, surpassing leading LLMs and significantly reducing the gap with specialized models, thereby establishing a robust foundation for a versatile and dependable drug discovery assistant.

📄 PDF Abstract BibTeX arXiv:2311.16208

Code (1)

idea-xl/instructmol 공식 구현 pytorch

Tasks

Drug DiscoveryMolecule Captioning

Similar Papers 제목 키워드 기반

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation

2025-12-25 · Jinqi Xiao, Qing Yan, Liming Jiang, Zichuan Liu 외 arxiv

Parameter-Efficient Fine-Tuning of Diffusion Transformers (DiTs) for diverse, multi-conditional tasks often suffers from task interference when using monolithic adapters like LoRA. The Mixture of Low-rank Experts (MoLE) …

parameter-efficient fine-tuningConditional Image Generation

M3imic: Learning a Versatile Whole-Body Controller for Multimodal Motion Mimicking

2026-06-03 · Zuxing Lu, Ziang Zheng, Yao Lyu, Jingyu Liu 외 arxiv

Building a general-purpose whole-body controller is essential for enabling diverse motion capabilities in humanoid robots across a wide range of downstream tasks, including locomotion and loco-manipulation. Different tas…

Reinforcement Learning

UnifiedVisionGPT: Streamlining Vision-Oriented AI through Generalized Multimodal Framework

2023-11-16 · Chris Kelly, Luhui Hu, Cindy Yang, Yu Tian 외

In the current landscape of artificial intelligence, foundation models serve as the bedrock for advancements in both language and vision domains. OpenAI GPT-4 has emerged as the pinnacle in large language models (LLMs), …

Versatile Medical Image Segmentation Learned from Multi-Source Datasets via Model Self-Disambiguation

2023-11-17 · CVPR 2024 1 · Xiaoyang Chen, Hao Zheng, Yuemeng Li, Yuncong Ma 외

A versatile medical image segmentation model applicable to images acquired with diverse equipment and protocols can facilitate model deployment and maintenance. However, building such a model typically demands a large, d…

Image SegmentationMedical Image SegmentationOrgan SegmentationSegmentation+1

QuadrupedGPT: Towards a Versatile Quadruped Agent in Open-ended Worlds

2024-06-24 · Yuting Mei, Ye Wang, Sipeng Zheng, Qin Jin

As robotic agents increasingly assist humans in reality, quadruped robots offer unique opportunities for interaction in complex scenarios due to their agile movement. However, building agents that can autonomously naviga…

Decision MakingNavigate