paper-with-me

홈 › Papers

ALMol: Aligned Language-Molecule Translation LLMs through Offline Preference Contrastive Optimisation

2024-05-14 · Dimitris Gkoumas

The field of chemistry and Artificial Intelligence (AI) intersection is an area of active research that aims to accelerate scientific discovery. The integration of large language models (LLMs) with scientific modalities has shown significant promise in this endeavour. However, challenges persist in effectively addressing training efficacy and the out-of-distribution problem, particularly as existing approaches rely on larger models and datasets. In this context, we focus on machine language-molecule translation and deploy a novel training approach called contrastive preference optimisation, which avoids generating translations that are merely adequate but not perfect. To ensure generalisability and mitigate memorisation effects, we conduct experiments using only 10% of the data. Our results demonstrate that our models achieve up to a 32% improvement compared to counterpart models. Finally, we introduce a fine-grained, domain-agnostic evaluation method to assess hallucination in LLMs and promote responsible use.

📄 PDF Abstract BibTeX arXiv:2405.08619

Code (0)

등록된 구현이 없습니다.

Tasks

Hallucinationscientific discoveryTranslation

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Empowering Molecule Discovery for Molecule-Caption Translation with Large Language Models: A ChatGPT Perspective

2023-06-11 · Jiatong Li, Yunqing Liu, Wenqi Fan, Xiao-Yong Wei 외

Molecule discovery plays a crucial role in various scientific fields, advancing the design of tailored materials and drugs. However, most of the existing methods heavily rely on domain experts, require excessive computat…

In-Context LearningMolecule CaptioningNatural Language UnderstandingText-based de novo Molecule Generation+1

Large Language Models are In-Context Molecule Learners

2024-03-07 · Jiatong Li, Wei Liu, Zhihao Ding, Wenqi Fan 외

Large Language Models (LLMs) have demonstrated exceptional performance in biochemical tasks, especially the molecule caption translation task, which aims to bridge the gap between molecules and natural language texts. Ho…

Cross-Modal RetrievalIn-Context LearningRe-RankingRetrieval+1

Multi-OphthaLingua: A Multilingual Benchmark for Assessing and Debiasing LLM Ophthalmological QA in LMICs

2024-12-18 · David Restrepo, Chenwei Wu, Zhengxu Tang, Zitao Shuai 외

Current ophthalmology clinical workflows are plagued by over-referrals, long waits, and complex and heterogeneous medical records. Large language models (LLMs) present a promising solution to automate various procedures …

Question AnsweringRAGRetrievalRetrieval-augmented Generation+1

BEnchmarking LLMs for Ophthalmology (BELO) for Ophthalmological Knowledge and Reasoning

2025-07-21 · Sahana Srinivasan, Xuguang Ai, Thaddaeus Wai Soon Lo, Aidan Gilson 외 arxiv

Current benchmarks evaluating large language models (LLMs) in ophthalmology are limited in scope and disproportionately prioritise accuracy. We introduce BELO (BEnchmarking LLMs for Ophthalmology), a standardized and com…

Emerging Opportunities of Using Large Language Models for Translation Between Drug Molecules and Indications

2024-02-14 · David Oniani, Jordan Hilsman, Chengxi Zang, Junmei Wang 외

A drug molecule is a substance that changes the organism's mental or physical state. Every approved drug has an indication, which refers to the therapeutic use of that drug for treating a particular medical condition. Wh…

Drug DiscoveryLanguage ModellingLarge Language ModelTranslation