paper-with-me

홈 › Papers

Less for More: Enhanced Feedback-aligned Mixed LLMs for Molecule Caption Generation and Fine-Grained NLI Evaluation

2024-05-22 · Dimitris Gkoumas, Maria Liakata

Scientific language models drive research innovation but require extensive fine-tuning on large datasets. This work enhances such models by improving their inference and evaluation capabilities with minimal or no additional training. Focusing on molecule caption generation, we explore post-training synergies between alignment fine-tuning and model merging in a cross-modal setup. We reveal intriguing insights into the behaviour and suitability of such methods while significantly surpassing state-of-the-art models. Moreover, we propose a novel atomic-level evaluation method leveraging off-the-shelf Natural Language Inference (NLI) models for use in the unseen chemical domain. Our experiments demonstrate that our evaluation operates at the right level of granularity, effectively handling multiple content units and subsentence reasoning, while widely adopted NLI methods consistently misalign with assessment criteria.

📄 PDF Abstract BibTeX arXiv:2405.13984

Code (0)

등록된 구현이 없습니다.

Tasks

Caption GenerationHallucinationNatural Language Inferencescientific discoveryTranslation

Similar Papers 제목 키워드 기반

Statistical Impossibility and Possibility of Aligning LLMs with Human Preferences: From Condorcet Paradox to Nash Equilibrium

2025-03-14 · Kaizhao Liu, Qi Long, Zhekun Shi, Weijie J. Su 외

Aligning large language models (LLMs) with diverse human preferences is critical for ensuring fairness and informed outcomes when deploying these models for decision-making. In this paper, we seek to uncover fundamental …

Fairness

M3HF: Multi-agent Reinforcement Learning from Multi-phase Human Feedback of Mixed Quality

2025-03-03 · Ziyan Wang, Zhicheng Zhang, Fei Fang, Yali Du

Designing effective reward functions in multi-agent reinforcement learning (MARL) is a significant challenge, often leading to suboptimal or misaligned behaviors in complex, coordinated environments. We introduce Multi-a…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learning

Integrating AI for Enhanced Feedback in Translation Revision- A Mixed-Methods Investigation of Student Engagement

2024-10-11 · Simin Xu, Yanfang Su, Kanglong Liu

Despite the well-established importance of feedback in education, the application of Artificial Intelligence (AI)-generated feedback, particularly from language models like ChatGPT, remains understudied in translation ed…

Translation

Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators

2026-05-12 · Heejin Do, Shashank Sonkar, Mrinmaya Sachan arxiv

Large language models (LLMs) can fluently generate student-like responses, making them attractive as simulated students for training and evaluating AI tutors and human educators. Yet such simulators are typically evaluat…

Reinforcement Learning

Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment

2024-12-06 · Ran Tian, Yilin Wu, Chenfeng Xu, Masayoshi Tomizuka 외

Visuomotor robot policies, increasingly pre-trained on large-scale datasets, promise significant advancements across robotics domains. However, aligning these policies with end-user preferences remains a challenge, parti…