paper-with-me

Papers

Optimizing Life Sciences Agents in Real-Time using Reinforcement Learning

2025-11-26 · Nihir Chadderwala arxiv

Generative AI agents in life sciences face a critical challenge: determining the optimal approach for diverse queries ranging from simple factoid questions to complex mechanistic reasoning. Traditional methods rely on fixed rules or expensive labeled training data, neither of which adapts to changing conditions or user preferences. We present a novel framework that combines AWS Strands Agents with Thompson Sampling contextual bandits to enable AI agents to learn optimal decision-making strategies from user feedback alone. Our system optimizes three key dimensions: generation strategy selection (direct vs. chain-of-thought), tool selection (literature search, drug databases, etc.), and domain routing (pharmacology, molecular biology, clinical specialists). Through empirical evaluation on life science queries, we demonstrate 15-30\% improvement in user satisfaction compared to random baselines, with clear learning patterns emerging after 20-30 queries. Our approach requires no ground truth labels, adapts continuously to user preferences, and provides a principled solution to the exploration-exploitation dilemma in agentic AI systems.

📄 PDF Abstract BibTeX arXiv:2512.03065

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

EMBL AI Librarian: Life-Sciences Knowledge Layer for AI Agents

2026-07-30 · Luigi Sigillo, Matteo Silvestri, Francesco Tabaro, Rajat Bhatnagar 외 arxiv

The web is increasingly accessed by AI agents rather than humans. Every agent needs knowledge, especially in the life-sciences, where agentic pipelines are growing fast. Access to the literature is a crucial part of that…

Open-Domain Question Answering

Attention Actor-Critic algorithm for Multi-Agent Constrained Co-operative Reinforcement Learning

2021-01-07 · P. Parnika, Raghuram Bharadwaj Diddigi, Sai Koti Reddy Danda, Shalabh Bhatnagar

In this work, we consider the problem of computing optimal actions for Reinforcement Learning (RL) agents in a co-operative setting, where the objective is to optimize a common goal. However, in many real-life applicatio…

reinforcement-learningReinforcement Learning (RL)

Computing in the Life Sciences: From Early Algorithms to Modern AI

2024-06-17 · Samuel A. Donkor, Matthew E. Walsh, Alexander J. Titus

Computing in the life sciences has undergone a transformative evolution, from early computational models in the 1950s to the applications of artificial intelligence (AI) and machine learning (ML) seen today. This paper h…

Decision Making

A Primal-Dual Solver for Large-Scale Tracking-by-Assignment

2020-04-14 · Stefan Haller, Mangal Prakash, Lisa Hutschenreiter, Tobias Pietzsch 외

We propose a fast approximate solver for the combinatorial problem known as tracking-by-assignment, which we apply to cell tracking. The latter plays a key role in discovery in many life sciences, especially in cell and …

Cell Tracking

Living Lab Evaluation for Life and Social Sciences Search Platforms -- LiLAS at CLEF 2021

2023-10-05 · Philipp Schaer, Johann Schaible, Leyla Jael Castro

Meta-evaluation studies of system performances in controlled offline evaluation campaigns, like TREC and CLEF, show a need for innovation in evaluating IR-systems. The field of academic search is no exception to this. Th…

Retrieval