paper-with-me

홈 › Papers

SynRL: Aligning Synthetic Clinical Trial Data with Human-preferred Clinical Endpoints Using Reinforcement Learning

2024-11-11 · Trisha Das, Zifeng Wang, Afrah Shafquat, Mandis Beigi, Jason Mezey, Jimeng Sun

Each year, hundreds of clinical trials are conducted to evaluate new medical interventions, but sharing patient records from these trials with other institutions can be challenging due to privacy concerns and federal regulations. To help mitigate privacy concerns, researchers have proposed methods for generating synthetic patient data. However, existing approaches for generating synthetic clinical trial data disregard the usage requirements of these data, including maintaining specific properties of clinical outcomes, and only use post hoc assessments that are not coupled with the data generation process. In this paper, we propose SynRL which leverages reinforcement learning to improve the performance of patient data generators by customizing the generated data to meet the user-specified requirements for synthetic data outcomes and endpoints. Our method includes a data value critic function to evaluate the quality of the generated data and uses reinforcement learning to align the data generator with the users' needs based on the critic's feedback. We performed experiments on four clinical trial datasets and demonstrated the advantages of SynRL in improving the quality of the generated synthetic data while keeping the privacy risks low. We also show that SynRL can be utilized as a general framework that can customize data generation of multiple types of synthetic data generators. Our code is available at https://anonymous.4open.science/r/SynRL-DB0F/.

📄 PDF Abstract BibTeX arXiv:2411.07317

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…
HOC 설명 없음

Similar Papers 제목 키워드 기반

Learning Transferable Temporal Primitives for Video Reasoning via Synthetic Videos

2026-03-18 · Songtao Jiang, Sibo Song, Chenyi Zhou, Yuan Wang 외 arxiv

The transition from image to video understanding requires vision-language models (VLMs) to shift from recognizing static patterns to reasoning over temporal dynamics such as motion trajectories, speed changes, and state …

Video Generation

Retrieval-Reasoning Large Language Model-based Synthetic Clinical Trial Generation

2024-10-16 · Zerui Xu, Fang Wu, Yuanyuan Zhang, Yue Zhao

Machine learning (ML) exhibits promise in the clinical domain. However, it is constrained by data scarcity and ethical considerations, as the generation of clinical trials presents significant challenges due to stringent…

Language ModelingLanguage ModellingLarge Language ModelRetrieval

TrialSynth: Generation of Synthetic Sequential Clinical Trial Data

2024-09-11 · Chufan Gao, Mandis Beigi, Afrah Shafquat, Jacob Aptekar 외

Analyzing data from past clinical trials is part of the ongoing effort to optimize the design, implementation, and execution of new clinical trials and more efficiently bring life-saving interventions to market. While th…

MatchMiner-AI: An Open-Source Solution for Cancer Clinical Trial Matching

2024-12-23 · Ethan Cerami, Pavel Trukhanov, Morgan A. Paul, Michael J. Hassett 외

Clinical trials drive improvements in cancer treatments and outcomes. However, most adults with cancer do not participate in trials, and trials often fail to enroll enough patients to answer their scientific questions. A…

AD-CDO: A Lightweight Ontology for Representing Eligibility Criteria in Alzheimer's Disease Clinical Trials

2025-11-20 · Zenan Sun, Rashmie Abeysinghe, Xiaojin Li, Xinyue Hu 외 arxiv

Objective This study introduces the Alzheimer's Disease Common Data Element Ontology for Clinical Trials (AD-CDO), a lightweight, semantically enriched ontology designed to represent and standardize key eligibility crite…