paper-with-me

홈 › Papers

Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Using a Large Language Model

2026-04-13 · Kihyuk Lee arxiv

Background: Large language models (LLMs) have been explored as tools for generating personalized exercise prescriptions, yet the consistency of outputs under identical conditions remains insufficiently examined. Objective: This study evaluated the intra-model consistency of LLM-generated exercise prescriptions using a repeated generation design. Methods: Six clinical scenarios were used to generate exercise prescriptions using Gemini 2.5 Flash (20 outputs per scenario; total n = 120). Consistency was assessed across three dimensions: (1) semantic consistency using SBERT-based cosine similarity, (2) structural consistency based on the FITT principle using an AI-as-a-judge approach, and (3) safety expression consistency, including inclusion rates and sentence-level quantification. Results: Semantic similarity was high across scenarios (mean cosine similarity: 0.879-0.939), with greater consistency in clinically constrained cases. Frequency showed consistent patterns, whereas variability was observed in quantitative components, particularly exercise intensity. Unclassifiable intensity expressions were observed in 10-25% of resistance training outputs. Safety-related expressions were included in 100% of outputs; however, safety sentence counts varied significantly across scenarios (H=86.18, p less than 0.001), with clinical cases generating more safety expressions than healthy adult cases. Conclusions: LLM-generated exercise prescriptions demonstrated high semantic consistency but showed variability in key quantitative components. Reliability depends substantially on prompt structure, and additional structural constraints and expert validation are needed before clinical deployment.

📄 PDF Abstract BibTeX arXiv:2604.11287

Code (0)

등록된 구현이 없습니다.

Tasks

Semantic Similarity

Similar Papers 제목 키워드 기반

Cross-Model Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Across Three Large Language Models

2026-04-21 · Kihyuk Lee arxiv

This study compared repeated generation consistency of exercise prescription outputs across three large language models (LLMs), specifically GPT-4.1, Claude Sonnet 4.6, and Gemini 2.5 Flash, under temperature=0 condition…

Semantic Similarity

Clinician-Directed Large Language Model Software Generation for Therapeutic Interventions in Physical Rehabilitation

2025-11-23 · Edward Kim, Yuri Cho, Jose Eduardo E. Lima, Julie Muccini 외 arxiv

Digital health interventions increasingly deliver home exercise programs via sensor-equipped devices such as smartphones, enabling remote monitoring of adherence and performance. However, current software is usually auth…

Dynamic Difficulty Adjustment in Virtual Reality Exergames through Experience-driven Procedural Content Generation

2021-08-19 · Tobias Huber, Silvan Mertes, Stanislava Rangelova, Simon Flutura 외

Virtual Reality (VR) games that feature physical activities have been shown to increase players' motivation to do physical exercise. However, for such exercises to have a positive healthcare effect, they have to be repea…

Deep Reinforcement Learning

DensityKV: Density-Guided KV Cache Compression for Long Video Generation

2026-08-28 · Wenqu Zhao, Xuemin Chi, Xin Zhang, Guoqing Ma 외 arxiv

Autoregressive video diffusion models enable streaming generation through sliding-window attention, but each generated block is conditioned on previously generated content, causing appearance and motion errors to propaga…

Video Generation

Generating Medical Prescriptions with Conditional Transformer

2023-10-30 · Samuel Belkadi, Nicolo Micheletti, Lifeng Han, Warren Del-Pinto 외

Access to real-world medication prescriptions is essential for medical research and healthcare quality improvement. However, access to real medication prescriptions is often limited due to the sensitive nature of the inf…

2kLanguage Modellingnamed-entity-recognitionNamed Entity Recognition+2