paper-with-me

홈 › Papers

MedCL-Bench: Benchmarking stability-efficiency trade-offs and scaling in biomedical continual learning

2026-03-17 · Min Zeng, Shuang Zhou, Zaifu Zhan, Rui Zhang arxiv

Medical language models must be updated as evidence and terminology evolve, yet sequential updating can trigger catastrophic forgetting. Although biomedical NLP has many static benchmarks, no unified, task-diverse benchmark exists for evaluating continual learning under standardized protocols, robustness to task order and compute-aware reporting. We introduce MedCL-Bench, which streams ten biomedical NLP datasets spanning five task families and evaluates eleven continual learning strategies across eight task orders, reporting retention, transfer, and GPU-hour cost. Across backbones and task orders, direct sequential fine-tuning on incoming tasks induces catastrophic forgetting, causing update-induced performance regressions on prior tasks. Continual learning methods occupy distinct retention-compute frontiers: parameter-isolation provides the best retention per GPU-hour, replay offers strong protection at higher cost, and regularization yields limited benefit. Forgetting is task-dependent, with multi-label topic classification most vulnerable and constrained-output tasks more robust. MedCL-Bench provides a reproducible framework for auditing model updates before deployment.

📄 PDF Abstract BibTeX arXiv:2603.16738

Code (0)

등록된 구현이 없습니다.

Tasks

Continual Learning

Similar Papers 제목 키워드 기반

Real-Time Performance Benchmarking of TinyML Models in Embedded Systems (PICO: Performance of Inference, CPU, and Operations)

2025-09-05 · Abhishek Dey, Saurabh Srivastava, Gaurav Singh, Robert G. Pettit arxiv

This paper presents PICO-TINYML-BENCHMARK, a modular and platform-agnostic framework for benchmarking the real-time performance of TinyML models on resource-constrained embedded systems. Evaluating key metrics such as in…

Keyword Spotting

Fast Benchmarking of Accuracy vs. Training Time with Cyclic Learning Rates

2022-06-02 · Jacob Portes, Davis Blalock, Cory Stephenson, Jonathan Frankle

Benchmarking the tradeoff between neural network accuracy and training time is computationally expensive. Here we show how a multiplicative cyclic learning rate schedule can be used to construct a tradeoff curve in a sin…

Benchmarking

Benchmarking Robustness of Contrastive Learning Models for Medical Image-Report Retrieval

2025-01-15 · Demetrio Deanda, Yuktha Priya Masupalli, Jeong Yang, Young Lee 외

Medical images and reports offer invaluable insights into patient health. The heterogeneity and complexity of these data hinder effective analysis. To bridge this gap, we investigate contrastive learning models for cross…

BenchmarkingContrastive LearningRetrieval

Does CLIP Benefit Visual Question Answering in the Medical Domain as Much as it Does in the General Domain?

2021-12-27 · Sedigheh Eslami, Gerard de Melo, Christoph Meinel

Contrastive Language--Image Pre-training (CLIP) has shown remarkable success in learning with cross-modal supervision from extensive amounts of image--text pairs collected online. Thus far, the effectiveness of CLIP has …

ArticlesMedical Visual Question AnsweringMeta-LearningQuestion Answering+3

MedCLIPSeg: Probabilistic Vision-Language Adaptation for Data-Efficient and Generalizable Medical Image Segmentation

2026-02-23 · Taha Koleilat, Hojat Asgariandehkordi, Omid Nejati Manzari, Berardino Barile 외 arxiv

Medical image segmentation remains challenging due to limited annotations for training, ambiguous anatomical features, and domain shifts. While vision-language models such as CLIP offer strong cross-modal representations…

Medical Image Segmentation