CUED at ProbSum 2023: Hierarchical Ensemble of Summarization Models
In this paper, we consider the challenge of summarizing patients' medical progress notes in a limited data setting. For the Problem List Summarization (shared task 1A) at the BioNLP Workshop 2023, we demonstrate that Clinical-T5 fine-tuned to 765 medical clinic notes outperforms other extractive, abstractive and zero-shot baselines, yielding reasonable baseline systems for medical note summarization. Further, we introduce Hierarchical Ensemble of Summarization Models (HESM), consisting of token-level ensembles of diverse fine-tuned Clinical-T5 models, followed by Minimum Bayes Risk (MBR) decoding. Our HESM approach lead to a considerable summarization performance boost, and when evaluated on held-out challenge data achieved a ROUGE-L of 32.77, which was the best-performing system at the top of the shared task leaderboard.
Code (1)
Similar Papers 제목 키워드 기반
Overview of the Problem List Summarization (ProbSum) 2023 Shared Task on Summarizing Patients' Active Diagnoses and Problems from Electronic Health Record Progress Notes
The BioNLP Workshop 2023 initiated the launch of a shared task on Problem List Summarization (ProbSum) in January 2023. The aim of this shared task is to attract future research efforts in building NLP models for real-wo…
Decision MakingDiagnosticCUED_speech at TREC 2020 Podcast Summarisation Track
In this paper, we describe our approach for the Podcast Summarisation challenge in TREC 2020. Given a podcast episode with its transcription, the goal is to generate a summary that captures the most important information…
Mapping de l'espace spectral vers l'espace visuel de la parole : les voyelles du fran\ccais en langue fran\ccaise parl\'ee compl\'et\'ee (Mapping of the spectral space to the visual speech space for French vowels cued in Cued Speech) [in French]
Cued-Agent: A Collaborative Multi-Agent System for Automatic Cued Speech Recognition
Cued Speech (CS) is a visual communication system that combines lip-reading with hand coding to facilitate communication for individuals with hearing impairments. Automatic CS Recognition (ACSR) aims to convert CS hand g…
Speech RecognitionCued Speech Generation Leveraging a Pre-trained Audiovisual Text-to-Speech Model
This paper presents a novel approach for the automatic generation of Cued Speech (ACSG), a visual communication system used by people with hearing impairment to better elicit the spoken language. We explore transfer lear…
text-to-speechText to SpeechTransfer Learning