paper-with-me

Papers

MedReadMe: A Systematic Study for Fine-grained Sentence Readability in Medical Domain

2024-05-03 · Chao Jiang, Wei Xu

Medical texts are notoriously challenging to read. Properly measuring their readability is the first step towards making them more accessible. In this paper, we present a systematic study on fine-grained readability measurements in the medical domain at both sentence-level and span-level. We introduce a new dataset MedReadMe, which consists of manually annotated readability ratings and fine-grained complex span annotation for 4,520 sentences, featuring two novel "Google-Easy" and "Google-Hard" categories. It supports our quantitative analysis, which covers 650 linguistic features and automatic complex word and jargon identification. Enabled by our high-quality annotation, we benchmark and improve several state-of-the-art sentence-level readability metrics for the medical domain specifically, which include unsupervised, supervised, and prompting-based methods using recently developed large language models (LLMs). Informed by our fine-grained complex span annotation, we find that adding a single feature, capturing the number of jargon spans, into existing readability formulas can significantly improve their correlation with human judgments. The data is available at tinyurl.com/medreadme-repo

📄 PDF Abstract BibTeX arXiv:2405.02144

Code (0)

등록된 구현이 없습니다.

Tasks

Sentence

Similar Papers 제목 키워드 기반

FinAI-BERT: A Transformer-Based Model for Sentence-Level Detection of AI Disclosures in Financial Reports

2025-06-29 · Muhammad Bilal Zafar

The proliferation of artificial intelligence (AI) in financial services has prompted growing demand for tools that can systematically detect AI-related disclosures in corporate filings. While prior approaches often rely …

Sentence

arXivEdits: Understanding the Human Revision Process in Scientific Writing

2022-10-26 · Chao Jiang, Wei Xu, Samuel Stevens

Scientific publications are the primary means to communicate research discoveries, where the writing quality is of crucial importance. However, prior work studying the human editing process in this domain mainly focused …

intent-classificationIntent ClassificationSentence

Aligning Large Language Model Behavior with Human Citation Preferences

2026-02-05 · Kenichiro Ando, Tatsuya Harada arxiv

Most services built on powerful large-scale language models (LLMs) add citations to their output to enhance credibility. Recent research has paid increasing attention to the question of what reference documents to link t…

ABCD-LINK: Annotation Bootstrapping for Cross-Document Fine-Grained Links

2025-09-01 · Serwar Basch, Ilia Kuznetsov, Tom Hope, Iryna Gurevych arxiv

Understanding fine-grained links between documents is crucial for many applications, yet progress is limited by the lack of efficient methods for data curation. To address this limitation, we introduce a domain-agnostic …

Weakly-supervised Domain Adaption for Aspect Extraction via Multi-level Interaction Transfer

2020-06-16 · Tao Liang, Wenya Wang, Fengmao Lv

Fine-grained aspect extraction is an essential sub-task in aspect based opinion analysis. It aims to identify the aspect terms (a.k.a. opinion targets) of a product or service in each sentence. However, expensive annotat…

Aspect ExtractionDomain AdaptationSentence