paper-with-me

홈 › Papers

A Dataset for Pharmacovigilance in German, French, and Japanese: Annotating Adverse Drug Reactions across Languages

2024-03-27 · Lisa Raithel, Hui-Syuan Yeh, Shuntaro Yada, Cyril Grouin, Thomas Lavergne, Aurélie Névéol, Patrick Paroubek, Philippe Thomas, Tomohiro Nishiyama, Sebastian Möller, Eiji Aramaki, Yuji Matsumoto, Roland Roller, Pierre Zweigenbaum

User-generated data sources have gained significance in uncovering Adverse Drug Reactions (ADRs), with an increasing number of discussions occurring in the digital world. However, the existing clinical corpora predominantly revolve around scientific articles in English. This work presents a multilingual corpus of texts concerning ADRs gathered from diverse sources, including patient fora, social media, and clinical reports in German, French, and Japanese. Our corpus contains annotations covering 12 entity types, four attribute types, and 13 relation types. It contributes to the development of real-world multilingual language models for healthcare. We provide statistics to highlight certain challenges associated with the corpus and conduct preliminary experiments resulting in strong baselines for extracting entities and relations between these entities, both within and across languages.

📄 PDF Abstract BibTeX arXiv:2403.18336

Code (2)

dfki-nlp/keepha_annotation_guidelines 공식 구현
dotkat-dotcome/keepha-adr 공식 구현 pytorch

Tasks

ArticlesAttributePharmacovigilance

Similar Papers 제목 키워드 기반

An AMR parser for English, French, German, Spanish and Japanese and a new AMR-annotated corpus

2015-06-01 · NAACL 2015 6 · V, Lucy erwende, Arul Menezes, Chris Quirk

Annotating tense, mood and voice for English, French and German

2017-07-01 · ACL 2017 7 · Anita Ramm, Sharid Lo{\'a}iciga, Annemarie Friedrich, Alex Fraser 외

Towards an Automatic Classification of Illustrative Examples in a Large Japanese-French Dictionary Obtained by OCR

2018-08-01 · COLING 2018 8 · Christian Boitet, Mathieu Mangeot, Mutsuko Tomokiyo

We work on improving the Cesselin, a large and open source Japanese-French bilingual dictionary digitalized by OCR, available on the web, and contributively improvable online. Labelling its examples (about 226000) would …

General ClassificationMachine TranslationOptical Character Recognition (OCR)

Multilingual Relative Clause Attachment Ambiguity Resolution in Large Language Models

2025-03-04 · So Young Lee, Russell Scheinberg, Amber Shore, Ameeta Agrawal

This study examines how large language models (LLMs) resolve relative clause (RC) attachment ambiguities and compares their performance to human sentence processing. Focusing on two linguistic factors, namely the length …

Sentence

A Survey on non-English Question Answering Dataset

2021-12-27 · Andreas Chandra, Affandy Fahrizain, Ibrahim, Simon Willyanto Laufried

Research in question answering datasets and models has gained a lot of attention in the research community. Many of them release their own question answering datasets as well as the models. There is tremendous progress t…

Cross-Lingual Question AnsweringQuestion AnsweringSurvey