paper-with-me

홈 › Papers

Multilingualization of Medical Terminology: Semantic and Structural Embedding Approaches

2020-05-01 · LREC 2020 5 · Long-Huei Chen, Kyo Kageura

The multilingualization of terminology is an essential step in the translation pipeline, to ensure the correct transfer of domain-specific concepts. Many institutions and language service providers construct and maintain multilingual terminologies, which constitute important assets. However, the curation of such multilingual resources requires significant human effort; though automatic multilingual term extraction methods have been proposed so far, they are of limited success as term translation cannot be satisfied by simply conveying meaning, but requires the terminologists and domain experts{'} knowledge to fit the term within the existing terminology. Here we propose a method to encode the structural property of a term by aligning their embeddings using graph convolutional networks trained from separate languages. We observe that the structural information can augment the semantic methods also explored in this work, and recognize the unique nature of terminologies allows our method to fully take advantage and produce superior results.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Term ExtractionTranslation

Methods 이 논문이 사용한 방법론

Graph Convolutional Networks 설명 없음

Similar Papers 제목 키워드 기반

Can Embeddings Adequately Represent Medical Terminology? New Large-Scale Medical Term Similarity Datasets Have the Answer!

2020-03-24 · Claudia Schulz, Damir Juric

A large number of embeddings trained on medical data have emerged, but it remains unclear how well they represent medical terminology, in particular whether the close relationship of semantically similar medical terms is…

Enriching Medcial Terminology Knowledge Bases via Pre-trained Language Model and Graph Convolutional Network

2019-09-02 · Jiaying Zhang, Zhixing Zhang, Huanhuan Zhang, Zhiyuan Ma 외

Enriching existing medical terminology knowledge bases (KBs) is an important and never-ending work for clinical research because new terminology alias may be continually added and standard terminologies may be newly rena…

Language ModelingLanguage Modelling

MedNorm: A Corpus and Embeddings for Cross-terminology Medical Concept Normalisation

2019-08-01 · WS 2019 8 · Maksim Belousov, William G. Dixon, Goran Nenadic

The medical concept normalisation task aims to map textual descriptions to standard terminologies such as SNOMED-CT or MedDRA. Existing publicly available datasets annotated using different terminologies cannot be simply…

Representation Learning

TermGPT: Multi-Level Contrastive Fine-Tuning for Terminology Adaptation in Legal and Financial Domain

2025-11-13 · Yidan Sun, Mengying Zhu, Feiyue Chen, Yangyang Wu 외 arxiv

Large language models (LLMs) have demonstrated impressive performance in text generation tasks; however, their embedding spaces often suffer from the isotropy problem, resulting in poor discrimination of domain-specific …

Contrastive LearningText Generation

Leveraging knowledge graphs to update scientific word embeddings using latent semantic imputation

2022-10-27 · Jason Hoelscher-Obermaier, Edward Stevinson, Valentin Stauber, Ivaylo Zhelev 외

The most interesting words in scientific texts will often be novel or rare. This presents a challenge for scientific word embedding models to determine quality embedding vectors for useful terms that are infrequent or ne…

ImputationKnowledge GraphsWord Embeddings