Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Cross-lingual Unified Medical Language System entity linking in online health communities

View through CrossRef
Abstract Objective In Hebrew online health communities, participants commonly write medical terms that appear as transliterated forms of a source term in English. Such transliterations introduce high variability in text and challenge text-analytics methods. To reduce their variability, medical terms must be normalized, such as linking them to Unified Medical Language System (UMLS) concepts. We present a method to identify both transliterated and translated Hebrew medical terms and link them with UMLS entities. Materials and Methods We investigate the effect of linking terms in Camoni, a popular Israeli online health community in Hebrew. Our method, MDTEL (Medical Deep Transliteration Entity Linking), includes (1) an attention-based recurrent neural network encoder-decoder to transliterate words and mapping UMLS from English to Hebrew, (2) an unsupervised method for creating a transliteration dataset in any language without manually labeled data, and (3) an efficient way to identify and link medical entities in the Hebrew corpus to UMLS concepts, by producing a high-recall list of candidate medical terms in the corpus, and then filtering the candidates to relevant medical terms. Results We carry out experiments on 3 disease-specific communities: diabetes, multiple sclerosis, and depression. MDTEL tagging and normalizing on Camoni posts achieved 99% accuracy, 92% recall, and 87% precision. When tagging and normalizing terms in queries from the Camoni search logs, UMLS-normalized queries improved search results in 46% of the cases. Conclusions Cross-lingual UMLS entity linking from Hebrew is possible and improves search performance across communities. Annotated datasets, annotation guidelines, and code are made available online (https://github.com/yonatanbitton/mdtel).
Title: Cross-lingual Unified Medical Language System entity linking in online health communities
Description:
Abstract Objective In Hebrew online health communities, participants commonly write medical terms that appear as transliterated forms of a source term in English.
Such transliterations introduce high variability in text and challenge text-analytics methods.
To reduce their variability, medical terms must be normalized, such as linking them to Unified Medical Language System (UMLS) concepts.
We present a method to identify both transliterated and translated Hebrew medical terms and link them with UMLS entities.
Materials and Methods We investigate the effect of linking terms in Camoni, a popular Israeli online health community in Hebrew.
Our method, MDTEL (Medical Deep Transliteration Entity Linking), includes (1) an attention-based recurrent neural network encoder-decoder to transliterate words and mapping UMLS from English to Hebrew, (2) an unsupervised method for creating a transliteration dataset in any language without manually labeled data, and (3) an efficient way to identify and link medical entities in the Hebrew corpus to UMLS concepts, by producing a high-recall list of candidate medical terms in the corpus, and then filtering the candidates to relevant medical terms.
Results We carry out experiments on 3 disease-specific communities: diabetes, multiple sclerosis, and depression.
MDTEL tagging and normalizing on Camoni posts achieved 99% accuracy, 92% recall, and 87% precision.
When tagging and normalizing terms in queries from the Camoni search logs, UMLS-normalized queries improved search results in 46% of the cases.
Conclusions Cross-lingual UMLS entity linking from Hebrew is possible and improves search performance across communities.
Annotated datasets, annotation guidelines, and code are made available online (https://github.
com/yonatanbitton/mdtel).

Related Results

Hubungan Perilaku Pola Makan dengan Kejadian Anak Obesitas
Hubungan Perilaku Pola Makan dengan Kejadian Anak Obesitas
<p><em><span style="font-size: 11.0pt; font-family: 'Times New Roman',serif; mso-fareast-font-family: 'Times New Roman'; mso-ansi-language: EN-US; mso-fareast-langua...
Burden of the Beast
Burden of the Beast
Introduction Throughout the COVID-19 pandemic, and its fluctuating waves of infections and the emergence of new variants, Indigenous populations in Australia and worldwide have re...
Effects of Age and Gender during Three Lingual Tasks on Peak Lingual Pressures in Healthy Adults
Effects of Age and Gender during Three Lingual Tasks on Peak Lingual Pressures in Healthy Adults
Purpose: This study examined the effects of age and gender during three intra-oral lingual tasks (elevation, protrusion, and depression) on peak lingual pressure in healthy adults....
Učinak poučavanja razrednomu jeziku u izobrazbi nastavnika njemačkoga
Učinak poučavanja razrednomu jeziku u izobrazbi nastavnika njemačkoga
The actual use of classroom language is principally limited to the classroom environment. As far as foreign language learning is concerned, the classroom often turns out to be the ...
PERILAKU SATUAN LINGUAL -(N)ING DALAM BAHASA JAWA (LINGUAL UNIT BEHAVIOR -(N)ING IN JAVANESE LANGUANGE)
PERILAKU SATUAN LINGUAL -(N)ING DALAM BAHASA JAWA (LINGUAL UNIT BEHAVIOR -(N)ING IN JAVANESE LANGUANGE)
Penelitian ini berjudul Perilaku Satuan Lingual (n)ing dalam Bahasa Jawa. Teori yang digunakan dalam kajian ini ialah kategori kata dan analisis konstituen. Pengumpulan data menggu...
Unsupervised Domain Adaptation of a Pretrained Cross-Lingual Language Model
Unsupervised Domain Adaptation of a Pretrained Cross-Lingual Language Model
Recent research indicates that pretraining cross-lingual language models on large-scale unlabeled texts yields significant performance improvements over various cross-lingual and l...

Back to Top