paper-with-me

Papers

Method of Tibetan Person Knowledge Extraction

2016-04-11 · Yuan Sun, Zhen Zhu

Person knowledge extraction is the foundation of the Tibetan knowledge graph construction, which provides support for Tibetan question answering system, information retrieval, information extraction and other researches, and promotes national unity and social stability. This paper proposes a SVM and template based approach to Tibetan person knowledge extraction. Through constructing the training corpus, we build the templates based the shallow parsing analysis of Tibetan syntactic, semantic features and verbs. Using the training corpus, we design a hierarchical SVM classifier to realize the entity knowledge extraction. Finally, experimental results prove the method has greater improvement in Tibetan person knowledge extraction.

📄 PDF Abstract BibTeX arXiv:1604.02843

Code (0)

등록된 구현이 없습니다.

Tasks

graph constructionInformation RetrievalQuestion AnsweringRetrievalUnity

Methods 이 논문이 사용한 방법론

SVM A Support Vector Machine, or SVM, is a non-parametric supervised learning model. For non-linear classification and regression, they utilise the kernel trick to map inputs…

Similar Papers 제목 키워드 기반

Sun-Shine: A Large Language Model for Tibetan Culture

2025-03-24 · Cheng Huang, Fan Gao, Nyima Tashi, Yutong Liu 외

Tibetan, a minority language in China, features a highly intricate grammatical structure, characterized by four verb tenses and a tense system with frequent irregularities, contributing to its extensive inflectional dive…

Language ModelingLanguage ModellingLarge Language ModelMachine Translation+2

Tibetan-TTS:Low-Resource Tibetan Speech Synthesis with Large Model Adaptation

2026-05-04 · Jiaxu He, Chao Wang, Jie Lian, Yuqing Cai 외 arxiv

Tibetan text-to-speech (TTS) has long been challenged by scarce speech resources, significant dialectal variation, and the complex mapping between written text and spoken pronunciation. To address these issues, this work…

Speech Synthesis

TreeProbe : A Tibetan Medicine Benchmark for Cultural Bias in LLMs

2026-08-01 · Jin Zhang, Linyu Li, Weili Jiang, Yuqing Cai 외 arxiv

Large language models are increasingly viewed as a potential means of mitigating global health inequities, yet their outputs often reflect dominant high-resource medical traditions and provide limited coverage of traditi…

Developing the Old Tibetan Treebank

2019-09-01 · RANLP 2019 9 · Christian Faggionato, Marieke Meelen

This paper presents a full procedure for the development of a segmented, POS-tagged and chunkparsed corpus of Old Tibetan. As an extremely low-resource language, Old Tibetan poses non-trivial problems in every step towar…

POS

A Aelf-supervised Tibetan-chinese Vocabulary Alignment Method Based On Adversarial Learning

2021-10-04 · Enshuai Hou, Jie Zhu

Tibetan is a low-resource language. In order to alleviate the shortage of parallel corpus between Tibetan and Chinese, this paper uses two monolingual corpora and a small number of seed dictionaries to learn the semi-sup…