Vietnamese Text Accent Restoration with Statistical Machine Translation
Code (0)
등록된 구현이 없습니다.
Tasks
Language ModellingMachine TranslationTranslationSimilar Papers 제목 키워드 기반
A study of Vietnamese readability assessing through semantic and statistical features
Determining the difficulty of a text involves assessing various textual features that may impact the reader's text comprehension, yet current research in Vietnamese has only focused on statistical features. This paper in…
Reading ComprehensionPhoWhisper: Automatic Speech Recognition for Vietnamese
We introduce PhoWhisper in five versions for Vietnamese automatic speech recognition. PhoWhisper's robustness is achieved through fine-tuning the Whisper model on an 844-hour dataset that encompasses diverse Vietnamese a…
Automatic Speech Recognitionspeech-recognitionSpeech RecognitionOn the Use of Machine Translation-Based Approaches for Vietnamese Diacritic Restoration
This paper presents an empirical study of two machine translation-based approaches for Vietnamese diacritic restoration problem, including phrase-based and neural-based machine translation models. This is the first work …
Machine TranslationTranslationWord EmbeddingsAccent Placement Models for Rigvedic Sanskrit Text
The Rigveda, among the oldest Indian texts in Vedic Sanskrit, employs a distinctive pitch-accent system : udātta, anudātta, svarita whose marks encode melodic and interpretive cues but are often absent from modern e-text…
parameter-efficient fine-tuningVietMed: A Dataset and Benchmark for Automatic Speech Recognition of Vietnamese in the Medical Domain
Due to privacy restrictions, there's a shortage of publicly available speech recognition datasets in the medical domain. In this work, we present VietMed - a Vietnamese speech recognition dataset in the medical domain co…
Language ModellingSpeech RecognitionUnsupervised Pre-trainingVietnamese Speech Recognition