paper-with-me

홈 › Papers

Vietnamese Text Accent Restoration with Statistical Machine Translation

2013-11-01 · PACLIC 2013 11 · Luan-Nghia Pham, Viet-Hong Tran, Vinh-Van Nguyen
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModellingMachine TranslationTranslation

Similar Papers 제목 키워드 기반

A study of Vietnamese readability assessing through semantic and statistical features

2024-11-07 · Hung Tuan Le, Long Truong To, Manh Trong Nguyen, Quyen Nguyen 외

Determining the difficulty of a text involves assessing various textual features that may impact the reader's text comprehension, yet current research in Vietnamese has only focused on statistical features. This paper in…

Reading Comprehension

PhoWhisper: Automatic Speech Recognition for Vietnamese

2024-03-27 · Thanh-Thien Le, Linh The Nguyen, Dat Quoc Nguyen

We introduce PhoWhisper in five versions for Vietnamese automatic speech recognition. PhoWhisper's robustness is achieved through fine-tuning the Whisper model on an 844-hour dataset that encompasses diverse Vietnamese a…

Automatic Speech Recognitionspeech-recognitionSpeech Recognition

On the Use of Machine Translation-Based Approaches for Vietnamese Diacritic Restoration

2017-09-20 · Thai-Hoang Pham, Xuan-Khoai Pham, Phuong Le-Hong

This paper presents an empirical study of two machine translation-based approaches for Vietnamese diacritic restoration problem, including phrase-based and neural-based machine translation models. This is the first work …

Machine TranslationTranslationWord Embeddings

Accent Placement Models for Rigvedic Sanskrit Text

2025-11-28 · Akhil Rajeev P, Annarao Kulkarni arxiv

The Rigveda, among the oldest Indian texts in Vedic Sanskrit, employs a distinctive pitch-accent system : udātta, anudātta, svarita whose marks encode melodic and interpretive cues but are often absent from modern e-text…

parameter-efficient fine-tuning

VietMed: A Dataset and Benchmark for Automatic Speech Recognition of Vietnamese in the Medical Domain

2024-04-08 · Khai Le-Duc

Due to privacy restrictions, there's a shortage of publicly available speech recognition datasets in the medical domain. In this work, we present VietMed - a Vietnamese speech recognition dataset in the medical domain co…

Language ModellingSpeech RecognitionUnsupervised Pre-trainingVietnamese Speech Recognition