paper-with-me

Papers

MYCanCor: A Video Corpus of spoken Malaysian Cantonese

2018-05-01 · LREC 2018 5 · Andreas Liesenfeld
📄 PDF Abstract BibTeX

Code (1)

liesenf/mycancor 공식 구현

Similar Papers 제목 키워드 기반

CantoMap: a Hong Kong Cantonese MapTask Corpus

2020-05-01 · LREC 2020 5 · Gr{\'e}goire Winterstein, Carmen Tang, Regine Lai

This work reports on the construction of a corpus of connected spoken Hong Kong Cantonese. The corpus aims at providing an additional resource for the study of modern (Hong Kong) Cantonese and also involves several contr…

HK-LegiCoST: Leveraging Non-Verbatim Transcripts for Speech Translation

2023-06-20 · Cihan Xiao, Henry Li Xinyuan, Jinyi Yang, Dongji Gao 외

We introduce HK-LegiCoST, a new three-way parallel corpus of Cantonese-English translations, containing 600+ hours of Cantonese audio, its standard traditional Chinese transcript, and English translation, segmented and a…

Cross-corpusSentencespeech-recognitionSpeech Recognition+1

Cifu: a Frequency Lexicon of Hong Kong Cantonese

2020-05-01 · LREC 2020 5 · Regine Lai, Gr{\'e}goire Winterstein

This paper introduces Cifu, a lexical database for Hong Kong Cantonese (HKC) that offers phonological and orthographic information, frequency measures, and lexical neighborhood information for lexical items in HKC. Cifu …

Diversity

WeCanTalk: A New Multi-language, Multi-modal Resource for Speaker Recognition

2022-06-01 · LREC 2022 6 · Karen Jones, Kevin Walker, Christopher Caruso, Jonathan Wright 외

The WeCanTalk (WCT) Corpus is a new multi-language, multi-modal resource for speaker recognition. The corpus contains Cantonese, Mandarin and English telephony and video speech data from over 200 multilingual speakers lo…

Speaker Recognition

Enriching Linguistic Representation in the Cantonese Wordnet and Building the New Cantonese Wordnet Corpus

2022-06-01 · LREC 2022 6 · Ut Seong Sio, Luís Morgado da Costa

This paper reports on the most recent improvements on the Cantonese Wordnet, a wordnet project started in 2019 (Sio and Morgado da Costa, 2019) with the aim of capturing and organizing lexico-semantic information of Hong…