paper-with-me

홈 › Papers

Database of Mandarin Neighborhood Statistics

2016-05-01 · LREC 2016 5 · Karl Neergaard, Hongzhi Xu, Chu-Ren Huang

In the design of controlled experiments with language stimuli, researchers from psycholinguistic, neurolinguistic, and related fields, require language resources that isolate variables known to affect language processing. This article describes a freely available database that provides word level statistics for words and nonwords of Mandarin, Chinese. The featured lexical statistics include subtitle corpus frequency, phonological neighborhood density, neighborhood frequency, and homophone density. The accompanying word descriptors include pinyin, ascii phonetic transcription (sampa), lexical tone, syllable structure, dominant PoS, and syllable, segment and pinyin lengths for each phonological word. It is designed for researchers particularly concerned with language processing of isolated words and made to accommodate multiple existing hypotheses concerning the structure of the Mandarin syllable. The database is divided into multiple files according to the desired search criteria: 1) the syllable segmentation schema used to calculate density measures, and 2) whether the search is for words or nonwords. The database is open to the research community at https://github.com/karlneergaard/Mandarin-Neighborhood-Statistics.

📄 PDF Abstract BibTeX

Code (1)

karlneergaard/Mandarin-Neighborhood-Statistics 공식 구현

Tasks

POS

Similar Papers 제목 키워드 기반

以多層感知器辨識情緒於國台客語料庫 (Use Multilayer Perceptron To Recognize Emotion in Mandarin,Taiwanese and Hakka Database) [In Chinese]

2016-10-01 · ROCLINGIJCLCLP 2016 10 · Chia-Hsien Chan, Chia-Ping Chen

Introducing MELI: the Mandarin-English Language Interview Corpus

2026-03-27 · Suyuan Liu, Molly Babel arxiv

We introduce the Mandarin-English Language Interview (MELI) Corpus, an open-source resource of 29.8 hours of speech from 51 Mandarin-English bilingual speakers. MELI combines matched sessions in Mandarin and English with…

A Generative Model of Natural Texture Surrogates

2015-05-28 · Niklas Ludtke, Debapriya Das, Lucas Theis, Matthias Bethge

Natural images can be viewed as patchworks of different textures, where the local image statistics is roughly stationary within a small neighborhood but otherwise varies from region to region. In order to model this vari…

Image Compressionmodel

A Multilingual Natural Stress Emotion Database

2012-05-01 · LREC 2012 5 · Xin Zuo, Tian Li, Pascale Fung

In this paper, we describe an ongoing effort in collecting and annotating a multilingual speech database of natural stress emotion from university students. The goal is to detect natural stress emotions and study the str…

Emotion RecognitionSpeech Synthesis

End-to-End Mandarin Tone Classification with Short Term Context Information

2021-04-12 · Jiyang Tang, Ming Li

In this paper, we propose an end-to-end Mandarin tone classification method from continuous speech utterances utilizing both the spectrogram and the short-term context information as the input. Both spectrograms and cont…

General Classification