paper-with-me

홈 › Papers

Fixed Suffix Dependency Ratio: Quantifying the Dual-Track Mechanism of Gender Assignment in Latvian Loanwords

2026-09-03 · Yelingyun Zhang, Atis Kapenieks arxiv

Existing research has repeatedly observed the tendency for English loanwords to cluster in the masculine gender across different recipient languages, yet the origin of this pattern remains difficult to determine, as fixed morphological rules and default assignments are frequently analysed together. This study proposes the Fixed Suffix Dependency Ratio (FSDR) to quantify the degree of reliance on fixed derivational suffixes across different genders, and to distinguish between morphological anchoring and free-choice in distribution. By examining 1,832 Latvian noun lemma types, the results reveal a significant FSDR asymmetry within the loanword system: feminine loanwords rely significantly more on fixed derivational suffixes, while masculine loanwords are more concentrated in the free-choice zone. This pattern exhibits loanword specificity and has become more pronounced in contemporary usage. FSDR therefore provides a quantitative framework for testing default gender and shows how masculine default can be activated and reinforced under language contact.

📄 PDF Abstract BibTeX arXiv:2609.03930

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A non-projective greedy dependency parser with bidirectional LSTMs

2017-07-11 · CONLL 2017 8 · David Vilares, Carlos Gómez-Rodríguez

The LyS-FASTPARSE team presents BIST-COVINGTON, a neural implementation of the Covington (2001) algorithm for non-projective dependency parsing. The bidirectional LSTM approach by Kipperwasser and Goldberg (2016) is used…

Dependency ParsingPOSPOS Tagging

Syst\`eme de pr\'ediction de n\'eologismes formels : le cas des N suffix\'es par -IER d\'enotant des artefacts (Prediction Device of Formal Neologisms : the Case of -IER Suffixed Nouns Denoting Artifacts) [in French]

2012-06-01 · JEPTALNRECITAL 2012 6 · Aur{\'e}lie Merlo

Estimating near-verbatim extraction risk in language models with decoding-constrained beam search

2026-03-26 · A. Feder Cooper, Mark A. Lemley, Christopher De Sa, Lea Duesterwald 외 arxiv

Recent work shows that standard greedy-decoding extraction methods for quantifying memorization in LLMs miss how extraction risk varies across sequences. Probabilistic extraction -- computing the probability of generatin…

DPad: Efficient Diffusion Language Models with Suffix Dropout

2025-08-19 · Xinhua Chen, Sitao Huang, Cong Guo, Chiyue Wei 외 arxiv

Diffusion-based Large Language Models (dLLMs) parallelize text generation by framing decoding as a denoising process, but suffer from high computational overhead since they predict all future suffix tokens at each step w…

Text Generation

Unsupervised Adverbial Identification in Modern Chinese Literature

2021-11-01 · EMNLP (LaTeCHCLfL, CLFL, LaTeCH) 2021 11 · Wenxiu Xie, John Lee, Fangqiong Zhan, Xiao Han 외

In many languages, adverbials can be derived from words of various parts-of-speech. In Chinese, the derivation may be marked either with the standard adverbial marker DI, or the non-standard marker DE. Since DE also serv…

Language ModelingLanguage Modelling