paper-with-me

Papers

Discovery of Discourse-Related Language Contrasts through Alignment Discrepancies in English-German Translation

2017-09-01 · WS 2017 9 · Ekaterina Lapshinova-Koltunski, Christian Hardmeier

In this paper, we analyse alignment discrepancies for discourse structures in English-German parallel data {--} sentence pairs, in which discourse structures in target or source texts have no alignment in the corresponding parallel sentences. The discourse-related structures are designed in form of linguistic patterns based on the information delivered by automatic part-of-speech and dependency annotation. In addition to alignment errors (existing structures left unaligned), these alignment discrepancies can be caused by language contrasts or through the phenomena of explicitation and implicitation in the translation process. We propose a new approach including new type of resources for corpus-based language contrast analysis and apply it to study and classify the contrasts found in our English-German parallel corpus. As unaligned discourse structures may also result in the loss of discourse information in the MT training data, we hope to deliver information in support of discourse-aware machine translation (MT).

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationSentenceTranslationWord Alignment

Similar Papers 제목 키워드 기반

Discourse-Related Language Contrasts in English-Croatian Human and Machine Translation

2018-10-01 · WS 2018 10 · Margita {\v{S}}o{\v{s}}tari{\'c}, Christian Hardmeier, Sara Stymne

We present an analysis of a number of coreference phenomena in English-Croatian human and machine translations. The aim is to shed light on the differences in the way these structurally different languages make use of di…

Machine TranslationNMTTranslationWord Alignment

The Pragmatic Persona: Discovering LLM Persona through Bridging Inference

2026-04-27 · Jisoo Yang, Jongwon Ryu, Minuk Ma, Trung X. Pham 외 arxiv

Large Language Models (LLMs) reveal inherent and distinctive personas through dialogue. However, most existing persona discovery approaches rely on surface-level lexical or stylistic cues, treating dialogue as a flat seq…

Knowledge Graphs

Discursive Circuits: How Do Language Models Understand Discourse Relations?

2025-10-13 · Yisong Miao, Min-Yen Kan arxiv

Which components in transformer language models are responsible for discourse understanding? We hypothesize that sparse computational graphs, termed as discursive circuits, control how models process discourse relations.…

DiscoPhon: Benchmarking the Unsupervised Discovery of Phoneme Inventories With Discrete Speech Units

2026-03-19 · Maxime Poli, Manel Khentout, Angelo Ortiz Tandazo, Ewan Dunbar 외 arxiv

We introduce DiscoPhon, a multilingual benchmark for evaluating unsupervised phoneme discovery from discrete speech units. DiscoPhon covers 6 dev and 6 test languages, chosen to span a wide range of phonemic contrasts. G…

Multilingual Extraction and Recognition of Implicit Discourse Relations in Speech and Text

2026-02-04 · Ahmed Ruby, Christian Hardmeier, Sara Stymne arxiv

Implicit discourse relation classification is a challenging task, as it requires inferring meaning from context. While contextual cues can be distributed across modalities and vary across languages, they are not always c…

Relation ClassificationCross-Lingual Transfer