Coherence Evaluation
2개 벤치마크 · 논문 34편 · 이 태스크의 논문 보기 →
Benchmarks
GCDC + RST - Accuracy
GCDC + RST - F1
Most implemented
Transformer Models for Text Coherence Assessment
ECoh: Turn-level Coherence Evaluation for Multilingual Dialogues
CoUDA: Coherence Evaluation via Unified Data Augmentation
Unlocking Structure Measuring: Introducing PDD, an Automatic Metric for Positional Discourse Coherence
The Problem of Coherence in Natural Language Explanations of Recommendations
Open-Domain Text Evaluation via Contrastive Distribution Methods
Papers
Watch Your Step: Information Injection in Diffusion Models via Shadow Timestep Embedding
Diffusion models have become the foundation of modern generative systems, with most research focusing primarily on improving generation efficiency and output quality. The timestep embedding component is a crucial part of…
Coherence EvaluationInteractive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition
Recent years have witnessed remarkable progress in automatic speech recognition (ASR), driven by advances in model architectures and large-scale training data. However, two important aspects remain underexplored. First, …
Coherence EvaluationSpeech RecognitionMME-CoF-Pro: Evaluating Reasoning Coherence in Video Generative Models with Text and Visual Hints
Video generative models show emerging reasoning behaviors. It is essential to ensure that generated events remain causally consistent across frames for reliable deployment, a property we define as reasoning coherence. To…
Coherence EvaluationJoint Modeling of Entities and Discourse Relations for Coherence Assessment
In linguistics, coherence can be achieved by different means, such as by maintaining reference to the same set of entities across sentences and by establishing discourse relations between them. However, most existing wor…
Coherence EvaluationLLMs and their Limited Theory of Mind: Evaluating Mental State Annotations in Situated Dialogue
What if large language models could not only infer human mindsets but also expose every blind spot in team dialogue such as discrepancies in the team members' joint understanding? We present a novel, two-step framework t…
Coherence EvaluationSpatial ReasoningSemantic-Augmented Latent Topic Modeling with LLM-in-the-Loop
Latent Dirichlet Allocation (LDA) is a prominent generative probabilistic model used for uncovering abstract topics within document collections. In this paper, we explore the effectiveness of augmenting topic models with…
Coherence EvaluationTopic Models