paper-with-me

Coherence Evaluation

2개 벤치마크 · 논문 34편 · 이 태스크의 논문 보기 →

Benchmarks

GCDC + RST - Accuracy

결과 8개

GCDC + RST - F1

결과 6개

Most implemented

Papers

Watch Your Step: Information Injection in Diffusion Models via Shadow Timestep Embedding

2026-05-01 · An Huang, Junggab Son, Zuobin Xiong arxiv

Diffusion models have become the foundation of modern generative systems, with most research focusing primarily on improving generation efficiency and output quality. The timestep embedding component is a crucial part of…

Coherence Evaluation

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition

2026-04-10 · Peng Wang, Yanqiao Zhu, Zixuan Jiang, Qinyuan Chen 외 arxiv

Recent years have witnessed remarkable progress in automatic speech recognition (ASR), driven by advances in model architectures and large-scale training data. However, two important aspects remain underexplored. First, …

Coherence EvaluationSpeech Recognition

MME-CoF-Pro: Evaluating Reasoning Coherence in Video Generative Models with Text and Visual Hints

2026-03-20 · Yu Qi, Xinyi Xu, Ziyu Guo, Siyuan Ma 외 arxiv

Video generative models show emerging reasoning behaviors. It is essential to ensure that generated events remain causally consistent across frames for reliable deployment, a property we define as reasoning coherence. To…

Coherence Evaluation

Joint Modeling of Entities and Discourse Relations for Coherence Assessment

2025-09-04 · Wei Liu, Michael Strube arxiv

In linguistics, coherence can be achieved by different means, such as by maintaining reference to the same set of entities across sentences and by establishing discourse relations between them. However, most existing wor…

Coherence Evaluation

LLMs and their Limited Theory of Mind: Evaluating Mental State Annotations in Situated Dialogue

2025-09-02 · Katharine Kowalyshyn, Matthias Scheutz arxiv

What if large language models could not only infer human mindsets but also expose every blind spot in team dialogue such as discrepancies in the team members' joint understanding? We present a novel, two-step framework t…

Coherence EvaluationSpatial Reasoning

Semantic-Augmented Latent Topic Modeling with LLM-in-the-Loop

2025-07-11 · Mengze Hong, Chen Jason Zhang, Di Jiang arxiv

Latent Dirichlet Allocation (LDA) is a prominent generative probabilistic model used for uncovering abstract topics within document collections. In this paper, we explore the effectiveness of augmenting topic models with…

Coherence EvaluationTopic Models

전체 34편 보기 →