paper-with-me

Papers

Boosting coherence of language models

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Naturality of long-term information structure -- coherence -- remains a challenge in language generation. Large language models have insufficiently learned such structure, as their long-form generations differ from natural text in measures of coherence. To alleviate this divergence, we propose coherence boosting, an inference procedure that increases the effect of distant context on next-token prediction. We show the benefits of coherence boosting with pretrained models by distributional analyses of generated ordinary text and dialog responses. We also find that coherence boosting with state-of-the-art models for various zero-shot NLP tasks yields performance gains with no additional training.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Similar Papers 제목 키워드 기반

Coherence boosting: When your pretrained language model is not paying enough attention

2021-10-15 · ACL 2022 5 · Nikolay Malkin, Zhen Wang, Nebojsa Jojic

Long-range semantic coherence remains a challenge in automatic language generation and understanding. We demonstrate that large language models have insufficiently learned the effect of distant words on next-token predic…

Language ModelingLanguage ModellingText Generation

Improving the Generalization Ability in Essay Coherence Evaluation through Monotonic Constraints

2023-07-25 · Chen Zheng, huan zhang, Yan Zhao, Yuxuan Lai

Coherence is a crucial aspect of evaluating text readability and can be assessed through two primary factors when evaluating an essay in a scoring scenario. The first factor is logical coherence, characterized by the app…

Coherence EvaluationregressionSentence

Leveraging Weak Cross-Modal Guidance for Coherence Modelling via Iterative Learning

2024-08-01 · Yi Bin, Junrong Liao, Yujuan Ding, Haoxuan Li 외

Cross-modal coherence modeling is essential for intelligent systems to help them organize and structure information, thereby understanding and creating content of the physical world coherently like human-beings. Previous…

EEG2Vision: A Multimodal EEG-Based Framework for 2D Visual Reconstruction in Cognitive Neuroscience

2026-04-09 · Emanuele Balloni, Emanuele Frontoni, Chiara Matti, Marina Paolanti 외 arxiv

Reconstructing visual stimuli from non-invasive electroencephalography (EEG) remains challenging due to its low spatial resolution and high noise, particularly under realistic low-density electrode configurations. To add…

Enhancing Coherence of Extractive Summarization with Multitask Learning

2023-05-22 · Renlong Jie, Xiaojun Meng, Lifeng Shang, Xin Jiang 외

This study proposes a multitask learning architecture for extractive summarization with coherence boosting. The architecture contains an extractive summarizer and coherent discriminator module. The coherent discriminator…

Extractive SummarizationSentence