paper-with-me

홈 › Papers

Evaluation of Text Generation: A Survey

2020-06-26 · Asli Celikyilmaz, Elizabeth Clark, Jianfeng Gao

The paper surveys evaluation methods of natural language generation (NLG) systems that have been developed in the last few years. We group NLG evaluation methods into three categories: (1) human-centric evaluation metrics, (2) automatic metrics that require no training, and (3) machine-learned metrics. For each category, we discuss the progress that has been made and the challenges still being faced, with a focus on the evaluation of recently proposed NLG tasks and neural NLG models. We then present two examples for task-specific NLG evaluations for automatic text summarization and long text generation, and conclude the paper by proposing future research directions.

📄 PDF Abstract BibTeX arXiv:2006.14799

Code (0)

등록된 구현이 없습니다.

Tasks

nlg evaluationSurveyText GenerationText Summarization

Similar Papers 제목 키워드 기반

SurveyX: Academic Survey Automation via Large Language Models

2025-02-20 · Xun Liang, Jiawei Yang, Yezhaohui Wang, Chen Tang 외

Large Language Models (LLMs) have demonstrated exceptional comprehension capabilities and a vast knowledge base, suggesting that LLMs can serve as efficient tools for automated survey generation. However, recent research…

Survey

SurveyEval: Towards Comprehensive Evaluation of LLM-Generated Academic Surveys

2025-12-02 · Jiahao Zhao, Shuaixing Zhang, Nan Xu, Lei Wang arxiv

LLM-based automatic survey systems are transforming how users acquire information from the web by integrating retrieval, organization, and content synthesis into end-to-end generation pipelines. While recent works focus …

Can We Catch the Elephant? A Survey of the Evolvement of Hallucination Evaluation on Natural Language Generation

2024-04-18 · Siya Qi, Yulan He, Zheng Yuan

Hallucination in Natural Language Generation (NLG) is like the elephant in the room, obvious but often overlooked until recent achievements significantly improved the fluency and grammaticality of generated text. As the …

HallucinationHallucination EvaluationSurveyText Generation

SurveyGen: Quality-Aware Scientific Survey Generation with Large Language Models

2025-08-25 · Tong Bao, Mir Tafseer Nayeem, Davood Rafiei, Chengzhi Zhang arxiv

Automatic survey generation has emerged as a key task in scientific document processing. While large language models (LLMs) have shown promise in generating survey texts, the lack of standardized evaluation datasets crit…

MVSS: A Unified Framework for Multi-View Structured Survey Generation

2026-01-14 · Yinqi Liu, Yueqi Zhu, Yongkang Zhang, Feiran Liu 외 arxiv

Scientific surveys require not only summarizing large bodies of literature, but also organizing them into clear and coherent conceptual structures. However, existing automatic survey generation methods typically focus on…

Text Generation