paper-with-me

홈 › Papers

The Iron(ic) Melting Pot: Reviewing Human Evaluation in Humour, Irony and Sarcasm Generation

2023-11-09 · Tyler Loakman, Aaron Maladry, Chenghua Lin

Human evaluation is often considered to be the gold standard method of evaluating a Natural Language Generation system. However, whilst its importance is accepted by the community at large, the quality of its execution is often brought into question. In this position paper, we argue that the generation of more esoteric forms of language - humour, irony and sarcasm - constitutes a subdomain where the characteristics of selected evaluator panels are of utmost importance, and every effort should be made to report demographic characteristics wherever possible, in the interest of transparency and replicability. We support these claims with an overview of each language form and an analysis of examples in terms of how their interpretation is affected by different participant variables. We additionally perform a critical survey of recent works in NLG to assess how well evaluation procedures are reported in this subdomain, and note a severe lack of open reporting of evaluator demographic information, and a significant reliance on crowdsourcing platforms for recruitment.

📄 PDF Abstract BibTeX arXiv:2311.05552

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Similar Papers 제목 키워드 기반

A machine learning approach to automation and uncertainty evaluation for self-validating thermocouples

2025-10-21 · Samuel Bilson, Andrew Thompson, Declan Tucker, Jonathan Pearce arxiv

Thermocouples are in widespread use in industry, but they are particularly susceptible to calibration drift in harsh environments. Self-validating thermocouples aim to address this issue by using a miniature phase-change…

Scalable Evaluation of Multi-Agent Reinforcement Learning with Melting Pot

2021-07-14 · Joel Z. Leibo, Edgar Duéñez-Guzmán, Alexander Sasha Vezhnevets, John P. Agapiou 외

Existing evaluation suites for multi-agent reinforcement learning (MARL) do not assess generalization to novel situations as their primary objective (unlike supervised-learning benchmarks). Our contribution, Melting Pot,…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Psychology-Driven Enhancement of Humour Translation

2025-07-12 · Yuchen Su, Yonghua Zhu, Yang Chen, Diana Benavides-Prado 외 arxiv

Humour translation plays a vital role as a bridge between different cultures, fostering understanding and communication. Although most existing Large Language Models (LLMs) are capable of general translation tasks, these…

Melting Pot 2.0

2022-11-24 · John P. Agapiou, Alexander Sasha Vezhnevets, Edgar A. Duéñez-Guzmán, Jayd Matyas 외

Multi-agent artificial intelligence research promises a path to develop intelligent technologies that are more human-like and more human-compatible than those produced by "solipsistic" approaches, which do not consider i…

Artificial LifeNavigate

Who's Laughing Now? An Overview of Computational Humour Generation and Explanation

2025-09-25 · Tyler Loakman, William Thorne, Chenghua Lin arxiv

The creation and perception of humour is a fundamental human trait, positioning its computational understanding as one of the most challenging tasks in natural language processing (NLP). As an abstract, creative, and fre…