paper-with-me

홈 › Papers

Evaluating the text quality, human likeness and tailoring component of PASS: A Dutch data-to-text system for soccer

2018-08-01 · COLING 2018 8 · Chris van der Lee, Bart Verduijn, Emiel Krahmer, S Wubben, er

We present an evaluation of PASS, a data-to-text system that generates Dutch soccer reports from match statistics which are automatically tailored towards fans of one club or the other. The evaluation in this paper consists of two studies. An intrinsic human-based evaluation of the system{'}s output is described in the first study. In this study it was found that compared to human-written texts, computer-generated texts were rated slightly lower on style-related text components (fluency and clarity) and slightly higher in terms of the correctness of given information. Furthermore, results from the first study showed that tailoring was accurately recognized in most cases, and that participants struggled with correctly identifying whether a text was written by a human or computer. The second study investigated if tailoring affects perceived text quality, for which no results were garnered. This lack of results might be due to negative preconceptions about computer-generated texts which were found in the first study.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Text Generation

Similar Papers 제목 키워드 기반

Towards Objective Evaluation of Socially-Situated Conversational Robots: Assessing Human-Likeness through Multimodal User Behaviors

2023-08-21 · Koji Inoue, Divesh Lala, Keiko Ochi, Tatsuya Kawahara 외

This paper tackles the challenging task of evaluating socially situated conversational robots and presents a novel objective evaluation approach that relies on multimodal user behaviors. In this study, our main focus is …

HLB: Benchmarking LLMs' Humanlikeness in Language Use

2024-09-24 · Xufeng Duan, Bei Xiao, Xuemei Tang, Zhenguang G. Cai

As synthetic data becomes increasingly prevalent in training language models, particularly through generated dialogue, concerns have emerged that these models may deviate from authentic human language patterns, potential…

Benchmarking

Towards Motion Turing Test: Evaluating Human-Likeness in Humanoid Robots

2026-03-06 · Mingzhe Li, Mengyin Liu, Zekai Wu, Xincheng Lin 외 arxiv

Humanoid robots have achieved significant progress in motion generation and control, exhibiting movements that appear increasingly natural and human-like. Inspired by the Turing Test, we propose the Motion Turing Test, a…

Spoken Humanoid Embodied Conversational Agents in Mobile Serious Games: A Usability Assessment

2023-09-14 · Danai Korre, Judy Robertson

This paper presents an empirical investigation of the extent to which spoken Humanoid Embodied Conversational Agents (HECAs) can foster usability in mobile serious game (MSG) applications. The aim of the research is to a…

GrowLoop: Self-Evolving Conversation Evaluation Seeded by Human

2026-05-26 · Yihang Lin, Yunze Gao, Zeyang Lin, Dongbo Li 외 arxiv

With the rapid advancement of large language models, evaluating human-likeness in open-ended conversation has become increasingly important. However, human-likeness is a form of tacit knowledge that humans perceive intui…