paper-with-me

Papers

Evaluation of Automatically Generated Pronoun Reference Questions

2017-09-01 · WS 2017 9 · Arief Yudha Satria, Takenobu Tokunaga

This study provides a detailed analysis of evaluation of English pronoun reference questions which are created automatically by machine. Pronoun reference questions are multiple choice questions that ask test takers to choose an antecedent of a target pronoun in a reading passage from four options. The evaluation was performed from two perspectives: the perspective of English teachers and that of English learners. Item analysis suggests that machine-generated questions achieve comparable quality with human-made questions. Correlation analysis revealed a strong correlation between the scores of machine-generated questions and that of human-made questions.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Multiple-choiceReading Comprehension

Similar Papers 제목 키워드 기반

Automatic Question Generation using Relative Pronouns and Adverbs

2018-07-01 · ACL 2018 7 · Payal Khullar, Konigari Rachna, Mukul Hase, Manish Shrivastava

This paper presents a system that automatically generates multiple, natural language questions using relative pronouns and relative adverbs from complex English sentences. Our system is syntax-based, runs on dependency p…

DescriptiveDialogue GenerationInformation RetrievalQuestion Answering+4

Asking and Answering Questions to Evaluate the Factual Consistency of Summaries

2020-04-08 · ACL 2020 6 · Alex Wang, Kyunghyun Cho, Mike Lewis

Practical applications of abstractive summarization models are limited by frequent factual inconsistencies with respect to their input. Existing automatic evaluation metrics for summarization are largely insensitive to s…

Abstractive Text Summarization

Rapidly Developing High-quality Instruction Data and Evaluation Benchmark for Large Language Models with Minimal Human Effort: A Case Study on Japanese

2024-03-06 · Yikun Sun, Zhen Wan, Nobuhiro Ueda, Sakiko Yahata 외

The creation of instruction data and evaluation benchmarks for serving Large language models often involves enormous human annotation. This issue becomes particularly pronounced when rapidly developing such resources for…

Generating Usage-related Questions for Preference Elicitation in Conversational Recommender Systems

2021-11-26 · Ivica Kostric, Krisztian Balog, Filip Radlinski

A key distinguishing feature of conversational recommender systems over traditional recommender systems is their ability to elicit user preferences using natural language. Currently, the predominant approach to preferenc…

Machine TranslationQuestion GenerationQuestion-GenerationRecommendation Systems

Can LLMs Detect Ambiguous Plural Reference? An Analysis of Split-Antecedent and Mereological Reference

2025-10-06 · Dang Anh, Rick Nouwen, Massimo Poesio arxiv

Our goal is to study how LLMs represent and interpret plural reference in ambiguous and unambiguous contexts. We ask the following research questions: (1) Do LLMs exhibit human-like preferences in representing plural ref…