paper-with-me

Papers

Author-in-the-Loop Response Generation and Evaluation: Integrating Author Expertise and Intent in Responses to Peer Review

2026-01-19 · Qian Ruan, Iryna Gurevych arxiv

Author response (rebuttal) writing is a critical stage of scientific peer review that demands substantial author effort. In practice, authors possess domain expertise, author-only information, and response strategies - concrete forms of author expertise and intent - and seek NLP assistance that integrates these signals into author response generation (ARG). Yet this author-in-the-loop paradigm lacks formal NLP formulation and systematic study: no dataset provides fine-grained author signals, existing ARG work lacks author inputs and controls, and no evaluation measures response reflection of author signals and effectiveness in addressing reviewer concerns. To fill these gaps, we introduce (i) Re3Align, the first large-scale dataset of aligned review-response-revision triplets, where revisions proxy author signals; (ii) REspGen, an author-in-the-loop ARG framework supporting flexible author input, multi-attribute control, and evaluation-guided refinement; and (iii) REspEval, a comprehensive evaluation suite with 20+ metrics spanning input utilization, controllability, response quality, and discourse. Experiments with SOTA LLMs demonstrate the benefits of author input and evaluation-guided refinement, the impact of input specificity on response quality, and controllability-quality trade-offs. We release our dataset, generation and evaluation tools.

📄 PDF Abstract BibTeX arXiv:2602.11173

Code (0)

등록된 구현이 없습니다.

Tasks

Response Generation

Similar Papers 제목 키워드 기반

STORIUM: A Dataset and Evaluation Platform for Machine-in-the-Loop Story Generation

2020-10-04 · EMNLP 2020 11 · Nader Akoury, Shufan Wang, Josh Whiting, Stephen Hood 외

Systems for story generation are asked to produce plausible and enjoyable stories given an input context. This task is underspecified, as a vast number of diverse stories can originate from a single input. The large outp…

Story Generation

Defend: Automated Rebuttals for Peer Review with Minimal Author Guidance

2026-03-28 · Jyotsana Khatri, Manasi Patwardhan arxiv

Rebuttal generation is a critical component of the peer review process for scientific papers, enabling authors to clarify misunderstandings, correct factual inaccuracies, and guide reviewers toward a more accurate evalua…

CoAuthorAI: A Human in the Loop System For Scientific Book Writing

2026-03-27 · Yangjie Tian, Xungang Gu, Yun Zhao, Jiale Yang 외 arxiv

Large language models (LLMs) are increasingly used in scientific writing but struggle with book-length tasks, often producing inconsistent structure and unreliable citations. We introduce CoAuthorAI, a human-in-the-loop …

Trick Me If You Can: Human-in-the-loop Generation of Adversarial Examples for Question Answering

2018-09-07 · TACL 2019 3 · Eric Wallace, Pedro Rodriguez, Shi Feng, Ikuya Yamada 외

Adversarial evaluation stress tests a model's understanding of natural language. While past approaches expose superficial patterns, the resulting adversarial examples are limited in complexity and diversity. We propose h…

DiversityInformation RetrievalQuestion AnsweringRetrieval

Evoke: Evoking Critical Thinking Abilities in LLMs via Reviewer-Author Prompt Editing

2023-10-20 · Xinyu Hu, Pengfei Tang, Simiao Zuo, Zihan Wang 외

Large language models (LLMs) have made impressive progress in natural language processing. These models rely on proper human instructions (or prompts) to generate suitable responses. However, the potential of LLMs are no…

Logical Fallacy Detection