paper-with-me

Papers

Learning from Mistakes: Self-correct Adversarial Training for Chinese Unnatural Text Correction

2024-12-23 · Xuan Feng, Tianlong Gu, Xiaoli Liu, Liang Chang

Unnatural text correction aims to automatically detect and correct spelling errors or adversarial perturbation errors in sentences. Existing methods typically rely on fine-tuning or adversarial training to correct errors, which have achieved significant success. However, these methods exhibit poor generalization performance due to the difference in data distribution between training data and real-world scenarios, known as the exposure bias problem. In this paper, we propose a self-correct adversarial training framework for \textbf{L}earn\textbf{I}ng from \textbf{MI}s\textbf{T}akes (\textbf{LIMIT}), which is a task- and model-independent framework to correct unnatural errors or mistakes. Specifically, we fully utilize errors generated by the model that are actively exposed during the inference phase, i.e., predictions that are inconsistent with the target. This training method not only simulates potential errors in real application scenarios, but also mitigates the exposure bias of the traditional training process. Meanwhile, we design a novel decoding intervention strategy to maintain semantic consistency. Extensive experimental results on Chinese unnatural text error correction datasets show that our proposed method can correct multiple forms of errors and outperforms the state-of-the-art text correction methods. In addition, extensive results on Chinese and English datasets validate that LIMIT can serve as a plug-and-play defense module and can extend to new models and datasets without further training.

📄 PDF Abstract BibTeX arXiv:2412.17279

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CEC-Zero: Zero-Supervision Character Error Correction with Self-Generated Rewards

2025-12-30 · Zhiming Lin, Kai Zhao, Sophie Zhang, Peilai Yu 외 arxiv

Large-scale Chinese spelling correction (CSC) remains critical for real-world text processing, yet existing LLMs and supervised methods lack robustness to novel errors and rely on costly annotations. We introduce CEC-Zer…

Reinforcement LearningSemantic Similarity

LLMs cannot find reasoning errors, but can correct them given the error location

2023-11-14 · Gladys Tyen, Hassan Mansoor, Victor Cărbune, Peter Chen 외

While self-correction has shown promise in improving LLM outputs in terms of style and quality (e.g. Chen et al., 2023b; Madaan et al., 2023), recent attempts to self-correct logical or reasoning errors often cause corre…

Internalized Self-Correction for Large Language Models

2024-12-21 · Nishanth Upadhyaya, Raghavendra Sridharamurthy

In this article, we introduce 'Internalized Self-Correction' (InSeC) for large language models (LLMs). While many approaches exist for self-reflection at inference time, we propose a novel method that combines ideas from…

Instruction Following

A Aelf-supervised Tibetan-chinese Vocabulary Alignment Method Based On Adversarial Learning

2021-10-04 · Enshuai Hou, Jie Zhu

Tibetan is a low-resource language. In order to alleviate the shortage of parallel corpus between Tibetan and Chinese, this paper uses two monolingual corpora and a small number of seed dictionaries to learn the semi-sup…

When Can LLMs Actually Correct Their Own Mistakes? A Critical Survey of Self-Correction of LLMs

2024-06-03 · Ryo Kamoi, Yusen Zhang, Nan Zhang, Jiawei Han 외

Self-correction is an approach to improving responses from large language models (LLMs) by refining the responses using LLMs during inference. Prior work has proposed various self-correction frameworks using different so…

Survey