paper-with-me

Papers

CEC-Zero: Zero-Supervision Character Error Correction with Self-Generated Rewards

2025-12-30 · Zhiming Lin, Kai Zhao, Sophie Zhang, Peilai Yu, Canran Xiao arxiv

Large-scale Chinese spelling correction (CSC) remains critical for real-world text processing, yet existing LLMs and supervised methods lack robustness to novel errors and rely on costly annotations. We introduce CEC-Zero, a zero-supervision reinforcement learning framework that addresses this by enabling LLMs to correct their own mistakes. CEC-Zero synthesizes errorful inputs from clean text, computes cluster-consensus rewards via semantic similarity and candidate agreement, and optimizes the policy with PPO. It outperforms supervised baselines by 10--13 F$_1$ points and strong LLM fine-tunes by 5--8 points across 9 benchmarks, with theoretical guarantees of unbiased rewards and convergence. CEC-Zero establishes a label-free paradigm for robust, scalable CSC, unlocking LLM potential in noisy text pipelines.

📄 PDF Abstract BibTeX arXiv:2512.23971

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningSemantic Similarity

Similar Papers 제목 키워드 기반

CEC-Zero: Chinese Error Correction Solution Based on LLM

2025-05-14 · Sophie Zhang, Zhiming Lin

Recent advancements in large language models (LLMs) demonstrate exceptional Chinese text processing capabilities, particularly in Chinese Spelling Correction (CSC). While LLMs outperform traditional BERT-based models in …

Domain GeneralizationReinforcement Learning (RL)Spelling Correction

Zero-shot Faithful Factual Error Correction

2023-05-13 · Kung-Hsiang Huang, Hou Pong Chan, Heng Ji

Faithfully correcting factual errors is critical for maintaining the integrity of textual knowledge bases and preventing hallucinations in sequence-to-sequence models. Drawing on humans' ability to identify and correct f…

A Systematic Analysis of Large Language Models with RAG-enabled Dynamic Prompting for Medical Error Detection and Correction

2025-11-25 · Farzad Ahmed, Joniel Augustine Jerome, Meliha Yetisgen, Özlem Uzuner arxiv

Objective: Clinical documentation contains factual, diagnostic, and management errors that can compromise patient safety. Large language models (LLMs) may help detect and correct such errors, but their behavior under dif…

Chinese Spelling Correction as Rephrasing Language Model

2023-08-17 · Linfeng Liu, Hongqiu Wu, Hai Zhao

This paper studies Chinese Spelling Correction (CSC), which aims to detect and correct the potential spelling errors in a given sentence. Current state-of-the-art methods regard CSC as a sequence tagging task and fine-tu…

Language ModelingLanguage ModellingmodelSentence+1

ZeroIDIR: Zero-Reference Illumination Degradation Image Restoration with Perturbed Consistency Diffusion Models

2026-05-12 · Hai Jiang, Zhen Liu, Yinjie Lei, Songchen Han 외 arxiv

In this paper, we propose a zero-reference diffusion-based framework, named ZeroIDIR, for illumination degradation image restoration, which decouples the restoration process into adaptive illumination correction and diff…

Image Restoration