CEC-Zero: Zero-Supervision Character Error Correction with Self-Generated Rewards
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Lin, Zhiming, Zhao, Kai, Zhang, Sophie, Yu, Peilai, Xiao, Canran |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
CEC-Zero: Chinese Error Correction Solution Based on LLM
par: Zhang, Sophie, et autres
Publié: (2025)
par: Zhang, Sophie, et autres
Publié: (2025)
Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision
par: He, Yinghui, et autres
Publié: (2026)
par: He, Yinghui, et autres
Publié: (2026)
A Training-free LLM-based Approach to General Chinese Character Error Correction
par: Zhou, Houquan, et autres
Publié: (2025)
par: Zhou, Houquan, et autres
Publié: (2025)
Writing-Zero: Bridge the Gap Between Non-verifiable Tasks and Verifiable Rewards
par: Jia, Ruipeng, et autres
Publié: (2025)
par: Jia, Ruipeng, et autres
Publié: (2025)
Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model
par: Liu, Runheng, et autres
Publié: (2026)
par: Liu, Runheng, et autres
Publié: (2026)
AlphaMath Almost Zero: Process Supervision without Process
par: Chen, Guoxin, et autres
Publié: (2024)
par: Chen, Guoxin, et autres
Publié: (2024)
Error Typing for Smarter Rewards: Improving Process Reward Models with Error-Aware Hierarchical Supervision
par: Pala, Tej Deep, et autres
Publié: (2025)
par: Pala, Tej Deep, et autres
Publié: (2025)
Absolute Zero: Reinforced Self-play Reasoning with Zero Data
par: Zhao, Andrew, et autres
Publié: (2025)
par: Zhao, Andrew, et autres
Publié: (2025)
SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation
par: Yang, Wenjie, et autres
Publié: (2025)
par: Yang, Wenjie, et autres
Publié: (2025)
R-Zero: Self-Evolving Reasoning LLM from Zero Data
par: Huang, Chengsong, et autres
Publié: (2025)
par: Huang, Chengsong, et autres
Publié: (2025)
Reuse Your Rewards: Reward Model Transfer for Zero-Shot Cross-Lingual Alignment
par: Wu, Zhaofeng, et autres
Publié: (2024)
par: Wu, Zhaofeng, et autres
Publié: (2024)
Triviality Corrected Endogenous Reward
par: Wang, Xinda, et autres
Publié: (2026)
par: Wang, Xinda, et autres
Publié: (2026)
PMF-CEC: Phoneme-augmented Multimodal Fusion for Context-aware ASR Error Correction with Error-specific Selective Decoding
par: He, Jiajun, et autres
Publié: (2025)
par: He, Jiajun, et autres
Publié: (2025)
Zero-shot Generative Linguistic Steganography
par: Lin, Ke, et autres
Publié: (2024)
par: Lin, Ke, et autres
Publié: (2024)
G-Zero: Self-Play for Open-Ended Generation from Zero Data
par: Huang, Chengsong, et autres
Publié: (2026)
par: Huang, Chengsong, et autres
Publié: (2026)
C-LLM: Learn to Check Chinese Spelling Errors Character by Character
par: Li, Kunting, et autres
Publié: (2024)
par: Li, Kunting, et autres
Publié: (2024)
Zero-RAG: Towards Retrieval-Augmented Generation with Zero Redundant Knowledge
par: Luo, Qi, et autres
Publié: (2025)
par: Luo, Qi, et autres
Publié: (2025)
ChatZero:Zero-shot Cross-Lingual Dialogue Generation via Pseudo-Target Language
par: Liu, Yongkang, et autres
Publié: (2024)
par: Liu, Yongkang, et autres
Publié: (2024)
Full-Step-DPO: Self-Supervised Preference Optimization with Step-wise Rewards for Mathematical Reasoning
par: Xu, Huimin, et autres
Publié: (2025)
par: Xu, Huimin, et autres
Publié: (2025)
Multilingual Zero Resource Speech Recognition Base on Self-Supervise Pre-Trained Acoustic Models
par: Wang, Haoyu, et autres
Publié: (2022)
par: Wang, Haoyu, et autres
Publié: (2022)
Investigating Zero-Shot Generalizability on Mandarin-English Code-Switched ASR and Speech-to-text Translation of Recent Foundation Models with Self-Supervision and Weak Supervision
par: Yang, Chih-Kai, et autres
Publié: (2023)
par: Yang, Chih-Kai, et autres
Publié: (2023)
Zero-shot Cross-Lingual Transfer for Synthetic Data Generation in Grammatical Error Detection
par: Latouche, Gaetan Lopez, et autres
Publié: (2024)
par: Latouche, Gaetan Lopez, et autres
Publié: (2024)
Generate then Refine: Data Augmentation for Zero-shot Intent Detection
par: Lin, I-Fan, et autres
Publié: (2024)
par: Lin, I-Fan, et autres
Publié: (2024)
Self-Prompting Large Language Models for Zero-Shot Open-Domain QA
par: Li, Junlong, et autres
Publié: (2022)
par: Li, Junlong, et autres
Publié: (2022)
Self-Evolved Reward Learning for LLMs
par: Huang, Chenghua, et autres
Publié: (2024)
par: Huang, Chenghua, et autres
Publié: (2024)
Reward-RAG: Enhancing RAG with Reward Driven Supervision
par: Nguyen, Thang, et autres
Publié: (2024)
par: Nguyen, Thang, et autres
Publié: (2024)
Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation
par: Xu, Zhichao, et autres
Publié: (2025)
par: Xu, Zhichao, et autres
Publié: (2025)
On Zero-Shot Counterspeech Generation by LLMs
par: Saha, Punyajoy, et autres
Publié: (2024)
par: Saha, Punyajoy, et autres
Publié: (2024)
Zero-Shot Stance Detection in the Wild: Dynamic Target Generation and Multi-Target Adaptation
par: Li, Aohua, et autres
Publié: (2026)
par: Li, Aohua, et autres
Publié: (2026)
Bring Your Own KG: Self-Supervised Program Synthesis for Zero-Shot KGQA
par: Agarwal, Dhruv, et autres
Publié: (2023)
par: Agarwal, Dhruv, et autres
Publié: (2023)
Taxonomy-Guided Zero-Shot Recommendations with LLMs
par: Liang, Yueqing, et autres
Publié: (2024)
par: Liang, Yueqing, et autres
Publié: (2024)
Generation-driven Contrastive Self-training for Zero-shot Text Classification with Instruction-following LLM
par: Zhang, Ruohong, et autres
Publié: (2023)
par: Zhang, Ruohong, et autres
Publié: (2023)
CorNav: Autonomous Agent with Self-Corrected Planning for Zero-Shot Vision-and-Language Navigation
par: Liang, Xiwen, et autres
Publié: (2023)
par: Liang, Xiwen, et autres
Publié: (2023)
Self-Improving for Zero-Shot Named Entity Recognition with Large Language Models
par: Xie, Tingyu, et autres
Publié: (2023)
par: Xie, Tingyu, et autres
Publié: (2023)
The Flip Side of RLHF: On-Policy Feedback for Reward Model Self-Supervised Improvement
par: Wang, Xiaobo, et autres
Publié: (2026)
par: Wang, Xiaobo, et autres
Publié: (2026)
GRAM-R$^2$: Self-Training Generative Foundation Reward Models for Reward Reasoning
par: Wang, Chenglong, et autres
Publié: (2025)
par: Wang, Chenglong, et autres
Publié: (2025)
Revealing and Mitigating the Challenge of Detecting Character Knowledge Errors in LLM Role-Playing
par: Zhang, Wenyuan, et autres
Publié: (2024)
par: Zhang, Wenyuan, et autres
Publié: (2024)
Generative Meta-Learning for Zero-Shot Relation Triplet Extraction
par: Li, Wanli, et autres
Publié: (2023)
par: Li, Wanli, et autres
Publié: (2023)
Solving a Million-Step LLM Task with Zero Errors
par: Meyerson, Elliot, et autres
Publié: (2025)
par: Meyerson, Elliot, et autres
Publié: (2025)
Distilling Implicit Multimodal Knowledge into Large Language Models for Zero-Resource Dialogue Generation
par: Zhang, Bo, et autres
Publié: (2024)
par: Zhang, Bo, et autres
Publié: (2024)
Documents similaires
-
CEC-Zero: Chinese Error Correction Solution Based on LLM
par: Zhang, Sophie, et autres
Publié: (2025) -
Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision
par: He, Yinghui, et autres
Publié: (2026) -
A Training-free LLM-based Approach to General Chinese Character Error Correction
par: Zhou, Houquan, et autres
Publié: (2025) -
Writing-Zero: Bridge the Gap Between Non-verifiable Tasks and Verifiable Rewards
par: Jia, Ruipeng, et autres
Publié: (2025) -
Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model
par: Liu, Runheng, et autres
Publié: (2026)