ChatLang-8: An LLM-Based Synthetic Data Generation Framework for Grammatical Error Correction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, Jeiyoon, Park, Chanjun, Lim, Heuiseok |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CharacterGPT: A Persona Reconstruction Framework for Role-Playing Agents
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024)
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
von: Park, Chanhee, et al.
Veröffentlicht: (2025)
von: Park, Chanhee, et al.
Veröffentlicht: (2025)
Analysis of Utterance Embeddings and Clustering Methods Related to Intent Induction for Task-Oriented Dialogue
von: Park, Jeiyoon, et al.
Veröffentlicht: (2022)
von: Park, Jeiyoon, et al.
Veröffentlicht: (2022)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
MultiDocFusion: Hierarchical and Multimodal Chunking Pipeline for Enhanced RAG on Long Industrial Documents
von: Shin, Joongmin, et al.
Veröffentlicht: (2026)
von: Shin, Joongmin, et al.
Veröffentlicht: (2026)
From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
Benchmark Profiling: Mechanistic Diagnosis of LLM Benchmarks
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
von: Kim, Dongjun, et al.
Veröffentlicht: (2025)
Alternative Speech: Complementary Method to Counter-Narrative for Better Discourse
von: Lee, Seungyoon, et al.
Veröffentlicht: (2024)
von: Lee, Seungyoon, et al.
Veröffentlicht: (2024)
CoME: An Unlearning-based Approach to Conflict-free Model Editing
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
FLEX: A Benchmark for Evaluating Robustness of Fairness in Large Language Models
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
von: Jung, Dahyun, et al.
Veröffentlicht: (2025)
Translation of Multifaceted Data without Re-Training of Machine Translation Systems
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2024)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2024)
Leveraging What's Overfixed: Post-Correction via LLM Grammatical Error Overcorrection
von: Park, Taehee, et al.
Veröffentlicht: (2025)
von: Park, Taehee, et al.
Veröffentlicht: (2025)
Towards the Development of Balanced Synthetic Data for Correcting Grammatical Errors in Arabic: An Approach Based on Error Tagging Model and Synthetic Data Generating Model
von: Alrehili, Ahlam, et al.
Veröffentlicht: (2025)
von: Alrehili, Ahlam, et al.
Veröffentlicht: (2025)
Understanding LLM Development Through Longitudinal Study: Insights from the Open Ko-LLM Leaderboard
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
Enhancing Grammatical Error Detection using BERT with Cleaned Lang-8 Dataset
von: Nihalani, Rahul, et al.
Veröffentlicht: (2024)
von: Nihalani, Rahul, et al.
Veröffentlicht: (2024)
InstaTrans: An Instruction-Aware Translation Framework for Non-English Instruction Datasets
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
von: Kim, Yungi, et al.
Veröffentlicht: (2024)
Model-Based Data-Centric AI: Bridging the Divide Between Academic Ideals and Industrial Pragmatism
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
von: Park, Chanjun, et al.
Veröffentlicht: (2024)
Zero-shot Cross-Lingual Transfer for Synthetic Data Generation in Grammatical Error Detection
von: Latouche, Gaetan Lopez, et al.
Veröffentlicht: (2024)
von: Latouche, Gaetan Lopez, et al.
Veröffentlicht: (2024)
Tibyan Corpus: Balanced and Comprehensive Error Coverage Corpus Using ChatGPT for Arabic Grammatical Error Correction
von: Alrehili, Ahlam, et al.
Veröffentlicht: (2024)
von: Alrehili, Ahlam, et al.
Veröffentlicht: (2024)
Improving Grammatical Error Correction via Contextual Data Augmentation
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
Corrections Meet Explanations: A Unified Framework for Explainable Grammatical Error Correction
von: Ye, Jingheng, et al.
Veröffentlicht: (2025)
von: Ye, Jingheng, et al.
Veröffentlicht: (2025)
GPT-3.5 for Grammatical Error Correction
von: Katinskaia, Anisia, et al.
Veröffentlicht: (2024)
von: Katinskaia, Anisia, et al.
Veröffentlicht: (2024)
Organic Data-Driven Approach for Turkish Grammatical Error Correction and LLMs
von: Ersoy, Asım, et al.
Veröffentlicht: (2024)
von: Ersoy, Asım, et al.
Veröffentlicht: (2024)
LLMCL-GEC: Advancing Grammatical Error Correction with LLM-Driven Curriculum Learning
von: Fang, Tao, et al.
Veröffentlicht: (2024)
von: Fang, Tao, et al.
Veröffentlicht: (2024)
Harnessing Rule-Based Reinforcement Learning for Enhanced Grammatical Error Correction
von: Li, Yilin, et al.
Veröffentlicht: (2025)
von: Li, Yilin, et al.
Veröffentlicht: (2025)
Data Augmentation for Spoken Grammatical Error Correction
von: Karanasou, Penny, et al.
Veröffentlicht: (2025)
von: Karanasou, Penny, et al.
Veröffentlicht: (2025)
Adapting LLMs for Minimal-edit Grammatical Error Correction
von: Staruch, Ryszard, et al.
Veröffentlicht: (2025)
von: Staruch, Ryszard, et al.
Veröffentlicht: (2025)
Loss-Aware Curriculum Learning for Chinese Grammatical Error Correction
von: Zhang, Ding, et al.
Veröffentlicht: (2024)
von: Zhang, Ding, et al.
Veröffentlicht: (2024)
Call for Rigor in Reporting Quality of Instruction Tuning Data
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
ZEBRA: Leveraging Model-Behavioral Knowledge for Zero-Annotation Preference Dataset Construction
von: Jung, Jeesu, et al.
Veröffentlicht: (2025)
von: Jung, Jeesu, et al.
Veröffentlicht: (2025)
NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Moon, Hyeonseok, et al.
Veröffentlicht: (2025)
Unveiling the Limits of Large Language Models in Inferring Pragmatic Meaning from Non-Verbal Responses
von: Eo, Sugyeong, et al.
Veröffentlicht: (2026)
von: Eo, Sugyeong, et al.
Veröffentlicht: (2026)
"When Data is Scarce, Prompt Smarter"... Approaches to Grammatical Error Correction in Low-Resource Settings
von: De, Somsubhra, et al.
Veröffentlicht: (2025)
von: De, Somsubhra, et al.
Veröffentlicht: (2025)
IMPARA-GED: Grammatical Error Detection is Boosting Reference-free Grammatical Error Quality Estimator
von: Sakai, Yusuke, et al.
Veröffentlicht: (2025)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2025)
Sensory-Aware Sequential Recommendation via Review-Distilled Representations
von: Yoon, Yeo Chan, et al.
Veröffentlicht: (2026)
von: Yoon, Yeo Chan, et al.
Veröffentlicht: (2026)
GECTurk WEB: An Explainable Online Platform for Turkish Grammatical Error Detection and Correction
von: Gebeşçe, Ali, et al.
Veröffentlicht: (2024)
von: Gebeşçe, Ali, et al.
Veröffentlicht: (2024)
KoGEC : Korean Grammatical Error Correction with Pre-trained Translation Models
von: Kim, Taeeun, et al.
Veröffentlicht: (2025)
von: Kim, Taeeun, et al.
Veröffentlicht: (2025)
Debate Only When Necessary: Adaptive Multiagent Collaboration for Efficient LLM Reasoning
von: Eo, Sugyeong, et al.
Veröffentlicht: (2025)
von: Eo, Sugyeong, et al.
Veröffentlicht: (2025)
LANGSAE EDITING: Improving Multilingual Information Retrieval via Post-hoc Language Identity Removal
von: Kim, Dongjun, et al.
Veröffentlicht: (2026)
von: Kim, Dongjun, et al.
Veröffentlicht: (2026)
Exploring Coding Spot: Understanding Parametric Contributions to LLM Coding Performance
von: Kim, Dongjun, et al.
Veröffentlicht: (2024)
von: Kim, Dongjun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CharacterGPT: A Persona Reconstruction Framework for Role-Playing Agents
von: Park, Jeiyoon, et al.
Veröffentlicht: (2024) -
MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
von: Park, Chanhee, et al.
Veröffentlicht: (2025) -
Analysis of Utterance Embeddings and Clustering Methods Related to Intent Induction for Task-Oriented Dialogue
von: Park, Jeiyoon, et al.
Veröffentlicht: (2022) -
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
von: Kim, Dongjun, et al.
Veröffentlicht: (2025) -
MultiDocFusion: Hierarchical and Multimodal Chunking Pipeline for Enhanced RAG on Long Industrial Documents
von: Shin, Joongmin, et al.
Veröffentlicht: (2026)