Time Is Effort: Estimating Human Post-Editing Time for Grammar Error Correction Tool Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | Vadehra, Ankit, Johnson, Bill, Saunders, Gene, Poupart, Pascal |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Critical Look At Tokenwise Reward-Guided Text Generation
di: Rashid, Ahmad, et al.
Pubblicazione: (2024)
di: Rashid, Ahmad, et al.
Pubblicazione: (2024)
Towards Cost-Effective Reward Guided Text Generation
di: Rashid, Ahmad, et al.
Pubblicazione: (2025)
di: Rashid, Ahmad, et al.
Pubblicazione: (2025)
Grammar as a Behavioral Biometric: Using Cognitively Motivated Grammar Models for Authorship Verification
di: Nini, Andrea, et al.
Pubblicazione: (2024)
di: Nini, Andrea, et al.
Pubblicazione: (2024)
Neural Grammatical Error Correction for Romanian
di: Cotet, Teodor-Mihai, et al.
Pubblicazione: (2026)
di: Cotet, Teodor-Mihai, et al.
Pubblicazione: (2026)
Time Sensitive Knowledge Editing through Efficient Finetuning
di: Ge, Xiou, et al.
Pubblicazione: (2024)
di: Ge, Xiou, et al.
Pubblicazione: (2024)
Latent Phase-Shift Rollback: Inference-Time Error Correction via Residual Stream Monitoring and KV-Cache Steering
di: Gupta, Manan, et al.
Pubblicazione: (2026)
di: Gupta, Manan, et al.
Pubblicazione: (2026)
Tools Fail: Detecting Silent Errors in Faulty Tools
di: Sun, Jimin, et al.
Pubblicazione: (2024)
di: Sun, Jimin, et al.
Pubblicazione: (2024)
MEMOIR: Lifelong Model Editing with Minimal Overwrite and Informed Retention for LLMs
di: Wang, Ke, et al.
Pubblicazione: (2025)
di: Wang, Ke, et al.
Pubblicazione: (2025)
Grammatical Error Correction for Low-Resource Languages: The Case of Zarma
di: Keita, Mamadou K., et al.
Pubblicazione: (2024)
di: Keita, Mamadou K., et al.
Pubblicazione: (2024)
Exploring Changes in Nation Perception with Nationality-Assigned Personas in LLMs
di: Kamruzzaman, Mahammed, et al.
Pubblicazione: (2024)
di: Kamruzzaman, Mahammed, et al.
Pubblicazione: (2024)
A Unified View of Delta Parameter Editing in Post-Trained Large-Scale Models
di: Tang, Qiaoyu, et al.
Pubblicazione: (2024)
di: Tang, Qiaoyu, et al.
Pubblicazione: (2024)
Learning to Reason Over Time: Timeline Self-Reflection for Improved Temporal Reasoning in Language Models
di: Bazaga, Adrián, et al.
Pubblicazione: (2025)
di: Bazaga, Adrián, et al.
Pubblicazione: (2025)
Grammar-Aligned Decoding
di: Park, Kanghee, et al.
Pubblicazione: (2024)
di: Park, Kanghee, et al.
Pubblicazione: (2024)
Assessing the Efficacy of Grammar Error Correction: A Human Evaluation Approach in the Japanese Context
di: Wang, Qiao, et al.
Pubblicazione: (2024)
di: Wang, Qiao, et al.
Pubblicazione: (2024)
Evaluating Judges as Evaluators: The JETTS Benchmark of LLM-as-Judges as Test-Time Scaling Evaluators
di: Zhou, Yilun, et al.
Pubblicazione: (2025)
di: Zhou, Yilun, et al.
Pubblicazione: (2025)
Confidence Estimation for Error Detection in Text-to-SQL Systems
di: Somov, Oleg, et al.
Pubblicazione: (2025)
di: Somov, Oleg, et al.
Pubblicazione: (2025)
Automatic Generation of Python Programs Using Context-Free Grammars
di: Yamani, Kamel, et al.
Pubblicazione: (2024)
di: Yamani, Kamel, et al.
Pubblicazione: (2024)
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction
di: Li, Xiaoyuan, et al.
Pubblicazione: (2024)
di: Li, Xiaoyuan, et al.
Pubblicazione: (2024)
FEval-TTC: Fair Evaluation Protocol for Test-Time Compute
di: Rumiantsev, Pavel, et al.
Pubblicazione: (2025)
di: Rumiantsev, Pavel, et al.
Pubblicazione: (2025)
Pareto Optimal Learning for Estimating Large Language Model Errors
di: Zhao, Theodore, et al.
Pubblicazione: (2023)
di: Zhao, Theodore, et al.
Pubblicazione: (2023)
CKnowEdit: A New Chinese Knowledge Editing Dataset for Linguistics, Facts, and Logic Error Correction in LLMs
di: Fang, Jizhan, et al.
Pubblicazione: (2024)
di: Fang, Jizhan, et al.
Pubblicazione: (2024)
Efficient Post-Training Pruning of Large Language Models with Statistical Correction
di: Yu, Peiqi, et al.
Pubblicazione: (2026)
di: Yu, Peiqi, et al.
Pubblicazione: (2026)
Correcting Suppressed Log-Probabilities in Language Models with Post-Transformer Adapters
di: Sanchez, Bryan
Pubblicazione: (2026)
di: Sanchez, Bryan
Pubblicazione: (2026)
Know When You're Wrong: Aligning Confidence with Correctness for LLM Error Detection
di: Xiaohu, Xie, et al.
Pubblicazione: (2026)
di: Xiaohu, Xie, et al.
Pubblicazione: (2026)
Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
di: Song, Yifan, et al.
Pubblicazione: (2024)
di: Song, Yifan, et al.
Pubblicazione: (2024)
Watermarking Language Models with Error Correcting Codes
di: Chao, Patrick, et al.
Pubblicazione: (2024)
di: Chao, Patrick, et al.
Pubblicazione: (2024)
PITA: Preference-Guided Inference-Time Alignment for LLM Post-Training
di: Bobbili, Sarat Chandra, et al.
Pubblicazione: (2025)
di: Bobbili, Sarat Chandra, et al.
Pubblicazione: (2025)
ShIOEnv: A Command Evaluation Environment for Grammar-Constrained Synthesis and Execution Behavior Modeling
di: Ragsdale, Jarrod, et al.
Pubblicazione: (2025)
di: Ragsdale, Jarrod, et al.
Pubblicazione: (2025)
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
di: Dumitran, Adrian-Marius, et al.
Pubblicazione: (2025)
LLMs in the Imaginarium: Tool Learning through Simulated Trial and Error
di: Wang, Boshi, et al.
Pubblicazione: (2024)
di: Wang, Boshi, et al.
Pubblicazione: (2024)
Can Prompts Rewind Time for LLMs? Evaluating the Effectiveness of Prompted Knowledge Cutoffs
di: Gao, Xin, et al.
Pubblicazione: (2025)
di: Gao, Xin, et al.
Pubblicazione: (2025)
ProRefine: Inference-Time Prompt Refinement with Textual Feedback
di: Pandita, Deepak, et al.
Pubblicazione: (2025)
di: Pandita, Deepak, et al.
Pubblicazione: (2025)
Predicting Compact Phrasal Rewrites with Large Language Models for ASR Post Editing
di: Zhang, Hao, et al.
Pubblicazione: (2025)
di: Zhang, Hao, et al.
Pubblicazione: (2025)
No Training Wheels: Steering Vectors for Bias Correction at Inference Time
di: Gupta, Aviral, et al.
Pubblicazione: (2025)
di: Gupta, Aviral, et al.
Pubblicazione: (2025)
Modeling Real-Time Interactive Conversations as Timed Diarized Transcripts
di: Tanzer, Garrett, et al.
Pubblicazione: (2024)
di: Tanzer, Garrett, et al.
Pubblicazione: (2024)
FedLog: Personalized Federated Classification with Less Communication and More Flexibility
di: Yu, Haolin, et al.
Pubblicazione: (2024)
di: Yu, Haolin, et al.
Pubblicazione: (2024)
TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning
di: Zou, Jiaru, et al.
Pubblicazione: (2025)
di: Zou, Jiaru, et al.
Pubblicazione: (2025)
Time-MMD: Multi-Domain Multimodal Dataset for Time Series Analysis
di: Liu, Haoxin, et al.
Pubblicazione: (2024)
di: Liu, Haoxin, et al.
Pubblicazione: (2024)
Post-OCR Text Correction for Bulgarian Historical Documents
di: Beshirov, Angel, et al.
Pubblicazione: (2024)
di: Beshirov, Angel, et al.
Pubblicazione: (2024)
TiSpell: A Semi-Masked Methodology for Tibetan Spelling Correction covering Multi-Level Error with Data Augmentation
di: Liu, Yutong, et al.
Pubblicazione: (2025)
di: Liu, Yutong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
A Critical Look At Tokenwise Reward-Guided Text Generation
di: Rashid, Ahmad, et al.
Pubblicazione: (2024) -
Towards Cost-Effective Reward Guided Text Generation
di: Rashid, Ahmad, et al.
Pubblicazione: (2025) -
Grammar as a Behavioral Biometric: Using Cognitively Motivated Grammar Models for Authorship Verification
di: Nini, Andrea, et al.
Pubblicazione: (2024) -
Neural Grammatical Error Correction for Romanian
di: Cotet, Teodor-Mihai, et al.
Pubblicazione: (2026) -
Time Sensitive Knowledge Editing through Efficient Finetuning
di: Ge, Xiou, et al.
Pubblicazione: (2024)