ScRPO: From Errors to Insights
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Lianrui, Lu, Dakuan, Shao, Jiawei, Li, Xuelong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Information Capacity: Evaluating the Efficiency of Large Language Models via Text Compression
por: Yuan, Cheng, et al.
Publicado: (2025)
por: Yuan, Cheng, et al.
Publicado: (2025)
Pipeline Parallelism is All You Need for Optimized Early-Exit Based Self-Speculative Decoding
por: Li, Ruanjun, et al.
Publicado: (2025)
por: Li, Ruanjun, et al.
Publicado: (2025)
The Law of Multi-Model Collaboration: Scaling Limits of Model Ensembling for Large Language Models
por: Lu, Dakuan, et al.
Publicado: (2025)
por: Lu, Dakuan, et al.
Publicado: (2025)
Silicon Bureaucracy and AI Test-Oriented Education: Contamination Sensitivity and Score Confidence in LLM Benchmarks
por: Song, Yiliang, et al.
Publicado: (2026)
por: Song, Yiliang, et al.
Publicado: (2026)
CreditAudit: 2$^\text{nd}$ Dimension for LLM Evaluation and Selection
por: Song, Yiliang, et al.
Publicado: (2026)
por: Song, Yiliang, et al.
Publicado: (2026)
Ruyi2 Technical Report
por: Song, Huan, et al.
Publicado: (2026)
por: Song, Huan, et al.
Publicado: (2026)
RPO-RAG: Aligning Small LLMs with Relation-aware Preference Optimization for Knowledge Graph Question Answering
por: Um, Kaehyun, et al.
Publicado: (2026)
por: Um, Kaehyun, et al.
Publicado: (2026)
Theoretical Foundations of Scaling Law in Familial Models
por: Song, Huan, et al.
Publicado: (2025)
por: Song, Huan, et al.
Publicado: (2025)
VLMQ: Token Saliency-Driven Post-Training Quantization for Vision-language Models
por: Xue, Yufei, et al.
Publicado: (2025)
por: Xue, Yufei, et al.
Publicado: (2025)
SentenceVAE: Enable Next-sentence Prediction for Large Language Models with Faster Speed, Higher Accuracy and Longer Context
por: An, Hongjun, et al.
Publicado: (2024)
por: An, Hongjun, et al.
Publicado: (2024)
SCP-116K: A High-Quality Problem-Solution Dataset and a Generalized Pipeline for Automated Extraction in the Higher Education Science Domain
por: Lu, Dakuan, et al.
Publicado: (2025)
por: Lu, Dakuan, et al.
Publicado: (2025)
The Death of Feature Engineering? BERT with Linguistic Features on SQuAD 2.0
por: Li, Jiawei, et al.
Publicado: (2024)
por: Li, Jiawei, et al.
Publicado: (2024)
Subtle Errors in Reasoning: Preference Learning via Error-injected Self-editing
por: Xu, Kaishuai, et al.
Publicado: (2024)
por: Xu, Kaishuai, et al.
Publicado: (2024)
A Study of In-Context-Learning-Based Text-to-SQL Errors
por: Shen, Jiawei, et al.
Publicado: (2025)
por: Shen, Jiawei, et al.
Publicado: (2025)
From LLM-anation to LLM-orchestrator: Coordinating Small Models for Data Labeling
por: Lu, Yao, et al.
Publicado: (2025)
por: Lu, Yao, et al.
Publicado: (2025)
On Linearizing Structured Data in Encoder-Decoder Language Models: Insights from Text-to-SQL
por: Shao, Yutong, et al.
Publicado: (2024)
por: Shao, Yutong, et al.
Publicado: (2024)
Error Typing for Smarter Rewards: Improving Process Reward Models with Error-Aware Hierarchical Supervision
por: Pala, Tej Deep, et al.
Publicado: (2025)
por: Pala, Tej Deep, et al.
Publicado: (2025)
On the Predictive Power of Representation Dispersion in Language Models
por: Li, Yanhong, et al.
Publicado: (2025)
por: Li, Yanhong, et al.
Publicado: (2025)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
por: Li, Yanhong, et al.
Publicado: (2025)
por: Li, Yanhong, et al.
Publicado: (2025)
Chunk-Distilled Language Modeling
por: Li, Yanhong, et al.
Publicado: (2024)
por: Li, Yanhong, et al.
Publicado: (2024)
ScIRGen: Synthesize Realistic and Large-Scale RAG Dataset for Scientific Research
por: Lin, Junyong, et al.
Publicado: (2025)
por: Lin, Junyong, et al.
Publicado: (2025)
ERAS: Evaluating the Robustness of Chinese NLP Models to Morphological Garden Path Errors
por: Li, Qinchan, et al.
Publicado: (2024)
por: Li, Qinchan, et al.
Publicado: (2024)
Logic-Regularized Verifier Elicits Reasoning from LLMs
por: Wang, Xinyu, et al.
Publicado: (2026)
por: Wang, Xinyu, et al.
Publicado: (2026)
Estimating the Error of Large Language Models at Pairwise Text Comparison
por: Li, Tianyi
Publicado: (2025)
por: Li, Tianyi
Publicado: (2025)
Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning
por: Gan, Zeyu, et al.
Publicado: (2025)
por: Gan, Zeyu, et al.
Publicado: (2025)
Loss-Aware Curriculum Learning for Chinese Grammatical Error Correction
por: Zhang, Ding, et al.
Publicado: (2024)
por: Zhang, Ding, et al.
Publicado: (2024)
DeepOmni: Towards Seamless and Smart Speech Interaction with Adaptive Modality-Specific MoE
por: Shao, Hang, et al.
Publicado: (2025)
por: Shao, Hang, et al.
Publicado: (2025)
Harnessing Rule-Based Reinforcement Learning for Enhanced Grammatical Error Correction
por: Li, Yilin, et al.
Publicado: (2025)
por: Li, Yilin, et al.
Publicado: (2025)
Enhancing LLM-Based Data Annotation with Error Decomposition
por: Xu, Zhen, et al.
Publicado: (2026)
por: Xu, Zhen, et al.
Publicado: (2026)
CL$^2$GEC: A Multi-Discipline Benchmark for Continual Learning in Chinese Literature Grammatical Error Correction
por: Qin, Shang, et al.
Publicado: (2025)
por: Qin, Shang, et al.
Publicado: (2025)
Word Order's Impacts: Insights from Reordering and Generation Analysis
por: Zhao, Qinghua, et al.
Publicado: (2024)
por: Zhao, Qinghua, et al.
Publicado: (2024)
E2CL: Exploration-based Error Correction Learning for Embodied Agents
por: Wang, Hanlin, et al.
Publicado: (2024)
por: Wang, Hanlin, et al.
Publicado: (2024)
ROME: Memorization Insights from Text, Logits and Representation
por: Li, Bo, et al.
Publicado: (2024)
por: Li, Bo, et al.
Publicado: (2024)
DSGram: Dynamic Weighting Sub-Metrics for Grammatical Error Correction in the Era of Large Language Models
por: Xie, Jinxiang, et al.
Publicado: (2024)
por: Xie, Jinxiang, et al.
Publicado: (2024)
Don't Fine-Tune, Decode: Syntax Error-Free Tool Use via Constrained Decoding
por: Zhang, Kexun, et al.
Publicado: (2023)
por: Zhang, Kexun, et al.
Publicado: (2023)
Corrections Meet Explanations: A Unified Framework for Explainable Grammatical Error Correction
por: Ye, Jingheng, et al.
Publicado: (2025)
por: Ye, Jingheng, et al.
Publicado: (2025)
Single-Pixel Vision-Language Model for Intrinsic Privacy-Preserving Behavioral Intelligence
por: An, Hongjun, et al.
Publicado: (2026)
por: An, Hongjun, et al.
Publicado: (2026)
Reassessing the Role of Chain-of-Thought in Sentiment Analysis: Insights and Limitations
por: Zheng, Kaiyuan, et al.
Publicado: (2025)
por: Zheng, Kaiyuan, et al.
Publicado: (2025)
Unveiling Effective In-Context Configurations for Image Captioning: An External & Internal Analysis
por: Li, Li, et al.
Publicado: (2025)
por: Li, Li, et al.
Publicado: (2025)
ErrorMap and ErrorAtlas: Charting the Failure Landscape of Large Language Models
por: Ashury-Tahan, Shir, et al.
Publicado: (2026)
por: Ashury-Tahan, Shir, et al.
Publicado: (2026)
Ejemplares similares
-
Information Capacity: Evaluating the Efficiency of Large Language Models via Text Compression
por: Yuan, Cheng, et al.
Publicado: (2025) -
Pipeline Parallelism is All You Need for Optimized Early-Exit Based Self-Speculative Decoding
por: Li, Ruanjun, et al.
Publicado: (2025) -
The Law of Multi-Model Collaboration: Scaling Limits of Model Ensembling for Large Language Models
por: Lu, Dakuan, et al.
Publicado: (2025) -
Silicon Bureaucracy and AI Test-Oriented Education: Contamination Sensitivity and Score Confidence in LLM Benchmarks
por: Song, Yiliang, et al.
Publicado: (2026) -
CreditAudit: 2$^\text{nd}$ Dimension for LLM Evaluation and Selection
por: Song, Yiliang, et al.
Publicado: (2026)