CoVeR: Conformal Calibration for Versatile and Reliable Autoregressive Next-Token Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Yuzhu, Wang, Yingjie, Liu, Shunyu, Jing, Yongcheng, Tao, Dacheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Distillation Traps and Guards: A Calibration Knob for LLM Distillability
von: Zhan, Weixiao, et al.
Veröffentlicht: (2026)
von: Zhan, Weixiao, et al.
Veröffentlicht: (2026)
A Theoretical Survey on Foundation Models
von: Fu, Shi, et al.
Veröffentlicht: (2024)
von: Fu, Shi, et al.
Veröffentlicht: (2024)
A Theoretical Perspective: How to Prevent Model Collapse in Self-consuming Training Loops
von: Fu, Shi, et al.
Veröffentlicht: (2025)
von: Fu, Shi, et al.
Veröffentlicht: (2025)
Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models
von: Kim, Minseo, et al.
Veröffentlicht: (2025)
von: Kim, Minseo, et al.
Veröffentlicht: (2025)
Physics-Guided Multimodal Transformers are the Necessary Foundation for the Next Generation of Meteorological Science
von: Han, Jing, et al.
Veröffentlicht: (2025)
von: Han, Jing, et al.
Veröffentlicht: (2025)
SPA-Cache: Singular Proxies for Adaptive Caching in Diffusion Language Models
von: Sun, Wenhao, et al.
Veröffentlicht: (2026)
von: Sun, Wenhao, et al.
Veröffentlicht: (2026)
Cautious Next Token Prediction
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
Efficient Differentiable Causal Discovery via Reliable Super-Structure Learning
von: Ma, Pingchuan, et al.
Veröffentlicht: (2026)
von: Ma, Pingchuan, et al.
Veröffentlicht: (2026)
Provable Long-Range Benefits of Next-Token Prediction
von: Cao, Xinyuan, et al.
Veröffentlicht: (2025)
von: Cao, Xinyuan, et al.
Veröffentlicht: (2025)
Drawback of Enforcing Equivariance and its Compensation via the Lens of Expressive Power
von: Chen, Yuzhu, et al.
Veröffentlicht: (2025)
von: Chen, Yuzhu, et al.
Veröffentlicht: (2025)
Intra-Trajectory Consistency for Reward Modeling
von: Zhou, Chaoyang, et al.
Veröffentlicht: (2025)
von: Zhou, Chaoyang, et al.
Veröffentlicht: (2025)
Towards Theoretical Understandings of Self-Consuming Generative Models
von: Fu, Shi, et al.
Veröffentlicht: (2024)
von: Fu, Shi, et al.
Veröffentlicht: (2024)
HRP: High-Rank Preheating for Superior LoRA Initialization
von: Chen, Yuzhu, et al.
Veröffentlicht: (2025)
von: Chen, Yuzhu, et al.
Veröffentlicht: (2025)
Factorize to Generalize: Retrieval-Guided Invariant-Dynamic Decomposition for Time Series Forecasting
von: Chi, Jinjin, et al.
Veröffentlicht: (2026)
von: Chi, Jinjin, et al.
Veröffentlicht: (2026)
Onboard Out-of-Calibration Detection of Deep Learning Models using Conformal Prediction
von: Bhattacharjee, Protim, et al.
Veröffentlicht: (2024)
von: Bhattacharjee, Protim, et al.
Veröffentlicht: (2024)
Next-Token Prediction Should be Ambiguity-Sensitive: A Meta-Learning Perspective
von: Gagnon, Leo, et al.
Veröffentlicht: (2025)
von: Gagnon, Leo, et al.
Veröffentlicht: (2025)
Enhancing Adversarial Robustness with Conformal Prediction: A Framework for Guaranteed Model Reliability
von: Bao, Jie, et al.
Veröffentlicht: (2025)
von: Bao, Jie, et al.
Veröffentlicht: (2025)
Next-Token Prediction and Regret Minimization
von: Mohri, Mehryar, et al.
Veröffentlicht: (2026)
von: Mohri, Mehryar, et al.
Veröffentlicht: (2026)
R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
von: Zhang, Jingyi, et al.
Veröffentlicht: (2025)
von: Zhang, Jingyi, et al.
Veröffentlicht: (2025)
Conformal Feedback Alignment: Quantifying Answer-Level Reliability for Robust LLM Alignment
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
LLMC: Benchmarking Large Language Model Quantization with a Versatile Compression Toolkit
von: Gong, Ruihao, et al.
Veröffentlicht: (2024)
von: Gong, Ruihao, et al.
Veröffentlicht: (2024)
A Law of Next-Token Prediction in Large Language Models
von: He, Hangfeng, et al.
Veröffentlicht: (2024)
von: He, Hangfeng, et al.
Veröffentlicht: (2024)
Reliable Real-Time Value at Risk Estimation via Quantile Regression Forest with Conformal Calibration
von: Wang, Du-Yi, et al.
Veröffentlicht: (2026)
von: Wang, Du-Yi, et al.
Veröffentlicht: (2026)
Black-Box Reliability Certification for AI Agents via Self-Consistency Sampling and Conformal Calibration
von: Mouzouni, Charafeddine
Veröffentlicht: (2026)
von: Mouzouni, Charafeddine
Veröffentlicht: (2026)
A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases
von: Lombardo, Gianfranco, et al.
Veröffentlicht: (2026)
von: Lombardo, Gianfranco, et al.
Veröffentlicht: (2026)
Genomic Next-Token Predictors are In-Context Learners
von: Breslow, Nathan, et al.
Veröffentlicht: (2025)
von: Breslow, Nathan, et al.
Veröffentlicht: (2025)
Graph-Augmented Reasoning: Evolving Step-by-Step Knowledge Graph Retrieval for LLM Reasoning
von: Wu, Wenjie, et al.
Veröffentlicht: (2025)
von: Wu, Wenjie, et al.
Veröffentlicht: (2025)
Mechanics of Next Token Prediction with Self-Attention
von: Li, Yingcong, et al.
Veröffentlicht: (2024)
von: Li, Yingcong, et al.
Veröffentlicht: (2024)
GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
von: Ghasemi, Narges, et al.
Veröffentlicht: (2025)
von: Ghasemi, Narges, et al.
Veröffentlicht: (2025)
High-Resolution Image Synthesis via Next-Token Prediction
von: Chen, Dengsheng, et al.
Veröffentlicht: (2024)
von: Chen, Dengsheng, et al.
Veröffentlicht: (2024)
Dynamics of Spontaneous Topic Changes in Next Token Prediction with Self-Attention
von: Jia, Mumin, et al.
Veröffentlicht: (2025)
von: Jia, Mumin, et al.
Veröffentlicht: (2025)
Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
Enabling Autoregressive Models to Fill In Masked Tokens
von: Israel, Daniel, et al.
Veröffentlicht: (2025)
von: Israel, Daniel, et al.
Veröffentlicht: (2025)
When Can We Reuse a Calibration Set for Multiple Conformal Predictions?
von: Balinsky, A. A., et al.
Veröffentlicht: (2025)
von: Balinsky, A. A., et al.
Veröffentlicht: (2025)
Manifold Trajectories in Next-Token Prediction: From Replicator Dynamics to Softmax Equilibrium
von: Lee-Jenkins, Christopher R.
Veröffentlicht: (2025)
von: Lee-Jenkins, Christopher R.
Veröffentlicht: (2025)
Beyond Next Token Prediction: Patch-Level Training for Large Language Models
von: Shao, Chenze, et al.
Veröffentlicht: (2024)
von: Shao, Chenze, et al.
Veröffentlicht: (2024)
Context Dependence and Reliability in Autoregressive Language Models
von: Sengupta, Poushali, et al.
Veröffentlicht: (2026)
von: Sengupta, Poushali, et al.
Veröffentlicht: (2026)
Enhancing Deep Neural Network Reliability with Refinement and Calibration
von: Hebbalaguppe, Ramya, et al.
Veröffentlicht: (2026)
von: Hebbalaguppe, Ramya, et al.
Veröffentlicht: (2026)
Explainable Graph Spectral Clustering For GloVe-like Text Embeddings
von: Kłopotek, Mieczysław A., et al.
Veröffentlicht: (2025)
von: Kłopotek, Mieczysław A., et al.
Veröffentlicht: (2025)
Continual Diffuser (CoD): Mastering Continual Offline Reinforcement Learning with Experience Rehearsal
von: Hu, Jifeng, et al.
Veröffentlicht: (2024)
von: Hu, Jifeng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Distillation Traps and Guards: A Calibration Knob for LLM Distillability
von: Zhan, Weixiao, et al.
Veröffentlicht: (2026) -
A Theoretical Survey on Foundation Models
von: Fu, Shi, et al.
Veröffentlicht: (2024) -
A Theoretical Perspective: How to Prevent Model Collapse in Self-consuming Training Loops
von: Fu, Shi, et al.
Veröffentlicht: (2025) -
Beyond Next-Token Prediction: A Performance Characterization of Diffusion versus Autoregressive Language Models
von: Kim, Minseo, et al.
Veröffentlicht: (2025) -
Physics-Guided Multimodal Transformers are the Necessary Foundation for the Next Generation of Meteorological Science
von: Han, Jing, et al.
Veröffentlicht: (2025)