Multi-Token Residual Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Yufeng, Bao, Zishuo, Wang, Qian, Zhang, Zeshen, Zhang, Haoqi, Peng, Bowen, Li, Ang, Chalamala, Rahul, Lu, Yucheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Distilling Token-Trained Models into Byte-Level Models
von: Bao, Zishuo, et al.
Veröffentlicht: (2026)
von: Bao, Zishuo, et al.
Veröffentlicht: (2026)
A Random Forest-based Prediction Model for Turning Points in Antagonistic Event-Group Competitions
von: Zhu, Zishuo
Veröffentlicht: (2024)
von: Zhu, Zishuo
Veröffentlicht: (2024)
Efficient Residual Learning with Mixture-of-Experts for Universal Dexterous Grasping
von: Huang, Ziye, et al.
Veröffentlicht: (2024)
von: Huang, Ziye, et al.
Veröffentlicht: (2024)
Learning Variable-Length Tokenization for Generative Recommendation
von: Wang, Minhao, et al.
Veröffentlicht: (2026)
von: Wang, Minhao, et al.
Veröffentlicht: (2026)
Residual Reweighted Conformal Prediction for Graph Neural Networks
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
LoLCATs: On Low-Rank Linearizing of Large Language Models
von: Zhang, Michael, et al.
Veröffentlicht: (2024)
von: Zhang, Michael, et al.
Veröffentlicht: (2024)
PrismFlow: Residual Dynamics for Flow Matching in Time-Series Generation
von: Zhang, Junru, et al.
Veröffentlicht: (2026)
von: Zhang, Junru, et al.
Veröffentlicht: (2026)
Data-Driven Extended Corresponding State Approach for Residual Property Prediction of Hydrofluoroolefins
von: Wang, Gang, et al.
Veröffentlicht: (2025)
von: Wang, Gang, et al.
Veröffentlicht: (2025)
Flow Perturbation++: Multi-Step Unbiased Jacobian Estimation for High-Dimensional Boltzmann Sampling
von: Peng, Xin, et al.
Veröffentlicht: (2026)
von: Peng, Xin, et al.
Veröffentlicht: (2026)
Sequential Regression for Continuous Value Prediction using Residual Quantization
von: Cui, Runpeng, et al.
Veröffentlicht: (2026)
von: Cui, Runpeng, et al.
Veröffentlicht: (2026)
AlignDistil: Token-Level Language Model Alignment as Adaptive Policy Distillation
von: Zhang, Songming, et al.
Veröffentlicht: (2025)
von: Zhang, Songming, et al.
Veröffentlicht: (2025)
Self-Distillation for Multi-Token Prediction
von: Zhao, Guoliang, et al.
Veröffentlicht: (2026)
von: Zhao, Guoliang, et al.
Veröffentlicht: (2026)
Predictive Auditing of Hidden Tokens in LLM APIs via Reasoning Length Estimation
von: Wang, Ziyao, et al.
Veröffentlicht: (2025)
von: Wang, Ziyao, et al.
Veröffentlicht: (2025)
Conformal Prediction with Cellwise Outliers: A Detect-then-Impute Approach
von: Peng, Qian, et al.
Veröffentlicht: (2025)
von: Peng, Qian, et al.
Veröffentlicht: (2025)
Guided Cooperation in Hierarchical Reinforcement Learning via Model-based Rollout
von: Wang, Haoran, et al.
Veröffentlicht: (2023)
von: Wang, Haoran, et al.
Veröffentlicht: (2023)
Reasoning Bias of Next Token Prediction Training
von: Lin, Pengxiao, et al.
Veröffentlicht: (2025)
von: Lin, Pengxiao, et al.
Veröffentlicht: (2025)
Amortized Predictability-aware Training Framework for Time Series Forecasting and Classification
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
Dynamic Model Selection for Trajectory Prediction via Pairwise Ranking and Meta-Features
von: Bowen, Lu
Veröffentlicht: (2025)
von: Bowen, Lu
Veröffentlicht: (2025)
Superlinear Multi-Step Attention
von: Huang, Yufeng
Veröffentlicht: (2026)
von: Huang, Yufeng
Veröffentlicht: (2026)
HG2P: Hippocampus-inspired High-reward Graph and Model-Free Q-Gradient Penalty for Path Planning and Motion Control
von: Wang, Haoran, et al.
Veröffentlicht: (2024)
von: Wang, Haoran, et al.
Veröffentlicht: (2024)
CodeBrain: Bridging Decoupled Tokenizer and Multi-Scale Architecture for EEG Foundation Model
von: Ma, Jingying, et al.
Veröffentlicht: (2025)
von: Ma, Jingying, et al.
Veröffentlicht: (2025)
PirateNets: Physics-informed Deep Learning with Residual Adaptive Networks
von: Wang, Sifan, et al.
Veröffentlicht: (2024)
von: Wang, Sifan, et al.
Veröffentlicht: (2024)
MergeDNA: Context-aware Genome Modeling with Dynamic Tokenization through Token Merging
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
Fast and Expressive Multi-Token Prediction with Probabilistic Circuits
von: Grivas, Andreas, et al.
Veröffentlicht: (2025)
von: Grivas, Andreas, et al.
Veröffentlicht: (2025)
Efficient Generative Modeling with Residual Vector Quantization-Based Tokens
von: Kim, Jaehyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jaehyeon, et al.
Veröffentlicht: (2024)
PRISM: Parallel Residual Iterative Sequence Model
von: Jiang, Jie, et al.
Veröffentlicht: (2026)
von: Jiang, Jie, et al.
Veröffentlicht: (2026)
Global-Lens Transformers: Adaptive Token Mixing for Dynamic Link Prediction
von: Zou, Tao, et al.
Veröffentlicht: (2025)
von: Zou, Tao, et al.
Veröffentlicht: (2025)
Trajeglish: Traffic Modeling as Next-Token Prediction
von: Philion, Jonah, et al.
Veröffentlicht: (2023)
von: Philion, Jonah, et al.
Veröffentlicht: (2023)
Theory Foundation of Physics-Enhanced Residual Learning
von: Liang, Shixiao, et al.
Veröffentlicht: (2025)
von: Liang, Shixiao, et al.
Veröffentlicht: (2025)
Joint Training of Multi-Token Prediction in Reinforcement Learning via Optimal Coefficient Calibration
von: Wang, Zili, et al.
Veröffentlicht: (2026)
von: Wang, Zili, et al.
Veröffentlicht: (2026)
Multi-Token Prediction via Self-Distillation
von: Kirchenbauer, John, et al.
Veröffentlicht: (2026)
von: Kirchenbauer, John, et al.
Veröffentlicht: (2026)
A General ReLearner: Empowering Spatiotemporal Prediction by Re-learning Input-label Residual
von: Ma, Jiaming, et al.
Veröffentlicht: (2026)
von: Ma, Jiaming, et al.
Veröffentlicht: (2026)
Creative Agents: Empowering Agents with Imagination for Creative Tasks
von: Cai, Penglin, et al.
Veröffentlicht: (2023)
von: Cai, Penglin, et al.
Veröffentlicht: (2023)
Cautious Next Token Prediction
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
TEFL: Prediction-Residual-Guided Rolling Forecasting for Multi-Horizon Time Series
von: Huang, Xiannan, et al.
Veröffentlicht: (2026)
von: Huang, Xiannan, et al.
Veröffentlicht: (2026)
From Residuals to Reasons: LLM-Guided Mechanism Inference from Tabular Data
von: Rezaei, Mohammad R., et al.
Veröffentlicht: (2026)
von: Rezaei, Mohammad R., et al.
Veröffentlicht: (2026)
Adaptive Tokenization: On the Hop-Overpriority Problem in Tokenized Graph Learning Models
von: Wang, Zhibiao, et al.
Veröffentlicht: (2025)
von: Wang, Zhibiao, et al.
Veröffentlicht: (2025)
Generative Verifiers: Reward Modeling as Next-Token Prediction
von: Zhang, Lunjun, et al.
Veröffentlicht: (2024)
von: Zhang, Lunjun, et al.
Veröffentlicht: (2024)
DeepForm: Reasoning Large Language Model for Communication System Formulation
von: Wu, Panlong, et al.
Veröffentlicht: (2025)
von: Wu, Panlong, et al.
Veröffentlicht: (2025)
Learning Diverse Bimanual Dexterous Manipulation Skills from Human Demonstrations
von: Zhou, Bohan, et al.
Veröffentlicht: (2024)
von: Zhou, Bohan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Distilling Token-Trained Models into Byte-Level Models
von: Bao, Zishuo, et al.
Veröffentlicht: (2026) -
A Random Forest-based Prediction Model for Turning Points in Antagonistic Event-Group Competitions
von: Zhu, Zishuo
Veröffentlicht: (2024) -
Efficient Residual Learning with Mixture-of-Experts for Universal Dexterous Grasping
von: Huang, Ziye, et al.
Veröffentlicht: (2024) -
Learning Variable-Length Tokenization for Generative Recommendation
von: Wang, Minhao, et al.
Veröffentlicht: (2026) -
Residual Reweighted Conformal Prediction for Graph Neural Networks
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)