Saved in:
| Main Authors: | Hao, Sophie, Merrill, William |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.16430 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast and Accurate Probing of In-Training LLMs' Downstream Performances
by: Liu, Zhichen, et al.
Published: (2026)
by: Liu, Zhichen, et al.
Published: (2026)
Smoothing DiLoCo with Primal Averaging for Faster Training of LLMs
by: Defazio, Aaron, et al.
Published: (2025)
by: Defazio, Aaron, et al.
Published: (2025)
Discovering Interpretable Algorithms by Decompiling Transformers to RASP
by: Huang, Xinting, et al.
Published: (2026)
by: Huang, Xinting, et al.
Published: (2026)
Learning to Solve Orienteering Problem with Time Windows and Variable Profits
by: Gao, Songqun, et al.
Published: (2026)
by: Gao, Songqun, et al.
Published: (2026)
Between Circuits and Chomsky: Pre-pretraining on Formal Languages Imparts Linguistic Biases
by: Hu, Michael Y., et al.
Published: (2025)
by: Hu, Michael Y., et al.
Published: (2025)
Compute-Optimal Quantization-Aware Training
by: Dremov, Aleksandr, et al.
Published: (2025)
by: Dremov, Aleksandr, et al.
Published: (2025)
Training LLMs for Honesty via Confessions
by: Joglekar, Manas, et al.
Published: (2025)
by: Joglekar, Manas, et al.
Published: (2025)
Path Regularization: A Near-Complete and Optimal Nonasymptotic Generalization Theory for Multilayer Neural Networks and Double Descent Phenomenon
by: Yu, Hao
Published: (2025)
by: Yu, Hao
Published: (2025)
Towards Effective Theory of LLMs: A Representation Learning Approach
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
Multi-Agent Reinforcement Learning for Dynamic Pricing: Balancing Profitability,Stability and Fairness
by: Amma, Krishna Kumar Neelakanta Pillai Santha Kumari
Published: (2026)
by: Amma, Krishna Kumar Neelakanta Pillai Santha Kumari
Published: (2026)
A Comedy of Estimators: On KL Regularization in RL Training of LLMs
by: Shah, Vedant, et al.
Published: (2025)
by: Shah, Vedant, et al.
Published: (2025)
Model Collapse Is Not a Bug but a Feature in Machine Unlearning for LLMs
by: Scholten, Yan, et al.
Published: (2025)
by: Scholten, Yan, et al.
Published: (2025)
Removing Sandbagging in LLMs by Training with Weak Supervision
by: Ryd, Emil, et al.
Published: (2026)
by: Ryd, Emil, et al.
Published: (2026)
Compute-Optimal LLMs Provably Generalize Better With Scale
by: Finzi, Marc, et al.
Published: (2025)
by: Finzi, Marc, et al.
Published: (2025)
Near-Optimal Online Deployment and Routing for Streaming LLMs
by: Li, Shaoang, et al.
Published: (2025)
by: Li, Shaoang, et al.
Published: (2025)
Are Language Models Actually Useful for Time Series Forecasting?
by: Tan, Mingtian, et al.
Published: (2024)
by: Tan, Mingtian, et al.
Published: (2024)
Pre-Training LLMs on a budget: A comparison of three optimizers
by: Schlotthauer, Joel, et al.
Published: (2025)
by: Schlotthauer, Joel, et al.
Published: (2025)
Optimal Corpus Aware Training for Neural Machine Translation
by: Liao, Yi-Hsiu, et al.
Published: (2025)
by: Liao, Yi-Hsiu, et al.
Published: (2025)
Training Free Guided Flow Matching with Optimal Control
by: Wang, Luran, et al.
Published: (2024)
by: Wang, Luran, et al.
Published: (2024)
Leveraging Fundamental Analysis for Stock Trend Prediction for Profit
by: Phan, John, et al.
Published: (2024)
by: Phan, John, et al.
Published: (2024)
Trained Persistent Memory for Frozen Decoder-Only LLMs
by: Jeong, Hong
Published: (2026)
by: Jeong, Hong
Published: (2026)
AdapterSwap: Continuous Training of LLMs with Data Removal and Access-Control Guarantees
by: Fleshman, William, et al.
Published: (2024)
by: Fleshman, William, et al.
Published: (2024)
Multi-Objective Bayesian Optimization for Networked Black-Box Systems: A Path to Greener Profits and Smarter Designs
by: Kudva, Akshay, et al.
Published: (2025)
by: Kudva, Akshay, et al.
Published: (2025)
Post-Training with Policy Gradients: Optimality and the Base Model Barrier
by: Mousavi-Hosseini, Alireza, et al.
Published: (2026)
by: Mousavi-Hosseini, Alireza, et al.
Published: (2026)
Efficient Pre-Training of LLMs through Truncated SVD Layers
by: Kamali, Kaivan, et al.
Published: (2026)
by: Kamali, Kaivan, et al.
Published: (2026)
Test-Time Training on Graphs with Large Language Models (LLMs)
by: Zhang, Jiaxin, et al.
Published: (2024)
by: Zhang, Jiaxin, et al.
Published: (2024)
FP4 All the Way: Fully Quantized Training of LLMs
by: Chmiel, Brian, et al.
Published: (2025)
by: Chmiel, Brian, et al.
Published: (2025)
Imbalanced Gradients in RL Post-Training of Multi-Task LLMs
by: Wu, Runzhe, et al.
Published: (2025)
by: Wu, Runzhe, et al.
Published: (2025)
Smart Profit-Aware Crop Advisory System: Kisan AI
by: Dwibedy, Debasis, et al.
Published: (2026)
by: Dwibedy, Debasis, et al.
Published: (2026)
ENOT: Expectile Regularization for Fast and Accurate Training of Neural Optimal Transport
by: Buzun, Nazar, et al.
Published: (2024)
by: Buzun, Nazar, et al.
Published: (2024)
Achieving binary weight and activation for LLMs using Post-Training Quantization
by: Song, Siqing, et al.
Published: (2025)
by: Song, Siqing, et al.
Published: (2025)
DAQ: Density-Aware Post-Training Weight-Only Quantization For LLMs
by: Luo, Yingsong, et al.
Published: (2024)
by: Luo, Yingsong, et al.
Published: (2024)
Sub-Scaling Laws: On the Role of Data Density and Training Strategies in LLMs
by: Chen, Zhengyu, et al.
Published: (2025)
by: Chen, Zhengyu, et al.
Published: (2025)
Training Long-Context LLMs Efficiently via Chunk-wise Optimization
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
End-to-End On-Device Quantization-Aware Training for LLMs at Inference Cost
by: Tan, Qitao, et al.
Published: (2025)
by: Tan, Qitao, et al.
Published: (2025)
Relational Learning in Pre-Trained Models: A Theory from Hypergraph Recovery Perspective
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
Balancing Profit and Fairness in Risk-Based Pricing Markets
by: Thibodeau, Jesse, et al.
Published: (2025)
by: Thibodeau, Jesse, et al.
Published: (2025)
Fairy2i: Training Complex LLMs from Real LLMs with All Parameters in $\{\pm 1, \pm i\}$
by: Wang, Feiyu, et al.
Published: (2025)
by: Wang, Feiyu, et al.
Published: (2025)
Reshaping Reasoning in LLMs: A Theoretical Analysis of RL Training Dynamics through Pattern Selection
by: Chen, Xingwu, et al.
Published: (2025)
by: Chen, Xingwu, et al.
Published: (2025)
How to Train Data-Efficient LLMs
by: Sachdeva, Noveen, et al.
Published: (2024)
by: Sachdeva, Noveen, et al.
Published: (2024)
Similar Items
-
Fast and Accurate Probing of In-Training LLMs' Downstream Performances
by: Liu, Zhichen, et al.
Published: (2026) -
Smoothing DiLoCo with Primal Averaging for Faster Training of LLMs
by: Defazio, Aaron, et al.
Published: (2025) -
Discovering Interpretable Algorithms by Decompiling Transformers to RASP
by: Huang, Xinting, et al.
Published: (2026) -
Learning to Solve Orienteering Problem with Time Windows and Variable Profits
by: Gao, Songqun, et al.
Published: (2026) -
Between Circuits and Chomsky: Pre-pretraining on Formal Languages Imparts Linguistic Biases
by: Hu, Michael Y., et al.
Published: (2025)