Saved in:
| Main Authors: | Teoh, Jayden, Tomar, Manan, Ahn, Kwangjun, Hu, Edward S., Pearce, Tim, Sharma, Pratyusha, Krishnamurthy, Akshay, Islam, Riashat, Lamb, Alex, Langford, John |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.05963 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Belief State Transformer
by: Hu, Edward S., et al.
Published: (2024)
by: Hu, Edward S., et al.
Published: (2024)
Efficient Joint Prediction of Multiple Future Tokens
by: Ahn, Kwangjun, et al.
Published: (2025)
by: Ahn, Kwangjun, et al.
Published: (2025)
Dion: Distributed Orthonormalized Updates
by: Ahn, Kwangjun, et al.
Published: (2025)
by: Ahn, Kwangjun, et al.
Published: (2025)
Dion2: A Simple Method to Shrink Matrix in Muon
by: Ahn, Kwangjun, et al.
Published: (2025)
by: Ahn, Kwangjun, et al.
Published: (2025)
Improving Sampling for Masked Diffusion Models via Information Gain
by: Yang, Kaisen, et al.
Published: (2026)
by: Yang, Kaisen, et al.
Published: (2026)
Learning Latent Dynamic Robust Representations for World Models
by: Sun, Ruixiang, et al.
Published: (2024)
by: Sun, Ruixiang, et al.
Published: (2024)
Video Occupancy Models
by: Tomar, Manan, et al.
Published: (2024)
by: Tomar, Manan, et al.
Published: (2024)
Generalizing Multi-Step Inverse Models for Representation Learning to Finite-Memory POMDPs
by: Wu, Lili, et al.
Published: (2024)
by: Wu, Lili, et al.
Published: (2024)
PcLast: Discovering Plannable Continuous Latent States
by: Koul, Anurag, et al.
Published: (2023)
by: Koul, Anurag, et al.
Published: (2023)
Adam with model exponential moving average is effective for nonconvex optimization
by: Ahn, Kwangjun, et al.
Published: (2024)
by: Ahn, Kwangjun, et al.
Published: (2024)
On Generalization Across Environments In Multi-Objective Reinforcement Learning
by: Teoh, Jayden, et al.
Published: (2025)
by: Teoh, Jayden, et al.
Published: (2025)
Improving Environment Novelty Quantification for Effective Unsupervised Environment Design
by: Teoh, Jayden, et al.
Published: (2025)
by: Teoh, Jayden, et al.
Published: (2025)
When Does Predictive Inverse Dynamics Outperform Behavior Cloning?
by: Schäfer, Lukas, et al.
Published: (2026)
by: Schäfer, Lukas, et al.
Published: (2026)
How to escape sharp minima with random perturbations
by: Ahn, Kwangjun, et al.
Published: (2023)
by: Ahn, Kwangjun, et al.
Published: (2023)
General framework for online-to-nonconvex conversion: Schedule-free SGD is also effective for nonconvex optimization
by: Ahn, Kwangjun, et al.
Published: (2024)
by: Ahn, Kwangjun, et al.
Published: (2024)
Does SGD really happen in tiny subspaces?
by: Song, Minhak, et al.
Published: (2024)
by: Song, Minhak, et al.
Published: (2024)
Learning Additively Compositional Latent Actions for Embodied AI
by: Wei, Hangxing, et al.
Published: (2026)
by: Wei, Hangxing, et al.
Published: (2026)
Next Concept Prediction in Discrete Latent Space Leads to Stronger Language Models
by: Liu, Yuliang, et al.
Published: (2026)
by: Liu, Yuliang, et al.
Published: (2026)
Compact Matrix Quantum Group Equivariant Neural Networks
by: Pearce-Crump, Edward
Published: (2023)
by: Pearce-Crump, Edward
Published: (2023)
Computational-Statistical Tradeoffs at the Next-Token Prediction Barrier: Autoregressive and Imitation Learning under Misspecification
by: Rohatgi, Dhruv, et al.
Published: (2025)
by: Rohatgi, Dhruv, et al.
Published: (2025)
Towards Principled Representation Learning from Videos for Reinforcement Learning
by: Misra, Dipendra, et al.
Published: (2024)
by: Misra, Dipendra, et al.
Published: (2024)
Through the River: Understanding the Benefit of Schedule-Free Methods for Language Model Training
by: Song, Minhak, et al.
Published: (2025)
by: Song, Minhak, et al.
Published: (2025)
Understanding Adam Optimizer via Online Learning of Updates: Adam is FTRL in Disguise
by: Ahn, Kwangjun, et al.
Published: (2024)
by: Ahn, Kwangjun, et al.
Published: (2024)
A Unified Approach to Controlling Implicit Regularization via Mirror Descent
by: Sun, Haoyuan, et al.
Published: (2023)
by: Sun, Haoyuan, et al.
Published: (2023)
On Discovering Algorithms for Adversarial Imitation Learning
by: Chirra, Shashank Reddy, et al.
Published: (2025)
by: Chirra, Shashank Reddy, et al.
Published: (2025)
SLOPE: Optimistic Potential Landscape Shaping for Model-based Reinforcement Learning
by: Li, Yao-Hui, et al.
Published: (2026)
by: Li, Yao-Hui, et al.
Published: (2026)
Rich-Observation Reinforcement Learning with Continuous Latent Dynamics
by: Song, Yuda, et al.
Published: (2024)
by: Song, Yuda, et al.
Published: (2024)
On the Sample Complexity of Imitation Learning for Smoothed Model Predictive Control
by: Pfrommer, Daniel, et al.
Published: (2023)
by: Pfrommer, Daniel, et al.
Published: (2023)
Reward Centering
by: Naik, Abhishek, et al.
Published: (2024)
by: Naik, Abhishek, et al.
Published: (2024)
Improved Sample Complexity of Imitation Learning for Barrier Model Predictive Control
by: Pfrommer, Daniel, et al.
Published: (2024)
by: Pfrommer, Daniel, et al.
Published: (2024)
Wavelet Predictive Representations for Non-Stationary Reinforcement Learning
by: Wang, Min, et al.
Published: (2025)
by: Wang, Min, et al.
Published: (2025)
Unified Auto-Encoding with Masked Diffusion
by: Hansen-Estruch, Philippe, et al.
Published: (2024)
by: Hansen-Estruch, Philippe, et al.
Published: (2024)
Reinforcement Learning under Latent Dynamics: Toward Statistical and Algorithmic Modularity
by: Amortila, Philip, et al.
Published: (2024)
by: Amortila, Philip, et al.
Published: (2024)
Linear attention is (maybe) all you need (to understand transformer optimization)
by: Ahn, Kwangjun, et al.
Published: (2023)
by: Ahn, Kwangjun, et al.
Published: (2023)
LoRA vs Full Fine-tuning: An Illusion of Equivalence
by: Shuttleworth, Reece, et al.
Published: (2024)
by: Shuttleworth, Reece, et al.
Published: (2024)
Compress to Impress: Efficient LLM Adaptation Using a Single Gradient Step on 100 Samples
by: Sreeram, Shiva, et al.
Published: (2025)
by: Sreeram, Shiva, et al.
Published: (2025)
Always Learning, Always Mixing: Efficient and Simple Data Mixing All The Time
by: Hu, Michael Y., et al.
Published: (2026)
by: Hu, Michael Y., et al.
Published: (2026)
Adaptive and Explainable AI Agents for Anomaly Detection in Critical IoT Infrastructure using LLM-Enhanced Contextual Reasoning
by: Sharma, Raghav, et al.
Published: (2025)
by: Sharma, Raghav, et al.
Published: (2025)
Small Language Models for Agentic Systems: A Survey of Architectures, Capabilities, and Deployment Trade offs
by: Sharma, Raghav, et al.
Published: (2025)
by: Sharma, Raghav, et al.
Published: (2025)
Learning Fused State Representations for Control from Multi-View Observations
by: Wang, Zeyu, et al.
Published: (2025)
by: Wang, Zeyu, et al.
Published: (2025)
Similar Items
-
The Belief State Transformer
by: Hu, Edward S., et al.
Published: (2024) -
Efficient Joint Prediction of Multiple Future Tokens
by: Ahn, Kwangjun, et al.
Published: (2025) -
Dion: Distributed Orthonormalized Updates
by: Ahn, Kwangjun, et al.
Published: (2025) -
Dion2: A Simple Method to Shrink Matrix in Muon
by: Ahn, Kwangjun, et al.
Published: (2025) -
Improving Sampling for Masked Diffusion Models via Information Gain
by: Yang, Kaisen, et al.
Published: (2026)