Better World Models Can Lead to Better Post-Training Performance
Fuente:
arXiv
Saved in:
| Main Authors: | Gupta, Prakhar, Conklin, Henry, Leslie, Sarah-Jane, Lee, Andrew |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bootstrapped Mixed Rewards for RL Post-Training: Injecting Canonical Action Order
by: Gupta, Prakhar, et al.
Published: (2025)
by: Gupta, Prakhar, et al.
Published: (2025)
Does Biomedical Training Lead to Better Medical Performance?
by: Dada, Amin, et al.
Published: (2024)
by: Dada, Amin, et al.
Published: (2024)
Train Faster, Perform Better: Modular Adaptive Training in Over-Parameterized Models
by: Shi, Yubin, et al.
Published: (2024)
by: Shi, Yubin, et al.
Published: (2024)
Information Structure in Mappings: An Approach to Learning, Representation, and Generalisation
by: Conklin, Henry
Published: (2025)
by: Conklin, Henry
Published: (2025)
Better Decisions through the Right Causal World Model
by: Dillies, Elisabeth, et al.
Published: (2025)
by: Dillies, Elisabeth, et al.
Published: (2025)
Revisiting the Relationship between Adversarial and Clean Training: Why Clean Training Can Make Adversarial Training Better
by: Zhou, MingWei, et al.
Published: (2025)
by: Zhou, MingWei, et al.
Published: (2025)
Do Transformer World Models Give Better Policy Gradients?
by: Ma, Michel, et al.
Published: (2024)
by: Ma, Michel, et al.
Published: (2024)
When Does Multimodality Lead to Better Time Series Forecasting?
by: Zhang, Xiyuan, et al.
Published: (2025)
by: Zhang, Xiyuan, et al.
Published: (2025)
When Better Eyes Lead to Blindness: A Diagnostic Study of the Information Bottleneck in CNN-LSTM Image Captioning Models
by: Gupta, Hitesh Kumar
Published: (2025)
by: Gupta, Hitesh Kumar
Published: (2025)
Inconsistencies In Consistency Models: Better ODE Solving Does Not Imply Better Samples
by: Vouitsis, Noël, et al.
Published: (2024)
by: Vouitsis, Noël, et al.
Published: (2024)
Explaining the Unexplained: Revealing Hidden Correlations for Better Interpretability
by: Jiang, Wen-Dong, et al.
Published: (2024)
by: Jiang, Wen-Dong, et al.
Published: (2024)
Can Neural Networks Learn Small Algebraic Worlds? An Investigation Into the Group-theoretic Structures Learned By Narrow Models Trained To Predict Group Operations
by: Kvinge, Henry, et al.
Published: (2026)
by: Kvinge, Henry, et al.
Published: (2026)
When Mean CE Fails: Median CE Can Better Track Language Model Quality
by: Guo, Hao, et al.
Published: (2026)
by: Guo, Hao, et al.
Published: (2026)
RSQ: Learning from Important Tokens Leads to Better Quantized LLMs
by: Sung, Yi-Lin, et al.
Published: (2025)
by: Sung, Yi-Lin, et al.
Published: (2025)
Mixtraining: A Better Trade-Off Between Compute and Performance
by: Li, Zexin, et al.
Published: (2025)
by: Li, Zexin, et al.
Published: (2025)
TimeBridge: Better Diffusion Prior Design with Bridge Models for Time Series Generation
by: Park, Jinseong, et al.
Published: (2024)
by: Park, Jinseong, et al.
Published: (2024)
Sparser, Better, Deeper, Stronger: Improving Sparse Training with Exact Orthogonal Initialization
by: Nowak, Aleksandra Irena, et al.
Published: (2024)
by: Nowak, Aleksandra Irena, et al.
Published: (2024)
Compute-Optimal LLMs Provably Generalize Better With Scale
by: Finzi, Marc, et al.
Published: (2025)
by: Finzi, Marc, et al.
Published: (2025)
Lifelong Knowledge Editing requires Better Regularization
by: Gupta, Akshat, et al.
Published: (2025)
by: Gupta, Akshat, et al.
Published: (2025)
Improving Large Models with Small models: Lower Costs and Better Performance
by: Chen, Dong, et al.
Published: (2024)
by: Chen, Dong, et al.
Published: (2024)
What Makes Looped Transformers Perform Better Than Non-Recursive Ones
by: Gong, Zixuan, et al.
Published: (2025)
by: Gong, Zixuan, et al.
Published: (2025)
Is Bigger Edit Batch Size Always Better? -- An Empirical Study on Model Editing with Llama-3
by: Yoon, Junsang, et al.
Published: (2024)
by: Yoon, Junsang, et al.
Published: (2024)
ANO : Faster is Better in Noisy Landscape
by: Kegreisz, Adrien
Published: (2025)
by: Kegreisz, Adrien
Published: (2025)
Towards Better Generalization and Interpretability in Unsupervised Concept-Based Models
by: De Santis, Francesco, et al.
Published: (2025)
by: De Santis, Francesco, et al.
Published: (2025)
Feature Distillation is the Better Choice for Model-Heterogeneous Federated Learning
by: Li, Yichen, et al.
Published: (2025)
by: Li, Yichen, et al.
Published: (2025)
This Looks Better than That: Better Interpretable Models with ProtoPNeXt
by: Willard, Frank, et al.
Published: (2024)
by: Willard, Frank, et al.
Published: (2024)
Understanding Task Representations in Neural Networks via Bayesian Ablation
by: Nam, Andrew, et al.
Published: (2025)
by: Nam, Andrew, et al.
Published: (2025)
Does "Do Differentiable Simulators Give Better Policy Gradients?'' Give Better Policy Gradients?
by: Onoda, Ku, et al.
Published: (2026)
by: Onoda, Ku, et al.
Published: (2026)
Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation
by: Manvi, Rohin, et al.
Published: (2024)
by: Manvi, Rohin, et al.
Published: (2024)
A Method on Searching Better Activation Functions
by: Sun, Haoyuan, et al.
Published: (2024)
by: Sun, Haoyuan, et al.
Published: (2024)
Better Generative Replay for Continual Federated Learning
by: Qi, Daiqing, et al.
Published: (2023)
by: Qi, Daiqing, et al.
Published: (2023)
Merging Smarter, Generalizing Better: Enhancing Model Merging on OOD Data
by: Zhang, Bingjie, et al.
Published: (2025)
by: Zhang, Bingjie, et al.
Published: (2025)
Learning Explainable and Better Performing Representations of POMDP Strategies
by: Bork, Alexander, et al.
Published: (2024)
by: Bork, Alexander, et al.
Published: (2024)
Rethinking Data Curation in LLM Training: Online Reweighting Offers Better Generalization than Offline Methods
by: Zhao, Wanru, et al.
Published: (2026)
by: Zhao, Wanru, et al.
Published: (2026)
Tree-Structured Parzen Estimator: Understanding Its Algorithm Components and Their Roles for Better Empirical Performance
by: Watanabe, Shuhei
Published: (2023)
by: Watanabe, Shuhei
Published: (2023)
Better Later Than Sooner: Neuro-Symbolic Knowledge Graph Construction via Ontology-grounded Post-extraction Correction
by: Loconte, Lorenzo, et al.
Published: (2026)
by: Loconte, Lorenzo, et al.
Published: (2026)
Technology-assisted Personalized Yoga for Better Health -- Challenges and Outlook
by: Kumar, Vivek, et al.
Published: (2025)
by: Kumar, Vivek, et al.
Published: (2025)
Training Multimodal Large Reasoning Models Needs Better Thoughts: A Three-Stage Framework for Long Chain-of-Thought Synthesis and Selection
by: Wang, Yizhi, et al.
Published: (2025)
by: Wang, Yizhi, et al.
Published: (2025)
Breaking Silos: Adaptive Model Fusion Unlocks Better Time Series Forecasting
by: Liu, Zhining, et al.
Published: (2025)
by: Liu, Zhining, et al.
Published: (2025)
Do Large Language Models Reason Causally Like Us? Even Better?
by: Dettki, Hanna M., et al.
Published: (2025)
by: Dettki, Hanna M., et al.
Published: (2025)
Similar Items
-
Bootstrapped Mixed Rewards for RL Post-Training: Injecting Canonical Action Order
by: Gupta, Prakhar, et al.
Published: (2025) -
Does Biomedical Training Lead to Better Medical Performance?
by: Dada, Amin, et al.
Published: (2024) -
Train Faster, Perform Better: Modular Adaptive Training in Over-Parameterized Models
by: Shi, Yubin, et al.
Published: (2024) -
Information Structure in Mappings: An Approach to Learning, Representation, and Generalisation
by: Conklin, Henry
Published: (2025) -
Better Decisions through the Right Causal World Model
by: Dillies, Elisabeth, et al.
Published: (2025)