Annotations Mitigate Post-Training Mode Collapse
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Springer, Jacob Mitchell, Advani, Madhu, Aichberger, Lukas, Bradley, Arwen, Malach, Eran, Saremi, Omid, Williamson, Sinead, Nakkiran, Preetum, Littwin, Etai, Raghunathan, Aditi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How JEPA Avoids Noisy Features: The Implicit Bias of Deep Linear Self Distillation Networks
von: Littwin, Etai, et al.
Veröffentlicht: (2024)
von: Littwin, Etai, et al.
Veröffentlicht: (2024)
To Infinity and Beyond: Tool-Use Unlocks Length Generalization in State Space Models
von: Malach, Eran, et al.
Veröffentlicht: (2025)
von: Malach, Eran, et al.
Veröffentlicht: (2025)
Step-by-Step Diffusion: An Elementary Tutorial
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2024)
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2024)
Vanishing Gradients in Reinforcement Finetuning of Language Models
von: Razin, Noam, et al.
Veröffentlicht: (2023)
von: Razin, Noam, et al.
Veröffentlicht: (2023)
Classifier-Free Guidance is a Predictor-Corrector
von: Bradley, Arwen, et al.
Veröffentlicht: (2024)
von: Bradley, Arwen, et al.
Veröffentlicht: (2024)
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
Trace Length is a Simple Uncertainty Signal in Reasoning Models
von: Devic, Siddartha, et al.
Veröffentlicht: (2025)
von: Devic, Siddartha, et al.
Veröffentlicht: (2025)
Mechanisms of Projective Composition of Diffusion Models
von: Bradley, Arwen, et al.
Veröffentlicht: (2025)
von: Bradley, Arwen, et al.
Veröffentlicht: (2025)
Composition and Control with Distilled Energy Diffusion Models and Sequential Monte Carlo
von: Thornton, James, et al.
Veröffentlicht: (2025)
von: Thornton, James, et al.
Veröffentlicht: (2025)
When can transformers reason with abstract symbols?
von: Boix-Adsera, Enric, et al.
Veröffentlicht: (2023)
von: Boix-Adsera, Enric, et al.
Veröffentlicht: (2023)
Rethinking JEPA: Compute-Efficient Video SSL with Frozen Teachers
von: Li, Xianhang, et al.
Veröffentlicht: (2025)
von: Li, Xianhang, et al.
Veröffentlicht: (2025)
Mitigating Bias in RAG: Controlling the Embedder
von: Kim, Taeyoun, et al.
Veröffentlicht: (2025)
von: Kim, Taeyoun, et al.
Veröffentlicht: (2025)
When is Multicalibration Post-Processing Necessary?
von: Hansen, Dutch, et al.
Veröffentlicht: (2024)
von: Hansen, Dutch, et al.
Veröffentlicht: (2024)
Enhancing JEPAs with Spatial Conditioning: Robust and Efficient Representation Learning
von: Littwin, Etai, et al.
Veröffentlicht: (2024)
von: Littwin, Etai, et al.
Veröffentlicht: (2024)
Sharpness-Aware Pretraining Mitigates Catastrophic Forgetting
von: Watts, Ishaan, et al.
Veröffentlicht: (2026)
von: Watts, Ishaan, et al.
Veröffentlicht: (2026)
Auto-Regressive Next-Token Predictors are Universal Learners
von: Malach, Eran
Veröffentlicht: (2023)
von: Malach, Eran
Veröffentlicht: (2023)
Sharpness-Aware Minimization Enhances Feature Quality via Balanced Learning
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2024)
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2024)
Understanding Catastrophic Forgetting in Language Models via Implicit Inference
von: Kotha, Suhas, et al.
Veröffentlicht: (2023)
von: Kotha, Suhas, et al.
Veröffentlicht: (2023)
Local Mechanisms of Compositional Generalization in Conditional Diffusion
von: Bradley, Arwen
Veröffentlicht: (2025)
von: Bradley, Arwen
Veröffentlicht: (2025)
The Power of Random Features and the Limits of Distribution-Free Gradient Descent
von: Karchmer, Ari, et al.
Veröffentlicht: (2025)
von: Karchmer, Ari, et al.
Veröffentlicht: (2025)
Text-Conditional JEPA for Learning Semantically Rich Visual Representations
von: Huang, Chen, et al.
Veröffentlicht: (2026)
von: Huang, Chen, et al.
Veröffentlicht: (2026)
UI-JEPA: Towards Active Perception of User Intent through Onscreen User Activity
von: Fu, Yicheng, et al.
Veröffentlicht: (2024)
von: Fu, Yicheng, et al.
Veröffentlicht: (2024)
Uncertainty Quantification for LLM Function-Calling
von: Ye, Zihuiwen, et al.
Veröffentlicht: (2026)
von: Ye, Zihuiwen, et al.
Veröffentlicht: (2026)
Repetition Improves Language Model Embeddings
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2024)
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2024)
Understanding the Influence of Synthetic Data for Text Embedders
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2025)
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2025)
Unlocking the Working Memory of Large Language Models for Latent Reasoning
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
Early Data Exposure Improves Robustness to Subsequent Fine-Tuning
von: Feng, Lawrence, et al.
Veröffentlicht: (2026)
von: Feng, Lawrence, et al.
Veröffentlicht: (2026)
Distillation Scaling Laws
von: Busbridge, Dan, et al.
Veröffentlicht: (2025)
von: Busbridge, Dan, et al.
Veröffentlicht: (2025)
Mitigating Modal Imbalance in Multimodal Reasoning
von: Wu, Chen Henry, et al.
Veröffentlicht: (2025)
von: Wu, Chen Henry, et al.
Veröffentlicht: (2025)
Mode-Conditioning Unlocks Superior Test-Time Scaling
von: Wu, Chen Henry, et al.
Veröffentlicht: (2025)
von: Wu, Chen Henry, et al.
Veröffentlicht: (2025)
Benign, Tempered, or Catastrophic: A Taxonomy of Overfitting
von: Mallinar, Neil, et al.
Veröffentlicht: (2022)
von: Mallinar, Neil, et al.
Veröffentlicht: (2022)
Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining
von: Zhao, Rosie, et al.
Veröffentlicht: (2025)
von: Zhao, Rosie, et al.
Veröffentlicht: (2025)
Differential Smoothing Mitigates Sharpening and Improves LLM Reasoning
von: Gai, Jingchu, et al.
Veröffentlicht: (2025)
von: Gai, Jingchu, et al.
Veröffentlicht: (2025)
Rethinking Uncertainty Estimation in LLMs: A Principled Single-Sequence Measure
von: Aichberger, Lukas, et al.
Veröffentlicht: (2024)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2024)
Self-Supervised Learning with Gaussian Processes
von: Duan, Yunshan, et al.
Veröffentlicht: (2025)
von: Duan, Yunshan, et al.
Veröffentlicht: (2025)
Posterior Uncertainty Quantification in Neural Networks using Data Augmentation
von: Wu, Luhuan, et al.
Veröffentlicht: (2024)
von: Wu, Luhuan, et al.
Veröffentlicht: (2024)
Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs
von: Zhong, Ziqian, et al.
Veröffentlicht: (2025)
von: Zhong, Ziqian, et al.
Veröffentlicht: (2025)
A Taxonomy of Transcendence
von: Abreu, Natalie, et al.
Veröffentlicht: (2025)
von: Abreu, Natalie, et al.
Veröffentlicht: (2025)
LLM Priors for ERM over Programs
von: Singhal, Shivam, et al.
Veröffentlicht: (2025)
von: Singhal, Shivam, et al.
Veröffentlicht: (2025)
How Reinforcement Learning After Next-Token Prediction Facilitates Learning
von: Tsilivis, Nikolaos, et al.
Veröffentlicht: (2025)
von: Tsilivis, Nikolaos, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
How JEPA Avoids Noisy Features: The Implicit Bias of Deep Linear Self Distillation Networks
von: Littwin, Etai, et al.
Veröffentlicht: (2024) -
To Infinity and Beyond: Tool-Use Unlocks Length Generalization in State Space Models
von: Malach, Eran, et al.
Veröffentlicht: (2025) -
Step-by-Step Diffusion: An Elementary Tutorial
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2024) -
Vanishing Gradients in Reinforcement Finetuning of Language Models
von: Razin, Noam, et al.
Veröffentlicht: (2023) -
Classifier-Free Guidance is a Predictor-Corrector
von: Bradley, Arwen, et al.
Veröffentlicht: (2024)