Weak-to-Strong Generalization is Nearly Inevitable (in Linear Models)
Fuente:
arXiv
Saved in:
| Main Authors: | Geng, Scott, Hansen, Dutch, Li, Jerry |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Linear to Nonlinear: Provable Weak-to-Strong Generalization through Feature Learning
by: Oh, Junsoo, et al.
Published: (2025)
by: Oh, Junsoo, et al.
Published: (2025)
When is Multicalibration Post-Processing Necessary?
by: Hansen, Dutch, et al.
Published: (2024)
by: Hansen, Dutch, et al.
Published: (2024)
Auditability and the Landscape of Distance to Multicalibration
by: Derhake, Nathan, et al.
Published: (2025)
by: Derhake, Nathan, et al.
Published: (2025)
Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher
by: Uzunoglu, Arda, et al.
Published: (2026)
by: Uzunoglu, Arda, et al.
Published: (2026)
On Weak-to-Strong Generalization and f-Divergence
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
Weak-to-Strong Generalization with Failure Trajectories: A Tree-based Approach to Elicit Optimal Policy in Strong Models
by: Ye, Ruimeng, et al.
Published: (2025)
by: Ye, Ruimeng, et al.
Published: (2025)
The Capabilities and Limitations of Weak-to-Strong Generalization: Generalization and Calibration
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
Generalizing Trust: Weak-to-Strong Trustworthiness in Language Models
by: Pawelczyk, Martin, et al.
Published: (2024)
by: Pawelczyk, Martin, et al.
Published: (2024)
On the Blessing of Pre-training in Weak-to-Strong Generalization
by: Yao, Wei, et al.
Published: (2026)
by: Yao, Wei, et al.
Published: (2026)
Weak-to-Strong Generalization Even in Random Feature Networks, Provably
by: Medvedev, Marko, et al.
Published: (2025)
by: Medvedev, Marko, et al.
Published: (2025)
Theoretical Analysis of Weak-to-Strong Generalization
by: Lang, Hunter, et al.
Published: (2024)
by: Lang, Hunter, et al.
Published: (2024)
Quantifying the Gain in Weak-to-Strong Generalization
by: Charikar, Moses, et al.
Published: (2024)
by: Charikar, Moses, et al.
Published: (2024)
Provable Weak-to-Strong Generalization via Benign Overfitting
by: Wu, David X., et al.
Published: (2024)
by: Wu, David X., et al.
Published: (2024)
On the Mechanisms of Weak-to-Strong Generalization: A Theoretical Perspective
by: Moniri, Behrad, et al.
Published: (2025)
by: Moniri, Behrad, et al.
Published: (2025)
Tensor Completion with Nearly Linear Samples Given Weak Side Information
by: Yu, Christina Lee, et al.
Published: (2020)
by: Yu, Christina Lee, et al.
Published: (2020)
Weak-to-Strong Generalization under Distribution Shifts
by: Jeon, Myeongho, et al.
Published: (2025)
by: Jeon, Myeongho, et al.
Published: (2025)
EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
by: Agrawal, Aakriti, et al.
Published: (2024)
by: Agrawal, Aakriti, et al.
Published: (2024)
On the Emergence of Weak-to-Strong Generalization: A Bias-Variance Perspective
by: Xu, Gengze, et al.
Published: (2025)
by: Xu, Gengze, et al.
Published: (2025)
Discrepancies are Virtue: Weak-to-Strong Generalization through Lens of Intrinsic Dimension
by: Dong, Yijun, et al.
Published: (2025)
by: Dong, Yijun, et al.
Published: (2025)
Representations Shape Weak-to-Strong Generalization: Theoretical Insights and Empirical Predictions
by: Xue, Yihao, et al.
Published: (2025)
by: Xue, Yihao, et al.
Published: (2025)
The Mechanism of Weak-to-Strong Generalization: Feature Elicitation from Latent Knowledge
by: Awano, Ryoya, et al.
Published: (2026)
by: Awano, Ryoya, et al.
Published: (2026)
Weak-to-Strong Generalization Through the Data-Centric Lens
by: Shin, Changho, et al.
Published: (2024)
by: Shin, Changho, et al.
Published: (2024)
Hallucination is Inevitable: An Innate Limitation of Large Language Models
by: Xu, Ziwei, et al.
Published: (2024)
by: Xu, Ziwei, et al.
Published: (2024)
Near-optimal Linear Predictive Clustering in Non-separable Spaces via MIP and QPBO Reductions
by: Liang, Jiazhou, et al.
Published: (2025)
by: Liang, Jiazhou, et al.
Published: (2025)
High-dimensional Analysis of Knowledge Distillation: Weak-to-Strong Generalization and Scaling Laws
by: Ildiz, M. Emrullah, et al.
Published: (2024)
by: Ildiz, M. Emrullah, et al.
Published: (2024)
EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
by: Agrawal, Aakriti, et al.
Published: (2025)
by: Agrawal, Aakriti, et al.
Published: (2025)
Improved Scaling Laws via Weak-to-Strong Generalization in Random Feature Ridge Regression
by: Wu, Diyuan, et al.
Published: (2026)
by: Wu, Diyuan, et al.
Published: (2026)
A Generalization Bound for Nearly-Linear Networks
by: Golikov, Eugene
Published: (2024)
by: Golikov, Eugene
Published: (2024)
Bayesian WeakS-to-Strong from Text Classification to Generation
by: Cui, Ziyun, et al.
Published: (2024)
by: Cui, Ziyun, et al.
Published: (2024)
Shapley Marginal Surplus for Strong Models
by: de Marchi, Daniel, et al.
Published: (2024)
by: de Marchi, Daniel, et al.
Published: (2024)
Multiplicity is an Inevitable and Inherent Challenge in Multimodal Learning
by: Chun, Sanghyuk
Published: (2025)
by: Chun, Sanghyuk
Published: (2025)
Minimizing False-Positive Attributions in Explanations of Non-Linear Models
by: Gjølbye, Anders, et al.
Published: (2025)
by: Gjølbye, Anders, et al.
Published: (2025)
Mixture of Weak & Strong Experts on Graphs
by: Zeng, Hanqing, et al.
Published: (2023)
by: Zeng, Hanqing, et al.
Published: (2023)
Multiple Weaks Win Single Strong: Large Language Models Ensemble Weak Reinforcement Learning Agents into a Supreme One
by: Song, Yiwen, et al.
Published: (2025)
by: Song, Yiwen, et al.
Published: (2025)
WST: Weak-to-Strong Knowledge Transfer via Reinforcement Learning
by: Ge, Haosen, et al.
Published: (2025)
by: Ge, Haosen, et al.
Published: (2025)
Position: The Inevitable End of One-Architecture-Fits-All-Domains in Time Series Forecasting
by: Ma, Qinwei, et al.
Published: (2026)
by: Ma, Qinwei, et al.
Published: (2026)
Weak-to-Strong Diffusion with Reflection
by: Bai, Lichen, et al.
Published: (2025)
by: Bai, Lichen, et al.
Published: (2025)
A Unifying Framework for Parallelizing Sequential Models with Linear Dynamical Systems
by: Gonzalez, Xavier, et al.
Published: (2025)
by: Gonzalez, Xavier, et al.
Published: (2025)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
by: Park, Jihwan, et al.
Published: (2025)
by: Park, Jihwan, et al.
Published: (2025)
Revisiting Weak-to-Strong Generalization in Theory and Practice: Reverse KL vs. Forward KL
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
Similar Items
-
From Linear to Nonlinear: Provable Weak-to-Strong Generalization through Feature Learning
by: Oh, Junsoo, et al.
Published: (2025) -
When is Multicalibration Post-Processing Necessary?
by: Hansen, Dutch, et al.
Published: (2024) -
Auditability and the Landscape of Distance to Multicalibration
by: Derhake, Nathan, et al.
Published: (2025) -
Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher
by: Uzunoglu, Arda, et al.
Published: (2026) -
On Weak-to-Strong Generalization and f-Divergence
by: Yao, Wei, et al.
Published: (2025)