On the Emergence of Weak-to-Strong Generalization: A Bias-Variance Perspective
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Gengze, Yao, Wei, Wang, Ziqiao, Liu, Yong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Capabilities and Limitations of Weak-to-Strong Generalization: Generalization and Calibration
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
On Weak-to-Strong Generalization and f-Divergence
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
On the Blessing of Pre-training in Weak-to-Strong Generalization
by: Yao, Wei, et al.
Published: (2026)
by: Yao, Wei, et al.
Published: (2026)
Revisiting Weak-to-Strong Generalization in Theory and Practice: Reverse KL vs. Forward KL
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
Theoretical Insights into Fine-Tuning Attention Mechanism: Generalization and Optimization
by: Yao, Xinhao, et al.
Published: (2024)
by: Yao, Xinhao, et al.
Published: (2024)
On the Mechanisms of Weak-to-Strong Generalization: A Theoretical Perspective
by: Moniri, Behrad, et al.
Published: (2025)
by: Moniri, Behrad, et al.
Published: (2025)
A Generalized Bias-Variance Decomposition for Bregman Divergences
by: Pfau, David
Published: (2025)
by: Pfau, David
Published: (2025)
The Bias-Variance Tradeoff in Data-Driven Optimization: A Local Misspecification Perspective
by: Lan, Haixiang, et al.
Published: (2025)
by: Lan, Haixiang, et al.
Published: (2025)
Exact Finite-Sample Variance Decomposition of Subagging: A Spectral Filtering Perspective
by: Su, Ye, et al.
Published: (2026)
by: Su, Ye, et al.
Published: (2026)
Understanding the Generalization of Bilevel Programming in Hyperparameter Optimization: A Tale of Bias-Variance Decomposition
by: Zhou, Yubo, et al.
Published: (2026)
by: Zhou, Yubo, et al.
Published: (2026)
A Bias-Variance-Covariance Decomposition of Kernel Scores for Generative Models
by: Gruber, Sebastian G., et al.
Published: (2023)
by: Gruber, Sebastian G., et al.
Published: (2023)
Generalization Bounds via Conditional $f$-Information
by: Wang, Ziqiao, et al.
Published: (2024)
by: Wang, Ziqiao, et al.
Published: (2024)
On the Emergence of Position Bias in Transformers
by: Wu, Xinyi, et al.
Published: (2025)
by: Wu, Xinyi, et al.
Published: (2025)
How Ensemble Learning Balances Accuracy and Overfitting: A Bias-Variance Perspective on Tabular Data
by: Mohammad, Zubair Ahmed
Published: (2025)
by: Mohammad, Zubair Ahmed
Published: (2025)
Theoretical Analysis of Weak-to-Strong Generalization
by: Lang, Hunter, et al.
Published: (2024)
by: Lang, Hunter, et al.
Published: (2024)
Quantifying the Gain in Weak-to-Strong Generalization
by: Charikar, Moses, et al.
Published: (2024)
by: Charikar, Moses, et al.
Published: (2024)
Weak-to-Strong Generalization with Failure Trajectories: A Tree-based Approach to Elicit Optimal Policy in Strong Models
by: Ye, Ruimeng, et al.
Published: (2025)
by: Ye, Ruimeng, et al.
Published: (2025)
Randomization Can Reduce Both Bias and Variance: A Case Study in Random Forests
by: Liu, Brian, et al.
Published: (2024)
by: Liu, Brian, et al.
Published: (2024)
Provable Weak-to-Strong Generalization via Benign Overfitting
by: Wu, David X., et al.
Published: (2024)
by: Wu, David X., et al.
Published: (2024)
Weak-to-Strong Generalization is Nearly Inevitable (in Linear Models)
by: Geng, Scott, et al.
Published: (2026)
by: Geng, Scott, et al.
Published: (2026)
Two Facets of SDE Under an Information-Theoretic Lens: Generalization of SGD via Training Trajectories and via Terminal States
by: Wang, Ziqiao, et al.
Published: (2022)
by: Wang, Ziqiao, et al.
Published: (2022)
Rethinking Semi-Supervised Imbalanced Node Classification from Bias-Variance Decomposition
by: Yan, Liang, et al.
Published: (2023)
by: Yan, Liang, et al.
Published: (2023)
Reducing Bias and Variance: Generative Semantic Guidance and Bi-Layer Ensemble for Image Clustering
by: Li, Feijiang, et al.
Published: (2026)
by: Li, Feijiang, et al.
Published: (2026)
Generalization in Federated Learning: A Conditional Mutual Information Framework
by: Wang, Ziqiao, et al.
Published: (2025)
by: Wang, Ziqiao, et al.
Published: (2025)
Weak-to-Strong Generalization under Distribution Shifts
by: Jeon, Myeongho, et al.
Published: (2025)
by: Jeon, Myeongho, et al.
Published: (2025)
Weak-to-Strong Generalization Even in Random Feature Networks, Provably
by: Medvedev, Marko, et al.
Published: (2025)
by: Medvedev, Marko, et al.
Published: (2025)
A Bias-Variance Decomposition for Ensembles over Multiple Synthetic Datasets
by: Räisä, Ossi, et al.
Published: (2024)
by: Räisä, Ossi, et al.
Published: (2024)
A Random Matrix Perspective of Echo State Networks: From Precise Bias--Variance Characterization to Optimal Regularization
by: Moakher, Yessin, et al.
Published: (2025)
by: Moakher, Yessin, et al.
Published: (2025)
Fine-Grained Dynamic Framework for Bias-Variance Joint Optimization on Data Missing Not at Random
by: Ha, Mingming, et al.
Published: (2024)
by: Ha, Mingming, et al.
Published: (2024)
The Mechanism of Weak-to-Strong Generalization: Feature Elicitation from Latent Knowledge
by: Awano, Ryoya, et al.
Published: (2026)
by: Awano, Ryoya, et al.
Published: (2026)
Generalizing Trust: Weak-to-Strong Trustworthiness in Language Models
by: Pawelczyk, Martin, et al.
Published: (2024)
by: Pawelczyk, Martin, et al.
Published: (2024)
Weak-to-Strong Generalization Through the Data-Centric Lens
by: Shin, Changho, et al.
Published: (2024)
by: Shin, Changho, et al.
Published: (2024)
On the Convergence of Experience Replay in Policy Optimization: Characterizing Bias, Variance, and Finite-Time Convergence
by: Zheng, Hua, et al.
Published: (2021)
by: Zheng, Hua, et al.
Published: (2021)
Evaluating Risks in Weak-to-Strong Alignment: A Bias-Variance Perspective
by: Osooli, Hamid, et al.
Published: (2026)
by: Osooli, Hamid, et al.
Published: (2026)
From Linear to Nonlinear: Provable Weak-to-Strong Generalization through Feature Learning
by: Oh, Junsoo, et al.
Published: (2025)
by: Oh, Junsoo, et al.
Published: (2025)
High-dimensional Analysis of Knowledge Distillation: Weak-to-Strong Generalization and Scaling Laws
by: Ildiz, M. Emrullah, et al.
Published: (2024)
by: Ildiz, M. Emrullah, et al.
Published: (2024)
Weak-to-Strong Elicitation via Mismatched Wrong Drafts
by: Deng, Wei
Published: (2026)
by: Deng, Wei
Published: (2026)
A Unified Stability Analysis of SAM vs SGD: Role of Data Coherence and Emergence of Simplicity Bias
by: Chang, Wei-Kai, et al.
Published: (2025)
by: Chang, Wei-Kai, et al.
Published: (2025)
Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher
by: Uzunoglu, Arda, et al.
Published: (2026)
by: Uzunoglu, Arda, et al.
Published: (2026)
On SkipGram Word Embedding Models with Negative Sampling: Unified Framework and Impact of Noise Distributions
by: Liu, Dezhi, et al.
Published: (2020)
by: Liu, Dezhi, et al.
Published: (2020)
Similar Items
-
The Capabilities and Limitations of Weak-to-Strong Generalization: Generalization and Calibration
by: Yao, Wei, et al.
Published: (2025) -
On Weak-to-Strong Generalization and f-Divergence
by: Yao, Wei, et al.
Published: (2025) -
On the Blessing of Pre-training in Weak-to-Strong Generalization
by: Yao, Wei, et al.
Published: (2026) -
Revisiting Weak-to-Strong Generalization in Theory and Practice: Reverse KL vs. Forward KL
by: Yao, Wei, et al.
Published: (2025) -
Theoretical Insights into Fine-Tuning Attention Mechanism: Generalization and Optimization
by: Yao, Xinhao, et al.
Published: (2024)