Quantifying the Gain in Weak-to-Strong Generalization
Fuente:
arXiv
Saved in:
| Main Authors: | Charikar, Moses, Pabbaraju, Chirag, Shiragur, Kirankumar |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Facets of Language Generation in the Limit
by: Charikar, Moses, et al.
Published: (2024)
by: Charikar, Moses, et al.
Published: (2024)
Pareto-optimal Non-uniform Language Generation
by: Charikar, Moses, et al.
Published: (2025)
by: Charikar, Moses, et al.
Published: (2025)
A Characterization of List Language Identification in the Limit
by: Charikar, Moses, et al.
Published: (2025)
by: Charikar, Moses, et al.
Published: (2025)
Testing with Non-identically Distributed Samples
by: Garg, Shivam, et al.
Published: (2023)
by: Garg, Shivam, et al.
Published: (2023)
Relating Misfit to Gain in Weak-to-Strong Generalization Beyond the Squared Loss
by: Mulgund, Abhijeet, et al.
Published: (2025)
by: Mulgund, Abhijeet, et al.
Published: (2025)
Membership Testing in Markov Equivalence Classes via Independence Query Oracles
by: Zhang, Jiaqi, et al.
Published: (2024)
by: Zhang, Jiaqi, et al.
Published: (2024)
Causal Discovery with Fewer Conditional Independence Tests
by: Shiragur, Kirankumar, et al.
Published: (2024)
by: Shiragur, Kirankumar, et al.
Published: (2024)
Agnostic Language Identification and Generation
by: Høgsgaard, Mikael Møller, et al.
Published: (2026)
by: Høgsgaard, Mikael Møller, et al.
Published: (2026)
Cost Efficient Fairness Audit Under Partial Feedback
by: Das, Nirjhar, et al.
Published: (2025)
by: Das, Nirjhar, et al.
Published: (2025)
Subset verification and search algorithms for causal DAGs
by: Choo, Davin, et al.
Published: (2023)
by: Choo, Davin, et al.
Published: (2023)
Embedding Probability Distributions into Low Dimensional $\ell_1$: Tree Ising Models via Truncated Metrics
by: Charikar, Moses, et al.
Published: (2023)
by: Charikar, Moses, et al.
Published: (2023)
The Optimal Sample Complexity of Multiclass and List Learning
by: Pabbaraju, Chirag
Published: (2026)
by: Pabbaraju, Chirag
Published: (2026)
COSMIR: Chain Orchestrated Structured Memory for Iterative Reasoning over Long Context
by: Gupta, Naman, et al.
Published: (2025)
by: Gupta, Naman, et al.
Published: (2025)
Learning Mixtures of Unknown Causal Interventions
by: Kumar, Abhinav, et al.
Published: (2024)
by: Kumar, Abhinav, et al.
Published: (2024)
Weak-to-Strong Generalization under Distribution Shifts
by: Jeon, Myeongho, et al.
Published: (2025)
by: Jeon, Myeongho, et al.
Published: (2025)
Generalizing Trust: Weak-to-Strong Trustworthiness in Language Models
by: Pawelczyk, Martin, et al.
Published: (2024)
by: Pawelczyk, Martin, et al.
Published: (2024)
Weak-to-Strong Generalization Through the Data-Centric Lens
by: Shin, Changho, et al.
Published: (2024)
by: Shin, Changho, et al.
Published: (2024)
Representations Shape Weak-to-Strong Generalization: Theoretical Insights and Empirical Predictions
by: Xue, Yihao, et al.
Published: (2025)
by: Xue, Yihao, et al.
Published: (2025)
Causal Discovery under Off-Target Interventions
by: Choo, Davin, et al.
Published: (2024)
by: Choo, Davin, et al.
Published: (2024)
Mixture of Weak & Strong Experts on Graphs
by: Zeng, Hanqing, et al.
Published: (2023)
by: Zeng, Hanqing, et al.
Published: (2023)
Bayesian WeakS-to-Strong from Text Classification to Generation
by: Cui, Ziyun, et al.
Published: (2024)
by: Cui, Ziyun, et al.
Published: (2024)
Revisiting Weak-to-Strong Generalization in Theory and Practice: Reverse KL vs. Forward KL
by: Yao, Wei, et al.
Published: (2025)
by: Yao, Wei, et al.
Published: (2025)
Preference Models assume Proportional Hazards of Utilities
by: Nagpal, Chirag
Published: (2025)
by: Nagpal, Chirag
Published: (2025)
Rethinking Explainability in the Era of Multimodal AI
by: Agarwal, Chirag
Published: (2025)
by: Agarwal, Chirag
Published: (2025)
WST: Weak-to-Strong Knowledge Transfer via Reinforcement Learning
by: Ge, Haosen, et al.
Published: (2025)
by: Ge, Haosen, et al.
Published: (2025)
When to Trust the Cheap Check: Weak and Strong Verification for Reasoning
by: Kiyani, Shayan, et al.
Published: (2026)
by: Kiyani, Shayan, et al.
Published: (2026)
A Characterization of List Regression
by: Pabbaraju, Chirag, et al.
Published: (2024)
by: Pabbaraju, Chirag, et al.
Published: (2024)
Multiple Weaks Win Single Strong: Large Language Models Ensemble Weak Reinforcement Learning Agents into a Supreme One
by: Song, Yiwen, et al.
Published: (2025)
by: Song, Yiwen, et al.
Published: (2025)
Reliable Weak-to-Strong Monitoring of LLM Agents
by: Kale, Neil, et al.
Published: (2025)
by: Kale, Neil, et al.
Published: (2025)
Weak-to-Strong Elicitation via Mismatched Wrong Drafts
by: Deng, Wei
Published: (2026)
by: Deng, Wei
Published: (2026)
Generative Modeling from Black-box Corruptions via Self-Consistent Stochastic Interpolants
by: Modi, Chirag, et al.
Published: (2025)
by: Modi, Chirag, et al.
Published: (2025)
On Giant's Shoulders: Effortless Weak to Strong by Dynamic Logits Fusion
by: Fan, Chenghao, et al.
Published: (2024)
by: Fan, Chenghao, et al.
Published: (2024)
The Amazing Agent Race: Strong Tool Users, Weak Navigators
by: Kim, Zae Myung, et al.
Published: (2026)
by: Kim, Zae Myung, et al.
Published: (2026)
Beyond Reasoning Gains: Mitigating General Capabilities Forgetting in Large Reasoning Models
by: Phan, Hoang, et al.
Published: (2025)
by: Phan, Hoang, et al.
Published: (2025)
How to Mitigate Overfitting in Weak-to-strong Generalization?
by: Shi, Junhao, et al.
Published: (2025)
by: Shi, Junhao, et al.
Published: (2025)
Co-Supervised Learning: Improving Weak-to-Strong Generalization with Hierarchical Mixture of Experts
by: Liu, Yuejiang, et al.
Published: (2024)
by: Liu, Yuejiang, et al.
Published: (2024)
Unbalanced Incomplete Multi-view Clustering via the Scheme of View Evolution: Weak Views are Meat; Strong Views do Eat
by: Fang, Xiang, et al.
Published: (2020)
by: Fang, Xiang, et al.
Published: (2020)
Analyzing Memorization in Large Language Models through the Lens of Model Attribution
by: Menta, Tarun Ram, et al.
Published: (2025)
by: Menta, Tarun Ram, et al.
Published: (2025)
A General Framework for Learning from Weak Supervision
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
How Do Latent Reasoning Methods Perform Under Weak and Strong Supervision?
by: Cui, Yingqian, et al.
Published: (2026)
by: Cui, Yingqian, et al.
Published: (2026)
Similar Items
-
Exploring Facets of Language Generation in the Limit
by: Charikar, Moses, et al.
Published: (2024) -
Pareto-optimal Non-uniform Language Generation
by: Charikar, Moses, et al.
Published: (2025) -
A Characterization of List Language Identification in the Limit
by: Charikar, Moses, et al.
Published: (2025) -
Testing with Non-identically Distributed Samples
by: Garg, Shivam, et al.
Published: (2023) -
Relating Misfit to Gain in Weak-to-Strong Generalization Beyond the Squared Loss
by: Mulgund, Abhijeet, et al.
Published: (2025)