Weak-to-Strong Generalization Through the Data-Centric Lens
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shin, Changho, Cooper, John, Sala, Frederic |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-Shot Robustification of Zero-Shot Models
von: Adila, Dyah, et al.
Veröffentlicht: (2023)
von: Adila, Dyah, et al.
Veröffentlicht: (2023)
OTTER: Effortless Label Distribution Adaptation of Zero-shot Models
von: Shin, Changho, et al.
Veröffentlicht: (2024)
von: Shin, Changho, et al.
Veröffentlicht: (2024)
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
von: Zhao, Jitian, et al.
Veröffentlicht: (2026)
von: Zhao, Jitian, et al.
Veröffentlicht: (2026)
Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights
von: Huang, Tzu-Heng, et al.
Veröffentlicht: (2025)
von: Huang, Tzu-Heng, et al.
Veröffentlicht: (2025)
Quantifying the Gain in Weak-to-Strong Generalization
von: Charikar, Moses, et al.
Veröffentlicht: (2024)
von: Charikar, Moses, et al.
Veröffentlicht: (2024)
Weak-to-Strong Generalization under Distribution Shifts
von: Jeon, Myeongho, et al.
Veröffentlicht: (2025)
von: Jeon, Myeongho, et al.
Veröffentlicht: (2025)
Generalizing Trust: Weak-to-Strong Trustworthiness in Language Models
von: Pawelczyk, Martin, et al.
Veröffentlicht: (2024)
von: Pawelczyk, Martin, et al.
Veröffentlicht: (2024)
Representations Shape Weak-to-Strong Generalization: Theoretical Insights and Empirical Predictions
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
MoRe Fine-Tuning with 10x Fewer Parameters
von: Tan, Wenxuan, et al.
Veröffentlicht: (2024)
von: Tan, Wenxuan, et al.
Veröffentlicht: (2024)
COSMOS: Predictable and Cost-Effective Adaptation of LLMs
von: Wang, Jiayu, et al.
Veröffentlicht: (2025)
von: Wang, Jiayu, et al.
Veröffentlicht: (2025)
SF(DA)$^2$: Source-free Domain Adaptation Through the Lens of Data Augmentation
von: Hwang, Uiwon, et al.
Veröffentlicht: (2024)
von: Hwang, Uiwon, et al.
Veröffentlicht: (2024)
Mixture of Weak & Strong Experts on Graphs
von: Zeng, Hanqing, et al.
Veröffentlicht: (2023)
von: Zeng, Hanqing, et al.
Veröffentlicht: (2023)
Evaluating Language Model Context Windows: A "Working Memory" Test and Inference-time Correction
von: Dsouza, Amanda, et al.
Veröffentlicht: (2024)
von: Dsouza, Amanda, et al.
Veröffentlicht: (2024)
Personalize Your LLM: Fake it then Align it
von: Zhang, Yijing, et al.
Veröffentlicht: (2025)
von: Zhang, Yijing, et al.
Veröffentlicht: (2025)
Time Series Forecasting Through the Lens of Dynamics
von: Brachet, Alexis-Raja, et al.
Veröffentlicht: (2025)
von: Brachet, Alexis-Raja, et al.
Veröffentlicht: (2025)
Time To Impeach LLM-as-a-Judge: Programs are the Future of Evaluation
von: Huang, Tzu-Heng, et al.
Veröffentlicht: (2025)
von: Huang, Tzu-Heng, et al.
Veröffentlicht: (2025)
Bayesian WeakS-to-Strong from Text Classification to Generation
von: Cui, Ziyun, et al.
Veröffentlicht: (2024)
von: Cui, Ziyun, et al.
Veröffentlicht: (2024)
Improving Generalization of Neural Vehicle Routing Problem Solvers Through the Lens of Model Architecture
von: Xiao, Yubin, et al.
Veröffentlicht: (2024)
von: Xiao, Yubin, et al.
Veröffentlicht: (2024)
Age Predictors Through the Lens of Generalization, Bias Mitigation, and Interpretability: Reflections on Causal Implications
von: Paul, Debdas, et al.
Veröffentlicht: (2026)
von: Paul, Debdas, et al.
Veröffentlicht: (2026)
Revisiting Weak-to-Strong Generalization in Theory and Practice: Reverse KL vs. Forward KL
von: Yao, Wei, et al.
Veröffentlicht: (2025)
von: Yao, Wei, et al.
Veröffentlicht: (2025)
Data Curation Through the Lens of Spectral Dynamics: Static Limits, Dynamic Acceleration, and Practical Oracles
von: Zhang, Yizhou, et al.
Veröffentlicht: (2025)
von: Zhang, Yizhou, et al.
Veröffentlicht: (2025)
Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer
von: Askin, Baris, et al.
Veröffentlicht: (2026)
von: Askin, Baris, et al.
Veröffentlicht: (2026)
GPT-2 Through the Lens of Vector Symbolic Architectures
von: Knittel, Johannes, et al.
Veröffentlicht: (2024)
von: Knittel, Johannes, et al.
Veröffentlicht: (2024)
Who's the (Multi-)Fairest of Them All: Rethinking Interpolation-Based Data Augmentation Through the Lens of Multicalibration
von: Halevy, Karina, et al.
Veröffentlicht: (2024)
von: Halevy, Karina, et al.
Veröffentlicht: (2024)
Quantifying Structure in CLIP Embeddings: A Statistical Framework for Concept Interpretation
von: Zhao, Jitian, et al.
Veröffentlicht: (2025)
von: Zhao, Jitian, et al.
Veröffentlicht: (2025)
Reframing Data Value for Large Language Models Through the Lens of Plausibility
von: Rammal, Mohamad Rida, et al.
Veröffentlicht: (2024)
von: Rammal, Mohamad Rida, et al.
Veröffentlicht: (2024)
The ALCHEmist: Automated Labeling 500x CHEaper Than LLM Data Annotators
von: Huang, Tzu-Heng, et al.
Veröffentlicht: (2024)
von: Huang, Tzu-Heng, et al.
Veröffentlicht: (2024)
Laplace Sample Information: Data Informativeness Through a Bayesian Lens
von: Kaiser, Johannes, et al.
Veröffentlicht: (2025)
von: Kaiser, Johannes, et al.
Veröffentlicht: (2025)
WST: Weak-to-Strong Knowledge Transfer via Reinforcement Learning
von: Ge, Haosen, et al.
Veröffentlicht: (2025)
von: Ge, Haosen, et al.
Veröffentlicht: (2025)
When to Trust the Cheap Check: Weak and Strong Verification for Reasoning
von: Kiyani, Shayan, et al.
Veröffentlicht: (2026)
von: Kiyani, Shayan, et al.
Veröffentlicht: (2026)
Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes
von: Bauer, Justin, et al.
Veröffentlicht: (2026)
von: Bauer, Justin, et al.
Veröffentlicht: (2026)
Is Free Self-Alignment Possible?
von: Adila, Dyah, et al.
Veröffentlicht: (2024)
von: Adila, Dyah, et al.
Veröffentlicht: (2024)
Promises and Pitfalls of Threshold-based Auto-labeling
von: Vishwakarma, Harit, et al.
Veröffentlicht: (2022)
von: Vishwakarma, Harit, et al.
Veröffentlicht: (2022)
Data-Centric Interpretability for LLM-based Multi-Agent Reinforcement Learning
von: Yan, John, et al.
Veröffentlicht: (2026)
von: Yan, John, et al.
Veröffentlicht: (2026)
DataMaster: Data-Centric Autonomous AI Research
von: Du, Yaxin, et al.
Veröffentlicht: (2026)
von: Du, Yaxin, et al.
Veröffentlicht: (2026)
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization
von: Chen, Xinjie, et al.
Veröffentlicht: (2025)
von: Chen, Xinjie, et al.
Veröffentlicht: (2025)
R&B: Domain Regrouping and Data Mixture Balancing for Efficient Foundation Model Training
von: Ge, Albert, et al.
Veröffentlicht: (2025)
von: Ge, Albert, et al.
Veröffentlicht: (2025)
On the Sparsity of the Strong Lottery Ticket Hypothesis
von: Natale, Emanuele, et al.
Veröffentlicht: (2024)
von: Natale, Emanuele, et al.
Veröffentlicht: (2024)
Towards Data-Centric AI: A Comprehensive Survey of Traditional, Reinforcement, and Generative Approaches for Tabular Data Transformation
von: Wang, Dongjie, et al.
Veröffentlicht: (2025)
von: Wang, Dongjie, et al.
Veröffentlicht: (2025)
Multiple Weaks Win Single Strong: Large Language Models Ensemble Weak Reinforcement Learning Agents into a Supreme One
von: Song, Yiwen, et al.
Veröffentlicht: (2025)
von: Song, Yiwen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Zero-Shot Robustification of Zero-Shot Models
von: Adila, Dyah, et al.
Veröffentlicht: (2023) -
OTTER: Effortless Label Distribution Adaptation of Zero-shot Models
von: Shin, Changho, et al.
Veröffentlicht: (2024) -
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
von: Zhao, Jitian, et al.
Veröffentlicht: (2026) -
Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights
von: Huang, Tzu-Heng, et al.
Veröffentlicht: (2025) -
Quantifying the Gain in Weak-to-Strong Generalization
von: Charikar, Moses, et al.
Veröffentlicht: (2024)