EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
Fuente:
arXiv
Saved in:
| Main Authors: | Agrawal, Aakriti, Ding, Mucong, Che, Zora, Deng, Chenghao, Satheesh, Anirudh, An, Bang, Bruss, Bayan, Langford, John, Huang, Furong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
by: Agrawal, Aakriti, et al.
Published: (2024)
by: Agrawal, Aakriti, et al.
Published: (2024)
SAFLEX: Self-Adaptive Augmentation via Feature Label Extrapolation
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
SAIL: Self-Improving Efficient Online Alignment of Large Language Models
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
Easy2Hard-Bench: Standardized Difficulty Labels for Profiling LLM Performance and Generalization
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
PoisonedParrot: Subtle Data Poisoning Attacks to Elicit Copyright-Infringing Content from Large Language Models
by: Panaitescu-Liess, Michael-Andrei, et al.
Published: (2025)
by: Panaitescu-Liess, Michael-Andrei, et al.
Published: (2025)
WAVES: Benchmarking the Robustness of Image Watermarks
by: An, Bang, et al.
Published: (2024)
by: An, Bang, et al.
Published: (2024)
PICore: Physics-Informed Unsupervised Coreset Selection for Data Efficient Neural Operator Training
by: Satheesh, Anirudh, et al.
Published: (2025)
by: Satheesh, Anirudh, et al.
Published: (2025)
Uncertainty-Aware Answer Selection for Improved Reasoning in Multi-LLM Systems
by: Agrawal, Aakriti, et al.
Published: (2025)
by: Agrawal, Aakriti, et al.
Published: (2025)
VeriGate: Verifier-Gated Step-Level Supervision for GRPO
by: Agrawal, Aakriti, et al.
Published: (2026)
by: Agrawal, Aakriti, et al.
Published: (2026)
Sketch-GNN: Scalable Graph Neural Networks with Sublinear Training Complexity
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
Provably Efficient Algorithms for S- and Non-Rectangular Robust MDPs with General Parameterization
by: Satheesh, Anirudh, et al.
Published: (2026)
by: Satheesh, Anirudh, et al.
Published: (2026)
Spectral Greedy Coresets for Graph Neural Networks
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
EnsemJudge: Enhancing Reliability in Chinese LLM-Generated Text Detection through Diverse Model Ensembles
by: Wang, Zhuoshang, et al.
Published: (2026)
by: Wang, Zhuoshang, et al.
Published: (2026)
Can Watermarking Large Language Models Prevent Copyrighted Text Generation and Hide Training Data?
by: Panaitescu-Liess, Michael-Andrei, et al.
Published: (2024)
by: Panaitescu-Liess, Michael-Andrei, et al.
Published: (2024)
Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings
by: Agrawal, Aakriti, et al.
Published: (2025)
by: Agrawal, Aakriti, et al.
Published: (2025)
EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors
by: Miyazato, Ryuhei, et al.
Published: (2026)
by: Miyazato, Ryuhei, et al.
Published: (2026)
A Constrained Multi-Agent Reinforcement Learning Approach to Autonomous Traffic Signal Control
by: Satheesh, Anirudh, et al.
Published: (2025)
by: Satheesh, Anirudh, et al.
Published: (2025)
Regret Analysis of Unichain Average Reward Constrained MDPs with General Parameterization
by: Satheesh, Anirudh, et al.
Published: (2026)
by: Satheesh, Anirudh, et al.
Published: (2026)
AI versus AI in Financial Crimes and Detection: GenAI Crime Waves to Co-Evolutionary AI
by: Kurshan, Eren, et al.
Published: (2024)
by: Kurshan, Eren, et al.
Published: (2024)
AegisLLM: Scaling Agentic Systems for Self-Reflective Defense in LLM Security
by: Cai, Zikui, et al.
Published: (2025)
by: Cai, Zikui, et al.
Published: (2025)
Learning When to Trust Which Teacher for Weakly Supervised ASR
by: Agrawal, Aakriti, et al.
Published: (2023)
by: Agrawal, Aakriti, et al.
Published: (2023)
TimeSqueeze: Dynamic Patching for Efficient Time Series Forecasting
by: Ankireddy, Sravan Kumar, et al.
Published: (2026)
by: Ankireddy, Sravan Kumar, et al.
Published: (2026)
W2S-AlignTree: Weak-to-Strong Inference-Time Alignment for Large Language Models via Monte Carlo Tree Search
by: Ding, Zhenyu, et al.
Published: (2025)
by: Ding, Zhenyu, et al.
Published: (2025)
Bridging the Divide: End-to-End Sequence-Graph Learning
by: Chen, Yuen, et al.
Published: (2025)
by: Chen, Yuen, et al.
Published: (2025)
cMALC-D: Contextual Multi-Agent LLM-Guided Curriculum Learning with Diversity-Based Context Blending
by: Satheesh, Anirudh, et al.
Published: (2025)
by: Satheesh, Anirudh, et al.
Published: (2025)
Distributionally Robust Self Paced Curriculum Reinforcement Learning
by: Satheesh, Anirudh, et al.
Published: (2025)
by: Satheesh, Anirudh, et al.
Published: (2025)
Auction-Based Regulation for Artificial Intelligence
by: Bornstein, Marco, et al.
Published: (2024)
by: Bornstein, Marco, et al.
Published: (2024)
Multiple Weaks Win Single Strong: Large Language Models Ensemble Weak Reinforcement Learning Agents into a Supreme One
by: Song, Yiwen, et al.
Published: (2025)
by: Song, Yiwen, et al.
Published: (2025)
Tuning-Free LLM Can Build A Strong Recommender Under Sparse Connectivity And Knowledge Gap Via Extracting Intent
by: Zheng, Wenqing, et al.
Published: (2025)
by: Zheng, Wenqing, et al.
Published: (2025)
Weak-to-Strong Jailbreaking on Large Language Models
by: Zhao, Xuandong, et al.
Published: (2024)
by: Zhao, Xuandong, et al.
Published: (2024)
Compositional Adversarial Training for Robust Visual Watermarking
by: Satheesh, Anirudh, et al.
Published: (2026)
by: Satheesh, Anirudh, et al.
Published: (2026)
Zero-shot Multivariate Time Series Forecasting Using Tabular Prior Fitted Networks
by: Jayawardhana, Mayuka, et al.
Published: (2026)
by: Jayawardhana, Mayuka, et al.
Published: (2026)
Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities
by: Che, Zora, et al.
Published: (2025)
by: Che, Zora, et al.
Published: (2025)
Improving Weak-to-Strong Generalization with Scalable Oversight and Ensemble Learning
by: Sang, Jitao, et al.
Published: (2024)
by: Sang, Jitao, et al.
Published: (2024)
Calibrated Dataset Condensation for Faster Hyperparameter Search
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
Beyond Worst-case Attacks: Robust RL with Adaptive Defense via Non-dominated Policies
by: Liu, Xiangyu, et al.
Published: (2024)
by: Liu, Xiangyu, et al.
Published: (2024)
Influencia de la presión y la temperatura en la reacción de hidroformilación en medio bifásico del 1-hexeno en régimen continuo con el complejo catalítico [(µ-Pz)(CO)(TFFTS)Rh]2 (I)
by: Arnoldo Bruss
Published: (2006)
by: Arnoldo Bruss
Published: (2006)
Modulation by context of a scene in monkey anterior inferotemporal cortex during a saccadic eye movement task
by: Bruss Lima
Published: (2003)
by: Bruss Lima
Published: (2003)
BEDTime: A Unified Benchmark for Automatically Describing Time Series
by: Sen, Medhasweta, et al.
Published: (2025)
by: Sen, Medhasweta, et al.
Published: (2025)
On Giant's Shoulders: Effortless Weak to Strong by Dynamic Logits Fusion
by: Fan, Chenghao, et al.
Published: (2024)
by: Fan, Chenghao, et al.
Published: (2024)
Similar Items
-
EnsemW2S: Enhancing Weak-to-Strong Generalization with Large Language Model Ensembles
by: Agrawal, Aakriti, et al.
Published: (2024) -
SAFLEX: Self-Adaptive Augmentation via Feature Label Extrapolation
by: Ding, Mucong, et al.
Published: (2024) -
SAIL: Self-Improving Efficient Online Alignment of Large Language Models
by: Ding, Mucong, et al.
Published: (2024) -
Easy2Hard-Bench: Standardized Difficulty Labels for Profiling LLM Performance and Generalization
by: Ding, Mucong, et al.
Published: (2024) -
PoisonedParrot: Subtle Data Poisoning Attacks to Elicit Copyright-Infringing Content from Large Language Models
by: Panaitescu-Liess, Michael-Andrei, et al.
Published: (2025)