The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yiqun, Li, Hao, Wang, Chenxu, Chen, Linyao, Zhang, Qiaosheng, Ye, Peng, Feng, Shi, Wang, Daling, Wang, Zhen, Wang, Xinrun, Xu, Jia, Bai, Lei, Ouyang, Wanli, Hu, Shuyue |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ICL-Router: In-Context Learned Model Representations for LLM Routing
by: Wang, Chenxu, et al.
Published: (2025)
by: Wang, Chenxu, et al.
Published: (2025)
Design First, Code Later: Aesthetically Pleasing Template-Free Slides Generation
by: Cui, Zhiyao, et al.
Published: (2026)
by: Cui, Zhiyao, et al.
Published: (2026)
LLMRouterBench: A Massive Benchmark and Unified Framework for LLM Routing
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Provably Efficient Information-Directed Sampling Algorithms for Multi-Agent Reinforcement Learning
by: Zhang, Qiaosheng, et al.
Published: (2024)
by: Zhang, Qiaosheng, et al.
Published: (2024)
Stop Overvaluing Multi-Agent Debate -- We Must Rethink Evaluation and Embrace Model Heterogeneity
by: Zhang, Hangfan, et al.
Published: (2025)
by: Zhang, Hangfan, et al.
Published: (2025)
Nature-Inspired Population-Based Evolution of Large Language Models
by: Zhang, Yiqun, et al.
Published: (2025)
by: Zhang, Yiqun, et al.
Published: (2025)
MTRouter: Cost-Aware Multi-Turn LLM Routing with History-Model Joint Embeddings
by: Zhang, Yiqun, et al.
Published: (2026)
by: Zhang, Yiqun, et al.
Published: (2026)
The Path of Self-Evolving Large Language Models: Achieving Data-Efficient Learning via Intrinsic Feedback
by: Zhang, Hangfan, et al.
Published: (2025)
by: Zhang, Hangfan, et al.
Published: (2025)
Graph Attention is Not Always Beneficial: A Theoretical Analysis of Graph Attention Mechanisms via Contextual Stochastic Block Models
by: Ma, Zhongtian, et al.
Published: (2024)
by: Ma, Zhongtian, et al.
Published: (2024)
Beyond Gemini-3-Pro: Revisiting LLM Routing and Aggregation at Scale
by: Tang, Shengji, et al.
Published: (2026)
by: Tang, Shengji, et al.
Published: (2026)
Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment
by: Li, Jiajia, et al.
Published: (2026)
by: Li, Jiajia, et al.
Published: (2026)
Ensemble Successor Representations for Task Generalization in Offline-to-Online Reinforcement Learning
by: Wang, Changhong, et al.
Published: (2024)
by: Wang, Changhong, et al.
Published: (2024)
Matrix Completion with Hypergraphs:Sharp Thresholds and Efficient Algorithms
by: Ma, Zhongtian, et al.
Published: (2024)
by: Ma, Zhongtian, et al.
Published: (2024)
Can LLMs Beat Humans in Debating? A Dynamic Multi-agent Framework for Competitive Debate
by: Zhang, Yiqun, et al.
Published: (2024)
by: Zhang, Yiqun, et al.
Published: (2024)
TOOL-ED: Enhancing Empathetic Response Generation with the Tool Calling Capability of LLM
by: Cao, Huiying, et al.
Published: (2024)
by: Cao, Huiying, et al.
Published: (2024)
Native-Resolution Image Synthesis
by: Wang, Zidong, et al.
Published: (2025)
by: Wang, Zidong, et al.
Published: (2025)
Learning Compact Representations of LLM Abilities via Item Response Theory
by: Chen, Jianhao, et al.
Published: (2025)
by: Chen, Jianhao, et al.
Published: (2025)
Misclassification Rate and Privacy-Utility Trade-offs in Graph Convolutional Networks via Subsampling Stability
by: Zhang, Yexin, et al.
Published: (2026)
by: Zhang, Yexin, et al.
Published: (2026)
Beyond GPT-5: Making LLMs Cheaper and Better via Performance-Efficiency Optimized Routing
by: Zhang, Yiqun, et al.
Published: (2025)
by: Zhang, Yiqun, et al.
Published: (2025)
PerPilot: Personalizing VLM-based Mobile Agents via Memory and Exploration
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
Leveraging Large Language Models for Enhanced Digital Twin Modeling: Trends, Methods, and Challenges
by: Yang, Linyao, et al.
Published: (2025)
by: Yang, Linyao, et al.
Published: (2025)
Are Smaller Open-Weight LLMs Closing the Gap to Proprietary Models for Biomedical Question Answering?
by: Stachura, Damian, et al.
Published: (2025)
by: Stachura, Damian, et al.
Published: (2025)
Adaptive Theory of Mind for LLM-based Multi-Agent Coordination
by: Mu, Chunjiang, et al.
Published: (2026)
by: Mu, Chunjiang, et al.
Published: (2026)
STICKERCONV: Generating Multimodal Empathetic Responses from Scratch
by: Zhang, Yiqun, et al.
Published: (2024)
by: Zhang, Yiqun, et al.
Published: (2024)
The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents
by: Wang, Xinrun, et al.
Published: (2026)
by: Wang, Xinrun, et al.
Published: (2026)
DEEPMED: Building a Medical DeepResearch Agent via Multi-hop Med-Search Data and Turn-Controlled Agentic Training & Inference
by: Wang, Zihan, et al.
Published: (2026)
by: Wang, Zihan, et al.
Published: (2026)
Avengers assemble: Avenger philanthropy as the new gift opportunity for nonprofit organizations
by: Sarah‐Louise Mitchell
Published: (2024)
by: Sarah‐Louise Mitchell
Published: (2024)
Do We Truly Need So Many Samples? Multi-LLM Repeated Sampling Efficiently Scales Test-Time Compute
by: Chen, Jianhao, et al.
Published: (2025)
by: Chen, Jianhao, et al.
Published: (2025)
How Many Visual Tokens Do Multimodal Language Models Need? Scaling Visual Token Pruning with F^3A
by: Huang, YiJie, et al.
Published: (2026)
by: Huang, YiJie, et al.
Published: (2026)
Community Detection in the Multi-View Stochastic Block Model
by: Zhang, Yexin, et al.
Published: (2024)
by: Zhang, Yexin, et al.
Published: (2024)
Gated Graph Attention Networks with Learnable Temperature
by: Ma, Zhongtian, et al.
Published: (2026)
by: Ma, Zhongtian, et al.
Published: (2026)
Single-Agent Scaling Fails Multi-Agent Intelligence: Towards Foundation Models with Native Multi-Agent Intelligence
by: Hu, Shuyue, et al.
Published: (2025)
by: Hu, Shuyue, et al.
Published: (2025)
PsyDraw: A Multi-Agent Multimodal System for Mental Health Screening in Left-Behind Children
by: Zhang, Yiqun, et al.
Published: (2024)
by: Zhang, Yiqun, et al.
Published: (2024)
HiFT: A Hierarchical Full Parameter Fine-Tuning Strategy
by: Liu, Yongkang, et al.
Published: (2024)
by: Liu, Yongkang, et al.
Published: (2024)
T-COL: Generating Counterfactual Explanations for General User Preferences on Variable Machine Learning Systems
by: Wang, Ming, et al.
Published: (2023)
by: Wang, Ming, et al.
Published: (2023)
A Scalable Multi-LLM Collaboration System with Retrieval-based Selection and Exploration-Exploitation-Driven Enhancement
by: Tang, Shengji, et al.
Published: (2025)
by: Tang, Shengji, et al.
Published: (2025)
Breaking the Compression Ceiling: Data-Free Pipeline for Ultra-Efficient Delta Compression
by: Wang, Xiaohui, et al.
Published: (2025)
by: Wang, Xiaohui, et al.
Published: (2025)
Why Do More Experts Fail? A Theoretical Analysis of Model Merging
by: Wang, Zijing, et al.
Published: (2025)
by: Wang, Zijing, et al.
Published: (2025)
SAD: A Large-Scale Strategic Argumentative Dialogue Dataset
by: Liu, Yongkang, et al.
Published: (2026)
by: Liu, Yongkang, et al.
Published: (2026)
Dynamic Base model Shift for Delta Compression
by: Huang, Chenyu, et al.
Published: (2025)
by: Huang, Chenyu, et al.
Published: (2025)
Similar Items
-
ICL-Router: In-Context Learned Model Representations for LLM Routing
by: Wang, Chenxu, et al.
Published: (2025) -
Design First, Code Later: Aesthetically Pleasing Template-Free Slides Generation
by: Cui, Zhiyao, et al.
Published: (2026) -
LLMRouterBench: A Massive Benchmark and Unified Framework for LLM Routing
by: Li, Hao, et al.
Published: (2026) -
Provably Efficient Information-Directed Sampling Algorithms for Multi-Agent Reinforcement Learning
by: Zhang, Qiaosheng, et al.
Published: (2024) -
Stop Overvaluing Multi-Agent Debate -- We Must Rethink Evaluation and Embrace Model Heterogeneity
by: Zhang, Hangfan, et al.
Published: (2025)