Herd: Using multiple, smaller LLMs to match the performances of proprietary, large LLMs via an intelligent composer
Fuente:
arXiv
Saved in:
| Main Authors: | Hari, Surya Narayanan, Liu, Rex, Thomson, Matt |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Leveraging Open-Source Large Language Models for encoding Social Determinants of Health using an Intelligent Router
by: Goel, Akul, et al.
Published: (2024)
by: Goel, Akul, et al.
Published: (2024)
Economic Evaluation of LLMs
by: Zellinger, Michael J., et al.
Published: (2025)
by: Zellinger, Michael J., et al.
Published: (2025)
Efficiently Deploying LLMs with Controlled Risk
by: Zellinger, Michael J., et al.
Published: (2024)
by: Zellinger, Michael J., et al.
Published: (2024)
Cost-Saving LLM Cascades with Early Abstention
by: Zellinger, Michael J., et al.
Published: (2025)
by: Zellinger, Michael J., et al.
Published: (2025)
Fail Fast, or Ask: Mitigating the Deficiencies of Reasoning LLMs with Human-in-the-Loop Systems Engineering
by: Zellinger, Michael J., et al.
Published: (2025)
by: Zellinger, Michael J., et al.
Published: (2025)
A theory of understanding for artificial intelligence: composability, catalysts, and learning
by: Zhang, Zijian, et al.
Published: (2024)
by: Zhang, Zijian, et al.
Published: (2024)
Knowledge-Guided Textual Reasoning for Explainable Video Anomaly Detection via LLMs
by: Lee, Hari
Published: (2025)
by: Lee, Hari
Published: (2025)
Rational Tuning of LLM Cascades via Probabilistic Modeling
by: Zellinger, Michael J., et al.
Published: (2025)
by: Zellinger, Michael J., et al.
Published: (2025)
LLMs Can Assist with Proposal Selection at Large User Facilities
by: Ding, Lijie, et al.
Published: (2025)
by: Ding, Lijie, et al.
Published: (2025)
Position: Ensuring mutual privacy is necessary for effective external evaluation of proprietary AI systems
by: Bucknall, Ben, et al.
Published: (2025)
by: Bucknall, Ben, et al.
Published: (2025)
Blueprint-Bench: Comparing spatial intelligence of LLMs, agents and image models
by: Petersson, Lukas, et al.
Published: (2025)
by: Petersson, Lukas, et al.
Published: (2025)
Relationship-Aware Safety Unlearning for Multimodal LLMs
by: Anilkumar, Vishnu Narayanan, et al.
Published: (2026)
by: Anilkumar, Vishnu Narayanan, et al.
Published: (2026)
Reinforcing privacy reasoning in LLMs via normative simulacra from fiction
by: Franchi, Matt, et al.
Published: (2026)
by: Franchi, Matt, et al.
Published: (2026)
CON-QA: Privacy-Preserving QA using cloud LLMs in Contract Domain
by: Singh, Ajeet Kumar, et al.
Published: (2025)
by: Singh, Ajeet Kumar, et al.
Published: (2025)
Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs
by: Cullen, Carissa, et al.
Published: (2026)
by: Cullen, Carissa, et al.
Published: (2026)
Counterfactual Evaluation Reveals Hidden Capability Profiles in Clinical LLMs and Agents
by: Turk, Matt
Published: (2026)
by: Turk, Matt
Published: (2026)
Cache What Lasts: Token Retention for Memory-Bounded KV Cache in LLMs
by: Bui, Ngoc, et al.
Published: (2025)
by: Bui, Ngoc, et al.
Published: (2025)
Promises and pitfalls of artificial intelligence for legal applications
by: Kapoor, Sayash, et al.
Published: (2024)
by: Kapoor, Sayash, et al.
Published: (2024)
The promise and limits of LLMs in constructing proofs and hints for logic problems in intelligent tutoring systems
by: Tithi, Sutapa Dey, et al.
Published: (2025)
by: Tithi, Sutapa Dey, et al.
Published: (2025)
Tele-LLMs: A Series of Specialized Large Language Models for Telecommunications
by: Maatouk, Ali, et al.
Published: (2024)
by: Maatouk, Ali, et al.
Published: (2024)
PEFA-AI: Advancing Open-source LLMs for RTL generation using Progressive Error Feedback Agentic-AI
by: Narayanan, Athma, et al.
Published: (2025)
by: Narayanan, Athma, et al.
Published: (2025)
Analytical and Empirical Study of Herding Effects in Recommendation Systems
by: Xie, Hong, et al.
Published: (2024)
by: Xie, Hong, et al.
Published: (2024)
Can LLMs perform structured graph reasoning?
by: Agrawal, Palaash, et al.
Published: (2024)
by: Agrawal, Palaash, et al.
Published: (2024)
The Art of Scaling Reinforcement Learning Compute for LLMs
by: Khatri, Devvrit, et al.
Published: (2025)
by: Khatri, Devvrit, et al.
Published: (2025)
AXCEL: Automated eXplainable Consistency Evaluation using LLMs
by: Sreekar, P Aditya, et al.
Published: (2024)
by: Sreekar, P Aditya, et al.
Published: (2024)
A Comprehensive Survey of Bias in LLMs: Current Landscape and Future Directions
by: Ranjan, Rajesh, et al.
Published: (2024)
by: Ranjan, Rajesh, et al.
Published: (2024)
Semantic Partial Grounding via LLMs
by: Canonaco, Giuseppe, et al.
Published: (2026)
by: Canonaco, Giuseppe, et al.
Published: (2026)
Can formal argumentative reasoning enhance LLMs performances?
by: Castagna, Federico, et al.
Published: (2024)
by: Castagna, Federico, et al.
Published: (2024)
An intelligent tutor for planning in large partially observable environments
by: Heindrich, Lovis, et al.
Published: (2023)
by: Heindrich, Lovis, et al.
Published: (2023)
Multi-Agent Learning Path Planning via LLMs
by: Xu, Haoxin, et al.
Published: (2026)
by: Xu, Haoxin, et al.
Published: (2026)
Auxiliary task demands mask the capabilities of smaller language models
by: Hu, Jennifer, et al.
Published: (2024)
by: Hu, Jennifer, et al.
Published: (2024)
LLMs can Realize Combinatorial Creativity: Generating Creative Ideas via LLMs for Scientific Research
by: Gu, Tianyang, et al.
Published: (2024)
by: Gu, Tianyang, et al.
Published: (2024)
Guided Persona-based AI Surveys: Can we replicate personal mobility preferences at scale using LLMs?
by: Tzachristas, Ioannis, et al.
Published: (2025)
by: Tzachristas, Ioannis, et al.
Published: (2025)
Aligning Paralinguistic Understanding and Generation in Speech LLMs via Multi-Task Reinforcement Learning
by: Chen, Jingxiang, et al.
Published: (2026)
by: Chen, Jingxiang, et al.
Published: (2026)
Jailbreak Detection in Clinical Training LLMs Using Feature-Based Predictive Models
by: Nguyen, Tri, et al.
Published: (2025)
by: Nguyen, Tri, et al.
Published: (2025)
MoleCode unlocks structural intelligence in large language models
by: Yan, Zhiyuan, et al.
Published: (2026)
by: Yan, Zhiyuan, et al.
Published: (2026)
Trust & Safety of LLMs and LLMs in Trust & Safety
by: You, Doohee, et al.
Published: (2024)
by: You, Doohee, et al.
Published: (2024)
Using LLMs to Discover Legal Factors
by: Gray, Morgan, et al.
Published: (2024)
by: Gray, Morgan, et al.
Published: (2024)
The Llama 3 Herd of Models
by: Grattafiori, Aaron, et al.
Published: (2024)
by: Grattafiori, Aaron, et al.
Published: (2024)
On the Roles of LLMs in Planning: Embedding LLMs into Planning Graphs
by: Zhuo, Hankz Hankui, et al.
Published: (2024)
by: Zhuo, Hankz Hankui, et al.
Published: (2024)
Similar Items
-
Leveraging Open-Source Large Language Models for encoding Social Determinants of Health using an Intelligent Router
by: Goel, Akul, et al.
Published: (2024) -
Economic Evaluation of LLMs
by: Zellinger, Michael J., et al.
Published: (2025) -
Efficiently Deploying LLMs with Controlled Risk
by: Zellinger, Michael J., et al.
Published: (2024) -
Cost-Saving LLM Cascades with Early Abstention
by: Zellinger, Michael J., et al.
Published: (2025) -
Fail Fast, or Ask: Mitigating the Deficiencies of Reasoning LLMs with Human-in-the-Loop Systems Engineering
by: Zellinger, Michael J., et al.
Published: (2025)