Mixture-of-Models: Unifying Heterogeneous Agents via N-Way Self-Evaluating Deliberation
Fuente:
arXiv
Guardado en:
| Autores principales: | Pecerskis, Tims, Smirnovs, Aivars |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Heterogeneous Multi-Agent Reinforcement Learning for Zero-Shot Scalable Collaboration
por: Guo, Xudong, et al.
Publicado: (2024)
por: Guo, Xudong, et al.
Publicado: (2024)
Latent World Models for Automated Driving: A Unified Taxonomy, Evaluation Framework, and Open Challenges
por: Zeng, Rongxiang, et al.
Publicado: (2026)
por: Zeng, Rongxiang, et al.
Publicado: (2026)
FORGE: Self-Evolving Agent Memory With No Weight Updates via Population Broadcast
por: Bogdanov, Igor, et al.
Publicado: (2026)
por: Bogdanov, Igor, et al.
Publicado: (2026)
Reliability and Effectiveness of Autonomous AI Agents in Supply Chain Management
por: Long, Carol Xuan, et al.
Publicado: (2026)
por: Long, Carol Xuan, et al.
Publicado: (2026)
IntersectionZoo: Eco-driving for Benchmarking Multi-Agent Contextual Reinforcement Learning
por: Jayawardana, Vindula, et al.
Publicado: (2024)
por: Jayawardana, Vindula, et al.
Publicado: (2024)
Multi-residual Mixture of Experts Learning for Cooperative Control in Multi-vehicle Systems
por: Jayawardana, Vindula, et al.
Publicado: (2025)
por: Jayawardana, Vindula, et al.
Publicado: (2025)
Hierarchical Policy-Gradient Reinforcement Learning for Multi-Agent Shepherding Control of Non-Cohesive Targets
por: Covone, Stefano, et al.
Publicado: (2025)
por: Covone, Stefano, et al.
Publicado: (2025)
Balancing Specialization and Centralization: A Multi-Agent Reinforcement Learning Benchmark for Sequential Industrial Control
por: Maus, Tom, et al.
Publicado: (2025)
por: Maus, Tom, et al.
Publicado: (2025)
EcoFair-CH-MARL: Scalable Constrained Hierarchical Multi-Agent RL with Real-Time Emission Budgets and Fairness Guarantees
por: Alqithami, Saad
Publicado: (2026)
por: Alqithami, Saad
Publicado: (2026)
Belief Engine: Configurable and Inspectable Stance Dynamics in Multi-Agent LLM Deliberation
por: Yang, Joshua C., et al.
Publicado: (2026)
por: Yang, Joshua C., et al.
Publicado: (2026)
Learning a Stable, Safe, Distributed Feedback Controller for a Heterogeneous Platoon of Autonomous Vehicles
por: Shaham, Michael H., et al.
Publicado: (2024)
por: Shaham, Michael H., et al.
Publicado: (2024)
Together We Rise: Optimizing Real-Time Multi-Robot Task Allocation using Coordinated Heterogeneous Plays
por: Pal, Aritra, et al.
Publicado: (2025)
por: Pal, Aritra, et al.
Publicado: (2025)
CityLight: A Neighborhood-inclusive Universal Model for Coordinated City-scale Traffic Signal Control
por: Zeng, Jinwei, et al.
Publicado: (2024)
por: Zeng, Jinwei, et al.
Publicado: (2024)
Interacting Large Language Model Agents. Interpretable Models and Social Learning
por: Jain, Adit, et al.
Publicado: (2024)
por: Jain, Adit, et al.
Publicado: (2024)
Learning Team-Based Navigation: A Review of Deep Reinforcement Learning Techniques for Multi-Agent Pathfinding
por: Chung, Jaehoon, et al.
Publicado: (2023)
por: Chung, Jaehoon, et al.
Publicado: (2023)
JCAS-MARL: Joint Communication and Sensing UAV Networks via Resource-Constrained Multi-Agent Reinforcement Learning
por: Guven, Islam, et al.
Publicado: (2026)
por: Guven, Islam, et al.
Publicado: (2026)
What Makes Local Updates Effective: The Role of Data Heterogeneity and Smoothness
por: Patel, Kumar Kshitij
Publicado: (2025)
por: Patel, Kumar Kshitij
Publicado: (2025)
Hierarchical LLM-Driven Control for HAPS-Assisted UAV Networks: Joint Optimization of Flight and Connectivity
por: Yan, Zijiang, et al.
Publicado: (2026)
por: Yan, Zijiang, et al.
Publicado: (2026)
Systematic Analyses of Reinforcement Learning Controllers in Signalized Urban Corridors
por: Song, Xiaofei, et al.
Publicado: (2026)
por: Song, Xiaofei, et al.
Publicado: (2026)
Multi-agent deep reinforcement learning with centralized training and decentralized execution for transportation infrastructure management
por: Saifullah, M., et al.
Publicado: (2024)
por: Saifullah, M., et al.
Publicado: (2024)
VAE-GAN Based Price Manipulation in Coordinated Local Energy Markets
por: Mukherjee, Biswarup, et al.
Publicado: (2025)
por: Mukherjee, Biswarup, et al.
Publicado: (2025)
LC-Opt: Benchmarking Reinforcement Learning and Agentic AI for End-to-End Liquid Cooling Optimization in Data Centers
por: Naug, Avisek, et al.
Publicado: (2025)
por: Naug, Avisek, et al.
Publicado: (2025)
DCcluster-Opt: Benchmarking Dynamic Multi-Objective Optimization for Geo-Distributed Data Center Workloads
por: Guillen-Perez, Antonio, et al.
Publicado: (2025)
por: Guillen-Perez, Antonio, et al.
Publicado: (2025)
Cooperative Cruising: Reinforcement Learning-Based Time-Headway Control for Increased Traffic Efficiency
por: Veksler, Yaron, et al.
Publicado: (2024)
por: Veksler, Yaron, et al.
Publicado: (2024)
Bayesian Critique-Tune-Based Reinforcement Learning with Adaptive Pressure for Multi-Intersection Traffic Signal Control
por: Duan, Wenchang, et al.
Publicado: (2024)
por: Duan, Wenchang, et al.
Publicado: (2024)
Improving Mixed-Criticality Scheduling with Reinforcement Learning
por: El-Mahdy, Muhammad, et al.
Publicado: (2025)
por: El-Mahdy, Muhammad, et al.
Publicado: (2025)
Carbon Footprint Reduction for Sustainable Data Centers in Real-Time
por: Sarkar, Soumyendu, et al.
Publicado: (2024)
por: Sarkar, Soumyendu, et al.
Publicado: (2024)
Decentralized Coordination of Distributed Energy Resources through Local Energy Markets and Deep Reinforcement Learning
por: May, Daniel, et al.
Publicado: (2024)
por: May, Daniel, et al.
Publicado: (2024)
Generalizing Cooperative Eco-driving via Multi-residual Task Learning
por: Jayawardana, Vindula, et al.
Publicado: (2024)
por: Jayawardana, Vindula, et al.
Publicado: (2024)
Unified Crew Planning and Replanning Optimization in Multi-Line Metro Systems Considering Workforce Heterogeneity
por: Chen, Qihang
Publicado: (2025)
por: Chen, Qihang
Publicado: (2025)
Locally Interdependent Multi-Agent MDP: Theoretical Framework for Decentralized Agents with Dynamic Dependencies
por: DeWeese, Alex, et al.
Publicado: (2024)
por: DeWeese, Alex, et al.
Publicado: (2024)
MoMA: A Mixture-of-Multimodal-Agents Architecture for Enhancing Clinical Prediction Modelling
por: Gao, Jifan, et al.
Publicado: (2025)
por: Gao, Jifan, et al.
Publicado: (2025)
Context, Reasoning, and Hierarchy: A Cost-Performance Study of Compound LLM Agent Design in an Adversarial POMDP
por: Bogdanov, Igor, et al.
Publicado: (2026)
por: Bogdanov, Igor, et al.
Publicado: (2026)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
por: Adibi, Arman, et al.
Publicado: (2024)
por: Adibi, Arman, et al.
Publicado: (2024)
DeMuon: A Decentralized Muon for Matrix Optimization over Graphs
por: He, Chuan, et al.
Publicado: (2025)
por: He, Chuan, et al.
Publicado: (2025)
Thinking Beyond Visibility: A Near-Optimal Policy Framework for Locally Interdependent Multi-Agent MDPs
por: DeWeese, Alex, et al.
Publicado: (2025)
por: DeWeese, Alex, et al.
Publicado: (2025)
Generalized Information Gathering Under Dynamics Uncertainty
por: Palafox, Fernando, et al.
Publicado: (2026)
por: Palafox, Fernando, et al.
Publicado: (2026)
Learn2Drive: A neural network-based framework for socially compliant automated vehicle control
por: Liu, Yuhui, et al.
Publicado: (2025)
por: Liu, Yuhui, et al.
Publicado: (2025)
From Abstraction to Reality: DARPA's Vision for Robust Sim-to-Real Autonomy
por: Noorani, Erfaun, et al.
Publicado: (2025)
por: Noorani, Erfaun, et al.
Publicado: (2025)
Towards Developing Socially Compliant Automated Vehicles: Advances, Expert Insights, and A Conceptual Framework
por: Dong, Yongqi, et al.
Publicado: (2025)
por: Dong, Yongqi, et al.
Publicado: (2025)
Ejemplares similares
-
Heterogeneous Multi-Agent Reinforcement Learning for Zero-Shot Scalable Collaboration
por: Guo, Xudong, et al.
Publicado: (2024) -
Latent World Models for Automated Driving: A Unified Taxonomy, Evaluation Framework, and Open Challenges
por: Zeng, Rongxiang, et al.
Publicado: (2026) -
FORGE: Self-Evolving Agent Memory With No Weight Updates via Population Broadcast
por: Bogdanov, Igor, et al.
Publicado: (2026) -
Reliability and Effectiveness of Autonomous AI Agents in Supply Chain Management
por: Long, Carol Xuan, et al.
Publicado: (2026) -
IntersectionZoo: Eco-driving for Benchmarking Multi-Agent Contextual Reinforcement Learning
por: Jayawardana, Vindula, et al.
Publicado: (2024)