Generative Evolutionary Meta-Solver (GEMS): Scalable Surrogate-Free Multi-Agent Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Sharma, Alakh, Trivedi, Gaurish, Bhandari, Kartikey Singh, Sinha, Yash, Kumar, Dhruv, Narang, Pratik, Challa, Jagat Sesh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Trust Regions Sell, But Who's Buying? Overlap Geometry as an Alternative Trust Region for Policy Optimization
by: Trivedi, Gaurish, et al.
Published: (2026)
by: Trivedi, Gaurish, et al.
Published: (2026)
HAEPO: History-Aggregated Exploratory Policy Optimization
by: Trivedi, Gaurish, et al.
Published: (2025)
by: Trivedi, Gaurish, et al.
Published: (2025)
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
by: Tang, Wenjie, et al.
Published: (2026)
by: Tang, Wenjie, et al.
Published: (2026)
Emergent Coordination in Multi-Agent Systems via Pressure Fields and Temporal Decay
by: Rodriguez, Roland
Published: (2026)
by: Rodriguez, Roland
Published: (2026)
FORMICA: Decision-Focused Learning for Communication-Free Multi-Robot Task Allocation
by: Lopez, Antonio, et al.
Published: (2026)
by: Lopez, Antonio, et al.
Published: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
by: Wang, Yuchen, et al.
Published: (2026)
by: Wang, Yuchen, et al.
Published: (2026)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
by: Hong, Yoosung
Published: (2026)
by: Hong, Yoosung
Published: (2026)
Evolutionary Data Theory: On the Similarities between Data Problems and Evolutionary Games
by: Wissgott, Philipp
Published: (2026)
by: Wissgott, Philipp
Published: (2026)
MALTopic: Multi-Agent LLM Topic Modeling Framework
by: Sharma, Yash
Published: (2026)
by: Sharma, Yash
Published: (2026)
EvoMAS: Evolutionary Generation of Multi-Agent Systems
by: Hu, Yuntong, et al.
Published: (2026)
by: Hu, Yuntong, et al.
Published: (2026)
EvoCUA: Evolving Computer Use Agents via Learning from Scalable Synthetic Experience
by: Xue, Taofeng, et al.
Published: (2026)
by: Xue, Taofeng, et al.
Published: (2026)
AgentSpawn: Adaptive Multi-Agent Collaboration Through Dynamic Spawning for Long-Horizon Code Generation
by: Costa, Igor
Published: (2026)
by: Costa, Igor
Published: (2026)
Semantic Risk-Aware Heuristic Planning for Robotic Navigation in Dynamic Environments: An LLM-Inspired Approach
by: Durrani, Hamza Ahmed, et al.
Published: (2026)
by: Durrani, Hamza Ahmed, et al.
Published: (2026)
FSFM: A Biologically-Inspired Framework for Selective Forgetting of Agent Memory
by: Gu, Yingjie, et al.
Published: (2026)
by: Gu, Yingjie, et al.
Published: (2026)
CoMoCAVs: Cohesive Decision-Guided Motion Planning for Connected and Autonomous Vehicles with Multi-Policy Reinforcement Learning
by: Hu, Pan
Published: (2025)
by: Hu, Pan
Published: (2025)
On the Fundamental Limitations of Decentralized Learnable Reward Shaping in Cooperative Multi-Agent Reinforcement Learning
by: Akella, Aditya
Published: (2025)
by: Akella, Aditya
Published: (2025)
Agentic, Context-Aware Risk Intelligence in the Internet of Value
by: Magableh, Basel, et al.
Published: (2026)
by: Magableh, Basel, et al.
Published: (2026)
Agent Capsules: Quality-Gated Granularity Control for Multi-Agent LLM Pipelines
by: Ray, Aninda
Published: (2026)
by: Ray, Aninda
Published: (2026)
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
by: Annapureddy, Sasank
Published: (2026)
by: Annapureddy, Sasank
Published: (2026)
A Scalable Communication Protocol for Networks of Large Language Models
by: Marro, Samuele, et al.
Published: (2024)
by: Marro, Samuele, et al.
Published: (2024)
MEMOA: Massive Mixtures of Online Agents via Mean-Field Decentralized Nash Equilibria
by: Yang, Xuwei, et al.
Published: (2026)
by: Yang, Xuwei, et al.
Published: (2026)
Exploring Robust Multi-Agent Workflows for Environmental Data Management
by: Guan, Boyuan, et al.
Published: (2026)
by: Guan, Boyuan, et al.
Published: (2026)
Memory-Augmented State Machine Prompting: A Novel LLM Agent Framework for Real-Time Strategy Games
by: Qi, Runnan, et al.
Published: (2025)
by: Qi, Runnan, et al.
Published: (2025)
A Simulation of Ageing and Care Accessibility in Italian Inner Areas
by: garrone, Roberto
Published: (2025)
by: garrone, Roberto
Published: (2025)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
by: Costa, Rimom
Published: (2025)
by: Costa, Rimom
Published: (2025)
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
by: Wang, Xiaohua, et al.
Published: (2026)
by: Wang, Xiaohua, et al.
Published: (2026)
Latent Cache Flow: Model-to-Model Communication Without Text
by: Rossi, Maximillian, et al.
Published: (2026)
by: Rossi, Maximillian, et al.
Published: (2026)
A Super-Learner with Large Language Models for Medical Emergency Advising
by: Aityan, Sergey K., et al.
Published: (2025)
by: Aityan, Sergey K., et al.
Published: (2025)
PRISM-Consult: A Panel-of-Experts Architecture for Clinician-Aligned Diagnosis
by: Levine, Lionel, et al.
Published: (2025)
by: Levine, Lionel, et al.
Published: (2025)
Advancing Multimodal Agent Reasoning with Long-Term Neuro-Symbolic Memory
by: Jiang, Rongjie, et al.
Published: (2026)
by: Jiang, Rongjie, et al.
Published: (2026)
Reinforcement Learning for Scalable and Trustworthy Intelligent Systems
by: Lan, Guangchen
Published: (2026)
by: Lan, Guangchen
Published: (2026)
SPECTra: Scalable Multi-Agent Reinforcement Learning with Permutation-Free Networks
by: Park, Hyunwoo, et al.
Published: (2025)
by: Park, Hyunwoo, et al.
Published: (2025)
Applying Cognitive Design Patterns to General LLM Agents
by: Wray, Robert E., et al.
Published: (2025)
by: Wray, Robert E., et al.
Published: (2025)
Privacy Preserving Multi Agent Path Finding
by: Lehman, Rotem Lev, et al.
Published: (2026)
by: Lehman, Rotem Lev, et al.
Published: (2026)
Good to Go: The LOOP Skill Engine That Hits 99% Success and Slashes Token Usage by 99% via One-Shot Recording and Deterministic Replay
by: Wang, Xiaohua, et al.
Published: (2026)
by: Wang, Xiaohua, et al.
Published: (2026)
Collaborative On-Sensor Array Cameras
by: Sun, Jipeng, et al.
Published: (2025)
by: Sun, Jipeng, et al.
Published: (2025)
Learning To Help: Training Models to Assist Legacy Devices
by: Wu, Yu, et al.
Published: (2024)
by: Wu, Yu, et al.
Published: (2024)
ScrapMem: A Bio-inspired Framework for On-device Personalized Agent Memory via Optical Forgetting
by: Chang, Jiale, et al.
Published: (2026)
by: Chang, Jiale, et al.
Published: (2026)
SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing
by: Meng, Zi, et al.
Published: (2026)
by: Meng, Zi, et al.
Published: (2026)
Right-to-Act: A Pre-Execution Non-Compensatory Decision Protocol for AI Systems
by: Lavi, Gadi
Published: (2026)
by: Lavi, Gadi
Published: (2026)
Similar Items
-
Trust Regions Sell, But Who's Buying? Overlap Geometry as an Alternative Trust Region for Policy Optimization
by: Trivedi, Gaurish, et al.
Published: (2026) -
HAEPO: History-Aggregated Exploratory Policy Optimization
by: Trivedi, Gaurish, et al.
Published: (2025) -
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
by: Tang, Wenjie, et al.
Published: (2026) -
Emergent Coordination in Multi-Agent Systems via Pressure Fields and Temporal Decay
by: Rodriguez, Roland
Published: (2026) -
FORMICA: Decision-Focused Learning for Communication-Free Multi-Robot Task Allocation
by: Lopez, Antonio, et al.
Published: (2026)