Advancing Multi-Agent RAG Systems with Minimalist Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Yihong, Ma, Liheng, Li, Muzhi, Zhou, Jiaming, Ding, Lei, Hao, Jianye, Leung, Ho-fung, King, Irwin, Zhang, Yingxue, Nie, Jian-Yun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Evidence to Trajectory: Abductive Reasoning Path Synthesis for Training Retrieval-Augmented Generation Agents
by: Li, Muzhi, et al.
Published: (2025)
by: Li, Muzhi, et al.
Published: (2025)
The Integration of Semantic and Structural Knowledge in Knowledge Graph Entity Typing
by: Li, Muzhi, et al.
Published: (2024)
by: Li, Muzhi, et al.
Published: (2024)
An Entity Linking Agent for Question Answering
by: Luo, Yajie, et al.
Published: (2025)
by: Luo, Yajie, et al.
Published: (2025)
Retrieval, Reasoning, Re-ranking: A Context-Enriched Framework for Knowledge Graph Completion
by: Li, Muzhi, et al.
Published: (2024)
by: Li, Muzhi, et al.
Published: (2024)
Context-aware Inductive Knowledge Graph Completion with Latent Type Constraints and Subgraph Reasoning
by: Li, Muzhi, et al.
Published: (2024)
by: Li, Muzhi, et al.
Published: (2024)
ADRA-Bank: A Modular Benchmark for Academic Deep Research Agents
by: Guo, Zhihan, et al.
Published: (2025)
by: Guo, Zhihan, et al.
Published: (2025)
It Takes Two: Your GRPO Is Secretly DPO
by: Wu, Yihong, et al.
Published: (2025)
by: Wu, Yihong, et al.
Published: (2025)
3D-MoE: A Mixture-of-Experts Multi-modal LLM for 3D Vision and Pose Diffusion via Rectified Flow
by: Ma, Yueen, et al.
Published: (2025)
by: Ma, Yueen, et al.
Published: (2025)
A Multi-Armed Bandit Approach to Online Selection and Evaluation of Generative Models
by: Hu, Xiaoyan, et al.
Published: (2024)
by: Hu, Xiaoyan, et al.
Published: (2024)
PAK-UCB Contextual Bandit: An Online Learning Approach to Prompt-Aware Selection of Generative Models and LLMs
by: Hu, Xiaoyan, et al.
Published: (2024)
by: Hu, Xiaoyan, et al.
Published: (2024)
An Information Theoretic Approach to Interaction-Grounded Learning
by: Hu, Xiaoyan, et al.
Published: (2024)
by: Hu, Xiaoyan, et al.
Published: (2024)
A Survey on Vision-Language-Action Models for Embodied AI
by: Ma, Yueen, et al.
Published: (2024)
by: Ma, Yueen, et al.
Published: (2024)
Omni-Thinker: Scaling Multi-Task RL in LLMs with Hybrid Reward and Task Scheduling
by: Li, Derek, et al.
Published: (2025)
by: Li, Derek, et al.
Published: (2025)
Enhancing Logical Reasoning in Large Language Models through Graph-based Synthetic Data
by: Zhou, Jiaming, et al.
Published: (2024)
by: Zhou, Jiaming, et al.
Published: (2024)
PromptWise: Online Learning for Cost-Aware Prompt Assignment in Generative Models
by: Hu, Xiaoyan, et al.
Published: (2025)
by: Hu, Xiaoyan, et al.
Published: (2025)
CKGConv: General Graph Convolution with Continuous Kernels
by: Ma, Liheng, et al.
Published: (2024)
by: Ma, Liheng, et al.
Published: (2024)
Multi-resolution Time-Series Transformer for Long-term Forecasting
by: Zhang, Yitian, et al.
Published: (2023)
by: Zhang, Yitian, et al.
Published: (2023)
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning
by: Wu, Qingyuan, et al.
Published: (2025)
by: Wu, Qingyuan, et al.
Published: (2025)
B3C: A Minimalist Approach to Offline Multi-Agent Reinforcement Learning
by: Kim, Woojun, et al.
Published: (2025)
by: Kim, Woojun, et al.
Published: (2025)
XP-MARL: Auxiliary Prioritization in Multi-Agent Reinforcement Learning to Address Non-Stationarity
by: Xu, Jianye, et al.
Published: (2024)
by: Xu, Jianye, et al.
Published: (2024)
VeritasFi: An Adaptable, Multi-tiered RAG Framework for Multi-modal Financial Question Answering
by: Tai, Zhenghan, et al.
Published: (2025)
by: Tai, Zhenghan, et al.
Published: (2025)
GraphPPD: Posterior Predictive Modelling for Graph-Level Inference
by: Pal, Soumyasundar, et al.
Published: (2025)
by: Pal, Soumyasundar, et al.
Published: (2025)
Effects of Cucumis and Cucurbita rootstocks on vegetative traits, yield and quality in ‘Tainan No. 1’ cucumber
by: Hsiu-fung Chao
Published: (2013)
by: Hsiu-fung Chao
Published: (2013)
SigmaRL: A Sample-Efficient and Generalizable Multi-Agent Reinforcement Learning Framework for Motion Planning
by: Xu, Jianye, et al.
Published: (2024)
by: Xu, Jianye, et al.
Published: (2024)
FinSage: A Multi-aspect RAG System for Financial Filings Question Answering
by: Wang, Xinyu, et al.
Published: (2025)
by: Wang, Xinyu, et al.
Published: (2025)
Collaborative Safe Formation Control for Coupled Multi-Agent Systems
by: Butler, Brooks A., et al.
Published: (2023)
by: Butler, Brooks A., et al.
Published: (2023)
Plain Transformers Can be Powerful Graph Learners
by: Ma, Liheng, et al.
Published: (2025)
by: Ma, Liheng, et al.
Published: (2025)
Exploring the Best Practices of Query Expansion with Large Language Models
by: Zhang, Le, et al.
Published: (2024)
by: Zhang, Le, et al.
Published: (2024)
Hierarchical Multi-Agent Reinforcement Learning-based Coordinated Spatial Reuse for Next Generation WLANs
by: Yu, Jiaming, et al.
Published: (2025)
by: Yu, Jiaming, et al.
Published: (2025)
MODULI: Unlocking Preference Generalization via Diffusion Models for Offline Multi-Objective Reinforcement Learning
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
One Demo Is All It Takes: Planning Domain Derivation with LLMs from A Single Demonstration
by: Huang, Jinbang, et al.
Published: (2025)
by: Huang, Jinbang, et al.
Published: (2025)
Event‐Triggered Fixed‐Time Prescribed Performance Containment Control for Switched Nonlinear Multi‐Agent Systems
by: Yingxue Hou, et al.
Published: (2026)
by: Yingxue Hou, et al.
Published: (2026)
Stable Reinforcement Learning for Efficient Reasoning
by: Dai, Muzhi, et al.
Published: (2025)
by: Dai, Muzhi, et al.
Published: (2025)
A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce
by: Xiong, Wei, et al.
Published: (2025)
by: Xiong, Wei, et al.
Published: (2025)
HM-RAG: Hierarchical Multi-Agent Multimodal Retrieval Augmented Generation
by: Liu, Pei, et al.
Published: (2025)
by: Liu, Pei, et al.
Published: (2025)
Unifying Graph Convolution and Contrastive Learning in Collaborative Filtering
by: Wu, Yihong, et al.
Published: (2024)
by: Wu, Yihong, et al.
Published: (2024)
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
by: Wang, Taiyi, et al.
Published: (2024)
by: Wang, Taiyi, et al.
Published: (2024)
Trust-based Consensus in Multi-Agent Reinforcement Learning Systems
by: Fung, Ho Long, et al.
Published: (2022)
by: Fung, Ho Long, et al.
Published: (2022)
S-GRPO: Early Exit via Reinforcement Learning in Reasoning Models
by: Dai, Muzhi, et al.
Published: (2025)
by: Dai, Muzhi, et al.
Published: (2025)
Belief States for Cooperative Multi-Agent Reinforcement Learning under Partial Observability
by: Pritz, Paul J., et al.
Published: (2025)
by: Pritz, Paul J., et al.
Published: (2025)
Similar Items
-
From Evidence to Trajectory: Abductive Reasoning Path Synthesis for Training Retrieval-Augmented Generation Agents
by: Li, Muzhi, et al.
Published: (2025) -
The Integration of Semantic and Structural Knowledge in Knowledge Graph Entity Typing
by: Li, Muzhi, et al.
Published: (2024) -
An Entity Linking Agent for Question Answering
by: Luo, Yajie, et al.
Published: (2025) -
Retrieval, Reasoning, Re-ranking: A Context-Enriched Framework for Knowledge Graph Completion
by: Li, Muzhi, et al.
Published: (2024) -
Context-aware Inductive Knowledge Graph Completion with Latent Type Constraints and Subgraph Reasoning
by: Li, Muzhi, et al.
Published: (2024)