An Autonomous Non-monolithic Agent with Multi-mode Exploration based on Options Framework
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, JaeYoon, Xuan, Junyu, Liang, Christy, Hussain, Farookh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Decoupling Exploration and Exploitation for Unsupervised Pre-training with Successor Features
von: Kim, JaeYoon, et al.
Veröffentlicht: (2024)
von: Kim, JaeYoon, et al.
Veröffentlicht: (2024)
A Non-Monolithic Policy Approach of Offline-to-Online Reinforcement Learning
von: Kim, JaeYoon, et al.
Veröffentlicht: (2024)
von: Kim, JaeYoon, et al.
Veröffentlicht: (2024)
Group-Aware Coordination Graph for Multi-Agent Reinforcement Learning
von: Duan, Wei, et al.
Veröffentlicht: (2024)
von: Duan, Wei, et al.
Veröffentlicht: (2024)
EffGen: Enabling Small Language Models as Capable Autonomous Agents
von: Srivastava, Gaurav, et al.
Veröffentlicht: (2026)
von: Srivastava, Gaurav, et al.
Veröffentlicht: (2026)
APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents
von: Li, Yibo, et al.
Veröffentlicht: (2026)
von: Li, Yibo, et al.
Veröffentlicht: (2026)
Heterogeneous Information-Bottleneck Coordination Graphs for Multi-Agent Reinforcement Learning
von: Duan, Wei, et al.
Veröffentlicht: (2026)
von: Duan, Wei, et al.
Veröffentlicht: (2026)
A Behavior-Aware Approach for Deep Reinforcement Learning in Non-stationary Environments without Known Change Points
von: Liu, Zihe, et al.
Veröffentlicht: (2024)
von: Liu, Zihe, et al.
Veröffentlicht: (2024)
OptionZero: Planning with Learned Options
von: Huang, Po-Wei, et al.
Veröffentlicht: (2025)
von: Huang, Po-Wei, et al.
Veröffentlicht: (2025)
Toward Autonomous UI Exploration: The UIExplorer Benchmark
von: Nica, Andrei Cristian, et al.
Veröffentlicht: (2025)
von: Nica, Andrei Cristian, et al.
Veröffentlicht: (2025)
Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration
von: Nimonkar, Chirayu, et al.
Veröffentlicht: (2025)
von: Nimonkar, Chirayu, et al.
Veröffentlicht: (2025)
Mitigating the Likelihood Paradox in Flow-based OOD Detection via Entropy Manipulation
von: Kim, Donghwan, et al.
Veröffentlicht: (2026)
von: Kim, Donghwan, et al.
Veröffentlicht: (2026)
Toward Autonomous Engineering Design: A Knowledge-Guided Multi-Agent Framework
von: Kumar, Varun, et al.
Veröffentlicht: (2025)
von: Kumar, Varun, et al.
Veröffentlicht: (2025)
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
von: Nair, Lakshmi, et al.
Veröffentlicht: (2025)
von: Nair, Lakshmi, et al.
Veröffentlicht: (2025)
Battling the Non-stationarity in Time Series Forecasting via Test-time Adaptation
von: Kim, HyunGi, et al.
Veröffentlicht: (2025)
von: Kim, HyunGi, et al.
Veröffentlicht: (2025)
PaperOrchestra: A Multi-Agent Framework for Automated AI Research Paper Writing
von: Song, Yiwen, et al.
Veröffentlicht: (2026)
von: Song, Yiwen, et al.
Veröffentlicht: (2026)
From Imitation to Exploration: End-to-end Autonomous Driving based on World Model
von: Li, Yueyuan, et al.
Veröffentlicht: (2024)
von: Li, Yueyuan, et al.
Veröffentlicht: (2024)
Diffusion Augmented Agents: A Framework for Efficient Exploration and Transfer Learning
von: Di Palo, Norman, et al.
Veröffentlicht: (2024)
von: Di Palo, Norman, et al.
Veröffentlicht: (2024)
Resolve Highway Conflict in Multi-Autonomous Vehicle Controls with Local State Attention
von: Ta, Xuan Duy, et al.
Veröffentlicht: (2025)
von: Ta, Xuan Duy, et al.
Veröffentlicht: (2025)
TCOD: Exploring Temporal Curriculum in On-Policy Distillation for Multi-turn Autonomous Agents
von: Wang, Jiaqi, et al.
Veröffentlicht: (2026)
von: Wang, Jiaqi, et al.
Veröffentlicht: (2026)
MESA: Cooperative Meta-Exploration in Multi-Agent Learning through Exploiting State-Action Space Structure
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2024)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2024)
Atomic Action Slicing: Planner-Aligned Options for Generalist VLA Agents
von: Tabakov, Stefan, et al.
Veröffentlicht: (2025)
von: Tabakov, Stefan, et al.
Veröffentlicht: (2025)
MIR: Efficient Exploration in Episodic Multi-Agent Reinforcement Learning via Mutual Intrinsic Reward
von: Chen, Kesheng, et al.
Veröffentlicht: (2025)
von: Chen, Kesheng, et al.
Veröffentlicht: (2025)
Comprehend, Divide, and Conquer: Feature Subspace Exploration via Multi-Agent Hierarchical Reinforcement Learning
von: Zhang, Weiliang, et al.
Veröffentlicht: (2025)
von: Zhang, Weiliang, et al.
Veröffentlicht: (2025)
ScholarPeer: A Context-Aware Multi-Agent Framework for Automated Peer Review
von: Goyal, Palash, et al.
Veröffentlicht: (2026)
von: Goyal, Palash, et al.
Veröffentlicht: (2026)
ORBIT: On-policy Exploration-Exploitation for Controllable Multi-Budget Reasoning
von: Liang, Kun, et al.
Veröffentlicht: (2026)
von: Liang, Kun, et al.
Veröffentlicht: (2026)
AlphaLab: Autonomous Multi-Agent Research Across Optimization Domains with Frontier LLMs
von: Hogan, Brendan R., et al.
Veröffentlicht: (2026)
von: Hogan, Brendan R., et al.
Veröffentlicht: (2026)
Meta-RL Induces Exploration in Language Agents
von: Jiang, Yulun, et al.
Veröffentlicht: (2025)
von: Jiang, Yulun, et al.
Veröffentlicht: (2025)
Incentivized Exploration of Non-Stationary Stochastic Bandits
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2024)
von: Chakraborty, Sourav, et al.
Veröffentlicht: (2024)
AMPED: Adaptive Multi-objective Projection for balancing Exploration and skill Diversification
von: Cho, Geonwoo, et al.
Veröffentlicht: (2025)
von: Cho, Geonwoo, et al.
Veröffentlicht: (2025)
Settling Decentralized Multi-Agent Coordinated Exploration by Novelty Sharing
von: Jiang, Haobin, et al.
Veröffentlicht: (2024)
von: Jiang, Haobin, et al.
Veröffentlicht: (2024)
Diversity-Enriched Option-Critic
von: Kamat, Anand, et al.
Veröffentlicht: (2020)
von: Kamat, Anand, et al.
Veröffentlicht: (2020)
Unveiling Options with Neural Decomposition
von: Alikhasi, Mahdi, et al.
Veröffentlicht: (2024)
von: Alikhasi, Mahdi, et al.
Veröffentlicht: (2024)
Trajectory Imputation in Multi-Agent Sports with Derivative-Accumulating Self-Ensemble
von: Choi, Han-Jun, et al.
Veröffentlicht: (2024)
von: Choi, Han-Jun, et al.
Veröffentlicht: (2024)
Preference-based Reinforcement Learning beyond Pairwise Comparisons: Benefits of Multiple Options
von: Lee, Joongkyu, et al.
Veröffentlicht: (2025)
von: Lee, Joongkyu, et al.
Veröffentlicht: (2025)
Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
von: Putta, Pranav, et al.
Veröffentlicht: (2024)
von: Putta, Pranav, et al.
Veröffentlicht: (2024)
RDAR: Reward-Driven Agent Relevance Estimation for Autonomous Driving
von: Bosio, Carlo, et al.
Veröffentlicht: (2025)
von: Bosio, Carlo, et al.
Veröffentlicht: (2025)
Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments
von: Su, Hongjin, et al.
Veröffentlicht: (2025)
von: Su, Hongjin, et al.
Veröffentlicht: (2025)
Reason for Future, Act for Now: A Principled Framework for Autonomous LLM Agents with Provable Sample Efficiency
von: Liu, Zhihan, et al.
Veröffentlicht: (2023)
von: Liu, Zhihan, et al.
Veröffentlicht: (2023)
REX: Rapid Exploration and eXploitation for AI Agents
von: Murthy, Rithesh, et al.
Veröffentlicht: (2023)
von: Murthy, Rithesh, et al.
Veröffentlicht: (2023)
Toward Efficient Exploration by Large Language Model Agents
von: Arumugam, Dilip, et al.
Veröffentlicht: (2025)
von: Arumugam, Dilip, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Decoupling Exploration and Exploitation for Unsupervised Pre-training with Successor Features
von: Kim, JaeYoon, et al.
Veröffentlicht: (2024) -
A Non-Monolithic Policy Approach of Offline-to-Online Reinforcement Learning
von: Kim, JaeYoon, et al.
Veröffentlicht: (2024) -
Group-Aware Coordination Graph for Multi-Agent Reinforcement Learning
von: Duan, Wei, et al.
Veröffentlicht: (2024) -
EffGen: Enabling Small Language Models as Capable Autonomous Agents
von: Srivastava, Gaurav, et al.
Veröffentlicht: (2026) -
APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents
von: Li, Yibo, et al.
Veröffentlicht: (2026)