Proximal Policy Gradient Arborescence for Quality Diversity Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Batra, Sumeet, Tjanaka, Bryon, Fontaine, Matthew C., Petrenko, Aleksei, Nikolaidis, Stefanos, Sukhatme, Gaurav |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Quality Diversity for Robot Learning: Limitations and Future Directions
by: Batra, Sumeet, et al.
Published: (2024)
by: Batra, Sumeet, et al.
Published: (2024)
Discount Model Search for Quality Diversity Optimization in High-Dimensional Measure Spaces
by: Tjanaka, Bryon, et al.
Published: (2026)
by: Tjanaka, Bryon, et al.
Published: (2026)
Zero-Shot Visual Generalization in Robot Manipulation
by: Batra, Sumeet, et al.
Published: (2025)
by: Batra, Sumeet, et al.
Published: (2025)
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
by: Batra, Sumeet, et al.
Published: (2024)
by: Batra, Sumeet, et al.
Published: (2024)
Density Descent for Diversity Optimization
by: Lee, David H., et al.
Published: (2023)
by: Lee, David H., et al.
Published: (2023)
AutoQD: Automatic Discovery of Diverse Behaviors with Quality-Diversity Optimization
by: Hedayatian, Saeed, et al.
Published: (2025)
by: Hedayatian, Saeed, et al.
Published: (2025)
Algorithmic Scenario Generation as Quality Diversity Optimization
by: Nikolaidis, Stefanos
Published: (2024)
by: Nikolaidis, Stefanos
Published: (2024)
Enabling Adaptive Agent Training in Open-Ended Simulators by Targeting Diversity
by: Costales, Robby, et al.
Published: (2024)
by: Costales, Robby, et al.
Published: (2024)
Red-Teaming Vision-Language-Action Models via Quality Diversity Prompt Generation for Robust Robot Policies
by: Srikanth, Siddharth, et al.
Published: (2026)
by: Srikanth, Siddharth, et al.
Published: (2026)
Collision Avoidance and Navigation for a Quadrotor Swarm Using End-to-end Deep Reinforcement Learning
by: Huang, Zhehui, et al.
Published: (2023)
by: Huang, Zhehui, et al.
Published: (2023)
QD-MAPPER: A Quality Diversity Framework to Automatically Evaluate Multi-Agent Path Finding Algorithms in Diverse Maps
by: Qian, Cheng, et al.
Published: (2024)
by: Qian, Cheng, et al.
Published: (2024)
Quality-Diversity Generative Sampling for Learning with Synthetic Data
by: Chang, Allen, et al.
Published: (2023)
by: Chang, Allen, et al.
Published: (2023)
WARPD: World model Assisted Reactive Policy Diffusion
by: Hegde, Shashank, et al.
Published: (2024)
by: Hegde, Shashank, et al.
Published: (2024)
Entropy-Preserving Reinforcement Learning
by: Petrenko, Aleksei, et al.
Published: (2026)
by: Petrenko, Aleksei, et al.
Published: (2026)
Reinforcement Learning for Long-Horizon Interactive LLM Agents
by: Chen, Kevin, et al.
Published: (2025)
by: Chen, Kevin, et al.
Published: (2025)
Contrastive Learning from Exploratory Actions: Leveraging Natural Interactions for Preference Elicitation
by: Dennler, Nathaniel, et al.
Published: (2025)
by: Dennler, Nathaniel, et al.
Published: (2025)
Rethinking Policy Diversity in Ensemble Policy Gradient in Large-Scale Reinforcement Learning
by: Shitanda, Naoki, et al.
Published: (2026)
by: Shitanda, Naoki, et al.
Published: (2026)
Optimization of Edge Directions and Weights for Mixed Guidance Graphs in Lifelong Multi-Agent Path Finding
by: Zhang, Yulun, et al.
Published: (2026)
by: Zhang, Yulun, et al.
Published: (2026)
Soft Quality-Diversity Optimization
by: Hedayatian, Saeed, et al.
Published: (2025)
by: Hedayatian, Saeed, et al.
Published: (2025)
Can Large Language Models Solve Robot Routing?
by: Huang, Zhehui, et al.
Published: (2024)
by: Huang, Zhehui, et al.
Published: (2024)
Efficient Deep Reinforcement Learning with Predictive Processing Proximal Policy Optimization
by: Küçükoğlu, Burcu, et al.
Published: (2022)
by: Küçükoğlu, Burcu, et al.
Published: (2022)
VoxAct-B: Voxel-Based Acting and Stabilizing Policy for Bimanual Manipulation
by: Liu, I-Chun Arthur, et al.
Published: (2024)
by: Liu, I-Chun Arthur, et al.
Published: (2024)
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
by: Melo, Luckeciano C., et al.
Published: (2025)
by: Melo, Luckeciano C., et al.
Published: (2025)
Policy Gradient Methods for Non-Markovian Reinforcement Learning
by: Kar, Avik, et al.
Published: (2026)
by: Kar, Avik, et al.
Published: (2026)
Policy Gradients for Cumulative Prospect Theory in Reinforcement Learning
by: Lepel, Olivier, et al.
Published: (2024)
by: Lepel, Olivier, et al.
Published: (2024)
Improving User Experience in Preference-Based Optimization of Reward Functions for Assistive Robots
by: Dennler, Nathaniel, et al.
Published: (2024)
by: Dennler, Nathaniel, et al.
Published: (2024)
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying
by: Nishimori, Soichiro, et al.
Published: (2026)
by: Nishimori, Soichiro, et al.
Published: (2026)
$K$-Level Policy Gradients for Multi-Agent Reinforcement Learning
by: Reddi, Aryaman, et al.
Published: (2025)
by: Reddi, Aryaman, et al.
Published: (2025)
Towards Interpretable Deep Generative Models via Causal Representation Learning
by: Moran, Gemma E., et al.
Published: (2025)
by: Moran, Gemma E., et al.
Published: (2025)
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
by: Barakat, Anas, et al.
Published: (2024)
by: Barakat, Anas, et al.
Published: (2024)
PG-Rainbow: Using Distributional Reinforcement Learning in Policy Gradient Methods
by: Jeon, WooJae, et al.
Published: (2024)
by: Jeon, WooJae, et al.
Published: (2024)
Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning
by: Reizinger, Patrik, et al.
Published: (2025)
by: Reizinger, Patrik, et al.
Published: (2025)
A Deep Reinforcement Learning Approach to Battery Management in Dairy Farming via Proximal Policy Optimization
by: Ali, Nawazish, et al.
Published: (2024)
by: Ali, Nawazish, et al.
Published: (2024)
Proximal Policy Distillation
by: Spigler, Giacomo
Published: (2024)
by: Spigler, Giacomo
Published: (2024)
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
by: Basu, Debabrota, et al.
Published: (2025)
by: Basu, Debabrota, et al.
Published: (2025)
LPPG-RL: Lexicographically Projected Policy Gradient Reinforcement Learning with Subproblem Exploration
by: Qiu, Ruiyu, et al.
Published: (2025)
by: Qiu, Ruiyu, et al.
Published: (2025)
The Definitive Guide to Policy Gradients in Deep Reinforcement Learning: Theory, Algorithms and Implementations
by: Lehmann, Matthias
Published: (2024)
by: Lehmann, Matthias
Published: (2024)
Scaling Policy Gradient Quality-Diversity with Massive Parallelization via Behavioral Variations
by: Mitsides, Konstantinos, et al.
Published: (2025)
by: Mitsides, Konstantinos, et al.
Published: (2025)
Compositional Coordination for Multi-Robot Teams with Large Language Models
by: Huang, Zhehui, et al.
Published: (2025)
by: Huang, Zhehui, et al.
Published: (2025)
Learning Branching Policies for MILPs with Proximal Policy Optimization
by: Mhamed, Abdelouahed Ben, et al.
Published: (2025)
by: Mhamed, Abdelouahed Ben, et al.
Published: (2025)
Similar Items
-
Quality Diversity for Robot Learning: Limitations and Future Directions
by: Batra, Sumeet, et al.
Published: (2024) -
Discount Model Search for Quality Diversity Optimization in High-Dimensional Measure Spaces
by: Tjanaka, Bryon, et al.
Published: (2026) -
Zero-Shot Visual Generalization in Robot Manipulation
by: Batra, Sumeet, et al.
Published: (2025) -
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
by: Batra, Sumeet, et al.
Published: (2024) -
Density Descent for Diversity Optimization
by: Lee, David H., et al.
Published: (2023)