Saved in:
| Main Authors: | Cozma, Andrei, Harris, Landon, Qi, Hairong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.14566 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Proximal Policy Optimization with Adaptive Exploration
by: Lixandru, Andrei
Published: (2024)
by: Lixandru, Andrei
Published: (2024)
Defect Detection in Tire X-Ray Images: Conventional Methods Meet Deep Structures
by: Cozma, Andrei, et al.
Published: (2024)
by: Cozma, Andrei, et al.
Published: (2024)
On-Policy Optimization of ANFIS Policies Using Proximal Policy Optimization
by: Shankar, Kaaustaaub, et al.
Published: (2025)
by: Shankar, Kaaustaaub, et al.
Published: (2025)
Reparameterization Proximal Policy Optimization
by: Zhong, Hai, et al.
Published: (2025)
by: Zhong, Hai, et al.
Published: (2025)
Cross-Scale MAE: A Tale of Multi-Scale Exploitation in Remote Sensing
by: Tang, Maofeng, et al.
Published: (2024)
by: Tang, Maofeng, et al.
Published: (2024)
Complexity-Regularized Proximal Policy Optimization
by: Serfilippi, Luca, et al.
Published: (2025)
by: Serfilippi, Luca, et al.
Published: (2025)
Beyond the Boundaries of Proximal Policy Optimization
by: Tan, Charlie B., et al.
Published: (2024)
by: Tan, Charlie B., et al.
Published: (2024)
OmniFed: A Modular Framework for Configurable Federated Learning from Edge to HPC
by: Tyagi, Sahil, et al.
Published: (2025)
by: Tyagi, Sahil, et al.
Published: (2025)
ESPO: Early-Stopping Proximal Policy Optimization
by: Li, Zihang, et al.
Published: (2026)
by: Li, Zihang, et al.
Published: (2026)
Learning Branching Policies for MILPs with Proximal Policy Optimization
by: Mhamed, Abdelouahed Ben, et al.
Published: (2025)
by: Mhamed, Abdelouahed Ben, et al.
Published: (2025)
Proximal Policy Distillation
by: Spigler, Giacomo
Published: (2024)
by: Spigler, Giacomo
Published: (2024)
Asymmetric Proximal Policy Optimization: mini-critics boost LLM reasoning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
CIM-PPO:Proximal Policy Optimization with Liu-Correntropy Induced Metric
by: Guo, Yunxiao, et al.
Published: (2021)
by: Guo, Yunxiao, et al.
Published: (2021)
A dynamical clipping approach with task feedback for Proximal Policy Optimization
by: Zhang, Ziqi, et al.
Published: (2023)
by: Zhang, Ziqi, et al.
Published: (2023)
Efficient Deep Reinforcement Learning with Predictive Processing Proximal Policy Optimization
by: Küçükoğlu, Burcu, et al.
Published: (2022)
by: Küçükoğlu, Burcu, et al.
Published: (2022)
ExO-PPO: an Extended Off-policy Proximal Policy Optimization Algorithm
by: Wang, Hanyong, et al.
Published: (2026)
by: Wang, Hanyong, et al.
Published: (2026)
Accelerating Proximal Policy Optimization Learning Using Task Prediction for Solving Environments with Delayed Rewards
by: Ahmad, Ahmad, et al.
Published: (2024)
by: Ahmad, Ahmad, et al.
Published: (2024)
AM-PPO: (Advantage) Alpha-Modulation with Proximal Policy Optimization
by: Sane, Soham
Published: (2025)
by: Sane, Soham
Published: (2025)
A Deep Reinforcement Learning Approach to Battery Management in Dairy Farming via Proximal Policy Optimization
by: Ali, Nawazish, et al.
Published: (2024)
by: Ali, Nawazish, et al.
Published: (2024)
Proximal Policy Optimization with Evolutionary Mutations
by: Czworkowski, Casimir, et al.
Published: (2026)
by: Czworkowski, Casimir, et al.
Published: (2026)
Proximal Policy Gradient Arborescence for Quality Diversity Reinforcement Learning
by: Batra, Sumeet, et al.
Published: (2023)
by: Batra, Sumeet, et al.
Published: (2023)
Assigning Credit with Partial Reward Decoupling in Multi-Agent Proximal Policy Optimization
by: Kapoor, Aditya, et al.
Published: (2024)
by: Kapoor, Aditya, et al.
Published: (2024)
PROMA: Projected Microbatch Accumulation for Reference-Free Proximal Policy Updates
by: Abrahamsen, Nilin
Published: (2026)
by: Abrahamsen, Nilin
Published: (2026)
Trust Region Reward Optimization and Proximal Inverse Reward Optimization Algorithm
by: Chen, Yang, et al.
Published: (2025)
by: Chen, Yang, et al.
Published: (2025)
Variational Delayed Policy Optimization
by: Wu, Qingyuan, et al.
Published: (2024)
by: Wu, Qingyuan, et al.
Published: (2024)
Binarizing Physics-Inspired GNNs for Combinatorial Optimization
by: Krutský, Martin, et al.
Published: (2025)
by: Krutský, Martin, et al.
Published: (2025)
Hierarchy-of-Groups Policy Optimization for Long-Horizon Agentic Tasks
by: He, Shuo, et al.
Published: (2026)
by: He, Shuo, et al.
Published: (2026)
Pretrain Value, Not Reward: Decoupled Value Policy Optimization
by: Huang, Chenghua, et al.
Published: (2025)
by: Huang, Chenghua, et al.
Published: (2025)
Orthogonalized Policy Optimization:Policy Optimization as Orthogonal Projection in Hilbert Space
by: Zixian, Wang
Published: (2026)
by: Zixian, Wang
Published: (2026)
Proximity Matters: Local Proximity Enhanced Balancing for Treatment Effect Estimation
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Wasserstein Policy Optimization
by: Pfau, David, et al.
Published: (2025)
by: Pfau, David, et al.
Published: (2025)
Reflective Policy Optimization
by: Gan, Yaozhong, et al.
Published: (2024)
by: Gan, Yaozhong, et al.
Published: (2024)
Belief-Based Offline Reinforcement Learning for Delay-Robust Policy Optimization
by: Zhan, Simon Sinong, et al.
Published: (2025)
by: Zhan, Simon Sinong, et al.
Published: (2025)
Auto-Unrolled Proximal Gradient Descent: An AutoML Approach to Interpretable Waveform Optimization
by: Kaplan, Ahmet
Published: (2026)
by: Kaplan, Ahmet
Published: (2026)
Efficient Dynamics Modeling in Interactive Environments with Koopman Theory
by: Mondal, Arnab Kumar, et al.
Published: (2023)
by: Mondal, Arnab Kumar, et al.
Published: (2023)
RLVR without Ineffective Samples: Group Prioritized Off-Policy Optimization for LLM Reasoning
by: Mao, Yixiu, et al.
Published: (2026)
by: Mao, Yixiu, et al.
Published: (2026)
How the Optimizer Shapes Learned Solutions in Equivariant Neural Networks
by: Stupariu, Teodor-Mihai, et al.
Published: (2026)
by: Stupariu, Teodor-Mihai, et al.
Published: (2026)
Optimizing Anytime Reasoning via Budget Relative Policy Optimization
by: Qi, Penghui, et al.
Published: (2025)
by: Qi, Penghui, et al.
Published: (2025)
HDNet: Physics-Inspired Neural Network for Flow Estimation based on Helmholtz Decomposition
by: Qi, Miao, et al.
Published: (2024)
by: Qi, Miao, et al.
Published: (2024)
Momentum Streams for Optimizer-Inspired Transformers
by: Gai, Jingchu, et al.
Published: (2026)
by: Gai, Jingchu, et al.
Published: (2026)
Similar Items
-
Proximal Policy Optimization with Adaptive Exploration
by: Lixandru, Andrei
Published: (2024) -
Defect Detection in Tire X-Ray Images: Conventional Methods Meet Deep Structures
by: Cozma, Andrei, et al.
Published: (2024) -
On-Policy Optimization of ANFIS Policies Using Proximal Policy Optimization
by: Shankar, Kaaustaaub, et al.
Published: (2025) -
Reparameterization Proximal Policy Optimization
by: Zhong, Hai, et al.
Published: (2025) -
Cross-Scale MAE: A Tale of Multi-Scale Exploitation in Remote Sensing
by: Tang, Maofeng, et al.
Published: (2024)