Complexity-Regularized Proximal Policy Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Serfilippi, Luca, Franceschelli, Giorgio, Corradi, Antonio, Musolesi, Mirco |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reinforcement Learning for Generative AI: State of the Art, Opportunities and Open Research Challenges
by: Franceschelli, Giorgio, et al.
Published: (2023)
by: Franceschelli, Giorgio, et al.
Published: (2023)
Do Agents Dream of Electric Sheep?: Improving Generalization in Reinforcement Learning through Generative Learning
by: Franceschelli, Giorgio, et al.
Published: (2024)
by: Franceschelli, Giorgio, et al.
Published: (2024)
DiffSampling: Enhancing Diversity and Accuracy in Neural Text Generation
by: Franceschelli, Giorgio, et al.
Published: (2025)
by: Franceschelli, Giorgio, et al.
Published: (2025)
Copyright in Generative Deep Learning
by: Franceschelli, Giorgio, et al.
Published: (2021)
by: Franceschelli, Giorgio, et al.
Published: (2021)
Creativity and Machine Learning: A Survey
by: Franceschelli, Giorgio, et al.
Published: (2021)
by: Franceschelli, Giorgio, et al.
Published: (2021)
DeepCreativity: Measuring Creativity with Deep Learning Techniques
by: Franceschelli, Giorgio, et al.
Published: (2022)
by: Franceschelli, Giorgio, et al.
Published: (2022)
Thinking Outside the (Gray) Box: A Context-Based Score for Assessing Value and Originality in Neural Text Generation
by: Franceschelli, Giorgio, et al.
Published: (2025)
by: Franceschelli, Giorgio, et al.
Published: (2025)
Creative Beam Search: LLM-as-a-Judge For Improving Response Generation
by: Franceschelli, Giorgio, et al.
Published: (2024)
by: Franceschelli, Giorgio, et al.
Published: (2024)
Training Foundation Models as Data Compression: On Information, Model Weights and Copyright Law
by: Franceschelli, Giorgio, et al.
Published: (2024)
by: Franceschelli, Giorgio, et al.
Published: (2024)
On the Creativity of AI Agents
by: Franceschelli, Giorgio, et al.
Published: (2026)
by: Franceschelli, Giorgio, et al.
Published: (2026)
On the Creativity of Large Language Models
by: Franceschelli, Giorgio, et al.
Published: (2023)
by: Franceschelli, Giorgio, et al.
Published: (2023)
Heterogeneous Knowledge for Augmented Modular Reinforcement Learning
by: Wolf, Lorenz, et al.
Published: (2023)
by: Wolf, Lorenz, et al.
Published: (2023)
Graph Reinforcement Learning for Combinatorial Optimization: A Survey and Unifying Perspective
by: Darvariu, Victor-Alexandru, et al.
Published: (2024)
by: Darvariu, Victor-Alexandru, et al.
Published: (2024)
Emergent Semantic Role Understanding in Language Models
by: Griffiths, Carla, et al.
Published: (2026)
by: Griffiths, Carla, et al.
Published: (2026)
CoMIX: A Multi-agent Reinforcement Learning Training Architecture for Efficient Decentralized Coordination and Independent Decision-Making
by: Minelli, Giovanni, et al.
Published: (2023)
by: Minelli, Giovanni, et al.
Published: (2023)
Investigating the Impact of Direct Punishment on the Emergence of Cooperation in Multi-Agent Reinforcement Learning Systems
by: Dasgupta, Nayana, et al.
Published: (2023)
by: Dasgupta, Nayana, et al.
Published: (2023)
Large Language Models are Effective Priors for Causal Graph Discovery
by: Darvariu, Victor-Alexandru, et al.
Published: (2024)
by: Darvariu, Victor-Alexandru, et al.
Published: (2024)
Tree Search in DAG Space with Model-based Reinforcement Learning for Causal Discovery
by: Darvariu, Victor-Alexandru, et al.
Published: (2023)
by: Darvariu, Victor-Alexandru, et al.
Published: (2023)
Reward Model Overoptimisation in Iterated RLHF
by: Wolf, Lorenz, et al.
Published: (2025)
by: Wolf, Lorenz, et al.
Published: (2025)
Partial Information Decomposition for Data Interpretability and Feature Selection
by: Westphal, Charles, et al.
Published: (2024)
by: Westphal, Charles, et al.
Published: (2024)
Moral Alignment for LLM Agents
by: Tennant, Elizaveta, et al.
Published: (2024)
by: Tennant, Elizaveta, et al.
Published: (2024)
Information-Theoretic State Variable Selection for Reinforcement Learning
by: Westphal, Charles, et al.
Published: (2024)
by: Westphal, Charles, et al.
Published: (2024)
On Distributional Reinforcement Learning in Chaotic Dynamical Systems
by: Rudd-Jones, James, et al.
Published: (2026)
by: Rudd-Jones, James, et al.
Published: (2026)
Quantum Chemistry Driven Molecular Inverse Design with Data-free Reinforcement Learning
by: Calcagno, Francesco, et al.
Published: (2025)
by: Calcagno, Francesco, et al.
Published: (2025)
Hybrid Approaches for Moral Value Alignment in AI Agents: a Manifesto
by: Tennant, Elizaveta, et al.
Published: (2023)
by: Tennant, Elizaveta, et al.
Published: (2023)
Dynamics of Moral Behavior in Heterogeneous Populations of Learning Agents
by: Tennant, Elizaveta, et al.
Published: (2024)
by: Tennant, Elizaveta, et al.
Published: (2024)
Graph Neural Modeling of Network Flows
by: Darvariu, Victor-Alexandru, et al.
Published: (2022)
by: Darvariu, Victor-Alexandru, et al.
Published: (2022)
On-Policy Optimization of ANFIS Policies Using Proximal Policy Optimization
by: Shankar, Kaaustaaub, et al.
Published: (2025)
by: Shankar, Kaaustaaub, et al.
Published: (2025)
Reparameterization Proximal Policy Optimization
by: Zhong, Hai, et al.
Published: (2025)
by: Zhong, Hai, et al.
Published: (2025)
(Ir)rationality in AI: State of the Art, Research Challenges and Open Questions
by: Macmillan-Scott, Olivia, et al.
Published: (2023)
by: Macmillan-Scott, Olivia, et al.
Published: (2023)
Beyond the Boundaries of Proximal Policy Optimization
by: Tan, Charlie B., et al.
Published: (2024)
by: Tan, Charlie B., et al.
Published: (2024)
Proximal Policy Optimization with Adaptive Exploration
by: Lixandru, Andrei
Published: (2024)
by: Lixandru, Andrei
Published: (2024)
Reinforcement Learning Discovers Efficient Decentralized Graph Path Search Strategies
by: Pisacane, Alexei, et al.
Published: (2024)
by: Pisacane, Alexei, et al.
Published: (2024)
Opponent Shaping in LLM Agents
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
by: Segura, Marta Emili Garcia, et al.
Published: (2025)
KIPPO: Koopman-Inspired Proximal Policy Optimization
by: Cozma, Andrei, et al.
Published: (2025)
by: Cozma, Andrei, et al.
Published: (2025)
ESPO: Early-Stopping Proximal Policy Optimization
by: Li, Zihang, et al.
Published: (2026)
by: Li, Zihang, et al.
Published: (2026)
Trust-based Consensus in Multi-Agent Reinforcement Learning Systems
by: Fung, Ho Long, et al.
Published: (2022)
by: Fung, Ho Long, et al.
Published: (2022)
Symmetric Behavior Regularized Policy Optimization
by: Zhu, Lingwei, et al.
Published: (2025)
by: Zhu, Lingwei, et al.
Published: (2025)
Learning Branching Policies for MILPs with Proximal Policy Optimization
by: Mhamed, Abdelouahed Ben, et al.
Published: (2025)
by: Mhamed, Abdelouahed Ben, et al.
Published: (2025)
Proximal Policy Distillation
by: Spigler, Giacomo
Published: (2024)
by: Spigler, Giacomo
Published: (2024)
Similar Items
-
Reinforcement Learning for Generative AI: State of the Art, Opportunities and Open Research Challenges
by: Franceschelli, Giorgio, et al.
Published: (2023) -
Do Agents Dream of Electric Sheep?: Improving Generalization in Reinforcement Learning through Generative Learning
by: Franceschelli, Giorgio, et al.
Published: (2024) -
DiffSampling: Enhancing Diversity and Accuracy in Neural Text Generation
by: Franceschelli, Giorgio, et al.
Published: (2025) -
Copyright in Generative Deep Learning
by: Franceschelli, Giorgio, et al.
Published: (2021) -
Creativity and Machine Learning: A Survey
by: Franceschelli, Giorgio, et al.
Published: (2021)