Solving Deep Reinforcement Learning Tasks with Evolution Strategies and Linear Policy Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Wong, Annie, de Nobel, Jacob, Bäck, Thomas, Plaat, Aske, Kononova, Anna V. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reasoning Capabilities of Large Language Models on Dynamic Tasks
by: Wong, Annie, et al.
Published: (2025)
by: Wong, Annie, et al.
Published: (2025)
State Design Matters: How Representations Shape Dynamic Reasoning in Large Language Models
by: Wong, Annie, et al.
Published: (2026)
by: Wong, Annie, et al.
Published: (2026)
Multi-Step Reasoning with Large Language Models, a Survey
by: Plaat, Aske, et al.
Published: (2024)
by: Plaat, Aske, et al.
Published: (2024)
Research Re: search & Re-search
by: Plaat, Aske
Published: (2024)
by: Plaat, Aske
Published: (2024)
Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learning
by: Spoor, Lindsay, et al.
Published: (2025)
by: Spoor, Lindsay, et al.
Published: (2025)
Mirror Mode in Fire Emblem: Beating Players at their own Game with Imitation and Reinforcement Learning
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
by: Smid, Yanna Elizabeth, et al.
Published: (2025)
Reinforcement Learning for Sustainable Energy: A Survey
by: Ponse, Koen, et al.
Published: (2024)
by: Ponse, Koen, et al.
Published: (2024)
Explicitly Disentangled Representations in Object-Centric Learning
by: Majellaro, Riccardo, et al.
Published: (2024)
by: Majellaro, Riccardo, et al.
Published: (2024)
CausalPlayground: Addressing Data-Generation Requirements in Cutting-Edge Causality Research
by: Sauter, Andreas W M, et al.
Published: (2024)
by: Sauter, Andreas W M, et al.
Published: (2024)
ACTIVA: Amortized Causal Effect Estimation via Transformer-based Variational Autoencoder
by: Sauter, Andreas, et al.
Published: (2025)
by: Sauter, Andreas, et al.
Published: (2025)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
by: Sauter, Andreas W. M., et al.
Published: (2024)
by: Sauter, Andreas W. M., et al.
Published: (2024)
Chargax: A JAX Accelerated EV Charging Simulator
by: Ponse, Koen, et al.
Published: (2025)
by: Ponse, Koen, et al.
Published: (2025)
Reset-free Reinforcement Learning with World Models
by: Yang, Zhao, et al.
Published: (2024)
by: Yang, Zhao, et al.
Published: (2024)
A Benchmark Study of Deep Reinforcement Learning Algorithms for the Container Stowage Planning Problem
by: Huang, Yunqi, et al.
Published: (2025)
by: Huang, Yunqi, et al.
Published: (2025)
MA-VAE: Multi-head Attention-based Variational Autoencoder Approach for Anomaly Detection in Multivariate Time-series Applied to Automotive Endurance Powertrain Testing
by: Correia, Lucas, et al.
Published: (2023)
by: Correia, Lucas, et al.
Published: (2023)
Online Model-based Anomaly Detection in Multivariate Time Series: Taxonomy, Survey, Research Challenges and Future Directions
by: Correia, Lucas, et al.
Published: (2024)
by: Correia, Lucas, et al.
Published: (2024)
Code Evolution Graphs: Understanding Large Language Model Driven Design of Algorithms
by: van Stein, Niki, et al.
Published: (2025)
by: van Stein, Niki, et al.
Published: (2025)
PATH: A Discrete-sequence Dataset for Evaluating Online Unsupervised Anomaly Detection Approaches for Multivariate Time Series
by: Correia, Lucas, et al.
Published: (2024)
by: Correia, Lucas, et al.
Published: (2024)
Analysis of Bluffing by DQN and CFR in Leduc Hold'em Poker
by: Zaciragic, Tarik, et al.
Published: (2025)
by: Zaciragic, Tarik, et al.
Published: (2025)
BLADE: Benchmark suite for LLM-driven Automated Design and Evolution of iterative optimisation heuristics
by: van Stein, Niki, et al.
Published: (2025)
by: van Stein, Niki, et al.
Published: (2025)
Abnormal Mutations: Evolution Strategies Don't Require Gaussianity
by: de Nobel, Jacob, et al.
Published: (2025)
by: de Nobel, Jacob, et al.
Published: (2025)
TeVAE: A Variational Autoencoder Approach for Discrete Online Anomaly Detection in Variable-state Multivariate Time-series Data
by: Correia, Lucas, et al.
Published: (2024)
by: Correia, Lucas, et al.
Published: (2024)
EduGym: An Environment and Notebook Suite for Reinforcement Learning Education
by: Moerland, Thomas M., et al.
Published: (2023)
by: Moerland, Thomas M., et al.
Published: (2023)
Assessing Reproducibility in Evolutionary Computation: A Case Study using Human- and LLM-based Assessment
by: Da Ros, Francesca, et al.
Published: (2026)
by: Da Ros, Francesca, et al.
Published: (2026)
From Performance to Understanding: A Vision for Explainable Automated Algorithm Design
by: van Stein, Niki, et al.
Published: (2025)
by: van Stein, Niki, et al.
Published: (2025)
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
by: Chen, Xi, et al.
Published: (2025)
by: Chen, Xi, et al.
Published: (2025)
Baba is LLM: Reasoning in a Game with Dynamic Rules
by: van Wetten, Fien, et al.
Published: (2025)
by: van Wetten, Fien, et al.
Published: (2025)
On the Effect of Regularization in Policy Mirror Descent
by: Kleuker, Jan Felix, et al.
Published: (2025)
by: Kleuker, Jan Felix, et al.
Published: (2025)
Prism: Policy Reuse via Interpretable Strategy Mapping in Reinforcement Learning
by: Pravetz, Thomas
Published: (2026)
by: Pravetz, Thomas
Published: (2026)
A Deep Dive into Effects of Structural Bias on CMA-ES Performance along Affine Trajectories
by: van Stein, Niki, et al.
Published: (2024)
by: van Stein, Niki, et al.
Published: (2024)
CoComposer: LLM Multi-agent Collaborative Music Composition
by: Xing, Peiwen, et al.
Published: (2025)
by: Xing, Peiwen, et al.
Published: (2025)
Planning with a Learned Policy Basis to Optimally Solve Complex Tasks
by: Infante, Guillermo, et al.
Published: (2024)
by: Infante, Guillermo, et al.
Published: (2024)
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks
by: Mayor, Walter, et al.
Published: (2025)
by: Mayor, Walter, et al.
Published: (2025)
Deep Reinforcement Learning for Inventory Networks: Toward Reliable Policy Optimization
by: Alvo, Matias, et al.
Published: (2023)
by: Alvo, Matias, et al.
Published: (2023)
PeersimGym: An Environment for Solving the Task Offloading Problem with Reinforcement Learning
by: Metelo, Frederico, et al.
Published: (2024)
by: Metelo, Frederico, et al.
Published: (2024)
TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing
by: Li, Yuanpeng, et al.
Published: (2026)
by: Li, Yuanpeng, et al.
Published: (2026)
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance
by: He, Jinmin, et al.
Published: (2025)
by: He, Jinmin, et al.
Published: (2025)
Explainable Benchmarking for Iterative Optimization Heuristics
by: van Stein, Niki, et al.
Published: (2024)
by: van Stein, Niki, et al.
Published: (2024)
Multi-Agent Influence Diagrams to Hybrid Threat Modeling
by: Vonk, Maarten C., et al.
Published: (2026)
by: Vonk, Maarten C., et al.
Published: (2026)
Proximal Curriculum with Task Correlations for Deep Reinforcement Learning
by: Tzannetos, Georgios, et al.
Published: (2024)
by: Tzannetos, Georgios, et al.
Published: (2024)
Similar Items
-
Reasoning Capabilities of Large Language Models on Dynamic Tasks
by: Wong, Annie, et al.
Published: (2025) -
State Design Matters: How Representations Shape Dynamic Reasoning in Large Language Models
by: Wong, Annie, et al.
Published: (2026) -
Multi-Step Reasoning with Large Language Models, a Survey
by: Plaat, Aske, et al.
Published: (2024) -
Research Re: search & Re-search
by: Plaat, Aske
Published: (2024) -
Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learning
by: Spoor, Lindsay, et al.
Published: (2025)