On-line Policy Improvement using Monte-Carlo Search
Fuente:
arXiv
Guardado en:
| Autores principales: | Tesauro, Gerald, Galperin, Gregory R. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning and Improving Backgammon Strategy
por: Galperin, Gregory R.
Publicado: (2025)
por: Galperin, Gregory R.
Publicado: (2025)
Learning Heuristics for Transit Network Design and Improvement with Deep Reinforcement Learning
por: Holliday, Andrew, et al.
Publicado: (2024)
por: Holliday, Andrew, et al.
Publicado: (2024)
Neural Architecture Search using Particle Swarm and Ant Colony Optimization
por: Lankford, Séamus, et al.
Publicado: (2024)
por: Lankford, Séamus, et al.
Publicado: (2024)
On the Improvement of Generalization and Stability of Forward-Only Learning via Neural Polarization
por: Terres-Escudero, Erik B., et al.
Publicado: (2024)
por: Terres-Escudero, Erik B., et al.
Publicado: (2024)
SEval-NAS: A Search-Agnostic Evaluation for Neural Architecture Search
por: Mih, Atah Nuh, et al.
Publicado: (2026)
por: Mih, Atah Nuh, et al.
Publicado: (2026)
Vector Policy Optimization: Training for Diversity Improves Test-Time Search
por: Bahlous-Boldi, Ryan, et al.
Publicado: (2026)
por: Bahlous-Boldi, Ryan, et al.
Publicado: (2026)
MIDAS: Mosaic Input-Specific Differentiable Architecture Search
por: Subbotko, Konstanty
Publicado: (2026)
por: Subbotko, Konstanty
Publicado: (2026)
Knowledge-aware Evolutionary Graph Neural Architecture Search
por: Wang, Chao, et al.
Publicado: (2024)
por: Wang, Chao, et al.
Publicado: (2024)
Learning to Reduce Search Space for Generalizable Neural Routing Solver
por: Zhou, Changliang, et al.
Publicado: (2025)
por: Zhou, Changliang, et al.
Publicado: (2025)
Evolutionary Architecture Search through Grammar-Based Sequence Alignment
por: Martín, Adri Gómez, et al.
Publicado: (2025)
por: Martín, Adri Gómez, et al.
Publicado: (2025)
Contrastive Concept-Tree Search for LLM-Assisted Algorithm Discovery
por: Leleu, Timothee, et al.
Publicado: (2026)
por: Leleu, Timothee, et al.
Publicado: (2026)
Discovering Effective Policies for Land-Use Planning with Neuroevolution
por: Young, Daniel, et al.
Publicado: (2023)
por: Young, Daniel, et al.
Publicado: (2023)
AM-PPO: (Advantage) Alpha-Modulation with Proximal Policy Optimization
por: Sane, Soham
Publicado: (2025)
por: Sane, Soham
Publicado: (2025)
Pareto-NRPA: A Novel Monte-Carlo Search Algorithm for Multi-Objective Optimization
por: Lallouet, Noé, et al.
Publicado: (2025)
por: Lallouet, Noé, et al.
Publicado: (2025)
Can We Optimize Deep RL Policy Weights as Trajectory Modeling?
por: Tang, Hongyao
Publicado: (2025)
por: Tang, Hongyao
Publicado: (2025)
Robust Lagrangian and Adversarial Policy Gradient for Robust Constrained Markov Decision Processes
por: Bossens, David M.
Publicado: (2023)
por: Bossens, David M.
Publicado: (2023)
Large-scale Multi-objective Feature Selection: A Multi-phase Search Space Shrinking Approach
por: Bidgoli, Azam Asilian, et al.
Publicado: (2024)
por: Bidgoli, Azam Asilian, et al.
Publicado: (2024)
QSLM: A Performance- and Memory-aware Quantization Framework with Tiered Search Strategy for Spike-driven Language Models
por: Putra, Rachmad Vidya Wicaksana, et al.
Publicado: (2026)
por: Putra, Rachmad Vidya Wicaksana, et al.
Publicado: (2026)
Space-Time Continuous PDE Forecasting using Equivariant Neural Fields
por: Knigge, David M., et al.
Publicado: (2024)
por: Knigge, David M., et al.
Publicado: (2024)
SpikeNAS: A Fast Memory-Aware Neural Architecture Search Framework for Spiking Neural Network-based Embedded AI Systems
por: Putra, Rachmad Vidya Wicaksana, et al.
Publicado: (2024)
por: Putra, Rachmad Vidya Wicaksana, et al.
Publicado: (2024)
Neural Policy Style Transfer
por: Fernandez-Fernandez, Raul, et al.
Publicado: (2024)
por: Fernandez-Fernandez, Raul, et al.
Publicado: (2024)
Machine Unlearning using Forgetting Neural Networks
por: Hatua, Amartya, et al.
Publicado: (2024)
por: Hatua, Amartya, et al.
Publicado: (2024)
Multi-scale Topology Optimization using Neural Networks
por: Chen, Hongrui, et al.
Publicado: (2024)
por: Chen, Hongrui, et al.
Publicado: (2024)
Recurrent Neural Networks Learn to Store and Generate Sequences using Non-Linear Representations
por: Csordás, Róbert, et al.
Publicado: (2024)
por: Csordás, Róbert, et al.
Publicado: (2024)
DEBUG-HD: Debugging TinyML models on-device using Hyper-Dimensional computing
por: Ghanathe, Nikhil P, et al.
Publicado: (2024)
por: Ghanathe, Nikhil P, et al.
Publicado: (2024)
Interpolating neural network: A novel unification of machine learning and interpolation theory
por: Park, Chanwook, et al.
Publicado: (2024)
por: Park, Chanwook, et al.
Publicado: (2024)
KHNNs: hypercomplex neural networks computations via Keras using TensorFlow and PyTorch
por: Niemczynowicz, Agnieszka, et al.
Publicado: (2024)
por: Niemczynowicz, Agnieszka, et al.
Publicado: (2024)
High Performance Im2win and Direct Convolutions using Three Tensor Layouts on SIMD Architectures
por: Fu, Xiang, et al.
Publicado: (2024)
por: Fu, Xiang, et al.
Publicado: (2024)
An Inverse Modeling Constrained Multi-Objective Evolutionary Algorithm Based on Decomposition
por: Farias, Lucas R. C., et al.
Publicado: (2024)
por: Farias, Lucas R. C., et al.
Publicado: (2024)
Diffusion Policies for Out-of-Distribution Generalization in Offline Reinforcement Learning
por: Ada, Suzan Ece, et al.
Publicado: (2023)
por: Ada, Suzan Ece, et al.
Publicado: (2023)
Evolutionary Extreme Learning Machine of ab-initio Energy Landscapes for Crystal Structure Prediction using Manta Ray Optimization with Levy Flight
por: Rubio-Solis, Adrian
Publicado: (2026)
por: Rubio-Solis, Adrian
Publicado: (2026)
Meta-Evolve: Continuous Robot Evolution for One-to-many Policy Transfer
por: Liu, Xingyu, et al.
Publicado: (2024)
por: Liu, Xingyu, et al.
Publicado: (2024)
Empirical Investigation into Configuring Echo State Networks for Representative Benchmark Problem Domains
por: Weborg, Brooke R., et al.
Publicado: (2025)
por: Weborg, Brooke R., et al.
Publicado: (2025)
Elastic Architecture Search for Efficient Language Models
por: Wang, Shang
Publicado: (2025)
por: Wang, Shang
Publicado: (2025)
Scaling Policy Gradient Quality-Diversity with Massive Parallelization via Behavioral Variations
por: Mitsides, Konstantinos, et al.
Publicado: (2025)
por: Mitsides, Konstantinos, et al.
Publicado: (2025)
PReLU: Yet Another Single-Layer Solution to the XOR Problem
por: Pinto, Rafael C., et al.
Publicado: (2024)
por: Pinto, Rafael C., et al.
Publicado: (2024)
A logical re-conception of neural networks: Hamiltonian bitwise part-whole architecture
por: Bowen, E, et al.
Publicado: (2026)
por: Bowen, E, et al.
Publicado: (2026)
Beyond Uniform Scaling: Exploring Depth Heterogeneity in Neural Architectures
por: T, Akash Guna R., et al.
Publicado: (2024)
por: T, Akash Guna R., et al.
Publicado: (2024)
A Transformer-based Neural Architecture Search Method
por: Wang, Shang, et al.
Publicado: (2025)
por: Wang, Shang, et al.
Publicado: (2025)
On Accelerating Edge AI: Optimizing Resource-Constrained Environments
por: Sander, Jacob, et al.
Publicado: (2025)
por: Sander, Jacob, et al.
Publicado: (2025)
Ejemplares similares
-
Learning and Improving Backgammon Strategy
por: Galperin, Gregory R.
Publicado: (2025) -
Learning Heuristics for Transit Network Design and Improvement with Deep Reinforcement Learning
por: Holliday, Andrew, et al.
Publicado: (2024) -
Neural Architecture Search using Particle Swarm and Ant Colony Optimization
por: Lankford, Séamus, et al.
Publicado: (2024) -
On the Improvement of Generalization and Stability of Forward-Only Learning via Neural Polarization
por: Terres-Escudero, Erik B., et al.
Publicado: (2024) -
SEval-NAS: A Search-Agnostic Evaluation for Neural Architecture Search
por: Mih, Atah Nuh, et al.
Publicado: (2026)