A Reinforcement Learning Calibration Benchmark of Tabular Q-Learning and One-Hot DQN on Enumerable Maze Navigation
Fuente:
Zenodo
Gespeichert in:
| 1. Verfasser: | MD Israfeel |
|---|---|
| Format: | Recurso digital |
| Sprache: | Englisch |
| Veröffentlicht: |
Zenodo
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
THE EMERGENCE OF PRAISE AS CONDITIONED REINFORCEMENT AS A FUNCTION OF OBSERVATION IN PRESCHOOL AND SCHOOL AGE CHILDREN
von: R. DOUGLAS GREER
Veröffentlicht: (2008)
von: R. DOUGLAS GREER
Veröffentlicht: (2008)
Journal of Intelligent Systems
Veröffentlicht: (2020)
Veröffentlicht: (2020)
A Reinforcement Learning Solution for Allocating Replicated Fragments in a Distributed Database
von: Abel Rodríguez Morff
Veröffentlicht: (2007)
von: Abel Rodríguez Morff
Veröffentlicht: (2007)
Behavioral variability: a unified notion and some criteria for experimental analysis
von: Rafael Moreno Rodríguez
Veröffentlicht: (2008)
von: Rafael Moreno Rodríguez
Veröffentlicht: (2008)
A Case Study of the Learning Styles in Low-Level Learners in a Private School in Bogotá
von: David Abella
Veröffentlicht: (2006)
von: David Abella
Veröffentlicht: (2006)
Adaptive Network Control with Reinforcement Learning for Edge-IoT
von: Rösner, Manuel
Veröffentlicht: (2026)
von: Rösner, Manuel
Veröffentlicht: (2026)
Generalization over environments in reinforcement learning
von: Andreas Matt
Veröffentlicht: (2003)
von: Andreas Matt
Veröffentlicht: (2003)
A systematic review of machine learning-enhanced metaheuristics for solving capacitated vehicle routing problems
von: Mauricio Maca-Chagüendo
Veröffentlicht: (2024)
von: Mauricio Maca-Chagüendo
Veröffentlicht: (2024)
Heuristique technique pour l'architecture des réacteurs à fusion
von: Couet, Antoine, et al.
Veröffentlicht: (2026)
von: Couet, Antoine, et al.
Veröffentlicht: (2026)
Benchmarking deep reinforcement learning strategies for the optimal scheduling of deficit irrigation systems - Supplementary dataset
von: Schütze, Niels, et al.
Veröffentlicht: (2026)
von: Schütze, Niels, et al.
Veröffentlicht: (2026)
Learning Future Structure in Predictive World Models
von: Yin, Yaming
Veröffentlicht: (2026)
von: Yin, Yaming
Veröffentlicht: (2026)
xgenius: LLM-Oriented Autonomous Research Platform for SLURM Clusters
von: Creus Castanyer, Roger
Veröffentlicht: (2026)
von: Creus Castanyer, Roger
Veröffentlicht: (2026)
Algorithmic Hysteresis: Structural Failure of Decision-Making Under Irreversible Latency
von: Tavella, Danilo
Veröffentlicht: (2026)
von: Tavella, Danilo
Veröffentlicht: (2026)
SAFIRL: Shielded RL with CBF/MPC on Franka-MuJoCo (v0.1.1)
von: Ozoglu, Nihan
Veröffentlicht: (2025)
von: Ozoglu, Nihan
Veröffentlicht: (2025)
AI-Driven Robotic Task Optimization
von: Budihal, Shivaraj
Veröffentlicht: (2025)
von: Budihal, Shivaraj
Veröffentlicht: (2025)
HDGC-hybrid task offloading framework using deep reinforcement learning and genetic algorithms for 6G edge cloud
von: Kaniezhil, Radhakrishnan, et al.
Veröffentlicht: (2026)
von: Kaniezhil, Radhakrishnan, et al.
Veröffentlicht: (2026)
Training AI Agents to Communicate Safely: Reinforcement Learning for Covert Channel Prevention in Inter-Agent Protocols
von: Maio, Anthony D.
Veröffentlicht: (2026)
von: Maio, Anthony D.
Veröffentlicht: (2026)
The effects of swimming exercise on recognition memory for objects and conditioned fear in rats
von: Julia Niehues da Cruz
Veröffentlicht: (2012)
von: Julia Niehues da Cruz
Veröffentlicht: (2012)
Modellfreies Lernen optimaler zeitdiskreter Regelungsstrategien für Fertigungsprozesse mit endlichem Zeithorizont
von: Dornheim, Johannes
Veröffentlicht: (2022)
von: Dornheim, Johannes
Veröffentlicht: (2022)
Commitment Floors for Tipping-Point Commons: Escaping Nash Traps in Multi-Agent Reinforcement Learning
von: Author, Anonymous
Veröffentlicht: (2026)
von: Author, Anonymous
Veröffentlicht: (2026)
Treating Causality as a Reading Rule — Repositioning Causality from "Cause" to Model Optimization —
von: 長嶺, 智
Veröffentlicht: (2025)
von: 長嶺, 智
Veröffentlicht: (2025)
Automatic design of the flexural strengthening of reinforced concrete beams using fiber reinforced polymers (FRP)
von: Rafael Alves de Souza
Veröffentlicht: (2012)
von: Rafael Alves de Souza
Veröffentlicht: (2012)
Response Acquisition with Delayed Conditioned Reinforcement
von: RODRIGO SOSA
Veröffentlicht: (2011)
von: RODRIGO SOSA
Veröffentlicht: (2011)
Timeout from reinforcement: restoring a balance between analysis and application
von: Timothy D. Hackenberg
Veröffentlicht: (2007)
von: Timothy D. Hackenberg
Veröffentlicht: (2007)
Evaluating functional differences between positive and negative reinforcement through preference for parameters of sound manipulation
von: Joseph M. Lambert
Veröffentlicht: (2019)
von: Joseph M. Lambert
Veröffentlicht: (2019)
Results in Control and Optimization
Veröffentlicht: (2021)
Veröffentlicht: (2021)
Machine Learning: Science and Technology
Veröffentlicht: (2020)
Veröffentlicht: (2020)
Journal of Automation and Intelligence
Veröffentlicht: (2025)
Veröffentlicht: (2025)
MULTI-AGENT COOPERATIVE CONTROL ARCHITECTURE FOR AUTONOMOUS INDUSTRIAL ROBOTS IN SMART MANUFACTURING ENVIRONMENTS
von: Min-Jae Kwon, et al.
Veröffentlicht: (2026)
von: Min-Jae Kwon, et al.
Veröffentlicht: (2026)
AgHealth+: Privacy-Aware Agentic AI for IoT Healthcare (PAAI Framework)
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2026)
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2026)
Addressing Exploration Challenges in Sparse Reward Reinforcement Learning Environments via Intrinsic Curiosity Modules and Reward Shaping
von: Elena Rossi
Veröffentlicht: (2026)
von: Elena Rossi
Veröffentlicht: (2026)
Robust Sensor Fusion for Autonomous Navigation in Dynamically Changing Environments
von: Liam O'Connor
Veröffentlicht: (2026)
von: Liam O'Connor
Veröffentlicht: (2026)
Evaluating Resilience of Deep Learning Models
von: Elvis Rojas
Veröffentlicht: (2020)
von: Elvis Rojas
Veröffentlicht: (2020)
Chapter Scheduling Optimization of Electric Ready Mixed Concrete Vehicles Using an Improved Model-Based Reinforcement Learning
von: Chen, Zhengyi, et al.
Veröffentlicht: (2024)
von: Chen, Zhengyi, et al.
Veröffentlicht: (2024)
Scalable and Efficient Distributed Training of Deep Learning Models via Hybrid Parallelism
von: Yuki Tanaka
Veröffentlicht: (2026)
von: Yuki Tanaka
Veröffentlicht: (2026)
Leveraging Graph Neural Networks for Enhanced Node Representation Learning in Sparse Interaction Networks
von: Sasha Petrova
Veröffentlicht: (2026)
von: Sasha Petrova
Veröffentlicht: (2026)
A Semigroup Theory of Governance: Operator Invariance, BayesianEpistemics, and Kernel Safety in Stochastic Systems
von: Zafar, Usman
Veröffentlicht: (2026)
von: Zafar, Usman
Veröffentlicht: (2026)
A Proximal Gradient Framework for Optimization Under Uncertainty with Applications to Sparse Data Representation
von: Jordan Smith
Veröffentlicht: (2026)
von: Jordan Smith
Veröffentlicht: (2026)
Proximal Gradient Methods for Non-Convex Optimization with Applications to Robust Computer Vision
von: Jamie Chen
Veröffentlicht: (2026)
von: Jamie Chen
Veröffentlicht: (2026)
Adaptive Regularization for Robust Optimization Under Data Distribution Shift
von: Yuki Tanaka
Veröffentlicht: (2026)
von: Yuki Tanaka
Veröffentlicht: (2026)
Ähnliche Einträge
-
THE EMERGENCE OF PRAISE AS CONDITIONED REINFORCEMENT AS A FUNCTION OF OBSERVATION IN PRESCHOOL AND SCHOOL AGE CHILDREN
von: R. DOUGLAS GREER
Veröffentlicht: (2008) -
Journal of Intelligent Systems
Veröffentlicht: (2020) -
A Reinforcement Learning Solution for Allocating Replicated Fragments in a Distributed Database
von: Abel Rodríguez Morff
Veröffentlicht: (2007) -
Behavioral variability: a unified notion and some criteria for experimental analysis
von: Rafael Moreno Rodríguez
Veröffentlicht: (2008) -
A Case Study of the Learning Styles in Low-Level Learners in a Private School in Bogotá
von: David Abella
Veröffentlicht: (2006)