Intrinsic Vicarious Conditioning for Deep Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Sanchez, Rodney A, Sahin, Ferat, Ororbia, Alex, Heard, Jamison |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Avoiding Death through Fear Intrinsic Conditioning
por: Sanchez, Rodney, et al.
Publicado: (2025)
por: Sanchez, Rodney, et al.
Publicado: (2025)
Formulating Reinforcement Learning for Human-Robot Collaboration through Off-Policy Evaluation
por: Singh, Saurav, et al.
Publicado: (2026)
por: Singh, Saurav, et al.
Publicado: (2026)
Human Comfortability Index Estimation in Industrial Human-Robot Collaboration Task
por: Savur, Celal, et al.
Publicado: (2023)
por: Savur, Celal, et al.
Publicado: (2023)
Robust Speech-Workload Estimation for Intelligent Human-Robot Systems
por: Fortune, Julian, et al.
Publicado: (2025)
por: Fortune, Julian, et al.
Publicado: (2025)
Optimizing Neurorobot Policy under Limited Demonstration Data through Preference Regret
por: Nguyen, Viet Dung, et al.
Publicado: (2026)
por: Nguyen, Viet Dung, et al.
Publicado: (2026)
Contrastive-Signal-Dependent Plasticity: Self-Supervised Learning in Spiking Neural Circuits
por: Ororbia, Alexander
Publicado: (2023)
por: Ororbia, Alexander
Publicado: (2023)
Deep Reinforcement Learning with Hybrid Intrinsic Reward Model
por: Yuan, Mingqi, et al.
Publicado: (2025)
por: Yuan, Mingqi, et al.
Publicado: (2025)
Class Incremental Continual Learning with Self-Organizing Maps and Variational Autoencoders Using Synthetic Replay
por: Thapa, Pujan, et al.
Publicado: (2025)
por: Thapa, Pujan, et al.
Publicado: (2025)
Adaptive Policy Synchronization for Scalable Reinforcement Learning
por: Lafuente-Mercado, Rodney
Publicado: (2025)
por: Lafuente-Mercado, Rodney
Publicado: (2025)
Domain Feature Collapse: Implications for Out-of-Distribution Detection and Solutions
por: Yang, Hong, et al.
Publicado: (2025)
por: Yang, Hong, et al.
Publicado: (2025)
Autotelic Agents with Intrinsically Motivated Goal-Conditioned Reinforcement Learning: a Short Survey
por: Colas, Cédric, et al.
Publicado: (2020)
por: Colas, Cédric, et al.
Publicado: (2020)
Image-Based Deep Reinforcement Learning with Intrinsically Motivated Stimuli: On the Execution of Complex Robotic Tasks
por: Valencia, David, et al.
Publicado: (2024)
por: Valencia, David, et al.
Publicado: (2024)
Beyond the Class Subspace: Teacher-Guided Training for Reliable Out-of-Distribution Detection in Single-Domain Models
por: Yang, Hong, et al.
Publicado: (2026)
por: Yang, Hong, et al.
Publicado: (2026)
Quantifying Memory Use in Reinforcement Learning with Temporal Range
por: Lafuente-Mercado, Rodney, et al.
Publicado: (2025)
por: Lafuente-Mercado, Rodney, et al.
Publicado: (2025)
Minimally Supervised Learning using Topological Projections in Self-Organizing Maps
por: Lyu, Zimeng, et al.
Publicado: (2024)
por: Lyu, Zimeng, et al.
Publicado: (2024)
Latent-Conditioned Policy Gradient for Multi-Objective Deep Reinforcement Learning
por: Kanazawa, Takuya, et al.
Publicado: (2023)
por: Kanazawa, Takuya, et al.
Publicado: (2023)
RLeXplore: Accelerating Research in Intrinsically-Motivated Reinforcement Learning
por: Yuan, Mingqi, et al.
Publicado: (2024)
por: Yuan, Mingqi, et al.
Publicado: (2024)
Neuro-mimetic Task-free Unsupervised Online Learning with Continual Self-Organizing Maps
por: Vaidya, Hitesh, et al.
Publicado: (2024)
por: Vaidya, Hitesh, et al.
Publicado: (2024)
Reward-Conditioned Reinforcement Learning
por: Nauman, Michal, et al.
Publicado: (2026)
por: Nauman, Michal, et al.
Publicado: (2026)
Directly Learning Stock Trading Strategies Through Profit Guided Loss Functions
por: Kar, Devroop, et al.
Publicado: (2025)
por: Kar, Devroop, et al.
Publicado: (2025)
Extending Spike-Timing Dependent Plasticity to Learning Synaptic Delays
por: Dominijanni, Marissa, et al.
Publicado: (2025)
por: Dominijanni, Marissa, et al.
Publicado: (2025)
Verification-Guided Shielding for Deep Reinforcement Learning
por: Corsi, Davide, et al.
Publicado: (2024)
por: Corsi, Davide, et al.
Publicado: (2024)
Adaptive Correlation-Weighted Intrinsic Rewards for Reinforcement Learning
por: Nguyen, Viet Bac, et al.
Publicado: (2026)
por: Nguyen, Viet Bac, et al.
Publicado: (2026)
Surprise-Adaptive Intrinsic Motivation for Unsupervised Reinforcement Learning
por: Hugessen, Adriana, et al.
Publicado: (2024)
por: Hugessen, Adriana, et al.
Publicado: (2024)
Meta-Representational Predictive Coding: Biomimetic Self-Supervised Learning
por: Ororbia, Alexander, et al.
Publicado: (2025)
por: Ororbia, Alexander, et al.
Publicado: (2025)
A Review of Neuroscience-Inspired Machine Learning
por: Ororbia, Alexander, et al.
Publicado: (2024)
por: Ororbia, Alexander, et al.
Publicado: (2024)
Automatic Grid Updates for Kolmogorov-Arnold Networks using Layer Histograms
por: Moody, Jamison, et al.
Publicado: (2025)
por: Moody, Jamison, et al.
Publicado: (2025)
Fostering Intrinsic Motivation in Reinforcement Learning with Pretrained Foundation Models
por: Andres, Alain, et al.
Publicado: (2024)
por: Andres, Alain, et al.
Publicado: (2024)
Efficient Reinforcement Learning for Large Language Models with Intrinsic Exploration
por: Sun, Yan, et al.
Publicado: (2025)
por: Sun, Yan, et al.
Publicado: (2025)
Deep Intrinsic Coregionalization Multi-Output Gaussian Process Surrogate with Active Learning
por: Chang, Chun-Yi, et al.
Publicado: (2025)
por: Chang, Chun-Yi, et al.
Publicado: (2025)
Intrinsic Benefits of Categorical Distributional Loss: Uncertainty-aware Regularized Exploration in Reinforcement Learning
por: Sun, Ke, et al.
Publicado: (2021)
por: Sun, Ke, et al.
Publicado: (2021)
Conditional Sequence Modeling for Safe Reinforcement Learning
por: Bai, Wensong, et al.
Publicado: (2026)
por: Bai, Wensong, et al.
Publicado: (2026)
Towards Interactive Reinforcement Learning with Intrinsic Feedback
por: Poole, Benjamin, et al.
Publicado: (2021)
por: Poole, Benjamin, et al.
Publicado: (2021)
LLM-Driven Intrinsic Motivation for Sparse Reward Reinforcement Learning
por: Quadros, André, et al.
Publicado: (2025)
por: Quadros, André, et al.
Publicado: (2025)
Mitigating Data Scarcity in Spaceflight Applications for Offline Reinforcement Learning Using Physics-Informed Deep Generative Models
por: Ballentine, Alex E., et al.
Publicado: (2026)
por: Ballentine, Alex E., et al.
Publicado: (2026)
PreND: Enhancing Intrinsic Motivation in Reinforcement Learning through Pre-trained Network Distillation
por: Davoodabadi, Mohammadamin, et al.
Publicado: (2024)
por: Davoodabadi, Mohammadamin, et al.
Publicado: (2024)
Neuroplastic Expansion in Deep Reinforcement Learning
por: Liu, Jiashun, et al.
Publicado: (2024)
por: Liu, Jiashun, et al.
Publicado: (2024)
Deep Reinforcement Learning with Swin Transformers
por: Meng, Li, et al.
Publicado: (2022)
por: Meng, Li, et al.
Publicado: (2022)
On the Robustness of Deep Learning-aided Symbol Detectors to Varying Conditions and Imperfect Channel Knowledge
por: Chen, Chin-Hung, et al.
Publicado: (2024)
por: Chen, Chin-Hung, et al.
Publicado: (2024)
Equivariant Goal Conditioned Contrastive Reinforcement Learning
por: Tangri, Arsh, et al.
Publicado: (2025)
por: Tangri, Arsh, et al.
Publicado: (2025)
Ejemplares similares
-
Avoiding Death through Fear Intrinsic Conditioning
por: Sanchez, Rodney, et al.
Publicado: (2025) -
Formulating Reinforcement Learning for Human-Robot Collaboration through Off-Policy Evaluation
por: Singh, Saurav, et al.
Publicado: (2026) -
Human Comfortability Index Estimation in Industrial Human-Robot Collaboration Task
por: Savur, Celal, et al.
Publicado: (2023) -
Robust Speech-Workload Estimation for Intelligent Human-Robot Systems
por: Fortune, Julian, et al.
Publicado: (2025) -
Optimizing Neurorobot Policy under Limited Demonstration Data through Preference Regret
por: Nguyen, Viet Dung, et al.
Publicado: (2026)