VDSC: Enhancing Exploration Timing with Value Discrepancy and State Counts
Fuente:
arXiv
Saved in:
| Main Authors: | Captari, Marius, Sasso, Remo, Sabatelli, Matthia |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning
by: Falzari, Massimiliano, et al.
Published: (2025)
by: Falzari, Massimiliano, et al.
Published: (2025)
On The Presence of Double-Descent in Deep Reinforcement Learning
by: Veselý, Viktor, et al.
Published: (2025)
by: Veselý, Viktor, et al.
Published: (2025)
Sparsity-Driven Plasticity in Multi-Task Reinforcement Learning
by: Todorov, Aleksandar, et al.
Published: (2025)
by: Todorov, Aleksandar, et al.
Published: (2025)
Mapping Transformer Leveraged Embeddings for Cross-Lingual Document Representation
by: Tashu, Tsegaye Misikir, et al.
Published: (2024)
by: Tashu, Tsegaye Misikir, et al.
Published: (2024)
Exploration with Foundation Models: Capabilities, Limitations, and Hybrid Approaches
by: Sasso, Remo, et al.
Published: (2025)
by: Sasso, Remo, et al.
Published: (2025)
Large-image Object Detection for Fine-grained Recognition of Punches Patterns in Medieval Panel Painting
by: Bruegger, Josh, et al.
Published: (2025)
by: Bruegger, Josh, et al.
Published: (2025)
Smaller Batches, Bigger Gains? Investigating the Impact of Batch Sizes on Reinforcement Learning Based Real-World Production Scheduling
by: Müller, Arthur, et al.
Published: (2024)
by: Müller, Arthur, et al.
Published: (2024)
Value of Information-Enhanced Exploration in Bootstrapped DQN
by: Plataniotis, Stergios, et al.
Published: (2025)
by: Plataniotis, Stergios, et al.
Published: (2025)
$ε$-Optimally Solving Zero-Sum POSGs
by: Escudie, Erwan, et al.
Published: (2024)
by: Escudie, Erwan, et al.
Published: (2024)
Video-Driven Graph Network-Based Simulators
by: Szewczyk, Franciszek, et al.
Published: (2024)
by: Szewczyk, Franciszek, et al.
Published: (2024)
Accelerating Reinforcement Learning with Value-Conditional State Entropy Exploration
by: Kim, Dongyoung, et al.
Published: (2023)
by: Kim, Dongyoung, et al.
Published: (2023)
Foundation Models as World Models: A Foundational Study in Text-Based GridWorlds
by: Sasso, Remo, et al.
Published: (2025)
by: Sasso, Remo, et al.
Published: (2025)
On the Generalisation of Koopman Representations for Chaotic System Control
by: Hjikakou, Kyriakos, et al.
Published: (2025)
by: Hjikakou, Kyriakos, et al.
Published: (2025)
Upside-Down Reinforcement Learning for More Interpretable Optimal Control
by: Cardenas-Cartagena, Juan, et al.
Published: (2024)
by: Cardenas-Cartagena, Juan, et al.
Published: (2024)
HyperSHAP: Shapley Values and Interactions for Explaining Hyperparameter Optimization
by: Wever, Marcel, et al.
Published: (2025)
by: Wever, Marcel, et al.
Published: (2025)
Adaptive Exploration for Data-Efficient General Value Function Evaluations
by: Jain, Arushi, et al.
Published: (2024)
by: Jain, Arushi, et al.
Published: (2024)
Value Bonuses using Ensemble Errors for Exploration in Reinforcement Learning
by: Wahab, Abdul, et al.
Published: (2026)
by: Wahab, Abdul, et al.
Published: (2026)
Online Drift Detection with Maximum Concept Discrepancy
by: Wan, Ke, et al.
Published: (2024)
by: Wan, Ke, et al.
Published: (2024)
Bias Detection via Maximum Subgroup Discrepancy
by: Němeček, Jiří, et al.
Published: (2025)
by: Němeček, Jiří, et al.
Published: (2025)
Discrepancy-Aware Graph Mask Auto-Encoder
by: Zheng, Ziyu, et al.
Published: (2025)
by: Zheng, Ziyu, et al.
Published: (2025)
Understanding Prediction Discrepancies in Machine Learning Classifiers
by: Renard, Xavier, et al.
Published: (2021)
by: Renard, Xavier, et al.
Published: (2021)
Neighboring State-based Exploration for Reinforcement Learning
by: Li, Yu-Teng, et al.
Published: (2022)
by: Li, Yu-Teng, et al.
Published: (2022)
Federated Unlearning in the Wild: Rethinking Fairness and Data Discrepancy
by: Huang, ZiHeng, et al.
Published: (2025)
by: Huang, ZiHeng, et al.
Published: (2025)
Towards Foundation Models for Zero-Shot Time Series Anomaly Detection: Leveraging Synthetic Data and Relative Context Discrepancy
by: Lan, Tian, et al.
Published: (2025)
by: Lan, Tian, et al.
Published: (2025)
An Analysis of Action-Value Temporal-Difference Methods That Learn State Values
by: Daley, Brett, et al.
Published: (2025)
by: Daley, Brett, et al.
Published: (2025)
VSFormer: Value and Shape-Aware Transformer with Prior-Enhanced Self-Attention for Multivariate Time Series Classification
by: Xi, Wenjie, et al.
Published: (2024)
by: Xi, Wenjie, et al.
Published: (2024)
Enhancing Cooperative Multi-Agent Reinforcement Learning with State Modelling and Adversarial Exploration
by: Kontogiannis, Andreas, et al.
Published: (2025)
by: Kontogiannis, Andreas, et al.
Published: (2025)
Every Rollout Counts: Optimal Resource Allocation for Efficient Test-Time Scaling
by: Wang, Xinglin, et al.
Published: (2025)
by: Wang, Xinglin, et al.
Published: (2025)
Mitigating Partial Observability in Sequential Decision Processes via the Lambda Discrepancy
by: Allen, Cameron, et al.
Published: (2024)
by: Allen, Cameron, et al.
Published: (2024)
Enhancing Model Interpretability with Local Attribution over Global Exploration
by: Zhu, Zhiyu, et al.
Published: (2024)
by: Zhu, Zhiyu, et al.
Published: (2024)
Enhance Exploration in Safe Reinforcement Learning with Contrastive Representation Learning
by: Doan, Duc Kien, et al.
Published: (2025)
by: Doan, Duc Kien, et al.
Published: (2025)
Active Inference with Reusable State-Dependent Value Profiles
by: Poschl, Jacob
Published: (2025)
by: Poschl, Jacob
Published: (2025)
Negative-Binomial Randomized Gamma Markov Processes for Heterogeneous Overdispersed Count Time Series
by: Huang, Rui, et al.
Published: (2024)
by: Huang, Rui, et al.
Published: (2024)
Incorporating Domain Differential Equations into Graph Convolutional Networks to Lower Generalization Discrepancy
by: Sun, Yue, et al.
Published: (2024)
by: Sun, Yue, et al.
Published: (2024)
The Structural Origin of Attention Sink: Variance Discrepancy, Super Neurons, and Dimension Disparity
by: Li, Siquan, et al.
Published: (2026)
by: Li, Siquan, et al.
Published: (2026)
Revisiting Multivariate Time Series Forecasting with Missing Values
by: Yang, Jie, et al.
Published: (2025)
by: Yang, Jie, et al.
Published: (2025)
The Road Less Traveled: Enhancing Exploration in LLMs via Sequential Sampling
by: Kang, Shijia, et al.
Published: (2025)
by: Kang, Shijia, et al.
Published: (2025)
Improving Offline-to-Online Reinforcement Learning with Q Conditioned State Entropy Exploration
by: Zhang, Ziqi, et al.
Published: (2023)
by: Zhang, Ziqi, et al.
Published: (2023)
Representation-Based Exploration for Language Models: From Test-Time to Post-Training
by: Tuyls, Jens, et al.
Published: (2025)
by: Tuyls, Jens, et al.
Published: (2025)
Worst-Case Regret Bounds for Exploration via Randomized Value Functions
by: Russo, Daniel
Published: (2019)
by: Russo, Daniel
Published: (2019)
Similar Items
-
Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning
by: Falzari, Massimiliano, et al.
Published: (2025) -
On The Presence of Double-Descent in Deep Reinforcement Learning
by: Veselý, Viktor, et al.
Published: (2025) -
Sparsity-Driven Plasticity in Multi-Task Reinforcement Learning
by: Todorov, Aleksandar, et al.
Published: (2025) -
Mapping Transformer Leveraged Embeddings for Cross-Lingual Document Representation
by: Tashu, Tsegaye Misikir, et al.
Published: (2024) -
Exploration with Foundation Models: Capabilities, Limitations, and Hybrid Approaches
by: Sasso, Remo, et al.
Published: (2025)