Unraveling the Rainbow: can value-based methods schedule?
Fuente:
arXiv
Saved in:
| Main Authors: | Corrêa, Arthur, Jesus, Alexandre, Nascimento, Paulo, Silva, Cristóvão, Moniz, Samuel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FiLMMeD: Feature-wise Linear Modulation for Cross-Problem Multi-Depot Vehicle Routing
by: Corrêa, Arthur, et al.
Published: (2026)
by: Corrêa, Arthur, et al.
Published: (2026)
TuneNSearch: a hybrid transfer learning and local search approach for solving vehicle routing problems
by: Corrêa, Arthur, et al.
Published: (2025)
by: Corrêa, Arthur, et al.
Published: (2025)
Evaluating Prompt Injection Defenses for Educational LLM Tutors: Security-Usability-Latency Trade-offs
by: Maiorano, Alexandre Cristovão
Published: (2026)
by: Maiorano, Alexandre Cristovão
Published: (2026)
Relevance-aware Algorithmic Recourse
by: Kim, Dongwhi, et al.
Published: (2024)
by: Kim, Dongwhi, et al.
Published: (2024)
Conformalized Selective Regression
by: Sokol, Anna, et al.
Published: (2024)
by: Sokol, Anna, et al.
Published: (2024)
An analysis of the noise schedule for score-based generative models
by: Strasman, Stanislas, et al.
Published: (2024)
by: Strasman, Stanislas, et al.
Published: (2024)
Reinforcement learning-based dynamic cleaning scheduling framework for solar energy system
by: An, Heungjo
Published: (2026)
by: An, Heungjo
Published: (2026)
Intersectional Divergence: Measuring Fairness in Regression
by: Germino, Joe, et al.
Published: (2025)
by: Germino, Joe, et al.
Published: (2025)
AnyLoss: Transforming Classification Metrics into Loss Functions
by: Han, Doheon, et al.
Published: (2024)
by: Han, Doheon, et al.
Published: (2024)
Optimistic critics can empower small actors
by: Mastikhina, Olya, et al.
Published: (2025)
by: Mastikhina, Olya, et al.
Published: (2025)
Rethinking Neural-based Matrix Inversion: Why can't, and Where can
by: Ji, Yuliang, et al.
Published: (2025)
by: Ji, Yuliang, et al.
Published: (2025)
Fast Explanations via Policy Gradient-Optimized Explainer
by: Pan, Deng, et al.
Published: (2024)
by: Pan, Deng, et al.
Published: (2024)
Automated Privacy-Preserving Techniques via Meta-Learning
by: Carvalho, Tânia, et al.
Published: (2024)
by: Carvalho, Tânia, et al.
Published: (2024)
Adaptive scheduling for adaptive sampling in POS taggers construction
by: Ferro, Manuel Vilares, et al.
Published: (2024)
by: Ferro, Manuel Vilares, et al.
Published: (2024)
Perturbative methods for non-parametric instrumental variable
by: Bu, Wei, et al.
Published: (2026)
by: Bu, Wei, et al.
Published: (2026)
Time Series Data Augmentation as an Imbalanced Learning Problem
by: Cerqueira, Vitor, et al.
Published: (2024)
by: Cerqueira, Vitor, et al.
Published: (2024)
Additive regularization schedule for neural architecture search
by: Potanin, Mark, et al.
Published: (2024)
by: Potanin, Mark, et al.
Published: (2024)
Lazy vs hasty: linearization in deep networks impacts learning schedule based on example difficulty
by: George, Thomas, et al.
Published: (2022)
by: George, Thomas, et al.
Published: (2022)
Hamiltonian Mechanics of Feature Learning: Bottleneck Structure in Leaky ResNets
by: Jacot, Arthur, et al.
Published: (2024)
by: Jacot, Arthur, et al.
Published: (2024)
Decision-focused learning for optimal PV-Battery scheduling
by: Depoortere, Joris, et al.
Published: (2026)
by: Depoortere, Joris, et al.
Published: (2026)
Version age-based client scheduling policy for federated learning
by: Hu, Xinyi, et al.
Published: (2024)
by: Hu, Xinyi, et al.
Published: (2024)
Disentangled Deep Smoothed Bootstrap for Fair Imbalanced Regression
by: Stocksieker, Samuel, et al.
Published: (2025)
by: Stocksieker, Samuel, et al.
Published: (2025)
Data Augmentation with Variational Autoencoder for Imbalanced Dataset
by: Stocksieker, Samuel, et al.
Published: (2024)
by: Stocksieker, Samuel, et al.
Published: (2024)
Boarding for ISS: Imbalanced Self-Supervised: Discovery of a Scaled Autoencoder for Mixed Tabular Datasets
by: Stocksieker, Samuel, et al.
Published: (2024)
by: Stocksieker, Samuel, et al.
Published: (2024)
Experiential-Informed Data Reconstruction for Fishery Sustainability and Policies in the Azores
by: Nogueira, Brenda, et al.
Published: (2023)
by: Nogueira, Brenda, et al.
Published: (2023)
Evolutionary scheduling of university activities based on consumption forecasts to minimise electricity costs
by: Ruddick, Julian, et al.
Published: (2022)
by: Ruddick, Julian, et al.
Published: (2022)
SPOT: Single-Shot Positioning via Trainable Near-Field Rainbow Beamforming
by: Cai, Yeyue, et al.
Published: (2025)
by: Cai, Yeyue, et al.
Published: (2025)
In value-based deep reinforcement learning, a pruned network is a good network
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Synthetic Data Outliers: Navigating Identity Disclosure
by: Trindade, Carolina, et al.
Published: (2024)
by: Trindade, Carolina, et al.
Published: (2024)
Differentially-Private Data Synthetisation for Efficient Re-Identification Risk Control
by: Carvalho, Tânia, et al.
Published: (2022)
by: Carvalho, Tânia, et al.
Published: (2022)
Sharing Knowledge without Sharing Data: Stitches can improve ensembles of disjointly trained models
by: Guijt, Arthur, et al.
Published: (2025)
by: Guijt, Arthur, et al.
Published: (2025)
What do near-optimal learning rate schedules look like?
by: Naganuma, Hiroki, et al.
Published: (2026)
by: Naganuma, Hiroki, et al.
Published: (2026)
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
LLMs can learn self-restraint through iterative self-reflection
by: Piché, Alexandre, et al.
Published: (2024)
by: Piché, Alexandre, et al.
Published: (2024)
Influence-based Attributions can be Manipulated
by: Yadav, Chhavi, et al.
Published: (2024)
by: Yadav, Chhavi, et al.
Published: (2024)
Rainbow-DemoRL: Combining Improvements in Demonstration-Augmented Reinforcement Learning
by: Bhatt, Dwait, et al.
Published: (2026)
by: Bhatt, Dwait, et al.
Published: (2026)
Post-Norm can Resharpen Attention
by: Zsámboki, Pál, et al.
Published: (2025)
by: Zsámboki, Pál, et al.
Published: (2025)
Handling missing values in clinical machine learning: Insights from an expert study
by: Stempfle, Lena, et al.
Published: (2024)
by: Stempfle, Lena, et al.
Published: (2024)
Stepsize anything: A unified learning rate schedule for budgeted-iteration training
by: Tang, Anda, et al.
Published: (2025)
by: Tang, Anda, et al.
Published: (2025)
HyperbolicLR: Epoch insensitive learning rate scheduler
by: Kim, Tae-Geun
Published: (2024)
by: Kim, Tae-Geun
Published: (2024)
Similar Items
-
FiLMMeD: Feature-wise Linear Modulation for Cross-Problem Multi-Depot Vehicle Routing
by: Corrêa, Arthur, et al.
Published: (2026) -
TuneNSearch: a hybrid transfer learning and local search approach for solving vehicle routing problems
by: Corrêa, Arthur, et al.
Published: (2025) -
Evaluating Prompt Injection Defenses for Educational LLM Tutors: Security-Usability-Latency Trade-offs
by: Maiorano, Alexandre Cristovão
Published: (2026) -
Relevance-aware Algorithmic Recourse
by: Kim, Dongwhi, et al.
Published: (2024) -
Conformalized Selective Regression
by: Sokol, Anna, et al.
Published: (2024)