Guardado en:
| Autores principales: | Tse, Hon Tik, Machado, Marlos C. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.00403 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reward-Aware Proto-Representations in Reinforcement Learning
por: Tse, Hon Tik, et al.
Publicado: (2025)
por: Tse, Hon Tik, et al.
Publicado: (2025)
Harnessing Discrete Representations For Continual Reinforcement Learning
por: Meyer, Edan, et al.
Publicado: (2023)
por: Meyer, Edan, et al.
Publicado: (2023)
Proper Laplacian Representation Learning
por: Gomez, Diego, et al.
Publicado: (2023)
por: Gomez, Diego, et al.
Publicado: (2023)
Averaging $n$-step Returns Reduces Variance in Reinforcement Learning
por: Daley, Brett, et al.
Publicado: (2024)
por: Daley, Brett, et al.
Publicado: (2024)
Trajectory-Aware Eligibility Traces for Off-Policy Reinforcement Learning
por: Daley, Brett, et al.
Publicado: (2023)
por: Daley, Brett, et al.
Publicado: (2023)
AGaLiTe: Approximate Gated Linear Transformers for Online Reinforcement Learning
por: Pramanik, Subhojeet, et al.
Publicado: (2023)
por: Pramanik, Subhojeet, et al.
Publicado: (2023)
A Study of Value-Aware Eigenoptions
por: Kotamreddy, Harshil, et al.
Publicado: (2025)
por: Kotamreddy, Harshil, et al.
Publicado: (2025)
Plastic Learning with Deep Fourier Features
por: Lewandowski, Alex, et al.
Publicado: (2024)
por: Lewandowski, Alex, et al.
Publicado: (2024)
Laplacian Representations for Decision-Time Planning
por: Shehmar, Dikshant, et al.
Publicado: (2026)
por: Shehmar, Dikshant, et al.
Publicado: (2026)
The Laplacian Keyboard: Beyond the Linear Span
por: Chandrasekar, Siddarth, et al.
Publicado: (2026)
por: Chandrasekar, Siddarth, et al.
Publicado: (2026)
Demystifying the Recency Heuristic in Temporal-Difference Learning
por: Daley, Brett, et al.
Publicado: (2024)
por: Daley, Brett, et al.
Publicado: (2024)
Deep Reinforcement Learning with Gradient Eligibility Traces
por: Elelimy, Esraa, et al.
Publicado: (2025)
por: Elelimy, Esraa, et al.
Publicado: (2025)
Deep Double Q-learning
por: Nagarajan, Prabhat, et al.
Publicado: (2025)
por: Nagarajan, Prabhat, et al.
Publicado: (2025)
An Analysis of Action-Value Temporal-Difference Methods That Learn State Values
por: Daley, Brett, et al.
Publicado: (2025)
por: Daley, Brett, et al.
Publicado: (2025)
The Cell Must Go On: Agar.io for Continual Reinforcement Learning
por: Mohamed, Mohamed A., et al.
Publicado: (2025)
por: Mohamed, Mohamed A., et al.
Publicado: (2025)
Protein Structure Prediction in the 3D HP Model Using Deep Reinforcement Learning
por: Espitia, Giovanny, et al.
Publicado: (2024)
por: Espitia, Giovanny, et al.
Publicado: (2024)
Directions of Curvature as an Explanation for Loss of Plasticity
por: Lewandowski, Alex, et al.
Publicado: (2023)
por: Lewandowski, Alex, et al.
Publicado: (2023)
Learning Continually by Spectral Regularization
por: Lewandowski, Alex, et al.
Publicado: (2024)
por: Lewandowski, Alex, et al.
Publicado: (2024)
Safety Optimized Reinforcement Learning via Multi-Objective Policy Optimization
por: Honari, Homayoun, et al.
Publicado: (2024)
por: Honari, Homayoun, et al.
Publicado: (2024)
Financial Default Prediction via Motif-preserving Graph Neural Network with Curriculum Learning
por: Wang, Daixin, et al.
Publicado: (2024)
por: Wang, Daixin, et al.
Publicado: (2024)
Causal Coordinated Concurrent Reinforcement Learning
por: Tse, Tim, et al.
Publicado: (2024)
por: Tse, Tim, et al.
Publicado: (2024)
Default Machine Learning Hyperparameters Do Not Provide Informative Initialization for Bayesian Optimization
por: Prieto, Nicolás Villagrán, et al.
Publicado: (2026)
por: Prieto, Nicolás Villagrán, et al.
Publicado: (2026)
Trustworthy Representation Learning via Information Funnels and Bottlenecks
por: de Freitas, João Machado, et al.
Publicado: (2022)
por: de Freitas, João Machado, et al.
Publicado: (2022)
Efficient Graph Optimization via Distance-Aware Graph Representation Learning
por: Liu, Dong, et al.
Publicado: (2024)
por: Liu, Dong, et al.
Publicado: (2024)
Optimizing Language Models for Inference Time Objectives using Reinforcement Learning
por: Tang, Yunhao, et al.
Publicado: (2025)
por: Tang, Yunhao, et al.
Publicado: (2025)
Multi-Objective and Mixed-Reward Reinforcement Learning via Reward-Decorrelated Policy Optimization
por: Bai, Yang, et al.
Publicado: (2026)
por: Bai, Yang, et al.
Publicado: (2026)
Relational Graph Modeling for Credit Default Prediction: Heterogeneous GNNs and Hybrid Ensemble Learning
por: Yang, Yvonne, et al.
Publicado: (2026)
por: Yang, Yvonne, et al.
Publicado: (2026)
Polychromic Objectives for Reinforcement Learning
por: Hamid, Jubayer Ibn, et al.
Publicado: (2025)
por: Hamid, Jubayer Ibn, et al.
Publicado: (2025)
Combining Multi-Objective Bayesian Optimization with Reinforcement Learning for TinyML
por: Deutel, Mark, et al.
Publicado: (2023)
por: Deutel, Mark, et al.
Publicado: (2023)
Enhancing Chess Reinforcement Learning with Graph Representation
por: Rigaux, Tomas, et al.
Publicado: (2024)
por: Rigaux, Tomas, et al.
Publicado: (2024)
SLA-MORL: SLA-Aware Multi-Objective Reinforcement Learning for HPC Resource Optimization
por: Mostafa, Seraj Al Mahmud, et al.
Publicado: (2025)
por: Mostafa, Seraj Al Mahmud, et al.
Publicado: (2025)
Interpretable Credit Default Prediction with Ensemble Learning and SHAP
por: Yang, Shiqi, et al.
Publicado: (2025)
por: Yang, Shiqi, et al.
Publicado: (2025)
TREX: Trajectory Explanations for Multi-Objective Reinforcement Learning
por: Rajapakse, Dilina, et al.
Publicado: (2026)
por: Rajapakse, Dilina, et al.
Publicado: (2026)
Pareto Set Learning for Multi-Objective Reinforcement Learning
por: Liu, Erlong, et al.
Publicado: (2025)
por: Liu, Erlong, et al.
Publicado: (2025)
Preference-based Multi-Objective Reinforcement Learning
por: Mu, Ni, et al.
Publicado: (2025)
por: Mu, Ni, et al.
Publicado: (2025)
On The Expressivity of Objective-Specification Formalisms in Reinforcement Learning
por: Subramani, Rohan, et al.
Publicado: (2023)
por: Subramani, Rohan, et al.
Publicado: (2023)
Deep Multi-Objective Reinforcement Learning for Utility-Based Infrastructural Maintenance Optimization
por: van Remmerden, Jesse, et al.
Publicado: (2024)
por: van Remmerden, Jesse, et al.
Publicado: (2024)
Optimistic Reinforcement Learning with Quantile Objectives
por: Alipour-Vaezi, Mohammad, et al.
Publicado: (2025)
por: Alipour-Vaezi, Mohammad, et al.
Publicado: (2025)
Multi-Objective Optimization Using Adaptive Distributed Reinforcement Learning
por: Tan, Jing, et al.
Publicado: (2024)
por: Tan, Jing, et al.
Publicado: (2024)
Scalable Multi-Objective and Meta Reinforcement Learning via Gradient Estimation
por: Zhang, Zhenshuo, et al.
Publicado: (2025)
por: Zhang, Zhenshuo, et al.
Publicado: (2025)
Ejemplares similares
-
Reward-Aware Proto-Representations in Reinforcement Learning
por: Tse, Hon Tik, et al.
Publicado: (2025) -
Harnessing Discrete Representations For Continual Reinforcement Learning
por: Meyer, Edan, et al.
Publicado: (2023) -
Proper Laplacian Representation Learning
por: Gomez, Diego, et al.
Publicado: (2023) -
Averaging $n$-step Returns Reduces Variance in Reinforcement Learning
por: Daley, Brett, et al.
Publicado: (2024) -
Trajectory-Aware Eligibility Traces for Off-Policy Reinforcement Learning
por: Daley, Brett, et al.
Publicado: (2023)