Distributional Reinforcement Learning with Dual Expectile-Quantile Regression
Fuente:
arXiv
Saved in:
| Main Authors: | Jullien, Sami, Deffayet, Romain, Renders, Jean-Michel, Groth, Paul, de Rijke, Maarten |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Task Adaptation from Skills: Information Geometry, Disentanglement, and New Objectives for Unsupervised Reinforcement Learning
by: Yang, Yucheng, et al.
Published: (2025)
by: Yang, Yucheng, et al.
Published: (2025)
Improved Exploration in GFlownets via Enhanced Epistemic Neural Networks
by: Muhammad, Sajan, et al.
Published: (2025)
by: Muhammad, Sajan, et al.
Published: (2025)
Active Causal Experimentalist (ACE): Learning Intervention Strategies via Direct Preference Optimization
by: Cooper, Patrick, et al.
Published: (2026)
by: Cooper, Patrick, et al.
Published: (2026)
Contract2Plan: Verified Contract-Grounded Retrieval-Augmented Optimization for BOM-Aware Procurement and Multi-Echelon Inventory Planning
by: Agarwal, Sahil
Published: (2026)
by: Agarwal, Sahil
Published: (2026)
Prediction-Based Markov Violation Scores for Detecting Non-Markovian Observations in Reinforcement Learning
by: Mysore, Naveen
Published: (2026)
by: Mysore, Naveen
Published: (2026)
Procedural Game Level Design with Deep Reinforcement Learning
by: Özkan, Miraç Buğra
Published: (2025)
by: Özkan, Miraç Buğra
Published: (2025)
Logging Policy Design for Off-Policy Evaluation
by: Douglas, Connor, et al.
Published: (2026)
by: Douglas, Connor, et al.
Published: (2026)
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
by: Aichmüller, Michael, et al.
Published: (2024)
by: Aichmüller, Michael, et al.
Published: (2024)
Reinforcement Learning for Portfolio Optimization with a Financial Goal and Defined Time Horizons
by: Leukam, Fermat, et al.
Published: (2025)
by: Leukam, Fermat, et al.
Published: (2025)
Sim-to-reality adaptation for Deep Reinforcement Learning applied to an underwater docking application
by: Chaarani, Alaaeddine, et al.
Published: (2026)
by: Chaarani, Alaaeddine, et al.
Published: (2026)
RHiOTS: A Framework for Evaluating Hierarchical Time Series Forecasting Algorithms
by: Roque, Luis, et al.
Published: (2024)
by: Roque, Luis, et al.
Published: (2024)
Ensuring Both Positivity and Stability Using Sector-Bounded Nonlinearity for Systems with Neural Network Controllers
by: Hedesh, Hamidreza Montazeri, et al.
Published: (2024)
by: Hedesh, Hamidreza Montazeri, et al.
Published: (2024)
Lifted Forward Planning in Relational Factored Markov Decision Processes with Concurrent Actions
by: Marwitz, Florian Andreas, et al.
Published: (2025)
by: Marwitz, Florian Andreas, et al.
Published: (2025)
Statistical Guarantees for Lifelong Reinforcement Learning using PAC-Bayes Theory
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
A Parallel Hybrid Action Space Reinforcement Learning Model for Real-world Adaptive Traffic Signal Control
by: Wang, Yuxuan, et al.
Published: (2025)
by: Wang, Yuxuan, et al.
Published: (2025)
ToolForge: A Data Synthesis Pipeline for Multi-Hop Search without Real-World APIs
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
A Dual Perspective of Reinforcement Learning for Imposing Policy Constraints
by: De Cooman, Bram, et al.
Published: (2024)
by: De Cooman, Bram, et al.
Published: (2024)
HTN Plan Repair Algorithms Compared: Strengths and Weaknesses of Different Methods
by: Zaidins, Paul, et al.
Published: (2025)
by: Zaidins, Paul, et al.
Published: (2025)
Integrating Explanations in Learning LTL Specifications from Demonstrations
by: Gupta, Ashutosh, et al.
Published: (2024)
by: Gupta, Ashutosh, et al.
Published: (2024)
Parallel Strategies for Best-First Generalized Planning
by: Fernández-Alburquerque, Alejandro, et al.
Published: (2024)
by: Fernández-Alburquerque, Alejandro, et al.
Published: (2024)
On Divergence Measures for Training GFlowNets
by: da Silva, Tiago, et al.
Published: (2024)
by: da Silva, Tiago, et al.
Published: (2024)
Predicting Future Actions of Reinforcement Learning Agents
by: Chung, Stephen, et al.
Published: (2024)
by: Chung, Stephen, et al.
Published: (2024)
Two-phase Optimization of Binary Sequences with Low Peak Sidelobe Level Value
by: Bošković, Borko, et al.
Published: (2021)
by: Bošković, Borko, et al.
Published: (2021)
Safe Reinforcement Learning with Preference-based Constraint Inference
by: Li, Chenglin, et al.
Published: (2026)
by: Li, Chenglin, et al.
Published: (2026)
Automated Generation of MDPs Using Logic Programming and LLMs for Robotic Applications
by: Saccon, Enrico, et al.
Published: (2025)
by: Saccon, Enrico, et al.
Published: (2025)
Learning to Select Goals in Automated Planning with Deep-Q Learning
by: Núñez-Molina, Carlos, et al.
Published: (2024)
by: Núñez-Molina, Carlos, et al.
Published: (2024)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
by: Sauter, Andreas W. M., et al.
Published: (2024)
by: Sauter, Andreas W. M., et al.
Published: (2024)
On inferring cumulative constraints
by: Sidorov, Konstantin
Published: (2026)
by: Sidorov, Konstantin
Published: (2026)
Calculating Ultra-Strong and Extended Solutions for Nine Men's Morris, Morabaraba, and Lasker Morris
by: Gévay, Gábor E., et al.
Published: (2014)
by: Gévay, Gábor E., et al.
Published: (2014)
Project Patti: Why can You Solve Diabolical Puzzles on one Sudoku Website but not Easy Puzzles on another Sudoku Website?
by: Eisenkolb-Vaithyanathan, Arman
Published: (2025)
by: Eisenkolb-Vaithyanathan, Arman
Published: (2025)
Comprehensive Taxonomies of Nature- and Bio-inspired Optimization: Inspiration versus Algorithmic Behavior, Critical Analysis and Recommendations (from 2020 to 2024)
by: Molina, Daniel, et al.
Published: (2020)
by: Molina, Daniel, et al.
Published: (2020)
Optimization of Activity Batching Policies in Business Processes
by: López-Pintado, Orlenys, et al.
Published: (2025)
by: López-Pintado, Orlenys, et al.
Published: (2025)
Requirements for Recognition and Rapid Response to Unfamiliar Events Outside of Agent Design Scope
by: Wray, Robert E., et al.
Published: (2025)
by: Wray, Robert E., et al.
Published: (2025)
VIGIL: A Reflective Runtime for Self-Healing Agents
by: Cruz, Christopher
Published: (2025)
by: Cruz, Christopher
Published: (2025)
Relevance Score: A Landmark-Like Heuristic for Planning
by: Kim, Oliver, et al.
Published: (2024)
by: Kim, Oliver, et al.
Published: (2024)
Petri Net Relaxation for Infeasibility Explanation and Sequential Task Planning
by: Le, Nguyen Cong Nhat, et al.
Published: (2026)
by: Le, Nguyen Cong Nhat, et al.
Published: (2026)
How To Discover Short, Shorter, and the Shortest Proofs of Unsatisfiability: A Branch-and-Bound Approach for Resolution Proof Length Minimization
by: Sidorov, Konstantin, et al.
Published: (2024)
by: Sidorov, Konstantin, et al.
Published: (2024)
Towards a Unified Framework for Sequential Decision Making
by: Núñez-Molina, Carlos, et al.
Published: (2023)
by: Núñez-Molina, Carlos, et al.
Published: (2023)
Umbrella Reinforcement Learning -- computationally efficient tool for hard non-linear problems
by: Nuzhin, Egor E., et al.
Published: (2024)
by: Nuzhin, Egor E., et al.
Published: (2024)
GIRL: Generative Imagination Reinforcement Learning via Information-Theoretic Hallucination Control
by: Hiremath, Prakul Sunil
Published: (2026)
by: Hiremath, Prakul Sunil
Published: (2026)
Similar Items
-
Task Adaptation from Skills: Information Geometry, Disentanglement, and New Objectives for Unsupervised Reinforcement Learning
by: Yang, Yucheng, et al.
Published: (2025) -
Improved Exploration in GFlownets via Enhanced Epistemic Neural Networks
by: Muhammad, Sajan, et al.
Published: (2025) -
Active Causal Experimentalist (ACE): Learning Intervention Strategies via Direct Preference Optimization
by: Cooper, Patrick, et al.
Published: (2026) -
Contract2Plan: Verified Contract-Grounded Retrieval-Augmented Optimization for BOM-Aware Procurement and Multi-Echelon Inventory Planning
by: Agarwal, Sahil
Published: (2026) -
Prediction-Based Markov Violation Scores for Detecting Non-Markovian Observations in Reinforcement Learning
by: Mysore, Naveen
Published: (2026)