Saved in:
| Main Authors: | Winkel, David, Strauß, Niklas, Bernhard, Maximilian, Li, Zongyue, Seidl, Thomas, Schubert, Matthias |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2409.18735 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Simplex Decomposition for Portfolio Allocation Constraints in Reinforcement Learning
by: Winkel, David, et al.
Published: (2024)
by: Winkel, David, et al.
Published: (2024)
DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions
by: Li, Zongyue, et al.
Published: (2025)
by: Li, Zongyue, et al.
Published: (2025)
Dying Clusters Is All You Need -- Deep Clustering With an Unknown Number of Clusters
by: Leiber, Collin, et al.
Published: (2024)
by: Leiber, Collin, et al.
Published: (2024)
Spatial-Aware Deep Reinforcement Learning for the Traveling Officer Problem
by: Strauß, Niklas, et al.
Published: (2024)
by: Strauß, Niklas, et al.
Published: (2024)
Efficient Task Transfer for HLS DSE
by: Ding, Zijian, et al.
Published: (2024)
by: Ding, Zijian, et al.
Published: (2024)
Memory Allocation in Resource-Constrained Reinforcement Learning
by: Tamborski, Massimiliano, et al.
Published: (2025)
by: Tamborski, Massimiliano, et al.
Published: (2025)
Optimizing Variational Quantum Circuits Using Metaheuristic Strategies in Reinforcement Learning
by: Kölle, Michael, et al.
Published: (2024)
by: Kölle, Michael, et al.
Published: (2024)
What-If Explanations Over Time: Counterfactuals for Time Series Classification
by: Schlegel, Udo, et al.
Published: (2026)
by: Schlegel, Udo, et al.
Published: (2026)
How to Allocate, How to Learn? Dynamic Rollout Allocation and Advantage Modulation for Policy Optimization
by: Fang, Yangyi, et al.
Published: (2026)
by: Fang, Yangyi, et al.
Published: (2026)
STree: Speculative Tree Decoding for Hybrid State-Space Models
by: Wu, Yangchao, et al.
Published: (2025)
by: Wu, Yangchao, et al.
Published: (2025)
Constrained Latent Action Policies for Model-Based Offline Reinforcement Learning
by: Alles, Marvin, et al.
Published: (2024)
by: Alles, Marvin, et al.
Published: (2024)
Conformal Constrained Policy Optimization for Cost-Effective LLM Agents
by: Si, Wenwen, et al.
Published: (2025)
by: Si, Wenwen, et al.
Published: (2025)
Detector-Evasive LLM Paraphrasing via Constrained Policy Optimization
by: Wang, Mingyi, et al.
Published: (2026)
by: Wang, Mingyi, et al.
Published: (2026)
Incentivizing Safer Actions in Policy Optimization for Constrained Reinforcement Learning
by: Hazra, Somnath, et al.
Published: (2025)
by: Hazra, Somnath, et al.
Published: (2025)
Federated Few-Shot Learning for Epileptic Seizure Detection Under Privacy Constraints
by: Sysoykova, Ekaterina, et al.
Published: (2025)
by: Sysoykova, Ekaterina, et al.
Published: (2025)
Entropy-Gated Selective Policy Optimization:Token-Level Gradient Allocation for Hybrid Training of Large Language Models
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
AROpt: An Optimization Method for Autoregressive Time Series Forecasting
by: Li, Zheng, et al.
Published: (2026)
by: Li, Zheng, et al.
Published: (2026)
Heuristic Methods are Good Teachers to Distill MLPs for Graph Link Prediction
by: Qin, Zongyue, et al.
Published: (2025)
by: Qin, Zongyue, et al.
Published: (2025)
Hierarchy-of-Groups Policy Optimization for Long-Horizon Agentic Tasks
by: He, Shuo, et al.
Published: (2026)
by: He, Shuo, et al.
Published: (2026)
When are LLMs Sufficient Policy Optimizers for Sequential RL Tasks?
by: Hatgis-Kessell, Stephane, et al.
Published: (2026)
by: Hatgis-Kessell, Stephane, et al.
Published: (2026)
Pruning the Way to Reliable Policies: A Multi-Objective Deep Q-Learning Approach to Critical Care
by: Shirali, Ali, et al.
Published: (2023)
by: Shirali, Ali, et al.
Published: (2023)
Stepwise Alignment for Constrained Language Model Policy Optimization
by: Wachi, Akifumi, et al.
Published: (2024)
by: Wachi, Akifumi, et al.
Published: (2024)
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
by: Biswas, Arpita, et al.
Published: (2023)
by: Biswas, Arpita, et al.
Published: (2023)
Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models
by: Peysakhovich, Alexander, et al.
Published: (2026)
by: Peysakhovich, Alexander, et al.
Published: (2026)
When Errors Can Be Beneficial: A Categorization of Imperfect Rewards for Policy Gradient
by: Shang, Shuning, et al.
Published: (2026)
by: Shang, Shuning, et al.
Published: (2026)
Handling Long and Richly Constrained Tasks through Constrained Hierarchical Reinforcement Learning
by: Lu, Yuxiao, et al.
Published: (2023)
by: Lu, Yuxiao, et al.
Published: (2023)
Latent Bayesian Optimization via Autoregressive Normalizing Flows
by: Lee, Seunghun, et al.
Published: (2025)
by: Lee, Seunghun, et al.
Published: (2025)
A Two-stage Reinforcement Learning-based Approach for Multi-entity Task Allocation
by: Gong, Aicheng, et al.
Published: (2024)
by: Gong, Aicheng, et al.
Published: (2024)
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance
by: He, Jinmin, et al.
Published: (2025)
by: He, Jinmin, et al.
Published: (2025)
Wasserstein Policy Optimization
by: Pfau, David, et al.
Published: (2025)
by: Pfau, David, et al.
Published: (2025)
Efficient Parking Search using Shared Fleet Data
by: Strauß, Niklas, et al.
Published: (2024)
by: Strauß, Niklas, et al.
Published: (2024)
Network-Constrained Policy Optimization for Adaptive Multi-agent Vehicle Routing
by: Arasteh, Fazel, et al.
Published: (2025)
by: Arasteh, Fazel, et al.
Published: (2025)
Preference-Based Gradient Estimation for ML-Guided Approximate Combinatorial Optimization
by: Mielke, Arman, et al.
Published: (2025)
by: Mielke, Arman, et al.
Published: (2025)
On-Policy Optimization of ANFIS Policies Using Proximal Policy Optimization
by: Shankar, Kaaustaaub, et al.
Published: (2025)
by: Shankar, Kaaustaaub, et al.
Published: (2025)
Optimizing Urban Service Allocation with Time-Constrained Restless Bandits
by: Mao, Yi, et al.
Published: (2025)
by: Mao, Yi, et al.
Published: (2025)
Hierarchical Mixture of Experts: Generalizable Learning for High-Level Synthesis
by: Li, Weikai, et al.
Published: (2024)
by: Li, Weikai, et al.
Published: (2024)
Balancing Immediate Revenue and Future Off-Policy Evaluation in Coupon Allocation
by: Nishimura, Naoki, et al.
Published: (2024)
by: Nishimura, Naoki, et al.
Published: (2024)
Model Ensembling for Constrained Optimization
by: Globus-Harris, Ira, et al.
Published: (2024)
by: Globus-Harris, Ira, et al.
Published: (2024)
Accelerating Proximal Policy Optimization Learning Using Task Prediction for Solving Environments with Delayed Rewards
by: Ahmad, Ahmad, et al.
Published: (2024)
by: Ahmad, Ahmad, et al.
Published: (2024)
Variance-Aware Prior-Based Tree Policies for Monte Carlo Tree Search
by: Weichart, Maximilian
Published: (2025)
by: Weichart, Maximilian
Published: (2025)
Similar Items
-
Simplex Decomposition for Portfolio Allocation Constraints in Reinforcement Learning
by: Winkel, David, et al.
Published: (2024) -
DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions
by: Li, Zongyue, et al.
Published: (2025) -
Dying Clusters Is All You Need -- Deep Clustering With an Unknown Number of Clusters
by: Leiber, Collin, et al.
Published: (2024) -
Spatial-Aware Deep Reinforcement Learning for the Traveling Officer Problem
by: Strauß, Niklas, et al.
Published: (2024) -
Efficient Task Transfer for HLS DSE
by: Ding, Zijian, et al.
Published: (2024)