Saved in:
| Main Authors: | Lee, Ken Ming, Barde, Paul, Cohen, Maxime C., Nowrouzezahrai, Derek |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.18449 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
by: Barde, Paul, et al.
Published: (2023)
by: Barde, Paul, et al.
Published: (2023)
Reinforcement Learning for Sequence Design Leveraging Protein Language Models
by: Subramanian, Jithendaraa, et al.
Published: (2024)
by: Subramanian, Jithendaraa, et al.
Published: (2024)
Simulation-Based Benchmarking of Reinforcement Learning Agents for Personalized Retail Promotions
by: Xia, Yu, et al.
Published: (2024)
by: Xia, Yu, et al.
Published: (2024)
GTA: Generative Trajectory Augmentation with Guidance for Offline Reinforcement Learning
by: Lee, Jaewoo, et al.
Published: (2024)
by: Lee, Jaewoo, et al.
Published: (2024)
TREX: Trajectory Explanations for Multi-Objective Reinforcement Learning
by: Rajapakse, Dilina, et al.
Published: (2026)
by: Rajapakse, Dilina, et al.
Published: (2026)
Toward Theoretical Insights into Diffusion Trajectory Distillation via Operator Merging
by: Gao, Weiguo, et al.
Published: (2025)
by: Gao, Weiguo, et al.
Published: (2025)
An Exploration of Clustering Algorithms for Customer Segmentation in the UK Retail Market
by: John, Jeen Mary, et al.
Published: (2024)
by: John, Jeen Mary, et al.
Published: (2024)
Trajectory Modeling via Random Utility Inverse Reinforcement Learning
by: Pitombeira-Neto, Anselmo R., et al.
Published: (2021)
by: Pitombeira-Neto, Anselmo R., et al.
Published: (2021)
Pivoting Retail Supply Chain with Deep Generative Techniques: Taxonomy, Survey and Insights
by: Wang, Yuan, et al.
Published: (2024)
by: Wang, Yuan, et al.
Published: (2024)
Comparative Analysis of Modern Machine Learning Models for Retail Sales Forecasting
by: Hobor, Luka, et al.
Published: (2025)
by: Hobor, Luka, et al.
Published: (2025)
Offline Reinforcement Learning with Generative Trajectory Policies
by: Feng, Xinsong, et al.
Published: (2025)
by: Feng, Xinsong, et al.
Published: (2025)
Offline Trajectory Optimization for Offline Reinforcement Learning
by: Zhao, Ziqi, et al.
Published: (2024)
by: Zhao, Ziqi, et al.
Published: (2024)
Single-Trajectory Distributionally Robust Reinforcement Learning
by: Liang, Zhipeng, et al.
Published: (2023)
by: Liang, Zhipeng, et al.
Published: (2023)
Learning to Construct Practical Agentic Systems
by: Kumar, Aditya, et al.
Published: (2026)
by: Kumar, Aditya, et al.
Published: (2026)
Policy-Based Trajectory Clustering in Offline Reinforcement Learning
by: Hu, Hao, et al.
Published: (2025)
by: Hu, Hao, et al.
Published: (2025)
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning
by: Sestini, Alessandro, et al.
Published: (2025)
by: Sestini, Alessandro, et al.
Published: (2025)
Offline Safe Reinforcement Learning Using Trajectory Classification
by: Gong, Ze, et al.
Published: (2024)
by: Gong, Ze, et al.
Published: (2024)
Consistency Trajectory Planning: High-Quality and Efficient Trajectory Optimization for Offline Model-Based Reinforcement Learning
by: Wang, Guanquan, et al.
Published: (2025)
by: Wang, Guanquan, et al.
Published: (2025)
Flow-based Policy With Distributional Reinforcement Learning in Trajectory Optimization
by: Hao, Ruijie, et al.
Published: (2026)
by: Hao, Ruijie, et al.
Published: (2026)
Offline Reinforcement Learning: Role of State Aggregation and Trajectory Data
by: Jia, Zeyu, et al.
Published: (2024)
by: Jia, Zeyu, et al.
Published: (2024)
A Practical Introduction to Deep Reinforcement Learning
by: Sun, Yinghan, et al.
Published: (2025)
by: Sun, Yinghan, et al.
Published: (2025)
Constraint-Guided Prediction Refinement via Deterministic Diffusion Trajectories
by: Dogoulis, Pantelis, et al.
Published: (2025)
by: Dogoulis, Pantelis, et al.
Published: (2025)
Offline Reinforcement Learning with Universal Horizon Models
by: Chung, Hojun, et al.
Published: (2026)
by: Chung, Hojun, et al.
Published: (2026)
Model-based Offline Reinforcement Learning with Lower Expectile Q-Learning
by: Park, Kwanyoung, et al.
Published: (2024)
by: Park, Kwanyoung, et al.
Published: (2024)
In-Dataset Trajectory Return Regularization for Offline Preference-based Reinforcement Learning
by: Tu, Songjun, et al.
Published: (2024)
by: Tu, Songjun, et al.
Published: (2024)
How Should We Meta-Learn Reinforcement Learning Algorithms?
by: Goldie, Alexander David, et al.
Published: (2025)
by: Goldie, Alexander David, et al.
Published: (2025)
Supervised Reinforcement Learning: From Expert Trajectories to Step-wise Reasoning
by: Deng, Yihe, et al.
Published: (2025)
by: Deng, Yihe, et al.
Published: (2025)
Graph-Attentive MAPPO for Dynamic Retail Pricing
by: Amma, Krishna Kumar Neelakanta Pillai Santha Kumari
Published: (2025)
by: Amma, Krishna Kumar Neelakanta Pillai Santha Kumari
Published: (2025)
Practical Insights into Knowledge Distillation for Pre-Trained Models
by: Alballa, Norah, et al.
Published: (2024)
by: Alballa, Norah, et al.
Published: (2024)
DiffStitch: Boosting Offline Reinforcement Learning with Diffusion-based Trajectory Stitching
by: Li, Guanghe, et al.
Published: (2024)
by: Li, Guanghe, et al.
Published: (2024)
Prioritized Trajectory Replay: A Replay Memory for Data-driven Reinforcement Learning
by: Liu, Jinyi, et al.
Published: (2023)
by: Liu, Jinyi, et al.
Published: (2023)
Generalizable Trajectory Prediction via Inverse Reinforcement Learning with Mamba-Graph Architecture
by: Li, Wenyun, et al.
Published: (2025)
by: Li, Wenyun, et al.
Published: (2025)
Meta-Statistical Learning: Supervised Learning of Statistical Estimators
by: Peyrard, Maxime, et al.
Published: (2025)
by: Peyrard, Maxime, et al.
Published: (2025)
Customized Load Profiles Synthesis for Electricity Customers Based on Conditional Diffusion Models
by: Wang, Zhenyi, et al.
Published: (2023)
by: Wang, Zhenyi, et al.
Published: (2023)
SPQR: Controlling Q-ensemble Independence with Spiked Random Model for Reinforcement Learning
by: Lee, Dohyeok, et al.
Published: (2024)
by: Lee, Dohyeok, et al.
Published: (2024)
Distilling Reinforcement Learning Algorithms for In-Context Model-Based Planning
by: Son, Jaehyeon, et al.
Published: (2025)
by: Son, Jaehyeon, et al.
Published: (2025)
DARTS: Distribution-Aware Active Rollout Trajectory Shaping for Accelerating LLM Reinforcement Learning
by: Wang, Yujie, et al.
Published: (2026)
by: Wang, Yujie, et al.
Published: (2026)
PathletRL++: Optimizing Trajectory Pathlet Extraction and Dictionary Formation via Reinforcement Learning
by: Alix, Gian, et al.
Published: (2024)
by: Alix, Gian, et al.
Published: (2024)
ARCLE: The Abstraction and Reasoning Corpus Learning Environment for Reinforcement Learning
by: Lee, Hosung, et al.
Published: (2024)
by: Lee, Hosung, et al.
Published: (2024)
Reward Generation via Large Vision-Language Model in Offline Reinforcement Learning
by: Lee, Younghwan, et al.
Published: (2025)
by: Lee, Younghwan, et al.
Published: (2025)
Similar Items
-
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
by: Barde, Paul, et al.
Published: (2023) -
Reinforcement Learning for Sequence Design Leveraging Protein Language Models
by: Subramanian, Jithendaraa, et al.
Published: (2024) -
Simulation-Based Benchmarking of Reinforcement Learning Agents for Personalized Retail Promotions
by: Xia, Yu, et al.
Published: (2024) -
GTA: Generative Trajectory Augmentation with Guidance for Offline Reinforcement Learning
by: Lee, Jaewoo, et al.
Published: (2024) -
TREX: Trajectory Explanations for Multi-Objective Reinforcement Learning
by: Rajapakse, Dilina, et al.
Published: (2026)