IQL-TD-MPC: Implicit Q-Learning for Hierarchical Model Predictive Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Chitnis, Rohan, Xu, Yingchen, Hashemi, Bobak, Lehnert, Lucas, Dogan, Urun, Zhu, Zheqing, Delalleau, Olivier |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AlignIQL: Policy Alignment in Implicit Q-Learning through Constrained Optimization
por: He, Longxiang, et al.
Publicado: (2024)
por: He, Longxiang, et al.
Publicado: (2024)
TD-MPC2: Scalable, Robust World Models for Continuous Control
por: Hansen, Nicklas, et al.
Publicado: (2023)
por: Hansen, Nicklas, et al.
Publicado: (2023)
DiSA-IQL: Offline Reinforcement Learning for Robust Soft Robot Control under Distribution Shifts
por: He, Linjin, et al.
Publicado: (2025)
por: He, Linjin, et al.
Publicado: (2025)
An Empirical Study of Deep Reinforcement Learning in Continuing Tasks
por: Wan, Yi, et al.
Publicado: (2025)
por: Wan, Yi, et al.
Publicado: (2025)
LeTac-MPC: Learning Model Predictive Control for Tactile-reactive Grasping
por: Xu, Zhengtong, et al.
Publicado: (2024)
por: Xu, Zhengtong, et al.
Publicado: (2024)
Pearl: A Production-ready Reinforcement Learning Agent
por: Zhu, Zheqing, et al.
Publicado: (2023)
por: Zhu, Zheqing, et al.
Publicado: (2023)
Investigating and Extending Homans' Social Exchange Theory with Large Language Model based Agents
por: Wang, Lei, et al.
Publicado: (2025)
por: Wang, Lei, et al.
Publicado: (2025)
Navigation with QPHIL: Quantizing Planner for Hierarchical Implicit Q-Learning
por: Canesse, Alexi, et al.
Publicado: (2024)
por: Canesse, Alexi, et al.
Publicado: (2024)
Smaller Abstract State Spaces Enable Cross-Scale Generalization in Reinforcement Learning
por: Mustakim, Nasehatul, et al.
Publicado: (2026)
por: Mustakim, Nasehatul, et al.
Publicado: (2026)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
por: Sikchi, Harshit, et al.
Publicado: (2023)
por: Sikchi, Harshit, et al.
Publicado: (2023)
When should we prefer Decision Transformers for Offline Reinforcement Learning?
por: Bhargava, Prajjwal, et al.
Publicado: (2023)
por: Bhargava, Prajjwal, et al.
Publicado: (2023)
Hierarchical Reinforcement Learning with Low-Level MPC for Multi-Agent Control
por: Studt, Max, et al.
Publicado: (2025)
por: Studt, Max, et al.
Publicado: (2025)
AC4MPC: Actor-Critic Reinforcement Learning for Nonlinear Model Predictive Control
por: Reiter, Rudolf, et al.
Publicado: (2024)
por: Reiter, Rudolf, et al.
Publicado: (2024)
Model Predictive Control-Guided Reinforcement Learning for Implicit Balancing
por: Madahi, Seyed Soroush Karimi, et al.
Publicado: (2025)
por: Madahi, Seyed Soroush Karimi, et al.
Publicado: (2025)
Rethinking the Design of Reinforcement Learning-Based Deep Research Agents
por: Wan, Yi, et al.
Publicado: (2025)
por: Wan, Yi, et al.
Publicado: (2025)
TMIQ: Quantifying Test and Measurement Domain Intelligence in Large Language Models
por: Olowe, Emmanuel A., et al.
Publicado: (2025)
por: Olowe, Emmanuel A., et al.
Publicado: (2025)
Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination
por: Spieler, Jonathan, et al.
Publicado: (2026)
por: Spieler, Jonathan, et al.
Publicado: (2026)
Value bounds and Convergence Analysis for Averages of LRP attributions
por: Binder, Alexander, et al.
Publicado: (2025)
por: Binder, Alexander, et al.
Publicado: (2025)
DeepSafeMPC: Deep Learning-Based Model Predictive Control for Safe Multi-Agent Reinforcement Learning
por: Wang, Xuefeng, et al.
Publicado: (2024)
por: Wang, Xuefeng, et al.
Publicado: (2024)
Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations
por: Spieler, Jonathan, et al.
Publicado: (2026)
por: Spieler, Jonathan, et al.
Publicado: (2026)
CA-AC-MPC: CUDA-Accelerated Actor-Critic Model Predictive Control
por: Buo, Antoonio, et al.
Publicado: (2026)
por: Buo, Antoonio, et al.
Publicado: (2026)
DR-MPC: Deep Residual Model Predictive Control for Real-world Social Navigation
por: Han, James R., et al.
Publicado: (2024)
por: Han, James R., et al.
Publicado: (2024)
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
por: Seo, Younggyo, et al.
Publicado: (2025)
por: Seo, Younggyo, et al.
Publicado: (2025)
Neural Network-Assisted Model Predictive Control for Implicit Balancing
por: Madahi, Seyed Soroush Karimi, et al.
Publicado: (2026)
por: Madahi, Seyed Soroush Karimi, et al.
Publicado: (2026)
GLaDiGAtor: Language-Model-Augmented Multi-Relation Graph Learning for Predicting Disease-Gene Associations
por: Kuzucu, Osman Onur, et al.
Publicado: (2026)
por: Kuzucu, Osman Onur, et al.
Publicado: (2026)
Grasp-MPC: Closed-Loop Visual Grasping via Value-Guided Model Predictive Control
por: Yamada, Jun, et al.
Publicado: (2025)
por: Yamada, Jun, et al.
Publicado: (2025)
Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making
por: Hori, Toshiaki, et al.
Publicado: (2025)
por: Hori, Toshiaki, et al.
Publicado: (2025)
Multi-State TD Target for Model-Free Reinforcement Learning
por: Wang, Wuhao, et al.
Publicado: (2024)
por: Wang, Wuhao, et al.
Publicado: (2024)
Implementation of Object-Oriented Design Patterns in Scalable Smart City Architectures
por: Ürün, Efe
Publicado: (2026)
por: Ürün, Efe
Publicado: (2026)
Robotic Arm Manipulation with Inverse Reinforcement Learning & TD-MPC
por: Hassan, Md Shoyib, et al.
Publicado: (2024)
por: Hassan, Md Shoyib, et al.
Publicado: (2024)
PIQL: Projective Implicit Q-Learning with Support Constraint for Offline Reinforcement Learning
por: Han, Xinchen, et al.
Publicado: (2025)
por: Han, Xinchen, et al.
Publicado: (2025)
Policy-Governed LLM Routing with Intent Matching for Instrument Laboratories
por: Olowe, Emmanuel A., et al.
Publicado: (2026)
por: Olowe, Emmanuel A., et al.
Publicado: (2026)
ComTraQ-MPC: Meta-Trained DQN-MPC Integration for Trajectory Tracking with Limited Active Localization Updates
por: Puthumanaillam, Gokul, et al.
Publicado: (2024)
por: Puthumanaillam, Gokul, et al.
Publicado: (2024)
Exploring Adaptive MCTS with TD Learning in miniXCOM
por: Saadat, Kimiya, et al.
Publicado: (2022)
por: Saadat, Kimiya, et al.
Publicado: (2022)
What Does Flow Matching Bring To TD Learning?
por: Agrawalla, Bhavya, et al.
Publicado: (2026)
por: Agrawalla, Bhavya, et al.
Publicado: (2026)
IQL‐OCDA: An intelligent Q‐learning‐based for optimal clustering and data‐aggregation for wireless sensor networks
por: Arwa N. Aledaily
Publicado: (2024)
por: Arwa N. Aledaily
Publicado: (2024)
High-order Interactions Modeling for Interpretable Multi-Agent Q-Learning
por: Xu, Qinyu, et al.
Publicado: (2025)
por: Xu, Qinyu, et al.
Publicado: (2025)
Implicit Maximum Likelihood Estimation for Real-time Generative Model Predictive Control
por: Lee, Grayson, et al.
Publicado: (2026)
por: Lee, Grayson, et al.
Publicado: (2026)
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents
por: Kuzmenko, Dmytro, et al.
Publicado: (2025)
por: Kuzmenko, Dmytro, et al.
Publicado: (2025)
TiMem: Temporal-Hierarchical Memory Consolidation for Long-Horizon Conversational Agents
por: Li, Kai, et al.
Publicado: (2026)
por: Li, Kai, et al.
Publicado: (2026)
Ejemplares similares
-
AlignIQL: Policy Alignment in Implicit Q-Learning through Constrained Optimization
por: He, Longxiang, et al.
Publicado: (2024) -
TD-MPC2: Scalable, Robust World Models for Continuous Control
por: Hansen, Nicklas, et al.
Publicado: (2023) -
DiSA-IQL: Offline Reinforcement Learning for Robust Soft Robot Control under Distribution Shifts
por: He, Linjin, et al.
Publicado: (2025) -
An Empirical Study of Deep Reinforcement Learning in Continuing Tasks
por: Wan, Yi, et al.
Publicado: (2025) -
LeTac-MPC: Learning Model Predictive Control for Tactile-reactive Grasping
por: Xu, Zhengtong, et al.
Publicado: (2024)