Chain-of-Thought Predictive Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Jia, Zhiwei, Thumuluri, Vineet, Liu, Fangchen, Chen, Linghao, Huang, Zhiao, Su, Hao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Body Transformer: Leveraging Robot Embodiment for Policy Learning
por: Sferrazza, Carmelo, et al.
Publicado: (2024)
por: Sferrazza, Carmelo, et al.
Publicado: (2024)
ExBody2: Advanced Expressive Humanoid Whole-Body Control
por: Ji, Mazeyu, et al.
Publicado: (2024)
por: Ji, Mazeyu, et al.
Publicado: (2024)
Conversational Orientation Reasoning: Egocentric-to-Allocentric Navigation with Multimodal Chain-of-Thought
por: Huang, Yu Ti
Publicado: (2025)
por: Huang, Yu Ti
Publicado: (2025)
Video2Policy: Scaling up Manipulation Tasks in Simulation through Internet Videos
por: Ye, Weirui, et al.
Publicado: (2025)
por: Ye, Weirui, et al.
Publicado: (2025)
Coupled Distributional Random Expert Distillation for World Model Online Imitation Learning
por: Li, Shangzhe, et al.
Publicado: (2025)
por: Li, Shangzhe, et al.
Publicado: (2025)
DrS: Learning Reusable Dense Rewards for Multi-Stage Tasks
por: Mu, Tongzhou, et al.
Publicado: (2024)
por: Mu, Tongzhou, et al.
Publicado: (2024)
Replication of Impedance Identification Experiments on a Reinforcement-Learning-Controlled Digital Twin of Human Elbows
por: Yu, Hao, et al.
Publicado: (2024)
por: Yu, Hao, et al.
Publicado: (2024)
Bootstrapped Model Predictive Control
por: Wang, Yuhang, et al.
Publicado: (2025)
por: Wang, Yuhang, et al.
Publicado: (2025)
SCoTT: Strategic Chain-of-Thought Tasking for Wireless-Aware Robot Navigation in Digital Twins
por: Djuhera, Aladin, et al.
Publicado: (2024)
por: Djuhera, Aladin, et al.
Publicado: (2024)
Implicit Maximum Likelihood Estimation for Real-time Generative Model Predictive Control
por: Lee, Grayson, et al.
Publicado: (2026)
por: Lee, Grayson, et al.
Publicado: (2026)
Exploiting Hybrid Policy in Reinforcement Learning for Interpretable Temporal Logic Manipulation
por: Zhang, Hao, et al.
Publicado: (2024)
por: Zhang, Hao, et al.
Publicado: (2024)
Dynamic Obstacle Avoidance through Uncertainty-Based Adaptive Planning with Diffusion
por: Punyamoorty, Vineet, et al.
Publicado: (2024)
por: Punyamoorty, Vineet, et al.
Publicado: (2024)
Integrating Learning-Based Manipulation and Physics-Based Locomotion for Whole-Body Badminton Robot Control
por: Wang, Haochen, et al.
Publicado: (2025)
por: Wang, Haochen, et al.
Publicado: (2025)
CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
por: Zhao, Qingqing, et al.
Publicado: (2025)
por: Zhao, Qingqing, et al.
Publicado: (2025)
Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination
por: Spieler, Jonathan, et al.
Publicado: (2026)
por: Spieler, Jonathan, et al.
Publicado: (2026)
Learning Adaptive Dexterous Grasping from Single Demonstrations
por: Shi, Liangzhi, et al.
Publicado: (2025)
por: Shi, Liangzhi, et al.
Publicado: (2025)
Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations
por: Spieler, Jonathan, et al.
Publicado: (2026)
por: Spieler, Jonathan, et al.
Publicado: (2026)
Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control
por: Lu, Chenhao, et al.
Publicado: (2024)
por: Lu, Chenhao, et al.
Publicado: (2024)
LaDi-WM: A Latent Diffusion-based World Model for Predictive Manipulation
por: Huang, Yuhang, et al.
Publicado: (2025)
por: Huang, Yuhang, et al.
Publicado: (2025)
Sample-Efficient Expert Query Control in Active Imitation Learning via Conformal Prediction
por: Firouzkouhi, Arad, et al.
Publicado: (2025)
por: Firouzkouhi, Arad, et al.
Publicado: (2025)
DR-MPC: Deep Residual Model Predictive Control for Real-world Social Navigation
por: Han, James R., et al.
Publicado: (2024)
por: Han, James R., et al.
Publicado: (2024)
Reverse Forward Curriculum Learning for Extreme Sample and Demonstration Efficiency in Reinforcement Learning
por: Tao, Stone, et al.
Publicado: (2024)
por: Tao, Stone, et al.
Publicado: (2024)
Q-Guided Stein Variational Model Predictive Control via RL-informed Policy Prior
por: Cai, Shizhe, et al.
Publicado: (2025)
por: Cai, Shizhe, et al.
Publicado: (2025)
Grasp-MPC: Closed-Loop Visual Grasping via Value-Guided Model Predictive Control
por: Yamada, Jun, et al.
Publicado: (2025)
por: Yamada, Jun, et al.
Publicado: (2025)
LodeStar: Long-horizon Dexterity via Synthetic Data Augmentation from Human Demonstrations
por: Wan, Weikang, et al.
Publicado: (2025)
por: Wan, Weikang, et al.
Publicado: (2025)
Learning Humanoid Standing-up Control across Diverse Postures
por: Huang, Tao, et al.
Publicado: (2025)
por: Huang, Tao, et al.
Publicado: (2025)
BehaviorGPT: Smart Agent Simulation for Autonomous Driving with Next-Patch Prediction
por: Zhou, Zikang, et al.
Publicado: (2024)
por: Zhou, Zikang, et al.
Publicado: (2024)
Boundary-to-Region Supervision for Offline Safe Reinforcement Learning
por: Su, Huikang, et al.
Publicado: (2025)
por: Su, Huikang, et al.
Publicado: (2025)
C2F-TP: A Coarse-to-Fine Denoising Framework for Uncertainty-Aware Trajectory Prediction
por: Wang, Zichen, et al.
Publicado: (2024)
por: Wang, Zichen, et al.
Publicado: (2024)
Policy Decorator: Model-Agnostic Online Refinement for Large Policy Model
por: Yuan, Xiu, et al.
Publicado: (2024)
por: Yuan, Xiu, et al.
Publicado: (2024)
DreamControl: Human-Inspired Whole-Body Humanoid Control for Scene Interaction via Guided Diffusion
por: Kalaria, Dvij, et al.
Publicado: (2025)
por: Kalaria, Dvij, et al.
Publicado: (2025)
Alpamayo-R1: Bridging Reasoning and Action Prediction for Generalizable Autonomous Driving in the Long Tail
por: NVIDIA, et al.
Publicado: (2025)
por: NVIDIA, et al.
Publicado: (2025)
Learning Long-Context Diffusion Policies via Past-Token Prediction
por: Torne, Marcel, et al.
Publicado: (2025)
por: Torne, Marcel, et al.
Publicado: (2025)
Perceptive Humanoid Parkour: Chaining Dynamic Human Skills via Motion Matching
por: Wu, Zhen, et al.
Publicado: (2026)
por: Wu, Zhen, et al.
Publicado: (2026)
Neural-Network-Driven Reward Prediction as a Heuristic: Advancing Q-Learning for Mobile Robot Path Planning
por: Ji, Yiming, et al.
Publicado: (2024)
por: Ji, Yiming, et al.
Publicado: (2024)
Self-Discovered Intention-aware Transformer for Multi-modal Vehicle Trajectory Prediction
por: Liu, Diyi, et al.
Publicado: (2026)
por: Liu, Diyi, et al.
Publicado: (2026)
AMO: Adaptive Motion Optimization for Hyper-Dexterous Humanoid Whole-Body Control
por: Li, Jialong, et al.
Publicado: (2025)
por: Li, Jialong, et al.
Publicado: (2025)
Towards Embodiment Scaling Laws in Robot Locomotion
por: Ai, Bo, et al.
Publicado: (2025)
por: Ai, Bo, et al.
Publicado: (2025)
Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms
por: Liu, Jingyi, et al.
Publicado: (2026)
por: Liu, Jingyi, et al.
Publicado: (2026)
Imagination Policy: Using Generative Point Cloud Models for Learning Manipulation Policies
por: Huang, Haojie, et al.
Publicado: (2024)
por: Huang, Haojie, et al.
Publicado: (2024)
Ejemplares similares
-
Body Transformer: Leveraging Robot Embodiment for Policy Learning
por: Sferrazza, Carmelo, et al.
Publicado: (2024) -
ExBody2: Advanced Expressive Humanoid Whole-Body Control
por: Ji, Mazeyu, et al.
Publicado: (2024) -
Conversational Orientation Reasoning: Egocentric-to-Allocentric Navigation with Multimodal Chain-of-Thought
por: Huang, Yu Ti
Publicado: (2025) -
Video2Policy: Scaling up Manipulation Tasks in Simulation through Internet Videos
por: Ye, Weirui, et al.
Publicado: (2025) -
Coupled Distributional Random Expert Distillation for World Model Online Imitation Learning
por: Li, Shangzhe, et al.
Publicado: (2025)