Data-Efficient Approach to Humanoid Control via Fine-Tuning a Pre-Trained GPT on Action Data
Fuente:
arXiv
Guardado en:
| Autores principales: | Padmanabhan, Siddharth, Miyazawa, Kazuki, Horii, Takato, Nagai, Takayuki |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Predictive Reachability for Embodiment Selection in Mobile Manipulation Behaviors
por: Feng, Xiaoxu, et al.
Publicado: (2024)
por: Feng, Xiaoxu, et al.
Publicado: (2024)
Creative Agents: Simulating the Systems Model of Creativity with Generative Agents
por: Imasato, Naomi, et al.
Publicado: (2024)
por: Imasato, Naomi, et al.
Publicado: (2024)
LiP-LLM: Integrating Linear Programming and dependency graph with Large Language Models for multi-robot task planning
por: Obata, Kazuma, et al.
Publicado: (2024)
por: Obata, Kazuma, et al.
Publicado: (2024)
Goal Estimation-based Adaptive Shared Control for Brain-Machine Interfaces Remote Robot Navigation
por: Muraoka, Tomoka, et al.
Publicado: (2024)
por: Muraoka, Tomoka, et al.
Publicado: (2024)
SADP: Subgoal-Aware Diffusion Policy for Explainable Robots Learned from Foundation Model Generated Demonstrations
por: Hu, Site, et al.
Publicado: (2026)
por: Hu, Site, et al.
Publicado: (2026)
Towards Adaptive Humanoid Control via Multi-Behavior Distillation and Reinforced Fine-Tuning
por: Zhao, Yingnan, et al.
Publicado: (2025)
por: Zhao, Yingnan, et al.
Publicado: (2025)
External Photoreflective Tactile Sensing Based on Surface Deformation Measurement
por: Yamamoto, Seiichi, et al.
Publicado: (2025)
por: Yamamoto, Seiichi, et al.
Publicado: (2025)
PPF: Pre-training and Preservative Fine-tuning of Humanoid Locomotion via Model-Assumption-based Regularization
por: Jung, Hyunyoung, et al.
Publicado: (2025)
por: Jung, Hyunyoung, et al.
Publicado: (2025)
Study of Emotion Concept Formation by Integrating Vision, Physiology, and Word Information using Multilayered Multimodal Latent Dirichlet Allocation
por: Tsurumaki, Kazuki, et al.
Publicado: (2024)
por: Tsurumaki, Kazuki, et al.
Publicado: (2024)
Agility Meets Stability: Versatile Humanoid Control with Heterogeneous Data
por: Pan, Yixuan, et al.
Publicado: (2025)
por: Pan, Yixuan, et al.
Publicado: (2025)
Integrating Offline Pre-Training with Online Fine-Tuning: A Reinforcement Learning Approach for Robot Social Navigation
por: Su, Run, et al.
Publicado: (2025)
por: Su, Run, et al.
Publicado: (2025)
Design and Control of a Small Humanoid Equipped with Flight Unit and Wheels for Multimodal Locomotion
por: Sugihara, Kazuki, et al.
Publicado: (2023)
por: Sugihara, Kazuki, et al.
Publicado: (2023)
D-CLING: Prior-Preserving Depth-Conditioned Fine-Tuning for Navigation Foundation Models
por: Nakaoka, Shintaro, et al.
Publicado: (2026)
por: Nakaoka, Shintaro, et al.
Publicado: (2026)
Unveiling the Impact of Data and Model Scaling on High-Level Control for Humanoid Robots
por: Wei, Yuxi, et al.
Publicado: (2025)
por: Wei, Yuxi, et al.
Publicado: (2025)
ULC: A Unified and Fine-Grained Controller for Humanoid Loco-Manipulation
por: Sun, Wandong, et al.
Publicado: (2025)
por: Sun, Wandong, et al.
Publicado: (2025)
Efficient Automatic Tuning for Data-driven Model Predictive Control via Meta-Learning
por: Li, Baoyu, et al.
Publicado: (2024)
por: Li, Baoyu, et al.
Publicado: (2024)
LEGS: Fine-Tuning Teleop-Free VLAs for Humanoid Loco-manipulation in an Embodied Gaussian Splatting World
por: Kim, Hojune, et al.
Publicado: (2026)
por: Kim, Hojune, et al.
Publicado: (2026)
Exciting Action: Investigating Efficient Exploration for Learning Musculoskeletal Humanoid Locomotion
por: Geiß, Henri-Jacques, et al.
Publicado: (2024)
por: Geiß, Henri-Jacques, et al.
Publicado: (2024)
SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows
por: Yang, Chenyu, et al.
Publicado: (2026)
por: Yang, Chenyu, et al.
Publicado: (2026)
System 0/1/2/3: Quad-process theory for multi-timescale embodied collective cognitive systems
por: Taniguchi, Tadahiro, et al.
Publicado: (2025)
por: Taniguchi, Tadahiro, et al.
Publicado: (2025)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
por: Xiao, Wei, et al.
Publicado: (2025)
por: Xiao, Wei, et al.
Publicado: (2025)
Dynamic Properties and Motion Reproducibility of a Compact Pneumatically Actuated Humanoid Upper Body for Data-Driven Control
por: Atsuta, Hiroshi, et al.
Publicado: (2026)
por: Atsuta, Hiroshi, et al.
Publicado: (2026)
Actions as Language: Fine-Tuning VLMs into VLAs Without Catastrophic Forgetting
por: Hancock, Asher J., et al.
Publicado: (2025)
por: Hancock, Asher J., et al.
Publicado: (2025)
GPT-Fabric: Smoothing and Folding Fabric by Leveraging Pre-Trained Foundation Models
por: Raval, Vedant, et al.
Publicado: (2024)
por: Raval, Vedant, et al.
Publicado: (2024)
Proceedings of the Dialogue Robot Competition 2023
por: Higashinaka, Ryuichiro, et al.
Publicado: (2023)
por: Higashinaka, Ryuichiro, et al.
Publicado: (2023)
Efficient and Compliant Control Framework for Versatile Human-Humanoid Collaborative Transportation
por: Kumbhar, Shubham S., et al.
Publicado: (2025)
por: Kumbhar, Shubham S., et al.
Publicado: (2025)
Data Efficient Behavior Cloning for Fine Manipulation via Continuity-based Corrective Labels
por: Deshpande, Abhay, et al.
Publicado: (2024)
por: Deshpande, Abhay, et al.
Publicado: (2024)
PvP: Data-Efficient Humanoid Robot Learning with Proprioceptive-Privileged Contrastive Representations
por: Yuan, Mingqi, et al.
Publicado: (2025)
por: Yuan, Mingqi, et al.
Publicado: (2025)
Synthetic Data Pipelines for Adaptive, Mission-Ready Militarized Humanoids
por: Habib, Mohammed Ayman, et al.
Publicado: (2025)
por: Habib, Mohammed Ayman, et al.
Publicado: (2025)
HumanoidGen: Data Generation for Bimanual Dexterous Manipulation via LLM Reasoning
por: Jing, Zhi, et al.
Publicado: (2025)
por: Jing, Zhi, et al.
Publicado: (2025)
GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data
por: Deng, Shengliang, et al.
Publicado: (2025)
por: Deng, Shengliang, et al.
Publicado: (2025)
ExFace: Expressive Facial Control for Humanoid Robots with Diffusion Transformers and Bootstrap Training
por: Zhang, Dong, et al.
Publicado: (2025)
por: Zhang, Dong, et al.
Publicado: (2025)
Fine-Tuned Convex Approximations of Probabilistic Reachable Sets under Data-driven Uncertainties
por: Wu, Pengcheng, et al.
Publicado: (2023)
por: Wu, Pengcheng, et al.
Publicado: (2023)
PHASOR: Phase-Anchored Universal Action Representations for Humanoid Embodiments
por: Kim, Kihyun, et al.
Publicado: (2026)
por: Kim, Kihyun, et al.
Publicado: (2026)
CHILD (Controller for Humanoid Imitation and Live Demonstration): a Whole-Body Humanoid Teleoperation System
por: Myers, Noboru, et al.
Publicado: (2025)
por: Myers, Noboru, et al.
Publicado: (2025)
HumanoidMimicGen: Data Generation for Loco-Manipulation via Whole-Body Planning
por: Lin, Kevin, et al.
Publicado: (2026)
por: Lin, Kevin, et al.
Publicado: (2026)
CO-RFT: Efficient Fine-Tuning of Vision-Language-Action Models through Chunked Offline Reinforcement Learning
por: Huang, Dongchi, et al.
Publicado: (2025)
por: Huang, Dongchi, et al.
Publicado: (2025)
General Humanoid Whole-Body Control via Pretraining and Fast Adaptation
por: Wang, Zepeng, et al.
Publicado: (2026)
por: Wang, Zepeng, et al.
Publicado: (2026)
Dynamic Whole-Body Dancing with Humanoid Robots -- A Model-Based Control Approach
por: Zhang, Shibowen, et al.
Publicado: (2026)
por: Zhang, Shibowen, et al.
Publicado: (2026)
HWC-Loco: A Hierarchical Whole-Body Control Approach to Robust Humanoid Locomotion
por: Lin, Sixu, et al.
Publicado: (2025)
por: Lin, Sixu, et al.
Publicado: (2025)
Ejemplares similares
-
Predictive Reachability for Embodiment Selection in Mobile Manipulation Behaviors
por: Feng, Xiaoxu, et al.
Publicado: (2024) -
Creative Agents: Simulating the Systems Model of Creativity with Generative Agents
por: Imasato, Naomi, et al.
Publicado: (2024) -
LiP-LLM: Integrating Linear Programming and dependency graph with Large Language Models for multi-robot task planning
por: Obata, Kazuma, et al.
Publicado: (2024) -
Goal Estimation-based Adaptive Shared Control for Brain-Machine Interfaces Remote Robot Navigation
por: Muraoka, Tomoka, et al.
Publicado: (2024) -
SADP: Subgoal-Aware Diffusion Policy for Explainable Robots Learned from Foundation Model Generated Demonstrations
por: Hu, Site, et al.
Publicado: (2026)