GROVE: A Generalized Reward for Learning Open-Vocabulary Physical Skill
Fuente:
arXiv
Salvato in:
| Autori principali: | Cui, Jieming, Liu, Tengyu, Meng, Ziyu, Yu, Jiale, Song, Ran, Zhang, Wei, Zhu, Yixin, Huang, Siyuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AnySkill: Learning Open-Vocabulary Physical Skill for Interactive Agents
di: Cui, Jieming, et al.
Pubblicazione: (2024)
di: Cui, Jieming, et al.
Pubblicazione: (2024)
Ag2Manip: Learning Novel Manipulation Skills with Agent-Agnostic Visual and Action Representations
di: Li, Puhao, et al.
Pubblicazione: (2024)
di: Li, Puhao, et al.
Pubblicazione: (2024)
Simultaneous Tactile-Visual Perception for Learning Multimodal Robot Manipulation
di: Li, Yuyang, et al.
Pubblicazione: (2025)
di: Li, Yuyang, et al.
Pubblicazione: (2025)
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
di: Jiang, Nan, et al.
Pubblicazione: (2025)
di: Jiang, Nan, et al.
Pubblicazione: (2025)
Grasp Multiple Objects with One Hand
di: Li, Yuyang, et al.
Pubblicazione: (2023)
di: Li, Yuyang, et al.
Pubblicazione: (2023)
Taccel: Scaling Up Vision-based Tactile Robotics via High-performance GPU Simulation
di: Li, Yuyang, et al.
Pubblicazione: (2025)
di: Li, Yuyang, et al.
Pubblicazione: (2025)
ManipTrans: Efficient Dexterous Bimanual Manipulation Transfer via Residual Learning
di: Li, Kailin, et al.
Pubblicazione: (2025)
di: Li, Kailin, et al.
Pubblicazione: (2025)
CLONE: Closed-Loop Whole-Body Humanoid Teleoperation for Long-Horizon Tasks
di: Li, Yixuan, et al.
Pubblicazione: (2025)
di: Li, Yixuan, et al.
Pubblicazione: (2025)
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing Scenes
di: Jiang, Jiajun, et al.
Pubblicazione: (2025)
di: Jiang, Jiajun, et al.
Pubblicazione: (2025)
SafeFall: Learning Protective Control for Humanoid Robots
di: Meng, Ziyu, et al.
Pubblicazione: (2025)
di: Meng, Ziyu, et al.
Pubblicazione: (2025)
OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies
di: Kong, Lingdong, et al.
Pubblicazione: (2024)
di: Kong, Lingdong, et al.
Pubblicazione: (2024)
SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding
di: Jia, Baoxiong, et al.
Pubblicazione: (2024)
di: Jia, Baoxiong, et al.
Pubblicazione: (2024)
Closed-Loop Open-Vocabulary Mobile Manipulation with GPT-4V
di: Zhi, Peiyuan, et al.
Pubblicazione: (2024)
di: Zhi, Peiyuan, et al.
Pubblicazione: (2024)
Re$^2$MoGen: Open-Vocabulary Motion Generation via LLM Reasoning and Physics-Aware Refinement
di: Zheng, Jiakun, et al.
Pubblicazione: (2026)
di: Zheng, Jiakun, et al.
Pubblicazione: (2026)
OpenOcc: Open Vocabulary 3D Scene Reconstruction via Occupancy Representation
di: Jiang, Haochen, et al.
Pubblicazione: (2024)
di: Jiang, Haochen, et al.
Pubblicazione: (2024)
StyleLoco: Generative Adversarial Distillation for Natural Humanoid Robot Locomotion
di: Ma, Le, et al.
Pubblicazione: (2025)
di: Ma, Le, et al.
Pubblicazione: (2025)
OpenGaussian: Towards Point-Level 3D Gaussian-based Open Vocabulary Understanding
di: Wu, Yanmin, et al.
Pubblicazione: (2024)
di: Wu, Yanmin, et al.
Pubblicazione: (2024)
SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields
di: Liu, Yu, et al.
Pubblicazione: (2024)
di: Liu, Yu, et al.
Pubblicazione: (2024)
MetaScenes: Towards Automated Replica Creation for Real-world 3D Scans
di: Yu, Huangyue, et al.
Pubblicazione: (2025)
di: Yu, Huangyue, et al.
Pubblicazione: (2025)
Open-Vocabulary Online Semantic Mapping for SLAM
di: Martins, Tomas Berriel, et al.
Pubblicazione: (2024)
di: Martins, Tomas Berriel, et al.
Pubblicazione: (2024)
LOVON: Legged Open-Vocabulary Object Navigator
di: Peng, Daojie, et al.
Pubblicazione: (2025)
di: Peng, Daojie, et al.
Pubblicazione: (2025)
Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation
di: Dong, Runpei, et al.
Pubblicazione: (2026)
di: Dong, Runpei, et al.
Pubblicazione: (2026)
Kinematify: Open-Vocabulary Synthesis of High-DoF Articulated Objects
di: Wang, Jiawei, et al.
Pubblicazione: (2025)
di: Wang, Jiawei, et al.
Pubblicazione: (2025)
Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction
di: Yu, Xuan, et al.
Pubblicazione: (2025)
di: Yu, Xuan, et al.
Pubblicazione: (2025)
GenFlowRL: Shaping Rewards with Generative Object-Centric Flow in Visual Reinforcement Learning
di: Yu, Kelin, et al.
Pubblicazione: (2025)
di: Yu, Kelin, et al.
Pubblicazione: (2025)
WildOS: Open-Vocabulary Object Search in the Wild
di: Shah, Hardik, et al.
Pubblicazione: (2026)
di: Shah, Hardik, et al.
Pubblicazione: (2026)
Dissecting Embodied Abilities in Multimodal Language Models through Skill-level Evaluation and Diagnosis
di: Qi, Yu, et al.
Pubblicazione: (2025)
di: Qi, Yu, et al.
Pubblicazione: (2025)
Masked Point-Entity Contrast for Open-Vocabulary 3D Scene Understanding
di: Wang, Yan, et al.
Pubblicazione: (2025)
di: Wang, Yan, et al.
Pubblicazione: (2025)
OVerSeeC: Open-Vocabulary Costmap Generation from Satellite Images and Natural Language
di: Rana, Rwik, et al.
Pubblicazione: (2026)
di: Rana, Rwik, et al.
Pubblicazione: (2026)
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
di: Sun, Haowen, et al.
Pubblicazione: (2026)
di: Sun, Haowen, et al.
Pubblicazione: (2026)
ModSkill: Physical Character Skill Modularization
di: Huang, Yiming, et al.
Pubblicazione: (2025)
di: Huang, Yiming, et al.
Pubblicazione: (2025)
Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Tracking
di: Ishaq, Ayesha, et al.
Pubblicazione: (2024)
di: Ishaq, Ayesha, et al.
Pubblicazione: (2024)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
di: Cai, Junhao, et al.
Pubblicazione: (2024)
di: Cai, Junhao, et al.
Pubblicazione: (2024)
Are Open-Vocabulary Models Ready for Detection of MEP Elements on Construction Sites
di: Abdalwhab, Abdalwhab, et al.
Pubblicazione: (2025)
di: Abdalwhab, Abdalwhab, et al.
Pubblicazione: (2025)
Leveraging Vision-Language Models for Open-Vocabulary Instance Segmentation and Tracking
di: Pätzold, Bastian, et al.
Pubblicazione: (2025)
di: Pätzold, Bastian, et al.
Pubblicazione: (2025)
OVGrasp: Open-Vocabulary Grasping Assistance via Multimodal Intent Detection
di: Hu, Chen, et al.
Pubblicazione: (2025)
di: Hu, Chen, et al.
Pubblicazione: (2025)
GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
di: Ma, Teli, et al.
Pubblicazione: (2024)
di: Ma, Teli, et al.
Pubblicazione: (2024)
FreeOcc: Training-Free Embodied Open-Vocabulary Occupancy Prediction
di: Jiang, Zeyu, et al.
Pubblicazione: (2026)
di: Jiang, Zeyu, et al.
Pubblicazione: (2026)
The Bare Necessities: Designing Simple, Effective Open-Vocabulary Scene Graphs
di: Kassab, Christina, et al.
Pubblicazione: (2024)
di: Kassab, Christina, et al.
Pubblicazione: (2024)
Adapt2Reward: Adapting Video-Language Models to Generalizable Robotic Rewards via Failure Prompts
di: Yang, Yanting, et al.
Pubblicazione: (2024)
di: Yang, Yanting, et al.
Pubblicazione: (2024)
Documenti analoghi
-
AnySkill: Learning Open-Vocabulary Physical Skill for Interactive Agents
di: Cui, Jieming, et al.
Pubblicazione: (2024) -
Ag2Manip: Learning Novel Manipulation Skills with Agent-Agnostic Visual and Action Representations
di: Li, Puhao, et al.
Pubblicazione: (2024) -
Simultaneous Tactile-Visual Perception for Learning Multimodal Robot Manipulation
di: Li, Yuyang, et al.
Pubblicazione: (2025) -
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
di: Jiang, Nan, et al.
Pubblicazione: (2025) -
Grasp Multiple Objects with One Hand
di: Li, Yuyang, et al.
Pubblicazione: (2023)