RoboAct-CLIP: Video-Driven Pre-training of Atomic Action Understanding for Robotics
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Zhiyuan, He, Yuxin, Sun, Yong, Shi, Junyu, Liu, Lijiang, Nie, Qiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ExoGait-MS: Learning Periodic Dynamics with Multi-Scale Graph Network for Exoskeleton Gait Recognition
by: Liu, Lijiang, et al.
Published: (2025)
by: Liu, Lijiang, et al.
Published: (2025)
RoboGene: Boosting VLA Pre-training via Diversity-Driven Agentic Framework for Real-World Task Generation
by: Zhang, Yixue, et al.
Published: (2026)
by: Zhang, Yixue, et al.
Published: (2026)
RoboGPT-R1: Enhancing Robot Planning with Reinforcement Learning
by: Liu, Jinrui, et al.
Published: (2025)
by: Liu, Jinrui, et al.
Published: (2025)
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training
by: Liu, Yi, et al.
Published: (2025)
by: Liu, Yi, et al.
Published: (2025)
MoGIC: Boosting Motion Generation via Intention Understanding and Visual Context
by: Shi, Junyu, et al.
Published: (2025)
by: Shi, Junyu, et al.
Published: (2025)
RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
Instruct2Act: From Human Instruction to Actions Sequencing and Execution via Robot Action Network for Robotic Manipulation
by: Sharma, Archit, et al.
Published: (2026)
by: Sharma, Archit, et al.
Published: (2026)
Towards Exploratory and Focused Manipulation with Bimanual Active Perception: A New Problem, Benchmark and Strategy
by: He, Yuxin, et al.
Published: (2026)
by: He, Yuxin, et al.
Published: (2026)
CLIP-Motion: Learning Reward Functions for Robotic Actions Using Consecutive Observations
by: Dang, Xuzhe, et al.
Published: (2023)
by: Dang, Xuzhe, et al.
Published: (2023)
Select before Act: Spatially Decoupled Action Repetition for Continuous Control
by: Nie, Buqing, et al.
Published: (2025)
by: Nie, Buqing, et al.
Published: (2025)
RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models
by: Wu, Hao, et al.
Published: (2026)
by: Wu, Hao, et al.
Published: (2026)
RoboOcc: Enhancing the Geometric and Semantic Scene Understanding for Robots
by: Zhang, Zhang, et al.
Published: (2025)
by: Zhang, Zhang, et al.
Published: (2025)
RoboBenchMart: Benchmarking Robots in Retail Environment
by: Soshin, Konstantin, et al.
Published: (2025)
by: Soshin, Konstantin, et al.
Published: (2025)
RoboPilot: Generalizable Dynamic Robotic Manipulation with Dual-thinking Modes
by: Liu, Xinyi, et al.
Published: (2025)
by: Liu, Xinyi, et al.
Published: (2025)
RoboCurate: Harnessing Diversity with Action-Verified Neural Trajectory for Robot Learning
by: Kim, Seungku, et al.
Published: (2026)
by: Kim, Seungku, et al.
Published: (2026)
RoboClaw: An Agentic Framework for Scalable Long-Horizon Robotic Tasks
by: Li, Ruiying, et al.
Published: (2026)
by: Li, Ruiying, et al.
Published: (2026)
MoReFun: Past-Movement Guided Motion Representation Learning for Future Motion Prediction and Understanding
by: Shi, Junyu, et al.
Published: (2024)
by: Shi, Junyu, et al.
Published: (2024)
RoboWits: Unexpected Challenges for Robotic Creative Problem Solving
by: Lin, Chunru, et al.
Published: (2026)
by: Lin, Chunru, et al.
Published: (2026)
RoboPARA: Dual-Arm Robot Planning with Parallel Allocation and Recomposition Across Tasks
by: Duan, Shiying, et al.
Published: (2025)
by: Duan, Shiying, et al.
Published: (2025)
RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation
by: Jiang, Feng, et al.
Published: (2026)
by: Jiang, Feng, et al.
Published: (2026)
Pre-training Auto-regressive Robotic Models with 4D Representations
by: Niu, Dantong, et al.
Published: (2025)
by: Niu, Dantong, et al.
Published: (2025)
Lifelike Agility and Play in Quadrupedal Robots using Reinforcement Learning and Generative Pre-trained Models
by: Han, Lei, et al.
Published: (2023)
by: Han, Lei, et al.
Published: (2023)
CrayonRobo: Object-Centric Prompt-Driven Vision-Language-Action Model for Robotic Manipulation
by: Li, Xiaoqi, et al.
Published: (2025)
by: Li, Xiaoqi, et al.
Published: (2025)
PoseDiff: A Unified Diffusion Model Bridging Robot Pose Estimation and Video-to-Action Control
by: Zhang, Haozhuo, et al.
Published: (2025)
by: Zhang, Haozhuo, et al.
Published: (2025)
RoboFAC: A Comprehensive Framework for Robotic Failure Analysis and Correction
by: Ye, Zewei, et al.
Published: (2025)
by: Ye, Zewei, et al.
Published: (2025)
RoboSSM: Scalable In-context Imitation Learning via State-Space Models
by: Yoo, Youngju, et al.
Published: (2025)
by: Yoo, Youngju, et al.
Published: (2025)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
by: Jiang, Guangqi, et al.
Published: (2024)
by: Jiang, Guangqi, et al.
Published: (2024)
RoboInspector: Unveiling the Unreliability of Policy Code for LLM-enabled Robotic Manipulation
by: Ying, Chenduo, et al.
Published: (2025)
by: Ying, Chenduo, et al.
Published: (2025)
EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
by: Wang, Ruixiang, et al.
Published: (2026)
by: Wang, Ruixiang, et al.
Published: (2026)
Robust Finetuning of Vision-Language-Action Robot Policies via Parameter Merging
by: Yadav, Yajat, et al.
Published: (2025)
by: Yadav, Yajat, et al.
Published: (2025)
A Central Motor System Inspired Pre-training Reinforcement Learning for Robotic Control
by: Zhang, Pei, et al.
Published: (2023)
by: Zhang, Pei, et al.
Published: (2023)
RoboEXP: Action-Conditioned Scene Graph via Interactive Exploration for Robotic Manipulation
by: Jiang, Hanxiao, et al.
Published: (2024)
by: Jiang, Hanxiao, et al.
Published: (2024)
RoboRAN: A Unified Robotics Framework for Reinforcement Learning-Based Autonomous Navigation
by: El-Hariry, Matteo, et al.
Published: (2025)
by: El-Hariry, Matteo, et al.
Published: (2025)
RoboPocket: Improve Robot Policies Instantly with Your Phone
by: Fang, Junjie, et al.
Published: (2026)
by: Fang, Junjie, et al.
Published: (2026)
RoboVIP: Multi-View Video Generation with Visual Identity Prompting Augments Robot Manipulation
by: Wang, Boyang, et al.
Published: (2026)
by: Wang, Boyang, et al.
Published: (2026)
Robo-DM: Data Management For Large Robot Datasets
by: Chen, Kaiyuan, et al.
Published: (2025)
by: Chen, Kaiyuan, et al.
Published: (2025)
RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots
by: Nasiriany, Soroush, et al.
Published: (2024)
by: Nasiriany, Soroush, et al.
Published: (2024)
RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins
by: Mu, Yao, et al.
Published: (2025)
by: Mu, Yao, et al.
Published: (2025)
Retrieval-Augmented Robots via Retrieve-Reason-Act
by: Temiraliev, Izat, et al.
Published: (2026)
by: Temiraliev, Izat, et al.
Published: (2026)
RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models
by: Kwok, Jacky, et al.
Published: (2025)
by: Kwok, Jacky, et al.
Published: (2025)
Similar Items
-
ExoGait-MS: Learning Periodic Dynamics with Multi-Scale Graph Network for Exoskeleton Gait Recognition
by: Liu, Lijiang, et al.
Published: (2025) -
RoboGene: Boosting VLA Pre-training via Diversity-Driven Agentic Framework for Real-World Task Generation
by: Zhang, Yixue, et al.
Published: (2026) -
RoboGPT-R1: Enhancing Robot Planning with Reinforcement Learning
by: Liu, Jinrui, et al.
Published: (2025) -
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training
by: Liu, Yi, et al.
Published: (2025) -
MoGIC: Boosting Motion Generation via Intention Understanding and Visual Context
by: Shi, Junyu, et al.
Published: (2025)