RoboOcc: Enhancing the Geometric and Semantic Scene Understanding for Robots
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Zhang, Zhang, Qiang, Cui, Wei, Shi, Shuai, Guo, Yijie, Han, Gang, Zhao, Wen, Ren, Hengle, Xu, Renjing, Tang, Jian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Whole-body Humanoid Robot Locomotion with Human Reference
by: Zhang, Qiang, et al.
Published: (2024)
by: Zhang, Qiang, et al.
Published: (2024)
Occupancy World Model for Robots
by: Zhang, Zhang, et al.
Published: (2025)
by: Zhang, Zhang, et al.
Published: (2025)
LiPS: Large-Scale Humanoid Robot Reinforcement Learning with Parallel-Series Structures
by: Zhang, Qiang, et al.
Published: (2025)
by: Zhang, Qiang, et al.
Published: (2025)
Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion
by: Zhang, Qiang, et al.
Published: (2025)
by: Zhang, Qiang, et al.
Published: (2025)
RobotPan: A 360$^\circ$ Surround-View Robotic Vision System for Embodied Perception
by: Ma, Jiahao, et al.
Published: (2026)
by: Ma, Jiahao, et al.
Published: (2026)
Trinity: A Modular Humanoid Robot AI System
by: Sun, Jingkai, et al.
Published: (2025)
by: Sun, Jingkai, et al.
Published: (2025)
HumanoidPano: Hybrid Spherical Panoramic-LiDAR Cross-Modal Perception for Humanoid Robots
by: Zhang, Qiang, et al.
Published: (2025)
by: Zhang, Qiang, et al.
Published: (2025)
ES-Parkour: Advanced Robot Parkour with Bio-inspired Event Camera and Spiking Neural Network
by: Zhang, Qiang, et al.
Published: (2025)
by: Zhang, Qiang, et al.
Published: (2025)
Query-based Semantic Gaussian Field for Scene Representation in Reinforcement Learning
by: Wang, Jiaxu, et al.
Published: (2024)
by: Wang, Jiaxu, et al.
Published: (2024)
RoboAct-CLIP: Video-Driven Pre-training of Atomic Action Understanding for Robotics
by: Zhang, Zhiyuan, et al.
Published: (2025)
by: Zhang, Zhiyuan, et al.
Published: (2025)
MeshMimic: Geometry-Aware Humanoid Motion Learning through 3D Scene Reconstruction
by: Zhang, Qiang, et al.
Published: (2026)
by: Zhang, Qiang, et al.
Published: (2026)
RoboAug: One Annotation to Hundreds of Scenes via Region-Contrastive Data Augmentation for Robotic Manipulation
by: Wang, Xinhua, et al.
Published: (2026)
by: Wang, Xinhua, et al.
Published: (2026)
Semantically Safe Robot Manipulation: From Semantic Scene Understanding to Motion Safeguards
by: Brunke, Lukas, et al.
Published: (2024)
by: Brunke, Lukas, et al.
Published: (2024)
Prompting Multi-Modal Tokens to Enhance End-to-End Autonomous Driving Imitation Learning with LLMs
by: Duan, Yiqun, et al.
Published: (2024)
by: Duan, Yiqun, et al.
Published: (2024)
MobileOcc: A Human-Aware Semantic Occupancy Dataset for Mobile Robots
by: Kim, Junseo, et al.
Published: (2025)
by: Kim, Junseo, et al.
Published: (2025)
RoboEye: Enhancing 2D Robotic Object Identification with Selective 3D Geometric Keypoint Matching
by: Zhang, Xingwu, et al.
Published: (2025)
by: Zhang, Xingwu, et al.
Published: (2025)
Semantic Co-Speech Gesture Synthesis and Real-Time Control for Humanoid Robots
by: Zhang, Gang
Published: (2025)
by: Zhang, Gang
Published: (2025)
RoboRetriever: Single-Camera Robot Object Retrieval via Active and Interactive Perception with Dynamic Scene Graph
by: Wang, Hecheng, et al.
Published: (2025)
by: Wang, Hecheng, et al.
Published: (2025)
Robo-ABC: Affordance Generalization Beyond Categories via Semantic Correspondence for Robot Manipulation
by: Ju, Yuanchen, et al.
Published: (2024)
by: Ju, Yuanchen, et al.
Published: (2024)
Mobile Robot Oriented Large-Scale Indoor Dataset for Dynamic Scene Understanding
by: Tang, Yifan, et al.
Published: (2024)
by: Tang, Yifan, et al.
Published: (2024)
RoboPearls: Editable Video Simulation for Robot Manipulation
by: Tang, Tao, et al.
Published: (2025)
by: Tang, Tao, et al.
Published: (2025)
OccRWKV: Rethinking Efficient 3D Semantic Occupancy Prediction with Linear Complexity
by: Wang, Junming, et al.
Published: (2024)
by: Wang, Junming, et al.
Published: (2024)
RoboAfford++: A Generative AI-Enhanced Dataset for Multimodal Affordance Learning in Robotic Manipulation and Navigation
by: Hao, Xiaoshuai, et al.
Published: (2025)
by: Hao, Xiaoshuai, et al.
Published: (2025)
DPL: Depth-only Perceptive Humanoid Locomotion via Realistic Depth Synthesis and Cross-Attention Terrain Reconstruction
by: Sun, Jingkai, et al.
Published: (2025)
by: Sun, Jingkai, et al.
Published: (2025)
OneOcc: Semantic Occupancy Prediction for Legged Robots with a Single Panoramic Camera
by: Shi, Hao, et al.
Published: (2025)
by: Shi, Hao, et al.
Published: (2025)
Robo-Dopamine: General Process Reward Modeling for High-Precision Robotic Manipulation
by: Tan, Huajie, et al.
Published: (2025)
by: Tan, Huajie, et al.
Published: (2025)
RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark
by: Lei, Huashuo, et al.
Published: (2026)
by: Lei, Huashuo, et al.
Published: (2026)
RoboRouter: Training-Free Policy Routing for Robotic Manipulation
by: Chen, Yiteng, et al.
Published: (2026)
by: Chen, Yiteng, et al.
Published: (2026)
RoboPARA: Dual-Arm Robot Planning with Parallel Allocation and Recomposition Across Tasks
by: Duan, Shiying, et al.
Published: (2025)
by: Duan, Shiying, et al.
Published: (2025)
RoboPaint: From Human Demonstration to Any Robot and Any View
by: Fan, Jiacheng, et al.
Published: (2026)
by: Fan, Jiacheng, et al.
Published: (2026)
OpenOcc: Open Vocabulary 3D Scene Reconstruction via Occupancy Representation
by: Jiang, Haochen, et al.
Published: (2024)
by: Jiang, Haochen, et al.
Published: (2024)
RoboCade: Gamifying Robot Data Collection
by: Mirchandani, Suvir, et al.
Published: (2025)
by: Mirchandani, Suvir, et al.
Published: (2025)
Semantic2D: Enabling Semantic Scene Understanding with 2D Lidar Alone
by: Xie, Zhanteng, et al.
Published: (2024)
by: Xie, Zhanteng, et al.
Published: (2024)
RoboEngine: Plug-and-Play Robot Data Augmentation with Semantic Robot Segmentation and Background Generation
by: Yuan, Chengbo, et al.
Published: (2025)
by: Yuan, Chengbo, et al.
Published: (2025)
RoboPocket: Improve Robot Policies Instantly with Your Phone
by: Fang, Junjie, et al.
Published: (2026)
by: Fang, Junjie, et al.
Published: (2026)
FedRC: A Rapid-Converged Hierarchical Federated Learning Framework in Street Scene Semantic Understanding
by: Kou, Wei-Bin, et al.
Published: (2024)
by: Kou, Wei-Bin, et al.
Published: (2024)
Fully Spiking Neural Network for Legged Robots
by: Jiang, Xiaoyang, et al.
Published: (2023)
by: Jiang, Xiaoyang, et al.
Published: (2023)
MapNav: A Novel Memory Representation via Annotated Semantic Maps for Vision-and-Language Navigation
by: Zhang, Lingfeng, et al.
Published: (2025)
by: Zhang, Lingfeng, et al.
Published: (2025)
VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation
by: Zhang, Haochen, et al.
Published: (2024)
by: Zhang, Haochen, et al.
Published: (2024)
Semantic Scene Segmentation for Robotics
by: Hurtado, Juana Valeria, et al.
Published: (2024)
by: Hurtado, Juana Valeria, et al.
Published: (2024)
Similar Items
-
Whole-body Humanoid Robot Locomotion with Human Reference
by: Zhang, Qiang, et al.
Published: (2024) -
Occupancy World Model for Robots
by: Zhang, Zhang, et al.
Published: (2025) -
LiPS: Large-Scale Humanoid Robot Reinforcement Learning with Parallel-Series Structures
by: Zhang, Qiang, et al.
Published: (2025) -
Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion
by: Zhang, Qiang, et al.
Published: (2025) -
RobotPan: A 360$^\circ$ Surround-View Robotic Vision System for Embodied Perception
by: Ma, Jiahao, et al.
Published: (2026)