Learning Structural Latent Points for Efficient Visual Representations in Robotic Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Yicheng, Wang, Jiaxu, He, Junhao, Gan, Zesen, Li, Junhao, Zhang, Qiang, Sun, Jingkai, Cao, Jiahang, Sun, Mingyuan, Yue, Xiangyu, Shao, Qiming |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MVISTA-4D: View-Consistent 4D World Model with Test-Time Action Inference for Robotic Manipulation
by: Wang, Jiaxu, et al.
Published: (2026)
by: Wang, Jiaxu, et al.
Published: (2026)
Query-based Semantic Gaussian Field for Scene Representation in Reinforcement Learning
by: Wang, Jiaxu, et al.
Published: (2024)
by: Wang, Jiaxu, et al.
Published: (2024)
MoSA: Motion-constrained Stress Adaptation for Mitigating Real-to-Sim Gap in Continuum Dynamics via Learning Residual Anisotropy
by: Wang, Jiaxu, et al.
Published: (2026)
by: Wang, Jiaxu, et al.
Published: (2026)
Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion
by: Zhang, Qiang, et al.
Published: (2025)
by: Zhang, Qiang, et al.
Published: (2025)
LiPS: Large-Scale Humanoid Robot Reinforcement Learning with Parallel-Series Structures
by: Zhang, Qiang, et al.
Published: (2025)
by: Zhang, Qiang, et al.
Published: (2025)
Trinity: A Modular Humanoid Robot AI System
by: Sun, Jingkai, et al.
Published: (2025)
by: Sun, Jingkai, et al.
Published: (2025)
DPL: Depth-only Perceptive Humanoid Locomotion via Realistic Depth Synthesis and Cross-Attention Terrain Reconstruction
by: Sun, Jingkai, et al.
Published: (2025)
by: Sun, Jingkai, et al.
Published: (2025)
Fully Spiking Neural Network for Legged Robots
by: Jiang, Xiaoyang, et al.
Published: (2023)
by: Jiang, Xiaoyang, et al.
Published: (2023)
ES-Parkour: Advanced Robot Parkour with Bio-inspired Event Camera and Spiking Neural Network
by: Zhang, Qiang, et al.
Published: (2025)
by: Zhang, Qiang, et al.
Published: (2025)
Physical Priors Augmented Event-Based 3D Reconstruction
by: Wang, Jiaxu, et al.
Published: (2024)
by: Wang, Jiaxu, et al.
Published: (2024)
Event Masked Autoencoder: Point-wise Action Recognition with Event-Based Cameras
by: Sun, Jingkai, et al.
Published: (2025)
by: Sun, Jingkai, et al.
Published: (2025)
DEGS: Deformable Event-based 3D Gaussian Splatting from RGB and Event Stream
by: He, Junhao, et al.
Published: (2025)
by: He, Junhao, et al.
Published: (2025)
EvGGS: A Collaborative Learning Framework for Event-based Generalizable Gaussian Splatting
by: Wang, Jiaxu, et al.
Published: (2024)
by: Wang, Jiaxu, et al.
Published: (2024)
DEL: Discrete Element Learner for Learning 3D Particle Dynamics with Neural Rendering
by: Wang, Jiaxu, et al.
Published: (2024)
by: Wang, Jiaxu, et al.
Published: (2024)
HumanoidPano: Hybrid Spherical Panoramic-LiDAR Cross-Modal Perception for Humanoid Robots
by: Zhang, Qiang, et al.
Published: (2025)
by: Zhang, Qiang, et al.
Published: (2025)
E2H: A Two-Stage Non-Invasive Neural Signal Driven Humanoid Robotic Whole-Body Control Framework
by: Duan, Yiqun, et al.
Published: (2024)
by: Duan, Yiqun, et al.
Published: (2024)
RoboDexVLM: Visual Language Model-Enabled Task Planning and Motion Control for Dexterous Robot Manipulation
by: Liu, Haichao, et al.
Published: (2025)
by: Liu, Haichao, et al.
Published: (2025)
Mamba Policy: Towards Efficient 3D Diffusion Policy with Hybrid Selective State Models
by: Cao, Jiahang, et al.
Published: (2024)
by: Cao, Jiahang, et al.
Published: (2024)
Occupancy World Model for Robots
by: Zhang, Zhang, et al.
Published: (2025)
by: Zhang, Zhang, et al.
Published: (2025)
PALo: Learning Posture-Aware Locomotion for Quadruped Robots
by: Miao, Xiangyu, et al.
Published: (2025)
by: Miao, Xiangyu, et al.
Published: (2025)
AtomVLA: Scalable Post-Training for Robotic Manipulation via Predictive Latent World Models
by: Sun, Xiaoquan, et al.
Published: (2026)
by: Sun, Xiaoquan, et al.
Published: (2026)
SPECI: Skill Prompts based Hierarchical Continual Imitation Learning for Robot Manipulation
by: Xu, Jingkai, et al.
Published: (2025)
by: Xu, Jingkai, et al.
Published: (2025)
Never-Ending Behavior-Cloning Agent for Robotic Manipulation
by: Liang, Wenqi, et al.
Published: (2024)
by: Liang, Wenqi, et al.
Published: (2024)
Modality-Composable Diffusion Policy via Inference-Time Distribution-level Composition
by: Cao, Jiahang, et al.
Published: (2025)
by: Cao, Jiahang, et al.
Published: (2025)
Language-Conditioned Open-Vocabulary Mobile Manipulation with Pretrained Models
by: Tan, Shen, et al.
Published: (2025)
by: Tan, Shen, et al.
Published: (2025)
DexRepNet++: Learning Dexterous Robotic Manipulation with Geometric and Spatial Hand-Object Representations
by: Liu, Qingtao, et al.
Published: (2026)
by: Liu, Qingtao, et al.
Published: (2026)
Multimodal Spiking Neural Network for Space Robotic Manipulation
by: Zhang, Liwen, et al.
Published: (2025)
by: Zhang, Liwen, et al.
Published: (2025)
Chasing Day and Night: Towards Robust and Efficient All-Day Object Detection Guided by an Event Camera
by: Cao, Jiahang, et al.
Published: (2023)
by: Cao, Jiahang, et al.
Published: (2023)
Continual Hand-Eye Calibration for Open-world Robotic Manipulation
by: Li, Fazeng, et al.
Published: (2026)
by: Li, Fazeng, et al.
Published: (2026)
LaVA-Man: Learning Visual Action Representations for Robot Manipulation
by: Zhu, Chaoran, et al.
Published: (2025)
by: Zhu, Chaoran, et al.
Published: (2025)
Latent Representations for Visual Proprioception in Inexpensive Robots
by: Sheikholeslami, Sahara, et al.
Published: (2025)
by: Sheikholeslami, Sahara, et al.
Published: (2025)
SpatialActor: Exploring Disentangled Spatial Representations for Robust Robotic Manipulation
by: Shi, Hao, et al.
Published: (2025)
by: Shi, Hao, et al.
Published: (2025)
DART: Dual-level Autonomous Robotic Topology for Efficient Exploration in Unknown Environments
by: Wang, Qiming, et al.
Published: (2025)
by: Wang, Qiming, et al.
Published: (2025)
Compose Your Policies! Improving Diffusion-based or Flow-based Robot Policies via Test-time Distribution-level Composition
by: Cao, Jiahang, et al.
Published: (2025)
by: Cao, Jiahang, et al.
Published: (2025)
Fast Visuomotor Policy for Robotic Manipulation
by: Jia, Jingkai, et al.
Published: (2025)
by: Jia, Jingkai, et al.
Published: (2025)
EmbodiedVSR: Dynamic Scene Graph-Guided Chain-of-Thought Reasoning for Visual Spatial Tasks
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
Beyond Viewpoint Generalization: What Multi-View Demonstrations Offer and How to Synthesize Them for Robot Manipulation?
by: Cai, Boyang, et al.
Published: (2026)
by: Cai, Boyang, et al.
Published: (2026)
MCVO: A Generic Visual Odometry for Arbitrarily Arranged Multi-Cameras
by: Yu, Huai, et al.
Published: (2024)
by: Yu, Huai, et al.
Published: (2024)
ManiVID-3D: Generalizable View-Invariant Reinforcement Learning for Robotic Manipulation via Disentangled 3D Representations
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
Empowering Embodied Manipulation: A Bimanual-Mobile Robot Manipulation Dataset for Household Tasks
by: Zhang, Tianle, et al.
Published: (2024)
by: Zhang, Tianle, et al.
Published: (2024)
Similar Items
-
MVISTA-4D: View-Consistent 4D World Model with Test-Time Action Inference for Robotic Manipulation
by: Wang, Jiaxu, et al.
Published: (2026) -
Query-based Semantic Gaussian Field for Scene Representation in Reinforcement Learning
by: Wang, Jiaxu, et al.
Published: (2024) -
MoSA: Motion-constrained Stress Adaptation for Mitigating Real-to-Sim Gap in Continuum Dynamics via Learning Residual Anisotropy
by: Wang, Jiaxu, et al.
Published: (2026) -
Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion
by: Zhang, Qiang, et al.
Published: (2025) -
LiPS: Large-Scale Humanoid Robot Reinforcement Learning with Parallel-Series Structures
by: Zhang, Qiang, et al.
Published: (2025)