ImitDiff: Transferring Foundation-Model Priors for Distraction Robust Visuomotor Policy
Fuente:
arXiv
Saved in:
| Main Authors: | Dong, Yuhang, Ge, Haizhou, Zeng, Yupei, Zhang, Jiangning, Tian, Beiwen, Zhu, Hongrui, Jia, Yufei, Wang, Ruixiang, Xue, Zhucun, Zhou, Guyue, Ma, Longhua, Tian, Guanzhong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Wavelet Policy: Imitation Learning in the Scale Domain with World Prior Memory
by: Yang, Changchuan, et al.
Published: (2025)
by: Yang, Changchuan, et al.
Published: (2025)
Locate n' Rotate: Two-stage Openable Part Detection with Foundation Model Priors
by: Li, Siqi, et al.
Published: (2024)
by: Li, Siqi, et al.
Published: (2024)
Learning Feature Inversion for Multi-class Anomaly Detection under General-purpose COCO-AD Benchmark
by: Zhang, Jiangning, et al.
Published: (2024)
by: Zhang, Jiangning, et al.
Published: (2024)
Bridging the Resource Gap: Deploying Advanced Imitation Learning Models onto Affordable Embedded Platforms
by: Ge, Haizhou, et al.
Published: (2024)
by: Ge, Haizhou, et al.
Published: (2024)
Dual-path Frequency Discriminators for Few-shot Anomaly Detection
by: Bai, Yuhu, et al.
Published: (2024)
by: Bai, Yuhu, et al.
Published: (2024)
Evolution of Video Generative Foundations
by: Hu, Teng, et al.
Published: (2026)
by: Hu, Teng, et al.
Published: (2026)
DISCOVERSE: Efficient Robot Simulation in Complex High-Fidelity Environments
by: Jia, Yufei, et al.
Published: (2025)
by: Jia, Yufei, et al.
Published: (2025)
Learning Multi-view Anomaly Detection with Efficient Adaptive Selection
by: He, Haoyang, et al.
Published: (2024)
by: He, Haoyang, et al.
Published: (2024)
Adaptive Surface Normal Constraint for Geometric Estimation from Monocular Images
by: Long, Xiaoxiao, et al.
Published: (2024)
by: Long, Xiaoxiao, et al.
Published: (2024)
H$^3$DP: Triply-Hierarchical Diffusion Policy for Visuomotor Learning
by: Lu, Yiyang, et al.
Published: (2025)
by: Lu, Yiyang, et al.
Published: (2025)
Key Patch Proposer: Key Patches Contain Rich Information
by: Xu, Jing, et al.
Published: (2024)
by: Xu, Jing, et al.
Published: (2024)
FILIC: Dual-Loop Force-Guided Imitation Learning with Impedance Torque Control for Contact-Rich Manipulation Tasks
by: Ge, Haizhou, et al.
Published: (2025)
by: Ge, Haizhou, et al.
Published: (2025)
FALCON: Actively Decoupled Visuomotor Policies for Loco-Manipulation with Foundation-Model-Based Coordination
by: He, Chengyang, et al.
Published: (2025)
by: He, Chengyang, et al.
Published: (2025)
Oracle-Guided Masked Contrastive Reinforcement Learning for Visuomotor Policies
by: Zhang, Yuhang, et al.
Published: (2025)
by: Zhang, Yuhang, et al.
Published: (2025)
FreqPolicy: Frequency Autoregressive Visuomotor Policy with Continuous Tokens
by: Zhong, Yiming, et al.
Published: (2025)
by: Zhong, Yiming, et al.
Published: (2025)
InstanceV: Instance-Level Video Generation
by: Chen, Yuheng, et al.
Published: (2025)
by: Chen, Yuheng, et al.
Published: (2025)
ImitAL: Learned Active Learning Strategy on Synthetic Data
by: Gonsior, Julius, et al.
Published: (2022)
by: Gonsior, Julius, et al.
Published: (2022)
Learning Adaptive Cross-Embodiment Visuomotor Policy with Contrastive Prompt Orchestration
by: Zhang, Yuhang, et al.
Published: (2026)
by: Zhang, Yuhang, et al.
Published: (2026)
Closed-Loop Visuomotor Control with Generative Expectation for Robotic Manipulation
by: Bu, Qingwen, et al.
Published: (2024)
by: Bu, Qingwen, et al.
Published: (2024)
FocalPolicy: Frequency-Optimized Chunking and Locally Anchored Flow Matching for Coherent Visuomotor Policy
by: He, Qian, et al.
Published: (2026)
by: He, Qian, et al.
Published: (2026)
Improving Autoregressive Visual Generation with Cluster-Oriented Token Prediction
by: Hu, Teng, et al.
Published: (2025)
by: Hu, Teng, et al.
Published: (2025)
Diff-Mosaic: Augmenting Realistic Representations in Infrared Small Target Detection via Diffusion Prior
by: Shi, Yukai, et al.
Published: (2024)
by: Shi, Yukai, et al.
Published: (2024)
Bridge Feature Matching and Cross-Modal Alignment with Mutual-filtering for Zero-shot Anomaly Detection
by: Bai, Yuhu, et al.
Published: (2025)
by: Bai, Yuhu, et al.
Published: (2025)
3D Equivariant Visuomotor Policy Learning via Spherical Projection
by: Hu, Boce, et al.
Published: (2025)
by: Hu, Boce, et al.
Published: (2025)
Spatial-Temporal Aware Visuomotor Diffusion Policy Learning
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
Maximizing Alignment with Minimal Feedback: Efficiently Learning Rewards for Visuomotor Robot Policy Alignment
by: Tian, Ran, et al.
Published: (2024)
by: Tian, Ran, et al.
Published: (2024)
Latency-aware Road Anomaly Segmentation in Videos: A Photorealistic Dataset and New Metrics
by: Tian, Beiwen, et al.
Published: (2024)
by: Tian, Beiwen, et al.
Published: (2024)
TIMotion: Temporal and Interactive Framework for Efficient Human-Human Motion Generation
by: Wang, Yabiao, et al.
Published: (2024)
by: Wang, Yabiao, et al.
Published: (2024)
UltraLBM-UNet: Ultralight Bidirectional Mamba-based Model for Skin Lesion Segmentation
by: Fan, Linxuan, et al.
Published: (2025)
by: Fan, Linxuan, et al.
Published: (2025)
AdaVideoRAG: Omni-Contextual Adaptive Retrieval-Augmented Efficient Long Video Understanding
by: Xue, Zhucun, et al.
Published: (2025)
by: Xue, Zhucun, et al.
Published: (2025)
DiffArtist: Towards Structure and Appearance Controllable Image Stylization
by: Jiang, Ruixiang, et al.
Published: (2024)
by: Jiang, Ruixiang, et al.
Published: (2024)
A Comprehensive Survey of Cross-Domain Policy Transfer for Embodied Agents
by: Niu, Haoyi, et al.
Published: (2024)
by: Niu, Haoyi, et al.
Published: (2024)
DexImit: Learning Bimanual Dexterous Manipulation from Monocular Human Videos
by: Mu, Juncheng, et al.
Published: (2026)
by: Mu, Juncheng, et al.
Published: (2026)
SeFA-Policy: Fast and Accurate Visuomotor Policy Learning with Selective Flow Alignment
by: Xue, Rong, et al.
Published: (2025)
by: Xue, Rong, et al.
Published: (2025)
Perception Stitching: Zero-Shot Perception Encoder Transfer for Visuomotor Robot Policies
by: Jian, Pingcheng, et al.
Published: (2024)
by: Jian, Pingcheng, et al.
Published: (2024)
ManiVID-3D: Generalizable View-Invariant Reinforcement Learning for Robotic Manipulation via Disentangled 3D Representations
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
GaussFly: Contrastive Reinforcement Learning for Visuomotor Policies in 3D Gaussian Fields
by: Zhang, Yuhang, et al.
Published: (2026)
by: Zhang, Yuhang, et al.
Published: (2026)
CLIP-AD: A Language-Guided Staged Dual-Path Model for Zero-shot Anomaly Detection
by: Chen, Xuhai, et al.
Published: (2023)
by: Chen, Xuhai, et al.
Published: (2023)
An Real-Sim-Real (RSR) Loop Framework for Generalizable Robotic Policy Transfer with Differentiable Simulation
by: Shi, Lu, et al.
Published: (2025)
by: Shi, Lu, et al.
Published: (2025)
Efficient Hybrid SE(3)-Equivariant Visuomotor Flow Policy via Spherical Harmonics for Robot Manipulation
by: Zhang, Qinglun, et al.
Published: (2026)
by: Zhang, Qinglun, et al.
Published: (2026)
Similar Items
-
Wavelet Policy: Imitation Learning in the Scale Domain with World Prior Memory
by: Yang, Changchuan, et al.
Published: (2025) -
Locate n' Rotate: Two-stage Openable Part Detection with Foundation Model Priors
by: Li, Siqi, et al.
Published: (2024) -
Learning Feature Inversion for Multi-class Anomaly Detection under General-purpose COCO-AD Benchmark
by: Zhang, Jiangning, et al.
Published: (2024) -
Bridging the Resource Gap: Deploying Advanced Imitation Learning Models onto Affordable Embedded Platforms
by: Ge, Haizhou, et al.
Published: (2024) -
Dual-path Frequency Discriminators for Few-shot Anomaly Detection
by: Bai, Yuhu, et al.
Published: (2024)