Policy Contrastive Decoding for Robotic Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Shihan, Luo, Xu, Zhang, Ji, Xie, Junlin, Song, Jingkuan, Shen, Heng Tao, Gao, Lianli |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
InSpire: Vision-Language-Action Models with Intrinsic Spatial Reasoning
by: Zhang, Ji, et al.
Published: (2025)
by: Zhang, Ji, et al.
Published: (2025)
Shortcut Learning in Generalist Robot Policies: The Role of Dataset Diversity and Fragmentation
by: Xing, Youguang, et al.
Published: (2025)
by: Xing, Youguang, et al.
Published: (2025)
Beyond the Majority: Long-tail Imitation Learning for Robotic Manipulation
by: Zhu, Junhong, et al.
Published: (2026)
by: Zhu, Junhong, et al.
Published: (2026)
Sim-and-Human Co-training for Data-Efficient and Generalizable Robotic Manipulation
by: Fang, Kaipeng, et al.
Published: (2026)
by: Fang, Kaipeng, et al.
Published: (2026)
DePT: Decoupled Prompt Tuning
by: Zhang, Ji, et al.
Published: (2023)
by: Zhang, Ji, et al.
Published: (2023)
A Closer Look at Conditional Prompt Tuning for Vision-Language Models
by: Zhang, Ji, et al.
Published: (2025)
by: Zhang, Ji, et al.
Published: (2025)
Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves
by: Wu, Shihan, et al.
Published: (2024)
by: Wu, Shihan, et al.
Published: (2024)
Benchmarking Few-shot Transferability of Pre-trained Models with Improved Evaluation Protocols
by: Luo, Xu, et al.
Published: (2026)
by: Luo, Xu, et al.
Published: (2026)
Language-Grounded Decoupled Action Representation for Robotic Manipulation
by: Weng, Wuding, et al.
Published: (2026)
by: Weng, Wuding, et al.
Published: (2026)
MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training
by: Yin, Zhenhan, et al.
Published: (2025)
by: Yin, Zhenhan, et al.
Published: (2025)
A Survey on Efficient Vision-Language-Action Models
by: Yu, Zhaoshu, et al.
Published: (2025)
by: Yu, Zhaoshu, et al.
Published: (2025)
Reliable Few-shot Learning under Dual Noises
by: Zhang, Ji, et al.
Published: (2025)
by: Zhang, Ji, et al.
Published: (2025)
AICL: Action In-Context Learning for Video Diffusion Model
by: Liu, Jianzhi, et al.
Published: (2024)
by: Liu, Jianzhi, et al.
Published: (2024)
Alleviating Hallucinations in Large Vision-Language Models through Hallucination-Induced Optimization
by: Lyu, Xinyu, et al.
Published: (2024)
by: Lyu, Xinyu, et al.
Published: (2024)
STDArm: Transferring Visuomotor Policies From Static Data Training to Dynamic Robot Manipulation
by: Duan, Yifan, et al.
Published: (2025)
by: Duan, Yifan, et al.
Published: (2025)
TIMI: Training-Free Image-to-3D Multi-Instance Generation with Spatial Fidelity
by: Cai, Xiao, et al.
Published: (2026)
by: Cai, Xiao, et al.
Published: (2026)
Attention Hijackers: Detect and Disentangle Attention Hijacking in LVLMs for Hallucination Mitigation
by: Chen, Beitao, et al.
Published: (2025)
by: Chen, Beitao, et al.
Published: (2025)
CFReID: Continual Few-shot Person Re-Identification
by: Ni, Hao, et al.
Published: (2025)
by: Ni, Hao, et al.
Published: (2025)
SafePTR: Token-Level Jailbreak Defense in Multimodal LLMs via Prune-then-Restore Mechanism
by: Chen, Beitao, et al.
Published: (2025)
by: Chen, Beitao, et al.
Published: (2025)
Prototype-based Aleatoric Uncertainty Quantification for Cross-modal Retrieval
by: Li, Hao, et al.
Published: (2023)
by: Li, Hao, et al.
Published: (2023)
From Channel Bias to Feature Redundancy: Uncovering the "Less is More" Principle in Few-Shot Learning
by: Zhang, Ji, et al.
Published: (2023)
by: Zhang, Ji, et al.
Published: (2023)
Unlocking Smarter Device Control: Foresighted Planning with a World Model-Driven Code Execution Approach
by: Yin, Xiaoran, et al.
Published: (2025)
by: Yin, Xiaoran, et al.
Published: (2025)
Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control
by: Gao, Yuxuan, et al.
Published: (2026)
by: Gao, Yuxuan, et al.
Published: (2026)
AttenA+: Rectifying Action Inequality in Robotic Foundation Models
by: Peng, Daojie, et al.
Published: (2026)
by: Peng, Daojie, et al.
Published: (2026)
Towards Forceful Robotic Foundation Models: a Literature Survey
by: Xie, William, et al.
Published: (2025)
by: Xie, William, et al.
Published: (2025)
Constrained Decoding for Safe Robot Navigation Foundation Models
by: Kapoor, Parv, et al.
Published: (2025)
by: Kapoor, Parv, et al.
Published: (2025)
FP3: A 3D Foundation Policy for Robotic Manipulation
by: Yang, Rujia, et al.
Published: (2025)
by: Yang, Rujia, et al.
Published: (2025)
Embodied Robot Manipulation in the Era of Foundation Models: Planning and Learning Perspectives
by: Bai, Shuanghao, et al.
Published: (2025)
by: Bai, Shuanghao, et al.
Published: (2025)
CoIN: A Benchmark of Continual Instruction tuNing for Multimodel Large Language Model
by: Chen, Cheng, et al.
Published: (2024)
by: Chen, Cheng, et al.
Published: (2024)
Cosmos-Surg-dVRK: World Foundation Model-based Automated Online Evaluation of Surgical Robot Policy Learning
by: Zbinden, Lukas, et al.
Published: (2025)
by: Zbinden, Lukas, et al.
Published: (2025)
BEACON: Cross-Domain Co-Training of Generative Robot Policies via Best-Effort Adaptation
by: Zhang, Antong, et al.
Published: (2026)
by: Zhang, Antong, et al.
Published: (2026)
Vision-Language Foundation Models as Effective Robot Imitators
by: Li, Xinghang, et al.
Published: (2023)
by: Li, Xinghang, et al.
Published: (2023)
Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning
by: Wang, Yixiao, et al.
Published: (2024)
by: Wang, Yixiao, et al.
Published: (2024)
SeMv-3D: Towards Concurrency of Semantic and Multi-view Consistency in General Text-to-3D Generation
by: Cai, Xiao, et al.
Published: (2024)
by: Cai, Xiao, et al.
Published: (2024)
ALF: Adaptive Label Finetuning for Scene Graph Generation
by: Chen, Qishen, et al.
Published: (2023)
by: Chen, Qishen, et al.
Published: (2023)
RISE: Self-Improving Robot Policy with Compositional World Model
by: Yang, Jiazhi, et al.
Published: (2026)
by: Yang, Jiazhi, et al.
Published: (2026)
Transferring Foundation Models for Generalizable Robotic Manipulation
by: Yang, Jiange, et al.
Published: (2023)
by: Yang, Jiange, et al.
Published: (2023)
A Robotic Skill Learning System Built Upon Diffusion Policies and Foundation Models
by: Ingelhag, Nils, et al.
Published: (2024)
by: Ingelhag, Nils, et al.
Published: (2024)
Are Foundation Models the Route to Full-Stack Transfer in Robotics?
by: Stulp, Freek, et al.
Published: (2026)
by: Stulp, Freek, et al.
Published: (2026)
Can Tabular Foundation Models Guide Exploration in Robot Policy Learning?
by: Ou, Buqing, et al.
Published: (2026)
by: Ou, Buqing, et al.
Published: (2026)
Similar Items
-
InSpire: Vision-Language-Action Models with Intrinsic Spatial Reasoning
by: Zhang, Ji, et al.
Published: (2025) -
Shortcut Learning in Generalist Robot Policies: The Role of Dataset Diversity and Fragmentation
by: Xing, Youguang, et al.
Published: (2025) -
Beyond the Majority: Long-tail Imitation Learning for Robotic Manipulation
by: Zhu, Junhong, et al.
Published: (2026) -
Sim-and-Human Co-training for Data-Efficient and Generalizable Robotic Manipulation
by: Fang, Kaipeng, et al.
Published: (2026) -
DePT: Decoupled Prompt Tuning
by: Zhang, Ji, et al.
Published: (2023)