CapsDT: Diffusion-Transformer for Capsule Robot Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | He, Xiting, Su, Mingwu, Jiang, Xinqi, Bai, Long, Lai, Jiewen, Ren, Hongliang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Contact-Aided Navigation of Flexible Robotic Endoscope Using Deep Reinforcement Learning in Dynamic Stomach
by: Ng, Chi Kit, et al.
Published: (2025)
by: Ng, Chi Kit, et al.
Published: (2025)
Tenma: Robust Cross-Embodiment Robot Manipulation with Diffusion Transformer
by: Davies, Travis, et al.
Published: (2025)
by: Davies, Travis, et al.
Published: (2025)
ForeDiffusion: Foresight-Conditioned Diffusion Policy via Future View Construction for Robot Manipulation
by: Xie, Weize, et al.
Published: (2026)
by: Xie, Weize, et al.
Published: (2026)
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction
by: Jiang, Shuo, et al.
Published: (2025)
by: Jiang, Shuo, et al.
Published: (2025)
Adaptive Diffusion Policy Optimization for Robotic Manipulation
by: Jiang, Huiyun, et al.
Published: (2025)
by: Jiang, Huiyun, et al.
Published: (2025)
Never-Ending Behavior-Cloning Agent for Robotic Manipulation
by: Liang, Wenqi, et al.
Published: (2024)
by: Liang, Wenqi, et al.
Published: (2024)
AC-DiT: Adaptive Coordination Diffusion Transformer for Mobile Manipulation
by: Chen, Sixiang, et al.
Published: (2025)
by: Chen, Sixiang, et al.
Published: (2025)
On-Device Diffusion Transformer Policy for Efficient Robot Manipulation
by: Wu, Yiming, et al.
Published: (2025)
by: Wu, Yiming, et al.
Published: (2025)
AnchorDP3: 3D Affordance Guided Sparse Diffusion Policy for Robotic Manipulation
by: Zhao, Ziyan, et al.
Published: (2025)
by: Zhao, Ziyan, et al.
Published: (2025)
Asynchronous Fast-Slow Vision-Language-Action Policies for Whole-Body Robotic Manipulation
by: Zou, Teqiang, et al.
Published: (2025)
by: Zou, Teqiang, et al.
Published: (2025)
World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
by: Jiang, Zhennan, et al.
Published: (2025)
by: Jiang, Zhennan, et al.
Published: (2025)
PC-Diffuser: Path-Consistent Capsule CBF Safety Filtering for Diffusion-Based Trajectory Planner
by: Ku, Eugene, et al.
Published: (2026)
by: Ku, Eugene, et al.
Published: (2026)
TMR-VLA:Vision-Language-Action Model for Magnetic Motion Control of Tri-leg Silicone-based Soft Robot
by: Tang, Ruijie, et al.
Published: (2026)
by: Tang, Ruijie, et al.
Published: (2026)
V$^2$-SfMLearner: Learning Monocular Depth and Ego-motion for Multimodal Wireless Capsule Endoscopy
by: Bai, Long, et al.
Published: (2024)
by: Bai, Long, et al.
Published: (2024)
SurgSora: Object-Aware Diffusion Model for Controllable Surgical Video Generation
by: Chen, Tong, et al.
Published: (2024)
by: Chen, Tong, et al.
Published: (2024)
ProgressVLA: Progress-Guided Diffusion Policy for Vision-Language Robotic Manipulation
by: Yan, Hongyu, et al.
Published: (2026)
by: Yan, Hongyu, et al.
Published: (2026)
DECO: Decoupled Multimodal Diffusion Transformer for Bimanual Dexterous Manipulation with a Plugin Tactile Adapter
by: Li, Xukun, et al.
Published: (2026)
by: Li, Xukun, et al.
Published: (2026)
A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI
by: Wong, Lik Hang Kenny, et al.
Published: (2025)
by: Wong, Lik Hang Kenny, et al.
Published: (2025)
Human-in-the-loop Online Rejection Sampling for Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2025)
by: Lu, Guanxing, et al.
Published: (2025)
Jacobian Exploratory Dual-Phase Reinforcement Learning for Dynamic Endoluminal Navigation of Deformable Continuum Robots
by: Tian, Yu, et al.
Published: (2025)
by: Tian, Yu, et al.
Published: (2025)
Learning Structural Latent Points for Efficient Visual Representations in Robotic Manipulation
by: Jiang, Yicheng, et al.
Published: (2026)
by: Jiang, Yicheng, et al.
Published: (2026)
S$^2$-Diffusion: Generalizing from Instance-level to Category-level Skills in Robot Manipulation
by: Yang, Quantao, et al.
Published: (2025)
by: Yang, Quantao, et al.
Published: (2025)
SDP: Spiking Diffusion Policy for Robotic Manipulation with Learnable Channel-Wise Membrane Thresholds
by: Hou, Zhixing, et al.
Published: (2024)
by: Hou, Zhixing, et al.
Published: (2024)
Escaping the Diversity Trap in Robotic Manipulation via Anchor-Centric Adaptation
by: Chen, Yanzhe, et al.
Published: (2026)
by: Chen, Yanzhe, et al.
Published: (2026)
Toward Generalist Neural Motion Planners for Robotic Manipulators: Challenges and Opportunities
by: Soleymanzadeh, Davood, et al.
Published: (2026)
by: Soleymanzadeh, Davood, et al.
Published: (2026)
Map Imagination Like Blind Humans: Group Diffusion Model for Robotic Map Generation
by: Song, Qijin, et al.
Published: (2024)
by: Song, Qijin, et al.
Published: (2024)
GSAM: A Generalizable and Safe Robotic Framework for Articulated Object Manipulation
by: Shao, Beichen, et al.
Published: (2026)
by: Shao, Beichen, et al.
Published: (2026)
OSSAR: Towards Open-Set Surgical Activity Recognition in Robot-assisted Surgery
by: Bai, Long, et al.
Published: (2024)
by: Bai, Long, et al.
Published: (2024)
ManiCM: Real-time 3D Diffusion Policy via Consistency Model for Robotic Manipulation
by: Lu, Guanxing, et al.
Published: (2024)
by: Lu, Guanxing, et al.
Published: (2024)
AdaWorldPolicy: World-Model-Driven Diffusion Policy with Online Adaptive Learning for Robotic Manipulation
by: Yuan, Ge, et al.
Published: (2026)
by: Yuan, Ge, et al.
Published: (2026)
DexterCap: An Affordable and Automated System for Capturing Dexterous Hand-Object Manipulation
by: Liang, Yutong, et al.
Published: (2026)
by: Liang, Yutong, et al.
Published: (2026)
EndoVLA: Dual-Phase Vision-Language-Action Model for Autonomous Tracking in Endoscopy
by: Ng, Chi Kit, et al.
Published: (2025)
by: Ng, Chi Kit, et al.
Published: (2025)
Object-Focus Actor for Data-efficient Robot Generalization Dexterous Manipulation
by: Li, Yihang, et al.
Published: (2025)
by: Li, Yihang, et al.
Published: (2025)
CroSTAta: Cross-State Transition Attention Transformer for Robotic Manipulation
by: Minelli, Giovanni, et al.
Published: (2025)
by: Minelli, Giovanni, et al.
Published: (2025)
Movement Primitive Diffusion: Learning Gentle Robotic Manipulation of Deformable Objects
by: Scheikl, Paul Maria, et al.
Published: (2023)
by: Scheikl, Paul Maria, et al.
Published: (2023)
Dynamics-Guided Diffusion Model for Sensor-less Robot Manipulator Design
by: Xu, Xiaomeng, et al.
Published: (2024)
by: Xu, Xiaomeng, et al.
Published: (2024)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
by: Lu, Guanxing, et al.
Published: (2025)
by: Lu, Guanxing, et al.
Published: (2025)
OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation
by: Chen, Xinzhe, et al.
Published: (2026)
by: Chen, Xinzhe, et al.
Published: (2026)
Transferring Foundation Models for Generalizable Robotic Manipulation
by: Yang, Jiange, et al.
Published: (2023)
by: Yang, Jiange, et al.
Published: (2023)
Affordance-based Robot Manipulation with Flow Matching
by: Zhang, Fan, et al.
Published: (2024)
by: Zhang, Fan, et al.
Published: (2024)
Similar Items
-
Contact-Aided Navigation of Flexible Robotic Endoscope Using Deep Reinforcement Learning in Dynamic Stomach
by: Ng, Chi Kit, et al.
Published: (2025) -
Tenma: Robust Cross-Embodiment Robot Manipulation with Diffusion Transformer
by: Davies, Travis, et al.
Published: (2025) -
ForeDiffusion: Foresight-Conditioned Diffusion Policy via Future View Construction for Robot Manipulation
by: Xie, Weize, et al.
Published: (2026) -
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction
by: Jiang, Shuo, et al.
Published: (2025) -
Adaptive Diffusion Policy Optimization for Robotic Manipulation
by: Jiang, Huiyun, et al.
Published: (2025)