Bi-VLA: Bilateral Control-Based Imitation Learning via Vision-Language Fusion for Action Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Kobayashi, Masato, Buamanee, Thanpimon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bi-LAT: Bilateral Control-Based Imitation Learning via Natural Language and Action Chunking with Transformers
by: Kobayashi, Takumi, et al.
Published: (2025)
by: Kobayashi, Takumi, et al.
Published: (2025)
Bi-ACT: Bilateral Control-Based Imitation Learning via Action Chunking with Transformer
by: Buamanee, Thanpimon, et al.
Published: (2024)
by: Buamanee, Thanpimon, et al.
Published: (2024)
Bi-HIL: Bilateral Control-Based Multimodal Hierarchical Imitation Learning via Subtask-Level Progress Rate and Keyframe Memory for Long-Horizon Contact-Rich Robotic Manipulation
by: Buamanee, Thanpimon, et al.
Published: (2026)
by: Buamanee, Thanpimon, et al.
Published: (2026)
DABI: Evaluation of Data Augmentation Methods Using Downsampling in Bilateral Control-Based Imitation Learning with Images
by: Kobayashi, Masato, et al.
Published: (2024)
by: Kobayashi, Masato, et al.
Published: (2024)
ALPHA-$α$ and Bi-ACT Are All You Need: Importance of Position and Force Information/Control for Imitation Learning of Unimanual and Bimanual Robotic Manipulation with Low-Cost System
by: Kobayashi, Masato, et al.
Published: (2024)
by: Kobayashi, Masato, et al.
Published: (2024)
ILBiT: Imitation Learning for Robot Using Position and Torque Information based on Bilateral Control with Transformer
by: Kobayashi, Masato, et al.
Published: (2024)
by: Kobayashi, Masato, et al.
Published: (2024)
Bi-AQUA: Bilateral Control-Based Imitation Learning for Underwater Robot Arms via Lighting-Aware Action Chunking with Transformers
by: Tsunoori, Takeru, et al.
Published: (2025)
by: Tsunoori, Takeru, et al.
Published: (2025)
CRL-VLA: Continual Vision-Language-Action Learning
by: Zeng, Qixin, et al.
Published: (2026)
by: Zeng, Qixin, et al.
Published: (2026)
OpenVLA: An Open-Source Vision-Language-Action Model
by: Kim, Moo Jin, et al.
Published: (2024)
by: Kim, Moo Jin, et al.
Published: (2024)
Tactile-VLA: Unlocking Vision-Language-Action Model's Physical Knowledge for Tactile Generalization
by: Huang, Jialei, et al.
Published: (2025)
by: Huang, Jialei, et al.
Published: (2025)
Error-Feedback Model for Output Correction in Bilateral Control-Based Imitation Learning
by: Sato, Hiroshi, et al.
Published: (2024)
by: Sato, Hiroshi, et al.
Published: (2024)
OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation
by: Hirose, Noriaki, et al.
Published: (2025)
by: Hirose, Noriaki, et al.
Published: (2025)
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
by: Shukor, Mustafa, et al.
Published: (2025)
by: Shukor, Mustafa, et al.
Published: (2025)
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
by: Xiong, Zheng, et al.
Published: (2025)
by: Xiong, Zheng, et al.
Published: (2025)
Imitation Learning with Limited Actions via Diffusion Planners and Deep Koopman Controllers
by: Bi, Jianxin, et al.
Published: (2024)
by: Bi, Jianxin, et al.
Published: (2024)
VLA-Touch: Enhancing Vision-Language-Action Models with Dual-Level Tactile Feedback
by: Bi, Jianxin, et al.
Published: (2025)
by: Bi, Jianxin, et al.
Published: (2025)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
by: Zhang, Hongyin, et al.
Published: (2025)
by: Zhang, Hongyin, et al.
Published: (2025)
Action-Constrained Imitation Learning
by: Yeh, Chia-Han, et al.
Published: (2025)
by: Yeh, Chia-Han, et al.
Published: (2025)
DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization
by: Lin, Sixu, et al.
Published: (2026)
by: Lin, Sixu, et al.
Published: (2026)
DyQ-VLA: Temporal-Dynamic-Aware Quantization for Embodied Vision-Language-Action Models
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models
by: Jiang, Yuhua, et al.
Published: (2025)
by: Jiang, Yuhua, et al.
Published: (2025)
Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models
by: Jin, Ruofan, et al.
Published: (2026)
by: Jin, Ruofan, et al.
Published: (2026)
TTF-VLA: Temporal Token Fusion via Pixel-Attention Integration for Vision-Language-Action Models
by: Liu, Chenghao, et al.
Published: (2025)
by: Liu, Chenghao, et al.
Published: (2025)
TwinVLA: Data-Efficient Bimanual Manipulation with Twin Single-Arm Vision-Language-Action Models
by: Im, Hokyun, et al.
Published: (2025)
by: Im, Hokyun, et al.
Published: (2025)
ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model
by: Zhou, Zhongyi, et al.
Published: (2025)
by: Zhou, Zhongyi, et al.
Published: (2025)
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
by: Yin, Cheng, et al.
Published: (2025)
by: Yin, Cheng, et al.
Published: (2025)
CubeDAgger: Interactive Imitation Learning for Dynamic Systems with Efficient yet Low-risk Interaction
by: Kobayashi, Taisuke
Published: (2025)
by: Kobayashi, Taisuke
Published: (2025)
VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching
by: Xu, Siyu, et al.
Published: (2025)
by: Xu, Siyu, et al.
Published: (2025)
Primary-Fine Decoupling for Action Generation in Robotic Imitation
by: Lei, Xiaohan, et al.
Published: (2026)
by: Lei, Xiaohan, et al.
Published: (2026)
Robotic Imitation of Human Actions
by: Spisak, Josua, et al.
Published: (2024)
by: Spisak, Josua, et al.
Published: (2024)
Diffusion-Based Imitation Learning for Social Pose Generation
by: Martin-Ozimek, Antonio Lech, et al.
Published: (2025)
by: Martin-Ozimek, Antonio Lech, et al.
Published: (2025)
$π_0$: A Vision-Language-Action Flow Model for General Robot Control
by: Black, Kevin, et al.
Published: (2024)
by: Black, Kevin, et al.
Published: (2024)
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
by: Liang, Zhixuan, et al.
Published: (2025)
by: Liang, Zhixuan, et al.
Published: (2025)
PointVLA: Injecting the 3D World into Vision-Language-Action Models
by: Li, Chengmeng, et al.
Published: (2025)
by: Li, Chengmeng, et al.
Published: (2025)
AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention
by: Xiao, Lei, et al.
Published: (2025)
by: Xiao, Lei, et al.
Published: (2025)
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
by: Sautenkov, Oleg, et al.
Published: (2025)
by: Sautenkov, Oleg, et al.
Published: (2025)
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
by: Yang, Ruihan, et al.
Published: (2025)
by: Yang, Ruihan, et al.
Published: (2025)
BEAST: Efficient Tokenization of B-Splines Encoded Action Sequences for Imitation Learning
by: Zhou, Hongyi, et al.
Published: (2025)
by: Zhou, Hongyi, et al.
Published: (2025)
Few-Shot Vision-Language Action-Incremental Policy Learning
by: Song, Mingchen, et al.
Published: (2025)
by: Song, Mingchen, et al.
Published: (2025)
Tactile Modality Fusion for Vision-Language-Action Models
by: Morissette, Charlotte, et al.
Published: (2026)
by: Morissette, Charlotte, et al.
Published: (2026)
Similar Items
-
Bi-LAT: Bilateral Control-Based Imitation Learning via Natural Language and Action Chunking with Transformers
by: Kobayashi, Takumi, et al.
Published: (2025) -
Bi-ACT: Bilateral Control-Based Imitation Learning via Action Chunking with Transformer
by: Buamanee, Thanpimon, et al.
Published: (2024) -
Bi-HIL: Bilateral Control-Based Multimodal Hierarchical Imitation Learning via Subtask-Level Progress Rate and Keyframe Memory for Long-Horizon Contact-Rich Robotic Manipulation
by: Buamanee, Thanpimon, et al.
Published: (2026) -
DABI: Evaluation of Data Augmentation Methods Using Downsampling in Bilateral Control-Based Imitation Learning with Images
by: Kobayashi, Masato, et al.
Published: (2024) -
ALPHA-$α$ and Bi-ACT Are All You Need: Importance of Position and Force Information/Control for Imitation Learning of Unimanual and Bimanual Robotic Manipulation with Low-Cost System
by: Kobayashi, Masato, et al.
Published: (2024)