FASTER: Rethinking Real-Time Flow VLAs
Fuente:
arXiv
Salvato in:
| Autori principali: | Lu, Yuxiang, Liu, Zhe, Fan, Xianzhe, Yang, Zhenya, Hou, Jinghua, Li, Junyi, Ding, Kaixin, Zhao, Hengshuang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DreamAvoid: Critical-Phase Test-Time Dreaming to Avoid Failures in VLA Policies
di: Fan, Xianzhe, et al.
Pubblicazione: (2026)
di: Fan, Xianzhe, et al.
Pubblicazione: (2026)
Any3D-VLA: Enhancing VLA Robustness via Diverse Point Clouds
di: Fan, Xianzhe, et al.
Pubblicazione: (2026)
di: Fan, Xianzhe, et al.
Pubblicazione: (2026)
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs
di: Fang, Yu, et al.
Pubblicazione: (2026)
di: Fang, Yu, et al.
Pubblicazione: (2026)
$π$-StepNFT: Wider Space Needs Finer Steps in Online RL for Flow-based VLAs
di: Wang, Siting, et al.
Pubblicazione: (2026)
di: Wang, Siting, et al.
Pubblicazione: (2026)
ABPolicy: Asynchronous B-Spline Flow Policy for Real-Time and Smooth Robotic Manipulation
di: Yang, Fan, et al.
Pubblicazione: (2026)
di: Yang, Fan, et al.
Pubblicazione: (2026)
Realtime-VLA FLASH: Speculative Inference Framework for Diffusion-based VLAs
di: Niu, Jiahui, et al.
Pubblicazione: (2026)
di: Niu, Jiahui, et al.
Pubblicazione: (2026)
GenieDrive: Towards Physics-Aware Driving World Model with 4D Occupancy Guided Video Generation
di: Yang, Zhenya, et al.
Pubblicazione: (2025)
di: Yang, Zhenya, et al.
Pubblicazione: (2025)
EndoFlow-SLAM: Real-Time Endoscopic SLAM with Flow-Constrained Gaussian Splatting
di: Wu, Taoyu, et al.
Pubblicazione: (2025)
di: Wu, Taoyu, et al.
Pubblicazione: (2025)
RC-NF: Robot-Conditioned Normalizing Flow for Real-Time Anomaly Detection in Robotic Manipulation
di: Zhou, Shijie, et al.
Pubblicazione: (2026)
di: Zhou, Shijie, et al.
Pubblicazione: (2026)
InsMapper: Exploring Inner-instance Information for Vectorized HD Mapping
di: Xu, Zhenhua, et al.
Pubblicazione: (2023)
di: Xu, Zhenhua, et al.
Pubblicazione: (2023)
UniLION: Towards Unified Autonomous Driving Model with Linear Group RNNs
di: Liu, Zhe, et al.
Pubblicazione: (2025)
di: Liu, Zhe, et al.
Pubblicazione: (2025)
Real-Time Metric-Semantic Mapping for Autonomous Navigation in Outdoor Environments
di: Jiao, Jianhao, et al.
Pubblicazione: (2024)
di: Jiao, Jianhao, et al.
Pubblicazione: (2024)
EmbodiedSAM: Online Segment Any 3D Thing in Real Time
di: Xu, Xiuwei, et al.
Pubblicazione: (2024)
di: Xu, Xiuwei, et al.
Pubblicazione: (2024)
VTAM: Video-Tactile-Action Models for Complex Physical Interaction Beyond VLAs
di: Yuan, Haoran, et al.
Pubblicazione: (2026)
di: Yuan, Haoran, et al.
Pubblicazione: (2026)
Dense Monocular Motion Segmentation Using Optical Flow and Pseudo Depth Map: A Zero-Shot Approach
di: Huang, Yuxiang, et al.
Pubblicazione: (2024)
di: Huang, Yuxiang, et al.
Pubblicazione: (2024)
PhysReaction: Physically Plausible Real-Time Humanoid Reaction Synthesis via Forward Dynamics Guided 4D Imitation
di: Liu, Yunze, et al.
Pubblicazione: (2024)
di: Liu, Yunze, et al.
Pubblicazione: (2024)
SimEndoGS: Efficient Data-driven Scene Simulation using Robotic Surgery Videos via Physics-embedded 3D Gaussians
di: Yang, Zhenya, et al.
Pubblicazione: (2024)
di: Yang, Zhenya, et al.
Pubblicazione: (2024)
FastViDAR: Real-Time Omnidirectional Depth Estimation via Alternative Hierarchical Attention
di: Zhao, Hangtian, et al.
Pubblicazione: (2025)
di: Zhao, Hangtian, et al.
Pubblicazione: (2025)
LION: Linear Group RNN for 3D Object Detection in Point Clouds
di: Liu, Zhe, et al.
Pubblicazione: (2024)
di: Liu, Zhe, et al.
Pubblicazione: (2024)
Fast-dVLA: Accelerating Discrete Diffusion VLA to Real-Time Performance
di: Song, Wenxuan, et al.
Pubblicazione: (2026)
di: Song, Wenxuan, et al.
Pubblicazione: (2026)
A-MFST: Adaptive Multi-Flow Sparse Tracker for Real-Time Tissue Tracking Under Occlusion
di: Chen, Yuxin, et al.
Pubblicazione: (2024)
di: Chen, Yuxin, et al.
Pubblicazione: (2024)
AsyncMDE: Real-Time Monocular Depth Estimation via Asynchronous Spatial Memory
di: Ma, Lianjie, et al.
Pubblicazione: (2026)
di: Ma, Lianjie, et al.
Pubblicazione: (2026)
Towards Real-Time Generation of Delay-Compensated Video Feeds for Outdoor Mobile Robot Teleoperation
di: Chakraborty, Neeloy, et al.
Pubblicazione: (2024)
di: Chakraborty, Neeloy, et al.
Pubblicazione: (2024)
mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs
di: Pai, Jonas, et al.
Pubblicazione: (2025)
di: Pai, Jonas, et al.
Pubblicazione: (2025)
UAV-Flow Colosseo: A Real-World Benchmark for Flying-on-a-Word UAV Imitation Learning
di: Wang, Xiangyu, et al.
Pubblicazione: (2025)
di: Wang, Xiangyu, et al.
Pubblicazione: (2025)
HUGSIM: A Real-Time, Photo-Realistic and Closed-Loop Simulator for Autonomous Driving
di: Zhou, Hongyu, et al.
Pubblicazione: (2024)
di: Zhou, Hongyu, et al.
Pubblicazione: (2024)
DriveGPT4: Interpretable End-to-end Autonomous Driving via Large Language Model
di: Xu, Zhenhua, et al.
Pubblicazione: (2023)
di: Xu, Zhenhua, et al.
Pubblicazione: (2023)
DVMNet++: Rethinking Relative Pose Estimation for Unseen Objects
di: Zhao, Chen, et al.
Pubblicazione: (2024)
di: Zhao, Chen, et al.
Pubblicazione: (2024)
ReMemNav: A Rethinking and Memory-Augmented Framework for Zero-Shot Object Navigation
di: Wu, Feng, et al.
Pubblicazione: (2026)
di: Wu, Feng, et al.
Pubblicazione: (2026)
Real-time Motion Segmentation with Event-based Normal Flow
di: Zhong, Sheng, et al.
Pubblicazione: (2026)
di: Zhong, Sheng, et al.
Pubblicazione: (2026)
R3DP: Real-Time 3D-Aware Policy for Embodied Manipulation
di: Zhang, Yuhao, et al.
Pubblicazione: (2026)
di: Zhang, Yuhao, et al.
Pubblicazione: (2026)
Chatting about Conditional Trajectory Prediction
di: Zhao, Yuxiang, et al.
Pubblicazione: (2026)
di: Zhao, Yuxiang, et al.
Pubblicazione: (2026)
Learning Camera Movement Control from Real-World Drone Videos
di: Hou, Yunzhong, et al.
Pubblicazione: (2024)
di: Hou, Yunzhong, et al.
Pubblicazione: (2024)
Real-Time Loop Closure Detection in Visual SLAM via NetVLAD and Faiss
di: Fan, Enguang
Pubblicazione: (2026)
di: Fan, Enguang
Pubblicazione: (2026)
Bunny-VisionPro: Real-Time Bimanual Dexterous Teleoperation for Imitation Learning
di: Ding, Runyu, et al.
Pubblicazione: (2024)
di: Ding, Runyu, et al.
Pubblicazione: (2024)
RealD$^2$iff: Bridging Real-World Gap in Robot Manipulation via Depth Diffusion
di: Liang, Xiujian, et al.
Pubblicazione: (2025)
di: Liang, Xiujian, et al.
Pubblicazione: (2025)
SaLF: Sparse Local Fields for Multi-Sensor Rendering in Real-Time
di: Chen, Yun, et al.
Pubblicazione: (2025)
di: Chen, Yun, et al.
Pubblicazione: (2025)
RoboGSim: A Real2Sim2Real Robotic Gaussian Splatting Simulator
di: Li, Xinhai, et al.
Pubblicazione: (2024)
di: Li, Xinhai, et al.
Pubblicazione: (2024)
HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction
di: Shi, Zhonghao, et al.
Pubblicazione: (2025)
di: Shi, Zhonghao, et al.
Pubblicazione: (2025)
GRS-SLAM3R: Real-Time Dense SLAM with Gated Recurrent State
di: Shen, Guole, et al.
Pubblicazione: (2025)
di: Shen, Guole, et al.
Pubblicazione: (2025)
Documenti analoghi
-
DreamAvoid: Critical-Phase Test-Time Dreaming to Avoid Failures in VLA Policies
di: Fan, Xianzhe, et al.
Pubblicazione: (2026) -
Any3D-VLA: Enhancing VLA Robustness via Diverse Point Clouds
di: Fan, Xianzhe, et al.
Pubblicazione: (2026) -
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs
di: Fang, Yu, et al.
Pubblicazione: (2026) -
$π$-StepNFT: Wider Space Needs Finer Steps in Online RL for Flow-based VLAs
di: Wang, Siting, et al.
Pubblicazione: (2026) -
ABPolicy: Asynchronous B-Spline Flow Policy for Real-Time and Smooth Robotic Manipulation
di: Yang, Fan, et al.
Pubblicazione: (2026)