RoboStream: Weaving Spatio-Temporal Reasoning with Memory in Vision-Language Models for Robotics
Fuente:
arXiv
Guardado en:
| Autores principales: | Huang, Yuzhi, Wu, Jie, Bu, Weijue, Xiong, Ziyi, Jiang, Gaoyang, Li, Ye, Ji, Kangye, Xie, Shuzhao, Huang, Yue, Wu, Chenglei, Jiang, Jingyan, Wang, Zhi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Conscious Gaze: Adaptive Attention Mechanisms for Hallucination Mitigation in Vision-Language Models
por: Bu, Weijue, et al.
Publicado: (2025)
por: Bu, Weijue, et al.
Publicado: (2025)
LightCL: Compact Continual Learning with Low Memory Footprint For Edge Device
por: Wang, Zeqing, et al.
Publicado: (2024)
por: Wang, Zeqing, et al.
Publicado: (2024)
ElegantVLA: Learning When to Think for Efficient Vision-Language-Action Models
por: Li, Ye, et al.
Publicado: (2026)
por: Li, Ye, et al.
Publicado: (2026)
Jump-teaching: Combating Sample Selection Bias via Temporal Disagreement
por: Ji, Kangye, et al.
Publicado: (2024)
por: Ji, Kangye, et al.
Publicado: (2024)
Memory Forcing: Spatio-Temporal Memory for Consistent Scene Generation on Minecraft
por: Huang, Junchao, et al.
Publicado: (2025)
por: Huang, Junchao, et al.
Publicado: (2025)
Membox: Weaving Topic Continuity into Long-Range Memory for LLM Agents
por: Tao, Dehao, et al.
Publicado: (2026)
por: Tao, Dehao, et al.
Publicado: (2026)
RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics
por: Zhou, Enshen, et al.
Publicado: (2025)
por: Zhou, Enshen, et al.
Publicado: (2025)
Spatial Policy: Guiding Visuomotor Robotic Manipulation with Spatial-Aware Modeling and Reasoning
por: Liu, Yijun, et al.
Publicado: (2025)
por: Liu, Yijun, et al.
Publicado: (2025)
RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics
por: Zhou, Enshen, et al.
Publicado: (2025)
por: Zhou, Enshen, et al.
Publicado: (2025)
MesonGS++: Post-training Compression of 3D Gaussian Splatting with Hyperparameter Searching
por: Xie, Shuzhao, et al.
Publicado: (2026)
por: Xie, Shuzhao, et al.
Publicado: (2026)
VGGT-DP: Generalizable Robot Control via Vision Foundation Models
por: Ge, Shijia, et al.
Publicado: (2025)
por: Ge, Shijia, et al.
Publicado: (2025)
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
por: Huang, Wenlong, et al.
Publicado: (2024)
por: Huang, Wenlong, et al.
Publicado: (2024)
RoboEXP: Action-Conditioned Scene Graph via Interactive Exploration for Robotic Manipulation
por: Jiang, Hanxiao, et al.
Publicado: (2024)
por: Jiang, Hanxiao, et al.
Publicado: (2024)
MemWeaver: Weaving Hybrid Memories for Traceable Long-Horizon Agentic Reasoning
por: Ye, Juexiang, et al.
Publicado: (2026)
por: Ye, Juexiang, et al.
Publicado: (2026)
EgoThinker: Unveiling Egocentric Reasoning with Spatio-Temporal CoT
por: Pei, Baoqi, et al.
Publicado: (2025)
por: Pei, Baoqi, et al.
Publicado: (2025)
RoboOmni: Proactive Robot Manipulation in Omni-modal Context
por: Wang, Siyin, et al.
Publicado: (2025)
por: Wang, Siyin, et al.
Publicado: (2025)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
por: Huang, Haifeng, et al.
Publicado: (2025)
por: Huang, Haifeng, et al.
Publicado: (2025)
Cognitive Weave: Synthesizing Abstracted Knowledge with a Spatio-Temporal Resonance Graph
por: Vishwakarma, Akash, et al.
Publicado: (2025)
por: Vishwakarma, Akash, et al.
Publicado: (2025)
Spatio-Temporal Self-Supervised Learning for Traffic Flow Prediction
por: Ji, Jiahao, et al.
Publicado: (2022)
por: Ji, Jiahao, et al.
Publicado: (2022)
RoboMamba: Efficient Vision-Language-Action Model for Robotic Reasoning and Manipulation
por: Liu, Jiaming, et al.
Publicado: (2024)
por: Liu, Jiaming, et al.
Publicado: (2024)
RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation
por: Jiang, Feng, et al.
Publicado: (2026)
por: Jiang, Feng, et al.
Publicado: (2026)
ReMEmbR: Building and Reasoning Over Long-Horizon Spatio-Temporal Memory for Robot Navigation
por: Anwar, Abrar, et al.
Publicado: (2024)
por: Anwar, Abrar, et al.
Publicado: (2024)
Geospatial Question Answering on Historical Maps Using Spatio-Temporal Knowledge Graphs and Large Language Models
por: Liu, Ziyi, et al.
Publicado: (2025)
por: Liu, Ziyi, et al.
Publicado: (2025)
RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark
por: Lei, Huashuo, et al.
Publicado: (2026)
por: Lei, Huashuo, et al.
Publicado: (2026)
WeaveTime: Stream from Earlier Frames into Emergent Memory in VideoLLMs
por: Zhang, Yulin, et al.
Publicado: (2026)
por: Zhang, Yulin, et al.
Publicado: (2026)
Earthfarseer: Versatile Spatio-Temporal Dynamical Systems Modeling in One Model
por: Wu, Hao, et al.
Publicado: (2023)
por: Wu, Hao, et al.
Publicado: (2023)
GUI-KV: Efficient GUI Agents via KV Cache with Spatio-Temporal Awareness
por: Huang, Kung-Hsiang, et al.
Publicado: (2025)
por: Huang, Kung-Hsiang, et al.
Publicado: (2025)
Semantic-Space Exploration and Exploitation in RLVR for LLM Reasoning
por: Huang, Fanding, et al.
Publicado: (2025)
por: Huang, Fanding, et al.
Publicado: (2025)
Outside Back Cover: Boosting C─O Bond Cleavage and Reverse Water‐Gas Shift Activity via Enriched in‐plane Sulfur Vacancies in Single‐Layer Molybdenum Disulfide (Angew. Chem. 17/2025)
por: Zhiyuan Zheng, et al.
Publicado: (2025)
por: Zhiyuan Zheng, et al.
Publicado: (2025)
Outside Back Cover: Boosting C─O Bond Cleavage and Reverse Water‐Gas Shift Activity via Enriched in‐plane Sulfur Vacancies in Single‐Layer Molybdenum Disulfide (Angew. Chem. Int. Ed. 17/2025)
por: Zhiyuan Zheng, et al.
Publicado: (2025)
por: Zhiyuan Zheng, et al.
Publicado: (2025)
DST-GTN: Dynamic Spatio-Temporal Graph Transformer Network for Traffic Forecasting
por: Huang, Songtao, et al.
Publicado: (2024)
por: Huang, Songtao, et al.
Publicado: (2024)
Test-Time Distillation for Continual Model Adaptation
por: Chen, Xiao, et al.
Publicado: (2025)
por: Chen, Xiao, et al.
Publicado: (2025)
Spatio-Temporal LLM: Reasoning about Environments and Actions
por: Zheng, Haozhen, et al.
Publicado: (2025)
por: Zheng, Haozhen, et al.
Publicado: (2025)
Mask-RadarNet: Enhancing Transformer With Spatial-Temporal Semantic Context for Radar Object Detection in Autonomous Driving
por: Wu, Yuzhi, et al.
Publicado: (2024)
por: Wu, Yuzhi, et al.
Publicado: (2024)
Modeling, Simulation, and Application of Spatio-Temporal Characteristics Detection in Incipient Slip
por: Li, Mingxuan, et al.
Publicado: (2025)
por: Li, Mingxuan, et al.
Publicado: (2025)
Robo-DM: Data Management For Large Robot Datasets
por: Chen, Kaiyuan, et al.
Publicado: (2025)
por: Chen, Kaiyuan, et al.
Publicado: (2025)
VLA-Pro: Cross-Task Procedural Memory Transfer for Vision-Language-Action Models
por: Si, Shengyu, et al.
Publicado: (2026)
por: Si, Shengyu, et al.
Publicado: (2026)
Detection of a Sparse Change in High-Dimensional Time Series
por: Huang, Jingyan
Publicado: (2025)
por: Huang, Jingyan
Publicado: (2025)
COSMIC: Clique-Oriented Semantic Multi-space Integration for Robust CLIP Test-Time Adaptation
por: Huang, Fanding, et al.
Publicado: (2025)
por: Huang, Fanding, et al.
Publicado: (2025)
RoboPARA: Dual-Arm Robot Planning with Parallel Allocation and Recomposition Across Tasks
por: Duan, Shiying, et al.
Publicado: (2025)
por: Duan, Shiying, et al.
Publicado: (2025)
Ejemplares similares
-
Conscious Gaze: Adaptive Attention Mechanisms for Hallucination Mitigation in Vision-Language Models
por: Bu, Weijue, et al.
Publicado: (2025) -
LightCL: Compact Continual Learning with Low Memory Footprint For Edge Device
por: Wang, Zeqing, et al.
Publicado: (2024) -
ElegantVLA: Learning When to Think for Efficient Vision-Language-Action Models
por: Li, Ye, et al.
Publicado: (2026) -
Jump-teaching: Combating Sample Selection Bias via Temporal Disagreement
por: Ji, Kangye, et al.
Publicado: (2024) -
Memory Forcing: Spatio-Temporal Memory for Consistent Scene Generation on Minecraft
por: Huang, Junchao, et al.
Publicado: (2025)