SpaceMind: A Modular and Self-Evolving Embodied Vision-Language Agent Framework for Autonomous On-orbit Servicing
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Aodi, Han, Haodong, Luo, Xubo, Wang, Ruisuo, He, Shan, Wan, Xue |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
StepNav: Structured Trajectory Priors for Efficient and Multimodal Visual Navigation
by: Luo, Xubo, et al.
Published: (2026)
by: Luo, Xubo, et al.
Published: (2026)
SpaceSense-Bench: A Large-Scale Multi-Modal Benchmark for Spacecraft Perception and Pose Estimation
by: Wu, Aodi, et al.
Published: (2026)
by: Wu, Aodi, et al.
Published: (2026)
Enhancing Vision-Language Models for Autonomous Driving through Task-Specific Prompting and Spatial Reasoning
by: Wu, Aodi, et al.
Published: (2025)
by: Wu, Aodi, et al.
Published: (2025)
Building Cooperative Embodied Agents Modularly with Large Language Models
by: Zhang, Hongxin, et al.
Published: (2023)
by: Zhang, Hongxin, et al.
Published: (2023)
JointLoc: A Real-time Visual Localization Framework for Planetary UAVs Based on Joint Relative and Absolute Pose Estimation
by: Luo, Xubo, et al.
Published: (2024)
by: Luo, Xubo, et al.
Published: (2024)
Embodied Arena: A Comprehensive, Unified, and Evolving Evaluation Platform for Embodied AI
by: Ni, Fei, et al.
Published: (2025)
by: Ni, Fei, et al.
Published: (2025)
Robo-Cortex: A Self-Evolving Embodied Agent via Dual-Grain Cognitive Memory and Autonomous Knowledge Induction
by: Chan, Nga Teng, et al.
Published: (2026)
by: Chan, Nga Teng, et al.
Published: (2026)
Endowing Embodied Agents with Spatial Reasoning Capabilities for Vision-and-Language Navigation
by: Bai, Qianqian, et al.
Published: (2025)
by: Bai, Qianqian, et al.
Published: (2025)
Agentic Robot: A Brain-Inspired Framework for Vision-Language-Action Models in Embodied Agents
by: Yang, Zhejian, et al.
Published: (2025)
by: Yang, Zhejian, et al.
Published: (2025)
SpaceMind: Camera-Guided Modality Fusion for Spatial Reasoning in Vision-Language Models
by: Zhao, Ruosen, et al.
Published: (2025)
by: Zhao, Ruosen, et al.
Published: (2025)
Embodied Hazard Mitigation using Vision-Language Models for Autonomous Mobile Robots
by: Sotomi, Oluwadamilola, et al.
Published: (2025)
by: Sotomi, Oluwadamilola, et al.
Published: (2025)
ManiTaskGen: A Comprehensive Task Generator for Benchmarking and Improving Vision-Language Agents on Embodied Decision-Making
by: Dai, Liu, et al.
Published: (2025)
by: Dai, Liu, et al.
Published: (2025)
Seed2Scale: A Self-Evolving Data Engine for Embodied AI via Small to Large Model Synergy and Multimodal Evaluation
by: Tai, Cong, et al.
Published: (2026)
by: Tai, Cong, et al.
Published: (2026)
Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models
by: Xu, Haiweng, et al.
Published: (2026)
by: Xu, Haiweng, et al.
Published: (2026)
RealMirror: A Comprehensive, Open-Source Vision-Language-Action Platform for Embodied AI
by: Tai, Cong, et al.
Published: (2025)
by: Tai, Cong, et al.
Published: (2025)
LEO-RobotAgent: A General-purpose Robotic Agent for Language-driven Embodied Operator
by: Chen, Lihuang, et al.
Published: (2025)
by: Chen, Lihuang, et al.
Published: (2025)
A Multi-Agent LLM Framework for Design Space Exploration in Autonomous Driving Systems
by: Shih, Po-An, et al.
Published: (2025)
by: Shih, Po-An, et al.
Published: (2025)
Affordance-Graphed Task Worlds: Self-Evolving Task Generation for Scalable Embodied Learning
by: Liu, Xiang, et al.
Published: (2026)
by: Liu, Xiang, et al.
Published: (2026)
AToM-Bot: Embodied Fulfillment of Unspoken Human Needs with Affective Theory of Mind
by: Ding, Wei, et al.
Published: (2024)
by: Ding, Wei, et al.
Published: (2024)
Learning to Anchor Visual Odometry: KAN-Based Pose Regression for Planetary Landing
by: Luo, Xubo, et al.
Published: (2025)
by: Luo, Xubo, et al.
Published: (2025)
PEPA: a Persistently Autonomous Embodied Agent with Personalities
by: Liu, Kaige, et al.
Published: (2026)
by: Liu, Kaige, et al.
Published: (2026)
Provable Ordering and Continuity in Vision-Language Pretraining for Generalizable Embodied Agents
by: Zhang, Zhizhen, et al.
Published: (2025)
by: Zhang, Zhizhen, et al.
Published: (2025)
Agentic Self-Evolutionary Replanning for Embodied Navigation
by: Li, Guoliang, et al.
Published: (2026)
by: Li, Guoliang, et al.
Published: (2026)
Mind to Hand: Purposeful Robotic Control via Embodied Reasoning
by: Tang, Peijun, et al.
Published: (2025)
by: Tang, Peijun, et al.
Published: (2025)
Modular Fault Diagnosis Framework for Complex Autonomous Driving Systems
by: Orf, Stefan, et al.
Published: (2024)
by: Orf, Stefan, et al.
Published: (2024)
HMR-1: Hierarchical Massage Robot with Vision-Language-Model for Embodied Healthcare
by: Xu, Rongtao, et al.
Published: (2026)
by: Xu, Rongtao, et al.
Published: (2026)
BBSEA: An Exploration of Brain-Body Synchronization for Embodied Agents
by: Yang, Sizhe, et al.
Published: (2024)
by: Yang, Sizhe, et al.
Published: (2024)
LAGEA: Language Guided Embodied Agents for Robotic Manipulation
by: Chowdhury, Abdul Monaf, et al.
Published: (2025)
by: Chowdhury, Abdul Monaf, et al.
Published: (2025)
NaviTrace: Evaluating Embodied Navigation of Vision-Language Models
by: Windecker, Tim, et al.
Published: (2025)
by: Windecker, Tim, et al.
Published: (2025)
Embodied Learning of Reward for Musculoskeletal Control with Vision Language Models
by: Soedarmadji, Saraswati, et al.
Published: (2025)
by: Soedarmadji, Saraswati, et al.
Published: (2025)
Embodied Co-Design for Rapidly Evolving Agents: Taxonomy, Frontiers, and Challenges
by: Wang, Yuxing, et al.
Published: (2025)
by: Wang, Yuxing, et al.
Published: (2025)
SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models
by: Dong, Xiangyu, et al.
Published: (2025)
by: Dong, Xiangyu, et al.
Published: (2025)
Efficient Coordination with the System-Level Shared State: An Embodied-AI Native Modular Framework
by: Deng, Yixuan, et al.
Published: (2026)
by: Deng, Yixuan, et al.
Published: (2026)
Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills
by: Yuan, Haoqi, et al.
Published: (2025)
by: Yuan, Haoqi, et al.
Published: (2025)
RoboOS: A Hierarchical Embodied Framework for Cross-Embodiment and Multi-Agent Collaboration
by: Tan, Huajie, et al.
Published: (2025)
by: Tan, Huajie, et al.
Published: (2025)
PhysiAgent: An Embodied Agent Framework in Physical World
by: Wang, Zhihao, et al.
Published: (2025)
by: Wang, Zhihao, et al.
Published: (2025)
Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control
by: Chen, William, et al.
Published: (2026)
by: Chen, William, et al.
Published: (2026)
VLP: Vision-Language Preference Learning for Embodied Manipulation
by: Liu, Runze, et al.
Published: (2025)
by: Liu, Runze, et al.
Published: (2025)
Survey of Vision-Language-Action Models for Embodied Manipulation
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
FRENETIX: A High-Performance and Modular Motion Planning Framework for Autonomous Driving
by: Trauth, Rainer, et al.
Published: (2024)
by: Trauth, Rainer, et al.
Published: (2024)
Similar Items
-
StepNav: Structured Trajectory Priors for Efficient and Multimodal Visual Navigation
by: Luo, Xubo, et al.
Published: (2026) -
SpaceSense-Bench: A Large-Scale Multi-Modal Benchmark for Spacecraft Perception and Pose Estimation
by: Wu, Aodi, et al.
Published: (2026) -
Enhancing Vision-Language Models for Autonomous Driving through Task-Specific Prompting and Spatial Reasoning
by: Wu, Aodi, et al.
Published: (2025) -
Building Cooperative Embodied Agents Modularly with Large Language Models
by: Zhang, Hongxin, et al.
Published: (2023) -
JointLoc: A Real-time Visual Localization Framework for Planetary UAVs Based on Joint Relative and Absolute Pose Estimation
by: Luo, Xubo, et al.
Published: (2024)