SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Dinging, Zhao, Yingxiu, Cheng, Xinrui, Lin, Kangheng, Peng, Hongbo, Li, Hongxing, Wang, Zixuan, Dai, Yuhong, Li, Haodong, Wang, Jia, Shi, Yukang, Zhao, Liang, Sun, Jianjian, Ge, Zheng, Zhang, Xiangyu, Lu, Weiming, Xiao, Jun, Zhuang, Yueting, Shen, Yongliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SpatialLadder: Progressive Training for Spatial Reasoning in Vision-Language Models
von: Li, Hongxing, et al.
Veröffentlicht: (2025)
von: Li, Hongxing, et al.
Veröffentlicht: (2025)
ViewSpatial-Bench: Evaluating Multi-perspective Spatial Localization in Vision-Language Models
von: Li, Dingming, et al.
Veröffentlicht: (2025)
von: Li, Dingming, et al.
Veröffentlicht: (2025)
Milestone-Guided Policy Learning for Long-Horizon Language Agents
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
GroundAct: Can LLM Agents Ground Actions in Environmental States?
von: Wang, Zixuan, et al.
Veröffentlicht: (2025)
von: Wang, Zixuan, et al.
Veröffentlicht: (2025)
SpatialFusion: Endowing Unified Image Generation with Intrinsic 3D Geometric Awareness
von: Qiu, Haiyi, et al.
Veröffentlicht: (2026)
von: Qiu, Haiyi, et al.
Veröffentlicht: (2026)
Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization
von: Zhang, Wenqi, et al.
Veröffentlicht: (2024)
von: Zhang, Wenqi, et al.
Veröffentlicht: (2024)
Code-A1: Adversarial Evolving of Code LLM and Test LLM via Reinforcement Learning
von: Wang, Aozhe, et al.
Veröffentlicht: (2026)
von: Wang, Aozhe, et al.
Veröffentlicht: (2026)
WebVR: Benchmarking Multimodal LLMs for WebPage Recreation from Videos via Human-Aligned Visual Rubrics
von: Dai, Yuhong, et al.
Veröffentlicht: (2026)
von: Dai, Yuhong, et al.
Veröffentlicht: (2026)
Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts
von: Xu, Haolei, et al.
Veröffentlicht: (2026)
von: Xu, Haolei, et al.
Veröffentlicht: (2026)
EvoEmpirBench: Dynamic Spatial Reasoning with Agent-ExpVer
von: Zhao, Pukun, et al.
Veröffentlicht: (2025)
von: Zhao, Pukun, et al.
Veröffentlicht: (2025)
Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency
von: Liu, Junming, et al.
Veröffentlicht: (2026)
von: Liu, Junming, et al.
Veröffentlicht: (2026)
Data-Copilot: Bridging Billions of Data and Humans with Autonomous Workflow
von: Zhang, Wenqi, et al.
Veröffentlicht: (2023)
von: Zhang, Wenqi, et al.
Veröffentlicht: (2023)
Evo-0: Vision-Language-Action Model with Implicit Spatial Understanding
von: Lin, Tao, et al.
Veröffentlicht: (2025)
von: Lin, Tao, et al.
Veröffentlicht: (2025)
2.5 Years in Class: A Multimodal Textbook for Vision-Language Pretraining
von: Zhang, Wenqi, et al.
Veröffentlicht: (2025)
von: Zhang, Wenqi, et al.
Veröffentlicht: (2025)
Self-Contrast: Better Reflection Through Inconsistent Solving Perspectives
von: Zhang, Wenqi, et al.
Veröffentlicht: (2024)
von: Zhang, Wenqi, et al.
Veröffentlicht: (2024)
EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
EvoTSE: Evolving Enrollment for Target Speaker Extraction
von: Liu, Zikai, et al.
Veröffentlicht: (2026)
von: Liu, Zikai, et al.
Veröffentlicht: (2026)
Spatial-SSRL: Enhancing Spatial Understanding via Self-Supervised Reinforcement Learning
von: Liu, Yuhong, et al.
Veröffentlicht: (2025)
von: Liu, Yuhong, et al.
Veröffentlicht: (2025)
Automatic Instruction Evolving for Large Language Models
von: Zeng, Weihao, et al.
Veröffentlicht: (2024)
von: Zeng, Weihao, et al.
Veröffentlicht: (2024)
EvoWiki: Evaluating LLMs on Evolving Knowledge
von: Tang, Wei, et al.
Veröffentlicht: (2024)
von: Tang, Wei, et al.
Veröffentlicht: (2024)
RAFT-UP: Robust Alignment for Spatial Transcriptomics with Explicit Control of Spatial Distortion
von: Wu, Yaqi, et al.
Veröffentlicht: (2026)
von: Wu, Yaqi, et al.
Veröffentlicht: (2026)
ClawGUI: A Unified Framework for Training, Evaluating, and Deploying GUI Agents
von: Tang, Fei, et al.
Veröffentlicht: (2026)
von: Tang, Fei, et al.
Veröffentlicht: (2026)
VideoRefer Suite: Advancing Spatial-Temporal Object Understanding with Video LLM
von: Yuan, Yuqian, et al.
Veröffentlicht: (2024)
von: Yuan, Yuqian, et al.
Veröffentlicht: (2024)
Slow Perception: Let's Perceive Geometric Figures Step-by-step
von: Wei, Haoran, et al.
Veröffentlicht: (2024)
von: Wei, Haoran, et al.
Veröffentlicht: (2024)
Geometrically-Constrained Agent for Spatial Reasoning
von: Chen, Zeren, et al.
Veröffentlicht: (2025)
von: Chen, Zeren, et al.
Veröffentlicht: (2025)
Neural Network-Assisted RIS Weight Optimization for Spatial Nulling in Distorted Reflector Antenna Systems
von: Li, Xinrui, et al.
Veröffentlicht: (2025)
von: Li, Xinrui, et al.
Veröffentlicht: (2025)
Mixed‐Mode Fracturing Characteristics of Asphalt Concrete at Low‐Temperature Considering Random Spatial Combinations of Aggregates and Voids
von: Mengzhang Chen, et al.
Veröffentlicht: (2025)
von: Mengzhang Chen, et al.
Veröffentlicht: (2025)
TaskBench: Benchmarking Large Language Models for Task Automation
von: Shen, Yongliang, et al.
Veröffentlicht: (2023)
von: Shen, Yongliang, et al.
Veröffentlicht: (2023)
Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning
von: Gao, Xianqiang, et al.
Veröffentlicht: (2026)
von: Gao, Xianqiang, et al.
Veröffentlicht: (2026)
EvoScene-VLA: Evolving Scene Beliefs Inside the Action Decoder for Chunked Robot Control
von: Zhang, Chushan, et al.
Veröffentlicht: (2026)
von: Zhang, Chushan, et al.
Veröffentlicht: (2026)
Spatial Blindness in Whole-Slide Multiple Instance Learning
von: Li, Xiangyu, et al.
Veröffentlicht: (2026)
von: Li, Xiangyu, et al.
Veröffentlicht: (2026)
TraceTrans: Translation and Spatial Tracing for Surgical Prediction
von: Luo, Xiyu, et al.
Veröffentlicht: (2025)
von: Luo, Xiyu, et al.
Veröffentlicht: (2025)
Structural-Temporal Coupling Anomaly Detection with Dynamic Graph Transformer
von: Zong, Chang, et al.
Veröffentlicht: (2025)
von: Zong, Chang, et al.
Veröffentlicht: (2025)
EvoCodeBench: A Human-Performance Benchmark for Self-Evolving LLM-Driven Coding Systems
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
von: Zhang, Wentao, et al.
Veröffentlicht: (2026)
EvoLM: Self-Evolving Language Models through Co-Evolved Discriminative Rubrics
von: Li, Shuyue Stella, et al.
Veröffentlicht: (2026)
von: Li, Shuyue Stella, et al.
Veröffentlicht: (2026)
Reconstructing 4D Spatial Intelligence: A Survey
von: Cao, Yukang, et al.
Veröffentlicht: (2025)
von: Cao, Yukang, et al.
Veröffentlicht: (2025)
Spatial Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model
von: Li, Fuhao, et al.
Veröffentlicht: (2025)
von: Li, Fuhao, et al.
Veröffentlicht: (2025)
Unhackable Temporal Rewarding for Scalable Video MLLMs
von: Yu, En, et al.
Veröffentlicht: (2025)
von: Yu, En, et al.
Veröffentlicht: (2025)
EvoTaxo: Building and Evolving Taxonomy from Social Media Streams
von: Li, Yiyang, et al.
Veröffentlicht: (2026)
von: Li, Yiyang, et al.
Veröffentlicht: (2026)
RIS-Aided Spatial Nulling: Algorithms, Analysis, and Nulling Limits
von: Li, Xinrui, et al.
Veröffentlicht: (2025)
von: Li, Xinrui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SpatialLadder: Progressive Training for Spatial Reasoning in Vision-Language Models
von: Li, Hongxing, et al.
Veröffentlicht: (2025) -
ViewSpatial-Bench: Evaluating Multi-perspective Spatial Localization in Vision-Language Models
von: Li, Dingming, et al.
Veröffentlicht: (2025) -
Milestone-Guided Policy Learning for Long-Horizon Language Agents
von: Wang, Zixuan, et al.
Veröffentlicht: (2026) -
GroundAct: Can LLM Agents Ground Actions in Environmental States?
von: Wang, Zixuan, et al.
Veröffentlicht: (2025) -
SpatialFusion: Endowing Unified Image Generation with Intrinsic 3D Geometric Awareness
von: Qiu, Haiyi, et al.
Veröffentlicht: (2026)