MASER: Modality-Adaptive Specialist Routing for Embodied 3D Spatial Intelligence
Fuente:
arXiv
Guardado en:
| Autores principales: | Raj, Hilton, AV, Vishnuram |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Embodied Spatial Intelligence: from Implicit Scene Modeling to Spatial Reasoning
por: Fang, Jiading
Publicado: (2025)
por: Fang, Jiading
Publicado: (2025)
BIP3D: Bridging 2D Images and 3D Perception for Embodied Intelligence
por: Lin, Xuewu, et al.
Publicado: (2024)
por: Lin, Xuewu, et al.
Publicado: (2024)
Towards Scalable Spatial Intelligence via 2D-to-3D Data Lifting
por: Miao, Xingyu, et al.
Publicado: (2025)
por: Miao, Xingyu, et al.
Publicado: (2025)
Reinforced Embodied Active Defense: Exploiting Adaptive Interaction for Robust Visual Perception in Adversarial 3D Environments
por: Yang, Xiao, et al.
Publicado: (2025)
por: Yang, Xiao, et al.
Publicado: (2025)
Embodied-R: Collaborative Framework for Activating Embodied Spatial Reasoning in Foundation Models via Reinforcement Learning
por: Zhao, Baining, et al.
Publicado: (2025)
por: Zhao, Baining, et al.
Publicado: (2025)
AmaraSpatial-10K: A Spatially and Semantically Aligned 3D Dataset for Spatial Computing and Embodied AI
por: Salehi, Mohammad Sadegh, et al.
Publicado: (2026)
por: Salehi, Mohammad Sadegh, et al.
Publicado: (2026)
SpatialPoint: Spatial-aware Point Prediction for Embodied Localization
por: Zhu, Qiming, et al.
Publicado: (2026)
por: Zhu, Qiming, et al.
Publicado: (2026)
Federated Cross-Modal Retrieval with Missing Modalities via Semantic Routing and Adapter Personalization
por: Zhou, Hefeng, et al.
Publicado: (2026)
por: Zhou, Hefeng, et al.
Publicado: (2026)
SPA: 3D Spatial-Awareness Enables Effective Embodied Representation
por: Zhu, Haoyi, et al.
Publicado: (2024)
por: Zhu, Haoyi, et al.
Publicado: (2024)
Vision-Language Navigation with Embodied Intelligence: A Survey
por: Gao, Peng, et al.
Publicado: (2024)
por: Gao, Peng, et al.
Publicado: (2024)
Dejavu: Towards Experience Feedback Learning for Embodied Intelligence
por: Wu, Shaokai, et al.
Publicado: (2025)
por: Wu, Shaokai, et al.
Publicado: (2025)
SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence
por: Wu, Haoning, et al.
Publicado: (2025)
por: Wu, Haoning, et al.
Publicado: (2025)
SPATIOROUTE: Dynamic Prompt Routing for Zero-Shot Spatial Reasoning
por: Chunhachatrachai, Pawat, et al.
Publicado: (2026)
por: Chunhachatrachai, Pawat, et al.
Publicado: (2026)
Aerial Vision-Language Navigation with a Unified Framework for Spatial, Temporal and Embodied Reasoning
por: Xu, Huilin, et al.
Publicado: (2025)
por: Xu, Huilin, et al.
Publicado: (2025)
Spatial 3D-LLM: Exploring Spatial Awareness in 3D Vision-Language Models
por: Wang, Xiaoyan, et al.
Publicado: (2025)
por: Wang, Xiaoyan, et al.
Publicado: (2025)
SpatialLLM: From Multi-modality Data to Urban Spatial Intelligence
por: Chen, Jiabin, et al.
Publicado: (2025)
por: Chen, Jiabin, et al.
Publicado: (2025)
g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks
por: Wang, Zihan, et al.
Publicado: (2024)
por: Wang, Zihan, et al.
Publicado: (2024)
MetaSpatial: Reinforcing 3D Spatial Reasoning in VLMs for the Metaverse
por: Pan, Zhenyu, et al.
Publicado: (2025)
por: Pan, Zhenyu, et al.
Publicado: (2025)
ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop
por: Hong, Yining, et al.
Publicado: (2026)
por: Hong, Yining, et al.
Publicado: (2026)
M3D-BFS: a Multi-stage Dynamic Fusion Strategy for Sample-Adaptive Multi-Modal Brain Network Analysis
por: Dong, Rui, et al.
Publicado: (2026)
por: Dong, Rui, et al.
Publicado: (2026)
SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning
por: Liu, Yuecheng, et al.
Publicado: (2025)
por: Liu, Yuecheng, et al.
Publicado: (2025)
SmartSpatial: Enhancing the 3D Spatial Arrangement Capabilities of Stable Diffusion Models and Introducing a Novel 3D Spatial Evaluation Framework
por: Huang, Mao Xun, et al.
Publicado: (2025)
por: Huang, Mao Xun, et al.
Publicado: (2025)
MosaicThinker: On-Device Visual Spatial Reasoning for Embodied AI via Iterative Construction of Space Representation
por: Wang, Haoming, et al.
Publicado: (2026)
por: Wang, Haoming, et al.
Publicado: (2026)
UrbanVideo-Bench: Benchmarking Vision-Language Models on Embodied Intelligence with Video Data in Urban Spaces
por: Zhao, Baining, et al.
Publicado: (2025)
por: Zhao, Baining, et al.
Publicado: (2025)
OmniEVA: Embodied Versatile Planner via Task-Adaptive 3D-Grounded and Embodiment-aware Reasoning
por: Liu, Yuecheng, et al.
Publicado: (2025)
por: Liu, Yuecheng, et al.
Publicado: (2025)
3DLLM-Mem: Long-Term Spatial-Temporal Memory for Embodied 3D Large Language Model
por: Hu, Wenbo, et al.
Publicado: (2025)
por: Hu, Wenbo, et al.
Publicado: (2025)
EmbodiedOcc: Embodied 3D Occupancy Prediction for Vision-based Online Scene Understanding
por: Wu, Yuqi, et al.
Publicado: (2024)
por: Wu, Yuqi, et al.
Publicado: (2024)
G3: An Effective and Adaptive Framework for Worldwide Geolocalization Using Large Multi-Modality Models
por: Jia, Pengyue, et al.
Publicado: (2024)
por: Jia, Pengyue, et al.
Publicado: (2024)
SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images
por: Liu, Zishan, et al.
Publicado: (2026)
por: Liu, Zishan, et al.
Publicado: (2026)
3D-Agent:Tri-Modal Multi-Agent Collaboration for Scalable 3D Object Annotation
por: Zhang, Jusheng, et al.
Publicado: (2026)
por: Zhang, Jusheng, et al.
Publicado: (2026)
Sensor-Adaptive Flood Mapping with Pre-trained Multi-Modal Transformers across SAR and Multispectral Modalities
por: Tanaka, Tomohiro, et al.
Publicado: (2025)
por: Tanaka, Tomohiro, et al.
Publicado: (2025)
MM-Mixing: Multi-Modal Mixing Alignment for 3D Understanding
por: Wang, Jiaze, et al.
Publicado: (2024)
por: Wang, Jiaze, et al.
Publicado: (2024)
VideoRouter: Query-Adaptive Dual Routing for Efficient Long-Video Understanding
por: Lin, Kuanwei, et al.
Publicado: (2026)
por: Lin, Kuanwei, et al.
Publicado: (2026)
SDA-PLANNER: State-Dependency Aware Adaptive Planner for Embodied Task Planning
por: Shen, Zichao, et al.
Publicado: (2025)
por: Shen, Zichao, et al.
Publicado: (2025)
RoomTour3D: Geometry-Aware Video-Instruction Tuning for Embodied Navigation
por: Han, Mingfei, et al.
Publicado: (2024)
por: Han, Mingfei, et al.
Publicado: (2024)
FRAME: Forensic Routing and Adaptive Multi-path Evidence Fusion for Image Manipulation Detection
por: Zhao, Kaixiang, et al.
Publicado: (2026)
por: Zhao, Kaixiang, et al.
Publicado: (2026)
Towards Balanced Multi-Modal Learning in 3D Human Pose Estimation
por: Qi, Mengshi, et al.
Publicado: (2025)
por: Qi, Mengshi, et al.
Publicado: (2025)
Physics-Inspired Modeling and Content Adaptive Routing in an Infrared Gas Leak Detection Network
por: Li, Dongsheng, et al.
Publicado: (2025)
por: Li, Dongsheng, et al.
Publicado: (2025)
EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models
por: Du, Mengfei, et al.
Publicado: (2024)
por: Du, Mengfei, et al.
Publicado: (2024)
VEQ: Modality-Adaptive Quantization for MoE Vision-Language Models
por: Qin, Guangshuo, et al.
Publicado: (2026)
por: Qin, Guangshuo, et al.
Publicado: (2026)
Ejemplares similares
-
Embodied Spatial Intelligence: from Implicit Scene Modeling to Spatial Reasoning
por: Fang, Jiading
Publicado: (2025) -
BIP3D: Bridging 2D Images and 3D Perception for Embodied Intelligence
por: Lin, Xuewu, et al.
Publicado: (2024) -
Towards Scalable Spatial Intelligence via 2D-to-3D Data Lifting
por: Miao, Xingyu, et al.
Publicado: (2025) -
Reinforced Embodied Active Defense: Exploiting Adaptive Interaction for Robust Visual Perception in Adversarial 3D Environments
por: Yang, Xiao, et al.
Publicado: (2025) -
Embodied-R: Collaborative Framework for Activating Embodied Spatial Reasoning in Foundation Models via Reinforcement Learning
por: Zhao, Baining, et al.
Publicado: (2025)