Where Bits Matter in World Model Planning: A Paired Mixed-Bit Study for Efficient Spatial Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ranganath, Suraj, Patnaik, Anish, Menon, Vaishak |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KV Cache Quantization for Self-Forcing Video Generation: A 33-Method Empirical Study
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026)
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026)
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024)
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024)
Sparse Imagination for Efficient Visual World Model Planning
von: Chun, Junha, et al.
Veröffentlicht: (2025)
von: Chun, Junha, et al.
Veröffentlicht: (2025)
MindJourney: Test-Time Scaling with World Models for Spatial Reasoning
von: Yang, Yuncong, et al.
Veröffentlicht: (2025)
von: Yang, Yuncong, et al.
Veröffentlicht: (2025)
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023)
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023)
Fast-ThinkAct: Efficient Vision-Language-Action Reasoning via Verbalizable Latent Planning
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2026)
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2026)
ExoPredicator: Learning Abstract Models of Dynamic Worlds for Robot Planning
von: Liang, Yichao, et al.
Veröffentlicht: (2025)
von: Liang, Yichao, et al.
Veröffentlicht: (2025)
Diffusion Models as Optimizers for Efficient Planning in Offline RL
von: Huang, Renming, et al.
Veröffentlicht: (2024)
von: Huang, Renming, et al.
Veröffentlicht: (2024)
VisualPredicator: Learning Abstract World Models with Neuro-Symbolic Predicates for Robot Planning
von: Liang, Yichao, et al.
Veröffentlicht: (2024)
von: Liang, Yichao, et al.
Veröffentlicht: (2024)
SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning
von: Liu, Yuecheng, et al.
Veröffentlicht: (2025)
von: Liu, Yuecheng, et al.
Veröffentlicht: (2025)
Knot So Simple: A Minimalistic Environment for Spatial Reasoning
von: Chen, Zizhao, et al.
Veröffentlicht: (2025)
von: Chen, Zizhao, et al.
Veröffentlicht: (2025)
Reason--Imagine--Act: Closed-Loop LLM Decision Making with World Models for Autonomous Driving
von: Sun, Zhengqi, et al.
Veröffentlicht: (2026)
von: Sun, Zhengqi, et al.
Veröffentlicht: (2026)
Embodied Spatial Intelligence: from Implicit Scene Modeling to Spatial Reasoning
von: Fang, Jiading
Veröffentlicht: (2025)
von: Fang, Jiading
Veröffentlicht: (2025)
Out of Sight, Still in Mind: Reasoning and Planning about Unobserved Objects with Video Tracking Enabled Memory Models
von: Huang, Yixuan, et al.
Veröffentlicht: (2023)
von: Huang, Yixuan, et al.
Veröffentlicht: (2023)
Spatial Policy: Guiding Visuomotor Robotic Manipulation with Spatial-Aware Modeling and Reasoning
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
von: Liu, Yijun, et al.
Veröffentlicht: (2025)
Composing Pre-Trained Object-Centric Representations for Robotics From "What" and "Where" Foundation Models
von: Shi, Junyao, et al.
Veröffentlicht: (2024)
von: Shi, Junyao, et al.
Veröffentlicht: (2024)
ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2025)
von: Huang, Chi-Pin, et al.
Veröffentlicht: (2025)
Navigation World Models
von: Bar, Amir, et al.
Veröffentlicht: (2024)
von: Bar, Amir, et al.
Veröffentlicht: (2024)
Where are we in the search for an Artificial Visual Cortex for Embodied Intelligence?
von: Majumdar, Arjun, et al.
Veröffentlicht: (2023)
von: Majumdar, Arjun, et al.
Veröffentlicht: (2023)
From Perception to Action: Spatial AI Agents and World Models
von: Felicia, Gloria, et al.
Veröffentlicht: (2026)
von: Felicia, Gloria, et al.
Veröffentlicht: (2026)
AdaWorld: Learning Adaptable World Models with Latent Actions
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
From Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priors
von: Zhang, Zhengshen, et al.
Veröffentlicht: (2025)
von: Zhang, Zhengshen, et al.
Veröffentlicht: (2025)
Robix: A Unified Model for Robot Interaction, Reasoning and Planning
von: Fang, Huang, et al.
Veröffentlicht: (2025)
von: Fang, Huang, et al.
Veröffentlicht: (2025)
Point Cloud Compression with Bits-back Coding
von: Hieu, Nguyen Quang, et al.
Veröffentlicht: (2024)
von: Hieu, Nguyen Quang, et al.
Veröffentlicht: (2024)
Embodied Tree of Thoughts: Deliberate Manipulation Planning with Embodied World Model
von: Xu, Wenjiang, et al.
Veröffentlicht: (2025)
von: Xu, Wenjiang, et al.
Veröffentlicht: (2025)
AutoWorld: Scaling Multi-Agent Traffic Simulation with Self-Supervised World Models
von: Pourkeshavatz, Mozhgan, et al.
Veröffentlicht: (2026)
von: Pourkeshavatz, Mozhgan, et al.
Veröffentlicht: (2026)
Focus On What Matters: Separated Models For Visual-Based RL Generalization
von: Zhang, Di, et al.
Veröffentlicht: (2024)
von: Zhang, Di, et al.
Veröffentlicht: (2024)
MoDem-V2: Visuo-Motor World Models for Real-World Robot Manipulation
von: Lancaster, Patrick, et al.
Veröffentlicht: (2023)
von: Lancaster, Patrick, et al.
Veröffentlicht: (2023)
Aether: Geometric-Aware Unified World Modeling
von: Aether Team, et al.
Veröffentlicht: (2025)
von: Aether Team, et al.
Veröffentlicht: (2025)
RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics
von: Zhou, Enshen, et al.
Veröffentlicht: (2025)
von: Zhou, Enshen, et al.
Veröffentlicht: (2025)
Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
Pixel-level Scene Understanding in One Token: Visual States Need What-is-Where Composition
von: Lee, Seokmin, et al.
Veröffentlicht: (2026)
von: Lee, Seokmin, et al.
Veröffentlicht: (2026)
OK-Robot: What Really Matters in Integrating Open-Knowledge Models for Robotics
von: Liu, Peiqi, et al.
Veröffentlicht: (2024)
von: Liu, Peiqi, et al.
Veröffentlicht: (2024)
Cosmos World Foundation Model Platform for Physical AI
von: NVIDIA, et al.
Veröffentlicht: (2025)
von: NVIDIA, et al.
Veröffentlicht: (2025)
World Simulation with Video Foundation Models for Physical AI
von: NVIDIA, et al.
Veröffentlicht: (2025)
von: NVIDIA, et al.
Veröffentlicht: (2025)
When Bits Break Recourse: Counterfactual-Faithful Quantization
von: Yahyati, Chaymae, et al.
Veröffentlicht: (2026)
von: Yahyati, Chaymae, et al.
Veröffentlicht: (2026)
Towards Unified World Models for Visual Navigation via Memory-Augmented Planning and Foresight
von: Dong, Yifei, et al.
Veröffentlicht: (2025)
von: Dong, Yifei, et al.
Veröffentlicht: (2025)
GWM: Towards Scalable Gaussian World Models for Robotic Manipulation
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
Real-World Robot Applications of Foundation Models: A Review
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2024)
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2024)
Towards Ambiguity-Free Spatial Foundation Model: Rethinking and Decoupling Depth Ambiguity
von: Xu, Xiaohao, et al.
Veröffentlicht: (2025)
von: Xu, Xiaohao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
KV Cache Quantization for Self-Forcing Video Generation: A 33-Method Empirical Study
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026) -
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
von: Bhatt, Neel P., et al.
Veröffentlicht: (2024) -
Sparse Imagination for Efficient Visual World Model Planning
von: Chun, Junha, et al.
Veröffentlicht: (2025) -
MindJourney: Test-Time Scaling with World Models for Spatial Reasoning
von: Yang, Yuncong, et al.
Veröffentlicht: (2025) -
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023)