WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shang, Yu, Li, Zhuohang, Ma, Yiding, Su, Weikang, Jin, Xin, Wang, Ziyou, Jin, Lei, Zhang, Xin, Tang, Yinzhou, Su, Haisheng, Gao, Chen, Wu, Wei, Liu, Xihui, Shah, Dhruv, Zhang, Zhaoxiang, Chen, Zhibo, Zhu, Jun, Tian, Yonghong, Chua, Tat-Seng, Zhu, Wenwu, Li, Yong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform
von: Shang, Yu, et al.
Veröffentlicht: (2026)
von: Shang, Yu, et al.
Veröffentlicht: (2026)
RoboScape: Physics-informed Embodied World Model
von: Shang, Yu, et al.
Veröffentlicht: (2025)
von: Shang, Yu, et al.
Veröffentlicht: (2025)
MoWM: Mixture-of-World-Models for Embodied Planning via Latent-to-Pixel Feature Modulation
von: Yu, Yangcheng, et al.
Veröffentlicht: (2025)
von: Yu, Yangcheng, et al.
Veröffentlicht: (2025)
Embodied AI: From LLMs to World Models
von: Feng, Tongtong, et al.
Veröffentlicht: (2025)
von: Feng, Tongtong, et al.
Veröffentlicht: (2025)
Compose Your Aesthetics: Empowering Text-to-Image Models with the Principles of Art
von: Jin, Zhe, et al.
Veröffentlicht: (2025)
von: Jin, Zhe, et al.
Veröffentlicht: (2025)
LongScape: Advancing Long-Horizon Embodied World Models with Context-Aware MoE
von: Shang, Yu, et al.
Veröffentlicht: (2025)
von: Shang, Yu, et al.
Veröffentlicht: (2025)
Self-evolving Embodied AI
von: Feng, Tongtong, et al.
Veröffentlicht: (2026)
von: Feng, Tongtong, et al.
Veröffentlicht: (2026)
Embodied-R: Collaborative Framework for Activating Embodied Spatial Reasoning in Foundation Models via Reinforcement Learning
von: Zhao, Baining, et al.
Veröffentlicht: (2025)
von: Zhao, Baining, et al.
Veröffentlicht: (2025)
Embody4D: A Generalist 4D World Model for Embodied AI
von: Tu, Peiyan, et al.
Veröffentlicht: (2026)
von: Tu, Peiyan, et al.
Veröffentlicht: (2026)
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
von: Zhu, Hanxin, et al.
Veröffentlicht: (2026)
von: Zhu, Hanxin, et al.
Veröffentlicht: (2026)
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
World Guidance: World Modeling in Condition Space for Action Generation
von: Su, Yue, et al.
Veröffentlicht: (2026)
von: Su, Yue, et al.
Veröffentlicht: (2026)
Ask-before-Plan: Proactive Language Agents for Real-World Planning
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
von: Chen, Hongjin, et al.
Veröffentlicht: (2026)
VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
von: Sun, Jingwen, et al.
Veröffentlicht: (2026)
U2UData+: A Scalable Swarm UAVs Autonomous Flight Dataset for Embodied Long-horizon Tasks
von: Feng, Tongtong, et al.
Veröffentlicht: (2025)
von: Feng, Tongtong, et al.
Veröffentlicht: (2025)
EgoSim: Egocentric World Simulator for Embodied Interaction Generation
von: Hao, Jinkun, et al.
Veröffentlicht: (2026)
von: Hao, Jinkun, et al.
Veröffentlicht: (2026)
World Reasoning Arena
von: PAN Team, et al.
Veröffentlicht: (2026)
von: PAN Team, et al.
Veröffentlicht: (2026)
Towards High-Consistency Embodied World Model with Multi-View Trajectory Videos
von: Su, Taiyi, et al.
Veröffentlicht: (2025)
von: Su, Taiyi, et al.
Veröffentlicht: (2025)
STYLE: Improving Domain Transferability of Asking Clarification Questions in Large Language Model Powered Conversational Agents
von: Chen, Yue, et al.
Veröffentlicht: (2024)
von: Chen, Yue, et al.
Veröffentlicht: (2024)
Understanding Long Videos via LLM-Powered Entity Relation Graphs
von: Chu, Meng, et al.
Veröffentlicht: (2025)
von: Chu, Meng, et al.
Veröffentlicht: (2025)
Universal Scene Graph Generation
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
Learning to Ask Critical Questions for Assisting Product Search
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
3D Magic Mirror: Clothing Reconstruction from a Single Image via a Causal Perspective
von: Zheng, Zhedong, et al.
Veröffentlicht: (2022)
von: Zheng, Zhedong, et al.
Veröffentlicht: (2022)
AEGIS: Exploring the Limit of World Knowledge Capabilities for Unified Mulitmodal Models
von: Lin, Jintao, et al.
Veröffentlicht: (2026)
von: Lin, Jintao, et al.
Veröffentlicht: (2026)
Embodied World Models Emerge from Navigational Task in Open-Ended Environments
von: Jin, Li, et al.
Veröffentlicht: (2025)
von: Jin, Li, et al.
Veröffentlicht: (2025)
On Generative Agents in Recommendation
von: Zhang, An, et al.
Veröffentlicht: (2023)
von: Zhang, An, et al.
Veröffentlicht: (2023)
Risky-Bench: Probing Agentic Safety Risks under Real-World Deployment
von: Zheng, Jingnan, et al.
Veröffentlicht: (2026)
von: Zheng, Jingnan, et al.
Veröffentlicht: (2026)
WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction
von: Liu, Chengzhi, et al.
Veröffentlicht: (2026)
von: Liu, Chengzhi, et al.
Veröffentlicht: (2026)
Agentic AIs Are the Missing Paradigm for Out-of-Distribution Generalization in Foundation Models
von: Wang, Xin, et al.
Veröffentlicht: (2026)
von: Wang, Xin, et al.
Veröffentlicht: (2026)
Out-of-Distribution Generalization in Graph Foundation Models
von: Li, Haoyang, et al.
Veröffentlicht: (2026)
von: Li, Haoyang, et al.
Veröffentlicht: (2026)
EmbodiedGen: Towards a Generative 3D World Engine for Embodied Intelligence
von: Wang, Xinjie, et al.
Veröffentlicht: (2025)
von: Wang, Xinjie, et al.
Veröffentlicht: (2025)
A Comprehensive Survey on World Models for Embodied AI
von: Li, Xinqing, et al.
Veröffentlicht: (2025)
von: Li, Xinqing, et al.
Veröffentlicht: (2025)
Structure-preserving Feature Alignment for Old Photo Colorization
von: Pang, Yingxue, et al.
Veröffentlicht: (2025)
von: Pang, Yingxue, et al.
Veröffentlicht: (2025)
RynnEC: Bringing MLLMs into Embodied World
von: Dang, Ronghao, et al.
Veröffentlicht: (2025)
von: Dang, Ronghao, et al.
Veröffentlicht: (2025)
Reasoning in Space via Grounding in the World
von: Chen, Yiming, et al.
Veröffentlicht: (2025)
von: Chen, Yiming, et al.
Veröffentlicht: (2025)
Towards Goal-oriented Intelligent Tutoring Systems in Online Education
von: Deng, Yang, et al.
Veröffentlicht: (2023)
von: Deng, Yang, et al.
Veröffentlicht: (2023)
Disentangling Masked Autoencoders for Unsupervised Domain Generalization
von: Zhang, An, et al.
Veröffentlicht: (2024)
von: Zhang, An, et al.
Veröffentlicht: (2024)
BiTAgent: A Task-Aware Modular Framework for Bidirectional Coupling between Multimodal Large Language Models and World Models
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2025)
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2025)
GigaWorld-0: World Models as Data Engine to Empower Embodied AI
von: GigaWorld Team, et al.
Veröffentlicht: (2025)
von: GigaWorld Team, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform
von: Shang, Yu, et al.
Veröffentlicht: (2026) -
RoboScape: Physics-informed Embodied World Model
von: Shang, Yu, et al.
Veröffentlicht: (2025) -
MoWM: Mixture-of-World-Models for Embodied Planning via Latent-to-Pixel Feature Modulation
von: Yu, Yangcheng, et al.
Veröffentlicht: (2025) -
Embodied AI: From LLMs to World Models
von: Feng, Tongtong, et al.
Veröffentlicht: (2025) -
Compose Your Aesthetics: Empowering Text-to-Image Models with the Principles of Art
von: Jin, Zhe, et al.
Veröffentlicht: (2025)