WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform
Fuente:
arXiv
Saved in:
| Main Authors: | Shang, Yu, Tang, Yinzhou, Ma, Yiding, Li, Zhuohang, Jin, Lei, Su, Weikang, Jin, Xin, Wang, Zhaolu, Wang, Ziyou, Zhang, Xin, Su, Haisheng, He, Weizhen, Wu, Wei, Duan, Haoyi, Wetzstein, Gordon, Liu, Xihui, Shah, Dhruv, Zhang, Zhaoxiang, Chen, Zhibo, Zhu, Jun, Tian, Yonghong, Chua, Tat-Seng, Zhu, Wenwu, Gao, Chen, Li, Yong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
by: Shang, Yu, et al.
Published: (2026)
by: Shang, Yu, et al.
Published: (2026)
MoWM: Mixture-of-World-Models for Embodied Planning via Latent-to-Pixel Feature Modulation
by: Yu, Yangcheng, et al.
Published: (2025)
by: Yu, Yangcheng, et al.
Published: (2025)
RoboScape: Physics-informed Embodied World Model
by: Shang, Yu, et al.
Published: (2025)
by: Shang, Yu, et al.
Published: (2025)
Embodied AI: From LLMs to World Models
by: Feng, Tongtong, et al.
Published: (2025)
by: Feng, Tongtong, et al.
Published: (2025)
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
by: Zhu, Hanxin, et al.
Published: (2026)
by: Zhu, Hanxin, et al.
Published: (2026)
World Guidance: World Modeling in Condition Space for Action Generation
by: Su, Yue, et al.
Published: (2026)
by: Su, Yue, et al.
Published: (2026)
Compose Your Aesthetics: Empowering Text-to-Image Models with the Principles of Art
by: Jin, Zhe, et al.
Published: (2025)
by: Jin, Zhe, et al.
Published: (2025)
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
by: Zhao, Baining, et al.
Published: (2026)
by: Zhao, Baining, et al.
Published: (2026)
VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model
by: Sun, Jingwen, et al.
Published: (2026)
by: Sun, Jingwen, et al.
Published: (2026)
LongScape: Advancing Long-Horizon Embodied World Models with Context-Aware MoE
by: Shang, Yu, et al.
Published: (2025)
by: Shang, Yu, et al.
Published: (2025)
AEGIS: Exploring the Limit of World Knowledge Capabilities for Unified Mulitmodal Models
by: Lin, Jintao, et al.
Published: (2026)
by: Lin, Jintao, et al.
Published: (2026)
World Reasoning Arena
by: PAN Team, et al.
Published: (2026)
by: PAN Team, et al.
Published: (2026)
BiTAgent: A Task-Aware Modular Framework for Bidirectional Coupling between Multimodal Large Language Models and World Models
by: Zhan, Yu-Wei, et al.
Published: (2025)
by: Zhan, Yu-Wei, et al.
Published: (2025)
WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models
by: Chen, Hongjin, et al.
Published: (2026)
by: Chen, Hongjin, et al.
Published: (2026)
AV-Unified: A Unified Framework for Audio-visual Scene Understanding
by: Li, Guangyao, et al.
Published: (2026)
by: Li, Guangyao, et al.
Published: (2026)
Self-evolving Embodied AI
by: Feng, Tongtong, et al.
Published: (2026)
by: Feng, Tongtong, et al.
Published: (2026)
Curriculum Graph Machine Learning: A Survey
by: Li, Haoyang, et al.
Published: (2023)
by: Li, Haoyang, et al.
Published: (2023)
Collected environmental change and nitrogen removal data
by: Dong, Liang, et al.
Published: (2025)
by: Dong, Liang, et al.
Published: (2025)
UniMamba: Unified Spatial-Channel Representation Learning with Group-Efficient Mamba for LiDAR-based 3D Object Detection
by: Jin, Xin, et al.
Published: (2025)
by: Jin, Xin, et al.
Published: (2025)
Physics-informed neural networks for unsteady incompressible flows with time-dependent moving boundaries
by: Zhu, Yongzheng, et al.
Published: (2023)
by: Zhu, Yongzheng, et al.
Published: (2023)
WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction
by: Liu, Chengzhi, et al.
Published: (2026)
by: Liu, Chengzhi, et al.
Published: (2026)
EvolvingAgent: Curriculum Self-evolving Agent with Continual World Model for Long-Horizon Tasks
by: Feng, Tongtong, et al.
Published: (2025)
by: Feng, Tongtong, et al.
Published: (2025)
MultiWorld: Scalable Multi-Agent Multi-View Video World Models
by: Wu, Haoyu, et al.
Published: (2026)
by: Wu, Haoyu, et al.
Published: (2026)
Ask-before-Plan: Proactive Language Agents for Real-World Planning
by: Zhang, Xuan, et al.
Published: (2024)
by: Zhang, Xuan, et al.
Published: (2024)
A Recipe for Efficient Sim-to-Real Transfer in Manipulation with Online Imitation-Pretrained World Models
by: Wang, Yilin, et al.
Published: (2025)
by: Wang, Yilin, et al.
Published: (2025)
A Stochastic Hybrid Approach to Decentralized Networked Control: Stochastic Network Delays and Poisson Pulsing Attacks
by: Zhang, Dandan, et al.
Published: (2024)
by: Zhang, Dandan, et al.
Published: (2024)
Inverting the wedge map and Gauss composition
by: Chua, Kok Seng
Published: (2024)
by: Chua, Kok Seng
Published: (2024)
Chebyshev polynomials and a refinement of the local residue/non-residue structure at a prime
by: Chua, Kok Seng
Published: (2026)
by: Chua, Kok Seng
Published: (2026)
PATCHEVAL: A New Benchmark for Evaluating LLMs on Patching Real-World Vulnerabilities
by: Wei, Zichao, et al.
Published: (2025)
by: Wei, Zichao, et al.
Published: (2025)
STT-Arena: A More Realistic Environment for Tool-Using with Spatio-Temporal Dynamics
by: Hui, Tingfeng, et al.
Published: (2026)
by: Hui, Tingfeng, et al.
Published: (2026)
Learning from Uncertain Data: From Possible Worlds to Possible Models
by: Zhu, Jiongli, et al.
Published: (2024)
by: Zhu, Jiongli, et al.
Published: (2024)
World-Coordinate Human Motion Retargeting via SAM 3D Body
by: Tu, Zhangzheng, et al.
Published: (2025)
by: Tu, Zhangzheng, et al.
Published: (2025)
Can LLMs Infer Personality from Real World Conversations?
by: Zhu, Jianfeng, et al.
Published: (2025)
by: Zhu, Jianfeng, et al.
Published: (2025)
Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models
by: Lu, Jiacheng, et al.
Published: (2026)
by: Lu, Jiacheng, et al.
Published: (2026)
Reasoning in Space via Grounding in the World
by: Chen, Yiming, et al.
Published: (2025)
by: Chen, Yiming, et al.
Published: (2025)
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
by: Peng, Baorui, et al.
Published: (2026)
by: Peng, Baorui, et al.
Published: (2026)
Probing compressed Higgsinos at the FASER experiment
by: Su, Shufang, et al.
Published: (2025)
by: Su, Shufang, et al.
Published: (2025)
SmartMLOps Studio: Design of an LLM-Integrated IDE with Automated MLOps Pipelines for Model Development and Monitoring
by: Jin, Jiawei, et al.
Published: (2025)
by: Jin, Jiawei, et al.
Published: (2025)
Agentic AIs Are the Missing Paradigm for Out-of-Distribution Generalization in Foundation Models
by: Wang, Xin, et al.
Published: (2026)
by: Wang, Xin, et al.
Published: (2026)
Automated Graph Machine Learning: Approaches, Libraries, Benchmarks and Directions
by: Wang, Xin, et al.
Published: (2022)
by: Wang, Xin, et al.
Published: (2022)
Similar Items
-
WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
by: Shang, Yu, et al.
Published: (2026) -
MoWM: Mixture-of-World-Models for Embodied Planning via Latent-to-Pixel Feature Modulation
by: Yu, Yangcheng, et al.
Published: (2025) -
RoboScape: Physics-informed Embodied World Model
by: Shang, Yu, et al.
Published: (2025) -
Embodied AI: From LLMs to World Models
by: Feng, Tongtong, et al.
Published: (2025) -
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
by: Zhu, Hanxin, et al.
Published: (2026)