4DWorldBench: A Comprehensive Evaluation Framework for 3D/4D World Generation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Yiting, Luo, Wei, Tu, Peiyan, Li, Haoran, Zhu, Hanxin, Yu, Zihao, Wang, Xingrui, Chen, Xinyi, Peng, Xinge, Li, Xin, Chen, Zhibo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Embody4D: A Generalist 4D World Model for Embodied AI
by: Tu, Peiyan, et al.
Published: (2026)
by: Tu, Peiyan, et al.
Published: (2026)
Sonic4D: Spatial Audio Generation for Immersive 4D Scene Exploration
by: Xie, Siyi, et al.
Published: (2025)
by: Xie, Siyi, et al.
Published: (2025)
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
by: Zhu, Hanxin, et al.
Published: (2026)
by: Zhu, Hanxin, et al.
Published: (2026)
IQA-Spider: Unifying Multi-Granularity Image Quality Assessment with Reasoning, Grounding and Referring
by: Peng, Xinge, et al.
Published: (2026)
by: Peng, Xinge, et al.
Published: (2026)
GSemSplat: Generalizable Semantic 3D Gaussian Splatting from Uncalibrated Image Pairs
by: Wang, Xingrui, et al.
Published: (2024)
by: Wang, Xingrui, et al.
Published: (2024)
P-4DGS: Predictive 4D Gaussian Splatting with 90$\times$ Compression
by: Wang, Henan, et al.
Published: (2025)
by: Wang, Henan, et al.
Published: (2025)
AR4D: Autoregressive 4D Generation from Monocular Videos
by: Zhu, Hanxin, et al.
Published: (2025)
by: Zhu, Hanxin, et al.
Published: (2025)
TIV-Diffusion: Towards Object-Centric Movement for Text-driven Image to Video Generation
by: Wang, Xingrui, et al.
Published: (2024)
by: Wang, Xingrui, et al.
Published: (2024)
CoNo: Consistency Noise Injection for Tuning-free Long Video Diffusion
by: Wang, Xingrui, et al.
Published: (2024)
by: Wang, Xingrui, et al.
Published: (2024)
GaussianSR: 3D Gaussian Super-Resolution with 2D Diffusion Priors
by: Yu, Xiqian, et al.
Published: (2024)
by: Yu, Xiqian, et al.
Published: (2024)
Video Quality Assessment Based on Swin TransformerV2 and Coarse to Fine Strategy
by: Yu, Zihao, et al.
Published: (2024)
by: Yu, Zihao, et al.
Published: (2024)
InternVQA: Advancing Compressed Video Quality Assessment with Distilling Large Foundation Model
by: Guan, Fengbin, et al.
Published: (2025)
by: Guan, Fengbin, et al.
Published: (2025)
QMamba: On First Exploration of Vision Mamba for Image Quality Assessment
by: Guan, Fengbin, et al.
Published: (2024)
by: Guan, Fengbin, et al.
Published: (2024)
SeD: Semantic-Aware Discriminator for Image Super-Resolution
by: Li, Bingchen, et al.
Published: (2024)
by: Li, Bingchen, et al.
Published: (2024)
LossAgent: Towards Any Optimization Objectives for Image Processing with LLM Agents
by: Li, Bingchen, et al.
Published: (2024)
by: Li, Bingchen, et al.
Published: (2024)
Hybrid Agents for Image Restoration
by: Li, Bingchen, et al.
Published: (2025)
by: Li, Bingchen, et al.
Published: (2025)
Is Vanilla MLP in Neural Radiance Field Enough for Few-shot View Synthesis?
by: Zhu, Hanxin, et al.
Published: (2024)
by: Zhu, Hanxin, et al.
Published: (2024)
CMC: Few-shot Novel View Synthesis via Cross-view Multiplane Consistency
by: Zhu, Hanxin, et al.
Published: (2024)
by: Zhu, Hanxin, et al.
Published: (2024)
Light Field Compression Based on Implicit Neural Representation
by: Wang, Henan, et al.
Published: (2024)
by: Wang, Henan, et al.
Published: (2024)
UCIP: A Universal Framework for Compressed Image Super-Resolution using Dynamic Prompt
by: Li, Xin, et al.
Published: (2024)
by: Li, Xin, et al.
Published: (2024)
Diffusion Models for Image Restoration and Enhancement: A Comprehensive Survey
by: Li, Xin, et al.
Published: (2023)
by: Li, Xin, et al.
Published: (2023)
Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising
by: Guo, Jun, et al.
Published: (2026)
by: Guo, Jun, et al.
Published: (2026)
Compositional 3D-aware Video Generation with LLM Director
by: Zhu, Hanxin, et al.
Published: (2024)
by: Zhu, Hanxin, et al.
Published: (2024)
Avatar4D: Synthesizing Domain-Specific 4D Humans for Real-World Pose Estimation
by: Bright, Jerrin, et al.
Published: (2025)
by: Bright, Jerrin, et al.
Published: (2025)
Priorformer: A UGC-VQA Method with content and distortion priors
by: Pei, Yajing, et al.
Published: (2024)
by: Pei, Yajing, et al.
Published: (2024)
LoViF 2026 The First Challenge on Holistic Quality Assessment for 4D World Model (PhyScore)
by: Luo, Wei, et al.
Published: (2026)
by: Luo, Wei, et al.
Published: (2026)
Dynamic Boronate Ester Chemistry Facilitating 3D Printing Interlayer Adhesion and Modular 4D Printing of Polylactic Acid
by: Wenjun Peng, et al.
Published: (2025)
by: Wenjun Peng, et al.
Published: (2025)
OrthoPhys: Physically Plausible Video Generation with Orthogonal-View Geometry Guidance
by: Wang, Cong, et al.
Published: (2026)
by: Wang, Cong, et al.
Published: (2026)
Instant4D: 4D Gaussian Splatting in Minutes
by: Luo, Zhanpeng, et al.
Published: (2025)
by: Luo, Zhanpeng, et al.
Published: (2025)
OmniQuality-R: Advancing Reward Models Through All-Encompassing Quality Assessment
by: Lu, Yiting, et al.
Published: (2025)
by: Lu, Yiting, et al.
Published: (2025)
Q-Adapt: Adapting LMM for Visual Quality Assessment with Progressive Instruction Tuning
by: Lu, Yiting, et al.
Published: (2025)
by: Lu, Yiting, et al.
Published: (2025)
CoCo4D: Comprehensive and Complex 4D Scene Generation
by: Zhou, Junwei, et al.
Published: (2025)
by: Zhou, Junwei, et al.
Published: (2025)
DiffCap-Bench: A Comprehensive, Challenging, Robust Benchmark for Image Difference Captioning
by: Wei, Yuancheng, et al.
Published: (2026)
by: Wei, Yuancheng, et al.
Published: (2026)
End-to-End Rate-Distortion Optimized 3D Gaussian Representation
by: Wang, Henan, et al.
Published: (2024)
by: Wang, Henan, et al.
Published: (2024)
Spatial4D-Bench: A Versatile 4D Spatial Intelligence Benchmark
by: Wang, Pan, et al.
Published: (2025)
by: Wang, Pan, et al.
Published: (2025)
Fusion4CA: Boosting 3D Object Detection via Comprehensive Image Exploitation
by: Luo, Kang, et al.
Published: (2026)
by: Luo, Kang, et al.
Published: (2026)
MemBench: Towards More Comprehensive Evaluation on the Memory of LLM-based Agents
by: Tan, Haoran, et al.
Published: (2025)
by: Tan, Haoran, et al.
Published: (2025)
Style4D-Bench: A Benchmark Suite for 4D Stylization
by: Chen, Beiqi, et al.
Published: (2025)
by: Chen, Beiqi, et al.
Published: (2025)
3D4D: An Interactive, Editable, 4D World Model via 3D Video Generation
by: He, Yunhong, et al.
Published: (2025)
by: He, Yunhong, et al.
Published: (2025)
Bench4Merge: A Comprehensive Benchmark for Merging in Realistic Dense Traffic with Micro-Interactive Vehicles
by: Wang, Zhengming, et al.
Published: (2024)
by: Wang, Zhengming, et al.
Published: (2024)
Similar Items
-
Embody4D: A Generalist 4D World Model for Embodied AI
by: Tu, Peiyan, et al.
Published: (2026) -
Sonic4D: Spatial Audio Generation for Immersive 4D Scene Exploration
by: Xie, Siyi, et al.
Published: (2025) -
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
by: Zhu, Hanxin, et al.
Published: (2026) -
IQA-Spider: Unifying Multi-Granularity Image Quality Assessment with Reasoning, Grounding and Referring
by: Peng, Xinge, et al.
Published: (2026) -
GSemSplat: Generalizable Semantic 3D Gaussian Splatting from Uncalibrated Image Pairs
by: Wang, Xingrui, et al.
Published: (2024)