Mirage2Matter: A Physically Grounded Gaussian World Model from Video
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Zhengqing, Li, Ziwen, Wang, Xin, Huang, Jiaxin, Ren, Zhenyang, Shao, Mingkai, Zhang, Hanlue, Huang, Tianyu, Cheng, Yongkang, Guo, Yandong, Lin, Runqi, Wang, Yuanyuan, Liu, Tongliang, Zhang, Kun, Gong, Mingming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HUGE-Bench: A Benchmark for High-Level UAV Vision-Language-Action Tasks
von: Guo, Jingyu, et al.
Veröffentlicht: (2026)
von: Guo, Jingyu, et al.
Veröffentlicht: (2026)
PosA-VLA: Enhancing Action Generation via Pose-Conditioned Anchor Attention
von: Li, Ziwen, et al.
Veröffentlicht: (2025)
von: Li, Ziwen, et al.
Veröffentlicht: (2025)
MLLM-For3D: Adapting Multimodal Large Language Model for 3D Reasoning Segmentation
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
UrbanGS: Semantic-Guided Gaussian Splatting for Urban Scene Reconstruction
von: Li, Ziwen, et al.
Veröffentlicht: (2024)
von: Li, Ziwen, et al.
Veröffentlicht: (2024)
KineVLA: Towards Kinematics-Aware Vision-Language-Action Models with Bi-Level Action Decomposition
von: Han, Gaoge, et al.
Veröffentlicht: (2026)
von: Han, Gaoge, et al.
Veröffentlicht: (2026)
SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025)
LightHarmony3D: Harmonizing Illumination and Shadows for Object Insertion in 3D Gaussian Splatting
von: Huang, Tianyu, et al.
Veröffentlicht: (2026)
von: Huang, Tianyu, et al.
Veröffentlicht: (2026)
Training-Free Robust Interactive Video Object Segmentation
von: Wei, Xiaoli, et al.
Veröffentlicht: (2024)
von: Wei, Xiaoli, et al.
Veröffentlicht: (2024)
OpenInsGaussian: Open-vocabulary Instance Gaussian Segmentation with Context-aware Cross-view Fusion
von: Huang, Tianyu, et al.
Veröffentlicht: (2025)
von: Huang, Tianyu, et al.
Veröffentlicht: (2025)
ProtoGS: Efficient and High-Quality Rendering with 3D Gaussian Prototypes
von: Gao, Zhengqing, et al.
Veröffentlicht: (2025)
von: Gao, Zhengqing, et al.
Veröffentlicht: (2025)
Physics from Video: Identifiability of Time-Invariant Second-Order ODEs under Minimal Trajectory Conditions
von: Wang, Yuanyuan, et al.
Veröffentlicht: (2026)
von: Wang, Yuanyuan, et al.
Veröffentlicht: (2026)
Identifiability and Asymptotics in Learning Homogeneous Linear ODE Systems from Discrete Observations
von: Wang, Yuanyuan, et al.
Veröffentlicht: (2022)
von: Wang, Yuanyuan, et al.
Veröffentlicht: (2022)
Intellectual Property Protection for 3D Gaussian Splatting Assets: A Survey
von: Zhao, Longjie, et al.
Veröffentlicht: (2026)
von: Zhao, Longjie, et al.
Veröffentlicht: (2026)
Mobile-VTON: High-Fidelity On-Device Virtual Try-On
von: Wan, Zhenchen, et al.
Veröffentlicht: (2026)
von: Wan, Zhenchen, et al.
Veröffentlicht: (2026)
RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing
von: Zhao, Longjie, et al.
Veröffentlicht: (2025)
von: Zhao, Longjie, et al.
Veröffentlicht: (2025)
AdLift: Lifting Adversarial Perturbations to Safeguard 3D Gaussian Splatting Assets Against Instruction-Driven Editing
von: Hong, Ziming, et al.
Veröffentlicht: (2025)
von: Hong, Ziming, et al.
Veröffentlicht: (2025)
Optimal Kernel Choice for Score Function-based Causal Discovery
von: Wang, Wenjie, et al.
Veröffentlicht: (2024)
von: Wang, Wenjie, et al.
Veröffentlicht: (2024)
Discovering and Reasoning of Causality in the Hidden World with Large Language Models
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
Open-Vocabulary Segmentation with Unpaired Mask-Text Supervision
von: Wang, Zhaoqing, et al.
Veröffentlicht: (2024)
von: Wang, Zhaoqing, et al.
Veröffentlicht: (2024)
PanoSLAM: Panoptic 3D Scene Reconstruction via Gaussian SLAM
von: Chen, Runnan, et al.
Veröffentlicht: (2024)
von: Chen, Runnan, et al.
Veröffentlicht: (2024)
Eliminating Catastrophic Overfitting Via Abnormal Adversarial Examples Regularization
von: Lin, Runqi, et al.
Veröffentlicht: (2024)
von: Lin, Runqi, et al.
Veröffentlicht: (2024)
Intern-GS: Vision Model Guided Sparse-View 3D Gaussian Splatting
von: Sun, Xiangyu, et al.
Veröffentlicht: (2025)
von: Sun, Xiangyu, et al.
Veröffentlicht: (2025)
ContactGaussian-WM: Learning Physics-Grounded World Model from Videos
von: Wang, Meizhong, et al.
Veröffentlicht: (2026)
von: Wang, Meizhong, et al.
Veröffentlicht: (2026)
Identifiability Analysis of Linear ODE Systems with Hidden Confounders
von: Wang, Yuanyuan, et al.
Veröffentlicht: (2024)
von: Wang, Yuanyuan, et al.
Veröffentlicht: (2024)
Generator Identification for Linear SDEs with Additive and Multiplicative Noise
von: Wang, Yuanyuan, et al.
Veröffentlicht: (2023)
von: Wang, Yuanyuan, et al.
Veröffentlicht: (2023)
OneWorld: Taming Scene Generation with 3D Unified Representation Autoencoder
von: Gao, Sensen, et al.
Veröffentlicht: (2026)
von: Gao, Sensen, et al.
Veröffentlicht: (2026)
DIDiffGes: Decoupled Semi-Implicit Diffusion Models for Real-time Gesture Generation from Speech
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
On the Over-Memorization During Natural, Robust and Catastrophic Overfitting
von: Lin, Runqi, et al.
Veröffentlicht: (2023)
von: Lin, Runqi, et al.
Veröffentlicht: (2023)
Hierarchical Collaborative Fusion for 3D Instance-aware Referring Expression Segmentation
von: Zhou, Keshen, et al.
Veröffentlicht: (2026)
von: Zhou, Keshen, et al.
Veröffentlicht: (2026)
Local Causal Discovery with Linear non-Gaussian Cyclic Models
von: Dai, Haoyue, et al.
Veröffentlicht: (2024)
von: Dai, Haoyue, et al.
Veröffentlicht: (2024)
An Interpretable AI Framework to Disentangle Self-Interacting and Cold Dark Matter in Galaxy Clusters: The CKAN Approach
von: Huang, Zhenyang, et al.
Veröffentlicht: (2025)
von: Huang, Zhenyang, et al.
Veröffentlicht: (2025)
Projection Pursuit Density Ratio Estimation
von: Wang, Meilin, et al.
Veröffentlicht: (2025)
von: Wang, Meilin, et al.
Veröffentlicht: (2025)
Unveiling Causal Reasoning in Large Language Models: Reality or Mirage?
von: Chi, Haoang, et al.
Veröffentlicht: (2025)
von: Chi, Haoang, et al.
Veröffentlicht: (2025)
On the Recoverability of Causal Relations from Temporally Aggregated I.I.D. Data
von: Fan, Shunxing, et al.
Veröffentlicht: (2024)
von: Fan, Shunxing, et al.
Veröffentlicht: (2024)
From 2D Alignment to 3D Plausibility: Unifying Heterogeneous 2D Priors and Penetration-Free Diffusion for Occlusion-Robust Two-Hand Reconstruction
von: Han, Gaoge, et al.
Veröffentlicht: (2025)
von: Han, Gaoge, et al.
Veröffentlicht: (2025)
OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies
von: Chen, Runnan, et al.
Veröffentlicht: (2024)
von: Chen, Runnan, et al.
Veröffentlicht: (2024)
TED-VITON: Transformer-Empowered Diffusion Models for Virtual Try-On
von: Wan, Zhenchen, et al.
Veröffentlicht: (2024)
von: Wan, Zhenchen, et al.
Veröffentlicht: (2024)
Enhanced Velocity Field Modeling for Gaussian Video Reconstruction
von: Li, Zhenyang, et al.
Veröffentlicht: (2025)
von: Li, Zhenyang, et al.
Veröffentlicht: (2025)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
von: Zhang, Haoze, et al.
Veröffentlicht: (2025)
von: Zhang, Haoze, et al.
Veröffentlicht: (2025)
Intersection theory, relative cohomology and the Feynman parametrization
von: Lu, Mingming, et al.
Veröffentlicht: (2024)
von: Lu, Mingming, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
HUGE-Bench: A Benchmark for High-Level UAV Vision-Language-Action Tasks
von: Guo, Jingyu, et al.
Veröffentlicht: (2026) -
PosA-VLA: Enhancing Action Generation via Pose-Conditioned Anchor Attention
von: Li, Ziwen, et al.
Veröffentlicht: (2025) -
MLLM-For3D: Adapting Multimodal Large Language Model for 3D Reasoning Segmentation
von: Huang, Jiaxin, et al.
Veröffentlicht: (2025) -
UrbanGS: Semantic-Guided Gaussian Splatting for Urban Scene Reconstruction
von: Li, Ziwen, et al.
Veröffentlicht: (2024) -
KineVLA: Towards Kinematics-Aware Vision-Language-Action Models with Bi-Level Action Decomposition
von: Han, Gaoge, et al.
Veröffentlicht: (2026)