Gespeichert in:
| Hauptverfasser: | Lillemark, Hansen Jin, Huang, Benhao, Zhan, Fangneng, Du, Yilun, Keller, Thomas Anderson |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.01075 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AAPMT: AGI Assessment Through Prompt and Metric Transformer
von: Huang, Benhao
Veröffentlicht: (2024)
von: Huang, Benhao
Veröffentlicht: (2024)
Flow Equivariant Recurrent Neural Networks
von: Keller, T. Anderson
Veröffentlicht: (2025)
von: Keller, T. Anderson
Veröffentlicht: (2025)
MixLight: Borrowing the Best of both Spherical Harmonics and Gaussian Models
von: Ji, Xinlong, et al.
Veröffentlicht: (2024)
von: Ji, Xinlong, et al.
Veröffentlicht: (2024)
Abstract 3D Perception for Spatial Intelligence in Vision-Language Models
von: Liu, Yifan, et al.
Veröffentlicht: (2025)
von: Liu, Yifan, et al.
Veröffentlicht: (2025)
GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
AdaWorld: Learning Adaptable World Models with Latent Actions
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2025)
AREA3D: Active Reconstruction Agent with Unified Feed-Forward 3D Perception and Vision-Language Guidance
von: Xu, Tianling, et al.
Veröffentlicht: (2025)
von: Xu, Tianling, et al.
Veröffentlicht: (2025)
Defining and Extracting generalizable interaction primitives from DNNs
von: Chen, Lu, et al.
Veröffentlicht: (2024)
von: Chen, Lu, et al.
Veröffentlicht: (2024)
Equilibrium Matching: Generative Modeling with Implicit Energy-Based Models
von: Wang, Runqian, et al.
Veröffentlicht: (2025)
von: Wang, Runqian, et al.
Veröffentlicht: (2025)
HAZARD Challenge: Embodied Decision Making in Dynamically Changing Environments
von: Zhou, Qinhong, et al.
Veröffentlicht: (2024)
von: Zhou, Qinhong, et al.
Veröffentlicht: (2024)
Equivariant Reinforcement Learning under Partial Observability
von: Nguyen, Hai, et al.
Veröffentlicht: (2024)
von: Nguyen, Hai, et al.
Veröffentlicht: (2024)
Long-Text-to-Image Generation via Compositional Prompt Decomposition
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2026)
von: Huang, Jen-Yuan, et al.
Veröffentlicht: (2026)
PPU-Bench:Real World Benchmark for Personalized Partial Unlearning in Vision Language Models
von: Guang, Jiahui, et al.
Veröffentlicht: (2026)
von: Guang, Jiahui, et al.
Veröffentlicht: (2026)
Compositional Generative Modeling: A Single Model is Not All You Need
von: Du, Yilun, et al.
Veröffentlicht: (2024)
von: Du, Yilun, et al.
Veröffentlicht: (2024)
Stream3D: Sequential Multi-View 3D Generation via Evidential Memory
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
Grounding Video Models to Actions through Goal Conditioned Exploration
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
Category-level Neural Field for Reconstruction of Partially Observed Objects in Indoor Environment
von: Lee, Taekbeom, et al.
Veröffentlicht: (2024)
von: Lee, Taekbeom, et al.
Veröffentlicht: (2024)
Equivariant Flow Matching for Point Cloud Assembly
von: Wang, Ziming, et al.
Veröffentlicht: (2025)
von: Wang, Ziming, et al.
Veröffentlicht: (2025)
MindJourney: Test-Time Scaling with World Models for Spatial Reasoning
von: Yang, Yuncong, et al.
Veröffentlicht: (2025)
von: Yang, Yuncong, et al.
Veröffentlicht: (2025)
DiffAge3D: Diffusion-based 3D-aware Face Aging
von: Wahid, Junaid, et al.
Veröffentlicht: (2024)
von: Wahid, Junaid, et al.
Veröffentlicht: (2024)
SOGS: Second-Order Anchor for Advanced 3D Gaussian Splatting
von: Zhang, Jiahui, et al.
Veröffentlicht: (2025)
von: Zhang, Jiahui, et al.
Veröffentlicht: (2025)
Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models
von: Chen, Kaijin, et al.
Veröffentlicht: (2026)
von: Chen, Kaijin, et al.
Veröffentlicht: (2026)
Variational Partial Group Convolutions for Input-Aware Partial Equivariance of Rotations and Color-Shifts
von: Kim, Hyunsu, et al.
Veröffentlicht: (2024)
von: Kim, Hyunsu, et al.
Veröffentlicht: (2024)
SPIE: Semantic and Structural Post-Training of Image Editing Diffusion Models with AI feedback
von: Benarous, Elior, et al.
Veröffentlicht: (2025)
von: Benarous, Elior, et al.
Veröffentlicht: (2025)
Video as the New Language for Real-World Decision Making
von: Yang, Sherry, et al.
Veröffentlicht: (2024)
von: Yang, Sherry, et al.
Veröffentlicht: (2024)
Toward Stable World Models: Measuring and Addressing World Instability in Generative Environments
von: Kwon, Soonwoo, et al.
Veröffentlicht: (2025)
von: Kwon, Soonwoo, et al.
Veröffentlicht: (2025)
General Neural Gauge Fields
von: Zhan, Fangneng, et al.
Veröffentlicht: (2023)
von: Zhan, Fangneng, et al.
Veröffentlicht: (2023)
seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models
von: Ghaemi, Hafez, et al.
Veröffentlicht: (2025)
von: Ghaemi, Hafez, et al.
Veröffentlicht: (2025)
Large-scale Reinforcement Learning for Diffusion Models
von: Zhang, Yinan, et al.
Veröffentlicht: (2024)
von: Zhang, Yinan, et al.
Veröffentlicht: (2024)
Visual Acoustic Fields
von: Li, Yuelei, et al.
Veröffentlicht: (2025)
von: Li, Yuelei, et al.
Veröffentlicht: (2025)
Dynamic Orchestration of Multi-Agent System for Real-World Multi-Image Agricultural VQA
von: Ke, Yan, et al.
Veröffentlicht: (2025)
von: Ke, Yan, et al.
Veröffentlicht: (2025)
Ctrl-VI: Controllable Video Synthesis via Variational Inference
von: Duan, Haoyi, et al.
Veröffentlicht: (2025)
von: Duan, Haoyi, et al.
Veröffentlicht: (2025)
COMBO: Compositional World Models for Embodied Multi-Agent Cooperation
von: Zhang, Hongxin, et al.
Veröffentlicht: (2024)
von: Zhang, Hongxin, et al.
Veröffentlicht: (2024)
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
von: Guo, Zirun, et al.
Veröffentlicht: (2025)
Do Vision-Language Models Have Internal World Models? Towards an Atomic Evaluation
von: Gao, Qiyue, et al.
Veröffentlicht: (2025)
von: Gao, Qiyue, et al.
Veröffentlicht: (2025)
DatasetNeRF: Efficient 3D-aware Data Factory with Generative Radiance Fields
von: Chi, Yu, et al.
Veröffentlicht: (2023)
von: Chi, Yu, et al.
Veröffentlicht: (2023)
FreGS: 3D Gaussian Splatting with Progressive Frequency Regularization
von: Zhang, Jiahui, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahui, et al.
Veröffentlicht: (2024)
MuSASplat: Efficient Sparse-View 3D Gaussian Splats via Lightweight Multi-Scale Adaptation
von: Xu, Muyu, et al.
Veröffentlicht: (2025)
von: Xu, Muyu, et al.
Veröffentlicht: (2025)
MIND: Benchmarking Memory Consistency and Action Control in World Models
von: Ye, Yixuan, et al.
Veröffentlicht: (2026)
von: Ye, Yixuan, et al.
Veröffentlicht: (2026)
3D-VLA: A 3D Vision-Language-Action Generative World Model
von: Zhen, Haoyu, et al.
Veröffentlicht: (2024)
von: Zhen, Haoyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AAPMT: AGI Assessment Through Prompt and Metric Transformer
von: Huang, Benhao
Veröffentlicht: (2024) -
Flow Equivariant Recurrent Neural Networks
von: Keller, T. Anderson
Veröffentlicht: (2025) -
MixLight: Borrowing the Best of both Spherical Harmonics and Gaussian Models
von: Ji, Xinlong, et al.
Veröffentlicht: (2024) -
Abstract 3D Perception for Spatial Intelligence in Vision-Language Models
von: Liu, Yifan, et al.
Veröffentlicht: (2025) -
GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)