Robust Dreamer: Deviation-Aware Latent Gaussian Memory for Action-Controlled AR Video Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Hanlin, Wei, Jiaxin, Song, Xibin, Wang, Yifu, Wang, Steve, Li, Hongdong, Ji, Pan, Lee, Gim Hee |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BachVid: Training-Free Video Generation with Consistent Background and Character
von: Yan, Han, et al.
Veröffentlicht: (2025)
von: Yan, Han, et al.
Veröffentlicht: (2025)
ChatSplat: 3D Conversational Gaussian Splatting
von: Chen, Hanlin, et al.
Veröffentlicht: (2024)
von: Chen, Hanlin, et al.
Veröffentlicht: (2024)
NeuSG: Neural Implicit Surface Reconstruction with 3D Gaussian Splatting Guidance
von: Chen, Hanlin, et al.
Veröffentlicht: (2023)
von: Chen, Hanlin, et al.
Veröffentlicht: (2023)
FreeSplat++: Generalizable 3D Gaussian Splatting for Efficient Indoor Scene Reconstruction
von: Wang, Yunsong, et al.
Veröffentlicht: (2025)
von: Wang, Yunsong, et al.
Veröffentlicht: (2025)
FreeSplat: Generalizable 3D Gaussian Splatting Towards Free-View Synthesis of Indoor Scenes
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
GOV-NeSF: Generalizable Open-Vocabulary Neural Semantic Fields
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
I3DM: Implicit 3D-aware Memory Retrieval and Injection for Consistent Video Scene Generation
von: Li, Jia, et al.
Veröffentlicht: (2026)
von: Li, Jia, et al.
Veröffentlicht: (2026)
Flow4DGS-SLAM: Optical Flow-Guided 4D Gaussian Splatting SLAM
von: Wang, Yunsong, et al.
Veröffentlicht: (2026)
von: Wang, Yunsong, et al.
Veröffentlicht: (2026)
GaussVideoDreamer: 3D Scene Generation with Video Diffusion and Inconsistency-Aware Gaussian Splatting
von: Hao, Junlin, et al.
Veröffentlicht: (2025)
von: Hao, Junlin, et al.
Veröffentlicht: (2025)
DOGS: Distributed-Oriented Gaussian Splatting for Large-Scale 3D Reconstruction Via Gaussian Consensus
von: Chen, Yu, et al.
Veröffentlicht: (2024)
von: Chen, Yu, et al.
Veröffentlicht: (2024)
Uni4D-LLM: A Unified SpatioTemporal-Aware VLM for 4D Understanding and Generation
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
VCR-GauS: View Consistent Depth-Normal Regularizer for Gaussian Surface Reconstruction
von: Chen, Hanlin, et al.
Veröffentlicht: (2024)
von: Chen, Hanlin, et al.
Veröffentlicht: (2024)
DiET-GS: Diffusion Prior and Event Stream-Assisted Motion Deblurring 3D Gaussian Splatting
von: Lee, Seungjun, et al.
Veröffentlicht: (2025)
von: Lee, Seungjun, et al.
Veröffentlicht: (2025)
VLA-4D: Embedding 4D Awareness into Vision-Language-Action Models for SpatioTemporally Coherent Robotic Manipulation
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
T$^3$-S2S: Training-free Triplet Tuning for Sketch to Scene Synthesis in Controllable Concept Art Generation
von: Sun, Zhenhong, et al.
Veröffentlicht: (2024)
von: Sun, Zhenhong, et al.
Veröffentlicht: (2024)
econSG: Efficient and Multi-view Consistent Open-Vocabulary 3D Semantic Gaussians
von: Zhang, Can, et al.
Veröffentlicht: (2025)
von: Zhang, Can, et al.
Veröffentlicht: (2025)
MotionScale: Reconstructing Appearance, Geometry, and Motion of Dynamic Scenes with Scalable 4D Gaussian Splatting
von: Zhou, Haoran, et al.
Veröffentlicht: (2026)
von: Zhou, Haoran, et al.
Veröffentlicht: (2026)
UNIKD: UNcertainty-filtered Incremental Knowledge Distillation for Neural Implicit Representation
von: Guo, Mengqi, et al.
Veröffentlicht: (2022)
von: Guo, Mengqi, et al.
Veröffentlicht: (2022)
NeuSDFusion: A Spatial-Aware Generative Model for 3D Shape Completion, Reconstruction, and Generation
von: Cui, Ruikai, et al.
Veröffentlicht: (2024)
von: Cui, Ruikai, et al.
Veröffentlicht: (2024)
g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
JADE: Joint-aware Latent Diffusion for 3D Human Generative Modeling
von: Ji, Haorui, et al.
Veröffentlicht: (2024)
von: Ji, Haorui, et al.
Veröffentlicht: (2024)
Generalizable Human Gaussians from Single-View Image
von: Chen, Jinnan, et al.
Veröffentlicht: (2024)
von: Chen, Jinnan, et al.
Veröffentlicht: (2024)
SPIRAL: Self-Evolving Action-Conditioned Video Generation via Reflective Planning Agents
von: Yang, Yu, et al.
Veröffentlicht: (2026)
von: Yang, Yu, et al.
Veröffentlicht: (2026)
Unified Geometry and Color Compression Framework for Point Clouds via Generative Diffusion Priors
von: Huang, Tianxin, et al.
Veröffentlicht: (2025)
von: Huang, Tianxin, et al.
Veröffentlicht: (2025)
4D3R: Motion-Aware Neural Reconstruction and Rendering of Dynamic Scenes from Monocular Videos
von: Guo, Mengqi, et al.
Veröffentlicht: (2025)
von: Guo, Mengqi, et al.
Veröffentlicht: (2025)
QuaDreamer: Controllable Panoramic Video Generation for Quadruped Robots
von: Wu, Sheng, et al.
Veröffentlicht: (2025)
von: Wu, Sheng, et al.
Veröffentlicht: (2025)
HumanDreamer: Generating Controllable Human-Motion Videos via Decoupled Generation
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
Statistical Confidence Rescoring for Robust 3D Scene Graph Generation from Multi-View Images
von: Yeo, Qi Xun, et al.
Veröffentlicht: (2025)
von: Yeo, Qi Xun, et al.
Veröffentlicht: (2025)
Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
Pandora3D: A Comprehensive Framework for High-Quality 3D Shape and Texture Generation
von: Yang, Jiayu, et al.
Veröffentlicht: (2025)
von: Yang, Jiayu, et al.
Veröffentlicht: (2025)
Segment Any Events with Language
von: Lee, Seungjun, et al.
Veröffentlicht: (2026)
von: Lee, Seungjun, et al.
Veröffentlicht: (2026)
DPBridge: Latent Diffusion Bridge for Dense Prediction
von: Ji, Haorui, et al.
Veröffentlicht: (2024)
von: Ji, Haorui, et al.
Veröffentlicht: (2024)
Make-A-Protagonist: Generic Video Editing with An Ensemble of Experts
von: Zhao, Yuyang, et al.
Veröffentlicht: (2023)
von: Zhao, Yuyang, et al.
Veröffentlicht: (2023)
MVGSR: Multi-View Consistency Gaussian Splatting for Robust Surface Reconstruction
von: Hou, Chenfeng, et al.
Veröffentlicht: (2025)
von: Hou, Chenfeng, et al.
Veröffentlicht: (2025)
SmileSplat: Generalizable Gaussian Splats for Unconstrained Sparse Images
von: Li, Yanyan, et al.
Veröffentlicht: (2024)
von: Li, Yanyan, et al.
Veröffentlicht: (2024)
Corruption-Aware Training of Latent Video Diffusion Models for Robust Text-to-Video Generation
von: Maduabuchi, Chika, et al.
Veröffentlicht: (2025)
von: Maduabuchi, Chika, et al.
Veröffentlicht: (2025)
BAG: Body-Aligned 3D Wearable Asset Generation
von: Luo, Zhongjin, et al.
Veröffentlicht: (2025)
von: Luo, Zhongjin, et al.
Veröffentlicht: (2025)
Syn-to-Real Unsupervised Domain Adaptation for Indoor 3D Object Detection
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
Enhancing Generalizability of Representation Learning for Data-Efficient 3D Scene Understanding
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
NavRAG: Generating User Demand Instructions for Embodied Navigation through Retrieval-Augmented LLM
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
BachVid: Training-Free Video Generation with Consistent Background and Character
von: Yan, Han, et al.
Veröffentlicht: (2025) -
ChatSplat: 3D Conversational Gaussian Splatting
von: Chen, Hanlin, et al.
Veröffentlicht: (2024) -
NeuSG: Neural Implicit Surface Reconstruction with 3D Gaussian Splatting Guidance
von: Chen, Hanlin, et al.
Veröffentlicht: (2023) -
FreeSplat++: Generalizable 3D Gaussian Splatting for Efficient Indoor Scene Reconstruction
von: Wang, Yunsong, et al.
Veröffentlicht: (2025) -
FreeSplat: Generalizable 3D Gaussian Splatting Towards Free-View Synthesis of Indoor Scenes
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)