Wonderland: Navigating 3D Scenes from a Single Image
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Hanwen, Cao, Junli, Goel, Vidit, Qian, Guocheng, Korolev, Sergei, Terzopoulos, Demetri, Plataniotis, Konstantinos N., Tulyakov, Sergey, Ren, Jian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Physical Understanding in Video Generation: A 3D Point Regularization Approach
von: Chen, Yunuo, et al.
Veröffentlicht: (2025)
von: Chen, Yunuo, et al.
Veröffentlicht: (2025)
Lightweight Predictive 3D Gaussian Splats
von: Cao, Junli, et al.
Veröffentlicht: (2024)
von: Cao, Junli, et al.
Veröffentlicht: (2024)
Prompting Medical Large Vision-Language Models to Diagnose Pathologies by Visual Question Answering
von: Guo, Danfeng, et al.
Veröffentlicht: (2024)
von: Guo, Danfeng, et al.
Veröffentlicht: (2024)
4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation
von: Wang, Chaoyang, et al.
Veröffentlicht: (2025)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2025)
Comp4D: LLM-Guided Compositional 4D Scene Generation
von: Xu, Dejia, et al.
Veröffentlicht: (2024)
von: Xu, Dejia, et al.
Veröffentlicht: (2024)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
TiP4GEN: Text to Immersive Panorama 4D Scene Generation
von: Xing, Ke, et al.
Veröffentlicht: (2025)
von: Xing, Ke, et al.
Veröffentlicht: (2025)
Inverse Attention Agents for Multi-Agent Systems
von: Long, Qian, et al.
Veröffentlicht: (2024)
von: Long, Qian, et al.
Veröffentlicht: (2024)
Diffusion4D: Fast Spatial-temporal Consistent 4D Generation via Video Diffusion Models
von: Liang, Hanwen, et al.
Veröffentlicht: (2024)
von: Liang, Hanwen, et al.
Veröffentlicht: (2024)
4Real: Towards Photorealistic 4D Scene Generation via Video Diffusion Models
von: Yu, Heng, et al.
Veröffentlicht: (2024)
von: Yu, Heng, et al.
Veröffentlicht: (2024)
AsCAN: Asymmetric Convolution-Attention Networks for Efficient Recognition and Generation
von: Kag, Anil, et al.
Veröffentlicht: (2024)
von: Kag, Anil, et al.
Veröffentlicht: (2024)
AToM: Amortized Text-to-Mesh using 2D Diffusion
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft
von: Long, Qian, et al.
Veröffentlicht: (2024)
von: Long, Qian, et al.
Veröffentlicht: (2024)
ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
The Need for Speed: Pruning Transformers with One Recipe
von: Khaki, Samir, et al.
Veröffentlicht: (2024)
von: Khaki, Samir, et al.
Veröffentlicht: (2024)
Active Inference and Reinforcement Learning: A unified inference on continuous state and action spaces under partial observability
von: Malekzadeh, Parvin, et al.
Veröffentlicht: (2022)
von: Malekzadeh, Parvin, et al.
Veröffentlicht: (2022)
Diffusion-DRF: Free, Rich, and Differentiable Reward for Video Diffusion Fine-Tuning
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
SF-V: Single Forward Video Generation Model
von: Zhang, Zhixing, et al.
Veröffentlicht: (2024)
von: Zhang, Zhixing, et al.
Veröffentlicht: (2024)
Learning Neural Force Manifolds for Sim2Real Robotic Symmetrical Paper Folding
von: Choi, Andrew, et al.
Veröffentlicht: (2023)
von: Choi, Andrew, et al.
Veröffentlicht: (2023)
A Human-centric Framework for Debating the Ethics of AI Consciousness Under Uncertainty
von: Ziheng, Zhou, et al.
Veröffentlicht: (2025)
von: Ziheng, Zhou, et al.
Veröffentlicht: (2025)
AC3D: Analyzing and Improving 3D Camera Control in Video Diffusion Transformers
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
von: Bahmani, Sherwin, et al.
Veröffentlicht: (2024)
Unstructured Moving Least Squares Material Point Methods: A Stable Kernel Approach With Continuous Gradient Reconstruction on General Unstructured Tessellations
von: Cao, Yadi, et al.
Veröffentlicht: (2023)
von: Cao, Yadi, et al.
Veröffentlicht: (2023)
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
Scalable Ranked Preference Optimization for Text-to-Image Generation
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
Visual Concept-driven Image Generation with Text-to-Image Diffusion Model
von: Rahman, Tanzila, et al.
Veröffentlicht: (2024)
von: Rahman, Tanzila, et al.
Veröffentlicht: (2024)
Producing Histopathology Phantom Images using Generative Adversarial Networks to improve Tumor Detection
von: Gautam, Vidit
Veröffentlicht: (2022)
von: Gautam, Vidit
Veröffentlicht: (2022)
SPAD : Spatially Aware Multiview Diffusers
von: Kant, Yash, et al.
Veröffentlicht: (2024)
von: Kant, Yash, et al.
Veröffentlicht: (2024)
AnimaMimic: Imitating 3D Animation from Video Priors
von: Xie, Tianyi, et al.
Veröffentlicht: (2025)
von: Xie, Tianyi, et al.
Veröffentlicht: (2025)
BitsFusion: 1.99 bits Weight Quantization of Diffusion Model
von: Sui, Yang, et al.
Veröffentlicht: (2024)
von: Sui, Yang, et al.
Veröffentlicht: (2024)
mBEST: Realtime Deformable Linear Object Detection Through Minimal Bending Energy Skeleton Pixel Traversals
von: Choi, Andrew, et al.
Veröffentlicht: (2023)
von: Choi, Andrew, et al.
Veröffentlicht: (2023)
A unified uncertainty-aware exploration: Combining epistemic and aleatory uncertainty
von: Malekzadeh, Parvin, et al.
Veröffentlicht: (2024)
von: Malekzadeh, Parvin, et al.
Veröffentlicht: (2024)
Uncertainty-aware transfer across tasks using hybrid model-based successor feature reinforcement learning
von: Malekzadeh, Parvin, et al.
Veröffentlicht: (2023)
von: Malekzadeh, Parvin, et al.
Veröffentlicht: (2023)
Wonderland / Leslie Sardiñas
von: Sardiñas, Leslie
von: Sardiñas, Leslie
Touring an Informational Wonderland.
von: Martin, William
Veröffentlicht: (1984)
von: Martin, William
Veröffentlicht: (1984)
PanFlow: Decoupled Motion Control for Panoramic Video Generation
von: Zhang, Cheng, et al.
Veröffentlicht: (2025)
von: Zhang, Cheng, et al.
Veröffentlicht: (2025)
Omni-ID: Holistic Identity Representation Designed for Generative Tasks
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
MaskControl: Spatio-Temporal Control for Masked Motion Synthesis
von: Pinyoanuntapong, Ekkasit, et al.
Veröffentlicht: (2024)
von: Pinyoanuntapong, Ekkasit, et al.
Veröffentlicht: (2024)
Cross-Slice Attention and Evidential Critical Loss for Uncertainty-Aware Prostate Cancer Detection
von: Hung, Alex Ling Yu, et al.
Veröffentlicht: (2024)
von: Hung, Alex Ling Yu, et al.
Veröffentlicht: (2024)
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models
von: Mi, Zhenxing, et al.
Veröffentlicht: (2025)
von: Mi, Zhenxing, et al.
Veröffentlicht: (2025)
EasyV2V: A High-quality Instruction-based Video Editing Framework
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards Physical Understanding in Video Generation: A 3D Point Regularization Approach
von: Chen, Yunuo, et al.
Veröffentlicht: (2025) -
Lightweight Predictive 3D Gaussian Splats
von: Cao, Junli, et al.
Veröffentlicht: (2024) -
Prompting Medical Large Vision-Language Models to Diagnose Pathologies by Visual Question Answering
von: Guo, Danfeng, et al.
Veröffentlicht: (2024) -
4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation
von: Wang, Chaoyang, et al.
Veröffentlicht: (2025) -
Comp4D: LLM-Guided Compositional 4D Scene Generation
von: Xu, Dejia, et al.
Veröffentlicht: (2024)