Controlling Space and Time with Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Watson, Daniel, Saxena, Saurabh, Li, Lala, Tagliasacchi, Andrea, Fleet, David J. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
360Anything: Geometry-Free Lifting of Images and Videos to 360°
by: Wu, Ziyi, et al.
Published: (2026)
by: Wu, Ziyi, et al.
Published: (2026)
RoMo: Robust Motion Segmentation Improves Structure from Motion
by: Goli, Lily, et al.
Published: (2024)
by: Goli, Lily, et al.
Published: (2024)
RobustNeRF: Ignoring Distractors with Robust Losses
by: Sabour, Sara, et al.
Published: (2023)
by: Sabour, Sara, et al.
Published: (2023)
High-Resolution Frame Interpolation with Patch-based Cascaded Diffusion
by: Hur, Junhwa, et al.
Published: (2024)
by: Hur, Junhwa, et al.
Published: (2024)
Radiant Foam: Real-Time Differentiable Ray Tracing
by: Govindarajan, Shrisudhan, et al.
Published: (2025)
by: Govindarajan, Shrisudhan, et al.
Published: (2025)
AC3D: Analyzing and Improving 3D Camera Control in Video Diffusion Transformers
by: Bahmani, Sherwin, et al.
Published: (2024)
by: Bahmani, Sherwin, et al.
Published: (2024)
Untwisting RoPE: Frequency Control for Shared Attention in DiTs
by: Mikaeili, Aryan, et al.
Published: (2026)
by: Mikaeili, Aryan, et al.
Published: (2026)
Directly Fine-Tuning Diffusion Models on Differentiable Rewards
by: Clark, Kevin, et al.
Published: (2023)
by: Clark, Kevin, et al.
Published: (2023)
Evaluating Alternatives to SFM Point Cloud Initialization for Gaussian Splatting
by: Foroutan, Yalda, et al.
Published: (2024)
by: Foroutan, Yalda, et al.
Published: (2024)
SpotlessSplats: Ignoring Distractors in 3D Gaussian Splatting
by: Sabour, Sara, et al.
Published: (2024)
by: Sabour, Sara, et al.
Published: (2024)
pixelSplat: 3D Gaussian Splats from Image Pairs for Scalable Generalizable 3D Reconstruction
by: Charatan, David, et al.
Published: (2023)
by: Charatan, David, et al.
Published: (2023)
Unsupervised Keypoints from Pretrained Diffusion Models
by: Hedlin, Eric, et al.
Published: (2023)
by: Hedlin, Eric, et al.
Published: (2023)
Power Foam: Unifying Real-Time Differentiable Ray Tracing and Rasterization
by: Govindarajan, Shrisudhan, et al.
Published: (2026)
by: Govindarajan, Shrisudhan, et al.
Published: (2026)
3D Gaussian Flats: Hybrid 2D/3D Photometric Scene Reconstruction
by: Taktasheva, Maria, et al.
Published: (2025)
by: Taktasheva, Maria, et al.
Published: (2025)
Volumetric Rendering with Baked Quadrature Fields
by: Sharma, Gopal, et al.
Published: (2023)
by: Sharma, Gopal, et al.
Published: (2023)
Hierarchical Transformers for Unsupervised 3D Shape Abstraction
by: Vora, Aditya, et al.
Published: (2025)
by: Vora, Aditya, et al.
Published: (2025)
Nexels: Neurally-Textured Surfels for Real-Time Novel View Synthesis with Sparse Geometries
by: Rong, Victor, et al.
Published: (2025)
by: Rong, Victor, et al.
Published: (2025)
Spherical Voronoi: Directional Appearance as a Differentiable Partition of the Sphere
by: Di Sario, Francesco, et al.
Published: (2025)
by: Di Sario, Francesco, et al.
Published: (2025)
VD3D: Taming Large Video Diffusion Transformers for 3D Camera Control
by: Bahmani, Sherwin, et al.
Published: (2024)
by: Bahmani, Sherwin, et al.
Published: (2024)
Lumiere: A Space-Time Diffusion Model for Video Generation
by: Bar-Tal, Omer, et al.
Published: (2024)
by: Bar-Tal, Omer, et al.
Published: (2024)
NoKSR: Kernel-Free Neural Surface Reconstruction via Point Cloud Serialization
by: Li, Zhen, et al.
Published: (2025)
by: Li, Zhen, et al.
Published: (2025)
Grow with the Flow: 4D Reconstruction of Growing Plants with Gaussian Flow Fields
by: Luo, Weihan, et al.
Published: (2026)
by: Luo, Weihan, et al.
Published: (2026)
Triangle Splatting for Real-Time Radiance Field Rendering
by: Held, Jan, et al.
Published: (2025)
by: Held, Jan, et al.
Published: (2025)
Video Interpolation with Diffusion Models
by: Jain, Siddhant, et al.
Published: (2024)
by: Jain, Siddhant, et al.
Published: (2024)
Griffin: Generative Reference and Layout Guided Image Composition
by: Mikaeili, Aryan, et al.
Published: (2025)
by: Mikaeili, Aryan, et al.
Published: (2025)
A Personalized Video-Based Hand Taxonomy: Application for Individuals with Spinal Cord Injury
by: Dousty, Mehdy, et al.
Published: (2024)
by: Dousty, Mehdy, et al.
Published: (2024)
Semantic Foam: Unifying Spatial and Semantic Scene Decomposition
by: Sharafeldin, Amr, et al.
Published: (2026)
by: Sharafeldin, Amr, et al.
Published: (2026)
FullCircle: Effortless 3D Reconstruction from Casual 360$^\circ$ Captures
by: Foroutan, Yalda, et al.
Published: (2026)
by: Foroutan, Yalda, et al.
Published: (2026)
UniDiff: Parameter-Efficient Adaptation of Diffusion Models for Land Cover Classification with Multi-Modal Remotely Sensed Imagery and Sparse Annotations
by: Hu, Yuzhen, et al.
Published: (2025)
by: Hu, Yuzhen, et al.
Published: (2025)
MResT: Multi-Resolution Sensing for Real-Time Control with Vision-Language Models
by: Saxena, Saumya, et al.
Published: (2024)
by: Saxena, Saumya, et al.
Published: (2024)
Lyra: Generative 3D Scene Reconstruction via Video Diffusion Model Self-Distillation
by: Bahmani, Sherwin, et al.
Published: (2025)
by: Bahmani, Sherwin, et al.
Published: (2025)
Control-DINO: Feature Space Conditioning for Controllable Image-to-Video Diffusion
by: Dominici, Edoardo A., et al.
Published: (2026)
by: Dominici, Edoardo A., et al.
Published: (2026)
LayerDiffusion: Layered Controlled Image Editing with Diffusion Models
by: Li, Pengzhi, et al.
Published: (2023)
by: Li, Pengzhi, et al.
Published: (2023)
Controllable Face Synthesis with Semantic Latent Diffusion Models
by: Ergasti, Alex, et al.
Published: (2024)
by: Ergasti, Alex, et al.
Published: (2024)
SMITE: Segment Me In TimE
by: Alimohammadi, Amirhossein, et al.
Published: (2024)
by: Alimohammadi, Amirhossein, et al.
Published: (2024)
LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation
by: Zheng, Guangcong, et al.
Published: (2023)
by: Zheng, Guangcong, et al.
Published: (2023)
Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise
by: Burgert, Ryan, et al.
Published: (2025)
by: Burgert, Ryan, et al.
Published: (2025)
Diffusion Hyperfeatures: Searching Through Time and Space for Semantic Correspondence
by: Luo, Grace, et al.
Published: (2023)
by: Luo, Grace, et al.
Published: (2023)
Audio-visual Controlled Video Diffusion with Masked Selective State Spaces Modeling for Natural Talking Head Generation
by: Hong, Fa-Ting, et al.
Published: (2025)
by: Hong, Fa-Ting, et al.
Published: (2025)
Distilling Diversity and Control in Diffusion Models
by: Gandikota, Rohit, et al.
Published: (2025)
by: Gandikota, Rohit, et al.
Published: (2025)
Similar Items
-
360Anything: Geometry-Free Lifting of Images and Videos to 360°
by: Wu, Ziyi, et al.
Published: (2026) -
RoMo: Robust Motion Segmentation Improves Structure from Motion
by: Goli, Lily, et al.
Published: (2024) -
RobustNeRF: Ignoring Distractors with Robust Losses
by: Sabour, Sara, et al.
Published: (2023) -
High-Resolution Frame Interpolation with Patch-based Cascaded Diffusion
by: Hur, Junhwa, et al.
Published: (2024) -
Radiant Foam: Real-Time Differentiable Ray Tracing
by: Govindarajan, Shrisudhan, et al.
Published: (2025)