Your Latent Mask is Wrong: Pixel-Equivalent Latent Compositing for Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bradbury, Rowan, Zhong, Dazhi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LatentHDR: Decoupling Exposure from Diffusion via Conditional Latent-to-Latent Mapping for Text/Image-to-Panoramic HDR
von: Fekri, Pedram, et al.
Veröffentlicht: (2026)
von: Fekri, Pedram, et al.
Veröffentlicht: (2026)
Single Mesh Diffusion Models with Field Latents for Texture Generation
von: Mitchel, Thomas W., et al.
Veröffentlicht: (2023)
von: Mitchel, Thomas W., et al.
Veröffentlicht: (2023)
Multimodal Latent Diffusion Model for Complex Sewing Pattern Generation
von: Liu, Shengqi, et al.
Veröffentlicht: (2024)
von: Liu, Shengqi, et al.
Veröffentlicht: (2024)
LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene Relighting
von: Xing, Xiaoyan, et al.
Veröffentlicht: (2024)
von: Xing, Xiaoyan, et al.
Veröffentlicht: (2024)
MESA: Text-Driven Terrain Generation Using Latent Diffusion and Global Copernicus Data
von: Borne--Pons, Paul, et al.
Veröffentlicht: (2025)
von: Borne--Pons, Paul, et al.
Veröffentlicht: (2025)
Local Scale Equivariance with Latent Deep Equilibrium Canonicalizer
von: Rahman, Md Ashiqur, et al.
Veröffentlicht: (2025)
von: Rahman, Md Ashiqur, et al.
Veröffentlicht: (2025)
Preserve Your Own Correlation: A Noise Prior for Video Diffusion Models
von: Ge, Songwei, et al.
Veröffentlicht: (2023)
von: Ge, Songwei, et al.
Veröffentlicht: (2023)
COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models
von: Daiya, Divyanshu, et al.
Veröffentlicht: (2024)
von: Daiya, Divyanshu, et al.
Veröffentlicht: (2024)
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
von: Guo, Yuwei, et al.
Veröffentlicht: (2023)
von: Guo, Yuwei, et al.
Veröffentlicht: (2023)
ODE-GS: Latent ODEs for Dynamic Scene Extrapolation with 3D Gaussian Splatting
von: Wang, Daniel, et al.
Veröffentlicht: (2025)
von: Wang, Daniel, et al.
Veröffentlicht: (2025)
SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing
von: Hong, Seokhyeon, et al.
Veröffentlicht: (2025)
von: Hong, Seokhyeon, et al.
Veröffentlicht: (2025)
An Object is Worth 64x64 Pixels: Generating 3D Object via Image Diffusion
von: Yan, Xingguang, et al.
Veröffentlicht: (2024)
von: Yan, Xingguang, et al.
Veröffentlicht: (2024)
Two Heads are Better than One: Geometric-Latent Attention for Point Cloud Classification and Segmentation
von: Cuevas-Velasquez, Hanz, et al.
Veröffentlicht: (2021)
von: Cuevas-Velasquez, Hanz, et al.
Veröffentlicht: (2021)
Click2Mask: Local Editing with Dynamic Mask Generation
von: Regev, Omer, et al.
Veröffentlicht: (2024)
von: Regev, Omer, et al.
Veröffentlicht: (2024)
End-to-End Training for Unified Tokenization and Latent Denoising
von: Duggal, Shivam, et al.
Veröffentlicht: (2026)
von: Duggal, Shivam, et al.
Veröffentlicht: (2026)
Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2025)
NeuralRemaster: Phase-Preserving Diffusion for Structure-Aligned Generation
von: Zeng, Yu, et al.
Veröffentlicht: (2025)
von: Zeng, Yu, et al.
Veröffentlicht: (2025)
Squeeze3D: Your 3D Generation Model is Secretly an Extreme Neural Compressor
von: Dagli, Rishit, et al.
Veröffentlicht: (2025)
von: Dagli, Rishit, et al.
Veröffentlicht: (2025)
TetraDiffusion: Tetrahedral Diffusion Models for 3D Shape Generation
von: Kalischek, Nikolai, et al.
Veröffentlicht: (2022)
von: Kalischek, Nikolai, et al.
Veröffentlicht: (2022)
RealOSR: Latent Guidance Boosts Diffusion-based Real-world Omnidirectional Image Super-Resolutions
von: Sheng, Xuhan, et al.
Veröffentlicht: (2024)
von: Sheng, Xuhan, et al.
Veröffentlicht: (2024)
Masked Extended Attention for Zero-Shot Virtual Try-On In The Wild
von: Orzech, Nadav, et al.
Veröffentlicht: (2024)
von: Orzech, Nadav, et al.
Veröffentlicht: (2024)
Flexible Motion In-betweening with Diffusion Models
von: Cohan, Setareh, et al.
Veröffentlicht: (2024)
von: Cohan, Setareh, et al.
Veröffentlicht: (2024)
Distilling Diffusion Models into Conditional GANs
von: Kang, Minguk, et al.
Veröffentlicht: (2024)
von: Kang, Minguk, et al.
Veröffentlicht: (2024)
Interpreting the Weight Space of Customized Diffusion Models
von: Dravid, Amil, et al.
Veröffentlicht: (2024)
von: Dravid, Amil, et al.
Veröffentlicht: (2024)
Diff-3DCap: Shape Captioning with Diffusion Models
von: Shu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Shu, Zhenyu, et al.
Veröffentlicht: (2025)
SliderSpace: Decomposing the Visual Capabilities of Diffusion Models
von: Gandikota, Rohit, et al.
Veröffentlicht: (2025)
von: Gandikota, Rohit, et al.
Veröffentlicht: (2025)
Transparent Image Layer Diffusion using Latent Transparency
von: Zhang, Lvmin, et al.
Veröffentlicht: (2024)
von: Zhang, Lvmin, et al.
Veröffentlicht: (2024)
Curved Diffusion: A Generative Model With Optical Geometry Control
von: Voynov, Andrey, et al.
Veröffentlicht: (2023)
von: Voynov, Andrey, et al.
Veröffentlicht: (2023)
Enhancing Image Layout Control with Loss-Guided Diffusion Models
von: Patel, Zakaria, et al.
Veröffentlicht: (2024)
von: Patel, Zakaria, et al.
Veröffentlicht: (2024)
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
von: Avrahami, Omri, et al.
Veröffentlicht: (2023)
von: Avrahami, Omri, et al.
Veröffentlicht: (2023)
Adaptive Hybrid Caching for Efficient Text-to-Video Diffusion Model Acceleration
von: Wei, Yuanxin, et al.
Veröffentlicht: (2025)
von: Wei, Yuanxin, et al.
Veröffentlicht: (2025)
VFusion3D: Learning Scalable 3D Generative Models from Video Diffusion Models
von: Han, Junlin, et al.
Veröffentlicht: (2024)
von: Han, Junlin, et al.
Veröffentlicht: (2024)
Leveraging Semantic Attribute Binding for Free-Lunch Color Control in Diffusion Models
von: Laria, Héctor, et al.
Veröffentlicht: (2025)
von: Laria, Héctor, et al.
Veröffentlicht: (2025)
DreamBlend: Advancing Personalized Fine-tuning of Text-to-Image Diffusion Models
von: Ram, Shwetha, et al.
Veröffentlicht: (2024)
von: Ram, Shwetha, et al.
Veröffentlicht: (2024)
L3DG: Latent 3D Gaussian Diffusion
von: Roessle, Barbara, et al.
Veröffentlicht: (2024)
von: Roessle, Barbara, et al.
Veröffentlicht: (2024)
TerraFusion: Joint Generation of Terrain Geometry and Texture Using Latent Diffusion Models
von: Higo, Kazuki, et al.
Veröffentlicht: (2025)
von: Higo, Kazuki, et al.
Veröffentlicht: (2025)
CookingDiffusion: Cooking Procedural Image Generation with Stable Diffusion
von: Wang, Yuan, et al.
Veröffentlicht: (2025)
von: Wang, Yuan, et al.
Veröffentlicht: (2025)
REED-VAE: RE-Encode Decode Training for Iterative Image Editing with Diffusion Models
von: Almog, Gal, et al.
Veröffentlicht: (2025)
von: Almog, Gal, et al.
Veröffentlicht: (2025)
Drag Your Noise: Interactive Point-based Editing via Diffusion Semantic Propagation
von: Liu, Haofeng, et al.
Veröffentlicht: (2024)
von: Liu, Haofeng, et al.
Veröffentlicht: (2024)
A General Implicit Framework for Fast NeRF Composition and Rendering
von: Gao, Xinyu, et al.
Veröffentlicht: (2023)
von: Gao, Xinyu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
LatentHDR: Decoupling Exposure from Diffusion via Conditional Latent-to-Latent Mapping for Text/Image-to-Panoramic HDR
von: Fekri, Pedram, et al.
Veröffentlicht: (2026) -
Single Mesh Diffusion Models with Field Latents for Texture Generation
von: Mitchel, Thomas W., et al.
Veröffentlicht: (2023) -
Multimodal Latent Diffusion Model for Complex Sewing Pattern Generation
von: Liu, Shengqi, et al.
Veröffentlicht: (2024) -
LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene Relighting
von: Xing, Xiaoyan, et al.
Veröffentlicht: (2024) -
MESA: Text-Driven Terrain Generation Using Latent Diffusion and Global Copernicus Data
von: Borne--Pons, Paul, et al.
Veröffentlicht: (2025)