DisCo-Diff: Enhancing Continuous Diffusion Models with Discrete Latents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Yilun, Corso, Gabriele, Jaakkola, Tommi, Vahdat, Arash, Kreis, Karsten |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Truncated Consistency Models
von: Lee, Sangyun, et al.
Veröffentlicht: (2024)
von: Lee, Sangyun, et al.
Veröffentlicht: (2024)
One-step Diffusion Models with $f$-Divergence Distribution Matching
von: Xu, Yilun, et al.
Veröffentlicht: (2025)
von: Xu, Yilun, et al.
Veröffentlicht: (2025)
DisCo: Disentangled Control for Realistic Human Dance Generation
von: Wang, Tan, et al.
Veröffentlicht: (2023)
von: Wang, Tan, et al.
Veröffentlicht: (2023)
Warped Diffusion: Solving Video Inverse Problems with Image Diffusion Models
von: Daras, Giannis, et al.
Veröffentlicht: (2024)
von: Daras, Giannis, et al.
Veröffentlicht: (2024)
O-DisCo-Edit: Object Distortion Control for Unified Realistic Video Editing
von: Chen, Yuqing, et al.
Veröffentlicht: (2025)
von: Chen, Yuqing, et al.
Veröffentlicht: (2025)
On Equivariance and Fast Sampling in Video Diffusion Models Trained with Warped Noise
von: Liu, Chao, et al.
Veröffentlicht: (2025)
von: Liu, Chao, et al.
Veröffentlicht: (2025)
Think While You Generate: Discrete Diffusion with Planned Denoising
von: Liu, Sulin, et al.
Veröffentlicht: (2024)
von: Liu, Sulin, et al.
Veröffentlicht: (2024)
DiLaDiff: Distilled Latent-Augmented Diffusion for Language Modeling
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2026)
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2026)
Fast Training of Diffusion Models with Masked Transformers
von: Zheng, Hongkai, et al.
Veröffentlicht: (2023)
von: Zheng, Hongkai, et al.
Veröffentlicht: (2023)
Elucidating Optimal Reward-Diversity Tradeoffs in Text-to-Image Diffusion Models
von: Jena, Rohit, et al.
Veröffentlicht: (2024)
von: Jena, Rohit, et al.
Veröffentlicht: (2024)
DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding
von: Cho, Jungbin, et al.
Veröffentlicht: (2024)
von: Cho, Jungbin, et al.
Veröffentlicht: (2024)
DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
Learning Diffusion Models with Flexible Representation Guidance
von: Wang, Chenyu, et al.
Veröffentlicht: (2025)
von: Wang, Chenyu, et al.
Veröffentlicht: (2025)
DisCo: Towards Distinct and Coherent Visual Encapsulation in Video MLLMs
von: Zhao, Jiahe, et al.
Veröffentlicht: (2025)
von: Zhao, Jiahe, et al.
Veröffentlicht: (2025)
Real-Time Fusion of Visual and Chart Data for Enhanced Maritime Vision
von: Kreis, Marten, et al.
Veröffentlicht: (2025)
von: Kreis, Marten, et al.
Veröffentlicht: (2025)
DiffiT: Diffusion Vision Transformers for Image Generation
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
Multi-student Diffusion Distillation for Better One-step Generators
von: Song, Yanke, et al.
Veröffentlicht: (2024)
von: Song, Yanke, et al.
Veröffentlicht: (2024)
CoCoDiff: Correspondence-Consistent Diffusion Model for Fine-grained Style Transfer
von: Nie, Wenbo, et al.
Veröffentlicht: (2026)
von: Nie, Wenbo, et al.
Veröffentlicht: (2026)
CoDiff: Conditional Diffusion Model for Collaborative 3D Object Detection
von: Huang, Zhe, et al.
Veröffentlicht: (2025)
von: Huang, Zhe, et al.
Veröffentlicht: (2025)
Img-Diff: Contrastive Data Synthesis for Multimodal Large Language Models
von: Jiao, Qirui, et al.
Veröffentlicht: (2024)
von: Jiao, Qirui, et al.
Veröffentlicht: (2024)
DisCo3D: Distilling Multi-View Consistency for 3D Scene Editing
von: Chi, Yufeng, et al.
Veröffentlicht: (2025)
von: Chi, Yufeng, et al.
Veröffentlicht: (2025)
Not-So-Optimal Transport Flows for 3D Point Cloud Generation
von: Hui, Ka-Hei, et al.
Veröffentlicht: (2025)
von: Hui, Ka-Hei, et al.
Veröffentlicht: (2025)
AnimateDiff-Lightning: Cross-Model Diffusion Distillation
von: Lin, Shanchuan, et al.
Veröffentlicht: (2024)
von: Lin, Shanchuan, et al.
Veröffentlicht: (2024)
AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
Fuse Your Latents: Video Editing with Multi-source Latent Diffusion Models
von: Lu, Tianyi, et al.
Veröffentlicht: (2023)
von: Lu, Tianyi, et al.
Veröffentlicht: (2023)
BlobGEN-Vid: Compositional Text-to-Video Generation with Blob Video Representations
von: Feng, Weixi, et al.
Veröffentlicht: (2025)
von: Feng, Weixi, et al.
Veröffentlicht: (2025)
AdaDiff: Adaptive Step Selection for Fast Diffusion Models
von: Zhang, Hui, et al.
Veröffentlicht: (2023)
von: Zhang, Hui, et al.
Veröffentlicht: (2023)
DiffGuard: Text-Based Safety Checker for Diffusion Models
von: Khader, Massine El, et al.
Veröffentlicht: (2024)
von: Khader, Massine El, et al.
Veröffentlicht: (2024)
DiffMorph: Text-less Image Morphing with Diffusion Models
von: Chatterjee, Shounak
Veröffentlicht: (2024)
von: Chatterjee, Shounak
Veröffentlicht: (2024)
MacDiff: Unified Skeleton Modeling with Masked Conditional Diffusion
von: Wu, Lehong, et al.
Veröffentlicht: (2024)
von: Wu, Lehong, et al.
Veröffentlicht: (2024)
Compositional Discrete Latent Code for High Fidelity, Productive Diffusion Models
von: Lavoie, Samuel, et al.
Veröffentlicht: (2025)
von: Lavoie, Samuel, et al.
Veröffentlicht: (2025)
Vision-Enhanced Time Series Forecasting via Latent Diffusion Models
von: Ruan, Weilin, et al.
Veröffentlicht: (2025)
von: Ruan, Weilin, et al.
Veröffentlicht: (2025)
DiffRIS: Enhancing Referring Remote Sensing Image Segmentation with Pre-trained Text-to-Image Diffusion Models
von: Dong, Zhe, et al.
Veröffentlicht: (2025)
von: Dong, Zhe, et al.
Veröffentlicht: (2025)
DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching
von: Zou, Chang, et al.
Veröffentlicht: (2026)
von: Zou, Chang, et al.
Veröffentlicht: (2026)
Enhancing Spatiotemporal Disease Progression Models via Latent Diffusion and Prior Knowledge
von: Puglisi, Lemuel, et al.
Veröffentlicht: (2024)
von: Puglisi, Lemuel, et al.
Veröffentlicht: (2024)
RealignDiff: Boosting Text-to-Image Diffusion Model with Coarse-to-fine Semantic Re-alignment
von: Jiang, Zutao, et al.
Veröffentlicht: (2023)
von: Jiang, Zutao, et al.
Veröffentlicht: (2023)
DiffAttn: Diffusion-Based Drivers' Visual Attention Prediction with LLM-Enhanced Semantic Reasoning
von: Liu, Weimin, et al.
Veröffentlicht: (2026)
von: Liu, Weimin, et al.
Veröffentlicht: (2026)
DiffCrossGait: Trajectory-Level Alignment for 2D-3D Cross-Modal Gait Recognition via Latent Diffusion
von: Lu, Zhiyang, et al.
Veröffentlicht: (2026)
von: Lu, Zhiyang, et al.
Veröffentlicht: (2026)
DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models
von: Yu, Zhengming, et al.
Veröffentlicht: (2026)
von: Yu, Zhengming, et al.
Veröffentlicht: (2026)
DivDiff: A Conditional Diffusion Model for Diverse Human Motion Prediction
von: Yu, Hua, et al.
Veröffentlicht: (2024)
von: Yu, Hua, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Truncated Consistency Models
von: Lee, Sangyun, et al.
Veröffentlicht: (2024) -
One-step Diffusion Models with $f$-Divergence Distribution Matching
von: Xu, Yilun, et al.
Veröffentlicht: (2025) -
DisCo: Disentangled Control for Realistic Human Dance Generation
von: Wang, Tan, et al.
Veröffentlicht: (2023) -
Warped Diffusion: Solving Video Inverse Problems with Image Diffusion Models
von: Daras, Giannis, et al.
Veröffentlicht: (2024) -
O-DisCo-Edit: Object Distortion Control for Unified Realistic Video Editing
von: Chen, Yuqing, et al.
Veröffentlicht: (2025)