Improving Vector-Quantized Image Modeling with Latent Consistency-Matching Diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen, Bac, Lai, Chieh-Hsin, Takida, Yuhta, Murata, Naoki, Uesaka, Toshimitsu, Ermon, Stefano, Mitsufuji, Yuki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GUDA: Counterfactual Group-wise Training Data Attribution for Diffusion Models via Unlearning
von: Murata, Naoki, et al.
Veröffentlicht: (2026)
von: Murata, Naoki, et al.
Veröffentlicht: (2026)
G2D2: Gradient-Guided Discrete Diffusion for Inverse Problem Solving
von: Murata, Naoki, et al.
Veröffentlicht: (2024)
von: Murata, Naoki, et al.
Veröffentlicht: (2024)
Improved Object-Centric Diffusion Learning with Registers and Contrastive Alignment
von: Nguyen, Bac, et al.
Veröffentlicht: (2026)
von: Nguyen, Bac, et al.
Veröffentlicht: (2026)
A Unified View of Score-Based and Drifting Models
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2026)
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2026)
Consistency Trajectory Models: Learning Probability Flow ODE Trajectory of Diffusion
von: Kim, Dongjun, et al.
Veröffentlicht: (2023)
von: Kim, Dongjun, et al.
Veröffentlicht: (2023)
Weighted Point Set Embedding for Multimodal Contrastive Learning Toward Optimal Similarity Metric
von: Uesaka, Toshimitsu, et al.
Veröffentlicht: (2024)
von: Uesaka, Toshimitsu, et al.
Veröffentlicht: (2024)
PaGoDA: Progressive Growing of a One-Step Generator from a Low-Resolution Diffusion Teacher
von: Kim, Dongjun, et al.
Veröffentlicht: (2024)
von: Kim, Dongjun, et al.
Veröffentlicht: (2024)
SAN: Inducing Metrizability of GAN with Discriminative Normalized Linear Layer
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
SONA: Learning Conditional, Unconditional, and Mismatching-Aware Discriminator
von: Takida, Yuhta, et al.
Veröffentlicht: (2025)
von: Takida, Yuhta, et al.
Veröffentlicht: (2025)
Theoretical Refinement of CLIP by Utilizing Linear Structure of Optimal Similarity
von: Yoshida, Naoki, et al.
Veröffentlicht: (2025)
von: Yoshida, Naoki, et al.
Veröffentlicht: (2025)
VCT: Training Consistency Models with Variational Noise Coupling
von: Silvestri, Gianluigi, et al.
Veröffentlicht: (2025)
von: Silvestri, Gianluigi, et al.
Veröffentlicht: (2025)
Noise Scheduling as Information-Guided Allocation in Diffusion Training
von: Raya, Gabriel, et al.
Veröffentlicht: (2026)
von: Raya, Gabriel, et al.
Veröffentlicht: (2026)
Distill, Forget, Repeat: A Framework for Continual Unlearning in Text-to-Image Diffusion Models
von: George, Naveen, et al.
Veröffentlicht: (2025)
von: George, Naveen, et al.
Veröffentlicht: (2025)
Denoising Multi-Beta VAE: Representation Learning for Disentanglement and Generation
von: Uppal, Anshuk, et al.
Veröffentlicht: (2025)
von: Uppal, Anshuk, et al.
Veröffentlicht: (2025)
Improving Classifier-Free Guidance in Masked Diffusion: Low-Dim Theoretical Insights with High-Dim Impact
von: Rojas, Kevin, et al.
Veröffentlicht: (2025)
von: Rojas, Kevin, et al.
Veröffentlicht: (2025)
Transformed Low-rank Adaptation via Tensor Decomposition and Its Applications to Text-to-image Models
von: Tao, Zerui, et al.
Veröffentlicht: (2025)
von: Tao, Zerui, et al.
Veröffentlicht: (2025)
$\textit{Jump Your Steps}$: Optimizing Sampling Schedule of Discrete Diffusion Models
von: Park, Yong-Hyun, et al.
Veröffentlicht: (2024)
von: Park, Yong-Hyun, et al.
Veröffentlicht: (2024)
CMT: Mid-Training for Efficient Learning of Consistency, Mean Flow, and Flow Map Models
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
HQ-VAE: Hierarchical Discrete Representation Learning with Variational Bayes
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
SoundCTM: Unifying Score-based and Consistency Models for Full-band Text-to-Sound Generation
von: Saito, Koichi, et al.
Veröffentlicht: (2024)
von: Saito, Koichi, et al.
Veröffentlicht: (2024)
The Principles of Diffusion Models
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2025)
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2025)
BigVSAN: Enhancing GAN-based Neural Vocoders with Slicing Adversarial Network
von: Shibuya, Takashi, et al.
Veröffentlicht: (2023)
von: Shibuya, Takashi, et al.
Veröffentlicht: (2023)
PAVAS: Physics-Aware Video-to-Audio Synthesis
von: Hyun-Bin, Oh, et al.
Veröffentlicht: (2025)
von: Hyun-Bin, Oh, et al.
Veröffentlicht: (2025)
Bellman Diffusion: Generative Modeling as Learning a Linear Operator in the Distribution Space
von: Li, Yangming, et al.
Veröffentlicht: (2024)
von: Li, Yangming, et al.
Veröffentlicht: (2024)
Understanding and Accelerating the Training of Masked Diffusion Language Models
von: Hong, Chunsan, et al.
Veröffentlicht: (2026)
von: Hong, Chunsan, et al.
Veröffentlicht: (2026)
Efficiency without Compromise: CLIP-aided Text-to-Image GANs with Increased Diversity
von: Kobayashi, Yuya, et al.
Veröffentlicht: (2025)
von: Kobayashi, Yuya, et al.
Veröffentlicht: (2025)
MoLA: Motion Generation and Editing with Latent Diffusion Enhanced by Adversarial Training
von: Uchida, Kengo, et al.
Veröffentlicht: (2024)
von: Uchida, Kengo, et al.
Veröffentlicht: (2024)
Distillation of Discrete Diffusion through Dimensional Correlations
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2024)
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2024)
Demystifying MaskGIT Sampler and Beyond: Adaptive Order Selection in Masked Diffusion
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2025)
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2025)
DisMix: Disentangling Mixtures of Musical Instruments for Source-level Pitch and Timbre Manipulation
von: Luo, Yin-Jyun, et al.
Veröffentlicht: (2024)
von: Luo, Yin-Jyun, et al.
Veröffentlicht: (2024)
MeanFlow Transformers with Representation Autoencoders
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
Forging and Removing Latent-Noise Diffusion Watermarks Using a Single Image
von: Jain, Anubhav, et al.
Veröffentlicht: (2025)
von: Jain, Anubhav, et al.
Veröffentlicht: (2025)
Concept-TRAK: Understanding how diffusion models learn concepts through concept-level attribution
von: Park, Yonghyun, et al.
Veröffentlicht: (2025)
von: Park, Yonghyun, et al.
Veröffentlicht: (2025)
DiffRoll: Diffusion-based Generative Music Transcription with Unsupervised Pretraining Capability
von: Cheuk, Kin Wai, et al.
Veröffentlicht: (2022)
von: Cheuk, Kin Wai, et al.
Veröffentlicht: (2022)
Blind Inverse Problem Solving Made Easy by Text-to-Image Latent Diffusion
von: Dontas, Michail, et al.
Veröffentlicht: (2024)
von: Dontas, Michail, et al.
Veröffentlicht: (2024)
Variable Bitrate Residual Vector Quantization for Audio Coding
von: Chae, Yunkee, et al.
Veröffentlicht: (2024)
von: Chae, Yunkee, et al.
Veröffentlicht: (2024)
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
von: Hiranaka, Ayano, et al.
Veröffentlicht: (2024)
von: Hiranaka, Ayano, et al.
Veröffentlicht: (2024)
TraSCE: Trajectory Steering for Concept Erasure
von: Jain, Anubhav, et al.
Veröffentlicht: (2024)
von: Jain, Anubhav, et al.
Veröffentlicht: (2024)
Classifier-Free Guidance inside the Attraction Basin May Cause Memorization
von: Jain, Anubhav, et al.
Veröffentlicht: (2024)
von: Jain, Anubhav, et al.
Veröffentlicht: (2024)
SequenceMatch: Imitation Learning for Autoregressive Sequence Modelling with Backtracking
von: Cundy, Chris, et al.
Veröffentlicht: (2023)
von: Cundy, Chris, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
GUDA: Counterfactual Group-wise Training Data Attribution for Diffusion Models via Unlearning
von: Murata, Naoki, et al.
Veröffentlicht: (2026) -
G2D2: Gradient-Guided Discrete Diffusion for Inverse Problem Solving
von: Murata, Naoki, et al.
Veröffentlicht: (2024) -
Improved Object-Centric Diffusion Learning with Registers and Contrastive Alignment
von: Nguyen, Bac, et al.
Veröffentlicht: (2026) -
A Unified View of Score-Based and Drifting Models
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2026) -
Consistency Trajectory Models: Learning Probability Flow ODE Trajectory of Diffusion
von: Kim, Dongjun, et al.
Veröffentlicht: (2023)