Improved Object-Centric Diffusion Learning with Registers and Contrastive Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen, Bac, Takida, Yuhta, Murata, Naoki, Lai, Chieh-Hsin, Uesaka, Toshimitsu, Ermon, Stefano, Mitsufuji, Yuki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
G2D2: Gradient-Guided Discrete Diffusion for Inverse Problem Solving
von: Murata, Naoki, et al.
Veröffentlicht: (2024)
von: Murata, Naoki, et al.
Veröffentlicht: (2024)
A Unified View of Score-Based and Drifting Models
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2026)
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2026)
Improving Vector-Quantized Image Modeling with Latent Consistency-Matching Diffusion
von: Nguyen, Bac, et al.
Veröffentlicht: (2024)
von: Nguyen, Bac, et al.
Veröffentlicht: (2024)
GUDA: Counterfactual Group-wise Training Data Attribution for Diffusion Models via Unlearning
von: Murata, Naoki, et al.
Veröffentlicht: (2026)
von: Murata, Naoki, et al.
Veröffentlicht: (2026)
PaGoDA: Progressive Growing of a One-Step Generator from a Low-Resolution Diffusion Teacher
von: Kim, Dongjun, et al.
Veröffentlicht: (2024)
von: Kim, Dongjun, et al.
Veröffentlicht: (2024)
Consistency Trajectory Models: Learning Probability Flow ODE Trajectory of Diffusion
von: Kim, Dongjun, et al.
Veröffentlicht: (2023)
von: Kim, Dongjun, et al.
Veröffentlicht: (2023)
SONA: Learning Conditional, Unconditional, and Mismatching-Aware Discriminator
von: Takida, Yuhta, et al.
Veröffentlicht: (2025)
von: Takida, Yuhta, et al.
Veröffentlicht: (2025)
Weighted Point Set Embedding for Multimodal Contrastive Learning Toward Optimal Similarity Metric
von: Uesaka, Toshimitsu, et al.
Veröffentlicht: (2024)
von: Uesaka, Toshimitsu, et al.
Veröffentlicht: (2024)
Noise Scheduling as Information-Guided Allocation in Diffusion Training
von: Raya, Gabriel, et al.
Veröffentlicht: (2026)
von: Raya, Gabriel, et al.
Veröffentlicht: (2026)
Denoising Multi-Beta VAE: Representation Learning for Disentanglement and Generation
von: Uppal, Anshuk, et al.
Veröffentlicht: (2025)
von: Uppal, Anshuk, et al.
Veröffentlicht: (2025)
PAVAS: Physics-Aware Video-to-Audio Synthesis
von: Hyun-Bin, Oh, et al.
Veröffentlicht: (2025)
von: Hyun-Bin, Oh, et al.
Veröffentlicht: (2025)
HQ-VAE: Hierarchical Discrete Representation Learning with Variational Bayes
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
SAN: Inducing Metrizability of GAN with Discriminative Normalized Linear Layer
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
von: Takida, Yuhta, et al.
Veröffentlicht: (2023)
VCT: Training Consistency Models with Variational Noise Coupling
von: Silvestri, Gianluigi, et al.
Veröffentlicht: (2025)
von: Silvestri, Gianluigi, et al.
Veröffentlicht: (2025)
$\textit{Jump Your Steps}$: Optimizing Sampling Schedule of Discrete Diffusion Models
von: Park, Yong-Hyun, et al.
Veröffentlicht: (2024)
von: Park, Yong-Hyun, et al.
Veröffentlicht: (2024)
MoLA: Motion Generation and Editing with Latent Diffusion Enhanced by Adversarial Training
von: Uchida, Kengo, et al.
Veröffentlicht: (2024)
von: Uchida, Kengo, et al.
Veröffentlicht: (2024)
CMT: Mid-Training for Efficient Learning of Consistency, Mean Flow, and Flow Map Models
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
Efficiency without Compromise: CLIP-aided Text-to-Image GANs with Increased Diversity
von: Kobayashi, Yuya, et al.
Veröffentlicht: (2025)
von: Kobayashi, Yuya, et al.
Veröffentlicht: (2025)
Forging and Removing Latent-Noise Diffusion Watermarks Using a Single Image
von: Jain, Anubhav, et al.
Veröffentlicht: (2025)
von: Jain, Anubhav, et al.
Veröffentlicht: (2025)
Theoretical Refinement of CLIP by Utilizing Linear Structure of Optimal Similarity
von: Yoshida, Naoki, et al.
Veröffentlicht: (2025)
von: Yoshida, Naoki, et al.
Veröffentlicht: (2025)
Concept-TRAK: Understanding how diffusion models learn concepts through concept-level attribution
von: Park, Yonghyun, et al.
Veröffentlicht: (2025)
von: Park, Yonghyun, et al.
Veröffentlicht: (2025)
MeanFlow Transformers with Representation Autoencoders
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
von: Hu, Zheyuan, et al.
Veröffentlicht: (2025)
Improving Classifier-Free Guidance in Masked Diffusion: Low-Dim Theoretical Insights with High-Dim Impact
von: Rojas, Kevin, et al.
Veröffentlicht: (2025)
von: Rojas, Kevin, et al.
Veröffentlicht: (2025)
Distill, Forget, Repeat: A Framework for Continual Unlearning in Text-to-Image Diffusion Models
von: George, Naveen, et al.
Veröffentlicht: (2025)
von: George, Naveen, et al.
Veröffentlicht: (2025)
Transformed Low-rank Adaptation via Tensor Decomposition and Its Applications to Text-to-image Models
von: Tao, Zerui, et al.
Veröffentlicht: (2025)
von: Tao, Zerui, et al.
Veröffentlicht: (2025)
GenWarp: Single Image to Novel Views with Semantic-Preserving Generative Warping
von: Seo, Junyoung, et al.
Veröffentlicht: (2024)
von: Seo, Junyoung, et al.
Veröffentlicht: (2024)
TraSCE: Trajectory Steering for Concept Erasure
von: Jain, Anubhav, et al.
Veröffentlicht: (2024)
von: Jain, Anubhav, et al.
Veröffentlicht: (2024)
Classifier-Free Guidance inside the Attraction Basin May Cause Memorization
von: Jain, Anubhav, et al.
Veröffentlicht: (2024)
von: Jain, Anubhav, et al.
Veröffentlicht: (2024)
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
von: Hiranaka, Ayano, et al.
Veröffentlicht: (2024)
von: Hiranaka, Ayano, et al.
Veröffentlicht: (2024)
BigVSAN: Enhancing GAN-based Neural Vocoders with Slicing Adversarial Network
von: Shibuya, Takashi, et al.
Veröffentlicht: (2023)
von: Shibuya, Takashi, et al.
Veröffentlicht: (2023)
The Principles of Diffusion Models
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2025)
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2025)
Bellman Diffusion: Generative Modeling as Learning a Linear Operator in the Distribution Space
von: Li, Yangming, et al.
Veröffentlicht: (2024)
von: Li, Yangming, et al.
Veröffentlicht: (2024)
Blind Inverse Problem Solving Made Easy by Text-to-Image Latent Diffusion
von: Dontas, Michail, et al.
Veröffentlicht: (2024)
von: Dontas, Michail, et al.
Veröffentlicht: (2024)
SoundCTM: Unifying Score-based and Consistency Models for Full-band Text-to-Sound Generation
von: Saito, Koichi, et al.
Veröffentlicht: (2024)
von: Saito, Koichi, et al.
Veröffentlicht: (2024)
Understanding and Accelerating the Training of Masked Diffusion Language Models
von: Hong, Chunsan, et al.
Veröffentlicht: (2026)
von: Hong, Chunsan, et al.
Veröffentlicht: (2026)
DisMix: Disentangling Mixtures of Musical Instruments for Source-level Pitch and Timbre Manipulation
von: Luo, Yin-Jyun, et al.
Veröffentlicht: (2024)
von: Luo, Yin-Jyun, et al.
Veröffentlicht: (2024)
Distillation of Discrete Diffusion through Dimensional Correlations
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2024)
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2024)
A Survey on Diffusion Models for Inverse Problems
von: Daras, Giannis, et al.
Veröffentlicht: (2024)
von: Daras, Giannis, et al.
Veröffentlicht: (2024)
Beyond Pairwise Preferences: Listwise Reward-Aware Alignment for Diffusion Models
von: Wang, Austin, et al.
Veröffentlicht: (2026)
von: Wang, Austin, et al.
Veröffentlicht: (2026)
Demystifying MaskGIT Sampler and Beyond: Adaptive Order Selection in Masked Diffusion
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2025)
von: Hayakawa, Satoshi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
G2D2: Gradient-Guided Discrete Diffusion for Inverse Problem Solving
von: Murata, Naoki, et al.
Veröffentlicht: (2024) -
A Unified View of Score-Based and Drifting Models
von: Lai, Chieh-Hsin, et al.
Veröffentlicht: (2026) -
Improving Vector-Quantized Image Modeling with Latent Consistency-Matching Diffusion
von: Nguyen, Bac, et al.
Veröffentlicht: (2024) -
GUDA: Counterfactual Group-wise Training Data Attribution for Diffusion Models via Unlearning
von: Murata, Naoki, et al.
Veröffentlicht: (2026) -
PaGoDA: Progressive Growing of a One-Step Generator from a Low-Resolution Diffusion Teacher
von: Kim, Dongjun, et al.
Veröffentlicht: (2024)