SAN: Inducing Metrizability of GAN with Discriminative Normalized Linear Layer
Fuente:
arXiv
Salvato in:
| Autori principali: | Takida, Yuhta, Imaizumi, Masaaki, Shibuya, Takashi, Lai, Chieh-Hsin, Uesaka, Toshimitsu, Murata, Naoki, Mitsufuji, Yuki |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SONA: Learning Conditional, Unconditional, and Mismatching-Aware Discriminator
di: Takida, Yuhta, et al.
Pubblicazione: (2025)
di: Takida, Yuhta, et al.
Pubblicazione: (2025)
Weighted Point Set Embedding for Multimodal Contrastive Learning Toward Optimal Similarity Metric
di: Uesaka, Toshimitsu, et al.
Pubblicazione: (2024)
di: Uesaka, Toshimitsu, et al.
Pubblicazione: (2024)
Improving Vector-Quantized Image Modeling with Latent Consistency-Matching Diffusion
di: Nguyen, Bac, et al.
Pubblicazione: (2024)
di: Nguyen, Bac, et al.
Pubblicazione: (2024)
BigVSAN: Enhancing GAN-based Neural Vocoders with Slicing Adversarial Network
di: Shibuya, Takashi, et al.
Pubblicazione: (2023)
di: Shibuya, Takashi, et al.
Pubblicazione: (2023)
GUDA: Counterfactual Group-wise Training Data Attribution for Diffusion Models via Unlearning
di: Murata, Naoki, et al.
Pubblicazione: (2026)
di: Murata, Naoki, et al.
Pubblicazione: (2026)
G2D2: Gradient-Guided Discrete Diffusion for Inverse Problem Solving
di: Murata, Naoki, et al.
Pubblicazione: (2024)
di: Murata, Naoki, et al.
Pubblicazione: (2024)
A Unified View of Score-Based and Drifting Models
di: Lai, Chieh-Hsin, et al.
Pubblicazione: (2026)
di: Lai, Chieh-Hsin, et al.
Pubblicazione: (2026)
Theoretical Refinement of CLIP by Utilizing Linear Structure of Optimal Similarity
di: Yoshida, Naoki, et al.
Pubblicazione: (2025)
di: Yoshida, Naoki, et al.
Pubblicazione: (2025)
Improved Object-Centric Diffusion Learning with Registers and Contrastive Alignment
di: Nguyen, Bac, et al.
Pubblicazione: (2026)
di: Nguyen, Bac, et al.
Pubblicazione: (2026)
PaGoDA: Progressive Growing of a One-Step Generator from a Low-Resolution Diffusion Teacher
di: Kim, Dongjun, et al.
Pubblicazione: (2024)
di: Kim, Dongjun, et al.
Pubblicazione: (2024)
HQ-VAE: Hierarchical Discrete Representation Learning with Variational Bayes
di: Takida, Yuhta, et al.
Pubblicazione: (2023)
di: Takida, Yuhta, et al.
Pubblicazione: (2023)
Consistency Trajectory Models: Learning Probability Flow ODE Trajectory of Diffusion
di: Kim, Dongjun, et al.
Pubblicazione: (2023)
di: Kim, Dongjun, et al.
Pubblicazione: (2023)
Efficiency without Compromise: CLIP-aided Text-to-Image GANs with Increased Diversity
di: Kobayashi, Yuya, et al.
Pubblicazione: (2025)
di: Kobayashi, Yuya, et al.
Pubblicazione: (2025)
Denoising Multi-Beta VAE: Representation Learning for Disentanglement and Generation
di: Uppal, Anshuk, et al.
Pubblicazione: (2025)
di: Uppal, Anshuk, et al.
Pubblicazione: (2025)
Demystifying MaskGIT Sampler and Beyond: Adaptive Order Selection in Masked Diffusion
di: Hayakawa, Satoshi, et al.
Pubblicazione: (2025)
di: Hayakawa, Satoshi, et al.
Pubblicazione: (2025)
Distillation of Discrete Diffusion through Dimensional Correlations
di: Hayakawa, Satoshi, et al.
Pubblicazione: (2024)
di: Hayakawa, Satoshi, et al.
Pubblicazione: (2024)
SoundCTM: Unifying Score-based and Consistency Models for Full-band Text-to-Sound Generation
di: Saito, Koichi, et al.
Pubblicazione: (2024)
di: Saito, Koichi, et al.
Pubblicazione: (2024)
Transformed Low-rank Adaptation via Tensor Decomposition and Its Applications to Text-to-image Models
di: Tao, Zerui, et al.
Pubblicazione: (2025)
di: Tao, Zerui, et al.
Pubblicazione: (2025)
VCT: Training Consistency Models with Variational Noise Coupling
di: Silvestri, Gianluigi, et al.
Pubblicazione: (2025)
di: Silvestri, Gianluigi, et al.
Pubblicazione: (2025)
Distill, Forget, Repeat: A Framework for Continual Unlearning in Text-to-Image Diffusion Models
di: George, Naveen, et al.
Pubblicazione: (2025)
di: George, Naveen, et al.
Pubblicazione: (2025)
$\textit{Jump Your Steps}$: Optimizing Sampling Schedule of Discrete Diffusion Models
di: Park, Yong-Hyun, et al.
Pubblicazione: (2024)
di: Park, Yong-Hyun, et al.
Pubblicazione: (2024)
Improving Classifier-Free Guidance in Masked Diffusion: Low-Dim Theoretical Insights with High-Dim Impact
di: Rojas, Kevin, et al.
Pubblicazione: (2025)
di: Rojas, Kevin, et al.
Pubblicazione: (2025)
PAVAS: Physics-Aware Video-to-Audio Synthesis
di: Hyun-Bin, Oh, et al.
Pubblicazione: (2025)
di: Hyun-Bin, Oh, et al.
Pubblicazione: (2025)
MoLA: Motion Generation and Editing with Latent Diffusion Enhanced by Adversarial Training
di: Uchida, Kengo, et al.
Pubblicazione: (2024)
di: Uchida, Kengo, et al.
Pubblicazione: (2024)
Noise Scheduling as Information-Guided Allocation in Diffusion Training
di: Raya, Gabriel, et al.
Pubblicazione: (2026)
di: Raya, Gabriel, et al.
Pubblicazione: (2026)
TraSCE: Trajectory Steering for Concept Erasure
di: Jain, Anubhav, et al.
Pubblicazione: (2024)
di: Jain, Anubhav, et al.
Pubblicazione: (2024)
Classifier-Free Guidance inside the Attraction Basin May Cause Memorization
di: Jain, Anubhav, et al.
Pubblicazione: (2024)
di: Jain, Anubhav, et al.
Pubblicazione: (2024)
DisMix: Disentangling Mixtures of Musical Instruments for Source-level Pitch and Timbre Manipulation
di: Luo, Yin-Jyun, et al.
Pubblicazione: (2024)
di: Luo, Yin-Jyun, et al.
Pubblicazione: (2024)
Concept-TRAK: Understanding how diffusion models learn concepts through concept-level attribution
di: Park, Yonghyun, et al.
Pubblicazione: (2025)
di: Park, Yonghyun, et al.
Pubblicazione: (2025)
Forging and Removing Latent-Noise Diffusion Watermarks Using a Single Image
di: Jain, Anubhav, et al.
Pubblicazione: (2025)
di: Jain, Anubhav, et al.
Pubblicazione: (2025)
Understanding and Accelerating the Training of Masked Diffusion Language Models
di: Hong, Chunsan, et al.
Pubblicazione: (2026)
di: Hong, Chunsan, et al.
Pubblicazione: (2026)
MMDisCo: Multi-Modal Discriminator-Guided Cooperative Diffusion for Joint Audio and Video Generation
di: Hayakawa, Akio, et al.
Pubblicazione: (2024)
di: Hayakawa, Akio, et al.
Pubblicazione: (2024)
DiffRoll: Diffusion-based Generative Music Transcription with Unsupervised Pretraining Capability
di: Cheuk, Kin Wai, et al.
Pubblicazione: (2022)
di: Cheuk, Kin Wai, et al.
Pubblicazione: (2022)
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
di: Hiranaka, Ayano, et al.
Pubblicazione: (2024)
di: Hiranaka, Ayano, et al.
Pubblicazione: (2024)
Bellman Diffusion: Generative Modeling as Learning a Linear Operator in the Distribution Space
di: Li, Yangming, et al.
Pubblicazione: (2024)
di: Li, Yangming, et al.
Pubblicazione: (2024)
Approximation of Permutation Invariant Polynomials by Transformers: Efficient Construction in Column-Size
di: Takeshita, Naoki, et al.
Pubblicazione: (2025)
di: Takeshita, Naoki, et al.
Pubblicazione: (2025)
Extended Wasserstein-GAN Approach to Causal Distribution Learning: Density-Free Estimation and Minimax Optimality
di: Tamano, Shu, et al.
Pubblicazione: (2026)
di: Tamano, Shu, et al.
Pubblicazione: (2026)
GenWarp: Single Image to Novel Views with Semantic-Preserving Generative Warping
di: Seo, Junyoung, et al.
Pubblicazione: (2024)
di: Seo, Junyoung, et al.
Pubblicazione: (2024)
CMT: Mid-Training for Efficient Learning of Consistency, Mean Flow, and Flow Map Models
di: Hu, Zheyuan, et al.
Pubblicazione: (2025)
di: Hu, Zheyuan, et al.
Pubblicazione: (2025)
Coherent Audio-Visual Editing via Conditional Audio Generation Following Video Edits
di: Ishii, Masato, et al.
Pubblicazione: (2025)
di: Ishii, Masato, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SONA: Learning Conditional, Unconditional, and Mismatching-Aware Discriminator
di: Takida, Yuhta, et al.
Pubblicazione: (2025) -
Weighted Point Set Embedding for Multimodal Contrastive Learning Toward Optimal Similarity Metric
di: Uesaka, Toshimitsu, et al.
Pubblicazione: (2024) -
Improving Vector-Quantized Image Modeling with Latent Consistency-Matching Diffusion
di: Nguyen, Bac, et al.
Pubblicazione: (2024) -
BigVSAN: Enhancing GAN-based Neural Vocoders with Slicing Adversarial Network
di: Shibuya, Takashi, et al.
Pubblicazione: (2023) -
GUDA: Counterfactual Group-wise Training Data Attribution for Diffusion Models via Unlearning
di: Murata, Naoki, et al.
Pubblicazione: (2026)