Adv-KD: Adversarial Knowledge Distillation for Faster Diffusion Sampling
Fuente:
arXiv
Salvato in:
| Autori principali: | Mekonnen, Kidist Amde, Dall'Asen, Nicola, Rota, Paolo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Conditioning GAN Without Training Dataset
di: Mekonnen, Kidist Amde
Pubblicazione: (2024)
di: Mekonnen, Kidist Amde
Pubblicazione: (2024)
Lightning Fast Video Anomaly Detection via Adversarial Knowledge Distillation
di: Croitoru, Florinel-Alin, et al.
Pubblicazione: (2022)
di: Croitoru, Florinel-Alin, et al.
Pubblicazione: (2022)
Vision-Language Meets the Skeleton: Progressively Distillation with Cross-Modal Knowledge for 3D Action Representation Learning
di: Chen, Yang, et al.
Pubblicazione: (2024)
di: Chen, Yang, et al.
Pubblicazione: (2024)
MIL-PF: Multiple Instance Learning on Precomputed Features for Mammography Classification
di: Jovišić, Nikola, et al.
Pubblicazione: (2026)
di: Jovišić, Nikola, et al.
Pubblicazione: (2026)
Low-Resolution Face Recognition via Adaptable Instance-Relation Distillation
di: Shi, Ruixin, et al.
Pubblicazione: (2024)
di: Shi, Ruixin, et al.
Pubblicazione: (2024)
Low-Resolution Object Recognition with Cross-Resolution Relational Contrastive Distillation
di: Zhang, Kangkai, et al.
Pubblicazione: (2024)
di: Zhang, Kangkai, et al.
Pubblicazione: (2024)
PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification
di: Luo, Qiuming, et al.
Pubblicazione: (2026)
di: Luo, Qiuming, et al.
Pubblicazione: (2026)
COMODO: Cross-Modal Video-to-IMU Distillation for Efficient Egocentric Human Activity Recognition
di: Chen, Baiyu, et al.
Pubblicazione: (2025)
di: Chen, Baiyu, et al.
Pubblicazione: (2025)
Zoomed In, Diffused Out: Towards Local Degradation-Aware Multi-Diffusion for Extreme Image Super-Resolution
di: Moser, Brian B., et al.
Pubblicazione: (2024)
di: Moser, Brian B., et al.
Pubblicazione: (2024)
Understanding the Fine-Grained Knowledge Capabilities of Vision-Language Models
di: Ghosh, Dhruba, et al.
Pubblicazione: (2026)
di: Ghosh, Dhruba, et al.
Pubblicazione: (2026)
Diffusion Model-Based Video Editing: A Survey
di: Sun, Wenhao, et al.
Pubblicazione: (2024)
di: Sun, Wenhao, et al.
Pubblicazione: (2024)
Diffusion Models, Image Super-Resolution And Everything: A Survey
di: Moser, Brian B., et al.
Pubblicazione: (2024)
di: Moser, Brian B., et al.
Pubblicazione: (2024)
IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video Generation
di: Lin, Yuanze, et al.
Pubblicazione: (2025)
di: Lin, Yuanze, et al.
Pubblicazione: (2025)
RDPM: Solve Diffusion Probabilistic Models via Recurrent Token Prediction
di: Wu, Xiaoping, et al.
Pubblicazione: (2024)
di: Wu, Xiaoping, et al.
Pubblicazione: (2024)
Identity Preserving 3D Head Stylization with Multiview Score Distillation
di: Bilecen, Bahri Batuhan, et al.
Pubblicazione: (2024)
di: Bilecen, Bahri Batuhan, et al.
Pubblicazione: (2024)
TP-Blend: Textual-Prompt Attention Pairing for Precise Object-Style Blending in Diffusion Models
di: Jin, Xin, et al.
Pubblicazione: (2026)
di: Jin, Xin, et al.
Pubblicazione: (2026)
David and Goliath: Small One-step Model Beats Large Diffusion with Score Post-training
di: Luo, Weijian, et al.
Pubblicazione: (2024)
di: Luo, Weijian, et al.
Pubblicazione: (2024)
Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning
di: Chen, Weifeng, et al.
Pubblicazione: (2023)
di: Chen, Weifeng, et al.
Pubblicazione: (2023)
A Systematic Review on Long-Tailed Learning
di: Zhang, Chongsheng, et al.
Pubblicazione: (2024)
di: Zhang, Chongsheng, et al.
Pubblicazione: (2024)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
di: Park, Inkyu, et al.
Pubblicazione: (2023)
di: Park, Inkyu, et al.
Pubblicazione: (2023)
GoodDrag: Towards Good Practices for Drag Editing with Diffusion Models
di: Zhang, Zewei, et al.
Pubblicazione: (2024)
di: Zhang, Zewei, et al.
Pubblicazione: (2024)
MST-Distill: Mixture of Specialized Teachers for Cross-Modal Knowledge Distillation
di: Li, Hui, et al.
Pubblicazione: (2025)
di: Li, Hui, et al.
Pubblicazione: (2025)
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
di: Sun, Zeyi, et al.
Pubblicazione: (2024)
di: Sun, Zeyi, et al.
Pubblicazione: (2024)
FastCache: Fast Caching for Diffusion Transformer Through Learnable Linear Approximation
di: Liu, Dong, et al.
Pubblicazione: (2025)
di: Liu, Dong, et al.
Pubblicazione: (2025)
DisCoM-KD: Cross-Modal Knowledge Distillation via Disentanglement Representation and Adversarial Learning
di: Ienco, Dino, et al.
Pubblicazione: (2024)
di: Ienco, Dino, et al.
Pubblicazione: (2024)
Spatial Knowledge Graph-Guided Multimodal Synthesis
di: Xue, Yida, et al.
Pubblicazione: (2025)
di: Xue, Yida, et al.
Pubblicazione: (2025)
MoDA: Modulation Adapter for Fine-Grained Visual Grounding in Instructional MLLMs
di: Barrios, Wayner, et al.
Pubblicazione: (2025)
di: Barrios, Wayner, et al.
Pubblicazione: (2025)
Who Brings the Frisbee: Probing Hidden Hallucination Factors in Large Vision-Language Model via Causality Analysis
di: Huang, Po-Hsuan, et al.
Pubblicazione: (2024)
di: Huang, Po-Hsuan, et al.
Pubblicazione: (2024)
Using AI to Summarize US Presidential Campaign TV Advertisement Videos, 1952-2012
di: Breuer, Adam, et al.
Pubblicazione: (2025)
di: Breuer, Adam, et al.
Pubblicazione: (2025)
RMAdapter: Reconstruction-based Multi-Modal Adapter for Vision-Language Models
di: Lin, Xiang, et al.
Pubblicazione: (2025)
di: Lin, Xiang, et al.
Pubblicazione: (2025)
Meta-CoT: Enhancing Granularity and Generalization in Image Editing
di: Zhang, Shiyi, et al.
Pubblicazione: (2026)
di: Zhang, Shiyi, et al.
Pubblicazione: (2026)
Are We Making Progress in Multimodal Domain Generalization? A Comprehensive Benchmark Study
di: Dong, Hao, et al.
Pubblicazione: (2026)
di: Dong, Hao, et al.
Pubblicazione: (2026)
Latent Space Probing for Adult Content Detection in Video Generative Models
di: Khatri, Alizishaan, et al.
Pubblicazione: (2026)
di: Khatri, Alizishaan, et al.
Pubblicazione: (2026)
InteractiveVideo: User-Centric Controllable Video Generation with Synergistic Multimodal Instructions
di: Zhang, Yiyuan, et al.
Pubblicazione: (2024)
di: Zhang, Yiyuan, et al.
Pubblicazione: (2024)
MagicMotion: Controllable Video Generation with Dense-to-Sparse Trajectory Guidance
di: Li, Quanhao, et al.
Pubblicazione: (2025)
di: Li, Quanhao, et al.
Pubblicazione: (2025)
Deciphering Personalization: Towards Fine-Grained Explainability in Natural Language for Personalized Image Generation Models
di: Wang, Haoming, et al.
Pubblicazione: (2025)
di: Wang, Haoming, et al.
Pubblicazione: (2025)
Improving Visual Representation Alignment Generation with GRPO
di: Mo, Shentong, et al.
Pubblicazione: (2026)
di: Mo, Shentong, et al.
Pubblicazione: (2026)
FlashMotion: Few-Step Controllable Video Generation with Trajectory Guidance
di: Li, Quanhao, et al.
Pubblicazione: (2026)
di: Li, Quanhao, et al.
Pubblicazione: (2026)
Scalable Object Relation Encoding for Better 3D Spatial Reasoning in Large Language Models
di: Zhou, Shengli, et al.
Pubblicazione: (2026)
di: Zhou, Shengli, et al.
Pubblicazione: (2026)
Towards Multi-Task Multi-Modal Models: A Video Generative Perspective
di: Yu, Lijun
Pubblicazione: (2024)
di: Yu, Lijun
Pubblicazione: (2024)
Documenti analoghi
-
Conditioning GAN Without Training Dataset
di: Mekonnen, Kidist Amde
Pubblicazione: (2024) -
Lightning Fast Video Anomaly Detection via Adversarial Knowledge Distillation
di: Croitoru, Florinel-Alin, et al.
Pubblicazione: (2022) -
Vision-Language Meets the Skeleton: Progressively Distillation with Cross-Modal Knowledge for 3D Action Representation Learning
di: Chen, Yang, et al.
Pubblicazione: (2024) -
MIL-PF: Multiple Instance Learning on Precomputed Features for Mammography Classification
di: Jovišić, Nikola, et al.
Pubblicazione: (2026) -
Low-Resolution Face Recognition via Adaptable Instance-Relation Distillation
di: Shi, Ruixin, et al.
Pubblicazione: (2024)