Improved Training Technique for Latent Consistency Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Dao, Quan, Doan, Khanh, Liu, Di, Le, Trung, Metaxas, Dimitris |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learning
por: Tran, Quyen, et al.
Publicado: (2022)
por: Tran, Quyen, et al.
Publicado: (2022)
Self-Corrected Flow Distillation for Consistent One-Step and Few-Step Text-to-Image Generation
por: Dao, Quan, et al.
Publicado: (2024)
por: Dao, Quan, et al.
Publicado: (2024)
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model
por: Dao, Quan, et al.
Publicado: (2026)
por: Dao, Quan, et al.
Publicado: (2026)
Hiding and Recovering Knowledge in Text-to-Image Diffusion Models via Learnable Prompts
por: Bui, Anh, et al.
Publicado: (2024)
por: Bui, Anh, et al.
Publicado: (2024)
Test-Time Spectrum-Aware Latent Steering for Zero-Shot Generalization in Vision-Language Models
por: Dafnis, Konstantinos M., et al.
Publicado: (2025)
por: Dafnis, Konstantinos M., et al.
Publicado: (2025)
Erasing Undesirable Concepts in Diffusion Models with Adversarial Preservation
por: Bui, Anh, et al.
Publicado: (2024)
por: Bui, Anh, et al.
Publicado: (2024)
Steering Rectified Flow Models in the Vector Field for Controlled Image Generation
por: Patel, Maitreya, et al.
Publicado: (2024)
por: Patel, Maitreya, et al.
Publicado: (2024)
DiMSUM: Diffusion Mamba -- A Scalable and Unified Spatial-Frequency Method for Image Generation
por: Phung, Hao, et al.
Publicado: (2024)
por: Phung, Hao, et al.
Publicado: (2024)
How to Trace Latent Generative Model Generated Images without Artificial Watermark?
por: Wang, Zhenting, et al.
Publicado: (2024)
por: Wang, Zhenting, et al.
Publicado: (2024)
Aligning Human Knowledge with Visual Concepts Towards Explainable Medical Image Classification
por: Gao, Yunhe, et al.
Publicado: (2024)
por: Gao, Yunhe, et al.
Publicado: (2024)
Provably Improving Generalization of Few-Shot Models with Synthetic Data
por: Nguyen, Lan-Cuong, et al.
Publicado: (2025)
por: Nguyen, Lan-Cuong, et al.
Publicado: (2025)
Anatomy-VLM: A Fine-grained Vision-Language Model for Medical Interpretation
por: Gu, Difei, et al.
Publicado: (2025)
por: Gu, Difei, et al.
Publicado: (2025)
Improved Training Technique for Shortcut Models
por: Nguyen, Anh, et al.
Publicado: (2025)
por: Nguyen, Anh, et al.
Publicado: (2025)
DICE: Discrete Inversion Enabling Controllable Editing for Multinomial Diffusion and Masked Generative Models
por: He, Xiaoxiao, et al.
Publicado: (2024)
por: He, Xiaoxiao, et al.
Publicado: (2024)
DC-Merge: Improving Model Merging with Directional Consistency
por: Zhang, Han-Chen, et al.
Publicado: (2026)
por: Zhang, Han-Chen, et al.
Publicado: (2026)
Gradient-Aligned Calibration for Post-Training Quantization of Diffusion Models
por: Hoang, Dung Anh, et al.
Publicado: (2026)
por: Hoang, Dung Anh, et al.
Publicado: (2026)
DIAGNOSIS: Detecting Unauthorized Data Usages in Text-to-image Diffusion Models
por: Wang, Zhenting, et al.
Publicado: (2023)
por: Wang, Zhenting, et al.
Publicado: (2023)
Stable Consistency Tuning: Understanding and Improving Consistency Models
por: Wang, Fu-Yun, et al.
Publicado: (2024)
por: Wang, Fu-Yun, et al.
Publicado: (2024)
Test-Time Instance-Specific Parameter Composition: A New Paradigm for Adaptive Generative Modeling
por: Tran, Minh-Tuan, et al.
Publicado: (2026)
por: Tran, Minh-Tuan, et al.
Publicado: (2026)
Seeing Farther and Smarter: Value-Guided Multi-Path Reflection for VLM Policy Optimization
por: Yang, Yanting, et al.
Publicado: (2026)
por: Yang, Yanting, et al.
Publicado: (2026)
CryoSAMU: Enhancing 3D Cryo-EM Density Maps of Protein Structures at Intermediate Resolution with Structure-Aware Multimodal U-Nets
por: Zhang, Chenwei, et al.
Publicado: (2025)
por: Zhang, Chenwei, et al.
Publicado: (2025)
Connective Viewpoints of Signal-to-Noise Diffusion Models
por: Doan, Khanh, et al.
Publicado: (2024)
por: Doan, Khanh, et al.
Publicado: (2024)
RadAlign: Advancing Radiology Report Generation with Vision-Language Concept Alignment
por: Gu, Difei, et al.
Publicado: (2025)
por: Gu, Difei, et al.
Publicado: (2025)
Explicit Eigenvalue Regularization Improves Sharpness-Aware Minimization
por: Luo, Haocheng, et al.
Publicado: (2025)
por: Luo, Haocheng, et al.
Publicado: (2025)
Spectrum-Aware Parameter Efficient Fine-Tuning for Diffusion Models
por: Zhang, Xinxi, et al.
Publicado: (2024)
por: Zhang, Xinxi, et al.
Publicado: (2024)
LatentMan: Generating Consistent Animated Characters using Image Diffusion Models
por: Eldesokey, Abdelrahman, et al.
Publicado: (2023)
por: Eldesokey, Abdelrahman, et al.
Publicado: (2023)
VCT: Training Consistency Models with Variational Noise Coupling
por: Silvestri, Gianluigi, et al.
Publicado: (2025)
por: Silvestri, Gianluigi, et al.
Publicado: (2025)
The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models via Visual Information Steering
por: Li, Zhuowei, et al.
Publicado: (2025)
por: Li, Zhuowei, et al.
Publicado: (2025)
AutoEdit: Automatic Hyperparameter Tuning for Image Editing
por: Pham, Chau, et al.
Publicado: (2025)
por: Pham, Chau, et al.
Publicado: (2025)
Latent Domain Modeling Improves Robustness to Geographic Shifts
por: Crasto, Ruth, et al.
Publicado: (2025)
por: Crasto, Ruth, et al.
Publicado: (2025)
Storybooth: Training-free Multi-Subject Consistency for Improved Visual Storytelling
por: Singh, Jaskirat, et al.
Publicado: (2025)
por: Singh, Jaskirat, et al.
Publicado: (2025)
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering
por: Chatzoudis, Gerasimos, et al.
Publicado: (2025)
por: Chatzoudis, Gerasimos, et al.
Publicado: (2025)
Improving Visual Reasoning with Iterative Evidence Refinement
por: Shi, Zeru, et al.
Publicado: (2026)
por: Shi, Zeru, et al.
Publicado: (2026)
CheXPO-v2: Preference Optimization for Chest X-ray VLMs with Knowledge Graph Consistency
por: Liang, Xiao, et al.
Publicado: (2025)
por: Liang, Xiao, et al.
Publicado: (2025)
Simple ReFlow: Improved Techniques for Fast Flow Models
por: Kim, Beomsu, et al.
Publicado: (2024)
por: Kim, Beomsu, et al.
Publicado: (2024)
Leveraging Latent Diffusion Models for Training-Free In-Distribution Data Augmentation for Surface Defect Detection
por: Girella, Federico, et al.
Publicado: (2024)
por: Girella, Federico, et al.
Publicado: (2024)
Improving robustness to corruptions with multiplicative weight perturbations
por: Trinh, Trung, et al.
Publicado: (2024)
por: Trinh, Trung, et al.
Publicado: (2024)
MetaAug: Meta-Data Augmentation for Post-Training Quantization
por: Pham, Cuong, et al.
Publicado: (2024)
por: Pham, Cuong, et al.
Publicado: (2024)
KOPPA: Improving Prompt-based Continual Learning with Key-Query Orthogonal Projection and Prototype-based One-Versus-All
por: Tran, Quyen, et al.
Publicado: (2023)
por: Tran, Quyen, et al.
Publicado: (2023)
Disentanglement in T-space for Faster and Distributed Training of Diffusion Models with Fewer Latent-states
por: Gupta, Samarth, et al.
Publicado: (2025)
por: Gupta, Samarth, et al.
Publicado: (2025)
Ejemplares similares
-
An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learning
por: Tran, Quyen, et al.
Publicado: (2022) -
Self-Corrected Flow Distillation for Consistent One-Step and Few-Step Text-to-Image Generation
por: Dao, Quan, et al.
Publicado: (2024) -
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model
por: Dao, Quan, et al.
Publicado: (2026) -
Hiding and Recovering Knowledge in Text-to-Image Diffusion Models via Learnable Prompts
por: Bui, Anh, et al.
Publicado: (2024) -
Test-Time Spectrum-Aware Latent Steering for Zero-Shot Generalization in Vision-Language Models
por: Dafnis, Konstantinos M., et al.
Publicado: (2025)