TLCM: Training-efficient Latent Consistency Model for Image Generation with 2-8 Steps
Fuente:
arXiv
Salvato in:
| Autori principali: | Xie, Qingsong, Liao, Zhenyi, Deng, Zhijie, chen, Chen, Lu, Haonan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improved Visual-Spatial Reasoning via R1-Zero-Like Training
di: Liao, Zhenyi, et al.
Pubblicazione: (2025)
di: Liao, Zhenyi, et al.
Pubblicazione: (2025)
UniCMs: A Unified Consistency Model For Efficient Multimodal Generation and Understanding
di: Xu, Chenkai, et al.
Pubblicazione: (2025)
di: Xu, Chenkai, et al.
Pubblicazione: (2025)
FaceScore: Benchmarking and Enhancing Face Quality in Human Generation
di: Liao, Zhenyi, et al.
Pubblicazione: (2024)
di: Liao, Zhenyi, et al.
Pubblicazione: (2024)
PainterNet: Adaptive Image Inpainting with Actual-Token Attention and Diverse Mask Control
di: Wang, Ruichen, et al.
Pubblicazione: (2024)
di: Wang, Ruichen, et al.
Pubblicazione: (2024)
LOVECon: Text-driven Training-Free Long Video Editing with ControlNet
di: Liao, Zhenyi, et al.
Pubblicazione: (2023)
di: Liao, Zhenyi, et al.
Pubblicazione: (2023)
APLA: Additional Perturbation for Latent Noise with Adversarial Training Enables Consistency
di: Yao, Yupu, et al.
Pubblicazione: (2023)
di: Yao, Yupu, et al.
Pubblicazione: (2023)
Efficient Text-driven Motion Generation via Latent Consistency Training
di: Hu, Mengxian, et al.
Pubblicazione: (2024)
di: Hu, Mengxian, et al.
Pubblicazione: (2024)
Layton: Latent Consistency Tokenizer for 1024-pixel Image Reconstruction and Generation by 256 Tokens
di: Xie, Qingsong, et al.
Pubblicazione: (2025)
di: Xie, Qingsong, et al.
Pubblicazione: (2025)
CookAnything: A Framework for Flexible and Consistent Multi-Step Recipe Image Generation
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
AdvDMD: Adversarial Reward Meets DMD For High-Quality Few-Step Generation
di: Wang, Xu, et al.
Pubblicazione: (2026)
di: Wang, Xu, et al.
Pubblicazione: (2026)
Consist-Retinex: One-Step Noise-Emphasized Consistency Training Accelerates High-Quality Retinex Enhancement
di: Xu, Jian, et al.
Pubblicazione: (2025)
di: Xu, Jian, et al.
Pubblicazione: (2025)
Advancing Text-to-3D Generation with Linearized Lookahead Variational Score Distillation
di: Lei, Yu, et al.
Pubblicazione: (2025)
di: Lei, Yu, et al.
Pubblicazione: (2025)
GvT: A Graph-based Vision Transformer with Talking-Heads Utilizing Sparsity, Trained from Scratch on Small Datasets
di: Shan, Dongjing, et al.
Pubblicazione: (2024)
di: Shan, Dongjing, et al.
Pubblicazione: (2024)
Bayesian Exploration of Pre-trained Models for Low-shot Image Classification
di: Miao, Yibo, et al.
Pubblicazione: (2024)
di: Miao, Yibo, et al.
Pubblicazione: (2024)
StereoDiffusion: Training-Free Stereo Image Generation Using Latent Diffusion Models
di: Wang, Lezhong, et al.
Pubblicazione: (2024)
di: Wang, Lezhong, et al.
Pubblicazione: (2024)
MCAD: Multi-teacher Cross-modal Alignment Distillation for efficient image-text retrieval
di: Lei, Youbo, et al.
Pubblicazione: (2023)
di: Lei, Youbo, et al.
Pubblicazione: (2023)
Degradation-Consistent Paired Training for Robust AI-Generated Image Detection
di: Yang, Zongyou, et al.
Pubblicazione: (2026)
di: Yang, Zongyou, et al.
Pubblicazione: (2026)
FissionVAE: Federated Non-IID Image Generation with Latent Space and Decoder Decomposition
di: Hu, Chen, et al.
Pubblicazione: (2024)
di: Hu, Chen, et al.
Pubblicazione: (2024)
Reward Guided Latent Consistency Distillation
di: Li, Jiachen, et al.
Pubblicazione: (2024)
di: Li, Jiachen, et al.
Pubblicazione: (2024)
SCott: Accelerating Diffusion Models with Stochastic Consistency Distillation
di: Liu, Hongjian, et al.
Pubblicazione: (2024)
di: Liu, Hongjian, et al.
Pubblicazione: (2024)
Enhance Multimodal Consistency and Coherence for Text-Image Plan Generation
di: Lu, Xiaoxin, et al.
Pubblicazione: (2025)
di: Lu, Xiaoxin, et al.
Pubblicazione: (2025)
MergeVQ: A Unified Framework for Visual Generation and Representation with Disentangled Token Merging and Quantization
di: Li, Siyuan, et al.
Pubblicazione: (2025)
di: Li, Siyuan, et al.
Pubblicazione: (2025)
LatentEdit: Adaptive Latent Control for Consistent Semantic Editing
di: Liu, Siyi, et al.
Pubblicazione: (2025)
di: Liu, Siyi, et al.
Pubblicazione: (2025)
StorySync: Training-Free Subject Consistency in Text-to-Image Generation via Region Harmonization
di: Gaur, Gopalji, et al.
Pubblicazione: (2025)
di: Gaur, Gopalji, et al.
Pubblicazione: (2025)
Trajectory Consistency for One-Step Generation on Euler Mean Flows
di: Li, Zhiqi, et al.
Pubblicazione: (2026)
di: Li, Zhiqi, et al.
Pubblicazione: (2026)
Training-Free Consistent Text-to-Image Generation
di: Tewel, Yoad, et al.
Pubblicazione: (2024)
di: Tewel, Yoad, et al.
Pubblicazione: (2024)
MUG-V 10B: High-efficiency Training Pipeline for Large Video Generation Models
di: Zhang, Yongshun, et al.
Pubblicazione: (2025)
di: Zhang, Yongshun, et al.
Pubblicazione: (2025)
OneActor: Consistent Character Generation via Cluster-Conditioned Guidance
di: Wang, Jiahao, et al.
Pubblicazione: (2024)
di: Wang, Jiahao, et al.
Pubblicazione: (2024)
Calibrating Biased Distribution in VFM-derived Latent Space via Cross-Domain Geometric Consistency
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
H2VU-Benchmark: A Comprehensive Benchmark for Hierarchical Holistic Video Understanding
di: Wu, Qi, et al.
Pubblicazione: (2025)
di: Wu, Qi, et al.
Pubblicazione: (2025)
Latent Action Control for Reasoning-Guided Unified Image Generation
di: Zhai, Fuxiang, et al.
Pubblicazione: (2026)
di: Zhai, Fuxiang, et al.
Pubblicazione: (2026)
A Unified Understanding of Adversarial Vulnerability Regarding Unimodal Models and Vision-Language Pre-training Models
di: Zheng, Haonan, et al.
Pubblicazione: (2024)
di: Zheng, Haonan, et al.
Pubblicazione: (2024)
Learning Patient-Specific Disease Dynamics with Latent Flow Matching for Longitudinal Imaging Generation
di: Chen, Hao, et al.
Pubblicazione: (2025)
di: Chen, Hao, et al.
Pubblicazione: (2025)
Diffusion Adversarial Post-Training for One-Step Video Generation
di: Lin, Shanchuan, et al.
Pubblicazione: (2025)
di: Lin, Shanchuan, et al.
Pubblicazione: (2025)
How to Trace Latent Generative Model Generated Images without Artificial Watermark?
di: Wang, Zhenting, et al.
Pubblicazione: (2024)
di: Wang, Zhenting, et al.
Pubblicazione: (2024)
Detecting AI-Generated Video via Frame Consistency
di: Ma, Long, et al.
Pubblicazione: (2024)
di: Ma, Long, et al.
Pubblicazione: (2024)
Chain-of-Restoration: Multi-Task Image Restoration Models are Zero-Shot Step-by-Step Universal Image Restorers
di: Cao, Jin, et al.
Pubblicazione: (2024)
di: Cao, Jin, et al.
Pubblicazione: (2024)
Thinking with Generated Images
di: Chern, Ethan, et al.
Pubblicazione: (2025)
di: Chern, Ethan, et al.
Pubblicazione: (2025)
Latent Guidance in Diffusion Models for Perceptual Evaluations
di: Saini, Shreshth, et al.
Pubblicazione: (2025)
di: Saini, Shreshth, et al.
Pubblicazione: (2025)
STEVE: A Step Verification Pipeline for Computer-use Agent Training
di: Lu, Fanbin, et al.
Pubblicazione: (2025)
di: Lu, Fanbin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Improved Visual-Spatial Reasoning via R1-Zero-Like Training
di: Liao, Zhenyi, et al.
Pubblicazione: (2025) -
UniCMs: A Unified Consistency Model For Efficient Multimodal Generation and Understanding
di: Xu, Chenkai, et al.
Pubblicazione: (2025) -
FaceScore: Benchmarking and Enhancing Face Quality in Human Generation
di: Liao, Zhenyi, et al.
Pubblicazione: (2024) -
PainterNet: Adaptive Image Inpainting with Actual-Token Attention and Diverse Mask Control
di: Wang, Ruichen, et al.
Pubblicazione: (2024) -
LOVECon: Text-driven Training-Free Long Video Editing with ControlNet
di: Liao, Zhenyi, et al.
Pubblicazione: (2023)