Phased Consistency Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Fu-Yun, Huang, Zhaoyang, Bergman, Alexander William, Shen, Dazhong, Gao, Peng, Lingelbach, Michael, Sun, Keqiang, Bian, Weikang, Song, Guanglu, Liu, Yu, Wang, Xiaogang, Li, Hongsheng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data
by: Wang, Fu-Yun, et al.
Published: (2024)
by: Wang, Fu-Yun, et al.
Published: (2024)
Be-Your-Outpainter: Mastering Video Outpainting through Input-Specific Adaptation
by: Wang, Fu-Yun, et al.
Published: (2024)
by: Wang, Fu-Yun, et al.
Published: (2024)
Rethinking the Spatial Inconsistency in Classifier-Free Diffusion Guidance
by: Shen, Dazhong, et al.
Published: (2024)
by: Shen, Dazhong, et al.
Published: (2024)
GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking
by: Bian, Weikang, et al.
Published: (2025)
by: Bian, Weikang, et al.
Published: (2025)
Deep Reward Supervisions for Tuning Text-to-Image Diffusion Models
by: Wu, Xiaoshi, et al.
Published: (2024)
by: Wu, Xiaoshi, et al.
Published: (2024)
BlinkVision: A Benchmark for Optical Flow, Scene Flow and Point Tracking Estimation using RGB Frames and Events
by: Li, Yijin, et al.
Published: (2024)
by: Li, Yijin, et al.
Published: (2024)
CoMat: Aligning Text-to-Image Diffusion Model with Image-to-Text Concept Matching
by: Jiang, Dongzhi, et al.
Published: (2024)
by: Jiang, Dongzhi, et al.
Published: (2024)
EasyRef: Omni-Generalized Group Image Reference for Diffusion Models via Multimodal LLM
by: Zong, Zhuofan, et al.
Published: (2024)
by: Zong, Zhuofan, et al.
Published: (2024)
MoVA: Adapting Mixture of Vision Experts to Multimodal Context
by: Zong, Zhuofan, et al.
Published: (2024)
by: Zong, Zhuofan, et al.
Published: (2024)
Real-time One-Step Diffusion-based Expressive Portrait Videos Generation
by: Guo, Hanzhong, et al.
Published: (2024)
by: Guo, Hanzhong, et al.
Published: (2024)
Diffusion-NPO: Negative Preference Optimization for Better Preference Aligned Generation of Diffusion Models
by: Wang, Fu-Yun, et al.
Published: (2025)
by: Wang, Fu-Yun, et al.
Published: (2025)
Stable Consistency Tuning: Understanding and Improving Consistency Models
by: Wang, Fu-Yun, et al.
Published: (2024)
by: Wang, Fu-Yun, et al.
Published: (2024)
Motion-I2V: Consistent and Controllable Image-to-Video Generation with Explicit Motion Modeling
by: Shi, Xiaoyu, et al.
Published: (2024)
by: Shi, Xiaoyu, et al.
Published: (2024)
Self-NPO: Data-Free Diffusion Model Enhancement via Truncated Diffusion Fine-Tuning
by: Wang, Fu-Yun, et al.
Published: (2025)
by: Wang, Fu-Yun, et al.
Published: (2025)
Rectified Diffusion: Straightness Is Not Your Need in Rectified Flow
by: Wang, Fu-Yun, et al.
Published: (2024)
by: Wang, Fu-Yun, et al.
Published: (2024)
Application of Data Mining and Machine Learning Methods to Industrial Heat Treatment Processes for Hardness Prediction
by: Lingelbach, Yannick
Published: (2024)
by: Lingelbach, Yannick
Published: (2024)
RelightMaster: Precise Video Relighting with Multi-plane Light Images
by: Bian, Weikang, et al.
Published: (2025)
by: Bian, Weikang, et al.
Published: (2025)
UniFormer: Unifying Convolution and Self-attention for Visual Recognition
by: Li, Kunchang, et al.
Published: (2022)
by: Li, Kunchang, et al.
Published: (2022)
ADT: Tuning Diffusion Models with Adversarial Supervision
by: Shen, Dazhong, et al.
Published: (2025)
by: Shen, Dazhong, et al.
Published: (2025)
Exploring the Role of Large Language Models in Prompt Encoding for Diffusion Models
by: Ma, Bingqi, et al.
Published: (2024)
by: Ma, Bingqi, et al.
Published: (2024)
A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose Embedding
by: Dong, Yitong, et al.
Published: (2024)
by: Dong, Yitong, et al.
Published: (2024)
MMEmb-R1: Reasoning-Enhanced Multimodal Embedding with Pair-Aware Selection and Adaptive Control
by: Wang, Yuchi, et al.
Published: (2026)
by: Wang, Yuchi, et al.
Published: (2026)
FouriScale: A Frequency Perspective on Training-Free High-Resolution Image Synthesis
by: Huang, Linjiang, et al.
Published: (2024)
by: Huang, Linjiang, et al.
Published: (2024)
AtomDisc: An Atom-level Tokenizer that Boosts Molecular LLMs and Reveals Structure--Property Associations
by: Zhang, Mingxu, et al.
Published: (2025)
by: Zhang, Mingxu, et al.
Published: (2025)
NGTM: Substructure-based Neural Graph Topic Model for Interpretable Graph Generation
by: Zhuang, Yuanxin, et al.
Published: (2025)
by: Zhuang, Yuanxin, et al.
Published: (2025)
MolEditRL: Structure-Preserving Molecular Editing via Discrete Diffusion and Reinforcement Learning
by: Zhuang, Yuanxin, et al.
Published: (2025)
by: Zhuang, Yuanxin, et al.
Published: (2025)
High-Fidelity Diffusion Face Swapping with ID-Constrained Facial Conditioning
by: He, Dailan, et al.
Published: (2025)
by: He, Dailan, et al.
Published: (2025)
HPSv3: Towards Wide-Spectrum Human Preference Score
by: Ma, Yuhang, et al.
Published: (2025)
by: Ma, Yuhang, et al.
Published: (2025)
Probability-Consistent Preference Optimization for Enhanced LLM Reasoning
by: Yang, Yunqiao, et al.
Published: (2025)
by: Yang, Yunqiao, et al.
Published: (2025)
Salt: Self-Consistent Distribution Matching with Cache-Aware Training for Fast Video Generation
by: Ge, Xingtong, et al.
Published: (2026)
by: Ge, Xingtong, et al.
Published: (2026)
Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning
by: Shao, Hao, et al.
Published: (2024)
by: Shao, Hao, et al.
Published: (2024)
See Further When Clear: Curriculum Consistency Model
by: Liu, Yunpeng, et al.
Published: (2024)
by: Liu, Yunpeng, et al.
Published: (2024)
Clean-image Backdoor Attacks
by: Rong, Dazhong, et al.
Published: (2024)
by: Rong, Dazhong, et al.
Published: (2024)
Ponymation: Learning Articulated 3D Animal Motions from Unlabeled Online Videos
by: Sun, Keqiang, et al.
Published: (2023)
by: Sun, Keqiang, et al.
Published: (2023)
ChemATP: A Training-Free Chemical Reasoning Framework for Large Language Models
by: Zhang, Mingxu, et al.
Published: (2025)
by: Zhang, Mingxu, et al.
Published: (2025)
IRDS: Interpretable RLVR Data Selection via Verifier-Coupled Sparse Autoencoder Coverage
by: Li, Yuhan, et al.
Published: (2026)
by: Li, Yuhan, et al.
Published: (2026)
Leveraging Static Relationships for Intra-Type and Inter-Type Message Passing in Video Question Answering
by: Liang, Lili, et al.
Published: (2025)
by: Liang, Lili, et al.
Published: (2025)
Optimal linear-payment auction design with aftermarket collaboration
by: Wang, Dazhong, et al.
Published: (2026)
by: Wang, Dazhong, et al.
Published: (2026)
ReinDriveGen: Reinforcement Post-Training for Out-of-Distribution Driving Scene Generation
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Low-Light Video Enhancement via Spatial-Temporal Consistent Decomposition
by: Xu, Xiaogang, et al.
Published: (2024)
by: Xu, Xiaogang, et al.
Published: (2024)
Similar Items
-
AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data
by: Wang, Fu-Yun, et al.
Published: (2024) -
Be-Your-Outpainter: Mastering Video Outpainting through Input-Specific Adaptation
by: Wang, Fu-Yun, et al.
Published: (2024) -
Rethinking the Spatial Inconsistency in Classifier-Free Diffusion Guidance
by: Shen, Dazhong, et al.
Published: (2024) -
GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking
by: Bian, Weikang, et al.
Published: (2025) -
Deep Reward Supervisions for Tuning Text-to-Image Diffusion Models
by: Wu, Xiaoshi, et al.
Published: (2024)