A Closer Look at Time Steps is Worthy of Triple Speed-Up for Diffusion Model Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Kai, Shi, Mingjia, Zhou, Yukun, Li, Zekai, Yuan, Zhihang, Shang, Yuzhang, Peng, Xiaojiang, Zhang, Hanwang, You, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Faster Vision Mamba is Rebuilt in Minutes via Merged Token Re-training
von: Shi, Mingjia, et al.
Veröffentlicht: (2024)
von: Shi, Mingjia, et al.
Veröffentlicht: (2024)
GSQ-Tuning: Group-Shared Exponents Integer in Fully Quantized Training for LLMs On-Device Fine-tuning
von: Zhou, Sifan, et al.
Veröffentlicht: (2025)
von: Zhou, Sifan, et al.
Veröffentlicht: (2025)
A Closer Look at Bias and Chain-of-Thought Faithfulness of Large (Vision) Language Models
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2025)
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2025)
Learning to Look Closer: A New Instance-Wise Loss for Small Cerebral Lesion Segmentation
von: Bouteille, Luc, et al.
Veröffentlicht: (2025)
von: Bouteille, Luc, et al.
Veröffentlicht: (2025)
Self-supervised ControlNet with Spatio-Temporal Mamba for Real-world Video Super-resolution
von: Shi, Shijun, et al.
Veröffentlicht: (2025)
von: Shi, Shijun, et al.
Veröffentlicht: (2025)
ElasticFlow: One-Step Physics-Consistent Policy with Elastic Time Horizons for Language-Guided Manipulation
von: Chen, Kewei, et al.
Veröffentlicht: (2026)
von: Chen, Kewei, et al.
Veröffentlicht: (2026)
Saliency-Aware Multi-Route Thinking: Revisiting Vision-Language Reasoning
von: Shi, Mingjia, et al.
Veröffentlicht: (2026)
von: Shi, Mingjia, et al.
Veröffentlicht: (2026)
Time Step Generating: A Universal Synthesized Deepfake Image Detector
von: Zeng, Ziyue, et al.
Veröffentlicht: (2024)
von: Zeng, Ziyue, et al.
Veröffentlicht: (2024)
Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
Single-Step Reconstruction-Free Anomaly Detection and Segmentation via Diffusion Models
von: Moradi, Mehrdad, et al.
Veröffentlicht: (2025)
von: Moradi, Mehrdad, et al.
Veröffentlicht: (2025)
PoTS: Proof-of-Training-Steps for Backdoor Detection in Large Language Models
von: Seddik, Issam, et al.
Veröffentlicht: (2025)
von: Seddik, Issam, et al.
Veröffentlicht: (2025)
Step-Aware Residual-Guided Diffusion for EEG Spatial Super-Resolution
von: Liu, Hongjun, et al.
Veröffentlicht: (2025)
von: Liu, Hongjun, et al.
Veröffentlicht: (2025)
A High-Speed Capable Spherical Robot
von: Zhang, Bixuan, et al.
Veröffentlicht: (2025)
von: Zhang, Bixuan, et al.
Veröffentlicht: (2025)
A Closer Look at Model Collapse: From a Generalization-to-Memorization Perspective
von: Shi, Lianghe, et al.
Veröffentlicht: (2025)
von: Shi, Lianghe, et al.
Veröffentlicht: (2025)
Surprise Calibration for Better In-Context Learning
von: Tan, Zhihang, et al.
Veröffentlicht: (2025)
von: Tan, Zhihang, et al.
Veröffentlicht: (2025)
Super Apriel: One Checkpoint, Many Speeds
von: Labs, SLAM, et al.
Veröffentlicht: (2026)
von: Labs, SLAM, et al.
Veröffentlicht: (2026)
LiqD: A Dynamic Liquid Level Detection Model under Tricky Small Containers
von: Ma, Yukun, et al.
Veröffentlicht: (2024)
von: Ma, Yukun, et al.
Veröffentlicht: (2024)
Exploring Diffusion with Test-Time Training on Efficient Image Restoration
von: Lu, Rongchang, et al.
Veröffentlicht: (2025)
von: Lu, Rongchang, et al.
Veröffentlicht: (2025)
Improving Angular Speed Uniformity by Piecewise Radical Reparameterization
von: Hong, Hoon, et al.
Veröffentlicht: (2024)
von: Hong, Hoon, et al.
Veröffentlicht: (2024)
On Scaling Up 3D Gaussian Splatting Training
von: Zhao, Hexu, et al.
Veröffentlicht: (2024)
von: Zhao, Hexu, et al.
Veröffentlicht: (2024)
Learning to Seek Evidence: A Verifiable Reasoning Agent with Causal Faithfulness Analysis
von: Huang, Yuhang, et al.
Veröffentlicht: (2025)
von: Huang, Yuhang, et al.
Veröffentlicht: (2025)
Make Optimization Once and for All with Fine-grained Guidance
von: Shi, Mingjia, et al.
Veröffentlicht: (2025)
von: Shi, Mingjia, et al.
Veröffentlicht: (2025)
Self-Supervised Learning of Time Series Representation via Diffusion Process and Imputation-Interpolation-Forecasting Mask
von: Senane, Zineb, et al.
Veröffentlicht: (2024)
von: Senane, Zineb, et al.
Veröffentlicht: (2024)
Mixing It Up: Exploring Mixer Networks for Irregular Multivariate Time Series Forecasting
von: Klötergens, Christian, et al.
Veröffentlicht: (2025)
von: Klötergens, Christian, et al.
Veröffentlicht: (2025)
Inference-Time Loss-Guided Colour Preservation in Diffusion Sampling
von: Ahuja, Angad Singh, et al.
Veröffentlicht: (2026)
von: Ahuja, Angad Singh, et al.
Veröffentlicht: (2026)
Efficient Diffusion Training through Parallelization with Truncated Karhunen-Loève Expansion
von: Ren, Yumeng, et al.
Veröffentlicht: (2025)
von: Ren, Yumeng, et al.
Veröffentlicht: (2025)
Interactive Text-to-SQL Generation via Editable Step-by-Step Explanations
von: Tian, Yuan, et al.
Veröffentlicht: (2023)
von: Tian, Yuan, et al.
Veröffentlicht: (2023)
Looking at Model Debiasing through the Lens of Anomaly Detection
von: Pastore, Vito Paolo, et al.
Veröffentlicht: (2024)
von: Pastore, Vito Paolo, et al.
Veröffentlicht: (2024)
QuEST: Low-bit Diffusion Model Quantization via Efficient Selective Finetuning
von: Wang, Haoxuan, et al.
Veröffentlicht: (2024)
von: Wang, Haoxuan, et al.
Veröffentlicht: (2024)
APT: Adaptive Personalized Training for Diffusion Models with Limited Data
von: Chae, JungWoo, et al.
Veröffentlicht: (2025)
von: Chae, JungWoo, et al.
Veröffentlicht: (2025)
VQ4DiT: Efficient Post-Training Vector Quantization for Diffusion Transformers
von: Deng, Juncan, et al.
Veröffentlicht: (2024)
von: Deng, Juncan, et al.
Veröffentlicht: (2024)
A Scalable Pipeline Combining Procedural 3D Graphics and Guided Diffusion for Photorealistic Synthetic Training Data Generation in White Button Mushroom Segmentation
von: Károly, Artúr I., et al.
Veröffentlicht: (2025)
von: Károly, Artúr I., et al.
Veröffentlicht: (2025)
Look Further: Socially-Compliant Navigation System in Residential Buildings
von: Shiba, Akira, et al.
Veröffentlicht: (2026)
von: Shiba, Akira, et al.
Veröffentlicht: (2026)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
von: Xu, Zhe
Veröffentlicht: (2026)
von: Xu, Zhe
Veröffentlicht: (2026)
Fast Diffusion Model For Seismic Data Noise Attenuation
von: Peng, Junheng, et al.
Veröffentlicht: (2024)
von: Peng, Junheng, et al.
Veröffentlicht: (2024)
REPA Works Until It Doesn't: Early-Stopped, Holistic Alignment Supercharges Diffusion Training
von: Wang, Ziqiao, et al.
Veröffentlicht: (2025)
von: Wang, Ziqiao, et al.
Veröffentlicht: (2025)
FastForward Pruning: Efficient LLM Pruning via Single-Step Reinforcement Learning
von: Yuan, Xin, et al.
Veröffentlicht: (2025)
von: Yuan, Xin, et al.
Veröffentlicht: (2025)
Evolving Programmatic Skill Networks
von: Shi, Haochen, et al.
Veröffentlicht: (2026)
von: Shi, Haochen, et al.
Veröffentlicht: (2026)
VDPP: Video Depth Post-Processing for Speed and Scalability
von: Yoon, Daewon, et al.
Veröffentlicht: (2026)
von: Yoon, Daewon, et al.
Veröffentlicht: (2026)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Faster Vision Mamba is Rebuilt in Minutes via Merged Token Re-training
von: Shi, Mingjia, et al.
Veröffentlicht: (2024) -
GSQ-Tuning: Group-Shared Exponents Integer in Fully Quantized Training for LLMs On-Device Fine-tuning
von: Zhou, Sifan, et al.
Veröffentlicht: (2025) -
A Closer Look at Bias and Chain-of-Thought Faithfulness of Large (Vision) Language Models
von: Balasubramanian, Sriram, et al.
Veröffentlicht: (2025) -
Learning to Look Closer: A New Instance-Wise Loss for Small Cerebral Lesion Segmentation
von: Bouteille, Luc, et al.
Veröffentlicht: (2025) -
Self-supervised ControlNet with Spatio-Temporal Mamba for Real-world Video Super-resolution
von: Shi, Shijun, et al.
Veröffentlicht: (2025)