TinyFusion: Diffusion Transformers Learned Shallow
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fang, Gongfan, Li, Kunjun, Ma, Xinyin, Wang, Xinchao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Remix-DiT: Mixing Diffusion Transformers for Multi-Expert Denoising
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Isomorphic Pruning for Vision Models
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
von: Fang, Gongfan, et al.
Veröffentlicht: (2024)
Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching
von: Ma, Xinyin, et al.
Veröffentlicht: (2024)
von: Ma, Xinyin, et al.
Veröffentlicht: (2024)
In-Video Instructions: Visual Signals as Generative Control
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
ConciseHint: Boosting Efficient Reasoning via Continuous Concise Hints during Generation
von: Tang, Siao, et al.
Veröffentlicht: (2025)
von: Tang, Siao, et al.
Veröffentlicht: (2025)
Collaborative Decoding Makes Visual Auto-Regressive Modeling Efficient
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)
SlimSAM: 0.1% Data Makes Segment Anything Slim
von: Chen, Zigeng, et al.
Veröffentlicht: (2023)
von: Chen, Zigeng, et al.
Veröffentlicht: (2023)
Q-ARVD: Quantizing Autoregressive Video Diffusion Models
von: Tang, Siao, et al.
Veröffentlicht: (2026)
von: Tang, Siao, et al.
Veröffentlicht: (2026)
OminiControl: Minimal and Universal Control for Diffusion Transformer
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2024)
Introducing Visual Perception Token into Multimodal Large Language Model
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
Diversity-Guided MLP Reduction for Efficient Large Vision Transformers
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
Kolmogorov-Arnold Transformer
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
Thinkless: LLM Learns When to Think
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
von: Fang, Gongfan, et al.
Veröffentlicht: (2025)
Tiny Machine Learning: Progress and Futures
von: Lin, Ji, et al.
Veröffentlicht: (2024)
von: Lin, Ji, et al.
Veröffentlicht: (2024)
MambaOut: Do We Really Need Mamba for Vision?
von: Yu, Weihao, et al.
Veröffentlicht: (2024)
von: Yu, Weihao, et al.
Veröffentlicht: (2024)
Neural Metamorphosis
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
von: Yang, Xingyi, et al.
Veröffentlicht: (2024)
DMax: Aggressive Parallel Decoding for dLLMs
von: Chen, Zigeng, et al.
Veröffentlicht: (2026)
von: Chen, Zigeng, et al.
Veröffentlicht: (2026)
LogTinyLLM: Tiny Large Language Models Based Contextual Log Anomaly Detection
von: Ocansey, Isaiah Thompson, et al.
Veröffentlicht: (2025)
von: Ocansey, Isaiah Thompson, et al.
Veröffentlicht: (2025)
InceptionNeXt: When Inception Meets ConvNeXt
von: Yu, Weihao, et al.
Veröffentlicht: (2023)
von: Yu, Weihao, et al.
Veröffentlicht: (2023)
TinyAlign: Boosting Lightweight Vision-Language Models by Mitigating Modal Alignment Bottlenecks
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
von: Hu, Yuanze, et al.
Veröffentlicht: (2025)
FedDiff: Diffusion Model Driven Federated Learning for Multi-Modal and Multi-Clients
von: Li, DaiXun, et al.
Veröffentlicht: (2023)
von: Li, DaiXun, et al.
Veröffentlicht: (2023)
MosaicFusion: Diffusion Models as Data Augmenters for Large Vocabulary Instance Segmentation
von: Xie, Jiahao, et al.
Veröffentlicht: (2023)
von: Xie, Jiahao, et al.
Veröffentlicht: (2023)
ShiftAddAug: Augment Multiplication-Free Tiny Neural Network with Hybrid Computation
von: Guo, Yipin, et al.
Veröffentlicht: (2024)
von: Guo, Yipin, et al.
Veröffentlicht: (2024)
Scaling Diffusion Transformers Efficiently via $μ$P
von: Zheng, Chenyu, et al.
Veröffentlicht: (2025)
von: Zheng, Chenyu, et al.
Veröffentlicht: (2025)
Taming Outlier Tokens in Diffusion Transformers
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2026)
SpeCa: Accelerating Diffusion Transformers with Speculative Feature Caching
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
Precipitation Nowcasting Using Diffusion Transformer with Causal Attention
von: Li, ChaoRong, et al.
Veröffentlicht: (2024)
von: Li, ChaoRong, et al.
Veröffentlicht: (2024)
PCaM: A Progressive Focus Attention-Based Information Fusion Method for Improving Vision Transformer Domain Adaptation
von: Zang, Zelin, et al.
Veröffentlicht: (2025)
von: Zang, Zelin, et al.
Veröffentlicht: (2025)
NoiseRater: Meta-Learned Noise Valuation for Diffusion Model Training
von: Wu, Fang, et al.
Veröffentlicht: (2026)
von: Wu, Fang, et al.
Veröffentlicht: (2026)
Dimple: Discrete Diffusion Multimodal Large Language Model with Parallel Decoding
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
von: Yu, Runpeng, et al.
Veröffentlicht: (2025)
Video Motion Transfer with Diffusion Transformers
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
FedAFD: Multimodal Federated Learning via Adversarial Fusion and Distillation
von: Tan, Min, et al.
Veröffentlicht: (2026)
von: Tan, Min, et al.
Veröffentlicht: (2026)
Towards Precise Scaling Laws for Video Diffusion Transformers
von: Yin, Yuanyang, et al.
Veröffentlicht: (2024)
von: Yin, Yuanyang, et al.
Veröffentlicht: (2024)
WiTUnet: A U-Shaped Architecture Integrating CNN and Transformer for Improved Feature Alignment and Local Information Fusion
von: Wang, Bin, et al.
Veröffentlicht: (2024)
von: Wang, Bin, et al.
Veröffentlicht: (2024)
DiffiT: Diffusion Vision Transformers for Image Generation
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
Fast Training of Diffusion Models with Masked Transformers
von: Zheng, Hongkai, et al.
Veröffentlicht: (2023)
von: Zheng, Hongkai, et al.
Veröffentlicht: (2023)
TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment
von: Li, Jiaming, et al.
Veröffentlicht: (2026)
von: Li, Jiaming, et al.
Veröffentlicht: (2026)
Accelerating Diffusion Transformers with Token-wise Feature Caching
von: Zou, Chang, et al.
Veröffentlicht: (2024)
von: Zou, Chang, et al.
Veröffentlicht: (2024)
Rethinking The Uniformity Metric in Self-Supervised Learning
von: Fang, Xianghong, et al.
Veröffentlicht: (2024)
von: Fang, Xianghong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Remix-DiT: Mixing Diffusion Transformers for Multi-Expert Denoising
von: Fang, Gongfan, et al.
Veröffentlicht: (2024) -
Isomorphic Pruning for Vision Models
von: Fang, Gongfan, et al.
Veröffentlicht: (2024) -
Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching
von: Ma, Xinyin, et al.
Veröffentlicht: (2024) -
In-Video Instructions: Visual Signals as Generative Control
von: Fang, Gongfan, et al.
Veröffentlicht: (2025) -
AsyncDiff: Parallelizing Diffusion Models by Asynchronous Denoising
von: Chen, Zigeng, et al.
Veröffentlicht: (2024)