Layer- and Timestep-Adaptive Differentiable Token Compression Ratios for Efficient Diffusion Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | You, Haoran, Barnes, Connelly, Zhou, Yuqian, Kang, Yan, Du, Zhenbang, Zhou, Wei, Zhang, Lingzhi, Nitzan, Yotam, Liu, Xiaoyang, Lin, Zhe, Shechtman, Eli, Amirghodsi, Sohrab, Lin, Yingyan Celine |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fine-grained Defocus Blur Control for Generative Image Models
von: Shrivastava, Ayush, et al.
Veröffentlicht: (2025)
von: Shrivastava, Ayush, et al.
Veröffentlicht: (2025)
UniSER: A Foundation Model for Unified Soft Effects Removal
von: Zhang, Jingdong, et al.
Veröffentlicht: (2025)
von: Zhang, Jingdong, et al.
Veröffentlicht: (2025)
Structure-Guided Image Completion with Image-level and Object-level Semantic Discriminators
von: Zheng, Haitian, et al.
Veröffentlicht: (2022)
von: Zheng, Haitian, et al.
Veröffentlicht: (2022)
Early-Bird Diffusion: Investigating and Leveraging Timestep-Aware Early-Bird Tickets in Diffusion Models for Efficient Training
von: Whalen, Lexington, et al.
Veröffentlicht: (2025)
von: Whalen, Lexington, et al.
Veröffentlicht: (2025)
Self-Evaluation Unlocks Any-Step Text-to-Image Generation
von: Yu, Xin, et al.
Veröffentlicht: (2025)
von: Yu, Xin, et al.
Veröffentlicht: (2025)
Early-Bird GCNs: Graph-Network Co-Optimization Towards More Efficient GCN Training and Inference via Drawing Early-Bird Lottery Tickets
von: You, Haoran, et al.
Veröffentlicht: (2021)
von: You, Haoran, et al.
Veröffentlicht: (2021)
ShiftAddViT: Mixture of Multiplication Primitives Towards Efficient Vision Transformer
von: You, Haoran, et al.
Veröffentlicht: (2023)
von: You, Haoran, et al.
Veröffentlicht: (2023)
Lazy Diffusion Transformer for Interactive Image Editing
von: Nitzan, Yotam, et al.
Veröffentlicht: (2024)
von: Nitzan, Yotam, et al.
Veröffentlicht: (2024)
Cautious Next Token Prediction
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
EDGE-LLM: Enabling Efficient Large Language Model Adaptation on Edge Devices via Layerwise Unified Compression and Adaptive Layer Tuning and Voting
von: Yu, Zhongzhi, et al.
Veröffentlicht: (2024)
von: Yu, Zhongzhi, et al.
Veröffentlicht: (2024)
Long-Context State-Space Video World Models
von: Po, Ryan, et al.
Veröffentlicht: (2025)
von: Po, Ryan, et al.
Veröffentlicht: (2025)
ShiftAddNAS: Hardware-Inspired Search for More Accurate and Efficient Neural Networks
von: You, Haoran, et al.
Veröffentlicht: (2022)
von: You, Haoran, et al.
Veröffentlicht: (2022)
SuperTickets: Drawing Task-Agnostic Lottery Tickets from Supernets via Jointly Architecture Searching and Parameter Pruning
von: You, Haoran, et al.
Veröffentlicht: (2022)
von: You, Haoran, et al.
Veröffentlicht: (2022)
When Linear Attention Meets Autoregressive Decoding: Towards More Effective and Efficient Linearized Large Language Models
von: You, Haoran, et al.
Veröffentlicht: (2024)
von: You, Haoran, et al.
Veröffentlicht: (2024)
GCoD: Graph Convolutional Network Acceleration via Dedicated Algorithm and Accelerator Co-Design
von: You, Haoran, et al.
Veröffentlicht: (2021)
von: You, Haoran, et al.
Veröffentlicht: (2021)
AutoGAN-Distiller: Searching to Compress Generative Adversarial Networks
von: Fu, Yonggan, et al.
Veröffentlicht: (2020)
von: Fu, Yonggan, et al.
Veröffentlicht: (2020)
DNA: Differentiable Network-Accelerator Co-Search
von: Zhang, Yongan, et al.
Veröffentlicht: (2020)
von: Zhang, Yongan, et al.
Veröffentlicht: (2020)
Castling-ViT: Compressing Self-Attention via Switching Towards Linear-Angular Attention at Vision Transformer Inference
von: You, Haoran, et al.
Veröffentlicht: (2022)
von: You, Haoran, et al.
Veröffentlicht: (2022)
ZipIR: Latent Pyramid Diffusion Transformer for High-Resolution Image Restoration
von: Yu, Yongsheng, et al.
Veröffentlicht: (2025)
von: Yu, Yongsheng, et al.
Veröffentlicht: (2025)
Generative Multimodal Pretraining with Discrete Diffusion Timestep Tokens
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
Gen-NeRF: Efficient and Generalizable Neural Radiance Fields via Algorithm-Hardware Co-Design
von: Fu, Yonggan, et al.
Veröffentlicht: (2023)
von: Fu, Yonggan, et al.
Veröffentlicht: (2023)
Distilling Diffusion Models into Conditional GANs
von: Kang, Minguk, et al.
Veröffentlicht: (2024)
von: Kang, Minguk, et al.
Veröffentlicht: (2024)
ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization
von: You, Haoran, et al.
Veröffentlicht: (2024)
von: You, Haoran, et al.
Veröffentlicht: (2024)
Timestep-Compressed Attack on Spiking Neural Networks through Timestep-Level Backpropagation
von: Kang, Donghwa, et al.
Veröffentlicht: (2025)
von: Kang, Donghwa, et al.
Veröffentlicht: (2025)
Low Bitrate High-Quality RVQGAN-based Discrete Speech Tokenizer
von: Shechtman, Slava, et al.
Veröffentlicht: (2024)
von: Shechtman, Slava, et al.
Veröffentlicht: (2024)
Fewer Denoising Steps or Cheaper Per-Step Inference: Towards Compute-Optimal Diffusion Model Deployment
von: Du, Zhenbang, et al.
Veröffentlicht: (2025)
von: Du, Zhenbang, et al.
Veröffentlicht: (2025)
Learning an Image Editing Model without Image Editing Pairs
von: Kumari, Nupur, et al.
Veröffentlicht: (2025)
von: Kumari, Nupur, et al.
Veröffentlicht: (2025)
FracTrain: Fractionally Squeezing Bit Savings Both Temporally and Spatially for Efficient DNN Training
von: Fu, Yonggan, et al.
Veröffentlicht: (2020)
von: Fu, Yonggan, et al.
Veröffentlicht: (2020)
ShiftAddNet: A Hardware-Inspired Deep Network
von: You, Haoran, et al.
Veröffentlicht: (2020)
von: You, Haoran, et al.
Veröffentlicht: (2020)
VLM-Guided Adaptive Negative Prompting for Creative Generation
von: Golan, Shelly, et al.
Veröffentlicht: (2025)
von: Golan, Shelly, et al.
Veröffentlicht: (2025)
Reasoning Physical Video Generation with Diffusion Timestep Tokens via Reinforcement Learning
von: Lin, Wang, et al.
Veröffentlicht: (2025)
von: Lin, Wang, et al.
Veröffentlicht: (2025)
Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
von: Huang, Xun, et al.
Veröffentlicht: (2025)
von: Huang, Xun, et al.
Veröffentlicht: (2025)
Jump Cut Smoothing for Talking Heads
von: Wang, Xiaojuan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaojuan, et al.
Veröffentlicht: (2024)
MixRT: Mixed Neural Representations For Real-Time NeRF Rendering
von: Li, Chaojian, et al.
Veröffentlicht: (2023)
von: Li, Chaojian, et al.
Veröffentlicht: (2023)
GauRast: Enhancing GPU Triangle Rasterizers to Accelerate 3D Gaussian Splatting
von: Li, Sixu, et al.
Veröffentlicht: (2025)
von: Li, Sixu, et al.
Veröffentlicht: (2025)
Auto-Agent-Distiller: Towards Efficient Deep Reinforcement Learning Agents via Neural Architecture Search
von: Fu, Yonggan, et al.
Veröffentlicht: (2020)
von: Fu, Yonggan, et al.
Veröffentlicht: (2020)
Raman spectroscopy at metal interfaces: A numerical study of the strong coupling regime
von: Zhou, Zeyu, et al.
Veröffentlicht: (2026)
von: Zhou, Zeyu, et al.
Veröffentlicht: (2026)
Max-Affine Spline Insights Into Deep Network Pruning
von: You, Haoran, et al.
Veröffentlicht: (2021)
von: You, Haoran, et al.
Veröffentlicht: (2021)
HW-NAS-Bench:Hardware-Aware Neural Architecture Search Benchmark
von: Li, Chaojian, et al.
Veröffentlicht: (2021)
von: Li, Chaojian, et al.
Veröffentlicht: (2021)
Instant-3D: Instant Neural Radiance Field Training Towards On-Device AR/VR 3D Reconstruction
von: Li, Sixu, et al.
Veröffentlicht: (2023)
von: Li, Sixu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Fine-grained Defocus Blur Control for Generative Image Models
von: Shrivastava, Ayush, et al.
Veröffentlicht: (2025) -
UniSER: A Foundation Model for Unified Soft Effects Removal
von: Zhang, Jingdong, et al.
Veröffentlicht: (2025) -
Structure-Guided Image Completion with Image-level and Object-level Semantic Discriminators
von: Zheng, Haitian, et al.
Veröffentlicht: (2022) -
Early-Bird Diffusion: Investigating and Leveraging Timestep-Aware Early-Bird Tickets in Diffusion Models for Efficient Training
von: Whalen, Lexington, et al.
Veröffentlicht: (2025) -
Self-Evaluation Unlocks Any-Step Text-to-Image Generation
von: Yu, Xin, et al.
Veröffentlicht: (2025)