Ditto: Accelerating Diffusion Model via Temporal Value Similarity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Sungbin, Lee, Hyunwuk, Cho, Wonho, Park, Mincheol, Ro, Won Woo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration
von: Lee, Hangyeol, et al.
Veröffentlicht: (2026)
von: Lee, Hangyeol, et al.
Veröffentlicht: (2026)
SQ-DM: Accelerating Diffusion Models with Aggressive Quantization and Temporal Sparsity
von: Fan, Zichen, et al.
Veröffentlicht: (2025)
von: Fan, Zichen, et al.
Veröffentlicht: (2025)
ViTCoD: Vision Transformer Acceleration via Dedicated Algorithm and Accelerator Co-Design
von: You, Haoran, et al.
Veröffentlicht: (2022)
von: You, Haoran, et al.
Veröffentlicht: (2022)
SF-MMCN: Low-Power Sever Flow Multi-Mode Diffusion Model Accelerator
von: Hsu, Huan-Ke, et al.
Veröffentlicht: (2024)
von: Hsu, Huan-Ke, et al.
Veröffentlicht: (2024)
Neo: Real-Time On-Device 3D Gaussian Splatting with Reuse-and-Update Sorting Acceleration
von: Oh, Changhun, et al.
Veröffentlicht: (2025)
von: Oh, Changhun, et al.
Veröffentlicht: (2025)
TIMERIPPLE: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space
von: Miao, Wenxuan, et al.
Veröffentlicht: (2025)
von: Miao, Wenxuan, et al.
Veröffentlicht: (2025)
GS-TG: 3D Gaussian Splatting Accelerator with Tile Grouping for Reducing Redundant Sorting while Preserving Rasterization Efficiency
von: Jo, Joongho, et al.
Veröffentlicht: (2025)
von: Jo, Joongho, et al.
Veröffentlicht: (2025)
Identifying Unnecessary 3D Gaussians using Clustering for Fast Rendering of 3D Gaussian Splatting
von: Jo, Joongho, et al.
Veröffentlicht: (2024)
von: Jo, Joongho, et al.
Veröffentlicht: (2024)
Stella Nera: A Differentiable Maddness-Based Hardware Accelerator for Efficient Approximate Matrix Multiplication
von: Schönleber, Jannis, et al.
Veröffentlicht: (2023)
von: Schönleber, Jannis, et al.
Veröffentlicht: (2023)
HARFLOW3D: A Latency-Oriented 3D-CNN Accelerator Toolflow for HAR on FPGA Devices
von: Toupas, Petros, et al.
Veröffentlicht: (2023)
von: Toupas, Petros, et al.
Veröffentlicht: (2023)
Vision Transformers on the Edge: A Comprehensive Survey of Model Compression and Acceleration Strategies
von: Saha, Shaibal, et al.
Veröffentlicht: (2025)
von: Saha, Shaibal, et al.
Veröffentlicht: (2025)
VR-Pipe: Streamlining Hardware Graphics Pipeline for Volume Rendering
von: Lee, Junseo, et al.
Veröffentlicht: (2025)
von: Lee, Junseo, et al.
Veröffentlicht: (2025)
A Parameterizable Convolution Accelerator for Embedded Deep Learning Applications
von: Mousouliotis, Panagiotis, et al.
Veröffentlicht: (2026)
von: Mousouliotis, Panagiotis, et al.
Veröffentlicht: (2026)
Primitive-Driven Acceleration of Hyperdimensional Computing for Real-Time Image Classification
von: Parikh, Dhruv, et al.
Veröffentlicht: (2026)
von: Parikh, Dhruv, et al.
Veröffentlicht: (2026)
MVQ:Towards Efficient DNN Compression and Acceleration with Masked Vector Quantization
von: Li, Shuaiting, et al.
Veröffentlicht: (2024)
von: Li, Shuaiting, et al.
Veröffentlicht: (2024)
Accelerating 3D Gaussian Splatting with Neural Sorting and Axis-Oriented Rasterization
von: Wang, Zhican, et al.
Veröffentlicht: (2025)
von: Wang, Zhican, et al.
Veröffentlicht: (2025)
DPU or GPU for Accelerating Neural Networks Inference -- Why not both? Split CNN Inference
von: Oztas, Ali Emre, et al.
Veröffentlicht: (2026)
von: Oztas, Ali Emre, et al.
Veröffentlicht: (2026)
SpNeRF: Memory Efficient Sparse Volumetric Neural Rendering Accelerator for Edge Devices
von: Zhang, Yipu, et al.
Veröffentlicht: (2025)
von: Zhang, Yipu, et al.
Veröffentlicht: (2025)
hARMS: A Hardware Acceleration Architecture for Real-Time Event-Based Optical Flow
von: Stumpp, Daniel C., et al.
Veröffentlicht: (2021)
von: Stumpp, Daniel C., et al.
Veröffentlicht: (2021)
Accelerating AI and Computer Vision for Satellite Pose Estimation on the Intel Myriad X Embedded SoC
von: Leon, Vasileios, et al.
Veröffentlicht: (2024)
von: Leon, Vasileios, et al.
Veröffentlicht: (2024)
CRISP: Hybrid Structured Sparsity for Class-aware Model Pruning
von: Aggarwal, Shivam, et al.
Veröffentlicht: (2023)
von: Aggarwal, Shivam, et al.
Veröffentlicht: (2023)
GRTX: Efficient Ray Tracing for 3D Gaussian-Based Rendering
von: Lee, Junseo, et al.
Veröffentlicht: (2026)
von: Lee, Junseo, et al.
Veröffentlicht: (2026)
Benchmarking Deep Learning Models on NVIDIA Jetson Nano for Real-Time Systems: An Empirical Investigation
von: Swaminathan, Tushar Prasanna, et al.
Veröffentlicht: (2024)
von: Swaminathan, Tushar Prasanna, et al.
Veröffentlicht: (2024)
AHCQ-SAM: Toward Accurate and Hardware-Compatible Post-Training Segment Anything Model Quantization
von: Zhang, Wenlun, et al.
Veröffentlicht: (2025)
von: Zhang, Wenlun, et al.
Veröffentlicht: (2025)
FG-Attn: Leveraging Fine-Grained Sparsity In Diffusion Transformers
von: Durvasula, Sankeerth, et al.
Veröffentlicht: (2025)
von: Durvasula, Sankeerth, et al.
Veröffentlicht: (2025)
LoRA-Edge: Tensor-Train-Assisted LoRA for Practical CNN Fine-Tuning on Edge Devices
von: Kwak, Hyunseok, et al.
Veröffentlicht: (2025)
von: Kwak, Hyunseok, et al.
Veröffentlicht: (2025)
ViM-Q: Scalable Algorithm-Hardware Co-Design for Vision Mamba Model Inference on FPGA
von: Lyu, Shengzhe, et al.
Veröffentlicht: (2026)
von: Lyu, Shengzhe, et al.
Veröffentlicht: (2026)
Ternary-Input Binary-Weight CNN Accelerator Design for Miniature Object Classification System with Query-Driven Spatial DVS
von: Li, Yuyang, et al.
Veröffentlicht: (2025)
von: Li, Yuyang, et al.
Veröffentlicht: (2025)
CDM-QTA: Quantized Training Acceleration for Efficient LoRA Fine-Tuning of Diffusion Model
von: Lu, Jinming, et al.
Veröffentlicht: (2025)
von: Lu, Jinming, et al.
Veröffentlicht: (2025)
Model Quantization and Hardware Acceleration for Vision Transformers: A Comprehensive Survey
von: Du, Dayou, et al.
Veröffentlicht: (2024)
von: Du, Dayou, et al.
Veröffentlicht: (2024)
Energy Efficient Exact and Approximate Systolic Array Architecture for Matrix Multiplication
von: Jaswal, Pragun, et al.
Veröffentlicht: (2025)
von: Jaswal, Pragun, et al.
Veröffentlicht: (2025)
Smaller, Faster, Cheaper: Architectural Designs for Efficient Machine Learning
von: Walton, Steven
Veröffentlicht: (2025)
von: Walton, Steven
Veröffentlicht: (2025)
SMOF: Streaming Modern CNNs on FPGAs with Smart Off-Chip Eviction
von: Toupas, Petros, et al.
Veröffentlicht: (2024)
von: Toupas, Petros, et al.
Veröffentlicht: (2024)
Performance Analysis of Edge and In-Sensor AI Processors: A Comparative Review
von: Capogrosso, Luigi, et al.
Veröffentlicht: (2026)
von: Capogrosso, Luigi, et al.
Veröffentlicht: (2026)
QUILL: An Algorithm-Architecture Co-Design for Cache-Local Deformable Attention
von: Oh, Hyunwoo, et al.
Veröffentlicht: (2025)
von: Oh, Hyunwoo, et al.
Veröffentlicht: (2025)
NeuralFuse: Learning to Recover the Accuracy of Access-Limited Neural Network Inference in Low-Voltage Regimes
von: Sun, Hao-Lun, et al.
Veröffentlicht: (2023)
von: Sun, Hao-Lun, et al.
Veröffentlicht: (2023)
Neuro-Channel Networks: A Multiplication-Free Architecture by Biological Signal Transmission
von: Mete, Emrah, et al.
Veröffentlicht: (2026)
von: Mete, Emrah, et al.
Veröffentlicht: (2026)
Mix-and-Match Pruning: Globally Guided Layer-Wise Sparsification of DNNs
von: Monachan, Danial, et al.
Veröffentlicht: (2026)
von: Monachan, Danial, et al.
Veröffentlicht: (2026)
Uni-Render: A Unified Accelerator for Real-Time Rendering Across Diverse Neural Renderers
von: Li, Chaojian, et al.
Veröffentlicht: (2025)
von: Li, Chaojian, et al.
Veröffentlicht: (2025)
Gen-NeRF: Efficient and Generalizable Neural Radiance Fields via Algorithm-Hardware Co-Design
von: Fu, Yonggan, et al.
Veröffentlicht: (2023)
von: Fu, Yonggan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration
von: Lee, Hangyeol, et al.
Veröffentlicht: (2026) -
SQ-DM: Accelerating Diffusion Models with Aggressive Quantization and Temporal Sparsity
von: Fan, Zichen, et al.
Veröffentlicht: (2025) -
ViTCoD: Vision Transformer Acceleration via Dedicated Algorithm and Accelerator Co-Design
von: You, Haoran, et al.
Veröffentlicht: (2022) -
SF-MMCN: Low-Power Sever Flow Multi-Mode Diffusion Model Accelerator
von: Hsu, Huan-Ke, et al.
Veröffentlicht: (2024) -
Neo: Real-Time On-Device 3D Gaussian Splatting with Reuse-and-Update Sorting Acceleration
von: Oh, Changhun, et al.
Veröffentlicht: (2025)