LPViT: Low-Power Semi-structured Pruning for Vision Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Kaixin, Wang, Zhe, Chen, Chunyun, Geng, Xue, Lin, Jie, Aly, Mohamed M. Sabry, Yang, Xulei, Wu, Min, Li, Xiaoli, Lin, Weisi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DM3D: Distortion-Minimized Weight Pruning for Lossless 3D Object Detection
by: Xu, Kaixin, et al.
Published: (2024)
by: Xu, Kaixin, et al.
Published: (2024)
From Algorithm to Hardware: A Survey on Efficient and Safe Deployment of Deep Neural Networks
by: Geng, Xue, et al.
Published: (2024)
by: Geng, Xue, et al.
Published: (2024)
Joint Architecture-Token-Bitwidth Multi-Axis Optimization of Vision Transformers for Semiconductor IC Packaging
by: Nguyen, Phat, et al.
Published: (2026)
by: Nguyen, Phat, et al.
Published: (2026)
Low Power Vision Transformer Accelerator with Hardware-Aware Pruning and Optimized Dataflow
by: Hsiung, Ching-Lin, et al.
Published: (2025)
by: Hsiung, Ching-Lin, et al.
Published: (2025)
LUT-DLA: Lookup Table as Efficient Extreme Low-Bit Deep Learning Accelerator
by: Li, Guoyu, et al.
Published: (2025)
by: Li, Guoyu, et al.
Published: (2025)
Q-Bench+: A Benchmark for Multi-modal Foundation Models on Low-level Vision from Single Images to Pairs
by: Zhang, Zicheng, et al.
Published: (2024)
by: Zhang, Zicheng, et al.
Published: (2024)
Compress Then Adapt? No, Do It Together via Task-aware Union of Subspaces
by: Ge, Jingze, et al.
Published: (2026)
by: Ge, Jingze, et al.
Published: (2026)
Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression
by: Zhang, Xi, et al.
Published: (2025)
by: Zhang, Xi, et al.
Published: (2025)
SEVEN: Pruning Transformer Model by Reserving Sentinels
by: Xiao, Jinying, et al.
Published: (2024)
by: Xiao, Jinying, et al.
Published: (2024)
Temporal Query Network for Efficient Multivariate Time Series Forecasting
by: Lin, Shengsheng, et al.
Published: (2025)
by: Lin, Shengsheng, et al.
Published: (2025)
An Efficient 3D Convolutional Neural Network with Channel-wise, Spatial-grouped, and Temporal Convolutions
by: Wang, Zhe, et al.
Published: (2025)
by: Wang, Zhe, et al.
Published: (2025)
Low-Cost Stereo Vision for Robust 3D Positioning of Thin Radiata Pine Branches in Autonomous Drone Pruning
by: Lin, Yida, et al.
Published: (2026)
by: Lin, Yida, et al.
Published: (2026)
DiffPCN: Latent Diffusion Model Based on Multi-view Depth Images for Point Cloud Completion
by: Li, Zijun, et al.
Published: (2025)
by: Li, Zijun, et al.
Published: (2025)
Explore the Hallucination on Low-level Perception for MLLMs
by: Sun, Yinan, et al.
Published: (2024)
by: Sun, Yinan, et al.
Published: (2024)
ResPrune: Text-Conditioned Subspace Reconstruction for Visual Token Pruning in Large Vision-Language Models
by: Li, Xu, et al.
Published: (2026)
by: Li, Xu, et al.
Published: (2026)
QAPruner: Quantization-Aware Vision Token Pruning for Multimodal Large Language Models
by: Wang, Xinhao, et al.
Published: (2026)
by: Wang, Xinhao, et al.
Published: (2026)
Half the Interference, Most of the Answer: Approximate Quantum Simulation via Path-Sum Pruning
by: Pehlivanoglu, Sinan, et al.
Published: (2026)
by: Pehlivanoglu, Sinan, et al.
Published: (2026)
Adaptive Computation Pruning for the Forgetting Transformer
by: Lin, Zhixuan, et al.
Published: (2025)
by: Lin, Zhixuan, et al.
Published: (2025)
School-based physical activity and health-related fitness in Mediterranean students: findings from the DELICIOUS project
by: Aly Mohamed, Mohamed
Published: (2025)
by: Aly Mohamed, Mohamed
Published: (2025)
A Timely Survey on Vision Transformer for Deepfake Detection
by: Wang, Zhikan, et al.
Published: (2024)
by: Wang, Zhikan, et al.
Published: (2024)
R4-CGQA: Retrieval-based Vision Language Models for Computer Graphics Image Quality Assessment
by: Li, Zhuangzi, et al.
Published: (2026)
by: Li, Zhuangzi, et al.
Published: (2026)
TSLANet: Rethinking Transformers for Time Series Representation Learning
by: Eldele, Emadeldeen, et al.
Published: (2024)
by: Eldele, Emadeldeen, et al.
Published: (2024)
PPT: Token Pruning and Pooling for Efficient Vision Transformers
by: Wu, Xinjian, et al.
Published: (2023)
by: Wu, Xinjian, et al.
Published: (2023)
ST-Prune: Training-Free Spatio-Temporal Token Pruning for Vision-Language Models in Autonomous Driving
by: Sha, Lin, et al.
Published: (2026)
by: Sha, Lin, et al.
Published: (2026)
CERSA: Cumulative Energy-Retaining Subspace Adaptation for Memory-Efficient Fine-Tuning
by: Ge, Jingze, et al.
Published: (2026)
by: Ge, Jingze, et al.
Published: (2026)
LLM-based Knowledge Pruning for Time Series Data Analytics on Edge-computing Devices
by: Jin, Ruibing, et al.
Published: (2024)
by: Jin, Ruibing, et al.
Published: (2024)
You Only Train Once: A Unified Framework for Both Full-Reference and No-Reference Image Quality Assessment
by: Yun, Yi Ke, et al.
Published: (2023)
by: Yun, Yi Ke, et al.
Published: (2023)
Improving Adversarial Robustness for 3D Point Cloud Recognition at Test-Time through Purified Self-Training
by: Lin, Jinpeng, et al.
Published: (2024)
by: Lin, Jinpeng, et al.
Published: (2024)
Experimental Investigation of an Incremental Contact Model for Hyperelastic Solids Using In-Situ Optical Interferometric Technique
by: Jiang, Chunyun, et al.
Published: (2024)
by: Jiang, Chunyun, et al.
Published: (2024)
Is Complexity Required for Neural Network Pruning? A Case Study on Global Magnitude Pruning
by: Gupta, Manas, et al.
Published: (2022)
by: Gupta, Manas, et al.
Published: (2022)
ADMM Based Semi-Structured Pattern Pruning Framework For Transformer
by: Wang, TianChen
Published: (2024)
by: Wang, TianChen
Published: (2024)
Exact Tensor Completion Powered by Slim Transforms
by: Ge, Li, et al.
Published: (2024)
by: Ge, Li, et al.
Published: (2024)
Systolic Array-based Architecture for Low-Bit Integerized Vision Transformers
by: Lin, Ching-Yi, et al.
Published: (2025)
by: Lin, Ching-Yi, et al.
Published: (2025)
On the Intractability of Chaotic Symbolic Walks: Toward a Non-Algebraic Post-Quantum Hardness Assumption
by: Bouke, Mohamed Aly
Published: (2025)
by: Bouke, Mohamed Aly
Published: (2025)
The Theory of the Unique Latent Pattern: A Formal Epistemic Framework for Structural Singularity in Complex Systems
by: Bouke, Mohamed Aly
Published: (2025)
by: Bouke, Mohamed Aly
Published: (2025)
The Hashed Fractal Key Recovery (HFKR) Problem: From Symbolic Path Inversion to Post-Quantum Cryptographic Keys
by: Bouke, Mohamed Aly
Published: (2025)
by: Bouke, Mohamed Aly
Published: (2025)
Fractal Attractors in Random Nonlinear Iterated Function Systems: Existence, Stability, and Dimensional Properties
by: Bouke, Mohamed Aly
Published: (2025)
by: Bouke, Mohamed Aly
Published: (2025)
GTPT: Group-based Token Pruning Transformer for Efficient Human Pose Estimation
by: Wang, Haonan, et al.
Published: (2024)
by: Wang, Haonan, et al.
Published: (2024)
Adaptive MLP Pruning for Large Vision Transformers
by: Shen, Chengchao
Published: (2026)
by: Shen, Chengchao
Published: (2026)
Robust Semi-Supervised Learning in Open Environments
by: Guo, Lan-Zhe, et al.
Published: (2024)
by: Guo, Lan-Zhe, et al.
Published: (2024)
Similar Items
-
DM3D: Distortion-Minimized Weight Pruning for Lossless 3D Object Detection
by: Xu, Kaixin, et al.
Published: (2024) -
From Algorithm to Hardware: A Survey on Efficient and Safe Deployment of Deep Neural Networks
by: Geng, Xue, et al.
Published: (2024) -
Joint Architecture-Token-Bitwidth Multi-Axis Optimization of Vision Transformers for Semiconductor IC Packaging
by: Nguyen, Phat, et al.
Published: (2026) -
Low Power Vision Transformer Accelerator with Hardware-Aware Pruning and Optimized Dataflow
by: Hsiung, Ching-Lin, et al.
Published: (2025) -
LUT-DLA: Lookup Table as Efficient Extreme Low-Bit Deep Learning Accelerator
by: Li, Guoyu, et al.
Published: (2025)