Comb, Prune, Distill: Towards Unified Pruning for Vision Model Compression
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Schmitt, Jonas, Liu, Ruiping, Zheng, Junwei, Zhang, Jiaming, Stiefelhagen, Rainer |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology
von: Jiang, Xin, et al.
Veröffentlicht: (2024)
von: Jiang, Xin, et al.
Veröffentlicht: (2024)
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping
von: Wei, Jiale, et al.
Veröffentlicht: (2024)
von: Wei, Jiale, et al.
Veröffentlicht: (2024)
Faster or Stronger: Towards Flexible Visual Place Recognition via Weighted Aggregation and Token Pruning
von: Zeng, Zichao, et al.
Veröffentlicht: (2026)
von: Zeng, Zichao, et al.
Veröffentlicht: (2026)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
HybriDLA: Hybrid Generation for Document Layout Analysis
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
Graph-based Document Structure Analysis
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
Scene-agnostic Pose Regression for Visual Localization
von: Zheng, Junwei, et al.
Veröffentlicht: (2025)
von: Zheng, Junwei, et al.
Veröffentlicht: (2025)
Open Panoramic Segmentation
von: Zheng, Junwei, et al.
Veröffentlicht: (2024)
von: Zheng, Junwei, et al.
Veröffentlicht: (2024)
Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model
von: Liu, Ruiping, et al.
Veröffentlicht: (2025)
von: Liu, Ruiping, et al.
Veröffentlicht: (2025)
Deformable Mamba for Wide Field of View Segmentation
von: Hu, Jie, et al.
Veröffentlicht: (2024)
von: Hu, Jie, et al.
Veröffentlicht: (2024)
SGR3 Model: Scene Graph Retrieval-Reasoning Model in 3D
von: Wang, Zirui, et al.
Veröffentlicht: (2026)
von: Wang, Zirui, et al.
Veröffentlicht: (2026)
More than the Sum: Panorama-Language Models for Adverse Omni-Scenes
von: Fan, Weijia, et al.
Veröffentlicht: (2026)
von: Fan, Weijia, et al.
Veröffentlicht: (2026)
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
mmWalk: Towards Multi-modal Multi-view Walking Assistance
von: Ying, Kedi, et al.
Veröffentlicht: (2025)
von: Ying, Kedi, et al.
Veröffentlicht: (2025)
LAPTOP-Diff: Layer Pruning and Normalized Distillation for Compressing Diffusion Models
von: Zhang, Dingkun, et al.
Veröffentlicht: (2024)
von: Zhang, Dingkun, et al.
Veröffentlicht: (2024)
DriveXQA: Cross-modal Visual Question Answering for Adverse Driving Scene Understanding
von: Tao, Mingzhe, et al.
Veröffentlicht: (2026)
von: Tao, Mingzhe, et al.
Veröffentlicht: (2026)
What if? Emulative Simulation with World Models for Situated Reasoning
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
EgoExoMem: Cross-View Memory Reasoning over Synchronized Egocentric and Exocentric Videos
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
von: Liu, Ruiping, et al.
Veröffentlicht: (2026)
RHO: Robust Holistic OSM-Based Metric Cross-View Geo-Localization
von: Zheng, Junwei, et al.
Veröffentlicht: (2026)
von: Zheng, Junwei, et al.
Veröffentlicht: (2026)
Towards Joint Quantization and Token Pruning of Vision-Language Models
von: Li, Xinqing, et al.
Veröffentlicht: (2026)
von: Li, Xinqing, et al.
Veröffentlicht: (2026)
ObjectFinder: An Open-Vocabulary Assistive System for Interactive Object Search by Blind People
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
Compressing Multi-Task Model for Autonomous Driving via Pruning and Knowledge Distillation
von: Wang, Jiayuan, et al.
Veröffentlicht: (2025)
von: Wang, Jiayuan, et al.
Veröffentlicht: (2025)
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
CHAOS: Chart Analysis with Outlier Samples
von: Moured, Omar, et al.
Veröffentlicht: (2025)
von: Moured, Omar, et al.
Veröffentlicht: (2025)
EPSD: Early Pruning with Self-Distillation for Efficient Model Compression
von: Chen, Dong, et al.
Veröffentlicht: (2024)
von: Chen, Dong, et al.
Veröffentlicht: (2024)
HiPrune: Hierarchical Attention for Efficient Token Pruning in Vision-Language Models
von: Liu, Jizhihui, et al.
Veröffentlicht: (2025)
von: Liu, Jizhihui, et al.
Veröffentlicht: (2025)
Towards Universal & Efficient Model Compression via Exponential Torque Pruning
von: Modi, Sarthak Ketanbhai, et al.
Veröffentlicht: (2025)
von: Modi, Sarthak Ketanbhai, et al.
Veröffentlicht: (2025)
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers
von: Ma, Ji, et al.
Veröffentlicht: (2025)
von: Ma, Ji, et al.
Veröffentlicht: (2025)
Topology-Aware Layer Pruning for Large Vision-Language Models
von: Zheng, Pengcheng, et al.
Veröffentlicht: (2026)
von: Zheng, Pengcheng, et al.
Veröffentlicht: (2026)
Towards Multi-Source Domain Generalization for Sleep Staging with Noisy Labels
von: Wang, Kening, et al.
Veröffentlicht: (2026)
von: Wang, Kening, et al.
Veröffentlicht: (2026)
ZOO-Prune: Training-Free Token Pruning via Zeroth-Order Gradient Estimation in Vision-Language Models
von: Kim, Youngeun, et al.
Veröffentlicht: (2025)
von: Kim, Youngeun, et al.
Veröffentlicht: (2025)
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
von: Wei, Yiping, et al.
Veröffentlicht: (2023)
von: Wei, Yiping, et al.
Veröffentlicht: (2023)
MULTIFLOW: Shifting Towards Task-Agnostic Vision-Language Pruning
von: Farina, Matteo, et al.
Veröffentlicht: (2024)
von: Farina, Matteo, et al.
Veröffentlicht: (2024)
Fourier Prompt Tuning for Modality-Incomplete Scene Segmentation
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
A Glimpse to Compress: Dynamic Visual Token Pruning for Large Vision-Language Models
von: Zeng, Quan-Sheng, et al.
Veröffentlicht: (2025)
von: Zeng, Quan-Sheng, et al.
Veröffentlicht: (2025)
IVC-Prune: Revealing the Implicit Visual Coordinates in LVLMs for Vision Token Pruning
von: Sun, Zhichao, et al.
Veröffentlicht: (2026)
von: Sun, Zhichao, et al.
Veröffentlicht: (2026)
TT-MPD: Test Time Model Pruning and Distillation
von: Wu, Haihang, et al.
Veröffentlicht: (2024)
von: Wu, Haihang, et al.
Veröffentlicht: (2024)
UTPTrack: Towards Simple and Unified Token Pruning for Visual Tracking
von: Wu, Hao, et al.
Veröffentlicht: (2026)
von: Wu, Hao, et al.
Veröffentlicht: (2026)
Model Compression using Progressive Channel Pruning
von: Guo, Jinyang, et al.
Veröffentlicht: (2025)
von: Guo, Jinyang, et al.
Veröffentlicht: (2025)
Pruning then Reweighting: Towards Data-Efficient Training of Diffusion Models
von: Li, Yize, et al.
Veröffentlicht: (2024)
von: Li, Yize, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology
von: Jiang, Xin, et al.
Veröffentlicht: (2024) -
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping
von: Wei, Jiale, et al.
Veröffentlicht: (2024) -
Faster or Stronger: Towards Flexible Visual Place Recognition via Weighted Aggregation and Token Pruning
von: Zeng, Zichao, et al.
Veröffentlicht: (2026) -
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
von: Chen, Yufan, et al.
Veröffentlicht: (2024) -
HybriDLA: Hybrid Generation for Document Layout Analysis
von: Chen, Yufan, et al.
Veröffentlicht: (2025)