Saved in:
| Main Authors: | Chen, Dong, Liu, Ning, Zhu, Yichen, Che, Zhengping, Ma, Rui, Zhang, Fachao, Mou, Xiaofeng, Chang, Yi, Tang, Jian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2402.00084 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
by: Zhu, Yichen, et al.
Published: (2024)
by: Zhu, Yichen, et al.
Published: (2024)
Compressing Multi-Task Model for Autonomous Driving via Pruning and Knowledge Distillation
by: Wang, Jiayuan, et al.
Published: (2025)
by: Wang, Jiayuan, et al.
Published: (2025)
Comb, Prune, Distill: Towards Unified Pruning for Vision Model Compression
by: Schmitt, Jonas, et al.
Published: (2024)
by: Schmitt, Jonas, et al.
Published: (2024)
Language-Conditioned Robotic Manipulation with Fast and Slow Thinking
by: Zhu, Minjie, et al.
Published: (2024)
by: Zhu, Minjie, et al.
Published: (2024)
Object-Centric Instruction Augmentation for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024)
by: Wen, Junjie, et al.
Published: (2024)
Pluggable Pruning with Contiguous Layer Distillation for Diffusion Transformers
by: Ma, Jian, et al.
Published: (2025)
by: Ma, Jian, et al.
Published: (2025)
EDT: An Efficient Diffusion Transformer Framework Inspired by Human-like Sketching
by: Chen, Xinwang, et al.
Published: (2024)
by: Chen, Xinwang, et al.
Published: (2024)
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation
by: Zhao, Yinuo, et al.
Published: (2024)
by: Zhao, Yinuo, et al.
Published: (2024)
LAPTOP-Diff: Layer Pruning and Normalized Distillation for Compressing Diffusion Models
by: Zhang, Dingkun, et al.
Published: (2024)
by: Zhang, Dingkun, et al.
Published: (2024)
SM$^3$: Self-Supervised Multi-task Modeling with Multi-view 2D Images for Articulated Objects
by: Wang, Haowen, et al.
Published: (2024)
by: Wang, Haowen, et al.
Published: (2024)
LiteVAR: Compressing Visual Autoregressive Modelling with Efficient Attention and Quantization
by: Xie, Rui, et al.
Published: (2024)
by: Xie, Rui, et al.
Published: (2024)
Model Compression using Progressive Channel Pruning
by: Guo, Jinyang, et al.
Published: (2025)
by: Guo, Jinyang, et al.
Published: (2025)
Towards Universal & Efficient Model Compression via Exponential Torque Pruning
by: Modi, Sarthak Ketanbhai, et al.
Published: (2025)
by: Modi, Sarthak Ketanbhai, et al.
Published: (2025)
Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions
by: Dong, ZiYi, et al.
Published: (2025)
by: Dong, ZiYi, et al.
Published: (2025)
Towards Efficient VLMs: Information-Theoretic Driven Compression via Adaptive Structural Pruning
by: Xu, Zhaoqi, et al.
Published: (2025)
by: Xu, Zhaoqi, et al.
Published: (2025)
Dual Distillation for Few-Shot Anomaly Detection
by: Dong, Le, et al.
Published: (2026)
by: Dong, Le, et al.
Published: (2026)
Cross-Self KV Cache Pruning for Efficient Vision-Language Inference
by: Pei, Xiaohuan, et al.
Published: (2024)
by: Pei, Xiaohuan, et al.
Published: (2024)
Efficient Point Cloud Classification via Offline Distillation Framework and Negative-Weight Self-Distillation Technique
by: Zheng, Qiang, et al.
Published: (2024)
by: Zheng, Qiang, et al.
Published: (2024)
2023 Low-Power Computer Vision Challenge (LPCVC) Summary
by: Chen, Leo, et al.
Published: (2024)
by: Chen, Leo, et al.
Published: (2024)
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024)
by: Wen, Junjie, et al.
Published: (2024)
Distill Gold from Massive Ores: Bi-level Data Pruning towards Efficient Dataset Distillation
by: Xu, Yue, et al.
Published: (2023)
by: Xu, Yue, et al.
Published: (2023)
HierarchicalPrune: Position-Aware Compression for Large-Scale Diffusion Models
by: Kwon, Young D., et al.
Published: (2025)
by: Kwon, Young D., et al.
Published: (2025)
High-Order Progressive Trajectory Matching for Medical Image Dataset Distillation
by: Dong, Le, et al.
Published: (2025)
by: Dong, Le, et al.
Published: (2025)
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers
by: Ma, Ji, et al.
Published: (2025)
by: Ma, Ji, et al.
Published: (2025)
TT-MPD: Test Time Model Pruning and Distillation
by: Wu, Haihang, et al.
Published: (2024)
by: Wu, Haihang, et al.
Published: (2024)
HiPrune: Hierarchical Attention for Efficient Token Pruning in Vision-Language Models
by: Liu, Jizhihui, et al.
Published: (2025)
by: Liu, Jizhihui, et al.
Published: (2025)
Disparity-based Stereo Image Compression with Aligned Cross-View Priors
by: Zhai, Yongqi, et al.
Published: (2022)
by: Zhai, Yongqi, et al.
Published: (2022)
PD$^{2}$GS: Part-Level Decoupling and Continuous Deformation of Articulated Objects via Gaussian Splatting
by: Wang, Haowen, et al.
Published: (2025)
by: Wang, Haowen, et al.
Published: (2025)
S2HPruner: Soft-to-Hard Distillation Bridges the Discretization Gap in Pruning
by: Lin, Weihao, et al.
Published: (2024)
by: Lin, Weihao, et al.
Published: (2024)
ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models
by: Ye, Wencheng, et al.
Published: (2025)
by: Ye, Wencheng, et al.
Published: (2025)
TernaryCLIP: Efficiently Compressing Vision-Language Models with Ternary Weights and Distilled Knowledge
by: Zhang, Shu-Hao, et al.
Published: (2025)
by: Zhang, Shu-Hao, et al.
Published: (2025)
Token Pruning for Caching Better: 9 Times Acceleration on Stable Diffusion for Free
by: Zhang, Evelyn, et al.
Published: (2024)
by: Zhang, Evelyn, et al.
Published: (2024)
Block-based Symmetric Pruning and Fusion for Efficient Vision Transformers
by: Hsieh, Yi-Kuan, et al.
Published: (2025)
by: Hsieh, Yi-Kuan, et al.
Published: (2025)
Efficient Learned Image Compression Through Knowledge Distillation
by: Allemand, Fabien, et al.
Published: (2025)
by: Allemand, Fabien, et al.
Published: (2025)
PUMA: Layer-Pruned Language Model for Efficient Unified Multimodal Retrieval with Modality-Adaptive Learning
by: Lyu, Yibo, et al.
Published: (2025)
by: Lyu, Yibo, et al.
Published: (2025)
Efficient Masked Image Compression with Position-Indexed Self-Attention
by: Dai, Chengjie, et al.
Published: (2025)
by: Dai, Chengjie, et al.
Published: (2025)
Trimming the Fat: Efficient Compression of 3D Gaussian Splats through Pruning
by: Ali, Muhammad Salman, et al.
Published: (2024)
by: Ali, Muhammad Salman, et al.
Published: (2024)
Singular Value Scaling: Efficient Generative Model Compression via Pruned Weights Refinement
by: Kim, Hyeonjin, et al.
Published: (2024)
by: Kim, Hyeonjin, et al.
Published: (2024)
Distilling the Knowledge in Data Pruning
by: Ben-Baruch, Emanuel, et al.
Published: (2024)
by: Ben-Baruch, Emanuel, et al.
Published: (2024)
How Should Video LLMs Output Time? An Analysis of Efficient Temporal Grounding Paradigms
by: Jin, Shengji, et al.
Published: (2026)
by: Jin, Shengji, et al.
Published: (2026)
Similar Items
-
LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
by: Zhu, Yichen, et al.
Published: (2024) -
Compressing Multi-Task Model for Autonomous Driving via Pruning and Knowledge Distillation
by: Wang, Jiayuan, et al.
Published: (2025) -
Comb, Prune, Distill: Towards Unified Pruning for Vision Model Compression
by: Schmitt, Jonas, et al.
Published: (2024) -
Language-Conditioned Robotic Manipulation with Fast and Slow Thinking
by: Zhu, Minjie, et al.
Published: (2024) -
Object-Centric Instruction Augmentation for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024)