MI-to-Mid Distilled Compression (M2M-DC): An Hybrid-Information-Guided-Block Pruning with Progressive Inner Slicing Approach to Model Compression
Fuente:
arXiv
Guardado en:
| Autores principales: | Levine, Lionel, Oskouie, Haniyeh Ehsani, Ghiasvand, Sajjad, Sarrafzadeh, Majid |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Exploring Cross-model Neuronal Correlations in the Context of Predicting Model Performance and Generalizability
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2024)
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2024)
Exploring the Impact of Dataset Statistical Effect Size on Model Performance and Data Sample Size Sufficiency
por: Hatamian, Arya, et al.
Publicado: (2025)
por: Hatamian, Arya, et al.
Publicado: (2025)
Few-Shot Adversarial Low-Rank Fine-Tuning of Vision-Language Models
por: Ghiasvand, Sajjad, et al.
Publicado: (2025)
por: Ghiasvand, Sajjad, et al.
Publicado: (2025)
MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation
por: Ghiasvand, Sajjad, et al.
Publicado: (2026)
por: Ghiasvand, Sajjad, et al.
Publicado: (2026)
Leveraging Large Language Models and Topic Modeling for Toxicity Classification
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2024)
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2024)
Interpretation of Neural Networks is Susceptible to Universal Adversarial Perturbations
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2022)
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2022)
Attack on Scene Flow using Point Clouds
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2024)
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2024)
PRISM: A Transformer-based Language Model of Structured Clinical Event Data
por: Levine, Lionel, et al.
Publicado: (2025)
por: Levine, Lionel, et al.
Publicado: (2025)
PRISM-Consult: A Panel-of-Experts Architecture for Clinician-Aligned Diagnosis
por: Levine, Lionel, et al.
Publicado: (2025)
por: Levine, Lionel, et al.
Publicado: (2025)
Test-Time Defense Against Adversarial Attacks via Stochastic Resonance of Latent Ensembles
por: Lao, Dong, et al.
Publicado: (2025)
por: Lao, Dong, et al.
Publicado: (2025)
Model Compression using Progressive Channel Pruning
por: Guo, Jinyang, et al.
Publicado: (2025)
por: Guo, Jinyang, et al.
Publicado: (2025)
Comb, Prune, Distill: Towards Unified Pruning for Vision Model Compression
por: Schmitt, Jonas, et al.
Publicado: (2024)
por: Schmitt, Jonas, et al.
Publicado: (2024)
EPSD: Early Pruning with Self-Distillation for Efficient Model Compression
por: Chen, Dong, et al.
Publicado: (2024)
por: Chen, Dong, et al.
Publicado: (2024)
LAPTOP-Diff: Layer Pruning and Normalized Distillation for Compressing Diffusion Models
por: Zhang, Dingkun, et al.
Publicado: (2024)
por: Zhang, Dingkun, et al.
Publicado: (2024)
Temporal Action Detection Model Compression by Progressive Block Drop
por: Chen, Xiaoyong, et al.
Publicado: (2025)
por: Chen, Xiaoyong, et al.
Publicado: (2025)
Decentralized Low-Rank Fine-Tuning of Large Language Models
por: Ghiasvand, Sajjad, et al.
Publicado: (2025)
por: Ghiasvand, Sajjad, et al.
Publicado: (2025)
pFedMMA: Personalized Federated Fine-Tuning with Multi-Modal Adapter for Vision-Language Models
por: Ghiasvand, Sajjad, et al.
Publicado: (2025)
por: Ghiasvand, Sajjad, et al.
Publicado: (2025)
Large Multimodal Model Compression via Efficient Pruning and Distillation at AntGroup
por: Wang, Maolin, et al.
Publicado: (2023)
por: Wang, Maolin, et al.
Publicado: (2023)
Compressing Multi-Task Model for Autonomous Driving via Pruning and Knowledge Distillation
por: Wang, Jiayuan, et al.
Publicado: (2025)
por: Wang, Jiayuan, et al.
Publicado: (2025)
Prune-then-Quantize or Quantize-then-Prune? Understanding the Impact of Compression Order in Joint Model Compression
por: Kim, Minjun, et al.
Publicado: (2026)
por: Kim, Minjun, et al.
Publicado: (2026)
Thanos: A Block-wise Pruning Algorithm for Efficient Large Language Model Compression
por: Ilin, Ivan, et al.
Publicado: (2025)
por: Ilin, Ivan, et al.
Publicado: (2025)
Prune-Quantize-Distill: An Ordered Pipeline for Efficient Neural Network Compression
por: Zhou, Longsheng, et al.
Publicado: (2026)
por: Zhou, Longsheng, et al.
Publicado: (2026)
PACE: Prune-And-Compress Ensemble Models
por: Akkerman, Fabian, et al.
Publicado: (2026)
por: Akkerman, Fabian, et al.
Publicado: (2026)
RDBE: Reasoning Distillation-Based Evaluation Enhances Automatic Essay Scoring
por: Mohammadkhani, Ali Ghiasvand
Publicado: (2024)
por: Mohammadkhani, Ali Ghiasvand
Publicado: (2024)
REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations
por: Ghiasvand, Sajjad, et al.
Publicado: (2026)
por: Ghiasvand, Sajjad, et al.
Publicado: (2026)
UniComp: A Unified Evaluation of Large Language Model Compression via Pruning, Quantization and Distillation
por: von Rad, Jonathan, et al.
Publicado: (2026)
por: von Rad, Jonathan, et al.
Publicado: (2026)
IG-Pruning: Input-Guided Block Pruning for Large Language Models
por: Qiao, Kangyu, et al.
Publicado: (2025)
por: Qiao, Kangyu, et al.
Publicado: (2025)
PC-LoRA: Low-Rank Adaptation for Progressive Model Compression with Knowledge Distillation
por: Hwang, Injoon, et al.
Publicado: (2024)
por: Hwang, Injoon, et al.
Publicado: (2024)
Advancing Compressed Video Action Recognition through Progressive Knowledge Distillation
por: Soufleri, Efstathia, et al.
Publicado: (2024)
por: Soufleri, Efstathia, et al.
Publicado: (2024)
Compressibility measurement of the thermal MI--BG transition in an optical lattice
por: Russ, Phil, et al.
Publicado: (2025)
por: Russ, Phil, et al.
Publicado: (2025)
Minitron-SSM: Efficient Hybrid Language Model Compression through Group-Aware SSM Pruning
por: Taghibakhshi, Ali, et al.
Publicado: (2025)
por: Taghibakhshi, Ali, et al.
Publicado: (2025)
Hybrid Pruning: In-Situ Compression of Self-Supervised Speech Models for Speaker Verification and Anti-Spoofing
por: Peng, Junyi, et al.
Publicado: (2025)
por: Peng, Junyi, et al.
Publicado: (2025)
Universal Graph Compression: Stochastic Block Models
por: Bhatt, Alankrita, et al.
Publicado: (2020)
por: Bhatt, Alankrita, et al.
Publicado: (2020)
SGLP: A Similarity Guided Fast Layer Partition Pruning for Compressing Large Deep Models
por: Li, Yuqi, et al.
Publicado: (2024)
por: Li, Yuqi, et al.
Publicado: (2024)
Shapley Pruning for Neural Network Compression
por: Adamczewski, Kamil, et al.
Publicado: (2024)
por: Adamczewski, Kamil, et al.
Publicado: (2024)
LRCP: Low-Rank Compressibility Guided Visual Token Pruning for Efficient LVLMs
por: Lu, Hongyu, et al.
Publicado: (2026)
por: Lu, Hongyu, et al.
Publicado: (2026)
Neural Texture Block Compression
por: Fujieda, Shin, et al.
Publicado: (2024)
por: Fujieda, Shin, et al.
Publicado: (2024)
PPC-GPT: Federated Task-Specific Compression of Large Language Models via Pruning and Chain-of-Thought Distillation
por: Fan, Tao, et al.
Publicado: (2025)
por: Fan, Tao, et al.
Publicado: (2025)
SliceGPT: Compress Large Language Models by Deleting Rows and Columns
por: Ashkboos, Saleh, et al.
Publicado: (2024)
por: Ashkboos, Saleh, et al.
Publicado: (2024)
HPM-KD: Hierarchical Progressive Multi-Teacher Framework for Knowledge Distillation and Efficient Model Compression
por: Haase, Gustavo Coelho, et al.
Publicado: (2025)
por: Haase, Gustavo Coelho, et al.
Publicado: (2025)
Ejemplares similares
-
Exploring Cross-model Neuronal Correlations in the Context of Predicting Model Performance and Generalizability
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2024) -
Exploring the Impact of Dataset Statistical Effect Size on Model Performance and Data Sample Size Sufficiency
por: Hatamian, Arya, et al.
Publicado: (2025) -
Few-Shot Adversarial Low-Rank Fine-Tuning of Vision-Language Models
por: Ghiasvand, Sajjad, et al.
Publicado: (2025) -
MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation
por: Ghiasvand, Sajjad, et al.
Publicado: (2026) -
Leveraging Large Language Models and Topic Modeling for Toxicity Classification
por: Oskouie, Haniyeh Ehsani, et al.
Publicado: (2024)