Towards Universal & Efficient Model Compression via Exponential Torque Pruning
Fuente:
arXiv
Guardado en:
| Autores principales: | Modi, Sarthak Ketanbhai, Lim, Zi Pong, Kuchhal, Shourya, Cao, Yushi, Cheng, Yupeng, Teo, Yon Shin, Lin, Shang-Wei, Li, Zhiming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards Calibration Enhanced Network by Inverse Adversarial Attack
por: Cheng, Yupeng, et al.
Publicado: (2025)
por: Cheng, Yupeng, et al.
Publicado: (2025)
SOSAE: Self-Organizing Sparse AutoEncoder
por: Modi, Sarthak Ketanbhai, et al.
Publicado: (2025)
por: Modi, Sarthak Ketanbhai, et al.
Publicado: (2025)
An Online Reference-Free Evaluation Framework for Flowchart Image-to-Code Generation
por: Nguyen, Giang Son, et al.
Publicado: (2026)
por: Nguyen, Giang Son, et al.
Publicado: (2026)
Learning More from Less: Exploiting Counterfactuals for Data-Efficient Chart Understanding
por: Bao, Jianzhu, et al.
Publicado: (2026)
por: Bao, Jianzhu, et al.
Publicado: (2026)
TT-Sparse: Learning Sparse Rule Models with Differentiable Truth Tables
por: Soegeng, Hans Farrell, et al.
Publicado: (2026)
por: Soegeng, Hans Farrell, et al.
Publicado: (2026)
LLMs for Relational Reasoning: How Far are We?
por: Li, Zhiming, et al.
Publicado: (2024)
por: Li, Zhiming, et al.
Publicado: (2024)
Pruning as a Defense: Reducing Memorization in Large Language Models
por: Gupta, Mansi, et al.
Publicado: (2025)
por: Gupta, Mansi, et al.
Publicado: (2025)
Towards Efficient Machine Learning Method for IoT DDoS Attack Detection
por: Modi, P
Publicado: (2024)
por: Modi, P
Publicado: (2024)
Exponential ergodicity and finite-dimensional approximation for Markovian lifts of stochastic Volterra equations
por: Hamaguchi, Yushi
Publicado: (2026)
por: Hamaguchi, Yushi
Publicado: (2026)
Comb, Prune, Distill: Towards Unified Pruning for Vision Model Compression
por: Schmitt, Jonas, et al.
Publicado: (2024)
por: Schmitt, Jonas, et al.
Publicado: (2024)
Towards Efficient VLMs: Information-Theoretic Driven Compression via Adaptive Structural Pruning
por: Xu, Zhaoqi, et al.
Publicado: (2025)
por: Xu, Zhaoqi, et al.
Publicado: (2025)
Exponential Random Graph models for Little Networks
por: Yon, George G. Vega, et al.
Publicado: (2019)
por: Yon, George G. Vega, et al.
Publicado: (2019)
FPGA Integrated Distributed Homogenous Clustered AODV Routing Analysis for Mobile Ad Hoc Networks
por: Arvind Kumar, et al.
Publicado: (2025)
por: Arvind Kumar, et al.
Publicado: (2025)
Locality-Aware Redundancy Pruning for LLM Depth Compression
por: Yun, Vincent-Daniel, et al.
Publicado: (2026)
por: Yun, Vincent-Daniel, et al.
Publicado: (2026)
Pruning then Reweighting: Towards Data-Efficient Training of Diffusion Models
por: Li, Yize, et al.
Publicado: (2024)
por: Li, Yize, et al.
Publicado: (2024)
Structured vs. Unstructured Pruning: An Exponential Gap
por: Ferre', Davide, et al.
Publicado: (2026)
por: Ferre', Davide, et al.
Publicado: (2026)
MOB-Net: Limb-modularized Uncertainty Torque Learning of Humanoids for Sensorless External Torque Estimation
por: Lim, Daegyu, et al.
Publicado: (2024)
por: Lim, Daegyu, et al.
Publicado: (2024)
SHRP: Specialized Head Routing and Pruning for Efficient Encoder Compression
por: Su, Zeli, et al.
Publicado: (2025)
por: Su, Zeli, et al.
Publicado: (2025)
Integrating Pruning with Quantization for Efficient Deep Neural Networks Compression
por: Makenali, Sara, et al.
Publicado: (2025)
por: Makenali, Sara, et al.
Publicado: (2025)
EPSD: Early Pruning with Self-Distillation for Efficient Model Compression
por: Chen, Dong, et al.
Publicado: (2024)
por: Chen, Dong, et al.
Publicado: (2024)
Load Restoration in Islanded Microgrids: Formulation and Solution Strategies
por: Bose, Shourya, et al.
Publicado: (2021)
por: Bose, Shourya, et al.
Publicado: (2021)
Language Models Entangle Language and Culture
por: Jain, Shourya, et al.
Publicado: (2026)
por: Jain, Shourya, et al.
Publicado: (2026)
EFPC: Towards Efficient and Flexible Prompt Compression
por: Cao, Yun-Hao, et al.
Publicado: (2025)
por: Cao, Yun-Hao, et al.
Publicado: (2025)
Tight Compression: Compressing CNN Through Fine-Grained Pruning and Weight Permutation for Efficient Implementation
por: Chen, Xizi, et al.
Publicado: (2021)
por: Chen, Xizi, et al.
Publicado: (2021)
Prune-then-Quantize or Quantize-then-Prune? Understanding the Impact of Compression Order in Joint Model Compression
por: Kim, Minjun, et al.
Publicado: (2026)
por: Kim, Minjun, et al.
Publicado: (2026)
Torque Hyperuniformity in Frictional Granular Matter - Theory and Experiments
por: Shang, Jin, et al.
Publicado: (2026)
por: Shang, Jin, et al.
Publicado: (2026)
On the Feasibility of Fidelity$^-$ for Graph Pruning
por: Shin, Yong-Min, et al.
Publicado: (2024)
por: Shin, Yong-Min, et al.
Publicado: (2024)
Prune-Quantize-Distill: An Ordered Pipeline for Efficient Neural Network Compression
por: Zhou, Longsheng, et al.
Publicado: (2026)
por: Zhou, Longsheng, et al.
Publicado: (2026)
Don't Let Micropayments Penalize You--Experience from the City University of Hong Kong
por: Ching, Steve H., et al.
Publicado: (2009)
por: Ching, Steve H., et al.
Publicado: (2009)
EvoPrune: Early-Stage Visual Token Pruning for Efficient MLLMs
por: Chen, Yuhao, et al.
Publicado: (2026)
por: Chen, Yuhao, et al.
Publicado: (2026)
Towards Robust Pruning: An Adaptive Knowledge-Retention Pruning Strategy for Language Models
por: Li, Jianwei, et al.
Publicado: (2023)
por: Li, Jianwei, et al.
Publicado: (2023)
DeepPrune: Parallel Scaling without Inter-trace Redundancy
por: Tu, Shangqing, et al.
Publicado: (2025)
por: Tu, Shangqing, et al.
Publicado: (2025)
Class-Aware Pruning for Efficient Neural Networks
por: Jiang, Mengnan, et al.
Publicado: (2023)
por: Jiang, Mengnan, et al.
Publicado: (2023)
Designing Magnetic Topological Insulator Trilayers for Highly-Efficient Spin-Orbit Torque Switching
por: Zhou, Ling-Jie, et al.
Publicado: (2026)
por: Zhou, Ling-Jie, et al.
Publicado: (2026)
Volumetric mediations: Atmospheres of crisis and unbelonging in humanitarian drone documentaries
por: Beryl Pong
Publicado: (2025)
por: Beryl Pong
Publicado: (2025)
C-SWAP: Explainability-Aware Structured Pruning for Efficient Neural Networks Compression
por: Bauvin, Baptiste, et al.
Publicado: (2025)
por: Bauvin, Baptiste, et al.
Publicado: (2025)
LRCP: Low-Rank Compressibility Guided Visual Token Pruning for Efficient LVLMs
por: Lu, Hongyu, et al.
Publicado: (2026)
por: Lu, Hongyu, et al.
Publicado: (2026)
Large Multimodal Model Compression via Efficient Pruning and Distillation at AntGroup
por: Wang, Maolin, et al.
Publicado: (2023)
por: Wang, Maolin, et al.
Publicado: (2023)
Automatic Joint Structured Pruning and Quantization for Efficient Neural Network Training and Compression
por: Qu, Xiaoyi, et al.
Publicado: (2025)
por: Qu, Xiaoyi, et al.
Publicado: (2025)
Focus on the Core: Efficient Attention via Pruned Token Compression for Document Classification
por: Yun, Jungmin, et al.
Publicado: (2024)
por: Yun, Jungmin, et al.
Publicado: (2024)
Ejemplares similares
-
Towards Calibration Enhanced Network by Inverse Adversarial Attack
por: Cheng, Yupeng, et al.
Publicado: (2025) -
SOSAE: Self-Organizing Sparse AutoEncoder
por: Modi, Sarthak Ketanbhai, et al.
Publicado: (2025) -
An Online Reference-Free Evaluation Framework for Flowchart Image-to-Code Generation
por: Nguyen, Giang Son, et al.
Publicado: (2026) -
Learning More from Less: Exploiting Counterfactuals for Data-Efficient Chart Understanding
por: Bao, Jianzhu, et al.
Publicado: (2026) -
TT-Sparse: Learning Sparse Rule Models with Differentiable Truth Tables
por: Soegeng, Hans Farrell, et al.
Publicado: (2026)