ONNXPruner: ONNX-Based General Model Pruning Adapter
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ren, Dongdong, Li, Wenbin, Ding, Tianyu, Wang, Lei, Fan, Qi, Huo, Jing, Pan, Hongbing, Gao, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Selective Quantization Tuner for ONNX Models
von: Louloudakis, Nikolaos, et al.
Veröffentlicht: (2025)
von: Louloudakis, Nikolaos, et al.
Veröffentlicht: (2025)
Enhancing Trust-Region Bayesian Optimization via Newton Methods
von: Chen, Quanlin, et al.
Veröffentlicht: (2025)
von: Chen, Quanlin, et al.
Veröffentlicht: (2025)
Probe Pruning: Accelerating LLMs through Dynamic Pruning via Model-Probing
von: Le, Qi, et al.
Veröffentlicht: (2025)
von: Le, Qi, et al.
Veröffentlicht: (2025)
Analysis of Failures and Risks in Deep Learning Model Converters: A Case Study in the ONNX Ecosystem
von: Jajal, Purvish, et al.
Veröffentlicht: (2023)
von: Jajal, Purvish, et al.
Veröffentlicht: (2023)
Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization
von: Wang, Tianyu, et al.
Veröffentlicht: (2026)
von: Wang, Tianyu, et al.
Veröffentlicht: (2026)
DiTOX: Fault Detection and Localization in the ONNX Optimizer
von: Louloudakis, Nikolaos, et al.
Veröffentlicht: (2025)
von: Louloudakis, Nikolaos, et al.
Veröffentlicht: (2025)
Model-Based Offline Reinforcement Learning with Adversarial Data Augmentation
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
The Structural Scalpel: Automated Contiguous Layer Pruning for Large Language Models
von: Lu, Yao, et al.
Veröffentlicht: (2025)
von: Lu, Yao, et al.
Veröffentlicht: (2025)
Towards Empowerment Gain through Causal Structure Learning in Model-Based RL
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
Block Circulant Adapter for Large Language Models
von: Ding, Xinyu, et al.
Veröffentlicht: (2025)
von: Ding, Xinyu, et al.
Veröffentlicht: (2025)
IDEA Prune: An Integrated Enlarge-and-Prune Pipeline in Generative Language Model Pretraining
von: Li, Yixiao, et al.
Veröffentlicht: (2025)
von: Li, Yixiao, et al.
Veröffentlicht: (2025)
ONNX-Net: Towards Universal Representations and Instant Performance Prediction for Neural Architectures
von: Qin, Shiwen, et al.
Veröffentlicht: (2025)
von: Qin, Shiwen, et al.
Veröffentlicht: (2025)
Federated Adapter on Foundation Models: An Out-Of-Distribution Approach
von: Yang, Yiyuan, et al.
Veröffentlicht: (2025)
von: Yang, Yiyuan, et al.
Veröffentlicht: (2025)
Federated Multimodal Learning with Dual Adapters and Selective Pruning for Communication and Computational Efficiency
von: Nguyen, Duy Phuong, et al.
Veröffentlicht: (2025)
von: Nguyen, Duy Phuong, et al.
Veröffentlicht: (2025)
Adaptive Pruning of Pretrained Transformer via Differential Inclusions
von: Ding, Yizhuo, et al.
Veröffentlicht: (2025)
von: Ding, Yizhuo, et al.
Veröffentlicht: (2025)
Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning
von: Xia, Mengzhou, et al.
Veröffentlicht: (2023)
von: Xia, Mengzhou, et al.
Veröffentlicht: (2023)
AIGC for Industrial Time Series: From Deep Generative Models to Large Generative Models
von: Ren, Lei, et al.
Veröffentlicht: (2024)
von: Ren, Lei, et al.
Veröffentlicht: (2024)
Reconstruct the Pruned Model without Any Retraining
von: Wang, Pingjie, et al.
Veröffentlicht: (2024)
von: Wang, Pingjie, et al.
Veröffentlicht: (2024)
Causal Information Prioritization for Efficient Reinforcement Learning
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
von: Cao, Hongye, et al.
Veröffentlicht: (2025)
SwiftPrune: Hessian-Free Weight Pruning for Large Language Models
von: Kang, Yuhan, et al.
Veröffentlicht: (2025)
von: Kang, Yuhan, et al.
Veröffentlicht: (2025)
FastMMoE: Accelerating Multimodal Large Language Models through Dynamic Expert Activation and Routing-Aware Token Pruning
von: Xia, Guoyang, et al.
Veröffentlicht: (2025)
von: Xia, Guoyang, et al.
Veröffentlicht: (2025)
Think Before You Prune: Self-Reflective Structured Pruning for Reasoning Language Models
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
Simple, Efficient and Scalable Structure-aware Adapter Boosts Protein Language Models
von: Tan, Yang, et al.
Veröffentlicht: (2024)
von: Tan, Yang, et al.
Veröffentlicht: (2024)
Dual-Personalizing Adapter for Federated Foundation Models
von: Yang, Yiyuan, et al.
Veröffentlicht: (2024)
von: Yang, Yiyuan, et al.
Veröffentlicht: (2024)
Non-transferable Pruning
von: Ding, Ruyi, et al.
Veröffentlicht: (2024)
von: Ding, Ruyi, et al.
Veröffentlicht: (2024)
Function-Guided Conditional Generation Using Protein Language Models with Adapters
von: Yang, Jason, et al.
Veröffentlicht: (2024)
von: Yang, Jason, et al.
Veröffentlicht: (2024)
A Generic Layer Pruning Method for Signal Modulation Recognition Deep Learning Models
von: Lu, Yao, et al.
Veröffentlicht: (2024)
von: Lu, Yao, et al.
Veröffentlicht: (2024)
Hadamard Adapter: An Extreme Parameter-Efficient Adapter Tuning Method for Pre-trained Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Winners with Confidence: Discrete Argmin Inference with an Application to Model Selection
von: Zhang, Tianyu, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyu, et al.
Veröffentlicht: (2024)
Online Estimation with Rolling Validation: Adaptive Nonparametric Estimation with Streaming Data
von: Zhang, Tianyu, et al.
Veröffentlicht: (2023)
von: Zhang, Tianyu, et al.
Veröffentlicht: (2023)
CausalTAD: Causal Implicit Generative Model for Debiased Online Trajectory Anomaly Detection
von: Li, Wenbin, et al.
Veröffentlicht: (2024)
von: Li, Wenbin, et al.
Veröffentlicht: (2024)
SLoPe: Double-Pruned Sparse Plus Lazy Low-Rank Adapter Pretraining of LLMs
von: Mozaffari, Mohammad, et al.
Veröffentlicht: (2024)
von: Mozaffari, Mohammad, et al.
Veröffentlicht: (2024)
Batch Loss Score for Dynamic Data Pruning
von: Zhou, Qing, et al.
Veröffentlicht: (2026)
von: Zhou, Qing, et al.
Veröffentlicht: (2026)
Archimedean Copula Inference via Taylor-Mode AD
von: Yang, Cambridge, et al.
Veröffentlicht: (2026)
von: Yang, Cambridge, et al.
Veröffentlicht: (2026)
Continual Low-Rank Adapters for LLM-based Generative Recommender Systems
von: Yoo, Hyunsik, et al.
Veröffentlicht: (2025)
von: Yoo, Hyunsik, et al.
Veröffentlicht: (2025)
Unbiased Dynamic Pruning for Efficient Group-Based Policy Optimization
von: Zhu, Haodong, et al.
Veröffentlicht: (2026)
von: Zhu, Haodong, et al.
Veröffentlicht: (2026)
Holistic Adversarially Robust Pruning
von: Zhao, Qi, et al.
Veröffentlicht: (2024)
von: Zhao, Qi, et al.
Veröffentlicht: (2024)
Beyond Efficiency: Molecular Data Pruning for Enhanced Generalization
von: Chen, Dingshuo, et al.
Veröffentlicht: (2024)
von: Chen, Dingshuo, et al.
Veröffentlicht: (2024)
Data Pruning in Generative Diffusion Models
von: Briq, Rania, et al.
Veröffentlicht: (2024)
von: Briq, Rania, et al.
Veröffentlicht: (2024)
Pruning Weights but Not Truth: Safeguarding Truthfulness While Pruning LLMs
von: Fu, Yao, et al.
Veröffentlicht: (2025)
von: Fu, Yao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Selective Quantization Tuner for ONNX Models
von: Louloudakis, Nikolaos, et al.
Veröffentlicht: (2025) -
Enhancing Trust-Region Bayesian Optimization via Newton Methods
von: Chen, Quanlin, et al.
Veröffentlicht: (2025) -
Probe Pruning: Accelerating LLMs through Dynamic Pruning via Model-Probing
von: Le, Qi, et al.
Veröffentlicht: (2025) -
Analysis of Failures and Risks in Deep Learning Model Converters: A Case Study in the ONNX Ecosystem
von: Jajal, Purvish, et al.
Veröffentlicht: (2023) -
Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization
von: Wang, Tianyu, et al.
Veröffentlicht: (2026)