FALCON: FLOP-Aware Combinatorial Optimization for Neural Network Pruning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Meng, Xiang, Chen, Wenyu, Benbaki, Riade, Mazumder, Rahul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Modeling with Categorical Features via Exact Fusion and Sparsity Regularisation
von: Behdin, Kayhan, et al.
Veröffentlicht: (2026)
von: Behdin, Kayhan, et al.
Veröffentlicht: (2026)
OSSCAR: One-Shot Structured Pruning in Vision and Language Models with Combinatorial Optimization
von: Meng, Xiang, et al.
Veröffentlicht: (2024)
von: Meng, Xiang, et al.
Veröffentlicht: (2024)
ALPS: Improved Optimization for Highly Sparse One-Shot Pruning for Large Language Models
von: Meng, Xiang, et al.
Veröffentlicht: (2024)
von: Meng, Xiang, et al.
Veröffentlicht: (2024)
Sparse Gaussian Graphical Models with Discrete Optimization: Computational and Statistical Perspectives
von: Behdin, Kayhan, et al.
Veröffentlicht: (2023)
von: Behdin, Kayhan, et al.
Veröffentlicht: (2023)
Preserving Deep Representations In One-Shot Pruning: A Hessian-Free Second-Order Optimization Framework
von: Lucas, Ryan, et al.
Veröffentlicht: (2024)
von: Lucas, Ryan, et al.
Veröffentlicht: (2024)
MOONSHOT : A Framework for Multi-Objective Pruning of Vision and Large Language Models
von: Afriat, Gabriel, et al.
Veröffentlicht: (2026)
von: Afriat, Gabriel, et al.
Veröffentlicht: (2026)
3BASiL: An Algorithmic Framework for Sparse plus Low-Rank Compression of LLMs
von: Makni, Mehdi, et al.
Veröffentlicht: (2026)
von: Makni, Mehdi, et al.
Veröffentlicht: (2026)
TSENOR: Highly-Efficient Algorithm for Finding Transposable N:M Sparse Masks
von: Meng, Xiang, et al.
Veröffentlicht: (2025)
von: Meng, Xiang, et al.
Veröffentlicht: (2025)
FAST: An Optimization Framework for Fast Additive Segmentation in Transparent ML
von: Liu, Brian, et al.
Veröffentlicht: (2024)
von: Liu, Brian, et al.
Veröffentlicht: (2024)
Computation of Least Trimmed Squares: A Branch-and-Bound framework with Hyperplane Arrangement Enhancements
von: Meng, Xiang, et al.
Veröffentlicht: (2026)
von: Meng, Xiang, et al.
Veröffentlicht: (2026)
MOSS: Multi-Objective Optimization for Stable Rule Sets
von: Liu, Brian, et al.
Veröffentlicht: (2025)
von: Liu, Brian, et al.
Veröffentlicht: (2025)
FLOP-Efficient Training: Early Stopping Based on Test-Time Compute Awareness
von: Amer, Hossam, et al.
Veröffentlicht: (2026)
von: Amer, Hossam, et al.
Veröffentlicht: (2026)
Combinatorial Optimization with Automated Graph Neural Networks
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
ADMM-Q: An Improved Hessian-based Weight Quantizer for Post-Training Quantization of Large Language Models
von: Lucas, Ryan, et al.
Veröffentlicht: (2026)
von: Lucas, Ryan, et al.
Veröffentlicht: (2026)
Decision-focused Graph Neural Networks for Combinatorial Optimization
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
Randomization Can Reduce Both Bias and Variance: A Case Study in Random Forests
von: Liu, Brian, et al.
Veröffentlicht: (2024)
von: Liu, Brian, et al.
Veröffentlicht: (2024)
Sparse NMF with Archetypal Regularization: Computational and Robustness Properties
von: Behdin, Kayhan, et al.
Veröffentlicht: (2021)
von: Behdin, Kayhan, et al.
Veröffentlicht: (2021)
SequentialAttention++ for Block Sparsification: Differentiable Pruning Meets Combinatorial Optimization
von: Yasuda, Taisuke, et al.
Veröffentlicht: (2024)
von: Yasuda, Taisuke, et al.
Veröffentlicht: (2024)
Robust Neural Pruning with Gradient Sampling Optimization for Residual Neural Networks
von: Yun, Juyoung
Veröffentlicht: (2023)
von: Yun, Juyoung
Veröffentlicht: (2023)
Neural networks can be FLOP-efficient integrators of 1D oscillatory integrands
von: Sinha, Anshuman, et al.
Veröffentlicht: (2024)
von: Sinha, Anshuman, et al.
Veröffentlicht: (2024)
Mixed Sparsity Training: Achieving 4$\times$ FLOP Reduction for Transformer Pretraining
von: Hu, Pihe, et al.
Veröffentlicht: (2024)
von: Hu, Pihe, et al.
Veröffentlicht: (2024)
Extracting Interpretable Models from Tree Ensembles: Computational and Statistical Perspectives
von: Liu, Brian, et al.
Veröffentlicht: (2025)
von: Liu, Brian, et al.
Veröffentlicht: (2025)
End-to-end Feature Selection Approach for Learning Skinny Trees
von: Ibrahim, Shibal, et al.
Veröffentlicht: (2023)
von: Ibrahim, Shibal, et al.
Veröffentlicht: (2023)
Problems with Chinchilla Approach 2: Systematic Biases in IsoFLOP Parabola Fits
von: Czech, Eric, et al.
Veröffentlicht: (2026)
von: Czech, Eric, et al.
Veröffentlicht: (2026)
Application-Specific Component-Aware Structured Pruning of Deep Neural Networks in Control via Soft Coefficient Optimization
von: Sundaram, Ganesh, et al.
Veröffentlicht: (2025)
von: Sundaram, Ganesh, et al.
Veröffentlicht: (2025)
Annealing Machine-assisted Learning of Graph Neural Network for Combinatorial Optimization
von: Loyola, Pablo, et al.
Veröffentlicht: (2025)
von: Loyola, Pablo, et al.
Veröffentlicht: (2025)
Probing Neural Combinatorial Optimization Models
von: Zhang, Zhiqin, et al.
Veröffentlicht: (2025)
von: Zhang, Zhiqin, et al.
Veröffentlicht: (2025)
Spectral Pruning for Recurrent Neural Networks
von: Furuya, Takashi, et al.
Veröffentlicht: (2021)
von: Furuya, Takashi, et al.
Veröffentlicht: (2021)
BOPO: Neural Combinatorial Optimization via Best-anchored and Objective-guided Preference Optimization
von: Liao, Zijun, et al.
Veröffentlicht: (2025)
von: Liao, Zijun, et al.
Veröffentlicht: (2025)
Leader Reward for POMO-Based Neural Combinatorial Optimization
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
Physics-Informed Neural Network for Concrete Manufacturing Process Optimization
von: Varghese, Sam, et al.
Veröffentlicht: (2024)
von: Varghese, Sam, et al.
Veröffentlicht: (2024)
Recurrent State Encoders for Efficient Neural Combinatorial Optimization
von: Dernedde, Tim, et al.
Veröffentlicht: (2025)
von: Dernedde, Tim, et al.
Veröffentlicht: (2025)
Multi-Action Self-Improvement for Neural Combinatorial Optimization
von: Luttmann, Laurin, et al.
Veröffentlicht: (2025)
von: Luttmann, Laurin, et al.
Veröffentlicht: (2025)
Supervised Robustness-preserving Data-free Neural Network Pruning
von: Meng, Mark Huasong, et al.
Veröffentlicht: (2022)
von: Meng, Mark Huasong, et al.
Veröffentlicht: (2022)
Neural Tractability via Structure: Learning-Augmented Algorithms for Graph Combinatorial Optimization
von: Li, Jialiang, et al.
Veröffentlicht: (2025)
von: Li, Jialiang, et al.
Veröffentlicht: (2025)
Rethinking Efficiency in Neural Combinatorial Optimization: Batched Preference Optimization with Mamba
von: Xu, Zhenxing, et al.
Veröffentlicht: (2026)
von: Xu, Zhenxing, et al.
Veröffentlicht: (2026)
Resource-Aware Neural Network Pruning Using Graph-based Reinforcement Learning
von: Balemans, Dieter, et al.
Veröffentlicht: (2025)
von: Balemans, Dieter, et al.
Veröffentlicht: (2025)
LatentMoE: Toward Optimal Accuracy per FLOP and Parameter in Mixture of Experts
von: Elango, Venmugil, et al.
Veröffentlicht: (2026)
von: Elango, Venmugil, et al.
Veröffentlicht: (2026)
Mutual Information Preserving Neural Network Pruning
von: Westphal, Charles, et al.
Veröffentlicht: (2024)
von: Westphal, Charles, et al.
Veröffentlicht: (2024)
Pruning and Quantization Impact on Graph Neural Networks
von: Khedri, Khatoon, et al.
Veröffentlicht: (2025)
von: Khedri, Khatoon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Modeling with Categorical Features via Exact Fusion and Sparsity Regularisation
von: Behdin, Kayhan, et al.
Veröffentlicht: (2026) -
OSSCAR: One-Shot Structured Pruning in Vision and Language Models with Combinatorial Optimization
von: Meng, Xiang, et al.
Veröffentlicht: (2024) -
ALPS: Improved Optimization for Highly Sparse One-Shot Pruning for Large Language Models
von: Meng, Xiang, et al.
Veröffentlicht: (2024) -
Sparse Gaussian Graphical Models with Discrete Optimization: Computational and Statistical Perspectives
von: Behdin, Kayhan, et al.
Veröffentlicht: (2023) -
Preserving Deep Representations In One-Shot Pruning: A Hessian-Free Second-Order Optimization Framework
von: Lucas, Ryan, et al.
Veröffentlicht: (2024)