MultiPruner: Balanced Structure Removal in Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Muñoz, J. Pablo, Yuan, Jinjie, Jain, Nilesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mamba-Shedder: Post-Transformer Compression for Efficient Selective Structured State Space Models
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025)
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025)
RTTC: Reward-Guided Collaborative Test-Time Compute
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025)
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025)
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
von: Sarkar, Nilesh, et al.
Veröffentlicht: (2026)
von: Sarkar, Nilesh, et al.
Veröffentlicht: (2026)
Towards Measuring Goal-Directedness in AI Systems
von: Xu, Dylan, et al.
Veröffentlicht: (2024)
von: Xu, Dylan, et al.
Veröffentlicht: (2024)
Regime Change Hypothesis: Foundations for Decoupled Dynamics in Neural Network Training
von: Pérez-Corral, Cristian, et al.
Veröffentlicht: (2026)
von: Pérez-Corral, Cristian, et al.
Veröffentlicht: (2026)
How Data Quality Affects Machine Learning Models for Credit Risk Assessment
von: Maurino, Andrea
Veröffentlicht: (2025)
von: Maurino, Andrea
Veröffentlicht: (2025)
Enhancing Multi-Agent Collaboration with Attention-Based Actor-Critic Policies
von: Belinchon, Hugo Garrido-Lestache, et al.
Veröffentlicht: (2025)
von: Belinchon, Hugo Garrido-Lestache, et al.
Veröffentlicht: (2025)
Tracking daily paths in home contexts with RSSI fingerprinting based on UWB through deep learning models
von: Polo-Rodríguez, Aurora, et al.
Veröffentlicht: (2025)
von: Polo-Rodríguez, Aurora, et al.
Veröffentlicht: (2025)
Deep one-gate per layer networks with skip connections are universal classifiers
von: Rojas, Raul
Veröffentlicht: (2025)
von: Rojas, Raul
Veröffentlicht: (2025)
Critical appraisal of artificial intelligence for rare-event recognition: principles and pharmacovigilance case studies
von: Noren, G. Niklas, et al.
Veröffentlicht: (2025)
von: Noren, G. Niklas, et al.
Veröffentlicht: (2025)
RegExplainer: Generating Explanations for Graph Neural Networks in Regression Tasks
von: Zhang, Jiaxing, et al.
Veröffentlicht: (2023)
von: Zhang, Jiaxing, et al.
Veröffentlicht: (2023)
Kolmogorov-Arnold Networks (KAN) for Time Series Classification and Robust Analysis
von: Dong, Chang, et al.
Veröffentlicht: (2024)
von: Dong, Chang, et al.
Veröffentlicht: (2024)
Artificial Inductive Bias for Synthetic Tabular Data Generation in Data-Scarce Scenarios
von: Apellániz, Patricia A., et al.
Veröffentlicht: (2024)
von: Apellániz, Patricia A., et al.
Veröffentlicht: (2024)
Deeper Understanding of Black-box Predictions via Generalized Influence Functions
von: Lyu, Hyeonsu, et al.
Veröffentlicht: (2023)
von: Lyu, Hyeonsu, et al.
Veröffentlicht: (2023)
TiVaT: A Transformer with a Single Unified Mechanism for Capturing Asynchronous Dependencies in Multivariate Time Series Forecasting
von: Ha, Junwoo, et al.
Veröffentlicht: (2024)
von: Ha, Junwoo, et al.
Veröffentlicht: (2024)
FedLoGe: Joint Local and Generic Federated Learning under Long-tailed Data
von: Xiao, Zikai, et al.
Veröffentlicht: (2024)
von: Xiao, Zikai, et al.
Veröffentlicht: (2024)
Synthetic Tabular Data Validation: A Divergence-Based Approach
von: Apellániz, Patricia A., et al.
Veröffentlicht: (2024)
von: Apellániz, Patricia A., et al.
Veröffentlicht: (2024)
The Long Delay to Arithmetic Generalization: When Learned Representations Outrun Behavior
von: Gonzalez, Laura Gomezjurado
Veröffentlicht: (2026)
von: Gonzalez, Laura Gomezjurado
Veröffentlicht: (2026)
Study of the Proper NNUE Dataset
von: Tan, Daniel, et al.
Veröffentlicht: (2024)
von: Tan, Daniel, et al.
Veröffentlicht: (2024)
Emotion-Gradient Metacognitive RSI (Part I): Theoretical Foundations and Single-Agent Architecture
von: Ando, Rintaro
Veröffentlicht: (2025)
von: Ando, Rintaro
Veröffentlicht: (2025)
Avoiding Death through Fear Intrinsic Conditioning
von: Sanchez, Rodney, et al.
Veröffentlicht: (2025)
von: Sanchez, Rodney, et al.
Veröffentlicht: (2025)
Simulation-Driven Railway Delay Prediction: An Imitation Learning Approach
von: Elliker, Clément, et al.
Veröffentlicht: (2025)
von: Elliker, Clément, et al.
Veröffentlicht: (2025)
A Comprehensive Survey of Deep Transfer Learning for Anomaly Detection in Industrial Time Series: Methods, Applications, and Directions
von: Yan, Peng, et al.
Veröffentlicht: (2023)
von: Yan, Peng, et al.
Veröffentlicht: (2023)
Interpretability-Guided Bi-objective Optimization: Aligning Accuracy and Explainability
von: Fouladi, Kasra, et al.
Veröffentlicht: (2026)
von: Fouladi, Kasra, et al.
Veröffentlicht: (2026)
Graceful task adaptation with a bi-hemispheric RL agent
von: Nicholas, Grant, et al.
Veröffentlicht: (2024)
von: Nicholas, Grant, et al.
Veröffentlicht: (2024)
An Axiomatic Approach to General Intelligence: SANC(E3) -- Self-organizing Active Network of Concepts with Energy E3
von: Kwon, Daesuk, et al.
Veröffentlicht: (2026)
von: Kwon, Daesuk, et al.
Veröffentlicht: (2026)
Social Learning through Interactions with Other Agents: A Survey
von: Hillier, Dylan, et al.
Veröffentlicht: (2024)
von: Hillier, Dylan, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Reward Discounting for Mitigating Reward Hacking
von: Singha, Disha
Veröffentlicht: (2026)
von: Singha, Disha
Veröffentlicht: (2026)
A Data-Driven Framework for Digital Transformation in Smart Cities: Integrating AI, Dashboards, and IoT Readiness
von: Lloret, Ángel, et al.
Veröffentlicht: (2025)
von: Lloret, Ángel, et al.
Veröffentlicht: (2025)
Synthetic Data Augmentation for Medical Audio Classification: A Preliminary Evaluation
von: McShannon, David, et al.
Veröffentlicht: (2026)
von: McShannon, David, et al.
Veröffentlicht: (2026)
Enhancing Feature Selection and Interpretability in AI Regression Tasks Through Feature Attribution
von: Hinterleitner, Alexander, et al.
Veröffentlicht: (2024)
von: Hinterleitner, Alexander, et al.
Veröffentlicht: (2024)
Unsupervised Anomaly Prediction with N-BEATS and Graph Neural Network in Multi-variate Semiconductor Process Time Series
von: Sorensen, Daniel, et al.
Veröffentlicht: (2025)
von: Sorensen, Daniel, et al.
Veröffentlicht: (2025)
Event Detection via Probability Density Function Regression
von: Peng, Clark, et al.
Veröffentlicht: (2024)
von: Peng, Clark, et al.
Veröffentlicht: (2024)
Learning Actionable World Models for Industrial Process Control
von: Yan, Peng, et al.
Veröffentlicht: (2025)
von: Yan, Peng, et al.
Veröffentlicht: (2025)
Deep Reinforcement Learning for Adverse Garage Scenario Generation
von: Li, Kai
Veröffentlicht: (2024)
von: Li, Kai
Veröffentlicht: (2024)
Digitizing Touch with an Artificial Multimodal Fingertip
von: Lambeta, Mike, et al.
Veröffentlicht: (2024)
von: Lambeta, Mike, et al.
Veröffentlicht: (2024)
Holistic Audit Dataset Generation for LLM Unlearning via Knowledge Graph Traversal and Redundancy Removal
von: Jiang, Weipeng, et al.
Veröffentlicht: (2025)
von: Jiang, Weipeng, et al.
Veröffentlicht: (2025)
Position: Tensor Networks are a Valuable Asset for Green AI
von: Memmel, Eva, et al.
Veröffentlicht: (2022)
von: Memmel, Eva, et al.
Veröffentlicht: (2022)
ConciseRL: Conciseness-Guided Reinforcement Learning for Efficient Reasoning Models
von: Dumitru, Razvan-Gabriel, et al.
Veröffentlicht: (2025)
von: Dumitru, Razvan-Gabriel, et al.
Veröffentlicht: (2025)
A Hierarchical conv-LSTM and LLM Integrated Model for Holistic Stock Forecasting
von: Chakraborty, Arya, et al.
Veröffentlicht: (2024)
von: Chakraborty, Arya, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Mamba-Shedder: Post-Transformer Compression for Efficient Selective Structured State Space Models
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025) -
RTTC: Reward-Guided Collaborative Test-Time Compute
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025) -
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
von: Sarkar, Nilesh, et al.
Veröffentlicht: (2026) -
Towards Measuring Goal-Directedness in AI Systems
von: Xu, Dylan, et al.
Veröffentlicht: (2024) -
Regime Change Hypothesis: Foundations for Decoupled Dynamics in Neural Network Training
von: Pérez-Corral, Cristian, et al.
Veröffentlicht: (2026)