Downsized and Compromised?: Assessing the Faithfulness of Model Compression
Fuente:
arXiv
Salvato in:
| Autori principali: | Kamal, Moumita, Talbert, Douglas A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Fusing Rewards and Preferences in Reinforcement Learning
di: Khorasani, Sadegh, et al.
Pubblicazione: (2025)
di: Khorasani, Sadegh, et al.
Pubblicazione: (2025)
Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales
di: Salfati, Samuel
Pubblicazione: (2026)
di: Salfati, Samuel
Pubblicazione: (2026)
Load and Renewable Energy Forecasting Using Deep Learning for Grid Stability
di: Sarkar, Kamal
Pubblicazione: (2025)
di: Sarkar, Kamal
Pubblicazione: (2025)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
di: Hedar, Abdel-Rahman, et al.
Pubblicazione: (2024)
di: Hedar, Abdel-Rahman, et al.
Pubblicazione: (2024)
Memory Bank Compression for Continual Adaptation of Large Language Models
di: Katraouras, Thomas, et al.
Pubblicazione: (2026)
di: Katraouras, Thomas, et al.
Pubblicazione: (2026)
FluidWorld: Reaction-Diffusion Dynamics as a Predictive Substrate for World Models
di: Polly, Fabien
Pubblicazione: (2026)
di: Polly, Fabien
Pubblicazione: (2026)
Scaling Laws for Neural Material Models
di: Trikha, Akshay, et al.
Pubblicazione: (2025)
di: Trikha, Akshay, et al.
Pubblicazione: (2025)
Seasoning Generative Models for a Generalization Aftertaste
di: Husain, Hisham, et al.
Pubblicazione: (2026)
di: Husain, Hisham, et al.
Pubblicazione: (2026)
Explanations Based on Item Response Theory (eXirt): A Model-Specific Method to Explain Tree-Ensemble Model in Trust Perspective
di: Ribeiro, José, et al.
Pubblicazione: (2022)
di: Ribeiro, José, et al.
Pubblicazione: (2022)
Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels
di: Rath, Plawan Kumar, et al.
Pubblicazione: (2026)
di: Rath, Plawan Kumar, et al.
Pubblicazione: (2026)
Measuring IIA Violations in Similarity Choices with Bayesian Models
di: Corrêa, Hugo Sales, et al.
Pubblicazione: (2025)
di: Corrêa, Hugo Sales, et al.
Pubblicazione: (2025)
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
di: Yousaf, Iqra
Pubblicazione: (2024)
di: Yousaf, Iqra
Pubblicazione: (2024)
Tempered Calculus for ML: Application to Hyperbolic Model Embedding
di: Nock, Richard, et al.
Pubblicazione: (2024)
di: Nock, Richard, et al.
Pubblicazione: (2024)
CASSANDRA: Programmatic and Probabilistic Learning and Inference for Stochastic World Modeling
di: Lymperopoulos, Panagiotis, et al.
Pubblicazione: (2026)
di: Lymperopoulos, Panagiotis, et al.
Pubblicazione: (2026)
TFMAdapter: Lightweight Instance-Level Adaptation of Foundation Models for Forecasting with Covariates
di: Dange, Afrin, et al.
Pubblicazione: (2025)
di: Dange, Afrin, et al.
Pubblicazione: (2025)
Securing Reliability: A Brief Overview on Enhancing In-Context Learning for Foundation Models
di: Huang, Yunpeng, et al.
Pubblicazione: (2024)
di: Huang, Yunpeng, et al.
Pubblicazione: (2024)
HGCN(O): A Self-Tuning GCN HyperModel Toolkit for Outcome Prediction in Event-Sequence Data
di: Wang, Fang, et al.
Pubblicazione: (2025)
di: Wang, Fang, et al.
Pubblicazione: (2025)
CoxSE: Exploring the Potential of Self-Explaining Neural Networks with Cox Proportional Hazards Model for Survival Analysis
di: Alabdallah, Abdallah, et al.
Pubblicazione: (2024)
di: Alabdallah, Abdallah, et al.
Pubblicazione: (2024)
Interestingness as an Inductive Heuristic for Future Compression Progress
di: Herrmann, Vincent, et al.
Pubblicazione: (2026)
di: Herrmann, Vincent, et al.
Pubblicazione: (2026)
LLM Vocabulary Compression for Low-Compute Environments
di: Vennam, Sreeram, et al.
Pubblicazione: (2024)
di: Vennam, Sreeram, et al.
Pubblicazione: (2024)
Playing Hex and Counter Wargames using Reinforcement Learning and Recurrent Neural Networks
di: Palma, Guilherme, et al.
Pubblicazione: (2025)
di: Palma, Guilherme, et al.
Pubblicazione: (2025)
FLrce: Resource-Efficient Federated Learning with Early-Stopping Strategy
di: Niu, Ziru, et al.
Pubblicazione: (2023)
di: Niu, Ziru, et al.
Pubblicazione: (2023)
2Mamba2Furious: Linear in Complexity, Competitive in Accuracy
di: Mongaras, Gabriel, et al.
Pubblicazione: (2026)
di: Mongaras, Gabriel, et al.
Pubblicazione: (2026)
Intervening to Learn and Compose Causally Disentangled Representations
di: Markham, Alex, et al.
Pubblicazione: (2025)
di: Markham, Alex, et al.
Pubblicazione: (2025)
Energy and Memory-Efficient Federated Learning With Ordered Layer Freezing
di: Niu, Ziru, et al.
Pubblicazione: (2025)
di: Niu, Ziru, et al.
Pubblicazione: (2025)
Learning To Help: Training Models to Assist Legacy Devices
di: Wu, Yu, et al.
Pubblicazione: (2024)
di: Wu, Yu, et al.
Pubblicazione: (2024)
The Bayesian Confidence (BACON) Estimator for Deep Neural Networks
di: Kee, Patrick D., et al.
Pubblicazione: (2024)
di: Kee, Patrick D., et al.
Pubblicazione: (2024)
Learned Relay Representations for Forward-Thinking Discrete Diffusion Models
di: Rozonoyer, Benjamin, et al.
Pubblicazione: (2026)
di: Rozonoyer, Benjamin, et al.
Pubblicazione: (2026)
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring
di: Heyman, Alex, et al.
Pubblicazione: (2025)
di: Heyman, Alex, et al.
Pubblicazione: (2025)
Kolmogorov Arnold Networks and Multi-Layer Perceptrons: A Paradigm Shift in Neural Modelling
di: Gaonkar, Aradhya, et al.
Pubblicazione: (2026)
di: Gaonkar, Aradhya, et al.
Pubblicazione: (2026)
Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models
di: Moghadasi, Mahdi Naser, et al.
Pubblicazione: (2026)
di: Moghadasi, Mahdi Naser, et al.
Pubblicazione: (2026)
I-GLIDE: Input Groups for Latent Health Indicators in Degradation Estimation
di: Thil, Lucas, et al.
Pubblicazione: (2025)
di: Thil, Lucas, et al.
Pubblicazione: (2025)
RobustModelMaker: Coupling Bootstrap Stability Selection with Leakage-Safe Nested Cross-Validation for Scientific Machine Learning
di: Barnard, Amanda S
Pubblicazione: (2026)
di: Barnard, Amanda S
Pubblicazione: (2026)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025)
di: Fadli, Samih
Pubblicazione: (2025)
FSC-Net: Fast-Slow Consolidation Networks for Continual Learning
di: Gorrim, Mohamed El
Pubblicazione: (2025)
di: Gorrim, Mohamed El
Pubblicazione: (2025)
Action-Dependent Optimality-Preserving Reward Shaping
di: Forbes, Grant C., et al.
Pubblicazione: (2025)
di: Forbes, Grant C., et al.
Pubblicazione: (2025)
Learning Stochastic Nonlinear Dynamics with Embedded Latent Transfer Operators
di: Ke, Naichang, et al.
Pubblicazione: (2025)
di: Ke, Naichang, et al.
Pubblicazione: (2025)
Multiple Token Divergence: Measuring and Steering In-Context Computation Density
di: Herrmann, Vincent, et al.
Pubblicazione: (2025)
di: Herrmann, Vincent, et al.
Pubblicazione: (2025)
Strengthening the Internal Adversarial Robustness in Lifted Neural Networks
di: Zach, Christopher
Pubblicazione: (2025)
di: Zach, Christopher
Pubblicazione: (2025)
Pruning Spurious Subgraphs for Graph Out-of-Distribution Generalization
di: Yao, Tianjun, et al.
Pubblicazione: (2025)
di: Yao, Tianjun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Fusing Rewards and Preferences in Reinforcement Learning
di: Khorasani, Sadegh, et al.
Pubblicazione: (2025) -
Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales
di: Salfati, Samuel
Pubblicazione: (2026) -
Load and Renewable Energy Forecasting Using Deep Learning for Grid Stability
di: Sarkar, Kamal
Pubblicazione: (2025) -
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
di: Hedar, Abdel-Rahman, et al.
Pubblicazione: (2024) -
Memory Bank Compression for Continual Adaptation of Large Language Models
di: Katraouras, Thomas, et al.
Pubblicazione: (2026)