NoWag: A Unified Framework for Shape Preserving Compression of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Lawrence, Chakrabarti, Inesh, Li, Yixiao, Wang, Mengdi, Zhao, Tuo, Yang, Lin F. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deep Reinforcement Learning from Hierarchical Preference Design
von: Bukharin, Alexander, et al.
Veröffentlicht: (2023)
von: Bukharin, Alexander, et al.
Veröffentlicht: (2023)
Quantifying and Understanding Uncertainty in Large Reasoning Models
von: Li, Yangyi, et al.
Veröffentlicht: (2026)
von: Li, Yangyi, et al.
Veröffentlicht: (2026)
Collaborative Compression for Large-Scale MoE Deployment on Edge
von: Chen, Yixiao, et al.
Veröffentlicht: (2025)
von: Chen, Yixiao, et al.
Veröffentlicht: (2025)
REMA: A Unified Reasoning Manifold Framework for Interpreting Large Language Model
von: Li, Bo, et al.
Veröffentlicht: (2025)
von: Li, Bo, et al.
Veröffentlicht: (2025)
EEGAgent: A Unified Framework for Automated EEG Analysis Using Large Language Models
von: Zhao, Sha, et al.
Veröffentlicht: (2025)
von: Zhao, Sha, et al.
Veröffentlicht: (2025)
IDEA Prune: An Integrated Enlarge-and-Prune Pipeline in Generative Language Model Pretraining
von: Li, Yixiao, et al.
Veröffentlicht: (2025)
von: Li, Yixiao, et al.
Veröffentlicht: (2025)
Extracting Heuristics from Large Language Models for Reward Shaping in Reinforcement Learning
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2024)
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2024)
Vanishing Contributions: A Unified Framework for Smooth and Iterative Model Compression
von: Nikiforos, Lorenzo, et al.
Veröffentlicht: (2025)
von: Nikiforos, Lorenzo, et al.
Veröffentlicht: (2025)
Adaptive Preference Scaling for Reinforcement Learning with Human Feedback
von: Hong, Ilgee, et al.
Veröffentlicht: (2024)
von: Hong, Ilgee, et al.
Veröffentlicht: (2024)
BAMDP Shaping: a Unified Framework for Intrinsic Motivation and Reward Shaping
von: Lidayan, Aly, et al.
Veröffentlicht: (2024)
von: Lidayan, Aly, et al.
Veröffentlicht: (2024)
Activation Sparsity Opportunities for Compressing General Large Language Models
von: Dhar, Nobel, et al.
Veröffentlicht: (2024)
von: Dhar, Nobel, et al.
Veröffentlicht: (2024)
Preserving Diversity in Supervised Fine-Tuning of Large Language Models
von: Li, Ziniu, et al.
Veröffentlicht: (2024)
von: Li, Ziniu, et al.
Veröffentlicht: (2024)
TokenSqueeze: Performance-Preserving Compression for Reasoning LLMs
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2025)
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2025)
Neural Diversity Regularizes Hallucinations in Language Models
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2025)
von: Chakrabarti, Kushal, et al.
Veröffentlicht: (2025)
A Systematic Study of Compression Ordering for Large Language Models
von: Chhawri, Shivansh, et al.
Veröffentlicht: (2025)
von: Chhawri, Shivansh, et al.
Veröffentlicht: (2025)
Large Language Model Compression with Global Rank and Sparsity Optimization
von: Zhou, Changhai, et al.
Veröffentlicht: (2025)
von: Zhou, Changhai, et al.
Veröffentlicht: (2025)
FragmentGPT: A Unified GPT Model for Fragment Growing, Linking, and Merging in Molecular Design
von: Liu, Xuefeng, et al.
Veröffentlicht: (2025)
von: Liu, Xuefeng, et al.
Veröffentlicht: (2025)
Distinguishable Deletion: Unifying Knowledge Erasure and Refusal for Large Language Model Unlearning
von: Yang, Puning, et al.
Veröffentlicht: (2026)
von: Yang, Puning, et al.
Veröffentlicht: (2026)
AmoebaLLM: Constructing Any-Shape Large Language Models for Efficient and Instant Deployment
von: Fu, Yonggan, et al.
Veröffentlicht: (2024)
von: Fu, Yonggan, et al.
Veröffentlicht: (2024)
On the Compressibility of Quantized Large Language Models
von: Mao, Yu, et al.
Veröffentlicht: (2024)
von: Mao, Yu, et al.
Veröffentlicht: (2024)
UniSD: Towards a Unified Self-Distillation Framework for Large Language Models
von: Jin, Yiqiao, et al.
Veröffentlicht: (2026)
von: Jin, Yiqiao, et al.
Veröffentlicht: (2026)
An Interpretable and Scalable Framework for Evaluating Large Language Models
von: Qu, Xinhao, et al.
Veröffentlicht: (2026)
von: Qu, Xinhao, et al.
Veröffentlicht: (2026)
CoSA: Compressed Sensing-Based Adaptation of Large Language Models
von: Wei, Songtao, et al.
Veröffentlicht: (2026)
von: Wei, Songtao, et al.
Veröffentlicht: (2026)
GWT: Scalable Optimizer State Compression for Large Language Model Training
von: Wen, Ziqing, et al.
Veröffentlicht: (2025)
von: Wen, Ziqing, et al.
Veröffentlicht: (2025)
PalmBench: A Comprehensive Benchmark of Compressed Large Language Models on Mobile Platforms
von: Li, Yilong, et al.
Veröffentlicht: (2024)
von: Li, Yilong, et al.
Veröffentlicht: (2024)
The Role of Emotional Stimuli and Intensity in Shaping Large Language Model Behavior
von: Patel, Ameen, et al.
Veröffentlicht: (2026)
von: Patel, Ameen, et al.
Veröffentlicht: (2026)
AutoDetect: Towards a Unified Framework for Automated Weakness Detection in Large Language Models
von: Cheng, Jiale, et al.
Veröffentlicht: (2024)
von: Cheng, Jiale, et al.
Veröffentlicht: (2024)
Model Compression and Efficient Inference for Large Language Models: A Survey
von: Wang, Wenxiao, et al.
Veröffentlicht: (2024)
von: Wang, Wenxiao, et al.
Veröffentlicht: (2024)
LOGIN: A Large Language Model Consulted Graph Neural Network Training Framework
von: Qiao, Yiran, et al.
Veröffentlicht: (2024)
von: Qiao, Yiran, et al.
Veröffentlicht: (2024)
Rethinking LoRA for Privacy-Preserving Federated Learning in Large Models
von: Liu, Jin, et al.
Veröffentlicht: (2026)
von: Liu, Jin, et al.
Veröffentlicht: (2026)
Pruning Large Language Models by Identifying and Preserving Functional Networks
von: Liu, Yiheng, et al.
Veröffentlicht: (2025)
von: Liu, Yiheng, et al.
Veröffentlicht: (2025)
Rethinking Key-Value Cache Compression Techniques for Large Language Model Serving
von: Gao, Wei, et al.
Veröffentlicht: (2025)
von: Gao, Wei, et al.
Veröffentlicht: (2025)
Lillama: Large Language Models Compression via Low-Rank Feature Distillation
von: Sy, Yaya, et al.
Veröffentlicht: (2024)
von: Sy, Yaya, et al.
Veröffentlicht: (2024)
LLMC: Benchmarking Large Language Model Quantization with a Versatile Compression Toolkit
von: Gong, Ruihao, et al.
Veröffentlicht: (2024)
von: Gong, Ruihao, et al.
Veröffentlicht: (2024)
Do Physics Foundation Models Learn Generalizable Physics? A Bias-Aware Benchmark Across Physical Regimes and Distribution Shifts
von: Chu, Mengdi, et al.
Veröffentlicht: (2026)
von: Chu, Mengdi, et al.
Veröffentlicht: (2026)
PromptBench: A Unified Library for Evaluation of Large Language Models
von: Zhu, Kaijie, et al.
Veröffentlicht: (2023)
von: Zhu, Kaijie, et al.
Veröffentlicht: (2023)
MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
von: Liu, Akide, et al.
Veröffentlicht: (2024)
von: Liu, Akide, et al.
Veröffentlicht: (2024)
A Practical Tensor-Network Compression Pipeline for Production-Scale Large Language Models
von: Kozyrev, Sergii, et al.
Veröffentlicht: (2026)
von: Kozyrev, Sergii, et al.
Veröffentlicht: (2026)
BI-DCGAN: A Theoretically Grounded Bayesian Framework for Efficient and Diverse GANs
von: Valizadeh, Mahsa, et al.
Veröffentlicht: (2025)
von: Valizadeh, Mahsa, et al.
Veröffentlicht: (2025)
Aioli: A Unified Optimization Framework for Language Model Data Mixing
von: Chen, Mayee F., et al.
Veröffentlicht: (2024)
von: Chen, Mayee F., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Deep Reinforcement Learning from Hierarchical Preference Design
von: Bukharin, Alexander, et al.
Veröffentlicht: (2023) -
Quantifying and Understanding Uncertainty in Large Reasoning Models
von: Li, Yangyi, et al.
Veröffentlicht: (2026) -
Collaborative Compression for Large-Scale MoE Deployment on Edge
von: Chen, Yixiao, et al.
Veröffentlicht: (2025) -
REMA: A Unified Reasoning Manifold Framework for Interpreting Large Language Model
von: Li, Bo, et al.
Veröffentlicht: (2025) -
EEGAgent: A Unified Framework for Automated EEG Analysis Using Large Language Models
von: Zhao, Sha, et al.
Veröffentlicht: (2025)