Guardado en:
| Autores principales: | Xv, Lin, Gao, Xian, Li, Ting, Fu, Yuzhuo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2510.19385 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ARA: Adaptive Rank Allocation for Efficient Large Language Model SVD Compression
por: Xv, Lin, et al.
Publicado: (2025)
por: Xv, Lin, et al.
Publicado: (2025)
DipSVD: Dual-importance Protected SVD for Efficient LLM Compression
por: Ding, Xuan, et al.
Publicado: (2025)
por: Ding, Xuan, et al.
Publicado: (2025)
Dobi-SVD: Differentiable SVD for LLM Compression and Some New Perspectives
por: Wang, Qinsi, et al.
Publicado: (2025)
por: Wang, Qinsi, et al.
Publicado: (2025)
IO-SVD: Input-Output Whitened SVD for Adaptive-Rank LLM Compression
por: Abbasi, Ali, et al.
Publicado: (2026)
por: Abbasi, Ali, et al.
Publicado: (2026)
AA-SVD : Anchored and Adaptive SVD for Large Language Model Compression
por: Sinha, Atul Kumar, et al.
Publicado: (2026)
por: Sinha, Atul Kumar, et al.
Publicado: (2026)
Zero Sum SVD: Balancing Loss Sensitivity for Low Rank LLM Compression
por: Abbasi, Ali, et al.
Publicado: (2026)
por: Abbasi, Ali, et al.
Publicado: (2026)
Different Prompts, Different Ranks: Prompt-aware Dynamic Rank Selection for SVD-based LLM Compression
por: Zhu, Hengyi, et al.
Publicado: (2026)
por: Zhu, Hengyi, et al.
Publicado: (2026)
SVD-LLM: Truncation-aware Singular Value Decomposition for Large Language Model Compression
por: Wang, Xin, et al.
Publicado: (2024)
por: Wang, Xin, et al.
Publicado: (2024)
Low-Rank Prehab: Preparing Neural Networks for SVD Compression
por: Qin, Haoran, et al.
Publicado: (2025)
por: Qin, Haoran, et al.
Publicado: (2025)
KQ-SVD: Compressing the KV Cache with Provable Guarantees on Attention Fidelity
por: Lesens, Damien, et al.
Publicado: (2025)
por: Lesens, Damien, et al.
Publicado: (2025)
SVD Contextual Sparsity Predictors for Fast LLM Inference
por: Serbin, Georgii, et al.
Publicado: (2026)
por: Serbin, Georgii, et al.
Publicado: (2026)
SVD-NO: Learning PDE Solution Operators with SVD Integral Kernels
por: Koren, Noam, et al.
Publicado: (2025)
por: Koren, Noam, et al.
Publicado: (2025)
SOLAR: SVD-Optimized Lifelong Attention for Recommendation
por: Zhang, Chenghao, et al.
Publicado: (2026)
por: Zhang, Chenghao, et al.
Publicado: (2026)
Concatenated Matrix SVD: Compression Bounds, Incremental Approximation, and Error-Constrained Clustering
por: Shamrai, Maksym
Publicado: (2026)
por: Shamrai, Maksym
Publicado: (2026)
List Sample Compression and Uniform Convergence
por: Hanneke, Steve, et al.
Publicado: (2024)
por: Hanneke, Steve, et al.
Publicado: (2024)
Enhancing Delta Compression in LLMs via SVD-based Quantization Error Minimization
por: Xiong, Boya, et al.
Publicado: (2025)
por: Xiong, Boya, et al.
Publicado: (2025)
Bilevel Optimization with Lower-Level Uniform Convexity: Theory and Algorithm
por: Wu, Yuman, et al.
Publicado: (2026)
por: Wu, Yuman, et al.
Publicado: (2026)
SAFE-SVD: Sensitivity-Aware Fidelity-Enforcing SVD for Physics Foundation Models
por: Hong, Chengjie, et al.
Publicado: (2026)
por: Hong, Chengjie, et al.
Publicado: (2026)
ExLLM: Experience-Enhanced LLM Optimization for Molecular Design and Beyond
por: Ran, Nian, et al.
Publicado: (2025)
por: Ran, Nian, et al.
Publicado: (2025)
PCA, SVD, and Centering of Data
por: Kim, Donggun, et al.
Publicado: (2023)
por: Kim, Donggun, et al.
Publicado: (2023)
Beyond SGD, Without SVD: Proximal Subspace Iteration LoRA with Diagonal Fractional K-FAC
por: Almansoori, Abdulla Jasem, et al.
Publicado: (2026)
por: Almansoori, Abdulla Jasem, et al.
Publicado: (2026)
SliceGPT: Compress Large Language Models by Deleting Rows and Columns
por: Ashkboos, Saleh, et al.
Publicado: (2024)
por: Ashkboos, Saleh, et al.
Publicado: (2024)
OnlineMate: An LLM-Based Multi-Agent Companion System for Cognitive Support in Online Learning
por: Gao, Xian, et al.
Publicado: (2025)
por: Gao, Xian, et al.
Publicado: (2025)
Uniform Convergence Beyond Glivenko-Cantelli
por: Devale, Tanmay, et al.
Publicado: (2025)
por: Devale, Tanmay, et al.
Publicado: (2025)
FlashSVD: Memory-Efficient Inference with Streaming for Low-Rank Models
por: Shao, Zishan, et al.
Publicado: (2025)
por: Shao, Zishan, et al.
Publicado: (2025)
Generalized Fisher-Weighted SVD: Scalable Kronecker-Factored Fisher Approximation for Compressing Large Language Models
por: Chekalina, Viktoriia, et al.
Publicado: (2025)
por: Chekalina, Viktoriia, et al.
Publicado: (2025)
Beyond Uniform Credit: Causal Credit Assignment for Policy Optimization
por: Khandoga, Mykola, et al.
Publicado: (2026)
por: Khandoga, Mykola, et al.
Publicado: (2026)
PV-Tuning: Beyond Straight-Through Estimation for Extreme LLM Compression
por: Malinovskii, Vladimir, et al.
Publicado: (2024)
por: Malinovskii, Vladimir, et al.
Publicado: (2024)
$k$-SVD with Gradient Descent
por: Jedra, Yassir, et al.
Publicado: (2025)
por: Jedra, Yassir, et al.
Publicado: (2025)
Uniformly Stable Algorithms for Adversarial Training and Beyond
por: Xiao, Jiancong, et al.
Publicado: (2024)
por: Xiao, Jiancong, et al.
Publicado: (2024)
SVDformer: Direction-Aware Spectral Graph Embedding Learning via SVD and Transformer
por: Fang, Jiayu, et al.
Publicado: (2025)
por: Fang, Jiayu, et al.
Publicado: (2025)
Decentralized Multi-Level Compositional Optimization Algorithms with Level-Independent Convergence Rate
por: Gao, Hongchang
Publicado: (2023)
por: Gao, Hongchang
Publicado: (2023)
Value-Compressed Sparse Column (VCSC): Sparse Matrix Storage for Redundant Data
por: Ruiter, Skyler, et al.
Publicado: (2023)
por: Ruiter, Skyler, et al.
Publicado: (2023)
LASER: Low-Rank Activation SVD for Efficient Recursion
por: Çakar, Ege, et al.
Publicado: (2026)
por: Çakar, Ege, et al.
Publicado: (2026)
Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents
por: Wang, Jiawei, et al.
Publicado: (2025)
por: Wang, Jiawei, et al.
Publicado: (2025)
Beyond Johnson-Lindenstrauss: Uniform Bounds for Sketched Bilinear Forms
por: Deb, Rohan, et al.
Publicado: (2025)
por: Deb, Rohan, et al.
Publicado: (2025)
Improving LLM Safety Alignment with Dual-Objective Optimization
por: Zhao, Xuandong, et al.
Publicado: (2025)
por: Zhao, Xuandong, et al.
Publicado: (2025)
tenSVD algorithm for compression
por: Gallo, Michele
Publicado: (2025)
por: Gallo, Michele
Publicado: (2025)
Learning More with Less: A Dynamic Dual-Level Down-Sampling Framework for Efficient Policy Optimization
por: Wang, Chao, et al.
Publicado: (2025)
por: Wang, Chao, et al.
Publicado: (2025)
Machine Learning-Enhanced Ant Colony Optimization for Column Generation
por: Xu, Hongjie, et al.
Publicado: (2024)
por: Xu, Hongjie, et al.
Publicado: (2024)
Ejemplares similares
-
ARA: Adaptive Rank Allocation for Efficient Large Language Model SVD Compression
por: Xv, Lin, et al.
Publicado: (2025) -
DipSVD: Dual-importance Protected SVD for Efficient LLM Compression
por: Ding, Xuan, et al.
Publicado: (2025) -
Dobi-SVD: Differentiable SVD for LLM Compression and Some New Perspectives
por: Wang, Qinsi, et al.
Publicado: (2025) -
IO-SVD: Input-Output Whitened SVD for Adaptive-Rank LLM Compression
por: Abbasi, Ali, et al.
Publicado: (2026) -
AA-SVD : Anchored and Adaptive SVD for Large Language Model Compression
por: Sinha, Atul Kumar, et al.
Publicado: (2026)