Q-GaLore: Quantized GaLore with INT4 Projection and Layer-Adaptive Low-Rank Gradients
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Zhenyu, Jaiswal, Ajay, Yin, Lu, Liu, Shiwei, Zhao, Jiawei, Tian, Yuandong, Wang, Zhangyang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
di: Zhao, Jiawei, et al.
Pubblicazione: (2024)
di: Zhao, Jiawei, et al.
Pubblicazione: (2024)
GaLore 2: Large-Scale LLM Pre-Training by Gradient Low-Rank Projection
di: Su, DiJia, et al.
Pubblicazione: (2025)
di: Su, DiJia, et al.
Pubblicazione: (2025)
Natural GaLore: Accelerating GaLore for memory-efficient LLM Training and Fine-tuning
di: Das, Arijit
Pubblicazione: (2024)
di: Das, Arijit
Pubblicazione: (2024)
GaLore$+$: Boosting Low-Rank Adaptation for LLMs with Cross-Head Projection
di: Liao, Xutao, et al.
Pubblicazione: (2024)
di: Liao, Xutao, et al.
Pubblicazione: (2024)
Subsampled Randomized Fourier GaLore for Adapting Foundation Models in Depth-Driven Liver Landmark Segmentation
di: Lin, Yun-Chen, et al.
Pubblicazione: (2025)
di: Lin, Yun-Chen, et al.
Pubblicazione: (2025)
From Low Rank Gradient Subspace Stabilization to Low-Rank Weights: Observations, Theories, and Applications
di: Jaiswal, Ajay, et al.
Pubblicazione: (2024)
di: Jaiswal, Ajay, et al.
Pubblicazione: (2024)
Junk DNA Hypothesis: Pruning Small Pre-Trained Weights Irreversibly and Monotonically Impairs "Difficult" Downstream Tasks in LLMs
di: Yin, Lu, et al.
Pubblicazione: (2023)
di: Yin, Lu, et al.
Pubblicazione: (2023)
Gray, Jane Mar. 4, 1851 [to Loring]
di: Gray, Jane Loring
Pubblicazione: (1851)
di: Gray, Jane Loring
Pubblicazione: (1851)
Gray, Jane Oct. 7, 1850 [to Loring]
di: Gray, Jane Loring
Pubblicazione: (1850)
di: Gray, Jane Loring
Pubblicazione: (1850)
Gray, Jane June 21, 1850 [to Loring]
di: Gray, Jane Loring
Pubblicazione: (1850)
di: Gray, Jane Loring
Pubblicazione: (1850)
The Loring--Schulz-Baldes Spectral Localizer Revisited
di: Berkolaiko, Gregory, et al.
Pubblicazione: (2025)
di: Berkolaiko, Gregory, et al.
Pubblicazione: (2025)
Family Lore, a Variant of Uncertain Significance, and CADASIL
di: Rhys Duarte, et al.
Pubblicazione: (2024)
di: Rhys Duarte, et al.
Pubblicazione: (2024)
The 80/20 Rule: Library Lore or Statistical Law?
di: Burrell, Quentin L.
Pubblicazione: (1985)
di: Burrell, Quentin L.
Pubblicazione: (1985)
R-Sparse: Rank-Aware Activation Sparsity for Efficient LLM Inference
di: Zhang, Zhenyu, et al.
Pubblicazione: (2025)
di: Zhang, Zhenyu, et al.
Pubblicazione: (2025)
MetaLore: Learning to Orchestrate Communication and Computation for Metaverse Synchronization
di: Ohri, Elif Ebru, et al.
Pubblicazione: (2025)
di: Ohri, Elif Ebru, et al.
Pubblicazione: (2025)
La autorregulación en el Mercado de VaLores
di: Mauricio Rosillo Rojas
Pubblicazione: (2008)
di: Mauricio Rosillo Rojas
Pubblicazione: (2008)
Folklorists of Educational Spaces: Material Lore in Classrooms With and Without Walls.
di: Merkel, Cecilia
Pubblicazione: (1999)
di: Merkel, Cecilia
Pubblicazione: (1999)
These Keys...Written Personal Narrative as Family Lore and Folk Object.
di: Sloan, Bernie
Pubblicazione: (1999)
di: Sloan, Bernie
Pubblicazione: (1999)
QuLore: An Adaptive Security Framework to Extend Quantum-Safe Communications to Real-World Networks
di: Sanz, Ane, et al.
Pubblicazione: (2025)
di: Sanz, Ane, et al.
Pubblicazione: (2025)
PixLore: A Dataset-driven Approach to Rich Image Captioning
di: Bonilla-Salvador, Diego, et al.
Pubblicazione: (2023)
di: Bonilla-Salvador, Diego, et al.
Pubblicazione: (2023)
Soy el número cuatro / Pittacus Lore ; Traducción Olga Martín Maldonado
di: Lore, Pittacus
Pubblicazione: (2010)
di: Lore, Pittacus
Pubblicazione: (2010)
El poder de seis / Pittacus Lore ; traducción de Gerardo Gambolini
di: Lore, Pittacus
Pubblicazione: (2012)
di: Lore, Pittacus
Pubblicazione: (2012)
Southern Appalachian Mountain Lore: Berea College's Weatherford-Hammond Mountain Collection
di: Perrin, Alfred H.
Pubblicazione: (1973)
di: Perrin, Alfred H.
Pubblicazione: (1973)
INT-FlashAttention: Enabling Flash Attention for INT8 Quantization
di: Chen, Shimao, et al.
Pubblicazione: (2024)
di: Chen, Shimao, et al.
Pubblicazione: (2024)
Toward INT4 Fixed-Point Training via Exploring Quantization Error for Gradients
di: Kim, Dohyung, et al.
Pubblicazione: (2024)
di: Kim, Dohyung, et al.
Pubblicazione: (2024)
In situ ATR‐FTIR spectroscopic study of metformin adsorption on gibbsite and Loring silt loam
di: Maheen Mehnaz, et al.
Pubblicazione: (2025)
di: Maheen Mehnaz, et al.
Pubblicazione: (2025)
FFN-SkipLLM: A Hidden Gem for Autoregressive Decoding with Adaptive Feed Forward Skipping
di: Jaiswal, Ajay, et al.
Pubblicazione: (2024)
di: Jaiswal, Ajay, et al.
Pubblicazione: (2024)
ParetoQ: Improving Scaling Laws in Extremely Low-bit LLM Quantization
di: Liu, Zechun, et al.
Pubblicazione: (2025)
di: Liu, Zechun, et al.
Pubblicazione: (2025)
Lore: Repurposing Git Commit Messages as a Structured Knowledge Protocol for AI Coding Agents
di: Stetsenko, Ivan
Pubblicazione: (2026)
di: Stetsenko, Ivan
Pubblicazione: (2026)
Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2026)
di: Kwon, JuneHyoung, et al.
Pubblicazione: (2026)
Unbiased Gradient Low-Rank Projection
di: Pan, Rui, et al.
Pubblicazione: (2025)
di: Pan, Rui, et al.
Pubblicazione: (2025)
La Gestión financiera / Jaime Loring Miró, Fuensanta Gal n Herrero, Teresa Montero Romero
di: Loring Miró, Jaime
di: Loring Miró, Jaime
Reseña de "Bartolomé Herrera y su tiempo" de Fernán Altuve-Febres Lores (compilador)
di: C.H. Sánchez Raygada
Pubblicazione: (2012)
di: C.H. Sánchez Raygada
Pubblicazione: (2012)
Condense, Don't Just Prune: Enhancing Efficiency and Performance in MoE Layer Pruning
di: Cao, Mingyu, et al.
Pubblicazione: (2024)
di: Cao, Mingyu, et al.
Pubblicazione: (2024)
LLaGA: Large Language and Graph Assistant
di: Chen, Runjin, et al.
Pubblicazione: (2024)
di: Chen, Runjin, et al.
Pubblicazione: (2024)
Full-Rank No More: Low-Rank Weight Training for Modern Speech Recognition Models
di: Fernandez-Lopez, Adriana, et al.
Pubblicazione: (2024)
di: Fernandez-Lopez, Adriana, et al.
Pubblicazione: (2024)
Comparison of Bayesian inference methods using the Loreli II database of hydro-radiative simulations of the 21-cm signal
di: Meriot, Romain, et al.
Pubblicazione: (2024)
di: Meriot, Romain, et al.
Pubblicazione: (2024)
Reseña de "Words of the Lagoon, Fishing and Marine Lore in the Palau District of Micronesia" de Johannes R. E.
di: Susana B. C. Devalle
Pubblicazione: (2004)
di: Susana B. C. Devalle
Pubblicazione: (2004)
FireQ: Fast INT4-FP8 Kernel and RoPE-aware Quantization for LLM Inference Acceleration
di: Baek, Daehyeon, et al.
Pubblicazione: (2025)
di: Baek, Daehyeon, et al.
Pubblicazione: (2025)
INT v.s. FP: A Comprehensive Study of Fine-Grained Low-bit Quantization Formats
di: Chen, Mengzhao, et al.
Pubblicazione: (2025)
di: Chen, Mengzhao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
di: Zhao, Jiawei, et al.
Pubblicazione: (2024) -
GaLore 2: Large-Scale LLM Pre-Training by Gradient Low-Rank Projection
di: Su, DiJia, et al.
Pubblicazione: (2025) -
Natural GaLore: Accelerating GaLore for memory-efficient LLM Training and Fine-tuning
di: Das, Arijit
Pubblicazione: (2024) -
GaLore$+$: Boosting Low-Rank Adaptation for LLMs with Cross-Head Projection
di: Liao, Xutao, et al.
Pubblicazione: (2024) -
Subsampled Randomized Fourier GaLore for Adapting Foundation Models in Depth-Driven Liver Landmark Segmentation
di: Lin, Yun-Chen, et al.
Pubblicazione: (2025)