Gespeichert in:
| Hauptverfasser: | Jovanović, Andrej, Iacob, Alex, Safaryan, Mher, Modoranu, Ionut-Vlad, Sani, Lorenzo, Shen, William F., Qiu, Xinchi, Alistarh, Dan, Lane, Nicholas D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.04396 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MatryoshkaLoRA: Learning Accurate Hierarchical Low-Rank Representations for LLM Fine-Tuning
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2026)
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2026)
LDAdam: Adaptive Optimization from Low-Dimensional Gradient Statistics
von: Robert, Thomas, et al.
Veröffentlicht: (2024)
von: Robert, Thomas, et al.
Veröffentlicht: (2024)
FFT-based Dynamic Subspace Selection for Low-Rank Adaptive Optimization of Large Language Models
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2025)
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2025)
The Iterative Optimal Brain Surgeon: Faster Sparse Recovery by Leveraging Second-Order Information
von: Wu, Diyuan, et al.
Veröffentlicht: (2024)
von: Wu, Diyuan, et al.
Veröffentlicht: (2024)
DASH: Faster Shampoo via Batched Block Preconditioning and Efficient Inverse-Root Solvers
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2026)
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2026)
MT-DAO: Multi-Timescale Distributed Adaptive Optimizers with Local Updates
von: Iacob, Alex, et al.
Veröffentlicht: (2025)
von: Iacob, Alex, et al.
Veröffentlicht: (2025)
Unified Scaling Laws for Compressed Representations
von: Panferov, Andrei, et al.
Veröffentlicht: (2025)
von: Panferov, Andrei, et al.
Veröffentlicht: (2025)
DES-LOC: Desynced Low Communication Adaptive Optimizers for Training Foundation Models
von: Iacob, Alex, et al.
Veröffentlicht: (2025)
von: Iacob, Alex, et al.
Veröffentlicht: (2025)
MicroAdam: Accurate Adaptive Optimization with Low Space Overhead and Provable Convergence
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2024)
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2024)
Error Feedback Can Accurately Compress Preconditioners
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2023)
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2023)
Towards Robust Scaling Laws for Optimizers
von: Volkova, Alexandra, et al.
Veröffentlicht: (2026)
von: Volkova, Alexandra, et al.
Veröffentlicht: (2026)
CAGE: Curvature-Aware Gradient Estimation For Accurate Quantization-Aware Training
von: Tabesh, Soroush, et al.
Veröffentlicht: (2025)
von: Tabesh, Soroush, et al.
Veröffentlicht: (2025)
DEPT: Decoupled Embeddings for Pre-training Language Models
von: Iacob, Alex, et al.
Veröffentlicht: (2024)
von: Iacob, Alex, et al.
Veröffentlicht: (2024)
LLM Unlearning via Neural Activation Redirection
von: Shen, William F., et al.
Veröffentlicht: (2025)
von: Shen, William F., et al.
Veröffentlicht: (2025)
Position: It's Time to Act on the Risk of Efficient Personalized Text Generation
von: Iofinova, Eugenia, et al.
Veröffentlicht: (2025)
von: Iofinova, Eugenia, et al.
Veröffentlicht: (2025)
GradSkip: Communication-Accelerated Local Gradient Methods with Better Computational Complexity
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2022)
von: Maranjyan, Artavazd, et al.
Veröffentlicht: (2022)
Sheaf HyperNetworks for Personalized Federated Learning
von: Nguyen, Bao, et al.
Veröffentlicht: (2024)
von: Nguyen, Bao, et al.
Veröffentlicht: (2024)
Pollen: High-throughput Federated Learning Simulation via Resource-Aware Client Placement
von: Sani, Lorenzo, et al.
Veröffentlicht: (2023)
von: Sani, Lorenzo, et al.
Veröffentlicht: (2023)
FedAnchor: Enhancing Federated Semi-Supervised Learning with Label Contrastive Loss for Unlabeled Clients
von: Qiu, Xinchi, et al.
Veröffentlicht: (2024)
von: Qiu, Xinchi, et al.
Veröffentlicht: (2024)
On Biased Compression for Distributed Learning
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2020)
von: Beznosikov, Aleksandr, et al.
Veröffentlicht: (2020)
SparsyFed: Sparse Adaptive Federated Training
von: Guastella, Adriano, et al.
Veröffentlicht: (2025)
von: Guastella, Adriano, et al.
Veröffentlicht: (2025)
Optimizers Qualitatively Alter Solutions And We Should Leverage This
von: Pascanu, Razvan, et al.
Veröffentlicht: (2025)
von: Pascanu, Razvan, et al.
Veröffentlicht: (2025)
Photon: Federated LLM Pre-Training
von: Sani, Lorenzo, et al.
Veröffentlicht: (2024)
von: Sani, Lorenzo, et al.
Veröffentlicht: (2024)
AbbIE: Autoregressive Block-Based Iterative Encoder for Efficient Sequence Modeling
von: Aleksandrov, Preslav, et al.
Veröffentlicht: (2025)
von: Aleksandrov, Preslav, et al.
Veröffentlicht: (2025)
The Future of Large Language Model Pre-training is Federated
von: Sani, Lorenzo, et al.
Veröffentlicht: (2024)
von: Sani, Lorenzo, et al.
Veröffentlicht: (2024)
Worldwide Federated Training of Language Models
von: Iacob, Alex, et al.
Veröffentlicht: (2024)
von: Iacob, Alex, et al.
Veröffentlicht: (2024)
Response-Conditioned Parallel-to-Sequential Orchestration for Multi-Agent Systems
von: Tastan, Nurbek, et al.
Veröffentlicht: (2026)
von: Tastan, Nurbek, et al.
Veröffentlicht: (2026)
SEAT: Sparse Entity-Aware Tuning for Knowledge Adaptation while Preserving Epistemic Abstention
von: Shen, William F., et al.
Veröffentlicht: (2025)
von: Shen, William F., et al.
Veröffentlicht: (2025)
Gradient-less Federated Gradient Boosting Trees with Learnable Learning Rates
von: Ma, Chenyang, et al.
Veröffentlicht: (2023)
von: Ma, Chenyang, et al.
Veröffentlicht: (2023)
Breaking Physical and Linguistic Borders: Multilingual Federated Prompt Tuning for Low-Resource Languages
von: Zhao, Wanru, et al.
Veröffentlicht: (2025)
von: Zhao, Wanru, et al.
Veröffentlicht: (2025)
Panza: Design and Analysis of a Fully-Local Personalized Text Writing Assistant
von: Nicolicioiu, Armand, et al.
Veröffentlicht: (2024)
von: Nicolicioiu, Armand, et al.
Veröffentlicht: (2024)
Hymnographic Indicators of the Armenian Renaissance
von: Mher Navoyan
Veröffentlicht: (2026)
von: Mher Navoyan
Veröffentlicht: (2026)
Speculative Decoding Speed-of-Light: Optimal Lower Bounds via Branching Random Walks
von: Pankratov, Sergey, et al.
Veröffentlicht: (2025)
von: Pankratov, Sergey, et al.
Veröffentlicht: (2025)
Simple Opinion Dynamics for No-Regret Learning
von: Lazarsfeld, John, et al.
Veröffentlicht: (2023)
von: Lazarsfeld, John, et al.
Veröffentlicht: (2023)
Behemoth: Benchmarking Unlearning in LLMs Using Fully Synthetic Data
von: Iofinova, Eugenia, et al.
Veröffentlicht: (2026)
von: Iofinova, Eugenia, et al.
Veröffentlicht: (2026)
LLMQ: Efficient Lower-Precision Pretraining for Consumer GPUs
von: Schultheis, Erik, et al.
Veröffentlicht: (2025)
von: Schultheis, Erik, et al.
Veröffentlicht: (2025)
Model Compression with Exact Budget Constraints via Riemannian Manifolds
von: Helcig, Michael, et al.
Veröffentlicht: (2026)
von: Helcig, Michael, et al.
Veröffentlicht: (2026)
Communication-Efficient Federated Learning With Data and Client Heterogeneity
von: Zakerinia, Hossein, et al.
Veröffentlicht: (2022)
von: Zakerinia, Hossein, et al.
Veröffentlicht: (2022)
Optimizing Classification of Infrequent Labels by Reducing Variability in Label Distribution
von: Agarwal, Ashutosh
Veröffentlicht: (2025)
von: Agarwal, Ashutosh
Veröffentlicht: (2025)
Deep Unrolling of Sparsity-Induced RDO for 3D Point Cloud Attribute Coding
von: Do, Tam Thuc, et al.
Veröffentlicht: (2025)
von: Do, Tam Thuc, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MatryoshkaLoRA: Learning Accurate Hierarchical Low-Rank Representations for LLM Fine-Tuning
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2026) -
LDAdam: Adaptive Optimization from Low-Dimensional Gradient Statistics
von: Robert, Thomas, et al.
Veröffentlicht: (2024) -
FFT-based Dynamic Subspace Selection for Low-Rank Adaptive Optimization of Large Language Models
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2025) -
The Iterative Optimal Brain Surgeon: Faster Sparse Recovery by Leveraging Second-Order Information
von: Wu, Diyuan, et al.
Veröffentlicht: (2024) -
DASH: Faster Shampoo via Batched Block Preconditioning and Efficient Inverse-Root Solvers
von: Modoranu, Ionut-Vlad, et al.
Veröffentlicht: (2026)