GIFT-SW: Gaussian noise Injected Fine-Tuning of Salient Weights for LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Zhelnin, Maxim, Moskvoretskii, Viktor, Shvetsov, Egor, Venediktov, Egor, Krylova, Mariya, Zuev, Aleksandr, Burnaev, Evgeny |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EBES: Easy Benchmarking for Event Sequences
by: Osin, Dmitry, et al.
Published: (2024)
by: Osin, Dmitry, et al.
Published: (2024)
MLEM: Generative and Contrastive Learning as Distinct Modalities for Event Sequences
by: Moskvoretskii, Viktor, et al.
Published: (2024)
by: Moskvoretskii, Viktor, et al.
Published: (2024)
Investigating the Impact of Quantization Methods on the Safety and Reliability of Large Language Models
by: Kharinaev, Artyom, et al.
Published: (2025)
by: Kharinaev, Artyom, et al.
Published: (2025)
From 2:4 to 8:16 sparsity patterns in LLMs for Outliers and Weights with Variance Correction
by: Maximov, Egor, et al.
Published: (2025)
by: Maximov, Egor, et al.
Published: (2025)
Motivating Next-Gen Accelerators with Flexible (N:M) Activation Sparsity via Benchmarking Lightweight Post-Training Sparsification Approaches
by: Alanova, Shirin, et al.
Published: (2025)
by: Alanova, Shirin, et al.
Published: (2025)
How to model Human Actions distribution with Event Sequence Data
by: Surkov, Egor, et al.
Published: (2025)
by: Surkov, Egor, et al.
Published: (2025)
SeqNAS: Neural Architecture Search for Event Sequence Classification
by: Udovichenko, Igor, et al.
Published: (2024)
by: Udovichenko, Igor, et al.
Published: (2024)
From Internal Representations to Text Quality: A Geometric Approach to LLM Evaluation
by: Yusupov, Viacheslav, et al.
Published: (2025)
by: Yusupov, Viacheslav, et al.
Published: (2025)
GIFT: Global stabilisation via Intrinsic Fine Tuning
by: Young, Rory, et al.
Published: (2026)
by: Young, Rory, et al.
Published: (2026)
Step Rejection Fine-Tuning: A Practical Distillation Recipe
by: Slinko, Igor, et al.
Published: (2026)
by: Slinko, Igor, et al.
Published: (2026)
QuantNAS for super resolution: searching for efficient quantization-friendly architectures against quantization noise
by: Shvetsov, Egor, et al.
Published: (2022)
by: Shvetsov, Egor, et al.
Published: (2022)
ELMUR: External Layer Memory with Update/Rewrite for Long-Horizon RL Problems
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
Stack Trace Deduplication: Faster, More Accurately, and in More Realistic Scenarios
by: Shibaev, Egor, et al.
Published: (2024)
by: Shibaev, Egor, et al.
Published: (2024)
The Density of Cross-Persistence Diagrams and Its Applications
by: Mironenko, Alexander, et al.
Published: (2026)
by: Mironenko, Alexander, et al.
Published: (2026)
BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
Memory Retention Is Not Enough to Master Memory Tasks in Reinforcement Learning
by: Shchendrigin, Oleg, et al.
Published: (2026)
by: Shchendrigin, Oleg, et al.
Published: (2026)
Recurrent Action Transformer with Memory
by: Cherepanov, Egor, et al.
Published: (2023)
by: Cherepanov, Egor, et al.
Published: (2023)
Re:Frame -- Retrieving Experience From Associative Memory
by: Zelezetsky, Daniil, et al.
Published: (2025)
by: Zelezetsky, Daniil, et al.
Published: (2025)
Enhancing PIBT via Multi-Action Operations
by: Yukhnevich, Egor, et al.
Published: (2025)
by: Yukhnevich, Egor, et al.
Published: (2025)
Memory, Benchmark & Robots: A Benchmark for Solving Complex Tasks with Reinforcement Learning
by: Cherepanov, Egor, et al.
Published: (2025)
by: Cherepanov, Egor, et al.
Published: (2025)
RTD-Lite: Scalable Topological Analysis for Comparing Weighted Graphs in Learning Tasks
by: Tulchinskii, Eduard, et al.
Published: (2025)
by: Tulchinskii, Eduard, et al.
Published: (2025)
Protecting Private Code in IDE Autocomplete using Differential Privacy
by: Grigorenko, Evgeny, et al.
Published: (2026)
by: Grigorenko, Evgeny, et al.
Published: (2026)
Beyond Early-Token Bias: Model-Specific and Language-Specific Position Effects in Multilingual LLMs
by: Menschikov, Mikhail, et al.
Published: (2025)
by: Menschikov, Mikhail, et al.
Published: (2025)
Drawing Pandas: A Benchmark for LLMs in Generating Plotting Code
by: Galimzyanov, Timur, et al.
Published: (2024)
by: Galimzyanov, Timur, et al.
Published: (2024)
Team-Based Self-Play With Dual Adaptive Weighting for Fine-Tuning LLMs
by: Li, Wu, et al.
Published: (2026)
by: Li, Wu, et al.
Published: (2026)
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
by: Cherepanov, Egor, et al.
Published: (2024)
by: Cherepanov, Egor, et al.
Published: (2024)
CAT: Causal Attention Tuning For Injecting Fine-grained Causal Knowledge into Large Language Models
by: Han, Kairong, et al.
Published: (2025)
by: Han, Kairong, et al.
Published: (2025)
KAGE-Bench: Fast Known-Axis Visual Generalization Evaluation for Reinforcement Learning
by: Cherepanov, Egor, et al.
Published: (2026)
by: Cherepanov, Egor, et al.
Published: (2026)
SCT: A Simple Baseline for Parameter-Efficient Fine-Tuning via Salient Channels
by: Zhao, Henry Hengyuan, et al.
Published: (2023)
by: Zhao, Henry Hengyuan, et al.
Published: (2023)
A New Perspective on Transformers in Online Reinforcement Learning for Continuous Control
by: Kachaev, Nikita, et al.
Published: (2025)
by: Kachaev, Nikita, et al.
Published: (2025)
Keep the General, Inject the Specific: Structured Dialogue Fine-Tuning for Knowledge Injection without Catastrophic Forgetting
by: Hong, Yijie, et al.
Published: (2025)
by: Hong, Yijie, et al.
Published: (2025)
Data filtering methods for training language models
by: Shevchenko, Egor, et al.
Published: (2026)
by: Shevchenko, Egor, et al.
Published: (2026)
GitGoodBench: A Novel Benchmark For Evaluating Agentic Performance On Git
by: Lindenbauer, Tobias, et al.
Published: (2025)
by: Lindenbauer, Tobias, et al.
Published: (2025)
S3: A Simple Strong Sample-effective Multimodal Dialog System
by: Rykov, Elisei, et al.
Published: (2024)
by: Rykov, Elisei, et al.
Published: (2024)
Tracing Persona Vectors Through LLM Pretraining
by: Moskvoretskii, Viktor, et al.
Published: (2026)
by: Moskvoretskii, Viktor, et al.
Published: (2026)
The One Where They Brain-Tune for Social Cognition: Multi-Modal Brain-Tuning on Friends
by: Policzer, Nico, et al.
Published: (2025)
by: Policzer, Nico, et al.
Published: (2025)
F3G-Avatar : Face Focused Full-body Gaussian Avatar
by: Menu, Willem, et al.
Published: (2026)
by: Menu, Willem, et al.
Published: (2026)
GIFT: Gradient-aware Immunization of diffusion models against malicious Fine-Tuning with safe concepts retention
by: Abdalla, Amro, et al.
Published: (2025)
by: Abdalla, Amro, et al.
Published: (2025)
Evolutionary Search for Automated Design of Uncertainty Quantification Methods
by: Seleznyov, Mikhail, et al.
Published: (2026)
by: Seleznyov, Mikhail, et al.
Published: (2026)
Small-scale photonic Kolmogorov-Arnold networks using standard telecom nonlinear modules
by: Calçado, Luca Nogueira, et al.
Published: (2026)
by: Calçado, Luca Nogueira, et al.
Published: (2026)
Similar Items
-
EBES: Easy Benchmarking for Event Sequences
by: Osin, Dmitry, et al.
Published: (2024) -
MLEM: Generative and Contrastive Learning as Distinct Modalities for Event Sequences
by: Moskvoretskii, Viktor, et al.
Published: (2024) -
Investigating the Impact of Quantization Methods on the Safety and Reliability of Large Language Models
by: Kharinaev, Artyom, et al.
Published: (2025) -
From 2:4 to 8:16 sparsity patterns in LLMs for Outliers and Weights with Variance Correction
by: Maximov, Egor, et al.
Published: (2025) -
Motivating Next-Gen Accelerators with Flexible (N:M) Activation Sparsity via Benchmarking Lightweight Post-Training Sparsification Approaches
by: Alanova, Shirin, et al.
Published: (2025)