Sculpting Subspaces: Constrained Full Fine-Tuning in LLMs for Continual Learning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Nayak, Nikhil Shivakumar, Killamsetty, Krishnateja, Han, Ligong, Bhandwaldar, Abhishek, Chanda, Prateek, Xu, Kai, Wang, Hao, Pareja, Aldo, Silkin, Oleg, Eyceoz, Mustafa, Srivastava, Akash |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Unveiling the Secret Recipe: A Guide For Supervised Fine-Tuning Small LLMs
par: Pareja, Aldo, et autres
Publié: (2024)
par: Pareja, Aldo, et autres
Publié: (2024)
Hopscotch: Discovering and Skipping Redundancies in Language Models
par: Eyceoz, Mustafa, et autres
Publié: (2025)
par: Eyceoz, Mustafa, et autres
Publié: (2025)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
par: Chanda, Prateek, et autres
Publié: (2025)
par: Chanda, Prateek, et autres
Publié: (2025)
DELIFT: Data Efficient Language model Instruction Fine Tuning
par: Agarwal, Ishika, et autres
Publié: (2024)
par: Agarwal, Ishika, et autres
Publié: (2024)
LAB: Large-Scale Alignment for ChatBots
par: Sudalairaj, Shivchander, et autres
Publié: (2024)
par: Sudalairaj, Shivchander, et autres
Publié: (2024)
Agent Factories for High Level Synthesis: How Far Can General-Purpose Coding Agents Go in Hardware Optimization?
par: Bhandwaldar, Abhishek, et autres
Publié: (2026)
par: Bhandwaldar, Abhishek, et autres
Publié: (2026)
SCoRe: Submodular Combinatorial Representation Learning
par: Majee, Anay, et autres
Publié: (2023)
par: Majee, Anay, et autres
Publié: (2023)
SQuat: Subspace-orthogonal KV Cache Quantization
par: Wang, Hao, et autres
Publié: (2025)
par: Wang, Hao, et autres
Publié: (2025)
Graph Attention for Heterogeneous Graphs with Positional Encoding
par: Nayak, Nikhil Shivakumar
Publié: (2025)
par: Nayak, Nikhil Shivakumar
Publié: (2025)
Mathematical Modeling of Option Pricing with an Extended Black-Scholes Framework
par: Nayak, Nikhil Shivakumar
Publié: (2025)
par: Nayak, Nikhil Shivakumar
Publié: (2025)
Mitigating Premature Exploitation in Particle-based Monte Carlo for Inference-Time Scaling
par: Giannone, Giorgio, et autres
Publié: (2025)
par: Giannone, Giorgio, et autres
Publié: (2025)
S2D2: Fast Decoding for Diffusion LLMs via Training-Free Self-Speculation
par: Han, Ligong, et autres
Publié: (2026)
par: Han, Ligong, et autres
Publié: (2026)
Spectrum-Aware Parameter Efficient Fine-Tuning for Diffusion Models
par: Zhang, Xinxi, et autres
Publié: (2024)
par: Zhang, Xinxi, et autres
Publié: (2024)
Zeroth-Order Fine-Tuning of LLMs in Random Subspaces
par: Yu, Ziming, et autres
Publié: (2024)
par: Yu, Ziming, et autres
Publié: (2024)
SNLP: Layer-Parallel Inference via Structured Newton Corrections
par: Han, Ligong, et autres
Publié: (2026)
par: Han, Ligong, et autres
Publié: (2026)
Application and Efficacy of diatom diversity indices for water quality evaluation of Chambal River System
par: Srivastava, Prateek
Publié: (2025)
par: Srivastava, Prateek
Publié: (2025)
TopoSculpt: Betti-Steered Topological Sculpting of 3D Fine-grained Tubular Shapes
par: Zhang, Minghui, et autres
Publié: (2025)
par: Zhang, Minghui, et autres
Publié: (2025)
Towards Interpretable Soft Prompts
par: Patel, Oam, et autres
Publié: (2025)
par: Patel, Oam, et autres
Publié: (2025)
Bayesian Coreset Optimization for Personalized Federated Learning
par: Chanda, Prateek, et autres
Publié: (2025)
par: Chanda, Prateek, et autres
Publié: (2025)
Med42 -- Evaluating Fine-Tuning Strategies for Medical LLMs: Full-Parameter vs. Parameter-Efficient Approaches
par: Christophe, Clément, et autres
Publié: (2024)
par: Christophe, Clément, et autres
Publié: (2024)
Extractive Schema Linking for Text-to-SQL
par: Glass, Michael, et autres
Publié: (2025)
par: Glass, Michael, et autres
Publié: (2025)
Efficient Orthogonal Fine-Tuning with Principal Subspace Adaptation
par: Wu, Fei, et autres
Publié: (2025)
par: Wu, Fei, et autres
Publié: (2025)
ROSA: Random Subspace Adaptation for Efficient Fine-Tuning
par: Hameed, Marawan Gamal Abdel, et autres
Publié: (2024)
par: Hameed, Marawan Gamal Abdel, et autres
Publié: (2024)
Parameter-Efficient Subspace Optimization for LLM Fine-Tuning
par: Lou, Yuchen, et autres
Publié: (2025)
par: Lou, Yuchen, et autres
Publié: (2025)
LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models
par: Sikdar, Prateek Kumar
Publié: (2026)
par: Sikdar, Prateek Kumar
Publié: (2026)
Memory-Efficient Backpropagation for Fine-Tuning LLMs on Resource-Constrained Mobile Devices
par: Song, Congzheng, et autres
Publié: (2025)
par: Song, Congzheng, et autres
Publié: (2025)
Dr. SoW: Density Ratio of Strong-over-weak LLMs for Reducing the Cost of Human Annotation in Preference Tuning
par: Xu, Guangxuan, et autres
Publié: (2024)
par: Xu, Guangxuan, et autres
Publié: (2024)
Differentially Private Subspace Fine-Tuning for Large Language Models
par: Zheng, Lele, et autres
Publié: (2026)
par: Zheng, Lele, et autres
Publié: (2026)
Beyond QA Pairs: Assessing Parameter-Efficient Fine-Tuning for Fact Embedding in LLMs
par: Ratnakar, Shivam, et autres
Publié: (2025)
par: Ratnakar, Shivam, et autres
Publié: (2025)
Keeping Code-Aware LLMs Fresh: Full Refresh, In-Context Deltas, and Incremental Fine-Tuning
par: Sharma, Pradeep Kumar, et autres
Publié: (2025)
par: Sharma, Pradeep Kumar, et autres
Publié: (2025)
Safety Subspaces are Not Linearly Distinct: A Fine-Tuning Case Study
par: Ponkshe, Kaustubh, et autres
Publié: (2025)
par: Ponkshe, Kaustubh, et autres
Publié: (2025)
Sculpting priors
par: Theiler, James
Publié: (2024)
par: Theiler, James
Publié: (2024)
Ramanujan Graphs and Interlacing Families
par: Srivastava, Nikhil
Publié: (2024)
par: Srivastava, Nikhil
Publié: (2024)
InfoSculpt: Sculpting the Latent Space for Generalized Category Discovery
par: Liao, Wenwen, et autres
Publié: (2026)
par: Liao, Wenwen, et autres
Publié: (2026)
INST-Sculpt: Interactive Stroke-based Neural SDF Sculpting
par: Rubab, Fizza, et autres
Publié: (2025)
par: Rubab, Fizza, et autres
Publié: (2025)
Gaze2Report: Radiology Report Generation via Visual-Gaze Prompt Tuning of LLMs
par: Konwer, Aishik, et autres
Publié: (2026)
par: Konwer, Aishik, et autres
Publié: (2026)
One-Pass to Reason: Token Duplication and Block-Sparse Mask for Efficient Fine-Tuning on Multi-Turn Reasoning
par: Goru, Ritesh, et autres
Publié: (2025)
par: Goru, Ritesh, et autres
Publié: (2025)
DePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning
par: Shi, Zhengxiang, et autres
Publié: (2023)
par: Shi, Zhengxiang, et autres
Publié: (2023)
RevFFN: Memory-Efficient Full-Parameter Fine-Tuning of Mixture-of-Experts LLMs with Reversible Blocks
par: Liu, Ningyuan, et autres
Publié: (2025)
par: Liu, Ningyuan, et autres
Publié: (2025)
Fine-Tuning Vision-Language Models for Multimodal Polymer Property Prediction
par: Vuong, An, et autres
Publié: (2025)
par: Vuong, An, et autres
Publié: (2025)
Documents similaires
-
Unveiling the Secret Recipe: A Guide For Supervised Fine-Tuning Small LLMs
par: Pareja, Aldo, et autres
Publié: (2024) -
Hopscotch: Discovering and Skipping Redundancies in Language Models
par: Eyceoz, Mustafa, et autres
Publié: (2025) -
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
par: Chanda, Prateek, et autres
Publié: (2025) -
DELIFT: Data Efficient Language model Instruction Fine Tuning
par: Agarwal, Ishika, et autres
Publié: (2024) -
LAB: Large-Scale Alignment for ChatBots
par: Sudalairaj, Shivchander, et autres
Publié: (2024)