GiVA: Gradient-Informed Bases for Vector-Based Adaptation
Fuente:
arXiv
Salvato in:
| Autori principali: | Gangwar, Neeraj, Deshmukh, Rishabh, Shavlovsky, Michael, Li, Hancao, Mittal, Vivek, Ying, Lexing, Kani, Nickvash |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Integrating Arithmetic Learning Improves Mathematical Reasoning in Smaller Models
di: Gangwar, Neeraj, et al.
Pubblicazione: (2025)
di: Gangwar, Neeraj, et al.
Pubblicazione: (2025)
E-Gen: Leveraging E-Graphs to Improve Continuous Representations of Symbolic Expressions
di: Zheng, Hongbo, et al.
Pubblicazione: (2025)
di: Zheng, Hongbo, et al.
Pubblicazione: (2025)
Parameter-Efficient Multi-Task Learning via Progressive Task-Specific Adaptation
di: Gangwar, Neeraj, et al.
Pubblicazione: (2025)
di: Gangwar, Neeraj, et al.
Pubblicazione: (2025)
Mathematical Derivation Graphs: A Relation Extraction Task in STEM Manuscripts
di: Prasad, Vishesh, et al.
Pubblicazione: (2024)
di: Prasad, Vishesh, et al.
Pubblicazione: (2024)
Building Robust and Scalable Multilingual ASR for Indian Languages
di: Gangwar, Arjun, et al.
Pubblicazione: (2025)
di: Gangwar, Arjun, et al.
Pubblicazione: (2025)
STEM-POM: Evaluating Language Models Math-Symbol Reasoning in Document Parsing
di: Zou, Jiaru, et al.
Pubblicazione: (2024)
di: Zou, Jiaru, et al.
Pubblicazione: (2024)
COS-DPO: Conditioned One-Shot Multi-Objective Fine-Tuning Framework
di: Ren, Yinuo, et al.
Pubblicazione: (2024)
di: Ren, Yinuo, et al.
Pubblicazione: (2024)
Hallucination-Resistant, Domain-Specific Research Assistant with Self-Evaluation and Vector-Grounded Retrieval
di: Bhavsar, Vivek, et al.
Pubblicazione: (2025)
di: Bhavsar, Vivek, et al.
Pubblicazione: (2025)
Lottery Ticket Adaptation: Mitigating Destructive Interference in LLMs
di: Panda, Ashwinee, et al.
Pubblicazione: (2024)
di: Panda, Ashwinee, et al.
Pubblicazione: (2024)
Schema Augmentation for Zero-Shot Domain Adaptation in Dialogue State Tracking
di: Richardson, Christopher, et al.
Pubblicazione: (2024)
di: Richardson, Christopher, et al.
Pubblicazione: (2024)
AMUSED: A Multi-Stream Vector Representation Method for Use in Natural Dialogue
di: Kumar, Gaurav, et al.
Pubblicazione: (2019)
di: Kumar, Gaurav, et al.
Pubblicazione: (2019)
Scaling Textual Gradients via Sampling-Based Momentum
di: Ding, Zixin, et al.
Pubblicazione: (2025)
di: Ding, Zixin, et al.
Pubblicazione: (2025)
PsyAgent: Constructing Human-like Agents Based on Psychological Modeling and Contextual Interaction
di: Meng, Zibin, et al.
Pubblicazione: (2026)
di: Meng, Zibin, et al.
Pubblicazione: (2026)
Data-Augmentation-Based Dialectal Adaptation for LLMs
di: Faisal, Fahim, et al.
Pubblicazione: (2024)
di: Faisal, Fahim, et al.
Pubblicazione: (2024)
Enhancing Large Language Model Performance with Gradient-Based Parameter Selection
di: Li, Haoling, et al.
Pubblicazione: (2024)
di: Li, Haoling, et al.
Pubblicazione: (2024)
KBAlign: Efficient Self Adaptation on Specific Knowledge Bases
di: Zeng, Zheni, et al.
Pubblicazione: (2024)
di: Zeng, Zheni, et al.
Pubblicazione: (2024)
SuperMerge: An Approach For Gradient-Based Model Merging
di: Yang, Haoyu, et al.
Pubblicazione: (2024)
di: Yang, Haoyu, et al.
Pubblicazione: (2024)
Annotation-Efficient Vision-Language Model Adaptation to the Polish Language Using the LLaVA Framework
di: Statkiewicz, Grzegorz, et al.
Pubblicazione: (2026)
di: Statkiewicz, Grzegorz, et al.
Pubblicazione: (2026)
GRAIL: Gradient-Based Adaptive Unlearning for Privacy and Copyright in LLMs
di: Kim, Kun-Woo, et al.
Pubblicazione: (2025)
di: Kim, Kun-Woo, et al.
Pubblicazione: (2025)
Beyond LoRA vs. Full Fine-Tuning: Gradient-Guided Optimizer Routing for LLM Adaptation
di: Tang, Haozhan, et al.
Pubblicazione: (2026)
di: Tang, Haozhan, et al.
Pubblicazione: (2026)
Enhancing Cache-Augmented Generation (CAG) with Adaptive Contextual Compression for Scalable Knowledge Integration
di: Agrawal, Rishabh, et al.
Pubblicazione: (2025)
di: Agrawal, Rishabh, et al.
Pubblicazione: (2025)
Multimodal Multihop Source Retrieval for Web Question Answering
di: Yarrabelly, Navya, et al.
Pubblicazione: (2025)
di: Yarrabelly, Navya, et al.
Pubblicazione: (2025)
C2-Faith: Benchmarking LLM Judges for Causal and Coverage Faithfulness in Chain-of-Thought Reasoning
di: Mittal, Avni, et al.
Pubblicazione: (2026)
di: Mittal, Avni, et al.
Pubblicazione: (2026)
ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding
di: Wang, Austin T., et al.
Pubblicazione: (2025)
di: Wang, Austin T., et al.
Pubblicazione: (2025)
Model Merging by Uncertainty-Based Gradient Matching
di: Daheim, Nico, et al.
Pubblicazione: (2023)
di: Daheim, Nico, et al.
Pubblicazione: (2023)
Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies
di: Mittal, Avni
Pubblicazione: (2026)
di: Mittal, Avni
Pubblicazione: (2026)
Did You Forget What I Asked? Prospective Memory Failures in Large Language Models
di: Mittal, Avni
Pubblicazione: (2026)
di: Mittal, Avni
Pubblicazione: (2026)
Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits
di: Ahmad, Areeb, et al.
Pubblicazione: (2025)
di: Ahmad, Areeb, et al.
Pubblicazione: (2025)
Scatter-Based Innovation Propagation in Large Language Models for Multi-Stage Process Adaptation
di: Su, Hong
Pubblicazione: (2025)
di: Su, Hong
Pubblicazione: (2025)
A Lightweight Framework for Trigger-Guided LoRA-Based Self-Adaptation in LLMs
di: Wei, Jiacheng, et al.
Pubblicazione: (2025)
di: Wei, Jiacheng, et al.
Pubblicazione: (2025)
Matrix-Transformation Based Low-Rank Adaptation (MTLoRA): A Brain-Inspired Method for Parameter-Efficient Fine-Tuning
di: Liang, Yao, et al.
Pubblicazione: (2024)
di: Liang, Yao, et al.
Pubblicazione: (2024)
GoRA: Gradient-driven Adaptive Low Rank Adaptation
di: He, Haonan, et al.
Pubblicazione: (2025)
di: He, Haonan, et al.
Pubblicazione: (2025)
SwitchLingua: The First Large-Scale Multilingual and Multi-Ethnic Code-Switching Dataset
di: Xie, Peng, et al.
Pubblicazione: (2025)
di: Xie, Peng, et al.
Pubblicazione: (2025)
Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs
di: Han, Pengrui, et al.
Pubblicazione: (2026)
di: Han, Pengrui, et al.
Pubblicazione: (2026)
Vector-ICL: In-context Learning with Continuous Vector Representations
di: Zhuang, Yufan, et al.
Pubblicazione: (2024)
di: Zhuang, Yufan, et al.
Pubblicazione: (2024)
Flexora: Flexible Low Rank Adaptation for Large Language Models
di: Wei, Chenxing, et al.
Pubblicazione: (2024)
di: Wei, Chenxing, et al.
Pubblicazione: (2024)
Vectoring Languages
di: Chen, Joseph
Pubblicazione: (2024)
di: Chen, Joseph
Pubblicazione: (2024)
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
di: Bhardwaj, Rishabh, et al.
Pubblicazione: (2024)
di: Bhardwaj, Rishabh, et al.
Pubblicazione: (2024)
"Hunt Takes Hare": Theming Games Through Game-Word Vector Translation
di: Younès, Rabii, et al.
Pubblicazione: (2024)
di: Younès, Rabii, et al.
Pubblicazione: (2024)
Deep Learning Based Named Entity Recognition Models for Recipes
di: Goel, Mansi, et al.
Pubblicazione: (2024)
di: Goel, Mansi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Integrating Arithmetic Learning Improves Mathematical Reasoning in Smaller Models
di: Gangwar, Neeraj, et al.
Pubblicazione: (2025) -
E-Gen: Leveraging E-Graphs to Improve Continuous Representations of Symbolic Expressions
di: Zheng, Hongbo, et al.
Pubblicazione: (2025) -
Parameter-Efficient Multi-Task Learning via Progressive Task-Specific Adaptation
di: Gangwar, Neeraj, et al.
Pubblicazione: (2025) -
Mathematical Derivation Graphs: A Relation Extraction Task in STEM Manuscripts
di: Prasad, Vishesh, et al.
Pubblicazione: (2024) -
Building Robust and Scalable Multilingual ASR for Indian Languages
di: Gangwar, Arjun, et al.
Pubblicazione: (2025)