Salvato in:
| Autori principali: | Fleshman, William, Khan, Aleem, Marone, Marc, Van Durme, Benjamin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2404.08417 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SEQR: Secure and Efficient QR-based LoRA Routing
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
LoRA-Augmented Generation (LAG) for Knowledge-Intensive Language Tasks
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
RE-Adapt: Reverse Engineered Adaptation of Large Language Models
di: Fleshman, William, et al.
Pubblicazione: (2024)
di: Fleshman, William, et al.
Pubblicazione: (2024)
SpectR: Dynamically Composing LM Experts with Spectral Routing
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
RE-AdaptIR: Improving Information Retrieval through Reverse Engineered Adaptation
di: Fleshman, William, et al.
Pubblicazione: (2024)
di: Fleshman, William, et al.
Pubblicazione: (2024)
mmBERT: A Modern Multilingual Encoder with Annealed Language Learning
di: Marone, Marc, et al.
Pubblicazione: (2025)
di: Marone, Marc, et al.
Pubblicazione: (2025)
"According to ...": Prompting Language Models Improves Quoting from Pre-Training Data
di: Weller, Orion, et al.
Pubblicazione: (2023)
di: Weller, Orion, et al.
Pubblicazione: (2023)
Generative Adapter: Contextualizing Language Models in Parameters with A Single Forward Pass
di: Chen, Tong, et al.
Pubblicazione: (2024)
di: Chen, Tong, et al.
Pubblicazione: (2024)
LLMs in the Imaginarium: Tool Learning through Simulated Trial and Error
di: Wang, Boshi, et al.
Pubblicazione: (2024)
di: Wang, Boshi, et al.
Pubblicazione: (2024)
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
di: Jiang, Dongwei, et al.
Pubblicazione: (2024)
di: Jiang, Dongwei, et al.
Pubblicazione: (2024)
Always Tell Me The Odds: Fine-grained Conditional Probability Estimation
di: Wang, Liaoyaqi, et al.
Pubblicazione: (2025)
di: Wang, Liaoyaqi, et al.
Pubblicazione: (2025)
Do Androids Know They're Only Dreaming of Electric Sheep?
di: CH-Wang, Sky, et al.
Pubblicazione: (2023)
di: CH-Wang, Sky, et al.
Pubblicazione: (2023)
Sample-Efficient Online Learning in LM Agents via Hindsight Trajectory Rewriting
di: Hu, Michael Y., et al.
Pubblicazione: (2025)
di: Hu, Michael Y., et al.
Pubblicazione: (2025)
MICE for CATs: Model-Internal Confidence Estimation for Calibrating Agents with Tools
di: Subramani, Nishant, et al.
Pubblicazione: (2025)
di: Subramani, Nishant, et al.
Pubblicazione: (2025)
How to Train Data-Efficient LLMs
di: Sachdeva, Noveen, et al.
Pubblicazione: (2024)
di: Sachdeva, Noveen, et al.
Pubblicazione: (2024)
Continuous Approximations for Improving Quantization Aware Training of LLMs
di: Li, He, et al.
Pubblicazione: (2024)
di: Li, He, et al.
Pubblicazione: (2024)
Seq vs Seq: An Open Suite of Paired Encoders and Decoders
di: Weller, Orion, et al.
Pubblicazione: (2025)
di: Weller, Orion, et al.
Pubblicazione: (2025)
Training-Free Bayesianization for Low-Rank Adapters of Large Language Models
di: Shi, Haizhou, et al.
Pubblicazione: (2024)
di: Shi, Haizhou, et al.
Pubblicazione: (2024)
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
di: Araujo, Vladimir, et al.
Pubblicazione: (2024)
di: Araujo, Vladimir, et al.
Pubblicazione: (2024)
Layer Swapping for Zero-Shot Cross-Lingual Transfer in Large Language Models
di: Bandarkar, Lucas, et al.
Pubblicazione: (2024)
di: Bandarkar, Lucas, et al.
Pubblicazione: (2024)
JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models
di: Dragomir, Alexandra, et al.
Pubblicazione: (2026)
di: Dragomir, Alexandra, et al.
Pubblicazione: (2026)
K-Merge: Online Continual Merging of Adapters for On-device Large Language Models
di: Shenaj, Donald, et al.
Pubblicazione: (2025)
di: Shenaj, Donald, et al.
Pubblicazione: (2025)
CuMA: Aligning LLMs with Sparse Cultural Values via Demographic-Aware Mixture of Adapters
di: Sun, Ao, et al.
Pubblicazione: (2026)
di: Sun, Ao, et al.
Pubblicazione: (2026)
Data-driven Clustering and Merging of Adapters for On-device Large Language Models
di: Bohdal, Ondrej, et al.
Pubblicazione: (2026)
di: Bohdal, Ondrej, et al.
Pubblicazione: (2026)
Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs
di: Pepper, Keenan, et al.
Pubblicazione: (2026)
di: Pepper, Keenan, et al.
Pubblicazione: (2026)
KV-Distill: Nearly Lossless Learnable Context Compression for LLMs
di: Chari, Vivek, et al.
Pubblicazione: (2025)
di: Chari, Vivek, et al.
Pubblicazione: (2025)
Efficient Exploration for LLMs
di: Dwaracherla, Vikranth, et al.
Pubblicazione: (2024)
di: Dwaracherla, Vikranth, et al.
Pubblicazione: (2024)
AutoScale: Scale-Aware Data Mixing for Pre-Training LLMs
di: Kang, Feiyang, et al.
Pubblicazione: (2024)
di: Kang, Feiyang, et al.
Pubblicazione: (2024)
Not All Adapters Matter: Selective Adapter Freezing for Memory-Efficient Fine-Tuning of Language Models
di: Son, Hyegang, et al.
Pubblicazione: (2024)
di: Son, Hyegang, et al.
Pubblicazione: (2024)
Predicting Training Re-evaluation Curves Enables Effective Data Curriculums for LLMs
di: Bergsma, Shane, et al.
Pubblicazione: (2025)
di: Bergsma, Shane, et al.
Pubblicazione: (2025)
Finer Parameter Steps for Low-Rank PEFT: A Controlled Study with CP Tensor Adapters
di: Wang, Xinjue, et al.
Pubblicazione: (2026)
di: Wang, Xinjue, et al.
Pubblicazione: (2026)
Ensembles of Low-Rank Expert Adapters
di: Li, Yinghao, et al.
Pubblicazione: (2025)
di: Li, Yinghao, et al.
Pubblicazione: (2025)
MoKA: Mixture of Kronecker Adapters
di: Sadeghi, Mohammadreza, et al.
Pubblicazione: (2025)
di: Sadeghi, Mohammadreza, et al.
Pubblicazione: (2025)
Random Initialization of Gated Sparse Adapters
di: Retault, Vi, et al.
Pubblicazione: (2025)
di: Retault, Vi, et al.
Pubblicazione: (2025)
Dodo: Dynamic Contextual Compression for Decoder-only LMs
di: Qin, Guanghui, et al.
Pubblicazione: (2023)
di: Qin, Guanghui, et al.
Pubblicazione: (2023)
PII-Scope: A Comprehensive Study on Training Data PII Extraction Attacks in LLMs
di: Nakka, Krishna Kanth, et al.
Pubblicazione: (2024)
di: Nakka, Krishna Kanth, et al.
Pubblicazione: (2024)
Dual-Personalizing Adapter for Federated Foundation Models
di: Yang, Yiyuan, et al.
Pubblicazione: (2024)
di: Yang, Yiyuan, et al.
Pubblicazione: (2024)
Learning Adapter Rank via Symmetry Breaking
di: Doyle, Cooper, et al.
Pubblicazione: (2025)
di: Doyle, Cooper, et al.
Pubblicazione: (2025)
Neutral Residues: Revisiting Adapters for Model Extension
di: Talla, Franck Signe, et al.
Pubblicazione: (2024)
di: Talla, Franck Signe, et al.
Pubblicazione: (2024)
Connecting the Dots: LLMs can Infer and Verbalize Latent Structure from Disparate Training Data
di: Treutlein, Johannes, et al.
Pubblicazione: (2024)
di: Treutlein, Johannes, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SEQR: Secure and Efficient QR-based LoRA Routing
di: Fleshman, William, et al.
Pubblicazione: (2025) -
LoRA-Augmented Generation (LAG) for Knowledge-Intensive Language Tasks
di: Fleshman, William, et al.
Pubblicazione: (2025) -
RE-Adapt: Reverse Engineered Adaptation of Large Language Models
di: Fleshman, William, et al.
Pubblicazione: (2024) -
SpectR: Dynamically Composing LM Experts with Spectral Routing
di: Fleshman, William, et al.
Pubblicazione: (2025) -
RE-AdaptIR: Improving Information Retrieval through Reverse Engineered Adaptation
di: Fleshman, William, et al.
Pubblicazione: (2024)