Context Parametrization with Compositional Adapters
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jukić, Josip, Tutek, Martin, Šnajder, Jan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Disentangling Latent Shifts of In-Context Learning with Weak Supervision
von: Jukić, Josip, et al.
Veröffentlicht: (2024)
von: Jukić, Josip, et al.
Veröffentlicht: (2024)
From Robustness to Improved Generalization and Calibration in Pre-trained Language Models
von: Jukić, Josip, et al.
Veröffentlicht: (2024)
von: Jukić, Josip, et al.
Veröffentlicht: (2024)
Out-of-Distribution Detection by Leveraging Between-Layer Transformation Smoothness
von: Jelenić, Fran, et al.
Veröffentlicht: (2023)
von: Jelenić, Fran, et al.
Veröffentlicht: (2023)
Characterizing Linguistic Shifts in Croatian News via Diachronic Word Embeddings
von: Dukić, David, et al.
Veröffentlicht: (2025)
von: Dukić, David, et al.
Veröffentlicht: (2025)
Improving Data and Parameter Efficiency of Neural Language Models Using Representation Analysis
von: Jukić, Josip
Veröffentlicht: (2025)
von: Jukić, Josip
Veröffentlicht: (2025)
Sequence Repetition Enhances Token Embeddings and Improves Sequence Labeling with Decoder-only Language Models
von: Kukić, Matija Luka, et al.
Veröffentlicht: (2026)
von: Kukić, Matija Luka, et al.
Veröffentlicht: (2026)
Supervised In-Context Fine-Tuning for Generative Sequence Labeling
von: Dukić, David, et al.
Veröffentlicht: (2025)
von: Dukić, David, et al.
Veröffentlicht: (2025)
Claim Check-Worthiness Detection: How Well do LLMs Grasp Annotation Guidelines?
von: Majer, Laura, et al.
Veröffentlicht: (2024)
von: Majer, Laura, et al.
Veröffentlicht: (2024)
Looking Right is Sometimes Right: Investigating the Capabilities of Decoder-only LLMs for Sequence Labeling
von: Dukić, David, et al.
Veröffentlicht: (2024)
von: Dukić, David, et al.
Veröffentlicht: (2024)
Are ELECTRA's Sentence Embeddings Beyond Repair? The Case of Semantic Textual Similarity
von: Rep, Ivan, et al.
Veröffentlicht: (2024)
von: Rep, Ivan, et al.
Veröffentlicht: (2024)
LLMs for Targeted Sentiment in News Headlines: Exploring the Descriptive-Prescriptive Dilemma
von: Juroš, Jana, et al.
Veröffentlicht: (2024)
von: Juroš, Jana, et al.
Veröffentlicht: (2024)
CATfOOD: Counterfactual Augmented Training for Improving Out-of-Domain Performance and Calibration
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2023)
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2023)
What Makes You CLIC: Detection of Croatian Clickbait Headlines
von: Anđelić, Marija, et al.
Veröffentlicht: (2025)
von: Anđelić, Marija, et al.
Veröffentlicht: (2025)
REVS: Unlearning Sensitive Information in Language Models via Rank Editing in the Vocabulary Space
von: Ashuach, Tomer, et al.
Veröffentlicht: (2024)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2024)
Leveraging Open Information Extraction for More Robust Domain Transfer of Event Trigger Detection
von: Dukić, David, et al.
Veröffentlicht: (2023)
von: Dukić, David, et al.
Veröffentlicht: (2023)
TakeLab Retriever: AI-Driven Search Engine for Articles from Croatian News Outlets
von: Dukić, David, et al.
Veröffentlicht: (2024)
von: Dukić, David, et al.
Veröffentlicht: (2024)
Measuring Chain of Thought Faithfulness by Unlearning Reasoning Steps
von: Tutek, Martin, et al.
Veröffentlicht: (2025)
von: Tutek, Martin, et al.
Veröffentlicht: (2025)
FedMosaic: Federated Retrieval-Augmented Generation via Parametric Adapters
von: Liang, Zhilin, et al.
Veröffentlicht: (2026)
von: Liang, Zhilin, et al.
Veröffentlicht: (2026)
Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
PragWorld: A Benchmark Evaluating LLMs' Local World Model under Minimal Linguistic Alterations and Conversational Dynamics
von: Vashistha, Sachin, et al.
Veröffentlicht: (2025)
von: Vashistha, Sachin, et al.
Veröffentlicht: (2025)
Reasoning Models Know What's Important, and Encode It in Their Activations
von: Nikankin, Yaniv, et al.
Veröffentlicht: (2026)
von: Nikankin, Yaniv, et al.
Veröffentlicht: (2026)
ManagerBench: Evaluating the Safety-Pragmatism Trade-off in Autonomous LLMs
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
von: Simhi, Adi, et al.
Veröffentlicht: (2025)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Old Habits Die Hard: How Conversational History Geometrically Traps LLMs
von: Simhi, Adi, et al.
Veröffentlicht: (2026)
von: Simhi, Adi, et al.
Veröffentlicht: (2026)
Fixed and Adaptive Simultaneous Machine Translation Strategies Using Adapters
von: Issam, Abderrahmane, et al.
Veröffentlicht: (2024)
von: Issam, Abderrahmane, et al.
Veröffentlicht: (2024)
Performance Trade-offs of Optimizing Small Language Models for E-Commerce
von: Licardo, Josip Tomo, et al.
Veröffentlicht: (2025)
von: Licardo, Josip Tomo, et al.
Veröffentlicht: (2025)
Context-Parametric Inversion: Why Instruction Finetuning Can Worsen Context Reliance
von: Goyal, Sachin, et al.
Veröffentlicht: (2024)
von: Goyal, Sachin, et al.
Veröffentlicht: (2024)
On the Efficacy of Sampling Adapters
von: Meister, Clara, et al.
Veröffentlicht: (2023)
von: Meister, Clara, et al.
Veröffentlicht: (2023)
Context Copying Modulation: The Role of Entropy Neurons in Managing Parametric and Contextual Knowledge Conflicts
von: Tighidet, Zineddine, et al.
Veröffentlicht: (2025)
von: Tighidet, Zineddine, et al.
Veröffentlicht: (2025)
CLIP-Adapter: Better Vision-Language Models with Feature Adapters
von: Gao, Peng, et al.
Veröffentlicht: (2021)
von: Gao, Peng, et al.
Veröffentlicht: (2021)
Findings of the BlackboxNLP 2025 Shared Task: Localizing Circuits and Causal Variables in Language Models
von: Arad, Dana, et al.
Veröffentlicht: (2025)
von: Arad, Dana, et al.
Veröffentlicht: (2025)
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
von: Araujo, Vladimir, et al.
Veröffentlicht: (2024)
von: Araujo, Vladimir, et al.
Veröffentlicht: (2024)
When Context Leads but Parametric Memory Follows in Large Language Models
von: Tao, Yufei, et al.
Veröffentlicht: (2024)
von: Tao, Yufei, et al.
Veröffentlicht: (2024)
Updating Parametric Knowledge with Context Distillation Retains Post-Training Capabilities
von: Padmanabhan, Shankar, et al.
Veröffentlicht: (2026)
von: Padmanabhan, Shankar, et al.
Veröffentlicht: (2026)
Limited-Resource Adapters Are Regularizers, Not Linguists
von: Fekete, Marcell, et al.
Veröffentlicht: (2025)
von: Fekete, Marcell, et al.
Veröffentlicht: (2025)
The Hidden Space of Transformer Language Adapters
von: Alabi, Jesujoba O., et al.
Veröffentlicht: (2024)
von: Alabi, Jesujoba O., et al.
Veröffentlicht: (2024)
Adapters Mixup: Mixing Parameter-Efficient Adapters to Enhance the Adversarial Robustness of Fine-tuned Pre-trained Text Classifiers
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuc, et al.
Veröffentlicht: (2024)
How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models
von: Kim, Minsung, et al.
Veröffentlicht: (2025)
von: Kim, Minsung, et al.
Veröffentlicht: (2025)
Task-Aware LoRA Adapter Composition via Similarity Retrieval in Vector Databases
von: Adsul, Riya, et al.
Veröffentlicht: (2026)
von: Adsul, Riya, et al.
Veröffentlicht: (2026)
Hadamard Adapter: An Extreme Parameter-Efficient Adapter Tuning Method for Pre-trained Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Disentangling Latent Shifts of In-Context Learning with Weak Supervision
von: Jukić, Josip, et al.
Veröffentlicht: (2024) -
From Robustness to Improved Generalization and Calibration in Pre-trained Language Models
von: Jukić, Josip, et al.
Veröffentlicht: (2024) -
Out-of-Distribution Detection by Leveraging Between-Layer Transformation Smoothness
von: Jelenić, Fran, et al.
Veröffentlicht: (2023) -
Characterizing Linguistic Shifts in Croatian News via Diachronic Word Embeddings
von: Dukić, David, et al.
Veröffentlicht: (2025) -
Improving Data and Parameter Efficiency of Neural Language Models Using Representation Analysis
von: Jukić, Josip
Veröffentlicht: (2025)