Limited-Resource Adapters Are Regularizers, Not Linguists
Fuente:
arXiv
Saved in:
| Main Authors: | Fekete, Marcell, Robinson, Nathaniel R., Lavrinovics, Ernests, Jean-Baptiste, E. Djeride, Dabre, Raj, Bjerva, Johannes, Lent, Heather |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MultiHal: Multilingual Dataset for Knowledge-Graph Grounded Evaluation of LLM Hallucinations
by: Lavrinovics, Ernests, et al.
Published: (2025)
by: Lavrinovics, Ernests, et al.
Published: (2025)
Linguistically Grounded Analysis of Language Models using Shapley Head Values
by: Fekete, Marcell, et al.
Published: (2024)
by: Fekete, Marcell, et al.
Published: (2024)
Knowledge Graphs, Large Language Models, and Hallucinations: An NLP Perspective
by: Lavrinovics, Ernests, et al.
Published: (2024)
by: Lavrinovics, Ernests, et al.
Published: (2024)
Text Embedding Inversion Security for Multilingual Language Models
by: Chen, Yiyi, et al.
Published: (2024)
by: Chen, Yiyi, et al.
Published: (2024)
When Discourse Pressures Conflict: Information Structure in Vision-Language Model Outputs
by: Fekete, Marcell, et al.
Published: (2026)
by: Fekete, Marcell, et al.
Published: (2026)
Against All Odds: Overcoming Typology, Script, and Language Confusion in Multilingual Embedding Inversion Attacks
by: Chen, Yiyi, et al.
Published: (2024)
by: Chen, Yiyi, et al.
Published: (2024)
CreoleVal: Multilingual Multitask Benchmarks for Creoles
by: Lent, Heather, et al.
Published: (2023)
by: Lent, Heather, et al.
Published: (2023)
Sociolinguistically Informed Interpretability: A Case Study on Hinglish Emotion Classification
by: Tatariya, Kushal, et al.
Published: (2024)
by: Tatariya, Kushal, et al.
Published: (2024)
Beyond Weaponization: NLP Security for Medium and Lower-Resourced Languages in Their Own Right
by: Lent, Heather
Published: (2025)
by: Lent, Heather
Published: (2025)
NLP Security and Ethics, in the Wild
by: Lent, Heather, et al.
Published: (2025)
by: Lent, Heather, et al.
Published: (2025)
Are Language Models Agnostic to Linguistically Grounded Perturbations? A Case Study of Indic Languages
by: Ghosh, Poulami, et al.
Published: (2024)
by: Ghosh, Poulami, et al.
Published: (2024)
Top-b: Entropic Regulation of Relative Probability Bands in Autoregressive Language Processes
by: Halder, Deepon, et al.
Published: (2026)
by: Halder, Deepon, et al.
Published: (2026)
How Good is Your Wikipedia? Auditing Data Quality for Low-resource and Multilingual NLP
by: Tatariya, Kushal, et al.
Published: (2024)
by: Tatariya, Kushal, et al.
Published: (2024)
How effective is Multi-source pivoting for Translation of Low Resource Indian Languages?
by: Gaikwad, Pranav, et al.
Published: (2024)
by: Gaikwad, Pranav, et al.
Published: (2024)
Patterns of Persistence and Diffusibility across the World's Languages
by: Chen, Yiyi, et al.
Published: (2024)
by: Chen, Yiyi, et al.
Published: (2024)
Pretraining Language Models Using Translationese
by: Doshi, Meet, et al.
Published: (2024)
by: Doshi, Meet, et al.
Published: (2024)
Scripts Through Time: A Survey of the Evolving Role of Transliteration in NLP
by: Jayakumar, Thanmay, et al.
Published: (2026)
by: Jayakumar, Thanmay, et al.
Published: (2026)
Connecting Ideas in 'Lower-Resource' Scenarios: NLP for National Varieties, Creoles and Other Low-resource Scenarios
by: Joshi, Aditya, et al.
Published: (2024)
by: Joshi, Aditya, et al.
Published: (2024)
An Empirical Study of In-context Learning in LLMs for Machine Translation
by: Chitale, Pranjal A., et al.
Published: (2024)
by: Chitale, Pranjal A., et al.
Published: (2024)
CycleDistill: Bootstrapping Machine Translation using LLMs with Cyclical Distillation
by: Halder, Deepon, et al.
Published: (2025)
by: Halder, Deepon, et al.
Published: (2025)
Multilingual Gradient Word-Order Typology from Universal Dependencies
by: Baylor, Emi, et al.
Published: (2024)
by: Baylor, Emi, et al.
Published: (2024)
Mark My Words: A Robust Multilingual Model for Punctuation in Text and Speech Transcripts
by: Pulipaka, Sidharth, et al.
Published: (2025)
by: Pulipaka, Sidharth, et al.
Published: (2025)
A Morphology-Based Investigation of Positional Encodings
by: Ghosh, Poulami, et al.
Published: (2024)
by: Ghosh, Poulami, et al.
Published: (2024)
How Good is Zero-Shot MT Evaluation for Low Resource Indian Languages?
by: Singh, Anushka, et al.
Published: (2024)
by: Singh, Anushka, et al.
Published: (2024)
Follow the Path: Reasoning over Knowledge Graph Paths to Improve Large Language Model Factuality
by: Zhang, Mike, et al.
Published: (2025)
by: Zhang, Mike, et al.
Published: (2025)
IndicRAGSuite: Large-Scale Datasets and a Benchmark for Indian Language RAG Systems
by: Prasanjith, Pasunuti, et al.
Published: (2025)
by: Prasanjith, Pasunuti, et al.
Published: (2025)
Characterizing Memorization in Diffusion Language Models: Generalized Extraction and Sampling Effects
by: Luo, Xiaoyu, et al.
Published: (2026)
by: Luo, Xiaoyu, et al.
Published: (2026)
The Reasoning Lingua Franca: A Double-Edged Sword for Multilingual AI
by: Saji, Alan, et al.
Published: (2025)
by: Saji, Alan, et al.
Published: (2025)
What is "Typological Diversity" in NLP?
by: Ploeger, Esther, et al.
Published: (2024)
by: Ploeger, Esther, et al.
Published: (2024)
Shared Path: Unraveling Memorization in Multilingual LLMs through Language Similarities
by: Luo, Xiaoyu, et al.
Published: (2025)
by: Luo, Xiaoyu, et al.
Published: (2025)
Do LLMs Really Memorize Personally Identifiable Information? Revisiting PII Leakage with a Cue-Controlled Memorization Framework
by: Luo, Xiaoyu, et al.
Published: (2026)
by: Luo, Xiaoyu, et al.
Published: (2026)
Malaysian English News Decoded: A Linguistic Resource for Named Entity and Relation Extraction
by: Chanthran, Mohan Raj, et al.
Published: (2024)
by: Chanthran, Mohan Raj, et al.
Published: (2024)
The Responsible Development of Automated Student Feedback with Generative AI
by: Lindsay, Euan D, et al.
Published: (2023)
by: Lindsay, Euan D, et al.
Published: (2023)
ALGEN: Few-shot Inversion Attacks on Textual Embeddings using Alignment and Generation
by: Chen, Yiyi, et al.
Published: (2025)
by: Chen, Yiyi, et al.
Published: (2025)
RiddleBench: A New Generative Reasoning Benchmark for LLMs
by: Halder, Deepon, et al.
Published: (2025)
by: Halder, Deepon, et al.
Published: (2025)
Can Large Language Models Code Like a Linguist?: A Case Study in Low Resource Sound Law Induction
by: Naik, Atharva, et al.
Published: (2024)
by: Naik, Atharva, et al.
Published: (2024)
PrahokBART: A Pre-trained Sequence-to-Sequence Model for Khmer Natural Language Generation
by: Kaing, Hour, et al.
Published: (2025)
by: Kaing, Hour, et al.
Published: (2025)
Pralekha: Cross-Lingual Document Alignment for Indic Languages
by: Suryanarayanan, Sanjay, et al.
Published: (2024)
by: Suryanarayanan, Sanjay, et al.
Published: (2024)
IndicIFEval: A Benchmark for Verifiable Instruction-Following Evaluation in 14 Indic Languages
by: Jayakumar, Thanmay, et al.
Published: (2026)
by: Jayakumar, Thanmay, et al.
Published: (2026)
On-Device LLMs for Home Assistant: Dual Role in Intent Detection and Response Generation
by: Birkmose, Rune, et al.
Published: (2025)
by: Birkmose, Rune, et al.
Published: (2025)
Similar Items
-
MultiHal: Multilingual Dataset for Knowledge-Graph Grounded Evaluation of LLM Hallucinations
by: Lavrinovics, Ernests, et al.
Published: (2025) -
Linguistically Grounded Analysis of Language Models using Shapley Head Values
by: Fekete, Marcell, et al.
Published: (2024) -
Knowledge Graphs, Large Language Models, and Hallucinations: An NLP Perspective
by: Lavrinovics, Ernests, et al.
Published: (2024) -
Text Embedding Inversion Security for Multilingual Language Models
by: Chen, Yiyi, et al.
Published: (2024) -
When Discourse Pressures Conflict: Information Structure in Vision-Language Model Outputs
by: Fekete, Marcell, et al.
Published: (2026)