Efficient Language Modeling for Low-Resource Settings with Hybrid RNN-Transformer Architectures
Fuente:
arXiv
Saved in:
| Main Authors: | Lindenmaier, Gabriel, Papay, Sean, Padó, Sebastian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Actor Identification in Discourse: A Challenge for LLMs?
by: Barić, Ana, et al.
Published: (2024)
by: Barić, Ana, et al.
Published: (2024)
Regular-pattern-sensitive CRFs for Distant Label Interactions
by: Papay, Sean, et al.
Published: (2024)
by: Papay, Sean, et al.
Published: (2024)
Multi-Dimensional Machine Translation Evaluation: Model Evaluation and Resource for Korean
by: Park, Dojun, et al.
Published: (2024)
by: Park, Dojun, et al.
Published: (2024)
Do Language Models Encode Knowledge of Linguistic Constraint Violations?
by: Hardy, et al.
Published: (2026)
by: Hardy, et al.
Published: (2026)
Are Humans as Brittle as Large Language Models?
by: Li, Jiahui, et al.
Published: (2025)
by: Li, Jiahui, et al.
Published: (2025)
Finding Sense in Nonsense with Generated Contexts: Perspectives from Humans and Language Models
by: Olsen, Katrina, et al.
Published: (2026)
by: Olsen, Katrina, et al.
Published: (2026)
Artwork Interpretation with Vision Language Models: A Case Study on Emotions and Emotion Symbols
by: Padó, Sebastian, et al.
Published: (2025)
by: Padó, Sebastian, et al.
Published: (2025)
Strategies for political-statement segmentation and labelling in unstructured text
by: Nikolaev, Dmitry, et al.
Published: (2025)
by: Nikolaev, Dmitry, et al.
Published: (2025)
Diverging Transformer Predictions for Human Sentence Processing: A Comprehensive Analysis of Agreement Attraction Effects
by: von der Malsburg, Titus, et al.
Published: (2026)
by: von der Malsburg, Titus, et al.
Published: (2026)
Approximate Attributions for Off-the-Shelf Siamese Transformers
by: Möller, Lucas, et al.
Published: (2024)
by: Möller, Lucas, et al.
Published: (2024)
Modeling Bilingual Sentence Processing: Evaluating RNN and Transformer Architectures for Cross-Language Structural Priming
by: Zhang, Demi, et al.
Published: (2024)
by: Zhang, Demi, et al.
Published: (2024)
Towards Understanding the Relationship between In-context Learning and Compositional Generalization
by: Han, Sungjun, et al.
Published: (2024)
by: Han, Sungjun, et al.
Published: (2024)
Efficient Extractive Summarization with MAMBA-Transformer Hybrids for Low-Resource Scenarios
by: Khayi, Nisrine Ait
Published: (2026)
by: Khayi, Nisrine Ait
Published: (2026)
Do Political Opinions Transfer Between Western Languages? An Analysis of Unaligned and Aligned Multilingual LLMs
by: Weeber, Franziska, et al.
Published: (2025)
by: Weeber, Franziska, et al.
Published: (2025)
iPOE: Interpretable Prompt Optimization via Explanations
by: Li, Jiahui, et al.
Published: (2026)
by: Li, Jiahui, et al.
Published: (2026)
Generalizability of Media Frames: Corpus creation and analysis across countries
by: Daffara, Agnese, et al.
Published: (2025)
by: Daffara, Agnese, et al.
Published: (2025)
Toeing the Party Line: Election Manifestos as a Key to Understand Political Discourse on Twitter
by: Maurer, Maximilian, et al.
Published: (2024)
by: Maurer, Maximilian, et al.
Published: (2024)
One Persona, Many Cues, Different Results: How Sociodemographic Cues Impact LLM Personalization
by: Weeber, Franziska, et al.
Published: (2026)
by: Weeber, Franziska, et al.
Published: (2026)
Beyond Marginal Distributions: A Framework to Evaluate the Representativeness of Demographic-Aligned LLMs
by: Williams, Tristan, et al.
Published: (2026)
by: Williams, Tristan, et al.
Published: (2026)
ARWKV: Pretrain is not what we need, an RNN-Attention-Based Language Model Born from Transformer
by: Yueyu, Lin, et al.
Published: (2025)
by: Yueyu, Lin, et al.
Published: (2025)
Unified Large Language Models for Misinformation Detection in Low-Resource Linguistic Settings
by: Islam, Muhammad, et al.
Published: (2025)
by: Islam, Muhammad, et al.
Published: (2025)
Hybrid and Unitary PEFT for Resource-Efficient Large Language Models
by: Qi, Haomin, et al.
Published: (2025)
by: Qi, Haomin, et al.
Published: (2025)
ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting
by: Fields, Clayton, et al.
Published: (2026)
by: Fields, Clayton, et al.
Published: (2026)
Explaining Pre-Trained Language Models with Attribution Scores: An Analysis in Low-Resource Settings
by: Zhou, Wei, et al.
Published: (2024)
by: Zhou, Wei, et al.
Published: (2024)
Explaining Caption-Image Interactions in CLIP Models with Second-Order Attributions
by: Möller, Lucas, et al.
Published: (2024)
by: Möller, Lucas, et al.
Published: (2024)
Comparative Analysis of Different Efficient Fine Tuning Methods of Large Language Models (LLMs) in Low-Resource Setting
by: Srinivasan, Krishna Prasad Varadarajan, et al.
Published: (2024)
by: Srinivasan, Krishna Prasad Varadarajan, et al.
Published: (2024)
A*-Thought: Efficient Reasoning via Bidirectional Compression for Low-Resource Settings
by: Xu, Xiaoang, et al.
Published: (2025)
by: Xu, Xiaoang, et al.
Published: (2025)
AutoMeet: a proof-of-concept study of genAI to automate meetings in automotive engineering
by: Baeuerle, Simon, et al.
Published: (2025)
by: Baeuerle, Simon, et al.
Published: (2025)
Subgroups of $U(d)$ Induce Natural RNN and Transformer Architectures
by: Nunley, Joshua
Published: (2026)
by: Nunley, Joshua
Published: (2026)
Beyond prompt brittleness: Evaluating the reliability and consistency of political worldviews in LLMs
by: Ceron, Tanise, et al.
Published: (2024)
by: Ceron, Tanise, et al.
Published: (2024)
Transformers for Low-Resource Languages: Is Féidir Linn!
by: Lankford, Séamus, et al.
Published: (2024)
by: Lankford, Séamus, et al.
Published: (2024)
LowCLIP: Adapting the CLIP Model Architecture for Low-Resource Languages in Multimodal Image Retrieval Task
by: Asgarov, Ali, et al.
Published: (2024)
by: Asgarov, Ali, et al.
Published: (2024)
Computational Analysis of Gender Depiction in the Comedias of Calderón de la Barca
by: Keith, Allison, et al.
Published: (2024)
by: Keith, Allison, et al.
Published: (2024)
Leveraging Digitized Newspapers to Collect Summarization Data in Low-Resource Languages
by: Dahan, Noam, et al.
Published: (2025)
by: Dahan, Noam, et al.
Published: (2025)
LakotaBERT: A Transformer-based Model for Low Resource Lakota Language
by: Parankusham, Kanishka, et al.
Published: (2025)
by: Parankusham, Kanishka, et al.
Published: (2025)
Synthetic Data Generation in Low-Resource Settings via Fine-Tuning of Large Language Models
by: Kaddour, Jean, et al.
Published: (2023)
by: Kaddour, Jean, et al.
Published: (2023)
Exploring NLP Benchmarks in an Extremely Low-Resource Setting
by: Nuha, Ulin, et al.
Published: (2025)
by: Nuha, Ulin, et al.
Published: (2025)
Irish-BLiMP: A Linguistic Benchmark for Evaluating Human and Language Model Performance in a Low-Resource Setting
by: McGiff, Josh, et al.
Published: (2025)
by: McGiff, Josh, et al.
Published: (2025)
Improved Visually Prompted Keyword Localisation in Real Low-Resource Settings
by: Nortje, Leanne, et al.
Published: (2024)
by: Nortje, Leanne, et al.
Published: (2024)
BabyLMs for isiXhosa: Data-Efficient Language Modelling in a Low-Resource Context
by: Matzopoulos, Alexis, et al.
Published: (2025)
by: Matzopoulos, Alexis, et al.
Published: (2025)
Similar Items
-
Actor Identification in Discourse: A Challenge for LLMs?
by: Barić, Ana, et al.
Published: (2024) -
Regular-pattern-sensitive CRFs for Distant Label Interactions
by: Papay, Sean, et al.
Published: (2024) -
Multi-Dimensional Machine Translation Evaluation: Model Evaluation and Resource for Korean
by: Park, Dojun, et al.
Published: (2024) -
Do Language Models Encode Knowledge of Linguistic Constraint Violations?
by: Hardy, et al.
Published: (2026) -
Are Humans as Brittle as Large Language Models?
by: Li, Jiahui, et al.
Published: (2025)