Geographic Adaptation of Pretrained Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Hofmann, Valentin, Glavaš, Goran, Ljubešić, Nikola, Pierrehumbert, Janet B., Schütze, Hinrich |
|---|---|
| Format: | Preprint |
| Publié: |
2022
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Derivational Morphology Reveals Analogical Generalization in Large Language Models
par: Hofmann, Valentin, et autres
Publié: (2024)
par: Hofmann, Valentin, et autres
Publié: (2024)
One Script Instead of Hundreds? On Pretraining Romanized Encoder Language Models
par: Ebing, Benedikt, et autres
Publié: (2026)
par: Ebing, Benedikt, et autres
Publié: (2026)
CLASSLA-web: Comparable Web Corpora of South Slavic Languages Enriched with Linguistic and Genre Annotation
par: Ljubešić, Nikola, et autres
Publié: (2024)
par: Ljubešić, Nikola, et autres
Publié: (2024)
Language Models on a Diet: Cost-Efficient Development of Encoders for Closely-Related Languages via Additional Pretraining
par: Ljubešić, Nikola, et autres
Publié: (2024)
par: Ljubešić, Nikola, et autres
Publié: (2024)
TransliCo: A Contrastive Learning Framework to Address the Script Barrier in Multilingual Pretrained Language Models
par: Liu, Yihong, et autres
Publié: (2024)
par: Liu, Yihong, et autres
Publié: (2024)
To Translate or Not to Translate: A Systematic Investigation of Translation-Based Cross-Lingual Transfer to Low-Resource Languages
par: Ebing, Benedikt, et autres
Publié: (2023)
par: Ebing, Benedikt, et autres
Publié: (2023)
TransMI: A Framework to Create Strong Baselines from Multilingual Pretrained Language Models for Transliterated Data
par: Liu, Yihong, et autres
Publié: (2024)
par: Liu, Yihong, et autres
Publié: (2024)
LLM Teacher-Student Framework for Text Classification With No Manually Annotated Data: A Case Study in IPTC News Topic Classification
par: Kuzman, Taja, et autres
Publié: (2024)
par: Kuzman, Taja, et autres
Publié: (2024)
LangSAMP: Language-Script Aware Multilingual Pretraining
par: Liu, Yihong, et autres
Publié: (2024)
par: Liu, Yihong, et autres
Publié: (2024)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
par: Paul, Indraneil, et autres
Publié: (2024)
par: Paul, Indraneil, et autres
Publié: (2024)
GLUScope: A Tool for Analyzing GLU Neurons in Transformer Language Models
par: Gerstner, Sebastian, et autres
Publié: (2026)
par: Gerstner, Sebastian, et autres
Publié: (2026)
MaLA-500: Massive Language Adaptation of Large Language Models
par: Lin, Peiqin, et autres
Publié: (2024)
par: Lin, Peiqin, et autres
Publié: (2024)
mPLM-Sim: Better Cross-Lingual Similarity and Transfer in Multilingual Pretrained Language Models
par: Lin, Peiqin, et autres
Publié: (2023)
par: Lin, Peiqin, et autres
Publié: (2023)
Probing Large Language Models for Scalar Adjective Lexical Semantics and Scalar Diversity Pragmatics
par: Lin, Fangru, et autres
Publié: (2024)
par: Lin, Fangru, et autres
Publié: (2024)
Identifying Primary Stress Across Related Languages and Dialects with Transformer-based Speech Encoder Models
par: Ljubešić, Nikola, et autres
Publié: (2025)
par: Ljubešić, Nikola, et autres
Publié: (2025)
Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
par: Geigle, Gregor, et autres
Publié: (2024)
par: Geigle, Gregor, et autres
Publié: (2024)
The Devil Is in the Word Alignment Details: On Translation-Based Cross-Lingual Transfer for Token Classification Tasks
par: Ebing, Benedikt, et autres
Publié: (2025)
par: Ebing, Benedikt, et autres
Publié: (2025)
Your Pretrained Model Tells the Difficulty Itself: A Self-Adaptive Curriculum Learning Paradigm for Natural Language Understanding
par: Feng, Qi, et autres
Publié: (2025)
par: Feng, Qi, et autres
Publié: (2025)
Mechanistic Understanding and Mitigation of Language Confusion in English-Centric Large Language Models
par: Nie, Ercong, et autres
Publié: (2025)
par: Nie, Ercong, et autres
Publié: (2025)
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
par: Geigle, Gregor, et autres
Publié: (2024)
par: Geigle, Gregor, et autres
Publié: (2024)
Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models
par: Liu, Yihong, et autres
Publié: (2026)
par: Liu, Yihong, et autres
Publié: (2026)
Decoding Climate Disagreement: A Graph Neural Network-Based Approach to Understanding Social Media Dynamics
par: Su, Ruiran, et autres
Publié: (2024)
par: Su, Ruiran, et autres
Publié: (2024)
OFA: A Framework of Initializing Unseen Subword Embeddings for Efficient Large-scale Multilingual Continued Pretraining
par: Liu, Yihong, et autres
Publié: (2023)
par: Liu, Yihong, et autres
Publié: (2023)
The ParlaSent Multilingual Training Dataset for Sentiment Identification in Parliamentary Proceedings
par: Mochtak, Michal, et autres
Publié: (2023)
par: Mochtak, Michal, et autres
Publié: (2023)
Breaking the Script Barrier in Multilingual Pre-Trained Language Models with Transliteration-Based Post-Training Alignment
par: Xhelili, Orgest, et autres
Publié: (2024)
par: Xhelili, Orgest, et autres
Publié: (2024)
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
par: Geigle, Gregor, et autres
Publié: (2023)
par: Geigle, Gregor, et autres
Publié: (2023)
NLNDE at SemEval-2023 Task 12: Adaptive Pretraining and Source Language Selection for Low-Resource Multilingual Sentiment Analysis
par: Wang, Mingyang, et autres
Publié: (2023)
par: Wang, Mingyang, et autres
Publié: (2023)
Can Large Language Models Generalize Procedures Across Representations?
par: Lin, Fangru, et autres
Publié: (2026)
par: Lin, Fangru, et autres
Publié: (2026)
A Recipe of Parallel Corpora Exploitation for Multilingual Large Language Models
par: Lin, Peiqin, et autres
Publié: (2024)
par: Lin, Peiqin, et autres
Publié: (2024)
Graph-enhanced Large Language Models in Asynchronous Plan Reasoning
par: Lin, Fangru, et autres
Publié: (2024)
par: Lin, Fangru, et autres
Publié: (2024)
Self-Distillation for Model Stacking Unlocks Cross-Lingual NLU in 200+ Languages
par: Schmidt, Fabian David, et autres
Publié: (2024)
par: Schmidt, Fabian David, et autres
Publié: (2024)
SpeechTaxi: On Multilingual Semantic Speech Classification
par: Keller, Lennart, et autres
Publié: (2024)
par: Keller, Lennart, et autres
Publié: (2024)
Knowledge Distillation vs. Pretraining from Scratch under a Fixed (Computation) Budget
par: Bui, Minh Duc, et autres
Publié: (2024)
par: Bui, Minh Duc, et autres
Publié: (2024)
GKnow: Measuring the Entanglement of Gender Bias and Factual Gender
par: Veloso, Leonor, et autres
Publié: (2026)
par: Veloso, Leonor, et autres
Publié: (2026)
Modular Sentence Encoders: Separating Language Specialization from Cross-Lingual Alignment
par: Huang, Yongxin, et autres
Publié: (2024)
par: Huang, Yongxin, et autres
Publié: (2024)
TransAlign: Machine Translation Encoders are Strong Word Aligners, Too
par: Ebing, Benedikt, et autres
Publié: (2025)
par: Ebing, Benedikt, et autres
Publié: (2025)
Supervised In-Context Fine-Tuning for Generative Sequence Labeling
par: Dukić, David, et autres
Publié: (2025)
par: Dukić, David, et autres
Publié: (2025)
ReCoVeR the Target Language: Language Steering without Sacrificing Task Performance
par: Sterz, Hannah, et autres
Publié: (2025)
par: Sterz, Hannah, et autres
Publié: (2025)
Mići Princ -- A Little Boy Teaching Speech Technologies the Chakavian Dialect
par: Ljubešić, Nikola, et autres
Publié: (2026)
par: Ljubešić, Nikola, et autres
Publié: (2026)
RET-LLM: Towards a General Read-Write Memory for Large Language Models
par: Modarressi, Ali, et autres
Publié: (2023)
par: Modarressi, Ali, et autres
Publié: (2023)
Documents similaires
-
Derivational Morphology Reveals Analogical Generalization in Large Language Models
par: Hofmann, Valentin, et autres
Publié: (2024) -
One Script Instead of Hundreds? On Pretraining Romanized Encoder Language Models
par: Ebing, Benedikt, et autres
Publié: (2026) -
CLASSLA-web: Comparable Web Corpora of South Slavic Languages Enriched with Linguistic and Genre Annotation
par: Ljubešić, Nikola, et autres
Publié: (2024) -
Language Models on a Diet: Cost-Efficient Development of Encoders for Closely-Related Languages via Additional Pretraining
par: Ljubešić, Nikola, et autres
Publié: (2024) -
TransliCo: A Contrastive Learning Framework to Address the Script Barrier in Multilingual Pretrained Language Models
par: Liu, Yihong, et autres
Publié: (2024)