Geographic Adaptation of Pretrained Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hofmann, Valentin, Glavaš, Goran, Ljubešić, Nikola, Pierrehumbert, Janet B., Schütze, Hinrich |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Derivational Morphology Reveals Analogical Generalization in Large Language Models
von: Hofmann, Valentin, et al.
Veröffentlicht: (2024)
von: Hofmann, Valentin, et al.
Veröffentlicht: (2024)
One Script Instead of Hundreds? On Pretraining Romanized Encoder Language Models
von: Ebing, Benedikt, et al.
Veröffentlicht: (2026)
von: Ebing, Benedikt, et al.
Veröffentlicht: (2026)
CLASSLA-web: Comparable Web Corpora of South Slavic Languages Enriched with Linguistic and Genre Annotation
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2024)
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2024)
Language Models on a Diet: Cost-Efficient Development of Encoders for Closely-Related Languages via Additional Pretraining
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2024)
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2024)
TransliCo: A Contrastive Learning Framework to Address the Script Barrier in Multilingual Pretrained Language Models
von: Liu, Yihong, et al.
Veröffentlicht: (2024)
von: Liu, Yihong, et al.
Veröffentlicht: (2024)
To Translate or Not to Translate: A Systematic Investigation of Translation-Based Cross-Lingual Transfer to Low-Resource Languages
von: Ebing, Benedikt, et al.
Veröffentlicht: (2023)
von: Ebing, Benedikt, et al.
Veröffentlicht: (2023)
TransMI: A Framework to Create Strong Baselines from Multilingual Pretrained Language Models for Transliterated Data
von: Liu, Yihong, et al.
Veröffentlicht: (2024)
von: Liu, Yihong, et al.
Veröffentlicht: (2024)
LLM Teacher-Student Framework for Text Classification With No Manually Annotated Data: A Case Study in IPTC News Topic Classification
von: Kuzman, Taja, et al.
Veröffentlicht: (2024)
von: Kuzman, Taja, et al.
Veröffentlicht: (2024)
LangSAMP: Language-Script Aware Multilingual Pretraining
von: Liu, Yihong, et al.
Veröffentlicht: (2024)
von: Liu, Yihong, et al.
Veröffentlicht: (2024)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
GLUScope: A Tool for Analyzing GLU Neurons in Transformer Language Models
von: Gerstner, Sebastian, et al.
Veröffentlicht: (2026)
von: Gerstner, Sebastian, et al.
Veröffentlicht: (2026)
MaLA-500: Massive Language Adaptation of Large Language Models
von: Lin, Peiqin, et al.
Veröffentlicht: (2024)
von: Lin, Peiqin, et al.
Veröffentlicht: (2024)
mPLM-Sim: Better Cross-Lingual Similarity and Transfer in Multilingual Pretrained Language Models
von: Lin, Peiqin, et al.
Veröffentlicht: (2023)
von: Lin, Peiqin, et al.
Veröffentlicht: (2023)
Probing Large Language Models for Scalar Adjective Lexical Semantics and Scalar Diversity Pragmatics
von: Lin, Fangru, et al.
Veröffentlicht: (2024)
von: Lin, Fangru, et al.
Veröffentlicht: (2024)
Identifying Primary Stress Across Related Languages and Dialects with Transformer-based Speech Encoder Models
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2025)
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2025)
Does Object Grounding Really Reduce Hallucination of Large Vision-Language Models?
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
The Devil Is in the Word Alignment Details: On Translation-Based Cross-Lingual Transfer for Token Classification Tasks
von: Ebing, Benedikt, et al.
Veröffentlicht: (2025)
von: Ebing, Benedikt, et al.
Veröffentlicht: (2025)
Your Pretrained Model Tells the Difficulty Itself: A Self-Adaptive Curriculum Learning Paradigm for Natural Language Understanding
von: Feng, Qi, et al.
Veröffentlicht: (2025)
von: Feng, Qi, et al.
Veröffentlicht: (2025)
Mechanistic Understanding and Mitigation of Language Confusion in English-Centric Large Language Models
von: Nie, Ercong, et al.
Veröffentlicht: (2025)
von: Nie, Ercong, et al.
Veröffentlicht: (2025)
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
von: Geigle, Gregor, et al.
Veröffentlicht: (2024)
Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models
von: Liu, Yihong, et al.
Veröffentlicht: (2026)
von: Liu, Yihong, et al.
Veröffentlicht: (2026)
Decoding Climate Disagreement: A Graph Neural Network-Based Approach to Understanding Social Media Dynamics
von: Su, Ruiran, et al.
Veröffentlicht: (2024)
von: Su, Ruiran, et al.
Veröffentlicht: (2024)
OFA: A Framework of Initializing Unseen Subword Embeddings for Efficient Large-scale Multilingual Continued Pretraining
von: Liu, Yihong, et al.
Veröffentlicht: (2023)
von: Liu, Yihong, et al.
Veröffentlicht: (2023)
The ParlaSent Multilingual Training Dataset for Sentiment Identification in Parliamentary Proceedings
von: Mochtak, Michal, et al.
Veröffentlicht: (2023)
von: Mochtak, Michal, et al.
Veröffentlicht: (2023)
Breaking the Script Barrier in Multilingual Pre-Trained Language Models with Transliteration-Based Post-Training Alignment
von: Xhelili, Orgest, et al.
Veröffentlicht: (2024)
von: Xhelili, Orgest, et al.
Veröffentlicht: (2024)
Babel-ImageNet: Massively Multilingual Evaluation of Vision-and-Language Representations
von: Geigle, Gregor, et al.
Veröffentlicht: (2023)
von: Geigle, Gregor, et al.
Veröffentlicht: (2023)
NLNDE at SemEval-2023 Task 12: Adaptive Pretraining and Source Language Selection for Low-Resource Multilingual Sentiment Analysis
von: Wang, Mingyang, et al.
Veröffentlicht: (2023)
von: Wang, Mingyang, et al.
Veröffentlicht: (2023)
Can Large Language Models Generalize Procedures Across Representations?
von: Lin, Fangru, et al.
Veröffentlicht: (2026)
von: Lin, Fangru, et al.
Veröffentlicht: (2026)
A Recipe of Parallel Corpora Exploitation for Multilingual Large Language Models
von: Lin, Peiqin, et al.
Veröffentlicht: (2024)
von: Lin, Peiqin, et al.
Veröffentlicht: (2024)
Graph-enhanced Large Language Models in Asynchronous Plan Reasoning
von: Lin, Fangru, et al.
Veröffentlicht: (2024)
von: Lin, Fangru, et al.
Veröffentlicht: (2024)
Self-Distillation for Model Stacking Unlocks Cross-Lingual NLU in 200+ Languages
von: Schmidt, Fabian David, et al.
Veröffentlicht: (2024)
von: Schmidt, Fabian David, et al.
Veröffentlicht: (2024)
SpeechTaxi: On Multilingual Semantic Speech Classification
von: Keller, Lennart, et al.
Veröffentlicht: (2024)
von: Keller, Lennart, et al.
Veröffentlicht: (2024)
Knowledge Distillation vs. Pretraining from Scratch under a Fixed (Computation) Budget
von: Bui, Minh Duc, et al.
Veröffentlicht: (2024)
von: Bui, Minh Duc, et al.
Veröffentlicht: (2024)
GKnow: Measuring the Entanglement of Gender Bias and Factual Gender
von: Veloso, Leonor, et al.
Veröffentlicht: (2026)
von: Veloso, Leonor, et al.
Veröffentlicht: (2026)
Modular Sentence Encoders: Separating Language Specialization from Cross-Lingual Alignment
von: Huang, Yongxin, et al.
Veröffentlicht: (2024)
von: Huang, Yongxin, et al.
Veröffentlicht: (2024)
TransAlign: Machine Translation Encoders are Strong Word Aligners, Too
von: Ebing, Benedikt, et al.
Veröffentlicht: (2025)
von: Ebing, Benedikt, et al.
Veröffentlicht: (2025)
Supervised In-Context Fine-Tuning for Generative Sequence Labeling
von: Dukić, David, et al.
Veröffentlicht: (2025)
von: Dukić, David, et al.
Veröffentlicht: (2025)
ReCoVeR the Target Language: Language Steering without Sacrificing Task Performance
von: Sterz, Hannah, et al.
Veröffentlicht: (2025)
von: Sterz, Hannah, et al.
Veröffentlicht: (2025)
Mići Princ -- A Little Boy Teaching Speech Technologies the Chakavian Dialect
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2026)
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2026)
RET-LLM: Towards a General Read-Write Memory for Large Language Models
von: Modarressi, Ali, et al.
Veröffentlicht: (2023)
von: Modarressi, Ali, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Derivational Morphology Reveals Analogical Generalization in Large Language Models
von: Hofmann, Valentin, et al.
Veröffentlicht: (2024) -
One Script Instead of Hundreds? On Pretraining Romanized Encoder Language Models
von: Ebing, Benedikt, et al.
Veröffentlicht: (2026) -
CLASSLA-web: Comparable Web Corpora of South Slavic Languages Enriched with Linguistic and Genre Annotation
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2024) -
Language Models on a Diet: Cost-Efficient Development of Encoders for Closely-Related Languages via Additional Pretraining
von: Ljubešić, Nikola, et al.
Veröffentlicht: (2024) -
TransliCo: A Contrastive Learning Framework to Address the Script Barrier in Multilingual Pretrained Language Models
von: Liu, Yihong, et al.
Veröffentlicht: (2024)