mALBERT: Is a Compact Multilingual BERT Model Still Worth It?
Fuente:
arXiv
Saved in:
| Main Authors: | Servan, Christophe, Ghannay, Sahar, Rosset, Sophie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Small Language Models are Good Too: An Empirical Study of Zero-Shot Classification
by: Lepagnol, Pierre, et al.
Published: (2024)
by: Lepagnol, Pierre, et al.
Published: (2024)
New Semantic Task for the French Spoken Language Understanding MEDIA Benchmark
by: Alavoine, Nadège, et al.
Published: (2024)
by: Alavoine, Nadège, et al.
Published: (2024)
Leveraging Information Retrieval to Enhance Spoken Language Understanding Prompts in Few-Shot Learning
by: Lepagnol, Pierre, et al.
Published: (2025)
by: Lepagnol, Pierre, et al.
Published: (2025)
LLM-based Atomic Propositions help weak extractors: Evaluation of a Propositioner for triplet extraction
by: Pommeret, Luc, et al.
Published: (2026)
by: Pommeret, Luc, et al.
Published: (2026)
A Benchmark Evaluation of Clinical Named Entity Recognition in French
by: Bannour, Nesrine, et al.
Published: (2024)
by: Bannour, Nesrine, et al.
Published: (2024)
Do BERT-Like Bidirectional Models Still Perform Better on Text Classification in the Era of LLMs?
by: Zhang, Junyan, et al.
Published: (2025)
by: Zhang, Junyan, et al.
Published: (2025)
EuroBERT: Scaling Multilingual Encoders for European Languages
by: Boizard, Nicolas, et al.
Published: (2025)
by: Boizard, Nicolas, et al.
Published: (2025)
Languages Still Left Behind: Toward a Better Multilingual Machine Translation Benchmark
by: Taguchi, Chihiro, et al.
Published: (2025)
by: Taguchi, Chihiro, et al.
Published: (2025)
BBPOS: BERT-based Part-of-Speech Tagging for Uzbek
by: Bobojonova, Latofat, et al.
Published: (2025)
by: Bobojonova, Latofat, et al.
Published: (2025)
Bug Destiny Prediction in Large Open-Source Software Repositories through Sentiment Analysis and BERT Topic Modeling
by: Pope, Sophie C., et al.
Published: (2025)
by: Pope, Sophie C., et al.
Published: (2025)
PolyBERT: Fine-Tuned Poly Encoder BERT-Based Model for Word Sense Disambiguation
by: Xia, Linhan, et al.
Published: (2025)
by: Xia, Linhan, et al.
Published: (2025)
MrBERT: Modern Multilingual Encoders via Vocabulary, Domain, and Dimensional Adaptation
by: Tamayo, Daniel, et al.
Published: (2026)
by: Tamayo, Daniel, et al.
Published: (2026)
KliniskVestBERT: BERT Model Specialised to Norwegian Clinical Texts
by: Autenried, Christian, et al.
Published: (2026)
by: Autenried, Christian, et al.
Published: (2026)
KuBERT: Central Kurdish BERT Model and Its Application for Sentiment Analysis
by: Awlla, Kozhin muhealddin, et al.
Published: (2025)
by: Awlla, Kozhin muhealddin, et al.
Published: (2025)
mHuBERT-147: A Compact Multilingual HuBERT Model
by: Boito, Marcely Zanon, et al.
Published: (2024)
by: Boito, Marcely Zanon, et al.
Published: (2024)
Is Peer-Reviewing Worth the Effort?
by: Church, Kenneth, et al.
Published: (2024)
by: Church, Kenneth, et al.
Published: (2024)
FLUX: Data Worth Training On
by: Gowtham, et al.
Published: (2026)
by: Gowtham, et al.
Published: (2026)
NeoBERT: A Next-Generation BERT
by: Breton, Lola Le, et al.
Published: (2025)
by: Breton, Lola Le, et al.
Published: (2025)
FRASIMED: a Clinical French Annotated Resource Produced through Crosslingual BERT-Based Annotation Projection
by: Zaghir, Jamil, et al.
Published: (2023)
by: Zaghir, Jamil, et al.
Published: (2023)
A Diversity Diet for a Healthier Model: A Case Study of French ModernBERT
by: Estève, Louis, et al.
Published: (2026)
by: Estève, Louis, et al.
Published: (2026)
Lessons Without Borders? Evaluating Cultural Alignment of LLMs Using Multilingual Story Moral Generation
by: Wu, Sophie, et al.
Published: (2026)
by: Wu, Sophie, et al.
Published: (2026)
BERT4FCA: A Method for Bipartite Link Prediction using Formal Concept Analysis and BERT
by: Peng, Siqi, et al.
Published: (2024)
by: Peng, Siqi, et al.
Published: (2024)
AI Hallucinations: A Misnomer Worth Clarifying
by: Maleki, Negar, et al.
Published: (2024)
by: Maleki, Negar, et al.
Published: (2024)
IM-BERT: Enhancing Robustness of BERT through the Implicit Euler Method
by: Kim, Mihyeon, et al.
Published: (2025)
by: Kim, Mihyeon, et al.
Published: (2025)
Analyzing Multi-Head Attention on Trojan BERT Models
by: Wang, Jingwei
Published: (2024)
by: Wang, Jingwei
Published: (2024)
How Much is Brain Data Worth for Machine Learning?
by: Lewis, Lane, et al.
Published: (2026)
by: Lewis, Lane, et al.
Published: (2026)
When Remembering and Planning are Worth it: Navigating under Change
by: Madani, Omid, et al.
Published: (2026)
by: Madani, Omid, et al.
Published: (2026)
LegalPro-BERT: Classification of Legal Provisions by fine-tuning BERT Large Language Model
by: Tewari, Amit
Published: (2024)
by: Tewari, Amit
Published: (2024)
mR3: Multilingual Rubric-Agnostic Reward Reasoning Models
by: Anugraha, David, et al.
Published: (2025)
by: Anugraha, David, et al.
Published: (2025)
Neural Architecture Search for Sentence Classification with BERT
by: Kenneweg, Philip, et al.
Published: (2024)
by: Kenneweg, Philip, et al.
Published: (2024)
Large Language Models Are Still Misled by Simple Bias Ensembles
by: Sun, Zhouhao, et al.
Published: (2025)
by: Sun, Zhouhao, et al.
Published: (2025)
RooseBERT: A New Deal For Political Language Modelling
by: Dore, Deborah, et al.
Published: (2025)
by: Dore, Deborah, et al.
Published: (2025)
Vision HgNN: An Electron-Micrograph is Worth Hypergraph of Hypernodes
by: Srinivas, Sakhinana Sagar, et al.
Published: (2024)
by: Srinivas, Sakhinana Sagar, et al.
Published: (2024)
Not All Personas Are Worth It: Culture-Reflective Persona Data Augmentation
by: Han, Ji-Eun, et al.
Published: (2025)
by: Han, Ji-Eun, et al.
Published: (2025)
Patent Language Model Pretraining with ModernBERT
by: Yousefiramandi, Amirhossein, et al.
Published: (2025)
by: Yousefiramandi, Amirhossein, et al.
Published: (2025)
Understanding Wikidata Qualifiers: An Analysis and Taxonomy
by: Falquet, Gilles, et al.
Published: (2026)
by: Falquet, Gilles, et al.
Published: (2026)
How Many Human Survey Respondents is a Large Language Model Worth? An Uncertainty Quantification Perspective
by: Huang, Chengpiao, et al.
Published: (2025)
by: Huang, Chengpiao, et al.
Published: (2025)
m3P: Towards Multimodal Multilingual Translation with Multimodal Prompt
by: Yang, Jian, et al.
Published: (2024)
by: Yang, Jian, et al.
Published: (2024)
A Time Series is Worth Five Experts: Heterogeneous Mixture of Experts for Traffic Flow Prediction
by: Wang, Guangyu, et al.
Published: (2024)
by: Wang, Guangyu, et al.
Published: (2024)
SMART: When is it Actually Worth Expanding a Speculative Tree?
by: Wang, Lifu, et al.
Published: (2026)
by: Wang, Lifu, et al.
Published: (2026)
Similar Items
-
Small Language Models are Good Too: An Empirical Study of Zero-Shot Classification
by: Lepagnol, Pierre, et al.
Published: (2024) -
New Semantic Task for the French Spoken Language Understanding MEDIA Benchmark
by: Alavoine, Nadège, et al.
Published: (2024) -
Leveraging Information Retrieval to Enhance Spoken Language Understanding Prompts in Few-Shot Learning
by: Lepagnol, Pierre, et al.
Published: (2025) -
LLM-based Atomic Propositions help weak extractors: Evaluation of a Propositioner for triplet extraction
by: Pommeret, Luc, et al.
Published: (2026) -
A Benchmark Evaluation of Clinical Named Entity Recognition in French
by: Bannour, Nesrine, et al.
Published: (2024)