TabiBERT: A Large-Scale ModernBERT Foundation Model and A Unified Benchmark for Turkish
Fuente:
arXiv
Saved in:
| Main Authors: | Türker, Melikşah, Kızıloğlu, A. Ebrar, Güngör, Onur, Üsküdarlı, Susan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TURNA: A Turkish Encoder-Decoder Language Model for Enhanced Understanding and Generation
by: Uludoğan, Gökçe, et al.
Published: (2024)
by: Uludoğan, Gökçe, et al.
Published: (2024)
Patent Language Model Pretraining with ModernBERT
by: Yousefiramandi, Amirhossein, et al.
Published: (2025)
by: Yousefiramandi, Amirhossein, et al.
Published: (2025)
Chinese ModernBERT with Whole-Word Masking
by: Zhao, Zeyu, et al.
Published: (2025)
by: Zhao, Zeyu, et al.
Published: (2025)
ModernBERT + ColBERT: Enhancing biomedical RAG through an advanced re-ranking retriever
by: Rivera, Eduardo Martínez, et al.
Published: (2025)
by: Rivera, Eduardo Martínez, et al.
Published: (2025)
ModernBERT is More Efficient than Conventional BERT for Chest CT Findings Classification in Japanese Radiology Reports
by: Yamagishi, Yosuke, et al.
Published: (2025)
by: Yamagishi, Yosuke, et al.
Published: (2025)
A Diversity Diet for a Healthier Model: A Case Study of French ModernBERT
by: Estève, Louis, et al.
Published: (2026)
by: Estève, Louis, et al.
Published: (2026)
llm-jp-modernbert: A ModernBERT Model Trained on a Large-Scale Japanese Corpus with Long Context Length
by: Sugiura, Issa, et al.
Published: (2025)
by: Sugiura, Issa, et al.
Published: (2025)
Clinical ModernBERT: An efficient and long context encoder for biomedical text
by: Lee, Simon A., et al.
Published: (2025)
by: Lee, Simon A., et al.
Published: (2025)
NorBERTo: A ModernBERT Model Trained for Portuguese with 331 Billion Tokens Corpus
by: Silva, Enzo S. N., et al.
Published: (2026)
by: Silva, Enzo S. N., et al.
Published: (2026)
Spatial ModernBERT: Spatial-Aware Transformer for Table and Key-Value Extraction in Financial Documents at Scale
by: Javis AI Team, et al.
Published: (2025)
by: Javis AI Team, et al.
Published: (2025)
VBART: The Turkish LLM
by: Turker, Meliksah, et al.
Published: (2024)
by: Turker, Meliksah, et al.
Published: (2024)
VNLP: Turkish NLP Package
by: Turker, Meliksah, et al.
Published: (2024)
by: Turker, Meliksah, et al.
Published: (2024)
ModernBERT or DeBERTaV3? Examining Architecture and Data Influence on Transformer Encoder Models Performance
by: Antoun, Wissam, et al.
Published: (2025)
by: Antoun, Wissam, et al.
Published: (2025)
BioClinical ModernBERT: A State-of-the-Art Long-Context Encoder for Biomedical and Clinical NLP
by: Sounack, Thomas, et al.
Published: (2025)
by: Sounack, Thomas, et al.
Published: (2025)
Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction
by: Saoud, Abdulkader, et al.
Published: (2024)
by: Saoud, Abdulkader, et al.
Published: (2024)
Dependency Annotation of Ottoman Turkish with Multilingual BERT
by: Özateş, Şaziye Betül, et al.
Published: (2024)
by: Özateş, Şaziye Betül, et al.
Published: (2024)
TurkBench: A Benchmark for Evaluating Turkish Large Language Models
by: Toraman, Çağrı, et al.
Published: (2026)
by: Toraman, Çağrı, et al.
Published: (2026)
SindBERT, the Sailor: Charting the Seas of Turkish NLP
by: Schmitt, Raphael, et al.
Published: (2025)
by: Schmitt, Raphael, et al.
Published: (2025)
Pretraining Finnish ModernBERTs
by: Reunamo, Akseli, et al.
Published: (2025)
by: Reunamo, Akseli, et al.
Published: (2025)
TurkColBERT: A Benchmark of Dense and Late-Interaction Models for Turkish Information Retrieval
by: Ezerceli, Özay, et al.
Published: (2025)
by: Ezerceli, Özay, et al.
Published: (2025)
NeoBERT: A Next-Generation BERT
by: Breton, Lola Le, et al.
Published: (2025)
by: Breton, Lola Le, et al.
Published: (2025)
FaBERT: Pre-training BERT on Persian Blogs
by: Masumi, Mostafa, et al.
Published: (2024)
by: Masumi, Mostafa, et al.
Published: (2024)
NusaBERT: Teaching IndoBERT to be Multilingual and Multicultural
by: Wongso, Wilson, et al.
Published: (2024)
by: Wongso, Wilson, et al.
Published: (2024)
SpikeBERT: A Language Spikformer Learned from BERT with Knowledge Distillation
by: Lv, Changze, et al.
Published: (2023)
by: Lv, Changze, et al.
Published: (2023)
m3BERT: A Modern, Multi-lingual, Matryoshka Bidirectional Encoder
by: Wang, Yaoxiang, et al.
Published: (2026)
by: Wang, Yaoxiang, et al.
Published: (2026)
Scaling HuBERT for African Languages: From Base to Large and XL
by: Caubrière, Antoine, et al.
Published: (2025)
by: Caubrière, Antoine, et al.
Published: (2025)
NeoDictaBERT: Pushing the Frontier of BERT models for Hebrew
by: Shmidman, Shaltiel, et al.
Published: (2025)
by: Shmidman, Shaltiel, et al.
Published: (2025)
KliniskVestBERT: BERT Model Specialised to Norwegian Clinical Texts
by: Autenried, Christian, et al.
Published: (2026)
by: Autenried, Christian, et al.
Published: (2026)
mHuBERT-147: A Compact Multilingual HuBERT Model
by: Boito, Marcely Zanon, et al.
Published: (2024)
by: Boito, Marcely Zanon, et al.
Published: (2024)
QiBERT -- Classifying Online Conversations Messages with BERT as a Feature
by: Ferreira-Saraiva, Bruno D., et al.
Published: (2024)
by: Ferreira-Saraiva, Bruno D., et al.
Published: (2024)
Pre-training technique to localize medical BERT and enhance biomedical BERT
by: Wada, Shoya, et al.
Published: (2020)
by: Wada, Shoya, et al.
Published: (2020)
KuBERT: Central Kurdish BERT Model and Its Application for Sentiment Analysis
by: Awlla, Kozhin muhealddin, et al.
Published: (2025)
by: Awlla, Kozhin muhealddin, et al.
Published: (2025)
AraModernBERT: Transtokenized Initialization and Long-Context Encoder Modeling for Arabic
by: Elshehy, Omar, et al.
Published: (2026)
by: Elshehy, Omar, et al.
Published: (2026)
Analysis of Argument Structure Constructions in the Large Language Model BERT
by: Ramezani, Pegah, et al.
Published: (2024)
by: Ramezani, Pegah, et al.
Published: (2024)
An investigation of structures responsible for gender bias in BERT and DistilBERT
by: Leteno, Thibaud, et al.
Published: (2024)
by: Leteno, Thibaud, et al.
Published: (2024)
EgyBERT: A Large Language Model Pretrained on Egyptian Dialect Corpora
by: Qarah, Faisal
Published: (2024)
by: Qarah, Faisal
Published: (2024)
mmBERT: A Modern Multilingual Encoder with Annealed Language Learning
by: Marone, Marc, et al.
Published: (2025)
by: Marone, Marc, et al.
Published: (2025)
ConfliBERT: A Language Model for Political Conflict
by: Brandt, Patrick T., et al.
Published: (2024)
by: Brandt, Patrick T., et al.
Published: (2024)
Leveraging IndoBERT and DistilBERT for Indonesian Emotion Classification in E-Commerce Reviews
by: Christian, William, et al.
Published: (2025)
by: Christian, William, et al.
Published: (2025)
ACL: Aligned Contrastive Learning Improves BERT and Multi-exit BERT Fine-tuning
by: Li, Liz, et al.
Published: (2026)
by: Li, Liz, et al.
Published: (2026)
Similar Items
-
TURNA: A Turkish Encoder-Decoder Language Model for Enhanced Understanding and Generation
by: Uludoğan, Gökçe, et al.
Published: (2024) -
Patent Language Model Pretraining with ModernBERT
by: Yousefiramandi, Amirhossein, et al.
Published: (2025) -
Chinese ModernBERT with Whole-Word Masking
by: Zhao, Zeyu, et al.
Published: (2025) -
ModernBERT + ColBERT: Enhancing biomedical RAG through an advanced re-ranking retriever
by: Rivera, Eduardo Martínez, et al.
Published: (2025) -
ModernBERT is More Efficient than Conventional BERT for Chest CT Findings Classification in Japanese Radiology Reports
by: Yamagishi, Yosuke, et al.
Published: (2025)