Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction
Fuente:
arXiv
Saved in:
| Main Authors: | Saoud, Abdulkader, Alomeyr, Mahmut, Kesgin, Himmet Toprak, Amasyali, Mehmet Fatih |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Iterative Mask Filling: An Effective Text Augmentation Method Using Masked Language Modeling
by: Kesgin, Himmet Toprak, et al.
Published: (2024)
by: Kesgin, Himmet Toprak, et al.
Published: (2024)
Investigating Semi-Supervised Learning Algorithms in Text Datasets
by: Kesgin, Himmet Toprak, et al.
Published: (2024)
by: Kesgin, Himmet Toprak, et al.
Published: (2024)
Advancing NLP Models with Strategic Text Augmentation: A Comprehensive Study of Augmentation Methods and Curriculum Strategies
by: Kesgin, Himmet Toprak, et al.
Published: (2024)
by: Kesgin, Himmet Toprak, et al.
Published: (2024)
Optimizing Large Language Models for Turkish: New Methodologies in Corpus Selection and Training
by: Kesgin, H. Toprak, et al.
Published: (2024)
by: Kesgin, H. Toprak, et al.
Published: (2024)
Introducing cosmosGPT: Monolingual Training for Turkish Language Models
by: Kesgin, H. Toprak, et al.
Published: (2024)
by: Kesgin, H. Toprak, et al.
Published: (2024)
Türkçe Dil Modellerinin Performans Karşılaştırması Performance Comparison of Turkish Language Models
by: Dogan, Eren, et al.
Published: (2024)
by: Dogan, Eren, et al.
Published: (2024)
Cosmos-LLaVA: Chatting with the Visual Cosmos-LLaVA: Görselle Sohbet Etmek
by: Zeer, Ahmed, et al.
Published: (2024)
by: Zeer, Ahmed, et al.
Published: (2024)
From text to multimodal: a survey of adversarial example generation in question answering systems
by: Yigit, Gulsum, et al.
Published: (2023)
by: Yigit, Gulsum, et al.
Published: (2023)
Theoretical research on generative diffusion models: an overview
by: Yeğin, Melike Nur, et al.
Published: (2024)
by: Yeğin, Melike Nur, et al.
Published: (2024)
VBART: The Turkish LLM
by: Turker, Meliksah, et al.
Published: (2024)
by: Turker, Meliksah, et al.
Published: (2024)
VNLP: Turkish NLP Package
by: Turker, Meliksah, et al.
Published: (2024)
by: Turker, Meliksah, et al.
Published: (2024)
Bridging Robustness and Generalization Against Word Substitution Attacks in NLP via the Growth Bound Matrix Approach
by: Bouri, Mohammed, et al.
Published: (2025)
by: Bouri, Mohammed, et al.
Published: (2025)
A Comprehensive Approach to Misspelling Correction with BERT and Levenshtein Distance
by: Naziri, Amirreza, et al.
Published: (2024)
by: Naziri, Amirreza, et al.
Published: (2024)
PersianPunc: A Large-Scale Dataset and BERT-Based Approach for Persian Punctuation Restoration
by: Kalahroodi, Mohammad Javad Ranjbar, et al.
Published: (2026)
by: Kalahroodi, Mohammad Javad Ranjbar, et al.
Published: (2026)
Patent Language Model Pretraining with ModernBERT
by: Yousefiramandi, Amirhossein, et al.
Published: (2025)
by: Yousefiramandi, Amirhossein, et al.
Published: (2025)
MedicalBERT: enhancing biomedical natural language processing using pretrained BERT-based model
by: Reddy, K. Sahit, et al.
Published: (2025)
by: Reddy, K. Sahit, et al.
Published: (2025)
LegalTurk Optimized BERT for Multi-Label Text Classification and NER
by: Zeidi, Farnaz, et al.
Published: (2024)
by: Zeidi, Farnaz, et al.
Published: (2024)
Toward Understanding BERT-Like Pre-Training for DNA Foundation Models
by: Liang, Chaoqi, et al.
Published: (2023)
by: Liang, Chaoqi, et al.
Published: (2023)
TURNA: A Turkish Encoder-Decoder Language Model for Enhanced Understanding and Generation
by: Uludoğan, Gökçe, et al.
Published: (2024)
by: Uludoğan, Gökçe, et al.
Published: (2024)
BERT-LSH: Reducing Absolute Compute For Attention
by: Li, Zezheng, et al.
Published: (2024)
by: Li, Zezheng, et al.
Published: (2024)
Teaching LLMs for Step-Level Automatic Math Correction via Reinforcement Learning
by: Li, Junsong, et al.
Published: (2025)
by: Li, Junsong, et al.
Published: (2025)
Injecting linguistic knowledge into BERT for Dialogue State Tracking
by: Feng, Xiaohan, et al.
Published: (2023)
by: Feng, Xiaohan, et al.
Published: (2023)
BERT Learns (and Teaches) Chemistry
by: Payne, Josh, et al.
Published: (2020)
by: Payne, Josh, et al.
Published: (2020)
Restoring Rhythm: Punctuation Restoration Using Transformer Models for Bangla, A Low-Resource Language
by: Mamun, Md Obyedullahil, et al.
Published: (2025)
by: Mamun, Md Obyedullahil, et al.
Published: (2025)
BERTCaps: BERT Capsule for Persian Multi-Domain Sentiment Analysis
by: Memari, Mohammadali, et al.
Published: (2024)
by: Memari, Mohammadali, et al.
Published: (2024)
BERT-JEPA: Reorganizing CLS Embeddings for Language-Invariant Semantics
by: Gillin, Taj, et al.
Published: (2026)
by: Gillin, Taj, et al.
Published: (2026)
Feature Structure Distillation with Centered Kernel Alignment in BERT Transferring
by: Jung, Hee-Jun, et al.
Published: (2022)
by: Jung, Hee-Jun, et al.
Published: (2022)
Harnessing Large Language Models: Fine-tuned BERT for Detecting Charismatic Leadership Tactics in Natural Language
by: Saeid, Yasser, et al.
Published: (2024)
by: Saeid, Yasser, et al.
Published: (2024)
Exploring the Reversal Curse and Other Deductive Logical Reasoning in BERT and GPT-Based Large Language Models
by: Wu, Da, et al.
Published: (2023)
by: Wu, Da, et al.
Published: (2023)
SMART: Automatically Scaling Down Language Models with Accuracy Guarantees for Reduced Processing Fees
by: Jo, Saehan, et al.
Published: (2024)
by: Jo, Saehan, et al.
Published: (2024)
Bridging the Bosphorus: Advancing Turkish Large Language Models through Strategies for Low-Resource Language Adaptation and Benchmarking
by: Acikgoz, Emre Can, et al.
Published: (2024)
by: Acikgoz, Emre Can, et al.
Published: (2024)
Cancer Diagnosis Categorization in Electronic Health Records Using Large Language Models and BioBERT: Model Performance Evaluation Study
by: Hashtarkhani, Soheil, et al.
Published: (2025)
by: Hashtarkhani, Soheil, et al.
Published: (2025)
Automatically Labeling Clinical Trial Outcomes: A Large-Scale Benchmark for Drug Development
by: Gao, Chufan, et al.
Published: (2024)
by: Gao, Chufan, et al.
Published: (2024)
Seeing Through VisualBERT: A Causal Adventure on Memetic Landscapes
by: Bandyopadhyay, Dibyanayan, et al.
Published: (2024)
by: Bandyopadhyay, Dibyanayan, et al.
Published: (2024)
Clinical ModernBERT: An efficient and long context encoder for biomedical text
by: Lee, Simon A., et al.
Published: (2025)
by: Lee, Simon A., et al.
Published: (2025)
2-Tier SimCSE: Elevating BERT for Robust Sentence Embeddings
by: Wang, Yumeng, et al.
Published: (2025)
by: Wang, Yumeng, et al.
Published: (2025)
Foundational Automatic Evaluators: Scaling Multi-Task Generative Evaluator Training for Reasoning-Centric Domains
by: Xu, Austin, et al.
Published: (2025)
by: Xu, Austin, et al.
Published: (2025)
AI-Generated Text Detection and Classification Based on BERT Deep Learning Algorithm
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Single layer tiny Co$^4$ outpaces GPT-2 and GPT-BERT
by: Zain, Noor Ul, et al.
Published: (2025)
by: Zain, Noor Ul, et al.
Published: (2025)
BERT-ASC: Auxiliary-Sentence Construction for Implicit Aspect Learning in Sentiment Analysis
by: Ahmed, Murtadha, et al.
Published: (2022)
by: Ahmed, Murtadha, et al.
Published: (2022)
Similar Items
-
Iterative Mask Filling: An Effective Text Augmentation Method Using Masked Language Modeling
by: Kesgin, Himmet Toprak, et al.
Published: (2024) -
Investigating Semi-Supervised Learning Algorithms in Text Datasets
by: Kesgin, Himmet Toprak, et al.
Published: (2024) -
Advancing NLP Models with Strategic Text Augmentation: A Comprehensive Study of Augmentation Methods and Curriculum Strategies
by: Kesgin, Himmet Toprak, et al.
Published: (2024) -
Optimizing Large Language Models for Turkish: New Methodologies in Corpus Selection and Training
by: Kesgin, H. Toprak, et al.
Published: (2024) -
Introducing cosmosGPT: Monolingual Training for Turkish Language Models
by: Kesgin, H. Toprak, et al.
Published: (2024)