OCRTurk: A Comprehensive OCR Benchmark for Turkish
Fuente:
arXiv
Salvato in:
| Autori principali: | Yılmaz, Deniz, Munis, Evren Ayberk, Toraman, Çağrı, Köse, Süha Kağan, Aktaş, Burak, Baytekin, Mehmet Can, Görür, Bilge Kaan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RAGTurk: Best Practices for Retrieval Augmented Generation in Turkish
di: Köse, Süha Kağan, et al.
Pubblicazione: (2026)
di: Köse, Süha Kağan, et al.
Pubblicazione: (2026)
BIRDTurk: Adaptation of the BIRD Text-to-SQL Dataset to Turkish
di: Aktaş, Burak, et al.
Pubblicazione: (2026)
di: Aktaş, Burak, et al.
Pubblicazione: (2026)
FIBER: A Multilingual Evaluation Resource for Factual Inference Bias
di: Munis, Evren Ayberk, et al.
Pubblicazione: (2025)
di: Munis, Evren Ayberk, et al.
Pubblicazione: (2025)
RAGSmith: A Framework for Finding the Optimal Composition of Retrieval-Augmented Generation Methods Across Datasets
di: Kartal, Muhammed Yusuf, et al.
Pubblicazione: (2025)
di: Kartal, Muhammed Yusuf, et al.
Pubblicazione: (2025)
Evaluating the Quality of Benchmark Datasets for Low-Resource Languages: A Case Study on Turkish
di: Cengiz, Ayşe Aysu, et al.
Pubblicazione: (2025)
di: Cengiz, Ayşe Aysu, et al.
Pubblicazione: (2025)
LlamaTurk: Adapting Open-Source Generative Large Language Models for Low-Resource Language
di: Toraman, Cagri
Pubblicazione: (2024)
di: Toraman, Cagri
Pubblicazione: (2024)
OpenEthics: A Comprehensive Ethical Evaluation of Open-Source Generative Large Language Models
di: Özen, Yıldırım, et al.
Pubblicazione: (2025)
di: Özen, Yıldırım, et al.
Pubblicazione: (2025)
ArgLLM-App: An Interactive System for Argumentative Reasoning with Large Language Models
di: Dejl, Adam, et al.
Pubblicazione: (2026)
di: Dejl, Adam, et al.
Pubblicazione: (2026)
TurkBench: A Benchmark for Evaluating Turkish Large Language Models
di: Toraman, Çağrı, et al.
Pubblicazione: (2026)
di: Toraman, Çağrı, et al.
Pubblicazione: (2026)
Bridging the Bosphorus: Advancing Turkish Large Language Models through Strategies for Low-Resource Language Adaptation and Benchmarking
di: Acikgoz, Emre Can, et al.
Pubblicazione: (2024)
di: Acikgoz, Emre Can, et al.
Pubblicazione: (2024)
Can Large Language Models perform Relation-based Argument Mining?
di: Gorur, Deniz, et al.
Pubblicazione: (2024)
di: Gorur, Deniz, et al.
Pubblicazione: (2024)
The Turkish validity and reliability of Addenbrooke's Cognitive Examination III
di: Mümüne Merve Parlak, et al.
Pubblicazione: (2024)
di: Mümüne Merve Parlak, et al.
Pubblicazione: (2024)
Context Aware Lemmatization and Morphological Tagging Method in Turkish
di: Sayallar, Cagri
Pubblicazione: (2025)
di: Sayallar, Cagri
Pubblicazione: (2025)
MiDe22: An Annotated Multi-Event Tweet Dataset for Misinformation Detection
di: Toraman, Cagri, et al.
Pubblicazione: (2022)
di: Toraman, Cagri, et al.
Pubblicazione: (2022)
VBART: The Turkish LLM
di: Turker, Meliksah, et al.
Pubblicazione: (2024)
di: Turker, Meliksah, et al.
Pubblicazione: (2024)
When Many-Shot Prompting Fails: An Empirical Study of LLM Code Translation
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
VNLP: Turkish NLP Package
di: Turker, Meliksah, et al.
Pubblicazione: (2024)
di: Turker, Meliksah, et al.
Pubblicazione: (2024)
Argumentative Large Language Models for Explainable and Contestable Claim Verification
di: Freedman, Gabriel, et al.
Pubblicazione: (2024)
di: Freedman, Gabriel, et al.
Pubblicazione: (2024)
Introducing TrGLUE and SentiTurca: A Comprehensive Benchmark for Turkish General Language Understanding and Sentiment Analysis
di: Altinok, Duygu
Pubblicazione: (2025)
di: Altinok, Duygu
Pubblicazione: (2025)
A Comprehensive Analysis of Static Word Embeddings for Turkish
di: Sarıtaş, Karahan, et al.
Pubblicazione: (2024)
di: Sarıtaş, Karahan, et al.
Pubblicazione: (2024)
Introducing cosmosGPT: Monolingual Training for Turkish Language Models
di: Kesgin, H. Toprak, et al.
Pubblicazione: (2024)
di: Kesgin, H. Toprak, et al.
Pubblicazione: (2024)
There Are No Silly Questions: Evaluation of Offline LLM Capabilities from a Turkish Perspective
di: Yilmaz, Edibe, et al.
Pubblicazione: (2026)
di: Yilmaz, Edibe, et al.
Pubblicazione: (2026)
PejorativITy: Disambiguating Pejorative Epithets to Improve Misogyny Detection in Italian Tweets
di: Muti, Arianna, et al.
Pubblicazione: (2024)
di: Muti, Arianna, et al.
Pubblicazione: (2024)
Retrieval- and Argumentation-Enhanced Multi-Agent LLMs for Judgmental Forecasting (Extended Version with Supplementary Material)
di: Gorur, Deniz, et al.
Pubblicazione: (2025)
di: Gorur, Deniz, et al.
Pubblicazione: (2025)
Türkçe Dil Modellerinin Performans Karşılaştırması Performance Comparison of Turkish Language Models
di: Dogan, Eren, et al.
Pubblicazione: (2024)
di: Dogan, Eren, et al.
Pubblicazione: (2024)
Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction
di: Saoud, Abdulkader, et al.
Pubblicazione: (2024)
di: Saoud, Abdulkader, et al.
Pubblicazione: (2024)
When Semantic Overlap Is Not Enough: Cross-Lingual Euphemism Transfer Between Turkish and English
di: Biyik, Hasan Can, et al.
Pubblicazione: (2026)
di: Biyik, Hasan Can, et al.
Pubblicazione: (2026)
Argumentatively Coherent Judgmental Forecasting
di: Gorur, Deniz, et al.
Pubblicazione: (2025)
di: Gorur, Deniz, et al.
Pubblicazione: (2025)
KITAB-Bench: A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understanding
di: Heakl, Ahmed, et al.
Pubblicazione: (2025)
di: Heakl, Ahmed, et al.
Pubblicazione: (2025)
Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation
di: Karakaş, Sercan, et al.
Pubblicazione: (2026)
di: Karakaş, Sercan, et al.
Pubblicazione: (2026)
Optimizing Large Language Models for Turkish: New Methodologies in Corpus Selection and Training
di: Kesgin, H. Toprak, et al.
Pubblicazione: (2024)
di: Kesgin, H. Toprak, et al.
Pubblicazione: (2024)
UGPhysics: A Comprehensive Benchmark for Undergraduate Physics Reasoning with Large Language Models
di: Xu, Xin, et al.
Pubblicazione: (2025)
di: Xu, Xin, et al.
Pubblicazione: (2025)
Privacy-Preserving Local Language Models for Longitudinal Data Retrieval in Chronic Dermatologic Disease: Implementation in Pemphigus Patients
di: Yilmaz, Abdurrahim, et al.
Pubblicazione: (2026)
di: Yilmaz, Abdurrahim, et al.
Pubblicazione: (2026)
Can Large Language Models Generate Effective Datasets for Emotion Recognition in Conversations?
di: Kaplan, Burak Can, et al.
Pubblicazione: (2025)
di: Kaplan, Burak Can, et al.
Pubblicazione: (2025)
PEaCE: A Chemistry-Oriented Dataset for Optical Character Recognition on Scientific Documents
di: Zhang, Nan, et al.
Pubblicazione: (2024)
di: Zhang, Nan, et al.
Pubblicazione: (2024)
Turkish Delights: a Dataset on Turkish Euphemisms
di: Biyik, Hasan Can, et al.
Pubblicazione: (2024)
di: Biyik, Hasan Can, et al.
Pubblicazione: (2024)
BreakFun: Jailbreaking LLMs via Schema Exploitation
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
Multimodal LLMs for OCR, OCR Post-Correction, and Named Entity Recognition in Historical Documents
di: Greif, Gavin, et al.
Pubblicazione: (2025)
di: Greif, Gavin, et al.
Pubblicazione: (2025)
Advancing NLP Models with Strategic Text Augmentation: A Comprehensive Study of Augmentation Methods and Curriculum Strategies
di: Kesgin, Himmet Toprak, et al.
Pubblicazione: (2024)
di: Kesgin, Himmet Toprak, et al.
Pubblicazione: (2024)
MRCEval: A Comprehensive, Challenging and Accessible Machine Reading Comprehension Benchmark
di: Ma, Shengkun, et al.
Pubblicazione: (2025)
di: Ma, Shengkun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
RAGTurk: Best Practices for Retrieval Augmented Generation in Turkish
di: Köse, Süha Kağan, et al.
Pubblicazione: (2026) -
BIRDTurk: Adaptation of the BIRD Text-to-SQL Dataset to Turkish
di: Aktaş, Burak, et al.
Pubblicazione: (2026) -
FIBER: A Multilingual Evaluation Resource for Factual Inference Bias
di: Munis, Evren Ayberk, et al.
Pubblicazione: (2025) -
RAGSmith: A Framework for Finding the Optimal Composition of Retrieval-Augmented Generation Methods Across Datasets
di: Kartal, Muhammed Yusuf, et al.
Pubblicazione: (2025) -
Evaluating the Quality of Benchmark Datasets for Low-Resource Languages: A Case Study on Turkish
di: Cengiz, Ayşe Aysu, et al.
Pubblicazione: (2025)