Estonian Native Large Language Model Benchmark
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lillepalu, Helena Grete, Alumäe, Tanel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Finetuning End-to-End Models for Estonian Conversational Spoken Language Translation
von: Sildam, Tiia, et al.
Veröffentlicht: (2024)
von: Sildam, Tiia, et al.
Veröffentlicht: (2024)
Optimizing Estonian TV Subtitles with Semi-supervised Learning and LLMs
von: Fedorchenko, Artem, et al.
Veröffentlicht: (2025)
von: Fedorchenko, Artem, et al.
Veröffentlicht: (2025)
Comparison of End-to-end Speech Assessment Models for the NOCASA 2025 Challenge
von: Žavoronkov, Aleksei, et al.
Veröffentlicht: (2025)
von: Žavoronkov, Aleksei, et al.
Veröffentlicht: (2025)
TalTech Systems for the Interspeech 2025 ML-SUPERB 2.0 Challenge
von: Alumäe, Tanel, et al.
Veröffentlicht: (2025)
von: Alumäe, Tanel, et al.
Veröffentlicht: (2025)
Multi-Source Evidence Fusion for Audio Question Answering
von: Olev, Aivo, et al.
Veröffentlicht: (2026)
von: Olev, Aivo, et al.
Veröffentlicht: (2026)
EstLLM: Enhancing Estonian Capabilities in Multilingual LLMs via Continued Pretraining and Post-Training
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2026)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2026)
Unveiling Audio Deepfake Origins: A Deep Metric learning And Conformer Network Approach With Ensemble Fusion
von: Kulkarni, Ajinkya, et al.
Veröffentlicht: (2025)
von: Kulkarni, Ajinkya, et al.
Veröffentlicht: (2025)
Do Compact SSL Backbones Matter for Audio Deepfake Detection? A Controlled Study with RAPTOR
von: Kulkarni, Ajinkya, et al.
Veröffentlicht: (2026)
von: Kulkarni, Ajinkya, et al.
Veröffentlicht: (2026)
Improving Estonian Text Simplification through Pretrained Language Models and Custom Datasets
von: Barbu, Eduard, et al.
Veröffentlicht: (2025)
von: Barbu, Eduard, et al.
Veröffentlicht: (2025)
Prune or Retrain: Optimizing the Vocabulary of Multilingual Models for Estonian
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2025)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2025)
Robust Training of Vector Quantized Bottleneck Models
von: Łańcucki, Adrian, et al.
Veröffentlicht: (2020)
von: Łańcucki, Adrian, et al.
Veröffentlicht: (2020)
An experimental and computational study of an Estonian single-person word naming
von: Lõo, Kaidi, et al.
Veröffentlicht: (2025)
von: Lõo, Kaidi, et al.
Veröffentlicht: (2025)
Comparison of Current Approaches to Lemmatization: A Case Study in Estonian
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
GliLem: Leveraging GliNER for Contextualized Lemmatization in Estonian
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
von: Dorkin, Aleksei, et al.
Veröffentlicht: (2024)
Leveraging Open-Source Large Language Models for Native Language Identification
von: Ng, Yee Man, et al.
Veröffentlicht: (2024)
von: Ng, Yee Man, et al.
Veröffentlicht: (2024)
Speech DF Arena: A Leaderboard for Speech DeepFake Detection Models
von: Dowerah, Sandipana, et al.
Veröffentlicht: (2025)
von: Dowerah, Sandipana, et al.
Veröffentlicht: (2025)
Creation of the Estonian Subjectivity Dataset: Assessing the Degree of Subjectivity on a Scale
von: Gailit, Karl Gustav, et al.
Veröffentlicht: (2025)
von: Gailit, Karl Gustav, et al.
Veröffentlicht: (2025)
GreekMMLU: A Native-Sourced Multitask Benchmark for Evaluating Language Models in Greek
von: Zhang, Yang, et al.
Veröffentlicht: (2026)
von: Zhang, Yang, et al.
Veröffentlicht: (2026)
Symbolic Grounding Reveals Representational Bottlenecks in Abstract Visual Reasoning
von: Vaishnav, Mohit, et al.
Veröffentlicht: (2026)
von: Vaishnav, Mohit, et al.
Veröffentlicht: (2026)
Estonian WinoGrande Dataset: Comparative Analysis of LLM Performance on Human and Machine Translation
von: Ojastu, Marii, et al.
Veröffentlicht: (2025)
von: Ojastu, Marii, et al.
Veröffentlicht: (2025)
Youtu-LLM: Unlocking the Native Agentic Potential for Lightweight Large Language Models
von: Lu, Junru, et al.
Veröffentlicht: (2025)
von: Lu, Junru, et al.
Veröffentlicht: (2025)
Autocorrect for Estonian texts: final report from project EKTB25
von: Luhtaru, Agnes, et al.
Veröffentlicht: (2024)
von: Luhtaru, Agnes, et al.
Veröffentlicht: (2024)
AI4Math: A Native Spanish Benchmark for University-Level Mathematical Reasoning in Large Language Models
von: Perez, Miguel Angel Peñaloza, et al.
Veröffentlicht: (2025)
von: Perez, Miguel Angel Peñaloza, et al.
Veröffentlicht: (2025)
TUMLU: A Unified and Native Language Understanding Benchmark for Turkic Languages
von: Isbarov, Jafar, et al.
Veröffentlicht: (2025)
von: Isbarov, Jafar, et al.
Veröffentlicht: (2025)
Native Design Bias: Studying the Impact of English Nativeness on Language Model Performance
von: Reusens, Manon, et al.
Veröffentlicht: (2024)
von: Reusens, Manon, et al.
Veröffentlicht: (2024)
BenchmarkCards: Standardized Documentation for Large Language Model Benchmarks
von: Sokol, Anna, et al.
Veröffentlicht: (2024)
von: Sokol, Anna, et al.
Veröffentlicht: (2024)
A Survey on Large Language Model Benchmarks
von: Ni, Shiwen, et al.
Veröffentlicht: (2025)
von: Ni, Shiwen, et al.
Veröffentlicht: (2025)
Benchmarking Linguistic Diversity of Large Language Models
von: Guo, Yanzhu, et al.
Veröffentlicht: (2024)
von: Guo, Yanzhu, et al.
Veröffentlicht: (2024)
MEGAVERSE: Benchmarking Large Language Models Across Languages, Modalities, Models and Tasks
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2023)
von: Ahuja, Sanchit, et al.
Veröffentlicht: (2023)
Lost in Benchmarks? Rethinking Large Language Model Benchmarking with Item Response Theory
von: Zhou, Hongli, et al.
Veröffentlicht: (2025)
von: Zhou, Hongli, et al.
Veröffentlicht: (2025)
Evalita-LLM: Benchmarking Large Language Models on Italian
von: Magnini, Bernardo, et al.
Veröffentlicht: (2025)
von: Magnini, Bernardo, et al.
Veröffentlicht: (2025)
Benchmarking and Rethinking Knowledge Editing for Large Language Models
von: He, Guoxiu, et al.
Veröffentlicht: (2025)
von: He, Guoxiu, et al.
Veröffentlicht: (2025)
Benchmarking Motivational Interviewing Competence of Large Language Models
von: Jha, Aishwariya, et al.
Veröffentlicht: (2026)
von: Jha, Aishwariya, et al.
Veröffentlicht: (2026)
Ebisu: Benchmarking Large Language Models in Japanese Finance
von: Peng, Xueqing, et al.
Veröffentlicht: (2026)
von: Peng, Xueqing, et al.
Veröffentlicht: (2026)
TRAM: Benchmarking Temporal Reasoning for Large Language Models
von: Wang, Yuqing, et al.
Veröffentlicht: (2023)
von: Wang, Yuqing, et al.
Veröffentlicht: (2023)
ArabianGPT: Native Arabic GPT-based Large Language Model
von: Koubaa, Anis, et al.
Veröffentlicht: (2024)
von: Koubaa, Anis, et al.
Veröffentlicht: (2024)
Benchmarking Benchmark Leakage in Large Language Models
von: Xu, Ruijie, et al.
Veröffentlicht: (2024)
von: Xu, Ruijie, et al.
Veröffentlicht: (2024)
Benchmarking Large Language Models for Cryptanalysis and Side-Channel Vulnerabilities
von: Maskey, Utsav, et al.
Veröffentlicht: (2025)
von: Maskey, Utsav, et al.
Veröffentlicht: (2025)
User Profile with Large Language Models: Construction, Updating, and Benchmarking
von: Prottasha, Nusrat Jahan, et al.
Veröffentlicht: (2025)
von: Prottasha, Nusrat Jahan, et al.
Veröffentlicht: (2025)
ProBench: Benchmarking Large Language Models in Competitive Programming
von: Yang, Lei, et al.
Veröffentlicht: (2025)
von: Yang, Lei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Finetuning End-to-End Models for Estonian Conversational Spoken Language Translation
von: Sildam, Tiia, et al.
Veröffentlicht: (2024) -
Optimizing Estonian TV Subtitles with Semi-supervised Learning and LLMs
von: Fedorchenko, Artem, et al.
Veröffentlicht: (2025) -
Comparison of End-to-end Speech Assessment Models for the NOCASA 2025 Challenge
von: Žavoronkov, Aleksei, et al.
Veröffentlicht: (2025) -
TalTech Systems for the Interspeech 2025 ML-SUPERB 2.0 Challenge
von: Alumäe, Tanel, et al.
Veröffentlicht: (2025) -
Multi-Source Evidence Fusion for Audio Question Answering
von: Olev, Aivo, et al.
Veröffentlicht: (2026)