Evaluating Metalinguistic Knowledge in Large Language Models across the World's Languages
Fuente:
arXiv
Saved in:
| Main Authors: | Arčon, Tjaša, Klemen, Matej, Robnik-Šikonja, Marko, Dobrovoljc, Kaja |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Corpus-Grounded Agentic LLMs for Multilingual Grammatical Analysis
by: Klemen, Matej, et al.
Published: (2025)
by: Klemen, Matej, et al.
Published: (2025)
Large language models for folktale type automation based on motifs: Cinderella case study
by: Arčon, Tjaša, et al.
Published: (2025)
by: Arčon, Tjaša, et al.
Published: (2025)
Building a Strong Instruction Language Model for a Less-Resourced Language
by: Vreš, Domen, et al.
Published: (2026)
by: Vreš, Domen, et al.
Published: (2026)
Challenges in Explaining Pretrained Clinical Text Classifiers
by: Miok, Kristian, et al.
Published: (2026)
by: Miok, Kristian, et al.
Published: (2026)
Neural spell-checker: Beyond words with synthetic data generation
by: Klemen, Matej, et al.
Published: (2024)
by: Klemen, Matej, et al.
Published: (2024)
Counting trees: A treebank-driven exploration of syntactic variation in speech and writing across languages
by: Dobrovoljc, Kaja
Published: (2025)
by: Dobrovoljc, Kaja
Published: (2025)
Code-mixed Sentiment and Hate-speech Prediction
by: Yadav, Anjali, et al.
Published: (2024)
by: Yadav, Anjali, et al.
Published: (2024)
Sarcasm Detection in a Less-Resourced Language
by: Đoković, Lazar, et al.
Published: (2024)
by: Đoković, Lazar, et al.
Published: (2024)
Tracking Semantic Change in Slovene: A Novel Dataset and Optimal Transport-Based Distance
by: Pranjić, Marko, et al.
Published: (2024)
by: Pranjić, Marko, et al.
Published: (2024)
Solving Word-Sense Disambiguation and Word-Sense Induction with Dictionary Examples
by: Škvorc, Tadej, et al.
Published: (2025)
by: Škvorc, Tadej, et al.
Published: (2025)
QFS-Composer: Query-focused summarization pipeline for less resourced languages
by: Đuranović, Vuk, et al.
Published: (2026)
by: Đuranović, Vuk, et al.
Published: (2026)
Review of Natural Language Processing in Pharmacology
by: Trajanov, Dimitar, et al.
Published: (2022)
by: Trajanov, Dimitar, et al.
Published: (2022)
Generative Model for Less-Resourced Language with 1 billion parameters
by: Vreš, Domen, et al.
Published: (2024)
by: Vreš, Domen, et al.
Published: (2024)
Improving LLMs for Machine Translation Using Synthetic Preference Data
by: Vajda, Dario, et al.
Published: (2025)
by: Vajda, Dario, et al.
Published: (2025)
Linguistic Characteristics of AI-Generated Text: A Survey
by: Terčon, Luka, et al.
Published: (2025)
by: Terčon, Luka, et al.
Published: (2025)
Real-time News Story Identification
by: Škvorc, Tadej, et al.
Published: (2025)
by: Škvorc, Tadej, et al.
Published: (2025)
Measuring Catastrophic Forgetting in Cross-Lingual Transfer Paradigms: Exploring Tuning Strategies
by: Koloski, Boshko, et al.
Published: (2023)
by: Koloski, Boshko, et al.
Published: (2023)
TT-XAI: Trustworthy Clinical Text Explanations via Keyword Distillation and LLM Reasoning
by: Miok, Kristian, et al.
Published: (2025)
by: Miok, Kristian, et al.
Published: (2025)
Incremental Graph Construction Enables Robust Spectral Clustering of Texts
by: Pranjić, Marko, et al.
Published: (2026)
by: Pranjić, Marko, et al.
Published: (2026)
Retrieval-augmented code completion for local projects using large language models
by: Hostnik, Marko, et al.
Published: (2024)
by: Hostnik, Marko, et al.
Published: (2024)
I am a Strange Dataset: Metalinguistic Tests for Language Models
by: Thrush, Tristan, et al.
Published: (2024)
by: Thrush, Tristan, et al.
Published: (2024)
Creating ConLangs to Probe the Metalinguistic Grammatical Knowledge of LLMs
by: Taguchi, Chihiro, et al.
Published: (2025)
by: Taguchi, Chihiro, et al.
Published: (2025)
Coconstructions in spoken data: UD annotation guidelines and first results
by: Pannitto, Ludovica, et al.
Published: (2026)
by: Pannitto, Ludovica, et al.
Published: (2026)
A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs
by: Allen, Bradley P., et al.
Published: (2025)
by: Allen, Bradley P., et al.
Published: (2025)
Carpe Diem: On the Evaluation of World Knowledge in Lifelong Language Models
by: Kim, Yujin, et al.
Published: (2023)
by: Kim, Yujin, et al.
Published: (2023)
Making Large Language Models into World Models with Precondition and Effect Knowledge
by: Xie, Kaige, et al.
Published: (2024)
by: Xie, Kaige, et al.
Published: (2024)
KoLA: Carefully Benchmarking World Knowledge of Large Language Models
by: Yu, Jifan, et al.
Published: (2023)
by: Yu, Jifan, et al.
Published: (2023)
"See the World, Discover Knowledge": A Chinese Factuality Evaluation for Large Vision Language Models
by: Gu, Jihao, et al.
Published: (2025)
by: Gu, Jihao, et al.
Published: (2025)
Evaluating Hydro-Science and Engineering Knowledge of Large Language Models
by: Hu, Shiruo, et al.
Published: (2025)
by: Hu, Shiruo, et al.
Published: (2025)
KGQuiz: Evaluating the Generalization of Encoded Knowledge in Large Language Models
by: Bai, Yuyang, et al.
Published: (2023)
by: Bai, Yuyang, et al.
Published: (2023)
Evaluating Large Language Model with Knowledge Oriented Language Specific Simple Question Answering
by: Jiang, Bowen, et al.
Published: (2025)
by: Jiang, Bowen, et al.
Published: (2025)
Examining the Robustness of Large Language Models across Language Complexity
by: Zhang, Jiayi
Published: (2025)
by: Zhang, Jiayi
Published: (2025)
Gender Bias in Large Language Models across Multiple Languages
by: Zhao, Jinman, et al.
Published: (2024)
by: Zhao, Jinman, et al.
Published: (2024)
Evaluating Text Creativity across Diverse Domains: A Dataset and Large Language Model Evaluator
by: Cao, Qian, et al.
Published: (2025)
by: Cao, Qian, et al.
Published: (2025)
Evaluating Knowledge-based Cross-lingual Inconsistency in Large Language Models
by: Xing, Xiaolin, et al.
Published: (2024)
by: Xing, Xiaolin, et al.
Published: (2024)
Automating Evaluation of Diffusion Model Unlearning with (Vision-) Language Model World Knowledge
by: Yeats, Eric, et al.
Published: (2025)
by: Yeats, Eric, et al.
Published: (2025)
RECKON: Large-scale Reference-based Efficient Knowledge Evaluation for Large Language Model
by: Zhang, Lin, et al.
Published: (2025)
by: Zhang, Lin, et al.
Published: (2025)
Knowledge Fusion of Large Language Models
by: Wan, Fanqi, et al.
Published: (2024)
by: Wan, Fanqi, et al.
Published: (2024)
Knowledge Sanitization of Large Language Models
by: Ishibashi, Yoichi, et al.
Published: (2023)
by: Ishibashi, Yoichi, et al.
Published: (2023)
Financial Knowledge Large Language Model
by: Yang, Cehao, et al.
Published: (2024)
by: Yang, Cehao, et al.
Published: (2024)
Similar Items
-
Towards Corpus-Grounded Agentic LLMs for Multilingual Grammatical Analysis
by: Klemen, Matej, et al.
Published: (2025) -
Large language models for folktale type automation based on motifs: Cinderella case study
by: Arčon, Tjaša, et al.
Published: (2025) -
Building a Strong Instruction Language Model for a Less-Resourced Language
by: Vreš, Domen, et al.
Published: (2026) -
Challenges in Explaining Pretrained Clinical Text Classifiers
by: Miok, Kristian, et al.
Published: (2026) -
Neural spell-checker: Beyond words with synthetic data generation
by: Klemen, Matej, et al.
Published: (2024)