Evaluating Metalinguistic Knowledge in Large Language Models across the World's Languages
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Arčon, Tjaša, Klemen, Matej, Robnik-Šikonja, Marko, Dobrovoljc, Kaja |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Corpus-Grounded Agentic LLMs for Multilingual Grammatical Analysis
von: Klemen, Matej, et al.
Veröffentlicht: (2025)
von: Klemen, Matej, et al.
Veröffentlicht: (2025)
Large language models for folktale type automation based on motifs: Cinderella case study
von: Arčon, Tjaša, et al.
Veröffentlicht: (2025)
von: Arčon, Tjaša, et al.
Veröffentlicht: (2025)
Building a Strong Instruction Language Model for a Less-Resourced Language
von: Vreš, Domen, et al.
Veröffentlicht: (2026)
von: Vreš, Domen, et al.
Veröffentlicht: (2026)
Challenges in Explaining Pretrained Clinical Text Classifiers
von: Miok, Kristian, et al.
Veröffentlicht: (2026)
von: Miok, Kristian, et al.
Veröffentlicht: (2026)
Neural spell-checker: Beyond words with synthetic data generation
von: Klemen, Matej, et al.
Veröffentlicht: (2024)
von: Klemen, Matej, et al.
Veröffentlicht: (2024)
Counting trees: A treebank-driven exploration of syntactic variation in speech and writing across languages
von: Dobrovoljc, Kaja
Veröffentlicht: (2025)
von: Dobrovoljc, Kaja
Veröffentlicht: (2025)
Code-mixed Sentiment and Hate-speech Prediction
von: Yadav, Anjali, et al.
Veröffentlicht: (2024)
von: Yadav, Anjali, et al.
Veröffentlicht: (2024)
Sarcasm Detection in a Less-Resourced Language
von: Đoković, Lazar, et al.
Veröffentlicht: (2024)
von: Đoković, Lazar, et al.
Veröffentlicht: (2024)
Tracking Semantic Change in Slovene: A Novel Dataset and Optimal Transport-Based Distance
von: Pranjić, Marko, et al.
Veröffentlicht: (2024)
von: Pranjić, Marko, et al.
Veröffentlicht: (2024)
Solving Word-Sense Disambiguation and Word-Sense Induction with Dictionary Examples
von: Škvorc, Tadej, et al.
Veröffentlicht: (2025)
von: Škvorc, Tadej, et al.
Veröffentlicht: (2025)
QFS-Composer: Query-focused summarization pipeline for less resourced languages
von: Đuranović, Vuk, et al.
Veröffentlicht: (2026)
von: Đuranović, Vuk, et al.
Veröffentlicht: (2026)
Review of Natural Language Processing in Pharmacology
von: Trajanov, Dimitar, et al.
Veröffentlicht: (2022)
von: Trajanov, Dimitar, et al.
Veröffentlicht: (2022)
Generative Model for Less-Resourced Language with 1 billion parameters
von: Vreš, Domen, et al.
Veröffentlicht: (2024)
von: Vreš, Domen, et al.
Veröffentlicht: (2024)
Improving LLMs for Machine Translation Using Synthetic Preference Data
von: Vajda, Dario, et al.
Veröffentlicht: (2025)
von: Vajda, Dario, et al.
Veröffentlicht: (2025)
Linguistic Characteristics of AI-Generated Text: A Survey
von: Terčon, Luka, et al.
Veröffentlicht: (2025)
von: Terčon, Luka, et al.
Veröffentlicht: (2025)
Real-time News Story Identification
von: Škvorc, Tadej, et al.
Veröffentlicht: (2025)
von: Škvorc, Tadej, et al.
Veröffentlicht: (2025)
Measuring Catastrophic Forgetting in Cross-Lingual Transfer Paradigms: Exploring Tuning Strategies
von: Koloski, Boshko, et al.
Veröffentlicht: (2023)
von: Koloski, Boshko, et al.
Veröffentlicht: (2023)
TT-XAI: Trustworthy Clinical Text Explanations via Keyword Distillation and LLM Reasoning
von: Miok, Kristian, et al.
Veröffentlicht: (2025)
von: Miok, Kristian, et al.
Veröffentlicht: (2025)
Incremental Graph Construction Enables Robust Spectral Clustering of Texts
von: Pranjić, Marko, et al.
Veröffentlicht: (2026)
von: Pranjić, Marko, et al.
Veröffentlicht: (2026)
Retrieval-augmented code completion for local projects using large language models
von: Hostnik, Marko, et al.
Veröffentlicht: (2024)
von: Hostnik, Marko, et al.
Veröffentlicht: (2024)
I am a Strange Dataset: Metalinguistic Tests for Language Models
von: Thrush, Tristan, et al.
Veröffentlicht: (2024)
von: Thrush, Tristan, et al.
Veröffentlicht: (2024)
Creating ConLangs to Probe the Metalinguistic Grammatical Knowledge of LLMs
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2025)
von: Taguchi, Chihiro, et al.
Veröffentlicht: (2025)
Coconstructions in spoken data: UD annotation guidelines and first results
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2026)
von: Pannitto, Ludovica, et al.
Veröffentlicht: (2026)
A Benchmark for the Detection of Metalinguistic Disagreements between LLMs and Knowledge Graphs
von: Allen, Bradley P., et al.
Veröffentlicht: (2025)
von: Allen, Bradley P., et al.
Veröffentlicht: (2025)
Carpe Diem: On the Evaluation of World Knowledge in Lifelong Language Models
von: Kim, Yujin, et al.
Veröffentlicht: (2023)
von: Kim, Yujin, et al.
Veröffentlicht: (2023)
Making Large Language Models into World Models with Precondition and Effect Knowledge
von: Xie, Kaige, et al.
Veröffentlicht: (2024)
von: Xie, Kaige, et al.
Veröffentlicht: (2024)
KoLA: Carefully Benchmarking World Knowledge of Large Language Models
von: Yu, Jifan, et al.
Veröffentlicht: (2023)
von: Yu, Jifan, et al.
Veröffentlicht: (2023)
"See the World, Discover Knowledge": A Chinese Factuality Evaluation for Large Vision Language Models
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
Evaluating Hydro-Science and Engineering Knowledge of Large Language Models
von: Hu, Shiruo, et al.
Veröffentlicht: (2025)
von: Hu, Shiruo, et al.
Veröffentlicht: (2025)
KGQuiz: Evaluating the Generalization of Encoded Knowledge in Large Language Models
von: Bai, Yuyang, et al.
Veröffentlicht: (2023)
von: Bai, Yuyang, et al.
Veröffentlicht: (2023)
Evaluating Large Language Model with Knowledge Oriented Language Specific Simple Question Answering
von: Jiang, Bowen, et al.
Veröffentlicht: (2025)
von: Jiang, Bowen, et al.
Veröffentlicht: (2025)
Examining the Robustness of Large Language Models across Language Complexity
von: Zhang, Jiayi
Veröffentlicht: (2025)
von: Zhang, Jiayi
Veröffentlicht: (2025)
Gender Bias in Large Language Models across Multiple Languages
von: Zhao, Jinman, et al.
Veröffentlicht: (2024)
von: Zhao, Jinman, et al.
Veröffentlicht: (2024)
Evaluating Text Creativity across Diverse Domains: A Dataset and Large Language Model Evaluator
von: Cao, Qian, et al.
Veröffentlicht: (2025)
von: Cao, Qian, et al.
Veröffentlicht: (2025)
Evaluating Knowledge-based Cross-lingual Inconsistency in Large Language Models
von: Xing, Xiaolin, et al.
Veröffentlicht: (2024)
von: Xing, Xiaolin, et al.
Veröffentlicht: (2024)
Automating Evaluation of Diffusion Model Unlearning with (Vision-) Language Model World Knowledge
von: Yeats, Eric, et al.
Veröffentlicht: (2025)
von: Yeats, Eric, et al.
Veröffentlicht: (2025)
RECKON: Large-scale Reference-based Efficient Knowledge Evaluation for Large Language Model
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
von: Zhang, Lin, et al.
Veröffentlicht: (2025)
Knowledge Fusion of Large Language Models
von: Wan, Fanqi, et al.
Veröffentlicht: (2024)
von: Wan, Fanqi, et al.
Veröffentlicht: (2024)
Knowledge Sanitization of Large Language Models
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2023)
von: Ishibashi, Yoichi, et al.
Veröffentlicht: (2023)
Financial Knowledge Large Language Model
von: Yang, Cehao, et al.
Veröffentlicht: (2024)
von: Yang, Cehao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Corpus-Grounded Agentic LLMs for Multilingual Grammatical Analysis
von: Klemen, Matej, et al.
Veröffentlicht: (2025) -
Large language models for folktale type automation based on motifs: Cinderella case study
von: Arčon, Tjaša, et al.
Veröffentlicht: (2025) -
Building a Strong Instruction Language Model for a Less-Resourced Language
von: Vreš, Domen, et al.
Veröffentlicht: (2026) -
Challenges in Explaining Pretrained Clinical Text Classifiers
von: Miok, Kristian, et al.
Veröffentlicht: (2026) -
Neural spell-checker: Beyond words with synthetic data generation
von: Klemen, Matej, et al.
Veröffentlicht: (2024)