Qorgau: Evaluating LLM Safety in Kazakh-Russian Bilingual Contexts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Goloburda, Maiya, Laiyk, Nurkhan, Turmakhan, Diana, Wang, Yuxia, Togmanov, Mukhammed, Mansurov, Jonibek, Sametov, Askhat, Mukhituly, Nurdaulet, Wang, Minghan, Orel, Daniil, Mujahid, Zain Muhammad, Koto, Fajri, Baldwin, Timothy, Nakov, Preslav |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KazMMLU: Evaluating Language Models on Kazakh, Russian, and Regional Knowledge of Kazakhstan
von: Togmanov, Mukhammed, et al.
Veröffentlicht: (2025)
von: Togmanov, Mukhammed, et al.
Veröffentlicht: (2025)
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2025)
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2025)
Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs
von: Goloburda, Maiya, et al.
Veröffentlicht: (2026)
von: Goloburda, Maiya, et al.
Veröffentlicht: (2026)
Sherkala-Chat: Building a State-of-the-Art LLM for Kazakh in a Moderately Resourced Setting
von: Koto, Fajri, et al.
Veröffentlicht: (2025)
von: Koto, Fajri, et al.
Veröffentlicht: (2025)
Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2026)
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2026)
UNCERTAINTY-LINE: Length-Invariant Estimation of Uncertainty for Large Language Models
von: Vashurin, Roman, et al.
Veröffentlicht: (2025)
von: Vashurin, Roman, et al.
Veröffentlicht: (2025)
Rethinking STS and NLI in Large Language Models
von: Wang, Yuxia, et al.
Veröffentlicht: (2023)
von: Wang, Yuxia, et al.
Veröffentlicht: (2023)
Cross-Cultural Transfer of Commonsense Reasoning in LLMs: Evidence from the Arab World
von: Almheiri, Saeed, et al.
Veröffentlicht: (2025)
von: Almheiri, Saeed, et al.
Veröffentlicht: (2025)
Can Machines Resonate with Humans? Evaluating the Emotional and Empathic Comprehension of LMs
von: Manzoor, Muhammad Arslan, et al.
Veröffentlicht: (2024)
von: Manzoor, Muhammad Arslan, et al.
Veröffentlicht: (2024)
Uncertainty Quantification for LLMs through Minimum Bayes Risk: Bridging Confidence and Consistency
von: Vashurin, Roman, et al.
Veröffentlicht: (2025)
von: Vashurin, Roman, et al.
Veröffentlicht: (2025)
A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis
von: Manzoor, Muhammad Arslan, et al.
Veröffentlicht: (2026)
von: Manzoor, Muhammad Arslan, et al.
Veröffentlicht: (2026)
Profiling News Media for Factuality and Bias Using LLMs and the Fact-Checking Methodology of Human Experts
von: Mujahid, Zain Muhammad, et al.
Veröffentlicht: (2025)
von: Mujahid, Zain Muhammad, et al.
Veröffentlicht: (2025)
CoDet-M4: Detecting Machine-Generated Code in Multi-Lingual, Multi-Generator and Multi-Domain Settings
von: Orel, Daniil, et al.
Veröffentlicht: (2025)
von: Orel, Daniil, et al.
Veröffentlicht: (2025)
Cracking the Code: Multi-domain LLM Evaluation on Real-World Professional Exams in Indonesia
von: Koto, Fajri
Veröffentlicht: (2024)
von: Koto, Fajri
Veröffentlicht: (2024)
Don't Throw Away Your Beams: Improving Consistency-based Uncertainties in LLMs via Beam Search
von: Fadeeva, Ekaterina, et al.
Veröffentlicht: (2025)
von: Fadeeva, Ekaterina, et al.
Veröffentlicht: (2025)
AICD Bench: A Challenging Benchmark for AI-Generated Code Detection
von: Orel, Daniil, et al.
Veröffentlicht: (2026)
von: Orel, Daniil, et al.
Veröffentlicht: (2026)
Arabic Dataset for LLM Safeguard Evaluation
von: Ashraf, Yasser, et al.
Veröffentlicht: (2024)
von: Ashraf, Yasser, et al.
Veröffentlicht: (2024)
Is Human-Like Text Liked by Humans? Multilingual Human Detection and Preference Against AI
von: Wang, Yuxia, et al.
Veröffentlicht: (2025)
von: Wang, Yuxia, et al.
Veröffentlicht: (2025)
OpenFactCheck: Building, Benchmarking Customized Fact-Checking Systems and Evaluating the Factuality of Claims and LLMs
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
$\texttt{Droid}$: A Resource Suite for AI-Generated Code Detection
von: Orel, Daniil, et al.
Veröffentlicht: (2025)
von: Orel, Daniil, et al.
Veröffentlicht: (2025)
GenAI Content Detection Task 1: English and Multilingual Machine-Generated Text Detection: AI vs. Human
von: Wang, Yuxia, et al.
Veröffentlicht: (2025)
von: Wang, Yuxia, et al.
Veröffentlicht: (2025)
IndoCulture: Exploring Geographically-Influenced Cultural Commonsense Reasoning Across Eleven Indonesian Provinces
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
UnsafeChain: Enhancing Reasoning Model Safety via Hard Cases
von: Tomar, Raj Vardhan, et al.
Veröffentlicht: (2025)
von: Tomar, Raj Vardhan, et al.
Veröffentlicht: (2025)
How Does Prefix Matter in Reasoning Model Tuning?
von: Tomar, Raj Vardhan, et al.
Veröffentlicht: (2026)
von: Tomar, Raj Vardhan, et al.
Veröffentlicht: (2026)
Data Laundering: Artificially Boosting Benchmark Results through Knowledge Distillation
von: Mansurov, Jonibek, et al.
Veröffentlicht: (2024)
von: Mansurov, Jonibek, et al.
Veröffentlicht: (2024)
SPIRIT: Patching Speech Language Models against Jailbreak Attacks
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2025)
OpenFactCheck: A Unified Framework for Factuality Evaluation of LLMs
von: Iqbal, Hasan, et al.
Veröffentlicht: (2024)
von: Iqbal, Hasan, et al.
Veröffentlicht: (2024)
Are Multilingual LLMs Culturally-Diverse Reasoners? An Investigation into Multicultural Proverbs and Sayings
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
Instruction-Guided Poetry Generation in Arabic and Its Dialects
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2026)
MuDRiC: Multi-Dialect Reasoning for Arabic Commonsense Validation
von: Elozeiri, Kareem, et al.
Veröffentlicht: (2025)
von: Elozeiri, Kareem, et al.
Veröffentlicht: (2025)
Loki: An Open-Source Tool for Fact Verification
von: Li, Haonan, et al.
Veröffentlicht: (2024)
von: Li, Haonan, et al.
Veröffentlicht: (2024)
Zero-shot Sentiment Analysis in Low-Resource Languages Using a Multilingual Sentiment Lexicon
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
AMIR-GRPO: Inducing Implicit Preference Signals into GRPO
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2026)
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2026)
Unveiling Cultural Blind Spots: Analyzing the Limitations of mLLMs in Procedural Text Comprehension
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
HALF: Harm-Aware LLM Fairness Evaluation Aligned with Deployment
von: Mekky, Ali, et al.
Veröffentlicht: (2025)
von: Mekky, Ali, et al.
Veröffentlicht: (2025)
Factuality of Large Language Models: A Survey
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
Sycophancy Hides Linearly in the Attention Heads
von: Genadi, Rifo, et al.
Veröffentlicht: (2026)
von: Genadi, Rifo, et al.
Veröffentlicht: (2026)
COMMUNITYNOTES: A Dataset for Exploring the Helpfulness of Fact-Checking Explanations
von: Xing, Rui, et al.
Veröffentlicht: (2025)
von: Xing, Rui, et al.
Veröffentlicht: (2025)
M4GT-Bench: Evaluation Benchmark for Black-Box Machine-Generated Text Detection
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
Parallel Tokenizers: Rethinking Vocabulary Design for Cross-Lingual Transfer
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
KazMMLU: Evaluating Language Models on Kazakh, Russian, and Regional Knowledge of Kazakhstan
von: Togmanov, Mukhammed, et al.
Veröffentlicht: (2025) -
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2025) -
Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs
von: Goloburda, Maiya, et al.
Veröffentlicht: (2026) -
Sherkala-Chat: Building a State-of-the-Art LLM for Kazakh in a Moderately Resourced Setting
von: Koto, Fajri, et al.
Veröffentlicht: (2025) -
Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2026)