Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
Fuente:
arXiv
Guardado en:
| Autores principales: | Laiyk, Nurkhan, Orel, Daniil, Joshi, Rituraj, Goloburda, Maiya, Wang, Yuxia, Nakov, Preslav, Koto, Fajri |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Qorgau: Evaluating LLM Safety in Kazakh-Russian Bilingual Contexts
por: Goloburda, Maiya, et al.
Publicado: (2025)
por: Goloburda, Maiya, et al.
Publicado: (2025)
Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs
por: Goloburda, Maiya, et al.
Publicado: (2026)
por: Goloburda, Maiya, et al.
Publicado: (2026)
KazMMLU: Evaluating Language Models on Kazakh, Russian, and Regional Knowledge of Kazakhstan
por: Togmanov, Mukhammed, et al.
Publicado: (2025)
por: Togmanov, Mukhammed, et al.
Publicado: (2025)
Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation
por: Laiyk, Nurkhan, et al.
Publicado: (2026)
por: Laiyk, Nurkhan, et al.
Publicado: (2026)
Sherkala-Chat: Building a State-of-the-Art LLM for Kazakh in a Moderately Resourced Setting
por: Koto, Fajri, et al.
Publicado: (2025)
por: Koto, Fajri, et al.
Publicado: (2025)
UNCERTAINTY-LINE: Length-Invariant Estimation of Uncertainty for Large Language Models
por: Vashurin, Roman, et al.
Publicado: (2025)
por: Vashurin, Roman, et al.
Publicado: (2025)
CoDet-M4: Detecting Machine-Generated Code in Multi-Lingual, Multi-Generator and Multi-Domain Settings
por: Orel, Daniil, et al.
Publicado: (2025)
por: Orel, Daniil, et al.
Publicado: (2025)
$\texttt{Droid}$: A Resource Suite for AI-Generated Code Detection
por: Orel, Daniil, et al.
Publicado: (2025)
por: Orel, Daniil, et al.
Publicado: (2025)
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
por: Pranida, Salsabila Zahirah, et al.
Publicado: (2025)
por: Pranida, Salsabila Zahirah, et al.
Publicado: (2025)
AICD Bench: A Challenging Benchmark for AI-Generated Code Detection
por: Orel, Daniil, et al.
Publicado: (2026)
por: Orel, Daniil, et al.
Publicado: (2026)
Cross-Cultural Transfer of Commonsense Reasoning in LLMs: Evidence from the Arab World
por: Almheiri, Saeed, et al.
Publicado: (2025)
por: Almheiri, Saeed, et al.
Publicado: (2025)
Uncertainty Quantification for LLMs through Minimum Bayes Risk: Bridging Confidence and Consistency
por: Vashurin, Roman, et al.
Publicado: (2025)
por: Vashurin, Roman, et al.
Publicado: (2025)
How Does Prefix Matter in Reasoning Model Tuning?
por: Tomar, Raj Vardhan, et al.
Publicado: (2026)
por: Tomar, Raj Vardhan, et al.
Publicado: (2026)
Rethinking STS and NLI in Large Language Models
por: Wang, Yuxia, et al.
Publicado: (2023)
por: Wang, Yuxia, et al.
Publicado: (2023)
Instruction-Guided Poetry Generation in Arabic and Its Dialects
por: Sadallah, Abdelrahman, et al.
Publicado: (2026)
por: Sadallah, Abdelrahman, et al.
Publicado: (2026)
UnsafeChain: Enhancing Reasoning Model Safety via Hard Cases
por: Tomar, Raj Vardhan, et al.
Publicado: (2025)
por: Tomar, Raj Vardhan, et al.
Publicado: (2025)
Cracking the Code: Multi-domain LLM Evaluation on Real-World Professional Exams in Indonesia
por: Koto, Fajri
Publicado: (2024)
por: Koto, Fajri
Publicado: (2024)
Don't Throw Away Your Beams: Improving Consistency-based Uncertainties in LLMs via Beam Search
por: Fadeeva, Ekaterina, et al.
Publicado: (2025)
por: Fadeeva, Ekaterina, et al.
Publicado: (2025)
Low-Resource Safety Failures Are Action Failures, Not Representation Failures
por: Aziz, Rashad, et al.
Publicado: (2026)
por: Aziz, Rashad, et al.
Publicado: (2026)
A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis
por: Manzoor, Muhammad Arslan, et al.
Publicado: (2026)
por: Manzoor, Muhammad Arslan, et al.
Publicado: (2026)
Unveiling Cultural Blind Spots: Analyzing the Limitations of mLLMs in Procedural Text Comprehension
por: Yari, Amir Hossein, et al.
Publicado: (2025)
por: Yari, Amir Hossein, et al.
Publicado: (2025)
Zero-shot Sentiment Analysis in Low-Resource Languages Using a Multilingual Sentiment Lexicon
por: Koto, Fajri, et al.
Publicado: (2024)
por: Koto, Fajri, et al.
Publicado: (2024)
Large Language Models are Few-Shot Training Example Generators: A Case Study in Fallacy Recognition
por: Alhindi, Tariq, et al.
Publicado: (2023)
por: Alhindi, Tariq, et al.
Publicado: (2023)
Exploring Language Model Generalization in Low-Resource Extractive QA
por: Sengupta, Saptarshi, et al.
Publicado: (2024)
por: Sengupta, Saptarshi, et al.
Publicado: (2024)
IndoCulture: Exploring Geographically-Influenced Cultural Commonsense Reasoning Across Eleven Indonesian Provinces
por: Koto, Fajri, et al.
Publicado: (2024)
por: Koto, Fajri, et al.
Publicado: (2024)
MuDRiC: Multi-Dialect Reasoning for Arabic Commonsense Validation
por: Elozeiri, Kareem, et al.
Publicado: (2025)
por: Elozeiri, Kareem, et al.
Publicado: (2025)
Toxicity Red-Teaming: Benchmarking LLM Safety in Singapore's Low-Resource Languages
por: Hu, Yujia, et al.
Publicado: (2025)
por: Hu, Yujia, et al.
Publicado: (2025)
AMIR-GRPO: Inducing Implicit Preference Signals into GRPO
por: Yari, Amir Hossein, et al.
Publicado: (2026)
por: Yari, Amir Hossein, et al.
Publicado: (2026)
Can Machines Resonate with Humans? Evaluating the Emotional and Empathic Comprehension of LMs
por: Manzoor, Muhammad Arslan, et al.
Publicado: (2024)
por: Manzoor, Muhammad Arslan, et al.
Publicado: (2024)
HALF: Harm-Aware LLM Fairness Evaluation Aligned with Deployment
por: Mekky, Ali, et al.
Publicado: (2025)
por: Mekky, Ali, et al.
Publicado: (2025)
LLMs as Cultural Archives: Cultural Commonsense Knowledge Graph Extraction
por: Tonga, Junior Cedric, et al.
Publicado: (2026)
por: Tonga, Junior Cedric, et al.
Publicado: (2026)
IndoSafety: Culturally Grounded Safety for LLMs in Indonesian Languages
por: Azmi, Muhammad Falensi, et al.
Publicado: (2025)
por: Azmi, Muhammad Falensi, et al.
Publicado: (2025)
Are Multilingual LLMs Culturally-Diverse Reasoners? An Investigation into Multicultural Proverbs and Sayings
por: Liu, Chen Cecilia, et al.
Publicado: (2023)
por: Liu, Chen Cecilia, et al.
Publicado: (2023)
A Survey of Confidence Estimation and Calibration in Large Language Models
por: Geng, Jiahui, et al.
Publicado: (2023)
por: Geng, Jiahui, et al.
Publicado: (2023)
Arabic Dataset for LLM Safeguard Evaluation
por: Ashraf, Yasser, et al.
Publicado: (2024)
por: Ashraf, Yasser, et al.
Publicado: (2024)
Con Instruction: Universal Jailbreaking of Multimodal Large Language Models via Non-Textual Modalities
por: Geng, Jiahui, et al.
Publicado: (2025)
por: Geng, Jiahui, et al.
Publicado: (2025)
Parallel Tokenizers: Rethinking Vocabulary Design for Cross-Lingual Transfer
por: Kautsar, Muhammad Dehan Al, et al.
Publicado: (2025)
por: Kautsar, Muhammad Dehan Al, et al.
Publicado: (2025)
Adapting Fake News Detection to the Era of Large Language Models
por: Su, Jinyan, et al.
Publicado: (2023)
por: Su, Jinyan, et al.
Publicado: (2023)
DUAL-Bench: Measuring Over-Refusal and Robustness in Vision-Language Models
por: Ren, Kaixuan, et al.
Publicado: (2025)
por: Ren, Kaixuan, et al.
Publicado: (2025)
Is Human-Like Text Liked by Humans? Multilingual Human Detection and Preference Against AI
por: Wang, Yuxia, et al.
Publicado: (2025)
por: Wang, Yuxia, et al.
Publicado: (2025)
Ejemplares similares
-
Qorgau: Evaluating LLM Safety in Kazakh-Russian Bilingual Contexts
por: Goloburda, Maiya, et al.
Publicado: (2025) -
Why Don't You Know? Evaluating the Impact of Uncertainty Sources on Uncertainty Quantification in LLMs
por: Goloburda, Maiya, et al.
Publicado: (2026) -
KazMMLU: Evaluating Language Models on Kazakh, Russian, and Regional Knowledge of Kazakhstan
por: Togmanov, Mukhammed, et al.
Publicado: (2025) -
Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation
por: Laiyk, Nurkhan, et al.
Publicado: (2026) -
Sherkala-Chat: Building a State-of-the-Art LLM for Kazakh in a Moderately Resourced Setting
por: Koto, Fajri, et al.
Publicado: (2025)