IndoSafety: Culturally Grounded Safety for LLMs in Indonesian Languages
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Azmi, Muhammad Falensi, Kautsar, Muhammad Dehan Al, Wicaksono, Alfan Farizki, Koto, Fajri |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Preserving Fairness and Safety in Quantized LLMs Through Critical Weight Protection
von: Hakim, Muhammad Alif Al, et al.
Veröffentlicht: (2026)
von: Hakim, Muhammad Alif Al, et al.
Veröffentlicht: (2026)
Simulating Training Data Leakage in Multiple-Choice Benchmarks for LLM Evaluation
von: Hidayat, Naila Shafirni, et al.
Veröffentlicht: (2025)
von: Hidayat, Naila Shafirni, et al.
Veröffentlicht: (2025)
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2026)
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2026)
Parallel Tokenizers: Rethinking Vocabulary Design for Cross-Lingual Transfer
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
What Do Indonesians Really Need from Language Technology? A Nationwide Survey
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
Evaluating Vision-Language and Large Language Models for Automated Student Assessment in Indonesian Classrooms
von: Aisyah, Nurul, et al.
Veröffentlicht: (2025)
von: Aisyah, Nurul, et al.
Veröffentlicht: (2025)
Grounding AI-in-Education Development in Teachers' Voices: Findings from a National Survey in Indonesia
von: Aisyah, Nurul, et al.
Veröffentlicht: (2026)
von: Aisyah, Nurul, et al.
Veröffentlicht: (2026)
University of Indonesia at SemEval-2025 Task 11: Evaluating State-of-the-Art Encoders for Multi-Label Emotion Detection
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2025)
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2025)
IndoCulture: Exploring Geographically-Influenced Cultural Commonsense Reasoning Across Eleven Indonesian Provinces
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
SEADialogues: A Multilingual Culturally Grounded Multi-turn Dialogue Dataset on Southeast Asian Languages
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
Role-Aware Language Models for Secure and Contextualized Access Control in Organizations
von: Almheiri, Saeed, et al.
Veröffentlicht: (2025)
von: Almheiri, Saeed, et al.
Veröffentlicht: (2025)
Vision Language Models are Confused Tourists
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2025)
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2025)
Unveiling Cultural Blind Spots: Analyzing the Limitations of mLLMs in Procedural Text Comprehension
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
Cracking the Code: Multi-domain LLM Evaluation on Real-World Professional Exams in Indonesia
von: Koto, Fajri
Veröffentlicht: (2024)
von: Koto, Fajri
Veröffentlicht: (2024)
Low-Resource Safety Failures Are Action Failures, Not Representation Failures
von: Aziz, Rashad, et al.
Veröffentlicht: (2026)
von: Aziz, Rashad, et al.
Veröffentlicht: (2026)
Pengembangan Model untuk Mendeteksi Kerusakan pada Terumbu Karang dengan Klasifikasi Citra
von: Muhammad, Fadhil, et al.
Veröffentlicht: (2023)
von: Muhammad, Fadhil, et al.
Veröffentlicht: (2023)
Sparse Autoencoders Can Capture Language-Specific Concepts Across Diverse Languages
von: Andrylie, Lyzander Marciano, et al.
Veröffentlicht: (2025)
von: Andrylie, Lyzander Marciano, et al.
Veröffentlicht: (2025)
LLMs as Cultural Archives: Cultural Commonsense Knowledge Graph Extraction
von: Tonga, Junior Cedric, et al.
Veröffentlicht: (2026)
von: Tonga, Junior Cedric, et al.
Veröffentlicht: (2026)
Unveiling the Influence of Amplifying Language-Specific Neurons
von: Rahmanisa, Inaya, et al.
Veröffentlicht: (2025)
von: Rahmanisa, Inaya, et al.
Veröffentlicht: (2025)
Cultural Benchmarking of LLMs in Standard and Dialectal Arabic Dialogues
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2026)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2026)
Are Multilingual LLMs Culturally-Diverse Reasoners? An Investigation into Multicultural Proverbs and Sayings
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
von: Pranida, Salsabila Zahirah, et al.
Veröffentlicht: (2025)
von: Pranida, Salsabila Zahirah, et al.
Veröffentlicht: (2025)
Cross-Cultural Transfer of Commonsense Reasoning in LLMs: Evidence from the Arab World
von: Almheiri, Saeed, et al.
Veröffentlicht: (2025)
von: Almheiri, Saeed, et al.
Veröffentlicht: (2025)
Qorgau: Evaluating LLM Safety in Kazakh-Russian Bilingual Contexts
von: Goloburda, Maiya, et al.
Veröffentlicht: (2025)
von: Goloburda, Maiya, et al.
Veröffentlicht: (2025)
Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2026)
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2026)
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2025)
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2025)
Cendol: Open Instruction-tuned Generative Large Language Models for Indonesian Languages
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2024)
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2024)
Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
Zero-shot Sentiment Analysis in Low-Resource Languages Using a Multilingual Sentiment Lexicon
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
Language Surgery in Multilingual Large Language Models
von: Lopo, Joanito Agili, et al.
Veröffentlicht: (2025)
von: Lopo, Joanito Agili, et al.
Veröffentlicht: (2025)
Listen, Correct, and Feed Back: Spoken Pedagogical Feedback Generation
von: Liang, Junhong, et al.
Veröffentlicht: (2026)
von: Liang, Junhong, et al.
Veröffentlicht: (2026)
AdaCultureSafe: Adaptive Cultural Safety Grounded by Cultural Knowledge in Large Language Models
von: Kang, Hankun, et al.
Veröffentlicht: (2026)
von: Kang, Hankun, et al.
Veröffentlicht: (2026)
XL-SafetyBench: A Country-Grounded Cross-Cultural Benchmark for LLM Safety and Cultural Sensitivity
von: Choi, Dasol, et al.
Veröffentlicht: (2026)
von: Choi, Dasol, et al.
Veröffentlicht: (2026)
IndoBERT-Sentiment: Context-Conditioned Sentiment Classification for Indonesian Text
von: Saputra, Muhammad Apriandito Arya, et al.
Veröffentlicht: (2026)
von: Saputra, Muhammad Apriandito Arya, et al.
Veröffentlicht: (2026)
UbuntuGuard: A Culturally-Grounded Policy Benchmark for Equitable AI Safety in African Languages
von: Abdullahi, Tassallah, et al.
Veröffentlicht: (2026)
von: Abdullahi, Tassallah, et al.
Veröffentlicht: (2026)
ThaiSafetyBench: Assessing Language Model Safety in Thai Cultural Contexts
von: Ukarapol, Trapoom, et al.
Veröffentlicht: (2026)
von: Ukarapol, Trapoom, et al.
Veröffentlicht: (2026)
IndoBERT-Relevancy: A Context-Conditioned Relevancy Classifier for Indonesian Text
von: Saputra, Muhammad Apriandito Arya, et al.
Veröffentlicht: (2026)
von: Saputra, Muhammad Apriandito Arya, et al.
Veröffentlicht: (2026)
Stuttering-Aware Automatic Speech Recognition for Indonesian Language
von: Muhammad, Fadhil, et al.
Veröffentlicht: (2026)
von: Muhammad, Fadhil, et al.
Veröffentlicht: (2026)
Assessing Socio-Cultural Alignment and Technical Safety of Sovereign LLMs
von: Chae, Kyubyung, et al.
Veröffentlicht: (2025)
von: Chae, Kyubyung, et al.
Veröffentlicht: (2025)
Commonsense Reasoning in Arab Culture
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Preserving Fairness and Safety in Quantized LLMs Through Critical Weight Protection
von: Hakim, Muhammad Alif Al, et al.
Veröffentlicht: (2026) -
Simulating Training Data Leakage in Multiple-Choice Benchmarks for LLM Evaluation
von: Hidayat, Naila Shafirni, et al.
Veröffentlicht: (2025) -
IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages
von: Hanif, Ikhlasul Akmal, et al.
Veröffentlicht: (2026) -
Parallel Tokenizers: Rethinking Vocabulary Design for Cross-Lingual Transfer
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025) -
What Do Indonesians Really Need from Language Technology? A Nationwide Survey
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)