LoraxBench: A Multitask, Multilingual Benchmark Suite for 20 Indonesian Languages
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Aji, Alham Fikri, Cohn, Trevor |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Language-Specific Latent Process Hinders Cross-Lingual Performance
von: Lim, Zheng Wei, et al.
Veröffentlicht: (2025)
von: Lim, Zheng Wei, et al.
Veröffentlicht: (2025)
SEA-SafeguardBench: Evaluating AI Safety in SEA Languages and Cultures
von: Tasawong, Panuthep, et al.
Veröffentlicht: (2025)
von: Tasawong, Panuthep, et al.
Veröffentlicht: (2025)
SEA-Guard: Culturally Grounded Multilingual Safeguard for Southeast Asia
von: Tasawong, Panuthep, et al.
Veröffentlicht: (2026)
von: Tasawong, Panuthep, et al.
Veröffentlicht: (2026)
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
NusaAksara: A Multimodal and Multilingual Benchmark for Preserving Indonesian Indigenous Scripts
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025)
The Privileged Students: On the Value of Initialization in Multilingual Knowledge Distillation
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2024)
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2024)
Multilinguality as Sense Adaptation
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
Improving Low-Resource Machine Translation via Round-Trip Reinforcement Learning
von: Attia, Ahmed, et al.
Veröffentlicht: (2026)
von: Attia, Ahmed, et al.
Veröffentlicht: (2026)
Data Laundering: Artificially Boosting Benchmark Results through Knowledge Distillation
von: Mansurov, Jonibek, et al.
Veröffentlicht: (2024)
von: Mansurov, Jonibek, et al.
Veröffentlicht: (2024)
Beyond Probabilities: Unveiling the Misalignment in Evaluating Large Language Models
von: Lyu, Chenyang, et al.
Veröffentlicht: (2024)
von: Lyu, Chenyang, et al.
Veröffentlicht: (2024)
Extracting General-use Transformers for Low-resource Languages via Knowledge Distillation
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2025)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2025)
Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition
von: Chevi, Rendi, et al.
Veröffentlicht: (2024)
von: Chevi, Rendi, et al.
Veröffentlicht: (2024)
COPAL-ID: Indonesian Language Reasoning with Local Culture and Nuances
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2023)
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2023)
Multicultural Spyfall: Assessing LLMs through Dynamic Multilingual Social Deduction Game
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2026)
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2026)
How Individual Traits and Language Styles Shape Preferences In Open-ended User-LLM Interaction: A Preliminary Study
von: Chevi, Rendi, et al.
Veröffentlicht: (2025)
von: Chevi, Rendi, et al.
Veröffentlicht: (2025)
Sense Representations Are Inducible Interfaces
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
LLM Olympiad: Why Model Evaluation Needs a Sealed Exam
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
Macaron: Controlled, Human-Written Benchmark for Multilingual and Multicultural Reasoning via Template-Filling
von: Elsetohy, Alaa, et al.
Veröffentlicht: (2026)
von: Elsetohy, Alaa, et al.
Veröffentlicht: (2026)
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
Can Multimodal LLMs do Visual Temporal Understanding and Reasoning? The answer is No!
von: Imam, Mohamed Fazli, et al.
Veröffentlicht: (2025)
von: Imam, Mohamed Fazli, et al.
Veröffentlicht: (2025)
Efficient and Interpretable Grammatical Error Correction with Mixture of Experts
von: Qorib, Muhammad Reza, et al.
Veröffentlicht: (2024)
von: Qorib, Muhammad Reza, et al.
Veröffentlicht: (2024)
Enabling Natural Zero-Shot Prompting on Encoder Models via Statement-Tuning
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2024)
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2024)
WangchanThaiInstruct: An instruction-following Dataset for Culture-Aware, Multitask, and Multi-domain Evaluation in Thai
von: Limkonchotiwat, Peerat, et al.
Veröffentlicht: (2025)
von: Limkonchotiwat, Peerat, et al.
Veröffentlicht: (2025)
When Alignment Hurts: Decoupling Representational Spaces in Multilingual Models
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2025)
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2025)
TextGames: Learning to Self-Play Text-Based Puzzle Games via Language Model Reasoning
von: Hudi, Frederikus, et al.
Veröffentlicht: (2025)
von: Hudi, Frederikus, et al.
Veröffentlicht: (2025)
DialectalArabicMMLU: Benchmarking Dialectal Capabilities in Arabic and Multilingual Language Models
von: Altakrori, Malik H., et al.
Veröffentlicht: (2025)
von: Altakrori, Malik H., et al.
Veröffentlicht: (2025)
Language Surgery in Multilingual Large Language Models
von: Lopo, Joanito Agili, et al.
Veröffentlicht: (2025)
von: Lopo, Joanito Agili, et al.
Veröffentlicht: (2025)
Thank You, Stingray: Multilingual Large Language Models Can Not (Yet) Disambiguate Cross-Lingual Word Sense
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2024)
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2024)
LinguAlchemy: Fusing Typological and Geographical Elements for Unseen Language Generalization
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
From Surveys to Narratives: Rethinking Cultural Value Adaptation in LLMs
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025)
LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions
von: Wu, Minghao, et al.
Veröffentlicht: (2023)
von: Wu, Minghao, et al.
Veröffentlicht: (2023)
Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation
von: Nur'aini, Khumaisa, et al.
Veröffentlicht: (2026)
von: Nur'aini, Khumaisa, et al.
Veröffentlicht: (2026)
LinguDistill: Recovering Linguistic Ability in Vision-Language Models via Selective Cross-Modal Distillation
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2026)
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2026)
SEACrowd: A Multilingual Multimodal Data Hub and Benchmark Suite for Southeast Asian Languages
von: Lovenia, Holy, et al.
Veröffentlicht: (2024)
von: Lovenia, Holy, et al.
Veröffentlicht: (2024)
Do Language Models Understand Honorific Systems in Javanese?
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2025)
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2025)
QLESS: A Quantized Approach for Data Valuation and Selection in Large Language Model Fine-Tuning
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
BenchMAX: A Comprehensive Multilingual Evaluation Suite for Large Language Models
von: Huang, Xu, et al.
Veröffentlicht: (2025)
von: Huang, Xu, et al.
Veröffentlicht: (2025)
Cendol: Open Instruction-tuned Generative Large Language Models for Indonesian Languages
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2024)
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2024)
SEADialogues: A Multilingual Culturally Grounded Multi-turn Dialogue Dataset on Southeast Asian Languages
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
Cultural Conditioning or Placebo? On the Effectiveness of Socio-Demographic Prompting
von: Mukherjee, Sagnik, et al.
Veröffentlicht: (2024)
von: Mukherjee, Sagnik, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Language-Specific Latent Process Hinders Cross-Lingual Performance
von: Lim, Zheng Wei, et al.
Veröffentlicht: (2025) -
SEA-SafeguardBench: Evaluating AI Safety in SEA Languages and Cultures
von: Tasawong, Panuthep, et al.
Veröffentlicht: (2025) -
SEA-Guard: Culturally Grounded Multilingual Safeguard for Southeast Asia
von: Tasawong, Panuthep, et al.
Veröffentlicht: (2026) -
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025) -
NusaAksara: A Multimodal and Multilingual Benchmark for Preserving Indonesian Indigenous Scripts
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025)