NusaAksara: A Multimodal and Multilingual Benchmark for Preserving Indonesian Indigenous Scripts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Adilazuarda, Muhammad Farid, Wijanarko, Musa Izzanardi, Susanto, Lucky, Nur'aini, Khumaisa, Wijaya, Derry, Aji, Alham Fikri |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Does Visual Rendering Bypass Tokenization? Investigating Script-Tokenizer Misalignment in Pixel-Based Language Models
von: Susanto, Lucky, et al.
Veröffentlicht: (2026)
von: Susanto, Lucky, et al.
Veröffentlicht: (2026)
Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation
von: Nur'aini, Khumaisa, et al.
Veröffentlicht: (2026)
von: Nur'aini, Khumaisa, et al.
Veröffentlicht: (2026)
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
von: Susanto, Lucky, et al.
Veröffentlicht: (2024)
A Multi-Labeled Dataset for Indonesian Discourse: Examining Toxicity, Polarization, and Demographics Information
von: Susanto, Lucky, et al.
Veröffentlicht: (2025)
von: Susanto, Lucky, et al.
Veröffentlicht: (2025)
LoraxBench: A Multitask, Multilingual Benchmark Suite for 20 Indonesian Languages
von: Aji, Alham Fikri, et al.
Veröffentlicht: (2025)
von: Aji, Alham Fikri, et al.
Veröffentlicht: (2025)
Environmental carrying capacity and environmental capacity in the issuance of recommendation and borrow-use permit of forest area
von: Wiwin, Nur'aini
Veröffentlicht: (2019)
von: Wiwin, Nur'aini
Veröffentlicht: (2019)
What Do Indonesians Really Need from Language Technology? A Nationwide Survey
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
From Surveys to Narratives: Rethinking Cultural Value Adaptation in LLMs
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2025)
Could We Have Had Better Multilingual LLMs If English Was Not the Central Language?
von: Diandaru, Ryandito, et al.
Veröffentlicht: (2024)
von: Diandaru, Ryandito, et al.
Veröffentlicht: (2024)
MLKV: Multi-Layer Key-Value Heads for Memory Efficient Transformer Decoding
von: Zuhri, Zayd Muhammad Kawakibi, et al.
Veröffentlicht: (2024)
von: Zuhri, Zayd Muhammad Kawakibi, et al.
Veröffentlicht: (2024)
LinguAlchemy: Fusing Typological and Geographical Elements for Unseen Language Generalization
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
QLESS: A Quantized Approach for Data Valuation and Selection in Large Language Model Fine-Tuning
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
von: Ananta, Moses, et al.
Veröffentlicht: (2025)
Cultural Conditioning or Placebo? On the Effectiveness of Socio-Demographic Prompting
von: Mukherjee, Sagnik, et al.
Veröffentlicht: (2024)
von: Mukherjee, Sagnik, et al.
Veröffentlicht: (2024)
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2025)
The Privileged Students: On the Value of Initialization in Multilingual Knowledge Distillation
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2024)
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2024)
Predicting LLM Correctness in Prosthodontics Using Metadata and Hallucination Signals
von: Susanto, Lucky, et al.
Veröffentlicht: (2025)
von: Susanto, Lucky, et al.
Veröffentlicht: (2025)
Improving Low-Resource Machine Translation via Round-Trip Reinforcement Learning
von: Attia, Ahmed, et al.
Veröffentlicht: (2026)
von: Attia, Ahmed, et al.
Veröffentlicht: (2026)
Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition
von: Chevi, Rendi, et al.
Veröffentlicht: (2024)
von: Chevi, Rendi, et al.
Veröffentlicht: (2024)
Do Language Models Understand Honorific Systems in Javanese?
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2025)
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2025)
Multilinguality as Sense Adaptation
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
Can Multimodal LLMs do Visual Temporal Understanding and Reasoning? The answer is No!
von: Imam, Mohamed Fazli, et al.
Veröffentlicht: (2025)
von: Imam, Mohamed Fazli, et al.
Veröffentlicht: (2025)
MetaMetrics: Calibrating Metrics For Generation Tasks Using Human Preferences
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
von: Winata, Genta Indra, et al.
Veröffentlicht: (2024)
MetaMetrics-MT: Tuning Meta-Metrics for Machine Translation via Human Preference Calibration
von: Anugraha, David, et al.
Veröffentlicht: (2024)
von: Anugraha, David, et al.
Veröffentlicht: (2024)
Multicultural Spyfall: Assessing LLMs through Dynamic Multilingual Social Deduction Game
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2026)
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2026)
Towards Measuring and Modeling "Culture" in LLMs: A Survey
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
von: Adilazuarda, Muhammad Farid, et al.
Veröffentlicht: (2024)
Data Laundering: Artificially Boosting Benchmark Results through Knowledge Distillation
von: Mansurov, Jonibek, et al.
Veröffentlicht: (2024)
von: Mansurov, Jonibek, et al.
Veröffentlicht: (2024)
Sense Representations Are Inducible Interfaces
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
LLM Olympiad: Why Model Evaluation Needs a Sealed Exam
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2026)
Extracting General-use Transformers for Low-resource Languages via Knowledge Distillation
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2025)
von: Cruz, Jan Christian Blaise, et al.
Veröffentlicht: (2025)
Beyond Probabilities: Unveiling the Misalignment in Evaluating Large Language Models
von: Lyu, Chenyang, et al.
Veröffentlicht: (2024)
von: Lyu, Chenyang, et al.
Veröffentlicht: (2024)
SEA-Guard: Culturally Grounded Multilingual Safeguard for Southeast Asia
von: Tasawong, Panuthep, et al.
Veröffentlicht: (2026)
von: Tasawong, Panuthep, et al.
Veröffentlicht: (2026)
COPAL-ID: Indonesian Language Reasoning with Local Culture and Nuances
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2023)
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2023)
Language-Specific Latent Process Hinders Cross-Lingual Performance
von: Lim, Zheng Wei, et al.
Veröffentlicht: (2025)
von: Lim, Zheng Wei, et al.
Veröffentlicht: (2025)
Aksara
Veröffentlicht: (2017)
Veröffentlicht: (2017)
Aksara
Veröffentlicht: (2021)
Veröffentlicht: (2021)
Beyond Turing: A Comparative Analysis of Approaches for Detecting Machine-Generated Text
von: Adilazuarda, Muhammad Farid
Veröffentlicht: (2023)
von: Adilazuarda, Muhammad Farid
Veröffentlicht: (2023)
DriveThru: a Document Extraction Platform and Benchmark Datasets for Indonesian Local Language Archives
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2024)
von: Farhansyah, Mohammad Rifqi, et al.
Veröffentlicht: (2024)
Efficient and Interpretable Grammatical Error Correction with Mixture of Experts
von: Qorib, Muhammad Reza, et al.
Veröffentlicht: (2024)
von: Qorib, Muhammad Reza, et al.
Veröffentlicht: (2024)
Enabling Natural Zero-Shot Prompting on Encoder Models via Statement-Tuning
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2024)
von: Elshabrawy, Ahmed, et al.
Veröffentlicht: (2024)
How Individual Traits and Language Styles Shape Preferences In Open-ended User-LLM Interaction: A Preliminary Study
von: Chevi, Rendi, et al.
Veröffentlicht: (2025)
von: Chevi, Rendi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Does Visual Rendering Bypass Tokenization? Investigating Script-Tokenizer Misalignment in Pixel-Based Language Models
von: Susanto, Lucky, et al.
Veröffentlicht: (2026) -
Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation
von: Nur'aini, Khumaisa, et al.
Veröffentlicht: (2026) -
IndoToxic2024: A Demographically-Enriched Dataset of Hate Speech and Toxicity Types for Indonesian Language
von: Susanto, Lucky, et al.
Veröffentlicht: (2024) -
A Multi-Labeled Dataset for Indonesian Discourse: Examining Toxicity, Polarization, and Demographics Information
von: Susanto, Lucky, et al.
Veröffentlicht: (2025) -
LoraxBench: A Multitask, Multilingual Benchmark Suite for 20 Indonesian Languages
von: Aji, Alham Fikri, et al.
Veröffentlicht: (2025)