Batayan: A Filipino NLP benchmark for evaluating Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Montalan, Jann Railey, Layacan, Jimson Paulo, Africa, David Demitri, Flores, Richell Isaiah, Lopez II, Michael T., Magsajo, Theresa Denise, Cayabyab, Anjanette, Tjhi, William Chandra |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Kalahi: A handcrafted, grassroots cultural LLM evaluation suite for Filipino
von: Montalan, Jann Railey, et al.
Veröffentlicht: (2024)
von: Montalan, Jann Railey, et al.
Veröffentlicht: (2024)
BURMESE-SAN: Burmese NLP Benchmark for Evaluating Large Language Models
von: Aung, Thura, et al.
Veröffentlicht: (2026)
von: Aung, Thura, et al.
Veröffentlicht: (2026)
SEA-HELM: Southeast Asian Holistic Evaluation of Language Models
von: Susanto, Yosephine, et al.
Veröffentlicht: (2025)
von: Susanto, Yosephine, et al.
Veröffentlicht: (2025)
SEA-BED: How Do Embedding Models Represent Southeast Asian Languages?
von: Ponwitayarat, Wuttikorn, et al.
Veröffentlicht: (2025)
von: Ponwitayarat, Wuttikorn, et al.
Veröffentlicht: (2025)
Identifying a Circuit for Verb Conjugation in GPT-2
von: Africa, David Demitri
Veröffentlicht: (2025)
von: Africa, David Demitri
Veröffentlicht: (2025)
LURE: Live-Usage Replay Evaluations for Reducing Evaluation Awareness
von: Ivanov, Igor, et al.
Veröffentlicht: (2026)
von: Ivanov, Igor, et al.
Veröffentlicht: (2026)
Steering Awareness: Detecting Activation Steering from Within
von: Rivera, Joshua Fonseca, et al.
Veröffentlicht: (2025)
von: Rivera, Joshua Fonseca, et al.
Veröffentlicht: (2025)
Does Self-Evaluation Enable Wireheading in Language Models?
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
Operaciones garantizadas internacionales: debate actual y posibles soluciones
von: Anjanette H. Raymond
Veröffentlicht: (2011)
von: Anjanette H. Raymond
Veröffentlicht: (2011)
Consistency Training while Mitigating Obfuscation via Rate Matching
von: Imran, Sohaib, et al.
Veröffentlicht: (2026)
von: Imran, Sohaib, et al.
Veröffentlicht: (2026)
Learning Dynamics of Meta-Learning in Small Model Pretraining
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
Investigating ReLoRA: Effects on the Learning Dynamics of Small Language Models
von: Weiss, Yuval, et al.
Veröffentlicht: (2025)
von: Weiss, Yuval, et al.
Veröffentlicht: (2025)
Learning Modular Exponentiation with Transformers
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
Training Vision-Language Process Reward Models for Test-Time Scaling in Multimodal Reasoning: Key Insights and Lessons Learned
von: Ong, Brandon, et al.
Veröffentlicht: (2025)
von: Ong, Brandon, et al.
Veröffentlicht: (2025)
Meta-Pretraining for Zero-Shot Cross-Lingual Named Entity Recognition in Low-Resource Philippine Languages
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
von: Africa, David Demitri, et al.
Veröffentlicht: (2025)
A Reinforcement Learning Inspired Latent Yield Based Adaptive Algorithm Switching Mechanism
von: Nair, Jayprakash S., et al.
Veröffentlicht: (2026)
von: Nair, Jayprakash S., et al.
Veröffentlicht: (2026)
Large Language Models for EEG: A Comprehensive Survey and Taxonomy
von: Babu, Naseem, et al.
Veröffentlicht: (2025)
von: Babu, Naseem, et al.
Veröffentlicht: (2025)
ThaiCoref: Thai Coreference Resolution Dataset
von: Trakuekul, Pontakorn, et al.
Veröffentlicht: (2024)
von: Trakuekul, Pontakorn, et al.
Veröffentlicht: (2024)
FiLLM -- A Filipino-optimized Large Language Model based on Southeast Asia Large Language Model (SEALLM)
von: Maminta, Carlos Jude G., et al.
Veröffentlicht: (2025)
von: Maminta, Carlos Jude G., et al.
Veröffentlicht: (2025)
Pico: A Modular Framework for Hypothesis-Driven Small Language Model Research
von: Martinez, Richard Diehl, et al.
Veröffentlicht: (2025)
von: Martinez, Richard Diehl, et al.
Veröffentlicht: (2025)
No Answer Needed: Predicting LLM Answer Accuracy from Question-Only Linear Probes
von: Cencerrado, Iván Vicente Moreno, et al.
Veröffentlicht: (2025)
von: Cencerrado, Iván Vicente Moreno, et al.
Veröffentlicht: (2025)
Inoculation Prompting: Eliciting traits from LLMs during training can suppress them at test-time
von: Tan, Daniel, et al.
Veröffentlicht: (2025)
von: Tan, Daniel, et al.
Veröffentlicht: (2025)
Thai Universal Dependency Treebank
von: Sriwirote, Panyut, et al.
Veröffentlicht: (2024)
von: Sriwirote, Panyut, et al.
Veröffentlicht: (2024)
Filipino labour in Hawaii
Veröffentlicht: (1927)
Veröffentlicht: (1927)
Reasoning Models Reason Well, Until They Don't
von: Rameshkumar, Revanth, et al.
Veröffentlicht: (2025)
von: Rameshkumar, Revanth, et al.
Veröffentlicht: (2025)
A single‐site feasibility randomised controlled trial comparing ‘my hypo compass’ short pyscho‐educational intervention with standard care alone in individuals with type 1 diabetes and impaired awareness of hypoglycaemia
von: Ayat Bashir, et al.
Veröffentlicht: (2024)
von: Ayat Bashir, et al.
Veröffentlicht: (2024)
A call for reporting of tumor‐specific outcomes in studies of DPYD genotyping
von: Jean De Dieu Ndayishimiye, et al.
Veröffentlicht: (2024)
von: Jean De Dieu Ndayishimiye, et al.
Veröffentlicht: (2024)
Compositional Analysis of Cultivated and Wild‐Harvested Boswellia sacra Frankincense Resin Essential Oils in Oman
von: Anjanette DeCarlo, et al.
Veröffentlicht: (2025)
von: Anjanette DeCarlo, et al.
Veröffentlicht: (2025)
Tapping into the Assets of First-Generation Students during Times of Transition
von: Hands, Africa S.
Veröffentlicht: (2020)
von: Hands, Africa S.
Veröffentlicht: (2020)
What's Your Type? An Examination of First-Year Doctoral Student Motivation
von: Hands, Africa S.
Veröffentlicht: (2020)
von: Hands, Africa S.
Veröffentlicht: (2020)
Public Libraries: Your Partner in Increasing College Literacy among Nontraditional Prospective Students
von: Hands, Africa S.
Veröffentlicht: (2023)
von: Hands, Africa S.
Veröffentlicht: (2023)
Successfully Serving the College Bound
von: Hands, Africa S.
Veröffentlicht: (2015)
von: Hands, Africa S.
Veröffentlicht: (2015)
Peer Genius Bar: Using the Wisdom of the Crowd to Learn Technology Tools
von: Hands, Africa S.
Veröffentlicht: (2023)
von: Hands, Africa S.
Veröffentlicht: (2023)
What Doctoral Student Motivation Tells Us about the Future of LIS Education
von: Hands, Africa S.
Veröffentlicht: (2018)
von: Hands, Africa S.
Veröffentlicht: (2018)
Voluntarism and Political Conflict in Barbados, 1814–33*
von: Isaiah Silvers
Veröffentlicht: (2026)
von: Isaiah Silvers
Veröffentlicht: (2026)
Antología de ensayos / Isaiah Berlin ; introducción Joaquín Abell n
von: Berlin, Isaiah
von: Berlin, Isaiah
Qué es la libertad política ?
von: Isaiah, Berlin
Veröffentlicht: (2006)
von: Isaiah, Berlin
Veröffentlicht: (2006)
The GATT, quantitative restrictions, and the balance of payments / Isaiah Frank
von: Frank, Isaiah
Veröffentlicht: (1987)
von: Frank, Isaiah
Veröffentlicht: (1987)
Cuatro ensayos sobre la libertad / Isaiah Berlin
von: Berlin, Isaiah
von: Berlin, Isaiah
Decadencia de las ideas utópicas en Occidente
von: Berlin, Isaiah
Veröffentlicht: (1986)
von: Berlin, Isaiah
Veröffentlicht: (1986)
Ähnliche Einträge
-
Kalahi: A handcrafted, grassroots cultural LLM evaluation suite for Filipino
von: Montalan, Jann Railey, et al.
Veröffentlicht: (2024) -
BURMESE-SAN: Burmese NLP Benchmark for Evaluating Large Language Models
von: Aung, Thura, et al.
Veröffentlicht: (2026) -
SEA-HELM: Southeast Asian Holistic Evaluation of Language Models
von: Susanto, Yosephine, et al.
Veröffentlicht: (2025) -
SEA-BED: How Do Embedding Models Represent Southeast Asian Languages?
von: Ponwitayarat, Wuttikorn, et al.
Veröffentlicht: (2025) -
Identifying a Circuit for Verb Conjugation in GPT-2
von: Africa, David Demitri
Veröffentlicht: (2025)