The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants
Fuente:
arXiv
Saved in:
| Main Authors: | Bandarkar, Lucas, Liang, Davis, Muller, Benjamin, Artetxe, Mikel, Shukla, Satya Narayan, Husa, Donald, Goyal, Naman, Krishnan, Abhinandan, Zettlemoyer, Luke, Khabsa, Madian |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
High Accuracy, Less Talk (HALT): Reliable LLMs through Capability-Aligned Finetuning
by: Franzmeyer, Tim, et al.
Published: (2025)
by: Franzmeyer, Tim, et al.
Published: (2025)
The Unreasonable Effectiveness of Model Merging for Cross-Lingual Transfer in LLMs
by: Bandarkar, Lucas, et al.
Published: (2025)
by: Bandarkar, Lucas, et al.
Published: (2025)
Emergent Abilities of Large Language Models under Continued Pretraining for Language Adaptation
by: Elhady, Ahmed, et al.
Published: (2025)
by: Elhady, Ahmed, et al.
Published: (2025)
WiCkeD: A Simple Method to Make Multiple Choice Benchmarks More Challenging
by: Elhady, Ahmed, et al.
Published: (2025)
by: Elhady, Ahmed, et al.
Published: (2025)
Cross-lingual Self-Consistency for Multilingual Reasoning with Language Models
by: Elhady, Ahmed, et al.
Published: (2026)
by: Elhady, Ahmed, et al.
Published: (2026)
Heteroscedastic Temporal Variational Autoencoder For Irregular Time Series
by: Shukla, Satya Narayan, et al.
Published: (2021)
by: Shukla, Satya Narayan, et al.
Published: (2021)
A comprehensive study of on-device NLP applications -- VQA, automated Form filling, Smart Replies for Linguistic Codeswitching
by: Goyal, Naman
Published: (2024)
by: Goyal, Naman
Published: (2024)
Cultural and Historical Identity in Amitav Ghosh’s River of Smoke: A Postcolonial Perspective
by: Satya Narayan
Published: (2021)
by: Satya Narayan
Published: (2021)
Depicting Culture and Identity in Amitav Ghosh’s The Shadow Lines
by: Satya Narayan
Published: (2017)
by: Satya Narayan
Published: (2017)
Cultural and Historical Identity in Amitav Ghosh’s River of Smoke: A Postcolonial Perspective
by: Satya Narayan
Published: (2021)
by: Satya Narayan
Published: (2021)
Large Reasoning Models Struggle to Transfer Parametric Knowledge Across Scripts
by: Bandarkar, Lucas, et al.
Published: (2026)
by: Bandarkar, Lucas, et al.
Published: (2026)
Knowledge Localization in Mixture-of-Experts LLMs Using Cross-Lingual Inconsistency
by: Bandarkar, Lucas, et al.
Published: (2026)
by: Bandarkar, Lucas, et al.
Published: (2026)
IndiaWeatherBench: A Dataset and Benchmark for Data-Driven Regional Weather Forecasting over India
by: Nguyen, Tung, et al.
Published: (2025)
by: Nguyen, Tung, et al.
Published: (2025)
Diversity-driven Data Selection for Language Model Tuning through Sparse Autoencoder
by: Yang, Xianjun, et al.
Published: (2025)
by: Yang, Xianjun, et al.
Published: (2025)
Extracting Small Translation Specialists from LLMs by Aggressively Pruning Experts
by: Martin, Liu O., et al.
Published: (2026)
by: Martin, Liu O., et al.
Published: (2026)
Parallel Graph Drawing Algorithm for Bipartite Planar Graphs
by: Jain, Naman
Published: (2024)
by: Jain, Naman
Published: (2024)
Beyond Reasoning Gains: Mitigating General Capabilities Forgetting in Large Reasoning Models
by: Phan, Hoang, et al.
Published: (2025)
by: Phan, Hoang, et al.
Published: (2025)
Crystalline representations and Wach modules in the imperfect residue field case
by: Abhinandan
Published: (2023)
by: Abhinandan
Published: (2023)
Syntomic complex and $p$-adic nearby cycles
by: Abhinandan
Published: (2023)
by: Abhinandan
Published: (2023)
Crystalline part of the Galois cohomology of crystalline representations
by: Abhinandan
Published: (2024)
by: Abhinandan
Published: (2024)
Crystalline representations and Wach modules in the relative case II
by: Abhinandan
Published: (2023)
by: Abhinandan
Published: (2023)
Crystalline representations and Wach modules in the relative case
by: Abhinandan
Published: (2021)
by: Abhinandan
Published: (2021)
Prismatic $F$-crystals and Wach modules
by: Abhinandan
Published: (2024)
by: Abhinandan
Published: (2024)
Debate Helps Weak Judges Reward Stronger Models
by: Elasky, Ethan, et al.
Published: (2026)
by: Elasky, Ethan, et al.
Published: (2026)
BertaQA: How Much Do Language Models Know About Local Culture?
by: Etxaniz, Julen, et al.
Published: (2024)
by: Etxaniz, Julen, et al.
Published: (2024)
Gender-specific Machine Translation with Large Language Models
by: Sánchez, Eduardo, et al.
Published: (2023)
by: Sánchez, Eduardo, et al.
Published: (2023)
Learning to Localize Objects Improves Spatial Reasoning in Visual-LLMs
by: Ranasinghe, Kanchana, et al.
Published: (2024)
by: Ranasinghe, Kanchana, et al.
Published: (2024)
Time series forecasting with high stakes: A field study of the air cargo industry
by: Garg, Abhinav, et al.
Published: (2024)
by: Garg, Abhinav, et al.
Published: (2024)
Comment on ‘The Relationship Between Quality of Discharge Teaching and Oral Nutritional Supplementation Adherence in Postoperative Patients With Gastric Cancer: A Chain Mediated Role of Readiness for Hospital Discharge and Medication Beliefs’
by: Abhinandan Patil
Published: (2026)
by: Abhinandan Patil
Published: (2026)
THE DISINHERITED: THE POLITICS OF CHRISTIAN CONVERSTION IN COLONIAL INDIA. By MouBanerjee. Cambridge: Harvard University Press. Pp. 368. Cloth $49.95.
by: Abhinandan Banerjee
Published: (2025)
by: Abhinandan Banerjee
Published: (2025)
Banana fibers camouflaging as a gut worm in a 6-month-old infant
by: Abhinandan Patil
Published: (2020)
by: Abhinandan Patil
Published: (2020)
Multilingual Routing in Mixture-of-Experts
by: Bandarkar, Lucas, et al.
Published: (2025)
by: Bandarkar, Lucas, et al.
Published: (2025)
Computational Tradeoffs in Image Synthesis: Diffusion, Masked-Token, and Next-Token Prediction
by: Kilian, Maciej, et al.
Published: (2024)
by: Kilian, Maciej, et al.
Published: (2024)
Comparing Hallucination Detection Metrics for Multilingual Generation
by: Kang, Haoqiang, et al.
Published: (2024)
by: Kang, Haoqiang, et al.
Published: (2024)
Multimodal RewardBench: Holistic Evaluation of Reward Models for Vision Language Models
by: Yasunaga, Michihiro, et al.
Published: (2025)
by: Yasunaga, Michihiro, et al.
Published: (2025)
(Mis)Fitting: A Survey of Scaling Laws
by: Li, Margaret, et al.
Published: (2025)
by: Li, Margaret, et al.
Published: (2025)
On the Equivalence of Regression and Classification
by: Jayadeva, et al.
Published: (2025)
by: Jayadeva, et al.
Published: (2025)
Tests of general relativity in pseudo-Newtonian approach
by: Goyal, Naman, et al.
Published: (2026)
by: Goyal, Naman, et al.
Published: (2026)
Preference Optimization with Multi-Sample Comparisons
by: Wang, Chaoqi, et al.
Published: (2024)
by: Wang, Chaoqi, et al.
Published: (2024)
Linguini: A benchmark for language-agnostic linguistic reasoning
by: Sánchez, Eduardo, et al.
Published: (2024)
by: Sánchez, Eduardo, et al.
Published: (2024)
Similar Items
-
High Accuracy, Less Talk (HALT): Reliable LLMs through Capability-Aligned Finetuning
by: Franzmeyer, Tim, et al.
Published: (2025) -
The Unreasonable Effectiveness of Model Merging for Cross-Lingual Transfer in LLMs
by: Bandarkar, Lucas, et al.
Published: (2025) -
Emergent Abilities of Large Language Models under Continued Pretraining for Language Adaptation
by: Elhady, Ahmed, et al.
Published: (2025) -
WiCkeD: A Simple Method to Make Multiple Choice Benchmarks More Challenging
by: Elhady, Ahmed, et al.
Published: (2025) -
Cross-lingual Self-Consistency for Multilingual Reasoning with Language Models
by: Elhady, Ahmed, et al.
Published: (2026)