IndicMMLU-Pro: Benchmarking Indic Large Language Models on Multi-Task Language Understanding
Fuente:
arXiv
Salvato in:
| Autori principali: | KJ, Sankalp, Kumar, Ashutosh, Balaji, Laxmaan, Kotecha, Nikunj, Jain, Vinija, Chadha, Aman, Bhaduri, Sreyoshi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Decoding the Diversity: A Review of the Indic AI Research Landscape
di: KJ, Sankalp, et al.
Pubblicazione: (2024)
di: KJ, Sankalp, et al.
Pubblicazione: (2024)
Hierarchical Prompting Taxonomy: A Universal Evaluation Framework for Large Language Models Aligned with Human Cognitive Principles
di: Budagam, Devichand, et al.
Pubblicazione: (2024)
di: Budagam, Devichand, et al.
Pubblicazione: (2024)
Multilingual State Space Models for Structured Question Answering in Indic Languages
di: Vats, Arpita, et al.
Pubblicazione: (2025)
di: Vats, Arpita, et al.
Pubblicazione: (2025)
Parameter Efficient Fine Tuning: A Comprehensive Analysis Across Applications
di: Balne, Charith Chandra Sai, et al.
Pubblicazione: (2024)
di: Balne, Charith Chandra Sai, et al.
Pubblicazione: (2024)
Simulating Meaning, Nevermore! Introducing ICR: A Semiotic-Hermeneutic Metric for Evaluating Meaning in LLM Text Summaries
di: Perez, Natalie, et al.
Pubblicazione: (2026)
di: Perez, Natalie, et al.
Pubblicazione: (2026)
MILU: A Multi-task Indic Language Understanding Benchmark
di: Verma, Sshubam, et al.
Pubblicazione: (2024)
di: Verma, Sshubam, et al.
Pubblicazione: (2024)
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications?
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
IndicParam: Benchmark to evaluate LLMs on low-resource Indic Languages
di: Maheshwari, Ayush, et al.
Pubblicazione: (2025)
di: Maheshwari, Ayush, et al.
Pubblicazione: (2025)
Guiding Vision-Language Model Selection for Visual Question-Answering Across Tasks, Domains, and Knowledge Types
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
IndicIFEval: A Benchmark for Verifiable Instruction-Following Evaluation in 14 Indic Languages
di: Jayakumar, Thanmay, et al.
Pubblicazione: (2026)
di: Jayakumar, Thanmay, et al.
Pubblicazione: (2026)
MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark
di: Wang, Yubo, et al.
Pubblicazione: (2024)
di: Wang, Yubo, et al.
Pubblicazione: (2024)
IndicGenBench: A Multilingual Benchmark to Evaluate Generation Capabilities of LLMs on Indic Languages
di: Singh, Harman, et al.
Pubblicazione: (2024)
di: Singh, Harman, et al.
Pubblicazione: (2024)
Born With a Silver Spoon? Investigating Socioeconomic Bias in Large Language Models
di: Singh, Smriti, et al.
Pubblicazione: (2024)
di: Singh, Smriti, et al.
Pubblicazione: (2024)
IndicEval-XL: Bridging Linguistic Diversity in Code Generation Across Indic Languages
di: Singh, Ujjwal, et al.
Pubblicazione: (2025)
di: Singh, Ujjwal, et al.
Pubblicazione: (2025)
Statistical Machine Translation for Indic Languages
di: Das, Sudhansu Bala, et al.
Pubblicazione: (2023)
di: Das, Sudhansu Bala, et al.
Pubblicazione: (2023)
IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languages
di: Nigam, Shubham Kumar, et al.
Pubblicazione: (2026)
di: Nigam, Shubham Kumar, et al.
Pubblicazione: (2026)
A Comprehensive Survey of Accelerated Generation Techniques in Large Language Models
di: Khoshnoodi, Mahsa, et al.
Pubblicazione: (2024)
di: Khoshnoodi, Mahsa, et al.
Pubblicazione: (2024)
IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages
di: Dawar, Aviral, et al.
Pubblicazione: (2026)
di: Dawar, Aviral, et al.
Pubblicazione: (2026)
Analysis of Indic Language Capabilities in LLMs
di: Vaidya, Aatman, et al.
Pubblicazione: (2025)
di: Vaidya, Aatman, et al.
Pubblicazione: (2025)
IndicSQuAD: A Comprehensive Multilingual Question Answering Dataset for Indic Languages
di: Endait, Sharvi, et al.
Pubblicazione: (2025)
di: Endait, Sharvi, et al.
Pubblicazione: (2025)
Neural FOXP2 -- Language Specific Neuron Steering for Targeted Language Improvement in LLMs
di: Saha, Anusa, et al.
Pubblicazione: (2026)
di: Saha, Anusa, et al.
Pubblicazione: (2026)
Improving Multilingual Neural Machine Translation System for Indic Languages
di: Das, Sudhansu Bala, et al.
Pubblicazione: (2022)
di: Das, Sudhansu Bala, et al.
Pubblicazione: (2022)
IndicRAGSuite: Large-Scale Datasets and a Benchmark for Indian Language RAG Systems
di: Prasanjith, Pasunuti, et al.
Pubblicazione: (2025)
di: Prasanjith, Pasunuti, et al.
Pubblicazione: (2025)
Safer in Translation? Presupposition Robustness in Indic Languages
di: Palnitkar, Aadi, et al.
Pubblicazione: (2025)
di: Palnitkar, Aadi, et al.
Pubblicazione: (2025)
Unicode Normalization and Grapheme Parsing of Indic Languages
di: Ansary, Nazmuddoha, et al.
Pubblicazione: (2023)
di: Ansary, Nazmuddoha, et al.
Pubblicazione: (2023)
How Culturally Aware are Vision-Language Models?
di: Burda-Lassen, Olena, et al.
Pubblicazione: (2024)
di: Burda-Lassen, Olena, et al.
Pubblicazione: (2024)
A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
di: Sahoo, Pranab, et al.
Pubblicazione: (2024)
Multi-Stage Training for Abusive Comment Detection in Indic Languages
di: Rastogi, Pranshu, et al.
Pubblicazione: (2026)
di: Rastogi, Pranshu, et al.
Pubblicazione: (2026)
Pralekha: Cross-Lingual Document Alignment for Indic Languages
di: Suryanarayanan, Sanjay, et al.
Pubblicazione: (2024)
di: Suryanarayanan, Sanjay, et al.
Pubblicazione: (2024)
Table Question Answering for Low-resourced Indic Languages
di: Pal, Vaishali, et al.
Pubblicazione: (2024)
di: Pal, Vaishali, et al.
Pubblicazione: (2024)
Overview of the 2023 ICON Shared Task on Gendered Abuse Detection in Indic Languages
di: Vaidya, Aatman, et al.
Pubblicazione: (2024)
di: Vaidya, Aatman, et al.
Pubblicazione: (2024)
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages?
di: Aravapalli, Akhilesh, et al.
Pubblicazione: (2024)
di: Aravapalli, Akhilesh, et al.
Pubblicazione: (2024)
Potential and Perils of Large Language Models as Judges of Unstructured Textual Data
di: Bedemariam, Rewina, et al.
Pubblicazione: (2025)
di: Bedemariam, Rewina, et al.
Pubblicazione: (2025)
L3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic Languages
di: Mirashi, Aishwarya, et al.
Pubblicazione: (2024)
di: Mirashi, Aishwarya, et al.
Pubblicazione: (2024)
MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
di: Xuan, Weihao, et al.
Pubblicazione: (2025)
di: Xuan, Weihao, et al.
Pubblicazione: (2025)
From Prejudice to Parity: A New Approach to Debiasing Large Language Model Word Embeddings
di: Rakshit, Aishik, et al.
Pubblicazione: (2024)
di: Rakshit, Aishik, et al.
Pubblicazione: (2024)
Exploring the Frontier of Vision-Language Models: A Survey of Current Methodologies and Future Directions
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
LLMsAgainstHate @ NLU of Devanagari Script Languages 2025: Hate Speech Detection and Target Identification in Devanagari Languages via Parameter Efficient Fine-Tuning of LLMs
di: Sidibomma, Rushendra, et al.
Pubblicazione: (2024)
di: Sidibomma, Rushendra, et al.
Pubblicazione: (2024)
Breaking Down the Defenses: A Comparative Survey of Attacks on Large Language Models
di: Chowdhury, Arijit Ghosh, et al.
Pubblicazione: (2024)
di: Chowdhury, Arijit Ghosh, et al.
Pubblicazione: (2024)
Personality Shapes Gender Bias in Persona-Conditioned LLM Narratives Across English and Hindi: An Empirical Investigation
di: Kumar, Tanay, et al.
Pubblicazione: (2026)
di: Kumar, Tanay, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Decoding the Diversity: A Review of the Indic AI Research Landscape
di: KJ, Sankalp, et al.
Pubblicazione: (2024) -
Hierarchical Prompting Taxonomy: A Universal Evaluation Framework for Large Language Models Aligned with Human Cognitive Principles
di: Budagam, Devichand, et al.
Pubblicazione: (2024) -
Multilingual State Space Models for Structured Question Answering in Indic Languages
di: Vats, Arpita, et al.
Pubblicazione: (2025) -
Parameter Efficient Fine Tuning: A Comprehensive Analysis Across Applications
di: Balne, Charith Chandra Sai, et al.
Pubblicazione: (2024) -
Simulating Meaning, Nevermore! Introducing ICR: A Semiotic-Hermeneutic Metric for Evaluating Meaning in LLM Text Summaries
di: Perez, Natalie, et al.
Pubblicazione: (2026)