IndicGEC: Powerful Models, or a Measurement Mirage?
Fuente:
arXiv
Guardado en:
| Autor principal: | Vajjala, Sowmya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Problem with Safety Classification is not just the Models
por: Vajjala, Sowmya
Publicado: (2025)
por: Vajjala, Sowmya
Publicado: (2025)
MATA: Mindful Assessment of the Telugu Abilities of Large Language Models
por: Kranti, Chalamalasetti, et al.
Publicado: (2025)
por: Kranti, Chalamalasetti, et al.
Publicado: (2025)
MetricalARGS: A Taxonomy for Studying Metrical Poetry with LLMs
por: Kranti, Chalamalasetti, et al.
Publicado: (2025)
por: Kranti, Chalamalasetti, et al.
Publicado: (2025)
Text Classification in the LLM Era -- Where do we stand?
por: Vajjala, Sowmya, et al.
Publicado: (2025)
por: Vajjala, Sowmya, et al.
Publicado: (2025)
Does Synthetic Data Help Named Entity Recognition for Low-Resource Languages?
por: Kamath, Gaurav, et al.
Publicado: (2025)
por: Kamath, Gaurav, et al.
Publicado: (2025)
Dravidian language family through Universal Dependencies lens
por: Rama, Taraka, et al.
Publicado: (2024)
por: Rama, Taraka, et al.
Publicado: (2024)
Minimal-Edit Instruction Tuning for Low-Resource Indic GEC
por: P, Akhil Rajeev
Publicado: (2025)
por: P, Akhil Rajeev
Publicado: (2025)
Annotation Errors and NER: A Study with OntoNotes 5.0
por: Bernier-Colborne, Gabriel, et al.
Publicado: (2024)
por: Bernier-Colborne, Gabriel, et al.
Publicado: (2024)
Scope Ambiguities in Large Language Models
por: Kamath, Gaurav, et al.
Publicado: (2024)
por: Kamath, Gaurav, et al.
Publicado: (2024)
Test Set Quality in Multilingual LLM Evaluation
por: Kranti, Chalamalasetti, et al.
Publicado: (2025)
por: Kranti, Chalamalasetti, et al.
Publicado: (2025)
Opportunities and Challenges of LLMs in Education: An NLP Perspective
por: Vajjala, Sowmya, et al.
Publicado: (2025)
por: Vajjala, Sowmya, et al.
Publicado: (2025)
LLMs in Education: Novel Perspectives, Challenges, and Opportunities
por: Alhafni, Bashar, et al.
Publicado: (2024)
por: Alhafni, Bashar, et al.
Publicado: (2024)
QUENCH: Measuring the gap between Indic and Non-Indic Contextual General Reasoning in LLMs
por: Khan, Mohammad Aflah, et al.
Publicado: (2024)
por: Khan, Mohammad Aflah, et al.
Publicado: (2024)
Refining Czech GEC: Insights from a Multi-Experiment Approach
por: Pechman, Petr, et al.
Publicado: (2025)
por: Pechman, Petr, et al.
Publicado: (2025)
IndicParam: Benchmark to evaluate LLMs on low-resource Indic Languages
por: Maheshwari, Ayush, et al.
Publicado: (2025)
por: Maheshwari, Ayush, et al.
Publicado: (2025)
The Mirage of Model Editing: Revisiting Evaluation in the Wild
por: Yang, Wanli, et al.
Publicado: (2025)
por: Yang, Wanli, et al.
Publicado: (2025)
KoGEC : Korean Grammatical Error Correction with Pre-trained Translation Models
por: Kim, Taeeun, et al.
Publicado: (2025)
por: Kim, Taeeun, et al.
Publicado: (2025)
IndicIFEval: A Benchmark for Verifiable Instruction-Following Evaluation in 14 Indic Languages
por: Jayakumar, Thanmay, et al.
Publicado: (2026)
por: Jayakumar, Thanmay, et al.
Publicado: (2026)
IndicMMLU-Pro: Benchmarking Indic Large Language Models on Multi-Task Language Understanding
por: KJ, Sankalp, et al.
Publicado: (2025)
por: KJ, Sankalp, et al.
Publicado: (2025)
COLA-GEC: A Bidirectional Framework for Enhancing Grammatical Acceptability and Error Correction
por: Yang, Xiangyu, et al.
Publicado: (2025)
por: Yang, Xiangyu, et al.
Publicado: (2025)
IndicGenBench: A Multilingual Benchmark to Evaluate Generation Capabilities of LLMs on Indic Languages
por: Singh, Harman, et al.
Publicado: (2024)
por: Singh, Harman, et al.
Publicado: (2024)
Analysis of Indic Language Capabilities in LLMs
por: Vaidya, Aatman, et al.
Publicado: (2025)
por: Vaidya, Aatman, et al.
Publicado: (2025)
Statistical Machine Translation for Indic Languages
por: Das, Sudhansu Bala, et al.
Publicado: (2023)
por: Das, Sudhansu Bala, et al.
Publicado: (2023)
IndicSQuAD: A Comprehensive Multilingual Question Answering Dataset for Indic Languages
por: Endait, Sharvi, et al.
Publicado: (2025)
por: Endait, Sharvi, et al.
Publicado: (2025)
IndicEval-XL: Bridging Linguistic Diversity in Code Generation Across Indic Languages
por: Singh, Ujjwal, et al.
Publicado: (2025)
por: Singh, Ujjwal, et al.
Publicado: (2025)
HITSZ's End-To-End Speech Translation Systems Combining Sequence-to-Sequence Auto Speech Recognition Model and Indic Large Language Model for IWSLT 2025 in Indic Track
por: Wei, Xuchen, et al.
Publicado: (2025)
por: Wei, Xuchen, et al.
Publicado: (2025)
Multi-Step Reasoning in Korean and the Emergent Mirage
por: Son, Guijin, et al.
Publicado: (2025)
por: Son, Guijin, et al.
Publicado: (2025)
LLMCL-GEC: Advancing Grammatical Error Correction with LLM-Driven Curriculum Learning
por: Fang, Tao, et al.
Publicado: (2024)
por: Fang, Tao, et al.
Publicado: (2024)
Safer in Translation? Presupposition Robustness in Indic Languages
por: Palnitkar, Aadi, et al.
Publicado: (2025)
por: Palnitkar, Aadi, et al.
Publicado: (2025)
Unicode Normalization and Grapheme Parsing of Indic Languages
por: Ansary, Nazmuddoha, et al.
Publicado: (2023)
por: Ansary, Nazmuddoha, et al.
Publicado: (2023)
LLM Judges Inconsistently Disagree Across Safety Criteria and Harm Categories
por: Vishnubhotla, Krishnapriya, et al.
Publicado: (2026)
por: Vishnubhotla, Krishnapriya, et al.
Publicado: (2026)
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages?
por: Aravapalli, Akhilesh, et al.
Publicado: (2024)
por: Aravapalli, Akhilesh, et al.
Publicado: (2024)
Indic-TunedLens: Interpreting Multilingual Models in Indian Languages
por: Panchal, Mihir, et al.
Publicado: (2026)
por: Panchal, Mihir, et al.
Publicado: (2026)
Pralekha: Cross-Lingual Document Alignment for Indic Languages
por: Suryanarayanan, Sanjay, et al.
Publicado: (2024)
por: Suryanarayanan, Sanjay, et al.
Publicado: (2024)
Table Question Answering for Low-resourced Indic Languages
por: Pal, Vaishali, et al.
Publicado: (2024)
por: Pal, Vaishali, et al.
Publicado: (2024)
Are Language Models Agnostic to Linguistically Grounded Perturbations? A Case Study of Indic Languages
por: Ghosh, Poulami, et al.
Publicado: (2024)
por: Ghosh, Poulami, et al.
Publicado: (2024)
L3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic Languages
por: Mirashi, Aishwarya, et al.
Publicado: (2024)
por: Mirashi, Aishwarya, et al.
Publicado: (2024)
L3Cube-IndicQuest: A Benchmark Question Answering Dataset for Evaluating Knowledge of LLMs in Indic Context
por: Rohera, Pritika, et al.
Publicado: (2024)
por: Rohera, Pritika, et al.
Publicado: (2024)
Introducing OmniGEC: A Silver Multilingual Dataset for Grammatical Error Correction
por: Kovalchuk, Roman, et al.
Publicado: (2025)
por: Kovalchuk, Roman, et al.
Publicado: (2025)
Towards Deployable OCR models for Indic languages
por: Mathew, Minesh, et al.
Publicado: (2022)
por: Mathew, Minesh, et al.
Publicado: (2022)
Ejemplares similares
-
The Problem with Safety Classification is not just the Models
por: Vajjala, Sowmya
Publicado: (2025) -
MATA: Mindful Assessment of the Telugu Abilities of Large Language Models
por: Kranti, Chalamalasetti, et al.
Publicado: (2025) -
MetricalARGS: A Taxonomy for Studying Metrical Poetry with LLMs
por: Kranti, Chalamalasetti, et al.
Publicado: (2025) -
Text Classification in the LLM Era -- Where do we stand?
por: Vajjala, Sowmya, et al.
Publicado: (2025) -
Does Synthetic Data Help Named Entity Recognition for Low-Resource Languages?
por: Kamath, Gaurav, et al.
Publicado: (2025)