Evaluating the Robustness of Machine Reading Comprehension Models to Low Resource Entity Renaming
Fuente:
arXiv
Saved in:
| Main Authors: | Siro, Clemencia, Ajayi, Tunde Oluwaseyi |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Judge: LLMs Designing and Applying Evaluation Rubrics
by: Siro, Clemencia, et al.
Published: (2026)
by: Siro, Clemencia, et al.
Published: (2026)
Rethinking the Evaluation of Dialogue Systems: Effects of User Feedback on Crowdworkers and LLMs
by: Siro, Clemencia, et al.
Published: (2024)
by: Siro, Clemencia, et al.
Published: (2024)
Context Does Matter: Implications for Crowdsourced Evaluation Labels in Task-Oriented Dialogue Systems
by: Siro, Clemencia, et al.
Published: (2024)
by: Siro, Clemencia, et al.
Published: (2024)
LLM-as-a-qualitative-judge: automating error analysis in natural language generation
by: Chirkova, Nadezhda, et al.
Published: (2025)
by: Chirkova, Nadezhda, et al.
Published: (2025)
AGENT-CQ: Automatic Generation and Evaluation of Clarifying Questions for Conversational Search with LLMs
by: Siro, Clemencia, et al.
Published: (2024)
by: Siro, Clemencia, et al.
Published: (2024)
Asking Multimodal Clarifying Questions in Mixed-Initiative Conversational Search
by: Yuan, Yifei, et al.
Published: (2024)
by: Yuan, Yifei, et al.
Published: (2024)
Do Images Clarify? A Study on the Effect of Images on Clarifying Questions in Conversational Search
by: Siro, Clemencia, et al.
Published: (2026)
by: Siro, Clemencia, et al.
Published: (2026)
Pay Attention to Real World Perturbations! Natural Robustness Evaluation in Machine Reading Comprehension
by: Wu, Yulong, et al.
Published: (2025)
by: Wu, Yulong, et al.
Published: (2025)
Reading Comprehension using Entity-based Memory Network
by: Wang, Xun, et al.
Published: (2016)
by: Wang, Xun, et al.
Published: (2016)
HumorRank: A Tournament-Based Leaderboard for Evaluating Humor Generation in Large Language Models
by: Ajayi, Edward, et al.
Published: (2026)
by: Ajayi, Edward, et al.
Published: (2026)
Unsupervised Named Entity Disambiguation for Low Resource Domains
by: Datta, Debarghya, et al.
Published: (2024)
by: Datta, Debarghya, et al.
Published: (2024)
Can LLMs Augment Low-Resource Reading Comprehension Datasets? Opportunities and Challenges
by: Samuel, Vinay, et al.
Published: (2023)
by: Samuel, Vinay, et al.
Published: (2023)
Multi-Turn Multi-Modal Question Clarification for Enhanced Conversational Understanding
by: Ramezan, Kimia, et al.
Published: (2025)
by: Ramezan, Kimia, et al.
Published: (2025)
Investigating Recent Large Language Models for Vietnamese Machine Reading Comprehension
by: Nguyen, Anh Duc, et al.
Published: (2025)
by: Nguyen, Anh Duc, et al.
Published: (2025)
Investigating a Benchmark for Training-set free Evaluation of Linguistic Capabilities in Machine Reading Comprehension
by: Schlegel, Viktor, et al.
Published: (2024)
by: Schlegel, Viktor, et al.
Published: (2024)
Efficient Biomedical Entity Linking: Clinical Text Standardization with Low-Resource Techniques
by: Achara, Akshit, et al.
Published: (2024)
by: Achara, Akshit, et al.
Published: (2024)
Does Synthetic Data Help Named Entity Recognition for Low-Resource Languages?
by: Kamath, Gaurav, et al.
Published: (2025)
by: Kamath, Gaurav, et al.
Published: (2025)
AfriEconQA: A Benchmark Dataset for African Economic Analysis based on World Bank Reports
by: Ajayi, Edward
Published: (2026)
by: Ajayi, Edward
Published: (2026)
Linguistic Entity Masking to Improve Cross-Lingual Representation of Multilingual Language Models for Low-Resource Languages
by: Fernando, Aloka, et al.
Published: (2025)
by: Fernando, Aloka, et al.
Published: (2025)
AgriCHN: A Comprehensive Cross-domain Resource for Chinese Agricultural Named Entity Recognition
by: Zeng, Lingxiao, et al.
Published: (2025)
by: Zeng, Lingxiao, et al.
Published: (2025)
How Good are LLMs at Relation Extraction under Low-Resource Scenario? Comprehensive Evaluation
by: Jinensibieke, Dawulie, et al.
Published: (2024)
by: Jinensibieke, Dawulie, et al.
Published: (2024)
Multilingual Multi-Aspect Explainability Analyses on Machine Reading Comprehension Models
by: Cui, Yiming, et al.
Published: (2021)
by: Cui, Yiming, et al.
Published: (2021)
Integrating a Heterogeneous Graph with Entity-aware Self-attention using Relative Position Labels for Reading Comprehension Model
by: Foolad, Shima, et al.
Published: (2023)
by: Foolad, Shima, et al.
Published: (2023)
MRCEval: A Comprehensive, Challenging and Accessible Machine Reading Comprehension Benchmark
by: Ma, Shengkun, et al.
Published: (2025)
by: Ma, Shengkun, et al.
Published: (2025)
Automatic Generation and Evaluation of Reading Comprehension Test Items with Large Language Models
by: Säuberli, Andreas, et al.
Published: (2024)
by: Säuberli, Andreas, et al.
Published: (2024)
HumorGen: Cognitive Synergy for Humor Generation in Large Language Models via Persona-Based Distillation
by: Ajayi, Edward, et al.
Published: (2026)
by: Ajayi, Edward, et al.
Published: (2026)
Generative AI for Named Entity Recognition in Low-Resource Language Nepali
by: Neupane, Sameer, et al.
Published: (2025)
by: Neupane, Sameer, et al.
Published: (2025)
A Hybrid Method for Low-Resource Named Entity Recognition
by: Duc, Do Minh, et al.
Published: (2026)
by: Duc, Do Minh, et al.
Published: (2026)
VlogQA: Task, Dataset, and Baseline Models for Vietnamese Spoken-Based Machine Reading Comprehension
by: Ngo, Thinh Phuoc, et al.
Published: (2024)
by: Ngo, Thinh Phuoc, et al.
Published: (2024)
RoBIn: A Transformer-Based Model For Risk Of Bias Inference With Machine Reading Comprehension
by: Dias, Abel Corrêa, et al.
Published: (2024)
by: Dias, Abel Corrêa, et al.
Published: (2024)
Machine Reading Comprehension using Case-based Reasoning
by: Thai, Dung, et al.
Published: (2023)
by: Thai, Dung, et al.
Published: (2023)
Benchmarking On-Device Machine Learning on Apple Silicon with MLX
by: Ajayi, Oluwaseun A., et al.
Published: (2025)
by: Ajayi, Oluwaseun A., et al.
Published: (2025)
Multi-Dimensional Machine Translation Evaluation: Model Evaluation and Resource for Korean
by: Park, Dojun, et al.
Published: (2024)
by: Park, Dojun, et al.
Published: (2024)
PANER: A Paraphrase-Augmented Framework for Low-Resource Named Entity Recognition
by: Rengarajan, Nanda Kumar, et al.
Published: (2025)
by: Rengarajan, Nanda Kumar, et al.
Published: (2025)
PICLe: Pseudo-Annotations for In-Context Learning in Low-Resource Named Entity Detection
by: Mamooler, Sepideh, et al.
Published: (2024)
by: Mamooler, Sepideh, et al.
Published: (2024)
Enhancing Pre-Trained Generative Language Models with Question Attended Span Extraction on Machine Reading Comprehension
by: Ai, Lin, et al.
Published: (2024)
by: Ai, Lin, et al.
Published: (2024)
Can GPT Redefine Medical Understanding? Evaluating GPT on Biomedical Machine Reading Comprehension
by: Vatsal, Shubham, et al.
Published: (2024)
by: Vatsal, Shubham, et al.
Published: (2024)
Dialectal and Low-Resource Machine Translation for Aromanian
by: Jerpelea, Alexandru-Iulius, et al.
Published: (2024)
by: Jerpelea, Alexandru-Iulius, et al.
Published: (2024)
Low-Resource Named Entity Recognition with Cross-Lingual, Character-Level Neural Conditional Random Fields
by: Cotterell, Ryan, et al.
Published: (2024)
by: Cotterell, Ryan, et al.
Published: (2024)
Evaluating LLMs Robustness in Less Resourced Languages with Proxy Models
by: Chrabąszcz, Maciej, et al.
Published: (2025)
by: Chrabąszcz, Maciej, et al.
Published: (2025)
Similar Items
-
Learning to Judge: LLMs Designing and Applying Evaluation Rubrics
by: Siro, Clemencia, et al.
Published: (2026) -
Rethinking the Evaluation of Dialogue Systems: Effects of User Feedback on Crowdworkers and LLMs
by: Siro, Clemencia, et al.
Published: (2024) -
Context Does Matter: Implications for Crowdsourced Evaluation Labels in Task-Oriented Dialogue Systems
by: Siro, Clemencia, et al.
Published: (2024) -
LLM-as-a-qualitative-judge: automating error analysis in natural language generation
by: Chirkova, Nadezhda, et al.
Published: (2025) -
AGENT-CQ: Automatic Generation and Evaluation of Clarifying Questions for Conversational Search with LLMs
by: Siro, Clemencia, et al.
Published: (2024)