Evaluating Dialect Robustness of Language Models via Conversation Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Srirag, Dipankar, Sahoo, Nihar Ranjan, Joshi, Aditya |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Predicting the Target Word of Game-playing Conversations using a Low-Rank Dialect Adapter for Decoder Models
by: Srirag, Dipankar, et al.
Published: (2024)
by: Srirag, Dipankar, et al.
Published: (2024)
Far Out: Evaluating Language Models on Slang in Australian and Indian English
by: Dilsiz, Deniz Kaya, et al.
Published: (2026)
by: Dilsiz, Deniz Kaya, et al.
Published: (2026)
PAREDA: A Multi-Accent Speech Dataset of Natural Language Processing Research Discussions
by: Jin, Sicheng, et al.
Published: (2026)
by: Jin, Sicheng, et al.
Published: (2026)
Nek Minit: Harnessing Pragmatic Metacognitive Prompting for Explainable Sarcasm Detection of Australian and Indian English
by: Singh, Ishmanbir, et al.
Published: (2025)
by: Singh, Ishmanbir, et al.
Published: (2025)
Experiences from Creating a Benchmark for Sentiment Classification for Varieties of English
by: Srirag, Dipankar, et al.
Published: (2024)
by: Srirag, Dipankar, et al.
Published: (2024)
TriageSim: A Conversational Emergency Triage Simulation Framework from Structured Electronic Health Records
by: Srirag, Dipankar, et al.
Published: (2026)
by: Srirag, Dipankar, et al.
Published: (2026)
BESSTIE: A Benchmark for Sentiment and Sarcasm Classification for Varieties of English
by: Srirag, Dipankar, et al.
Published: (2024)
by: Srirag, Dipankar, et al.
Published: (2024)
BharatBBQ: A Multilingual Bias Benchmark for Question Answering in the Indian Context
by: Tomar, Aditya, et al.
Published: (2025)
by: Tomar, Aditya, et al.
Published: (2025)
Mathematics Isn't Culture-Free: Probing Cultural Gaps via Entity and Scenario Perturbations
by: Tomar, Aditya, et al.
Published: (2025)
by: Tomar, Aditya, et al.
Published: (2025)
Open-DeBias: Toward Mitigating Open-Set Bias in Language Models
by: Rani, Arti, et al.
Published: (2025)
by: Rani, Arti, et al.
Published: (2025)
A Taxonomy-Driven Case Study of Australian Web Resources Against Technology-Facilitated Abuse
by: Srirag, Dipankar, et al.
Published: (2025)
by: Srirag, Dipankar, et al.
Published: (2025)
Natural Language Processing for Dialects of a Language: A Survey
by: Joshi, Aditya, et al.
Published: (2024)
by: Joshi, Aditya, et al.
Published: (2024)
Harnessing Test-time Adaptation for NLU tasks Involving Dialects of English
by: Nguyen, Duke, et al.
Published: (2025)
by: Nguyen, Duke, et al.
Published: (2025)
IndiBias: A Benchmark Dataset to Measure Social Biases in Language Models for Indian Context
by: Sahoo, Nihar Ranjan, et al.
Published: (2024)
by: Sahoo, Nihar Ranjan, et al.
Published: (2024)
What am I missing here?: Evaluating Large Language Models for Masked Sentence Prediction
by: Wyatt, Charlie, et al.
Published: (2025)
by: Wyatt, Charlie, et al.
Published: (2025)
DialUp! Modeling the Language Continuum by Adapting Models to Dialects and Dialects to Models
by: Bafna, Niyati, et al.
Published: (2025)
by: Bafna, Niyati, et al.
Published: (2025)
SLURP-TN : Resource for Tunisian Dialect Spoken Language Understanding
by: Elleuch, Haroun, et al.
Published: (2026)
by: Elleuch, Haroun, et al.
Published: (2026)
Many Dialects, Many Languages, One Cultural Lens: Evaluating Multilingual VLMs for Bengali Culture Understanding Across Historically Linked Languages and Regional Dialects
by: Sayeedi, Nurul Labib, et al.
Published: (2026)
by: Sayeedi, Nurul Labib, et al.
Published: (2026)
JEEM: Vision-Language Understanding in Four Arabic Dialects
by: Kadaoui, Karima, et al.
Published: (2025)
by: Kadaoui, Karima, et al.
Published: (2025)
Assessing Dialect Fairness and Robustness of Large Language Models in Reasoning Tasks
by: Lin, Fangru, et al.
Published: (2024)
by: Lin, Fangru, et al.
Published: (2024)
KoDialogBench: Evaluating Conversational Understanding of Language Models with Korean Dialogue Benchmark
by: Jang, Seongbo, et al.
Published: (2024)
by: Jang, Seongbo, et al.
Published: (2024)
Voice Conversion Improves Cross-Domain Robustness for Spoken Arabic Dialect Identification
by: Abdullah, Badr M., et al.
Published: (2025)
by: Abdullah, Badr M., et al.
Published: (2025)
DialectalArabicMMLU: Benchmarking Dialectal Capabilities in Arabic and Multilingual Language Models
by: Altakrori, Malik H., et al.
Published: (2025)
by: Altakrori, Malik H., et al.
Published: (2025)
ILLUMINER: Instruction-tuned Large Language Models as Few-shot Intent Classifier and Slot Filler
by: Mirza, Paramita, et al.
Published: (2024)
by: Mirza, Paramita, et al.
Published: (2024)
LangLingual: A Personalised, Exercise-oriented English Language Learning Tool Leveraging Large Language Models
by: Gupta, Sammriddh, et al.
Published: (2025)
by: Gupta, Sammriddh, et al.
Published: (2025)
AlcLaM: Arabic Dialectal Language Model
by: Ahmed, Murtadha, et al.
Published: (2024)
by: Ahmed, Murtadha, et al.
Published: (2024)
DialectGen: Benchmarking and Improving Dialect Robustness in Multimodal Generation
by: Zhou, Yu, et al.
Published: (2025)
by: Zhou, Yu, et al.
Published: (2025)
Idiom Understanding as a Tool to Measure the Dialect Gap
by: Beauchemin, David, et al.
Published: (2025)
by: Beauchemin, David, et al.
Published: (2025)
Improving Sequence-to-Sequence Models for Abstractive Text Summarization Using Meta Heuristic Approaches
by: Saxena, Aditya, et al.
Published: (2024)
by: Saxena, Aditya, et al.
Published: (2024)
Large Language Models Discriminate Against Speakers of German Dialects
by: Bui, Minh Duc, et al.
Published: (2025)
by: Bui, Minh Duc, et al.
Published: (2025)
Low-Resource Dialect Adaptation of Large Language Models: A French Dialect Case-Study
by: Khan, Eeham, et al.
Published: (2025)
by: Khan, Eeham, et al.
Published: (2025)
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection
by: Chan, Fai Leui, et al.
Published: (2024)
by: Chan, Fai Leui, et al.
Published: (2024)
Understanding Refusal in Language Models with Sparse Autoencoders
by: Yeo, Wei Jie, et al.
Published: (2025)
by: Yeo, Wei Jie, et al.
Published: (2025)
Narrative Theory-Driven LLM Methods for Automatic Story Generation and Understanding: A Survey
by: Liu, David Y., et al.
Published: (2026)
by: Liu, David Y., et al.
Published: (2026)
From Local Concepts to Universals: Evaluating the Multicultural Understanding of Vision-Language Models
by: Bhatia, Mehar, et al.
Published: (2024)
by: Bhatia, Mehar, et al.
Published: (2024)
Dialectal Toxicity Detection: Evaluating LLM-as-a-Judge Consistency Across Language Varieties
by: Faisal, Fahim, et al.
Published: (2024)
by: Faisal, Fahim, et al.
Published: (2024)
Jet Quenching in Heavy-Ion Collisions at RHIC and the LHC experiments
by: Sahoo, Nihar Ranjan
Published: (2025)
by: Sahoo, Nihar Ranjan
Published: (2025)
INDIC DIALECT: A Multi Task Benchmark to Evaluate and Translate in Indian Language Dialects
by: Sharma, Tarun, et al.
Published: (2026)
by: Sharma, Tarun, et al.
Published: (2026)
Examining Language Modeling Assumptions Using an Annotated Literary Dialect Corpus
by: Messner, Craig, et al.
Published: (2024)
by: Messner, Craig, et al.
Published: (2024)
Exploring Bengali Religious Dialect Biases in Large Language Models with Evaluation Perspectives
by: Wasi, Azmine Toushik, et al.
Published: (2024)
by: Wasi, Azmine Toushik, et al.
Published: (2024)
Similar Items
-
Predicting the Target Word of Game-playing Conversations using a Low-Rank Dialect Adapter for Decoder Models
by: Srirag, Dipankar, et al.
Published: (2024) -
Far Out: Evaluating Language Models on Slang in Australian and Indian English
by: Dilsiz, Deniz Kaya, et al.
Published: (2026) -
PAREDA: A Multi-Accent Speech Dataset of Natural Language Processing Research Discussions
by: Jin, Sicheng, et al.
Published: (2026) -
Nek Minit: Harnessing Pragmatic Metacognitive Prompting for Explainable Sarcasm Detection of Australian and Indian English
by: Singh, Ishmanbir, et al.
Published: (2025) -
Experiences from Creating a Benchmark for Sentiment Classification for Varieties of English
by: Srirag, Dipankar, et al.
Published: (2024)