Exploring Robustness of Multilingual LLMs on Real-World Noisy Data
Fuente:
arXiv
Saved in:
| Main Authors: | Aliakbarzadeh, Amirhossein, Flek, Lucie, Karimi, Akbar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Robustness of LLMs to Paraphrasing Based on Sociodemographic Factors
by: Arora, Pulkit, et al.
Published: (2025)
by: Arora, Pulkit, et al.
Published: (2025)
ArithmAttack: Evaluating Robustness of LLMs to Noisy Context in Math Problem Solving
by: Abedin, Zain Ul, et al.
Published: (2025)
by: Abedin, Zain Ul, et al.
Published: (2025)
Encoder Fine-tuning with Stochastic Sampling Outperforms Open-weight GPT in Astronomy Knowledge Extraction
by: Rawat, Shivam, et al.
Published: (2025)
by: Rawat, Shivam, et al.
Published: (2025)
Multi-Hop Reasoning for Question Answering with Hyperbolic Representations
by: Welz, Simon, et al.
Published: (2025)
by: Welz, Simon, et al.
Published: (2025)
Label-Consistent Data Generation for Aspect-Based Sentiment Analysis Using LLM Agents
by: Monfared, Mohammad H. A., et al.
Published: (2026)
by: Monfared, Mohammad H. A., et al.
Published: (2026)
More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persists
by: Alavi, Khashayar, et al.
Published: (2025)
by: Alavi, Khashayar, et al.
Published: (2025)
Can LLM Agents Identify Spoken Dialects like a Linguist?
by: Bystrich, Tobias, et al.
Published: (2026)
by: Bystrich, Tobias, et al.
Published: (2026)
Improving Low-Resource Dialect Classification Using Retrieval-based Voice Conversion
by: Fischbach, Lea, et al.
Published: (2025)
by: Fischbach, Lea, et al.
Published: (2025)
Do Multilingual Large Language Models Mitigate Stereotype Bias?
by: Nie, Shangrui, et al.
Published: (2024)
by: Nie, Shangrui, et al.
Published: (2024)
Probing the Robustness of Theory of Mind in Large Language Models
by: Nickel, Christian, et al.
Published: (2024)
by: Nickel, Christian, et al.
Published: (2024)
Evaluating Robustness of LLMs in Question Answering on Multilingual Noisy OCR Data
by: Piryani, Bhawna, et al.
Published: (2025)
by: Piryani, Bhawna, et al.
Published: (2025)
Pitfalls of Conversational LLMs on News Debiasing
by: Schlicht, Ipek Baris, et al.
Published: (2024)
by: Schlicht, Ipek Baris, et al.
Published: (2024)
On the Limitations of Language Targeted Pruning: Investigating the Calibration Language Impact in Multilingual LLM Pruning
by: Kurz, Simon, et al.
Published: (2024)
by: Kurz, Simon, et al.
Published: (2024)
Reasoning Primitives in Hybrid and Non-Hybrid LLMs: Do Architectural Differences Yield Advantages in State-Tracking and Recall?
by: Rawat, Shivam, et al.
Published: (2026)
by: Rawat, Shivam, et al.
Published: (2026)
IKnow: Instruction-Knowledge-Aware Continual Pretraining for Effective Domain Adaptation
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
Survey-to-Behavior: Downstream Alignment of Human Values in LLMs via Survey Questions
by: Nie, Shangrui, et al.
Published: (2025)
by: Nie, Shangrui, et al.
Published: (2025)
Disparities in Multilingual LLM-Based Healthcare Q&A
by: Schlicht, Ipek Baris, et al.
Published: (2025)
by: Schlicht, Ipek Baris, et al.
Published: (2025)
The Muddy Waters of Modeling Empathy in Language: The Practical Impacts of Theoretical Constructs
by: Lahnala, Allison, et al.
Published: (2025)
by: Lahnala, Allison, et al.
Published: (2025)
A Critical Reflection and Forward Perspective on Empathy and Natural Language Processing
by: Lahnala, Allison, et al.
Published: (2022)
by: Lahnala, Allison, et al.
Published: (2022)
Funzac at CoMeDi Shared Task: Modeling Annotator Disagreement from Word-In-Context Perspectives
by: Sarumi, Olufunke O., et al.
Published: (2025)
by: Sarumi, Olufunke O., et al.
Published: (2025)
Can Stories Help LLMs Reason? Curating Information Space Through Narrative
by: Javadi, Vahid Sadiri, et al.
Published: (2024)
by: Javadi, Vahid Sadiri, et al.
Published: (2024)
Do LLMs Provide Consistent Answers to Health-Related Questions across Languages?
by: Schlicht, Ipek Baris, et al.
Published: (2025)
by: Schlicht, Ipek Baris, et al.
Published: (2025)
Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models
by: Nickel, Christian, et al.
Published: (2026)
by: Nickel, Christian, et al.
Published: (2026)
PERSPECTRA: A Scalable and Configurable Pluralist Benchmark of Perspectives from Arguments
by: Nie, Shangrui, et al.
Published: (2026)
by: Nie, Shangrui, et al.
Published: (2026)
Tucano 2 Cool: Better Open Source LLMs for Portuguese
by: Corrêa, Nicholas Kluge, et al.
Published: (2026)
by: Corrêa, Nicholas Kluge, et al.
Published: (2026)
Reinforcement Learning Amplifies Emergent Misalignment from Harmless Rewards
by: Jørgenvåg, Magnus, et al.
Published: (2026)
by: Jørgenvåg, Magnus, et al.
Published: (2026)
How to Learn in a Noisy World? Self-Correcting the Real-World Data Noise in Machine Translation
by: Meng, Yan, et al.
Published: (2024)
by: Meng, Yan, et al.
Published: (2024)
Unraveling Babel: Exploring Multilingual Activation Patterns of LLMs and Their Applications
by: Liu, Weize, et al.
Published: (2024)
by: Liu, Weize, et al.
Published: (2024)
Unifying the Extremes: Developing a Unified Model for Detecting and Predicting Extremist Traits and Radicalization
by: Lahnala, Allison, et al.
Published: (2025)
by: Lahnala, Allison, et al.
Published: (2025)
ISCA: A Framework for Interview-Style Conversational Agents
by: Welch, Charles, et al.
Published: (2025)
by: Welch, Charles, et al.
Published: (2025)
Raising Bars, Not Parameters: LilMoo Compact Language Model for Hindi
by: Fatimah, Shiza, et al.
Published: (2026)
by: Fatimah, Shiza, et al.
Published: (2026)
USDC: A Dataset of $\underline{U}$ser $\underline{S}$tance and $\underline{D}$ogmatism in Long $\underline{C}$onversations
by: Marreddy, Mounika, et al.
Published: (2024)
by: Marreddy, Mounika, et al.
Published: (2024)
Reasoning-Guided Claim Normalization for Noisy Multilingual Social Media Posts
by: Sharma, Manan, et al.
Published: (2025)
by: Sharma, Manan, et al.
Published: (2025)
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Multimodal Educational Tasks
by: Marius, Dumitran Adrian, et al.
Published: (2025)
by: Marius, Dumitran Adrian, et al.
Published: (2025)
Toward Robust Multilingual Adaptation of LLMs for Low-Resource Languages
by: Li, Haolin, et al.
Published: (2025)
by: Li, Haolin, et al.
Published: (2025)
HPLT 3.0: Very Large-Scale Multilingual Resources for LLMs and MT. Mono- and Bi-lingual Data, Multilingual Evaluation, and Pre-Trained Models
by: Oepen, Stephan, et al.
Published: (2025)
by: Oepen, Stephan, et al.
Published: (2025)
LEMONADE: A Large Multilingual Expert-Annotated Abstractive Event Dataset for the Real World
by: Semnani, Sina J., et al.
Published: (2025)
by: Semnani, Sina J., et al.
Published: (2025)
HEALTH-PARIKSHA: Assessing RAG Models for Health Chatbots in Real-World Multilingual Settings
by: Gumma, Varun, et al.
Published: (2024)
by: Gumma, Varun, et al.
Published: (2024)
Towards Automated Fact-Checking of Real-World Claims: Exploring Task Formulation and Assessment with LLMs
by: Sahitaj, Premtim, et al.
Published: (2025)
by: Sahitaj, Premtim, et al.
Published: (2025)
XDoGE: Multilingual Data Reweighting to Enhance Language Inclusivity in LLMs
by: Lacunza, Iñaki, et al.
Published: (2025)
by: Lacunza, Iñaki, et al.
Published: (2025)
Similar Items
-
Exploring Robustness of LLMs to Paraphrasing Based on Sociodemographic Factors
by: Arora, Pulkit, et al.
Published: (2025) -
ArithmAttack: Evaluating Robustness of LLMs to Noisy Context in Math Problem Solving
by: Abedin, Zain Ul, et al.
Published: (2025) -
Encoder Fine-tuning with Stochastic Sampling Outperforms Open-weight GPT in Astronomy Knowledge Extraction
by: Rawat, Shivam, et al.
Published: (2025) -
Multi-Hop Reasoning for Question Answering with Hyperbolic Representations
by: Welz, Simon, et al.
Published: (2025) -
Label-Consistent Data Generation for Aspect-Based Sentiment Analysis Using LLM Agents
by: Monfared, Mohammad H. A., et al.
Published: (2026)