Evaluating Robustness of Large Language Models Against Multilingual Typographical Errors
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhao, Raoyuan, Liu, Yihong, Altinger, Lena, Schütze, Hinrich, Hedderich, Michael A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Large Reasoning Models Are (Not Yet) Multilingual Latent Reasoners
di: Liu, Yihong, et al.
Pubblicazione: (2026)
di: Liu, Yihong, et al.
Pubblicazione: (2026)
A Comprehensive Evaluation of Multilingual Chain-of-Thought Reasoning: Performance, Consistency, and Faithfulness Across Languages
di: Zhao, Raoyuan, et al.
Pubblicazione: (2025)
di: Zhao, Raoyuan, et al.
Pubblicazione: (2025)
Crosslingual On-Policy Self-Distillation for Multilingual Reasoning
di: Liu, Yihong, et al.
Pubblicazione: (2026)
di: Liu, Yihong, et al.
Pubblicazione: (2026)
ReverseMath: Answer Inversion for Scalable and Verifiable Mathematical Problem Generation
di: Zhao, Raoyuan, et al.
Pubblicazione: (2026)
di: Zhao, Raoyuan, et al.
Pubblicazione: (2026)
Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models
di: Liu, Yihong, et al.
Pubblicazione: (2026)
di: Liu, Yihong, et al.
Pubblicazione: (2026)
Beyond Input Understanding: Diagnosing Multilingual Mathematical Reasoning with Directed Acyclic Trace Graphs
di: Zhang, Jiaqiao, et al.
Pubblicazione: (2026)
di: Zhang, Jiaqiao, et al.
Pubblicazione: (2026)
Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing
di: Zhao, Raoyuan, et al.
Pubblicazione: (2025)
di: Zhao, Raoyuan, et al.
Pubblicazione: (2025)
MAKIEval: A Multilingual Automatic WiKidata-based Framework for Cultural Awareness Evaluation for LLMs
di: Zhao, Raoyuan, et al.
Pubblicazione: (2025)
di: Zhao, Raoyuan, et al.
Pubblicazione: (2025)
Breaking the Script Barrier in Multilingual Pre-Trained Language Models with Transliteration-Based Post-Training Alignment
di: Xhelili, Orgest, et al.
Pubblicazione: (2024)
di: Xhelili, Orgest, et al.
Pubblicazione: (2024)
SYNTHEVAL: Hybrid Behavioral Testing of NLP Models with Synthetic CheckLists
di: Zhao, Raoyuan, et al.
Pubblicazione: (2024)
di: Zhao, Raoyuan, et al.
Pubblicazione: (2024)
TransliCo: A Contrastive Learning Framework to Address the Script Barrier in Multilingual Pretrained Language Models
di: Liu, Yihong, et al.
Pubblicazione: (2024)
di: Liu, Yihong, et al.
Pubblicazione: (2024)
TransMI: A Framework to Create Strong Baselines from Multilingual Pretrained Language Models for Transliterated Data
di: Liu, Yihong, et al.
Pubblicazione: (2024)
di: Liu, Yihong, et al.
Pubblicazione: (2024)
OFA: A Framework of Initializing Unseen Subword Embeddings for Efficient Large-scale Multilingual Continued Pretraining
di: Liu, Yihong, et al.
Pubblicazione: (2023)
di: Liu, Yihong, et al.
Pubblicazione: (2023)
LangSAMP: Language-Script Aware Multilingual Pretraining
di: Liu, Yihong, et al.
Pubblicazione: (2024)
di: Liu, Yihong, et al.
Pubblicazione: (2024)
A Recipe of Parallel Corpora Exploitation for Multilingual Large Language Models
di: Lin, Peiqin, et al.
Pubblicazione: (2024)
di: Lin, Peiqin, et al.
Pubblicazione: (2024)
Lost in Multilinguality: Dissecting Cross-lingual Factual Inconsistency in Transformer Language Models
di: Wang, Mingyang, et al.
Pubblicazione: (2025)
di: Wang, Mingyang, et al.
Pubblicazione: (2025)
Enhancing Robustness of Autoregressive Language Models against Orthographic Attacks via Pixel-based Approach
di: Yang, Han, et al.
Pubblicazione: (2025)
di: Yang, Han, et al.
Pubblicazione: (2025)
Why Better Cross-Lingual Alignment Fails for Better Cross-Lingual Transfer: Case of Encoders
di: Veitsman, Yana, et al.
Pubblicazione: (2026)
di: Veitsman, Yana, et al.
Pubblicazione: (2026)
How Programming Concepts and Neurons Are Shared in Code Language Models
di: Kargaran, Amir Hossein, et al.
Pubblicazione: (2025)
di: Kargaran, Amir Hossein, et al.
Pubblicazione: (2025)
Left, Right, or Center? Evaluating LLM Framing in News Classification and Generation
di: Kennedy, Molly, et al.
Pubblicazione: (2026)
di: Kennedy, Molly, et al.
Pubblicazione: (2026)
HYPEROFA: Expanding LLM Vocabulary to New Languages via Hypernetwork-Based Embedding Initialization
di: Özeren, Enes, et al.
Pubblicazione: (2025)
di: Özeren, Enes, et al.
Pubblicazione: (2025)
Your Pretrained Model Tells the Difficulty Itself: A Self-Adaptive Curriculum Learning Paradigm for Natural Language Understanding
di: Feng, Qi, et al.
Pubblicazione: (2025)
di: Feng, Qi, et al.
Pubblicazione: (2025)
Human Uncertainty-Aware Data Selection and Automatic Labeling in Visual Question Answering
di: Lan, Jian, et al.
Pubblicazione: (2025)
di: Lan, Jian, et al.
Pubblicazione: (2025)
Reasoning Robustness of LLMs to Adversarial Typographical Errors
di: Gan, Esther, et al.
Pubblicazione: (2024)
di: Gan, Esther, et al.
Pubblicazione: (2024)
Relational Linearity is a Predictor of Hallucinations
di: Lu, Yuetian, et al.
Pubblicazione: (2026)
di: Lu, Yuetian, et al.
Pubblicazione: (2026)
Decomposed Prompting: Probing Multilingual Linguistic Structure Knowledge in Large Language Models
di: Nie, Ercong, et al.
Pubblicazione: (2024)
di: Nie, Ercong, et al.
Pubblicazione: (2024)
Exploring the Role of Transliteration in In-Context Learning for Low-resource Languages Written in Non-Latin Scripts
di: Ma, Chunlan, et al.
Pubblicazione: (2024)
di: Ma, Chunlan, et al.
Pubblicazione: (2024)
Calibration Is Not Enough: Evaluating Confidence Estimation Under Language Variations
di: Xia, Yuxi, et al.
Pubblicazione: (2026)
di: Xia, Yuxi, et al.
Pubblicazione: (2026)
Mechanistic Understanding and Mitigation of Language Confusion in English-Centric Large Language Models
di: Nie, Ercong, et al.
Pubblicazione: (2025)
di: Nie, Ercong, et al.
Pubblicazione: (2025)
MoSECroT: Model Stitching with Static Word Embeddings for Crosslingual Zero-shot Transfer
di: Ye, Haotian, et al.
Pubblicazione: (2024)
di: Ye, Haotian, et al.
Pubblicazione: (2024)
Refusal Direction is Universal Across Safety-Aligned Languages
di: Wang, Xinpeng, et al.
Pubblicazione: (2025)
di: Wang, Xinpeng, et al.
Pubblicazione: (2025)
On the Entity-Level Alignment in Crosslingual Consistency
di: Liu, Yihong, et al.
Pubblicazione: (2025)
di: Liu, Yihong, et al.
Pubblicazione: (2025)
Understanding In-Context Machine Translation for Low-Resource Languages: A Case Study on Manchu
di: Pei, Renhao, et al.
Pubblicazione: (2025)
di: Pei, Renhao, et al.
Pubblicazione: (2025)
What's the Difference? Supporting Users in Identifying the Effects of Prompt and Model Changes Through Token Patterns
di: Hedderich, Michael A., et al.
Pubblicazione: (2025)
di: Hedderich, Michael A., et al.
Pubblicazione: (2025)
GLUScope: A Tool for Analyzing GLU Neurons in Transformer Language Models
di: Gerstner, Sebastian, et al.
Pubblicazione: (2026)
di: Gerstner, Sebastian, et al.
Pubblicazione: (2026)
Tracing Multilingual Factual Knowledge Acquisition in Pretraining
di: Liu, Yihong, et al.
Pubblicazione: (2025)
di: Liu, Yihong, et al.
Pubblicazione: (2025)
mPLM-Sim: Better Cross-Lingual Similarity and Transfer in Multilingual Pretrained Language Models
di: Lin, Peiqin, et al.
Pubblicazione: (2023)
di: Lin, Peiqin, et al.
Pubblicazione: (2023)
On Relation-Specific Neurons in Large Language Models
di: Liu, Yihong, et al.
Pubblicazione: (2025)
di: Liu, Yihong, et al.
Pubblicazione: (2025)
RET-LLM: Towards a General Read-Write Memory for Large Language Models
di: Modarressi, Ali, et al.
Pubblicazione: (2023)
di: Modarressi, Ali, et al.
Pubblicazione: (2023)
EMMA-500: Enhancing Massively Multilingual Adaptation of Large Language Models
di: Ji, Shaoxiong, et al.
Pubblicazione: (2024)
di: Ji, Shaoxiong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Large Reasoning Models Are (Not Yet) Multilingual Latent Reasoners
di: Liu, Yihong, et al.
Pubblicazione: (2026) -
A Comprehensive Evaluation of Multilingual Chain-of-Thought Reasoning: Performance, Consistency, and Faithfulness Across Languages
di: Zhao, Raoyuan, et al.
Pubblicazione: (2025) -
Crosslingual On-Policy Self-Distillation for Multilingual Reasoning
di: Liu, Yihong, et al.
Pubblicazione: (2026) -
ReverseMath: Answer Inversion for Scalable and Verifiable Mathematical Problem Generation
di: Zhao, Raoyuan, et al.
Pubblicazione: (2026) -
Evaluating Contextually Mediated Factual Recall in Multilingual Large Language Models
di: Liu, Yihong, et al.
Pubblicazione: (2026)