Found in Translation: Measuring Multilingual LLM Consistency as Simple as Translate then Evaluate
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gupta, Ashim, Mehta, Maitrey, Xu, Zhichao, Srikumar, Vivek |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion
von: Mehta, Maitrey, et al.
Veröffentlicht: (2026)
von: Mehta, Maitrey, et al.
Veröffentlicht: (2026)
Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation
von: Gupta, Ashim, et al.
Veröffentlicht: (2025)
von: Gupta, Ashim, et al.
Veröffentlicht: (2025)
Beyond Perplexity: Multi-dimensional Safety Evaluation of LLM Compression
von: Xu, Zhichao, et al.
Veröffentlicht: (2024)
von: Xu, Zhichao, et al.
Veröffentlicht: (2024)
Promptly Predicting Structures: The Return of Inference
von: Mehta, Maitrey, et al.
Veröffentlicht: (2024)
von: Mehta, Maitrey, et al.
Veröffentlicht: (2024)
State Space Models are Strong Text Rerankers
von: Xu, Zhichao, et al.
Veröffentlicht: (2024)
von: Xu, Zhichao, et al.
Veröffentlicht: (2024)
LACONIC: Dense-Level Effectiveness for Scalable Sparse Retrieval via a Two-Phase Training Curriculum
von: Xu, Zhichao, et al.
Veröffentlicht: (2026)
von: Xu, Zhichao, et al.
Veröffentlicht: (2026)
An Empirical Investigation of Matrix Factorization Methods for Pre-trained Transformers
von: Gupta, Ashim, et al.
Veröffentlicht: (2024)
von: Gupta, Ashim, et al.
Veröffentlicht: (2024)
Whispers of Doubt Amidst Echoes of Triumph in NLP Robustness
von: Gupta, Ashim, et al.
Veröffentlicht: (2023)
von: Gupta, Ashim, et al.
Veröffentlicht: (2023)
In-Context Example Ordering Guided by Label Distributions
von: Xu, Zhichao, et al.
Veröffentlicht: (2024)
von: Xu, Zhichao, et al.
Veröffentlicht: (2024)
LLM-Symbolic Integration for Robust Temporal Tabular Reasoning
von: Kulkarni, Atharv, et al.
Veröffentlicht: (2025)
von: Kulkarni, Atharv, et al.
Veröffentlicht: (2025)
Quantifying the Impact of Translation Errors on Multilingual LLM Evaluation
von: Thellmann, Klaudia-Doris, et al.
Veröffentlicht: (2026)
von: Thellmann, Klaudia-Doris, et al.
Veröffentlicht: (2026)
Mufu: Multilingual Fused Learning for Low-Resource Translation with LLM
von: Lim, Zheng Wei, et al.
Veröffentlicht: (2024)
von: Lim, Zheng Wei, et al.
Veröffentlicht: (2024)
Unequal Voices: How LLMs Construct Constrained Queer Narratives
von: Ghosal, Atreya, et al.
Veröffentlicht: (2025)
von: Ghosal, Atreya, et al.
Veröffentlicht: (2025)
Distillation versus Contrastive Learning: How to Train Your Rerankers
von: Xu, Zhichao, et al.
Veröffentlicht: (2025)
von: Xu, Zhichao, et al.
Veröffentlicht: (2025)
Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
von: Kreutzer, Julia, et al.
Veröffentlicht: (2025)
von: Kreutzer, Julia, et al.
Veröffentlicht: (2025)
Sāmayik: A Benchmark and Dataset for English-Sanskrit Translation
von: Maheshwari, Ayush, et al.
Veröffentlicht: (2023)
von: Maheshwari, Ayush, et al.
Veröffentlicht: (2023)
Reinforcing Code Generation: Improving Text-to-SQL with Execution-Based Learning
von: Kulkarni, Atharv, et al.
Veröffentlicht: (2025)
von: Kulkarni, Atharv, et al.
Veröffentlicht: (2025)
Backdoor Attack on Multilingual Machine Translation
von: Wang, Jun, et al.
Veröffentlicht: (2024)
von: Wang, Jun, et al.
Veröffentlicht: (2024)
Seed-X: Building Strong Multilingual Translation LLM with 7B Parameters
von: Cheng, Shanbo, et al.
Veröffentlicht: (2025)
von: Cheng, Shanbo, et al.
Veröffentlicht: (2025)
"Be My Cheese?": Assessing Cultural Nuance in Multilingual LLM Translations
von: Van Doren, Madison, et al.
Veröffentlicht: (2025)
von: Van Doren, Madison, et al.
Veröffentlicht: (2025)
Translation as a Scalable Proxy for Multilingual Evaluation
von: Issaka, Sheriff, et al.
Veröffentlicht: (2026)
von: Issaka, Sheriff, et al.
Veröffentlicht: (2026)
EEG-to-Text Translation: A Model for Deciphering Human Brain Activity
von: Murad, Saydul Akbar, et al.
Veröffentlicht: (2025)
von: Murad, Saydul Akbar, et al.
Veröffentlicht: (2025)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
von: Choenni, Rochelle, et al.
Veröffentlicht: (2024)
Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation
von: Wang, Xintong, et al.
Veröffentlicht: (2025)
von: Wang, Xintong, et al.
Veröffentlicht: (2025)
InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis
von: Bentham, Oliver, et al.
Veröffentlicht: (2026)
von: Bentham, Oliver, et al.
Veröffentlicht: (2026)
Multilingual LLM Prompting Strategies for Medical English-Vietnamese Machine Translation
von: Vo, Nhu, et al.
Veröffentlicht: (2025)
von: Vo, Nhu, et al.
Veröffentlicht: (2025)
xTower: A Multilingual LLM for Explaining and Correcting Translation Errors
von: Treviso, Marcos, et al.
Veröffentlicht: (2024)
von: Treviso, Marcos, et al.
Veröffentlicht: (2024)
Copy First, Translate Later: Interpreting Translation Dynamics in Multilingual Pretraining
von: Körner, Felicia, et al.
Veröffentlicht: (2026)
von: Körner, Felicia, et al.
Veröffentlicht: (2026)
Leveraging LLM For Synchronizing Information Across Multilingual Tables
von: Khincha, Siddharth, et al.
Veröffentlicht: (2025)
von: Khincha, Siddharth, et al.
Veröffentlicht: (2025)
Conditions for Catastrophic Forgetting in Multilingual Translation
von: Liu, Danni, et al.
Veröffentlicht: (2025)
von: Liu, Danni, et al.
Veröffentlicht: (2025)
Ensuring Consistency for In-Image Translation
von: Fu, Chengpeng, et al.
Veröffentlicht: (2024)
von: Fu, Chengpeng, et al.
Veröffentlicht: (2024)
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection
von: Chan, Fai Leui, et al.
Veröffentlicht: (2024)
von: Chan, Fai Leui, et al.
Veröffentlicht: (2024)
Asymmetric Conflict and Synergy in Post-training for LLM-based Multilingual Machine Translation
von: Zheng, Tong, et al.
Veröffentlicht: (2025)
von: Zheng, Tong, et al.
Veröffentlicht: (2025)
How and Where to Translate? The Impact of Translation Strategies in Cross-lingual LLM Prompting
von: Gupta, Aman, et al.
Veröffentlicht: (2025)
von: Gupta, Aman, et al.
Veröffentlicht: (2025)
Beyond Translation: LLM-Based Data Generation for Multilingual Fact-Checking
von: Chung, Yi-Ling, et al.
Veröffentlicht: (2025)
von: Chung, Yi-Ling, et al.
Veröffentlicht: (2025)
Is Multilingual LLM Watermarking Truly Multilingual? Scaling Robustness to 100+ Languages via Back-Translation
von: Mohamed, Asim, et al.
Veröffentlicht: (2025)
von: Mohamed, Asim, et al.
Veröffentlicht: (2025)
Science Across Languages: Assessing LLM Multilingual Translation of Scientific Papers
von: Kleidermacher, Hannah Calzi, et al.
Veröffentlicht: (2025)
von: Kleidermacher, Hannah Calzi, et al.
Veröffentlicht: (2025)
PMMT: Preference Alignment in Multilingual Machine Translation via LLM Distillation
von: Sun, Shuqiao, et al.
Veröffentlicht: (2024)
von: Sun, Shuqiao, et al.
Veröffentlicht: (2024)
TAPO: Translation Augmented Policy Optimization for Multilingual Mathematical Reasoning
von: Huang, Xu, et al.
Veröffentlicht: (2026)
von: Huang, Xu, et al.
Veröffentlicht: (2026)
Understanding the Logic of Direct Preference Alignment through Logic
von: Richardson, Kyle, et al.
Veröffentlicht: (2024)
von: Richardson, Kyle, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion
von: Mehta, Maitrey, et al.
Veröffentlicht: (2026) -
Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation
von: Gupta, Ashim, et al.
Veröffentlicht: (2025) -
Beyond Perplexity: Multi-dimensional Safety Evaluation of LLM Compression
von: Xu, Zhichao, et al.
Veröffentlicht: (2024) -
Promptly Predicting Structures: The Return of Inference
von: Mehta, Maitrey, et al.
Veröffentlicht: (2024) -
State Space Models are Strong Text Rerankers
von: Xu, Zhichao, et al.
Veröffentlicht: (2024)