Automated evaluation of LLMs for effective machine translation of Mandarin Chinese to English
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yue, Beard, Rodney, Hawkins, John, Chandra, Rohitash |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluation of Google Translate for Mandarin Chinese translation using sentiment and semantic analysis
by: Wang, Xuechun, et al.
Published: (2024)
by: Wang, Xuechun, et al.
Published: (2024)
Machine Learning for Detection and Analysis of Novel LLM Jailbreaks
by: Hawkins, John, et al.
Published: (2025)
by: Hawkins, John, et al.
Published: (2025)
An evaluation of LLMs and Google Translate for translation of selected Indian languages via sentiment and semantic analyses
by: Chandra, Rohitash, et al.
Published: (2025)
by: Chandra, Rohitash, et al.
Published: (2025)
Recursive deep learning framework for forecasting the decadal world economic outlook
by: Wang, Tianyi, et al.
Published: (2023)
by: Wang, Tianyi, et al.
Published: (2023)
Abusive text transformation using LLMs
by: Chandra, Rohitash, et al.
Published: (2025)
by: Chandra, Rohitash, et al.
Published: (2025)
Abusive music and song transformation using GenAI and LLMs
by: Choi, Jiyang, et al.
Published: (2026)
by: Choi, Jiyang, et al.
Published: (2026)
An evaluation of LLMs for political bias in Western media: Israel-Hamas and Ukraine-Russia wars
by: Chandra, Rohitash, et al.
Published: (2026)
by: Chandra, Rohitash, et al.
Published: (2026)
A longitudinal sentiment analysis of Sinophobia during COVID-19 using large language models
by: Wang, Chen, et al.
Published: (2024)
by: Wang, Chen, et al.
Published: (2024)
An evaluation of LLMs for generating movie reviews: GPT-4o, Gemini-2.0 and DeepSeek-V3
by: Sands, Brendan, et al.
Published: (2025)
by: Sands, Brendan, et al.
Published: (2025)
Longitudinal Abuse and Sentiment Analysis of Hollywood Movie Dialogues using Language Models
by: Chandra, Rohitash, et al.
Published: (2025)
by: Chandra, Rohitash, et al.
Published: (2025)
Polyphone Disambiguation in Mandarin Chinese with Semi-Supervised Learning
by: Shi, Yi, et al.
Published: (2021)
by: Shi, Yi, et al.
Published: (2021)
Language models for longitudinal analysis of abusive content in Billboard Music Charts
by: Chandra, Rohitash, et al.
Published: (2025)
by: Chandra, Rohitash, et al.
Published: (2025)
Ukrainian-to-English folktale corpus: Parallel corpus creation and augmentation for machine translation in low-resource languages
by: Burda-Lassen, Olena
Published: (2024)
by: Burda-Lassen, Olena
Published: (2024)
Large language model for Bible sentiment analysis: Sermon on the Mount
by: Vora, Mahek, et al.
Published: (2024)
by: Vora, Mahek, et al.
Published: (2024)
Can professional translators identify machine-generated text?
by: Farrell, Michael
Published: (2026)
by: Farrell, Michael
Published: (2026)
Improving AGI Evaluation: A Data Science Perspective
by: Hawkins, John
Published: (2025)
by: Hawkins, John
Published: (2025)
Toward domain-specific machine translation and quality estimation systems
by: Sharami, Javad Pourmostafa Roshan
Published: (2026)
by: Sharami, Javad Pourmostafa Roshan
Published: (2026)
Can postgraduate translation students identify machine-generated text?
by: Farrell, Michael
Published: (2025)
by: Farrell, Michael
Published: (2025)
The Role of Handling Attributive Nouns in Improving Chinese-To-English Machine Translation
by: Wang, Lisa, et al.
Published: (2024)
by: Wang, Lisa, et al.
Published: (2024)
Code-Based English Models Surprising Performance on Chinese QA Pair Extraction Task
by: Zheng, Linghan, et al.
Published: (2024)
by: Zheng, Linghan, et al.
Published: (2024)
Investigating the potential of Sparse Mixtures-of-Experts for multi-domain neural machine translation
by: Chirkova, Nadezhda, et al.
Published: (2024)
by: Chirkova, Nadezhda, et al.
Published: (2024)
Agent-Driven Large Language Models for Mandarin Lyric Generation
by: Liu, Hong-Hsiang, et al.
Published: (2024)
by: Liu, Hong-Hsiang, et al.
Published: (2024)
Quantity Convergence, Quality Divergence: Disentangling Fluency and Accuracy in L2 Mandarin Prosody
by: Shi, Yuqi, et al.
Published: (2026)
by: Shi, Yuqi, et al.
Published: (2026)
Enhanced Review Detection and Recognition: A Platform-Agnostic Approach with Application to Online Commerce
by: Karmakar, Priyabrata, et al.
Published: (2024)
by: Karmakar, Priyabrata, et al.
Published: (2024)
C-FAITH: A Chinese Fine-Grained Benchmark for Automated Hallucination Evaluation
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
Bayesian neural networks via MCMC: a Python-based tutorial
by: Chandra, Rohitash, et al.
Published: (2023)
by: Chandra, Rohitash, et al.
Published: (2023)
Echoes of Automation: The Increasing Use of LLMs in Newsmaking
by: Ansari, Abolfazl, et al.
Published: (2025)
by: Ansari, Abolfazl, et al.
Published: (2025)
Automated test generation to evaluate tool-augmented LLMs as conversational AI agents
by: Arcadinho, Samuel, et al.
Published: (2024)
by: Arcadinho, Samuel, et al.
Published: (2024)
Uncovering the Fragility of Trustworthy LLMs through Chinese Textual Ambiguity
by: Wu, Xinwei, et al.
Published: (2025)
by: Wu, Xinwei, et al.
Published: (2025)
Flames: Benchmarking Value Alignment of LLMs in Chinese
by: Huang, Kexin, et al.
Published: (2023)
by: Huang, Kexin, et al.
Published: (2023)
BreezyVoice: Adapting TTS for Taiwanese Mandarin with Enhanced Polyphone Disambiguation -- Challenges and Insights
by: Hsu, Chan-Jan, et al.
Published: (2025)
by: Hsu, Chan-Jan, et al.
Published: (2025)
Trivial Vocabulary Bans Improve LLM Reasoning More Than Deep Linguistic Constraints
by: Jehu-Appiah, Rodney
Published: (2026)
by: Jehu-Appiah, Rodney
Published: (2026)
Enigme: Generative Text Puzzles for Evaluating Reasoning in Language Models
by: Hawkins, John
Published: (2025)
by: Hawkins, John
Published: (2025)
An Analysis on Automated Metrics for Evaluating Japanese-English Chat Translation
by: Rusli, Andre, et al.
Published: (2024)
by: Rusli, Andre, et al.
Published: (2024)
Bridging the Data Gap: Creating a Hindi Text Summarization Dataset from the English XSUM
by: Katwe, Praveenkumar, et al.
Published: (2026)
by: Katwe, Praveenkumar, et al.
Published: (2026)
Position: LLMs Can be Good Tutors in English Education
by: Ye, Jingheng, et al.
Published: (2025)
by: Ye, Jingheng, et al.
Published: (2025)
Clinical knowledge in LLMs does not translate to human interactions
by: Bean, Andrew M., et al.
Published: (2025)
by: Bean, Andrew M., et al.
Published: (2025)
Benchmarking the Detection of LLMs-Generated Modern Chinese Poetry
by: Wang, Shanshan, et al.
Published: (2025)
by: Wang, Shanshan, et al.
Published: (2025)
CTourLLM: Enhancing LLMs with Chinese Tourism Knowledge
by: Wei, Qikai, et al.
Published: (2024)
by: Wei, Qikai, et al.
Published: (2024)
Accommodation and Epistemic Vigilance: A Pragmatic Account of Why LLMs Fail to Challenge Harmful Beliefs
by: Cheng, Myra, et al.
Published: (2026)
by: Cheng, Myra, et al.
Published: (2026)
Similar Items
-
Evaluation of Google Translate for Mandarin Chinese translation using sentiment and semantic analysis
by: Wang, Xuechun, et al.
Published: (2024) -
Machine Learning for Detection and Analysis of Novel LLM Jailbreaks
by: Hawkins, John, et al.
Published: (2025) -
An evaluation of LLMs and Google Translate for translation of selected Indian languages via sentiment and semantic analyses
by: Chandra, Rohitash, et al.
Published: (2025) -
Recursive deep learning framework for forecasting the decadal world economic outlook
by: Wang, Tianyi, et al.
Published: (2023) -
Abusive text transformation using LLMs
by: Chandra, Rohitash, et al.
Published: (2025)