English Please: Evaluating Machine Translation with Large Language Models for Multilingual Bug Reports
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Patil, Avinash, Tao, Siru, Jadon, Aryan |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
GitBugs: Bug Reports for Duplicate Detection, Retrieval Augmented Generation, Triage, and More
par: Patil, Avinash, et autres
Publié: (2025)
par: Patil, Avinash, et autres
Publié: (2025)
Advancing Reasoning in Large Language Models: Promising Methods and Approaches
par: Patil, Avinash, et autres
Publié: (2025)
par: Patil, Avinash, et autres
Publié: (2025)
A Blueprint for AI-Driven Software Quality: Integrating LLMs with Established Standards
par: Patil, Avinash
Publié: (2025)
par: Patil, Avinash
Publié: (2025)
Exploring Large Language Models for Translating Romanian Computational Problems into English
par: Dumitran, Adrian Marius, et autres
Publié: (2025)
par: Dumitran, Adrian Marius, et autres
Publié: (2025)
When Bugs Linger: A Study of Anomalous Resolution Time Outliers and Their Themes
par: Patil, Avinash
Publié: (2025)
par: Patil, Avinash
Publié: (2025)
Exploring Large Language Models in Resolving Environment-Related Crash Bugs: Localizing and Repairing
par: Du, Xueying, et autres
Publié: (2023)
par: Du, Xueying, et autres
Publié: (2023)
Enhancing Domain-Specific Retrieval-Augmented Generation: Synthetic Data Generation and Evaluation using Reasoning Models
par: Jadon, Aryan, et autres
Publié: (2025)
par: Jadon, Aryan, et autres
Publié: (2025)
Progressive Code Integration for Abstractive Bug Report Summarization
par: Karim, Shaira Sadia, et autres
Publié: (2025)
par: Karim, Shaira Sadia, et autres
Publié: (2025)
Machine Translation Testing via Syntactic Tree Pruning
par: Zhang, Quanjun, et autres
Publié: (2024)
par: Zhang, Quanjun, et autres
Publié: (2024)
Coffee: Boost Your Code LLMs by Fixing Bugs with Feedback
par: Moon, Seungjun, et autres
Publié: (2023)
par: Moon, Seungjun, et autres
Publié: (2023)
Mix-of-Language-Experts Architecture for Multilingual Programming
par: Zong, Yifan, et autres
Publié: (2025)
par: Zong, Yifan, et autres
Publié: (2025)
CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios
par: Huang, Shiting, et autres
Publié: (2025)
par: Huang, Shiting, et autres
Publié: (2025)
What's Wrong with Your Code Generated by Large Language Models? An Extensive Study
par: Dou, Shihan, et autres
Publié: (2024)
par: Dou, Shihan, et autres
Publié: (2024)
Less is More: Adaptive Program Repair with Bug Localization and Preference Learning
par: Dai, Zhenlong, et autres
Publié: (2025)
par: Dai, Zhenlong, et autres
Publié: (2025)
HumanEval Pro and MBPP Pro: Evaluating Large Language Models on Self-invoking Code Generation
par: Yu, Zhaojian, et autres
Publié: (2024)
par: Yu, Zhaojian, et autres
Publié: (2024)
Interpretable Online Log Analysis Using Large Language Models with Prompt Strategies
par: Liu, Yilun, et autres
Publié: (2023)
par: Liu, Yilun, et autres
Publié: (2023)
BugPilot: Complex Bug Generation for Efficient Learning of SWE Skills
par: Sonwane, Atharv, et autres
Publié: (2025)
par: Sonwane, Atharv, et autres
Publié: (2025)
Adapting Large Language Models to Log Analysis with Interpretable Domain Knowledge
par: Ji, Yuhe, et autres
Publié: (2024)
par: Ji, Yuhe, et autres
Publié: (2024)
M2rc-Eval: Massively Multilingual Repository-level Code Completion Evaluation
par: Liu, Jiaheng, et autres
Publié: (2024)
par: Liu, Jiaheng, et autres
Publié: (2024)
Generating and Evaluating Sustainable Procurement Criteria for the Swiss Public Sector using In-Context Prompting with Large Language Models
par: Gao, Yingqiang, et autres
Publié: (2026)
par: Gao, Yingqiang, et autres
Publié: (2026)
Evaluation of the Programming Skills of Large Language Models
par: Heitz, Luc Bryan, et autres
Publié: (2024)
par: Heitz, Luc Bryan, et autres
Publié: (2024)
AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators
par: Chou, Jason, et autres
Publié: (2025)
par: Chou, Jason, et autres
Publié: (2025)
SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability Analysis
par: Li, Yansong, et autres
Publié: (2025)
par: Li, Yansong, et autres
Publié: (2025)
Towards a Middleware for Large Language Models
par: Guran, Narcisa, et autres
Publié: (2024)
par: Guran, Narcisa, et autres
Publié: (2024)
Calibration of Large Language Models on Code Summarization
par: Virk, Yuvraj, et autres
Publié: (2024)
par: Virk, Yuvraj, et autres
Publié: (2024)
BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?
par: Chen, Guoxin, et autres
Publié: (2026)
par: Chen, Guoxin, et autres
Publié: (2026)
TestExplora: Benchmarking LLMs for Proactive Bug Discovery via Repository-Level Test Generation
par: Liu, Steven, et autres
Publié: (2026)
par: Liu, Steven, et autres
Publié: (2026)
Is Your AI-Generated Code Really Safe? Evaluating Large Language Models on Secure Code Generation with CodeSecEval
par: Wang, Jiexin, et autres
Publié: (2024)
par: Wang, Jiexin, et autres
Publié: (2024)
ASTRAL: Automated Safety Testing of Large Language Models
par: Ugarte, Miriam, et autres
Publié: (2025)
par: Ugarte, Miriam, et autres
Publié: (2025)
A Survey of AIOps in the Era of Large Language Models
par: Zhang, Lingzhe, et autres
Publié: (2025)
par: Zhang, Lingzhe, et autres
Publié: (2025)
Large Language Models for IT Automation Tasks: Are We There Yet?
par: Hassan, Md Mahadi, et autres
Publié: (2025)
par: Hassan, Md Mahadi, et autres
Publié: (2025)
Evaluation and Improvement of Fault Detection for Large Language Models
par: Hu, Qiang, et autres
Publié: (2024)
par: Hu, Qiang, et autres
Publié: (2024)
Investigating Markers and Drivers of Gender Bias in Machine Translations
par: Barclay, Peter J, et autres
Publié: (2024)
par: Barclay, Peter J, et autres
Publié: (2024)
Lost in Translation: A Study of Bugs Introduced by Large Language Models while Translating Code
par: Pan, Rangeet, et autres
Publié: (2023)
par: Pan, Rangeet, et autres
Publié: (2023)
TimeMachine-bench: A Benchmark for Evaluating Model Capabilities in Repository-Level Migration Tasks
par: Fujii, Ryo, et autres
Publié: (2026)
par: Fujii, Ryo, et autres
Publié: (2026)
A Comparative Study on Large Language Models for Log Parsing
par: Astekin, Merve, et autres
Publié: (2024)
par: Astekin, Merve, et autres
Publié: (2024)
Assessing the Latent Automated Program Repair Capabilities of Large Language Models using Round-Trip Translation
par: Ruiz, Fernando Vallecillos, et autres
Publié: (2024)
par: Ruiz, Fernando Vallecillos, et autres
Publié: (2024)
Enhancing Translation Validation of Compiler Transformations with Large Language Models
par: Wang, Yanzhao, et autres
Publié: (2024)
par: Wang, Yanzhao, et autres
Publié: (2024)
Evaluating the Generalization Capabilities of Large Language Models on Code Reasoning
par: Yang, Rem, et autres
Publié: (2025)
par: Yang, Rem, et autres
Publié: (2025)
ICE-Score: Instructing Large Language Models to Evaluate Code
par: Zhuo, Terry Yue
Publié: (2023)
par: Zhuo, Terry Yue
Publié: (2023)
Documents similaires
-
GitBugs: Bug Reports for Duplicate Detection, Retrieval Augmented Generation, Triage, and More
par: Patil, Avinash, et autres
Publié: (2025) -
Advancing Reasoning in Large Language Models: Promising Methods and Approaches
par: Patil, Avinash, et autres
Publié: (2025) -
A Blueprint for AI-Driven Software Quality: Integrating LLMs with Established Standards
par: Patil, Avinash
Publié: (2025) -
Exploring Large Language Models for Translating Romanian Computational Problems into English
par: Dumitran, Adrian Marius, et autres
Publié: (2025) -
When Bugs Linger: A Study of Anomalous Resolution Time Outliers and Their Themes
par: Patil, Avinash
Publié: (2025)