JMigBench: A Benchmark for Evaluating LLMs on Source Code Migration (Java 8 to Java 11)
Fuente:
arXiv
Guardado en:
| Autores principales: | Amin, Nishil, Fei, Zhiwei, Li, Xiang, Petke, Justyna, Ye, He |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MigrationBench: Repository-Level Code Migration Benchmark from Java 8
por: Liu, Linbo, et al.
Publicado: (2025)
por: Liu, Linbo, et al.
Publicado: (2025)
A Comprehensive Survey of Benchmarks for Automated Improvement of Software's Non-Functional Properties
por: Blot, Aymeric, et al.
Publicado: (2022)
por: Blot, Aymeric, et al.
Publicado: (2022)
ScarfBench: A Benchmark for Cross-Framework Application Migration in Enterprise Java
por: Pavuluri, Advait, et al.
Publicado: (2026)
por: Pavuluri, Advait, et al.
Publicado: (2026)
FreshBrew: A Benchmark for Evaluating AI Agents on Java Code Migration
por: May, Victor, et al.
Publicado: (2025)
por: May, Victor, et al.
Publicado: (2025)
Detection of Technical Debt in Java Source Code
por: Hai, Nam Le, et al.
Publicado: (2024)
por: Hai, Nam Le, et al.
Publicado: (2024)
HotBugs.jar: A Benchmark of Hot Fixes for Time-Critical Bugs
por: Hanna, Carol, et al.
Publicado: (2025)
por: Hanna, Carol, et al.
Publicado: (2025)
Reinforcement Learning for Mutation Operator Selection in Automated Program Repair
por: Hanna, Carol, et al.
Publicado: (2023)
por: Hanna, Carol, et al.
Publicado: (2023)
Environment-in-the-Loop: Rethinking Code Migration with LLM-based Agents
por: Li, Xiang, et al.
Publicado: (2026)
por: Li, Xiang, et al.
Publicado: (2026)
GitBug-Java: A Reproducible Benchmark of Recent Java Bugs
por: Silva, André, et al.
Publicado: (2024)
por: Silva, André, et al.
Publicado: (2024)
Generating Accurate OpenAPI Descriptions from Java Source Code
por: Lercher, Alexander, et al.
Publicado: (2024)
por: Lercher, Alexander, et al.
Publicado: (2024)
Concolic Testing of JavaScript using Sparkplug
por: Li, Zhe, et al.
Publicado: (2024)
por: Li, Zhe, et al.
Publicado: (2024)
Unmasking the Genuine Type Inference Capabilities of LLMs for Java Code Snippets
por: Dong, Yiwen, et al.
Publicado: (2025)
por: Dong, Yiwen, et al.
Publicado: (2025)
Acknowledging Good Java Code with Code Perfumes
por: Straubinger, Philipp, et al.
Publicado: (2024)
por: Straubinger, Philipp, et al.
Publicado: (2024)
jscefr: A Framework to Evaluate the Code Proficiency for JavaScript
por: Ragkhitwetsagul, Chaiyong, et al.
Publicado: (2024)
por: Ragkhitwetsagul, Chaiyong, et al.
Publicado: (2024)
Serializing Java Objects in Plain Code
por: Wachter, Julian, et al.
Publicado: (2024)
por: Wachter, Julian, et al.
Publicado: (2024)
Quality Evaluation of COBOL to Java Code Transformation
por: Froimovich, Shmulik, et al.
Publicado: (2025)
por: Froimovich, Shmulik, et al.
Publicado: (2025)
Hot Fixing Software: A Comprehensive Review of Terminology, Techniques, and Applications
por: Hanna, Carol, et al.
Publicado: (2024)
por: Hanna, Carol, et al.
Publicado: (2024)
Compilation of Commit Changes within Java Source Code Repositories
por: Schott, Stefan, et al.
Publicado: (2024)
por: Schott, Stefan, et al.
Publicado: (2024)
JavaBench: A Benchmark of Object-Oriented Code Generation for Evaluating Large Language Models
por: Cao, Jialun, et al.
Publicado: (2024)
por: Cao, Jialun, et al.
Publicado: (2024)
Demystifying and Assessing Code Understandability in Java Decompilation
por: Qin, Ruixin, et al.
Publicado: (2024)
por: Qin, Ruixin, et al.
Publicado: (2024)
A Soundness and Precision Benchmark for Java Debloating Tools
por: Klauke, Jonas, et al.
Publicado: (2025)
por: Klauke, Jonas, et al.
Publicado: (2025)
Scalable Thread-Safety Analysis of Java Classes with CodeQL
por: Jåtten, Bjørnar Haugstad, et al.
Publicado: (2025)
por: Jåtten, Bjørnar Haugstad, et al.
Publicado: (2025)
Leveraging LLMs for Automated Translation of Legacy Code: A Case Study on PL/SQL to Java Transformation
por: Solovyeva, Lola, et al.
Publicado: (2025)
por: Solovyeva, Lola, et al.
Publicado: (2025)
JavaVFC: Java Vulnerability Fixing Commits from Open-source Software
por: Bui, Tan, et al.
Publicado: (2024)
por: Bui, Tan, et al.
Publicado: (2024)
Zero-shot Evaluation of Deep Learning for Java Code Clone Detection
por: Heinze, Thomas S.
Publicado: (2026)
por: Heinze, Thomas S.
Publicado: (2026)
SWE-Bench+: Enhanced Coding Benchmark for LLMs
por: Aleithan, Reem, et al.
Publicado: (2024)
por: Aleithan, Reem, et al.
Publicado: (2024)
Systematic Detection of Energy Regression and Corresponding Code Patterns in Java Projects
por: Bechet, François, et al.
Publicado: (2026)
por: Bechet, François, et al.
Publicado: (2026)
Formal Methods Meets Readability: Auto-Documenting JML Java Code
por: Abad, Juan Carlos Recio, et al.
Publicado: (2025)
por: Abad, Juan Carlos Recio, et al.
Publicado: (2025)
Automated Refactoring of Legacy JavaScript Code to ES6 Modules
por: Paltoglou, Katerina, et al.
Publicado: (2021)
por: Paltoglou, Katerina, et al.
Publicado: (2021)
Demonstrating ARG-V's Generation of Realistic Java Benchmarks for SV-COMP
por: Moloney, Charles, et al.
Publicado: (2026)
por: Moloney, Charles, et al.
Publicado: (2026)
State-Of-The-Practice in Quality Assurance in Java-Based Open Source Software Development
por: Khatami, Ali, et al.
Publicado: (2023)
por: Khatami, Ali, et al.
Publicado: (2023)
Roseau: Fast, Accurate, Source-based API Breaking Change Analysis in Java
por: Latappy, Corentin, et al.
Publicado: (2025)
por: Latappy, Corentin, et al.
Publicado: (2025)
A Systematic Evaluation of Environmental Flakiness in JavaScript Tests
por: Hashemi, Negar, et al.
Publicado: (2026)
por: Hashemi, Negar, et al.
Publicado: (2026)
Detecting and Evaluating Order-Dependent Flaky Tests in JavaScript
por: Hashemi, Negar, et al.
Publicado: (2025)
por: Hashemi, Negar, et al.
Publicado: (2025)
An Empirical Study on Method-Level Performance Evolution in Open-Source Java Projects
por: Shahedi, Kaveh, et al.
Publicado: (2025)
por: Shahedi, Kaveh, et al.
Publicado: (2025)
Generating Java Methods: An Empirical Assessment of Four AI-Based Code Assistants
por: Corso, Vincenzo, et al.
Publicado: (2024)
por: Corso, Vincenzo, et al.
Publicado: (2024)
An Empirical Study of Java Code Improvements Based on Stack Overflow Answer Edits
por: Wiratsin, In-on, et al.
Publicado: (2025)
por: Wiratsin, In-on, et al.
Publicado: (2025)
Unveiling Practical Shortcomings of Patch Overfitting Detection Techniques
por: Williams, David, et al.
Publicado: (2026)
por: Williams, David, et al.
Publicado: (2026)
Test-based Patch Clustering for Automatically-Generated Patches Assessment
por: Martinez, Matias, et al.
Publicado: (2022)
por: Martinez, Matias, et al.
Publicado: (2022)
Empirical and Sustainability Aspects of Software Engineering Research in the Era of Large Language Models: A Reflection
por: Williams, David, et al.
Publicado: (2025)
por: Williams, David, et al.
Publicado: (2025)
Ejemplares similares
-
MigrationBench: Repository-Level Code Migration Benchmark from Java 8
por: Liu, Linbo, et al.
Publicado: (2025) -
A Comprehensive Survey of Benchmarks for Automated Improvement of Software's Non-Functional Properties
por: Blot, Aymeric, et al.
Publicado: (2022) -
ScarfBench: A Benchmark for Cross-Framework Application Migration in Enterprise Java
por: Pavuluri, Advait, et al.
Publicado: (2026) -
FreshBrew: A Benchmark for Evaluating AI Agents on Java Code Migration
por: May, Victor, et al.
Publicado: (2025) -
Detection of Technical Debt in Java Source Code
por: Hai, Nam Le, et al.
Publicado: (2024)