LLM-as-a-Judge for Software Engineering: Literature Review, Vision, and the Road Ahead
Fuente:
arXiv
Guardado en:
| Autores principales: | He, Junda, Shi, Jieke, Zhuo, Terry Yue, Treude, Christoph, Sun, Jiamou, Xing, Zhenchang, Du, Xiaoning, Lo, David |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
From Code to Courtroom: LLMs as the New Software Judges
por: He, Junda, et al.
Publicado: (2025)
por: He, Junda, et al.
Publicado: (2025)
LLM-Based Multi-Agent Systems for Software Engineering: Literature Review, Vision and the Road Ahead
por: He, Junda, et al.
Publicado: (2024)
por: He, Junda, et al.
Publicado: (2024)
Efficient and Green Large Language Models for Software Engineering: Literature Review, Vision, and the Road Ahead
por: Shi, Jieke, et al.
Publicado: (2024)
por: Shi, Jieke, et al.
Publicado: (2024)
Identifying and Mitigating API Misuse in Large Language Models
por: Zhuo, Terry Yue, et al.
Publicado: (2025)
por: Zhuo, Terry Yue, et al.
Publicado: (2025)
Artificial Intelligence for Software Architecture: Literature Review and the Road Ahead
por: Bucaioni, Alessio, et al.
Publicado: (2025)
por: Bucaioni, Alessio, et al.
Publicado: (2025)
Compiling Code LLMs into Lightweight Executables
por: Shi, Jieke, et al.
Publicado: (2026)
por: Shi, Jieke, et al.
Publicado: (2026)
Qualitative Data Analysis in Software Engineering: Techniques and Teaching Insights
por: Treude, Christoph
Publicado: (2024)
por: Treude, Christoph
Publicado: (2024)
Large Language Model for Vulnerability Detection and Repair: Literature Review and the Road Ahead
por: Zhou, Xin, et al.
Publicado: (2024)
por: Zhou, Xin, et al.
Publicado: (2024)
Synthesizing Efficient and Permissive Programmatic Runtime Shields for Neural Policies
por: Shi, Jieke, et al.
Publicado: (2024)
por: Shi, Jieke, et al.
Publicado: (2024)
PTMPicker: Facilitating Efficient Pretrained Model Selection for Application Developers
por: Liu, Pei, et al.
Publicado: (2025)
por: Liu, Pei, et al.
Publicado: (2025)
Quantum Artificial Intelligence for Software Engineering: the Road Ahead
por: Wang, Xinyi, et al.
Publicado: (2025)
por: Wang, Xinyi, et al.
Publicado: (2025)
Accountable Agents in Software Engineering: An Analysis of Terms of Service and a Research Roadmap
por: Treude, Christoph
Publicado: (2026)
por: Treude, Christoph
Publicado: (2026)
Gender Influence on Student Teams' Online Communication in Software Engineering Education
por: Garcia, Rita, et al.
Publicado: (2025)
por: Garcia, Rita, et al.
Publicado: (2025)
LLM-Assisted Empirical Software Engineering: Systematic Literature Review and Research Agenda
por: Gomes, Victoria, et al.
Publicado: (2026)
por: Gomes, Victoria, et al.
Publicado: (2026)
GenAI Is No Silver Bullet for Qualitative Research in Software Engineering
por: Ernst, Neil A., et al.
Publicado: (2026)
por: Ernst, Neil A., et al.
Publicado: (2026)
Foundation Models for Software Engineering of Cyber-Physical Systems: the Road Ahead
por: Lu, Chengjie, et al.
Publicado: (2025)
por: Lu, Chengjie, et al.
Publicado: (2025)
Automated Soap Opera Testing Directed by LLMs and Scenario Knowledge: Feasibility, Challenges, and Road Ahead
por: Su, Yanqi, et al.
Publicado: (2024)
por: Su, Yanqi, et al.
Publicado: (2024)
Greening Large Language Models of Code
por: Shi, Jieke, et al.
Publicado: (2023)
por: Shi, Jieke, et al.
Publicado: (2023)
A Functional Software Reference Architecture for LLM-Integrated Systems
por: Bucaioni, Alessio, et al.
Publicado: (2025)
por: Bucaioni, Alessio, et al.
Publicado: (2025)
Novice Developers' Perspectives on Adopting LLMs for Software Development: A Systematic Literature Review
por: Ferino, Samuel, et al.
Publicado: (2025)
por: Ferino, Samuel, et al.
Publicado: (2025)
Towards Reliable LLM-Driven Fuzz Testing: Vision and Road Ahead
por: Cheng, Yiran, et al.
Publicado: (2025)
por: Cheng, Yiran, et al.
Publicado: (2025)
Generative AI and Empirical Software Engineering: A Paradigm Shift
por: Treude, Christoph, et al.
Publicado: (2025)
por: Treude, Christoph, et al.
Publicado: (2025)
Rethinking Artifact Evaluation for Software Engineering in the Age of Generative AI
por: Treude, Christoph, et al.
Publicado: (2026)
por: Treude, Christoph, et al.
Publicado: (2026)
Human in the Loop for Fuzz Testing: Literature Review and the Road Ahead
por: Yu, Jiongchi, et al.
Publicado: (2026)
por: Yu, Jiongchi, et al.
Publicado: (2026)
Assessing and Advancing Benchmarks for Evaluating Large Language Models in Software Engineering Tasks
por: Hu, Xing, et al.
Publicado: (2025)
por: Hu, Xing, et al.
Publicado: (2025)
Context Engineering for AI Agents in Open-Source Software
por: Mohsenimofidi, Seyedmoein, et al.
Publicado: (2025)
por: Mohsenimofidi, Seyedmoein, et al.
Publicado: (2025)
Software Engineering for Large Language Models: Research Status, Challenges and the Road Ahead
por: Rao, Hongzhou, et al.
Publicado: (2025)
por: Rao, Hongzhou, et al.
Publicado: (2025)
Finding Safety Violations of AI-Enabled Control Systems through the Lens of Synthesized Proxy Programs
por: Shi, Jieke, et al.
Publicado: (2024)
por: Shi, Jieke, et al.
Publicado: (2024)
Quantum Software Engineering: Roadmap and Challenges Ahead
por: Murillo, Juan M., et al.
Publicado: (2024)
por: Murillo, Juan M., et al.
Publicado: (2024)
Do Chase Your Tail! Missing Key Aspects Augmentation in Textual Vulnerability Descriptions of Long-tail Software through Feature Inference
por: Han, Linyi, et al.
Publicado: (2024)
por: Han, Linyi, et al.
Publicado: (2024)
When Prompt Engineering Meets Software Engineering: CNL-P as Natural and Robust "APIs'' for Human-AI Interaction
por: Xing, Zhenchang, et al.
Publicado: (2025)
por: Xing, Zhenchang, et al.
Publicado: (2025)
LLMAID: Identifying AI Capabilities in Android Apps with LLMs
por: Liu, Pei, et al.
Publicado: (2025)
por: Liu, Pei, et al.
Publicado: (2025)
Curiosity-Driven Testing for Sequential Decision-Making Process
por: He, Junda, et al.
Publicado: (2025)
por: He, Junda, et al.
Publicado: (2025)
Large Language Models for Software Engineering: A Systematic Literature Review
por: Hou, Xinyi, et al.
Publicado: (2023)
por: Hou, Xinyi, et al.
Publicado: (2023)
Interacting with AI Reasoning Models: Harnessing "Thoughts" for AI-Driven Software Engineering
por: Treude, Christoph, et al.
Publicado: (2025)
por: Treude, Christoph, et al.
Publicado: (2025)
AI Slop and the Software Commons
por: Baltes, Sebastian, et al.
Publicado: (2026)
por: Baltes, Sebastian, et al.
Publicado: (2026)
How Developers Interact with AI: A Taxonomy of Human-AI Collaboration in Software Engineering
por: Treude, Christoph, et al.
Publicado: (2025)
por: Treude, Christoph, et al.
Publicado: (2025)
APIDocBooster: An Extract-Then-Abstract Framework Leveraging Large Language Models for Augmenting API Documentation
por: Yang, Chengran, et al.
Publicado: (2023)
por: Yang, Chengran, et al.
Publicado: (2023)
AgentSZZ: Teaching the LLM Agent to Play Detective with Bug-Inducing Commits
por: Lyu, Yunbo, et al.
Publicado: (2026)
por: Lyu, Yunbo, et al.
Publicado: (2026)
Bot-Driven Development: From Simple Automation to Autonomous Software Development Bots
por: Treude, Christoph, et al.
Publicado: (2024)
por: Treude, Christoph, et al.
Publicado: (2024)
Ejemplares similares
-
From Code to Courtroom: LLMs as the New Software Judges
por: He, Junda, et al.
Publicado: (2025) -
LLM-Based Multi-Agent Systems for Software Engineering: Literature Review, Vision and the Road Ahead
por: He, Junda, et al.
Publicado: (2024) -
Efficient and Green Large Language Models for Software Engineering: Literature Review, Vision, and the Road Ahead
por: Shi, Jieke, et al.
Publicado: (2024) -
Identifying and Mitigating API Misuse in Large Language Models
por: Zhuo, Terry Yue, et al.
Publicado: (2025) -
Artificial Intelligence for Software Architecture: Literature Review and the Road Ahead
por: Bucaioni, Alessio, et al.
Publicado: (2025)