HalluJudge: A Reference-Free Hallucination Detection for Context Misalignment in Code Review Automation
Fuente:
arXiv
Guardado en:
| Autores principales: | Tantithamthavorn, Kla, Lin, Hong Yi, Thongtanunam, Patanamon, Charoenwet, Wachiraphan, Jeong, Minwoo, Wu, Ming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection
por: Charoenwet, Wachiraphan, et al.
Publicado: (2026)
por: Charoenwet, Wachiraphan, et al.
Publicado: (2026)
Improving Automated Code Reviews: Learning from Experience
por: Lin, Hong Yi, et al.
Publicado: (2024)
por: Lin, Hong Yi, et al.
Publicado: (2024)
Toward Effective Secure Code Reviews: An Empirical Study of Security-Related Coding Weaknesses
por: Charoenwet, Wachiraphan, et al.
Publicado: (2023)
por: Charoenwet, Wachiraphan, et al.
Publicado: (2023)
Enhancing Code Review through Fuzzing and Likely Invariants
por: Charoenwet, Wachiraphan, et al.
Publicado: (2025)
por: Charoenwet, Wachiraphan, et al.
Publicado: (2025)
An Empirical Study of Static Analysis Tools for Secure Code Review
por: Charoenwet, Wachiraphan, et al.
Publicado: (2024)
por: Charoenwet, Wachiraphan, et al.
Publicado: (2024)
Leveraging Reviewer Experience in Code Review Comment Generation
por: Lin, Hong Yi, et al.
Publicado: (2024)
por: Lin, Hong Yi, et al.
Publicado: (2024)
Code Ownership: The Principles, Differences, and Their Associations with Software Quality
por: Thongtanunam, Patanamon, et al.
Publicado: (2024)
por: Thongtanunam, Patanamon, et al.
Publicado: (2024)
Hallucinations in Code Change to Natural Language Generation: Prevalence and Evaluation of Detection Metrics
por: Liu, Chunhua, et al.
Publicado: (2025)
por: Liu, Chunhua, et al.
Publicado: (2025)
What Types of Code Review Comments Do Developers Most Frequently Resolve?
por: Goldman, Saul, et al.
Publicado: (2025)
por: Goldman, Saul, et al.
Publicado: (2025)
Too Noisy To Learn: Enhancing Data Quality for Code Review Comment Generation
por: Liu, Chunhua, et al.
Publicado: (2025)
por: Liu, Chunhua, et al.
Publicado: (2025)
Automatically Recommend Code Updates: Are We There Yet?
por: Liu, Yue, et al.
Publicado: (2022)
por: Liu, Yue, et al.
Publicado: (2022)
Encoding Version History Context for Better Code Representation
por: Nguyen, Huy, et al.
Publicado: (2024)
por: Nguyen, Huy, et al.
Publicado: (2024)
Enhancing Neural Code Representation with Additional Context
por: Nguyen, Huy, et al.
Publicado: (2025)
por: Nguyen, Huy, et al.
Publicado: (2025)
CodeReviewQA: The Code Review Comprehension Assessment for Large Language Models
por: Lin, Hong Yi, et al.
Publicado: (2025)
por: Lin, Hong Yi, et al.
Publicado: (2025)
Fine-grained Approaches for Confidence Calibration of LLMs in Automated Code Revision
por: Lin, Hong Yi, et al.
Publicado: (2026)
por: Lin, Hong Yi, et al.
Publicado: (2026)
RovoDev Code Reviewer: A Large-Scale Online Evaluation of LLM-based Code Review Automation at Atlassian
por: Tantithamthavorn, Kla, et al.
Publicado: (2026)
por: Tantithamthavorn, Kla, et al.
Publicado: (2026)
Should Code Models Learn Pedagogically? A Preliminary Evaluation of Curriculum Learning for Real-World Software Engineering Tasks
por: Khant, Kyi Shin, et al.
Publicado: (2025)
por: Khant, Kyi Shin, et al.
Publicado: (2025)
A Systematic Literature Review on Reasons and Approaches for Accurate Effort Estimations in Agile
por: Pasuksmit, Jirat, et al.
Publicado: (2024)
por: Pasuksmit, Jirat, et al.
Publicado: (2024)
Exploring the Potential of Large Language Models in Fine-Grained Review Comment Classification
por: Nguyen, Linh, et al.
Publicado: (2025)
por: Nguyen, Linh, et al.
Publicado: (2025)
Following Dragons: Code Review-Guided Fuzzing
por: Luu, Viet Hoang, et al.
Publicado: (2026)
por: Luu, Viet Hoang, et al.
Publicado: (2026)
Issue-Oriented Agent-Based Framework for Automated Review Comment Generation
por: Li, Shuochuan, et al.
Publicado: (2025)
por: Li, Shuochuan, et al.
Publicado: (2025)
Practitioners' Challenges and Perceptions of CI Build Failure Predictions at Atlassian
por: Hong, Yang, et al.
Publicado: (2024)
por: Hong, Yang, et al.
Publicado: (2024)
AI-Assisted Code Review as a Scaffold for Code Quality and Self-Regulated Learning: An Experience Report
por: Oliveira, Eduardo, et al.
Publicado: (2026)
por: Oliveira, Eduardo, et al.
Publicado: (2026)
Fine-Tuning and Prompt Engineering for Large Language Models-based Code Review Automation
por: Pornprasit, Chanathip, et al.
Publicado: (2024)
por: Pornprasit, Chanathip, et al.
Publicado: (2024)
Context-Augmented Code Generation Using Programming Knowledge Graphs
por: Seddik, Shahd, et al.
Publicado: (2026)
por: Seddik, Shahd, et al.
Publicado: (2026)
Adversarial Attacks on Code Models with Discriminative Graph Patterns
por: Nguyen, Thanh-Dat, et al.
Publicado: (2023)
por: Nguyen, Thanh-Dat, et al.
Publicado: (2023)
Human-In-The-Loop Software Development Agents: Challenges and Future Directions
por: Pasuksmit, Jirat, et al.
Publicado: (2025)
por: Pasuksmit, Jirat, et al.
Publicado: (2025)
When AI Models Become Dependencies: Studying the Evolution of Pre-Trained Model Reuse in Downstream Software Systems
por: Banyongrakkul, Peerachai, et al.
Publicado: (2026)
por: Banyongrakkul, Peerachai, et al.
Publicado: (2026)
From Release to Adoption: Challenges in Reusing Pre-trained AI Models for Downstream Developers
por: Banyongrakkul, Peerachai, et al.
Publicado: (2025)
por: Banyongrakkul, Peerachai, et al.
Publicado: (2025)
A Systematic Survey on Debugging Techniques for Machine Learning Systems
por: Nguyen, Thanh-Dat, et al.
Publicado: (2025)
por: Nguyen, Thanh-Dat, et al.
Publicado: (2025)
Breaking Changes in Software Ecosystems: A Systematic Literature Review
por: Chen, Juntao, et al.
Publicado: (2026)
por: Chen, Juntao, et al.
Publicado: (2026)
Comparing Human and LLM Generated Code: The Jury is Still Out!
por: Licorish, Sherlock A., et al.
Publicado: (2025)
por: Licorish, Sherlock A., et al.
Publicado: (2025)
Requirements-Driven Automated Software Testing: A Systematic Review
por: Wang, Fanyu, et al.
Publicado: (2025)
por: Wang, Fanyu, et al.
Publicado: (2025)
Human-In-the-Loop Software Development Agents
por: Takerngsaksiri, Wannita, et al.
Publicado: (2024)
por: Takerngsaksiri, Wannita, et al.
Publicado: (2024)
Automatic Programming: Large Language Models and Beyond
por: Lyu, Michael R., et al.
Publicado: (2024)
por: Lyu, Michael R., et al.
Publicado: (2024)
LLM-as-a-Judge for Reference-less Automatic Code Validation and Refinement for Natural Language to Bash in IT Automation
por: Vo, Ngoc Phuoc An, et al.
Publicado: (2025)
por: Vo, Ngoc Phuoc An, et al.
Publicado: (2025)
Ethics in AI through the Practitioner's View: A Grounded Theory Literature Review
por: Pant, Aastha, et al.
Publicado: (2022)
por: Pant, Aastha, et al.
Publicado: (2022)
Automated Code Review In Practice
por: Cihan, Umut, et al.
Publicado: (2024)
por: Cihan, Umut, et al.
Publicado: (2024)
Blended PC Peer Review Model: Process and Reflection
por: Tantithamthavorn, Chakkrit, et al.
Publicado: (2025)
por: Tantithamthavorn, Chakkrit, et al.
Publicado: (2025)
Students' Perspective on AI Code Completion: Benefits and Challenges
por: Takerngsaksiri, Wannita, et al.
Publicado: (2023)
por: Takerngsaksiri, Wannita, et al.
Publicado: (2023)
Ejemplares similares
-
AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection
por: Charoenwet, Wachiraphan, et al.
Publicado: (2026) -
Improving Automated Code Reviews: Learning from Experience
por: Lin, Hong Yi, et al.
Publicado: (2024) -
Toward Effective Secure Code Reviews: An Empirical Study of Security-Related Coding Weaknesses
por: Charoenwet, Wachiraphan, et al.
Publicado: (2023) -
Enhancing Code Review through Fuzzing and Likely Invariants
por: Charoenwet, Wachiraphan, et al.
Publicado: (2025) -
An Empirical Study of Static Analysis Tools for Secure Code Review
por: Charoenwet, Wachiraphan, et al.
Publicado: (2024)