ETF: An Entity Tracing Framework for Hallucination Detection in Code Summaries
Fuente:
arXiv
Saved in:
| Main Authors: | Maharaj, Kishan, Munigala, Vitobha, Tamilselvam, Srikanth G., Kumar, Prince, Sen, Sayandeep, Kodeswaran, Palani, Mishra, Abhijit, Bhattacharyya, Pushpak |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ConCodeEval: Evaluating Large Language Models for Code Constraints in Domain-Specific Languages
by: Kammakomati, Mehant, et al.
Published: (2024)
by: Kammakomati, Mehant, et al.
Published: (2024)
Robustness and Reasoning Fidelity of Large Language Models in Long-Context Code Question Answering
by: Maharaj, Kishan, et al.
Published: (2026)
by: Maharaj, Kishan, et al.
Published: (2026)
DocCGen: Document-based Controlled Code Generation
by: Pimparkhede, Sameer, et al.
Published: (2024)
by: Pimparkhede, Sameer, et al.
Published: (2024)
Emotion Entanglement and Bayesian Inference for Multi-Dimensional Emotion Understanding
by: Kotaprolu, Hemanth, et al.
Published: (2026)
by: Kotaprolu, Hemanth, et al.
Published: (2026)
Codellm-Devkit: A Framework for Contextualizing Code LLMs with Program Analysis Insights
by: Krishna, Rahul, et al.
Published: (2024)
by: Krishna, Rahul, et al.
Published: (2024)
Lost in Transcription: How Speech-to-Text Errors Derail Code Understanding
by: Havare, Jayant, et al.
Published: (2026)
by: Havare, Jayant, et al.
Published: (2026)
Understand the Implication: Learning to Think for Pragmatic Understanding
by: Sravanthi, Settaluri Lakshmi, et al.
Published: (2025)
by: Sravanthi, Settaluri Lakshmi, et al.
Published: (2025)
A Code Comprehension Benchmark for Large Language Models for Code
by: Havare, Jayant, et al.
Published: (2025)
by: Havare, Jayant, et al.
Published: (2025)
CodeSAM: Source Code Representation Learning by Infusing Self-Attention with Multi-Code-View Graphs
by: Mathai, Alex, et al.
Published: (2024)
by: Mathai, Alex, et al.
Published: (2024)
Enabling Communication via APIs for Mainframe Applications
by: Kanvar, Vini, et al.
Published: (2024)
by: Kanvar, Vini, et al.
Published: (2024)
Mental Disorder Classification via Temporal Representation of Text
by: Kumar, Raja, et al.
Published: (2024)
by: Kumar, Raja, et al.
Published: (2024)
Yaksha-Prashna: Understanding eBPF Bytecode Network Function Behavior
by: Singh, Animesh, et al.
Published: (2026)
by: Singh, Animesh, et al.
Published: (2026)
Read between the lines -- Functionality Extraction From READMEs
by: Kumar, Prince, et al.
Published: (2024)
by: Kumar, Prince, et al.
Published: (2024)
ScarfBench: A Benchmark for Cross-Framework Application Migration in Enterprise Java
by: Pavuluri, Advait, et al.
Published: (2026)
by: Pavuluri, Advait, et al.
Published: (2026)
Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries
by: Nirujan, Hinduja, et al.
Published: (2026)
by: Nirujan, Hinduja, et al.
Published: (2026)
BINGO! Simple Optimizers Win Big if Problems Collapse to a Few Buckets
by: Ganguly, Kishan Kumar, et al.
Published: (2025)
by: Ganguly, Kishan Kumar, et al.
Published: (2025)
How Low Can You Go? The Data-Light SE Challenge
by: Ganguly, Kishan Kumar, et al.
Published: (2025)
by: Ganguly, Kishan Kumar, et al.
Published: (2025)
Zoom, Don't Wander: Why Regional Search Outperforms Pareto Reasoning and Global Optimization in Budget-Constrained SBSE
by: Ganguly, Kishan Kumar, et al.
Published: (2026)
by: Ganguly, Kishan Kumar, et al.
Published: (2026)
From Verification to Herding: Exploiting Software's Sparsity of Influence
by: Menzies, Tim, et al.
Published: (2026)
by: Menzies, Tim, et al.
Published: (2026)
Code Hallucination
by: Rahman, Mirza Masfiqur, et al.
Published: (2024)
by: Rahman, Mirza Masfiqur, et al.
Published: (2024)
Context-aware Code Summary Generation
by: Su, Chia-Yi, et al.
Published: (2024)
by: Su, Chia-Yi, et al.
Published: (2024)
Designing and Implementing Robust Test Automation Frameworks using Cucumber BDD and Java
by: Srinivas, Srikanth, et al.
Published: (2025)
by: Srinivas, Srikanth, et al.
Published: (2025)
A Vulnerability Code Intent Summary Dataset
by: Huang, Yifan, et al.
Published: (2025)
by: Huang, Yifan, et al.
Published: (2025)
De-Hallucinator: Mitigating LLM Hallucinations in Code Generation Tasks via Iterative Grounding
by: Eghbali, Aryaz, et al.
Published: (2024)
by: Eghbali, Aryaz, et al.
Published: (2024)
MetricSynth: Framework for Aggregating DORA and KPI Metrics Across Multi-Platform Engineering
by: Jain, Pallav, et al.
Published: (2025)
by: Jain, Pallav, et al.
Published: (2025)
From Coverage to Causes: Data-Centric Fuzzing for JavaScript Engines
by: Ganguly, Kishan Kumar, et al.
Published: (2025)
by: Ganguly, Kishan Kumar, et al.
Published: (2025)
Studying Vulnerable Code Entities in R
by: Zhao, Zixiao, et al.
Published: (2024)
by: Zhao, Zixiao, et al.
Published: (2024)
Hallucinations in Code Change to Natural Language Generation: Prevalence and Evaluation of Detection Metrics
by: Liu, Chunhua, et al.
Published: (2025)
by: Liu, Chunhua, et al.
Published: (2025)
An Empirical Analysis of Static Analysis Methods for Detection and Mitigation of Code Library Hallucinations
by: Miranda-Pena, Clarissa, et al.
Published: (2026)
by: Miranda-Pena, Clarissa, et al.
Published: (2026)
Detecting and Correcting Hallucinations in LLM-Generated Code via Deterministic AST Analysis
by: Khati, Dipin, et al.
Published: (2026)
by: Khati, Dipin, et al.
Published: (2026)
RepoSummary: Feature-Oriented Summarization and Documentation Generation for Code Repositories
by: Zhu, Yifeng, et al.
Published: (2025)
by: Zhu, Yifeng, et al.
Published: (2025)
Reducing Hallucinations in LLM-Generated Code via Semantic Triangulation
by: Dai, Yihan, et al.
Published: (2025)
by: Dai, Yihan, et al.
Published: (2025)
HalluJudge: A Reference-Free Hallucination Detection for Context Misalignment in Code Review Automation
by: Tantithamthavorn, Kla, et al.
Published: (2026)
by: Tantithamthavorn, Kla, et al.
Published: (2026)
Trace Sampling 2.0: Code Knowledge Enhanced Span-level Sampling for Distributed Tracing
by: Wu, Yulun, et al.
Published: (2025)
by: Wu, Yulun, et al.
Published: (2025)
CodeHalu: Investigating Code Hallucinations in LLMs via Execution-based Verification
by: Tian, Yuchen, et al.
Published: (2024)
by: Tian, Yuchen, et al.
Published: (2024)
Towards Mitigating API Hallucination in Code Generated by LLMs with Hierarchical Dependency Aware
by: Chen, Yujia, et al.
Published: (2025)
by: Chen, Yujia, et al.
Published: (2025)
Code Clone Detection via an AlphaFold-Inspired Framework
by: Jia, Changguo, et al.
Published: (2025)
by: Jia, Changguo, et al.
Published: (2025)
TraceCoder: A Trace-Driven Multi-Agent Framework for Automated Debugging of LLM-Generated Code
by: Huang, Jiangping, et al.
Published: (2026)
by: Huang, Jiangping, et al.
Published: (2026)
TraceRAG: A LLM-Based Framework for Explainable Android Malware Detection and Behavior Analysis
by: Zhang, Guangyu, et al.
Published: (2025)
by: Zhang, Guangyu, et al.
Published: (2025)
CodeMirage: Hallucinations in Code Generated by Large Language Models
by: Agarwal, Vibhor, et al.
Published: (2024)
by: Agarwal, Vibhor, et al.
Published: (2024)
Similar Items
-
ConCodeEval: Evaluating Large Language Models for Code Constraints in Domain-Specific Languages
by: Kammakomati, Mehant, et al.
Published: (2024) -
Robustness and Reasoning Fidelity of Large Language Models in Long-Context Code Question Answering
by: Maharaj, Kishan, et al.
Published: (2026) -
DocCGen: Document-based Controlled Code Generation
by: Pimparkhede, Sameer, et al.
Published: (2024) -
Emotion Entanglement and Bayesian Inference for Multi-Dimensional Emotion Understanding
by: Kotaprolu, Hemanth, et al.
Published: (2026) -
Codellm-Devkit: A Framework for Contextualizing Code LLMs with Program Analysis Insights
by: Krishna, Rahul, et al.
Published: (2024)