CoHalLo: code hallucination localization via probing hidden layer vector
Fuente:
arXiv
Saved in:
| Main Authors: | Jia, Nan, Sang, Wangchao, Lin, Pengfei, Chen, Xiangping, Huang, Yuan, Liu, Yi, Li, Mingliang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Are your comments outdated? Towards automatically detecting code-comment consistency
by: Huang, Yuan, et al.
Published: (2024)
by: Huang, Yuan, et al.
Published: (2024)
Comment Traps: How Defective Commented-out Code Augment Defects in AI-Assisted Code Generation
by: Huang, Yuan, et al.
Published: (2025)
by: Huang, Yuan, et al.
Published: (2025)
Exposing the hidden layers and interplay in the quantum software stack
by: Stirbu, Vlad, et al.
Published: (2024)
by: Stirbu, Vlad, et al.
Published: (2024)
Generative Software Engineering
by: Huang, Yuan, et al.
Published: (2024)
by: Huang, Yuan, et al.
Published: (2024)
Can Adjusting Hyperparameters Lead to Green Deep Learning: An Empirical Study on Correlations between Hyperparameters and Energy Consumption of Deep Learning Models
by: Wang, Taoran, et al.
Published: (2026)
by: Wang, Taoran, et al.
Published: (2026)
ACL-Verbatim: hallucination-free question answering for research
by: Recski, Gábor, et al.
Published: (2026)
by: Recski, Gábor, et al.
Published: (2026)
Mirror Matrix on the Wall: coding and vector notation as tools for introspection
by: Araújo, Leonardo
Published: (2024)
by: Araújo, Leonardo
Published: (2024)
Multi-CoLoR: Context-Aware Localization and Reasoning across Multi-Language Codebases
by: Vats, Indira, et al.
Published: (2026)
by: Vats, Indira, et al.
Published: (2026)
MoCo: Fuzzing Deep Learning Libraries via Assembling Code
by: Ji, Pin, et al.
Published: (2024)
by: Ji, Pin, et al.
Published: (2024)
Revisiting Evolutionary Program Repair via Code Language Model
by: Wang, Yunan, et al.
Published: (2024)
by: Wang, Yunan, et al.
Published: (2024)
ExploraCoder: Advancing code generation for multiple unseen APIs via planning and chained exploration
by: Wang, Yunkun, et al.
Published: (2024)
by: Wang, Yunkun, et al.
Published: (2024)
Code Clone Detection via an AlphaFold-Inspired Framework
by: Jia, Changguo, et al.
Published: (2025)
by: Jia, Changguo, et al.
Published: (2025)
LLM-based Unit Test Generation via Property Retrieval
by: Zhang, Zhe, et al.
Published: (2024)
by: Zhang, Zhe, et al.
Published: (2024)
Zero-Shot Code Representation Learning via Prompt Tuning
by: Cui, Nan, et al.
Published: (2024)
by: Cui, Nan, et al.
Published: (2024)
CoReQA: Uncovering Potentials of Language Models in Code Repository Question Answering
by: Chen, Jialiang, et al.
Published: (2025)
by: Chen, Jialiang, et al.
Published: (2025)
Enhancing Software Maintenance: A Learning to Rank Approach for Co-changed Method Identification
by: Jia, Yiping, et al.
Published: (2024)
by: Jia, Yiping, et al.
Published: (2024)
InferLog: Accelerating LLM Inference for Online Log Parsing via ICL-oriented Prefix Caching
by: Wang, Yilun, et al.
Published: (2025)
by: Wang, Yilun, et al.
Published: (2025)
LoCoML: A Framework for Real-World ML Inference Pipelines
by: Maddireddy, Kritin, et al.
Published: (2025)
by: Maddireddy, Kritin, et al.
Published: (2025)
Mint: Cost-Efficient Tracing with All Requests Collection via Commonality and Variability Analysis
by: Huang, Haiyu, et al.
Published: (2024)
by: Huang, Haiyu, et al.
Published: (2024)
FaaSRCA: Full Lifecycle Root Cause Analysis for Serverless Applications
by: Huang, Jin, et al.
Published: (2024)
by: Huang, Jin, et al.
Published: (2024)
CoCoEvo: Co-Evolution of Programs and Test Cases to Enhance Code Generation
by: Li, Kefan, et al.
Published: (2025)
by: Li, Kefan, et al.
Published: (2025)
Low-code and no-code with BESSER to create and deploy smart web applications
by: Alfonso, Iván, et al.
Published: (2026)
by: Alfonso, Iván, et al.
Published: (2026)
CONNECTOR: Enhancing the Traceability of Decentralized Bridge Applications via Automatic Cross-chain Transaction Association
by: Lin, Dan, et al.
Published: (2024)
by: Lin, Dan, et al.
Published: (2024)
Improving Deep Assertion Generation via Fine-Tuning Retrieval-Augmented Pre-trained Language Models
by: Zhang, Quanjun, et al.
Published: (2025)
by: Zhang, Quanjun, et al.
Published: (2025)
May the Feedback Be with You! Unlocking the Power of Feedback-Driven Deep Learning Framework Fuzzing via LLMs
by: Yang, Shaoyu, et al.
Published: (2025)
by: Yang, Shaoyu, et al.
Published: (2025)
LoCoBench-Agent: An Interactive Benchmark for LLM Agents in Long-Context Software Engineering
by: Qiu, Jielin, et al.
Published: (2025)
by: Qiu, Jielin, et al.
Published: (2025)
SmartState: Detecting State-Reverting Vulnerabilities in Smart Contracts via Fine-Grained State-Dependency Analysis
by: Liao, Zeqin, et al.
Published: (2024)
by: Liao, Zeqin, et al.
Published: (2024)
Cloud-OpsBench: A Reproducible Benchmark for Agentic Root Cause Analysis in Cloud Systems
by: Wang, Yilun, et al.
Published: (2026)
by: Wang, Yilun, et al.
Published: (2026)
Measuring how changes in code readability attributes affect code quality evaluation by Large Language Models
by: Simoes, Igor Regis da Silva, et al.
Published: (2025)
by: Simoes, Igor Regis da Silva, et al.
Published: (2025)
LoCoBench: A Benchmark for Long-Context Large Language Models in Complex Software Engineering
by: Qiu, Jielin, et al.
Published: (2025)
by: Qiu, Jielin, et al.
Published: (2025)
CoEdPilot: Recommending Code Edits with Learned Prior Edit Relevance, Project-wise Awareness, and Interactive Nature
by: Liu, Chenyan, et al.
Published: (2024)
by: Liu, Chenyan, et al.
Published: (2024)
SLA-Awareness for AI-assisted coding
by: Thangarajah, Kishanthan, et al.
Published: (2025)
by: Thangarajah, Kishanthan, et al.
Published: (2025)
ClarEval: A Benchmark for Evaluating Clarification Skills of Code Agents under Ambiguous Instructions
by: Li, Jialin, et al.
Published: (2026)
by: Li, Jialin, et al.
Published: (2026)
Context-aware Code Summary Generation
by: Su, Chia-Yi, et al.
Published: (2024)
by: Su, Chia-Yi, et al.
Published: (2024)
FuSeBMC v4: Improving code coverage with smart seeds via BMC, fuzzing and static analysis
by: Alshmrany, Kaled M., et al.
Published: (2022)
by: Alshmrany, Kaled M., et al.
Published: (2022)
TENET: Leveraging Tests Beyond Validation for Code Generation
by: Hu, Yiran, et al.
Published: (2025)
by: Hu, Yiran, et al.
Published: (2025)
Requirements-Based Test Generation: A Comprehensive Survey
by: Yang, Zhenzhen, et al.
Published: (2025)
by: Yang, Zhenzhen, et al.
Published: (2025)
An Empirical Framework for Evaluating Semantic Preservation Using Hugging Face
by: Jia, Nan, et al.
Published: (2025)
by: Jia, Nan, et al.
Published: (2025)
SWE-PolyBench: A multi-language benchmark for repository level evaluation of coding agents
by: Rashid, Muhammad Shihab, et al.
Published: (2025)
by: Rashid, Muhammad Shihab, et al.
Published: (2025)
Statistical complexity of software systems represented as multi-layer networks
by: Žižka, Jan
Published: (2025)
by: Žižka, Jan
Published: (2025)
Similar Items
-
Are your comments outdated? Towards automatically detecting code-comment consistency
by: Huang, Yuan, et al.
Published: (2024) -
Comment Traps: How Defective Commented-out Code Augment Defects in AI-Assisted Code Generation
by: Huang, Yuan, et al.
Published: (2025) -
Exposing the hidden layers and interplay in the quantum software stack
by: Stirbu, Vlad, et al.
Published: (2024) -
Generative Software Engineering
by: Huang, Yuan, et al.
Published: (2024) -
Can Adjusting Hyperparameters Lead to Green Deep Learning: An Empirical Study on Correlations between Hyperparameters and Energy Consumption of Deep Learning Models
by: Wang, Taoran, et al.
Published: (2026)