Towards More Trustworthy and Interpretable LLMs for Code through Syntax-Grounded Explanations
Fuente:
arXiv
Salvato in:
| Autori principali: | Palacio, David N., Rodriguez-Cardenas, Daniel, Velasco, Alejandro, Khati, Dipin, Moran, Kevin, Poshyvanyk, Denys |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation
di: Khati, Dipin, et al.
Pubblicazione: (2025)
di: Khati, Dipin, et al.
Pubblicazione: (2025)
A Causal Perspective on Measuring, Explaining and Mitigating Smells in LLM-Generated Code
di: Velasco, Alejandro, et al.
Pubblicazione: (2025)
di: Velasco, Alejandro, et al.
Pubblicazione: (2025)
Detecting and Correcting Hallucinations in LLM-Generated Code via Deterministic AST Analysis
di: Khati, Dipin, et al.
Pubblicazione: (2026)
di: Khati, Dipin, et al.
Pubblicazione: (2026)
Tricky$^2$: Towards a Benchmark for Evaluating Human and LLM Error Interactions
di: Granger, Cole, et al.
Pubblicazione: (2026)
di: Granger, Cole, et al.
Pubblicazione: (2026)
Mapping the Trust Terrain: LLMs in Software Engineering -- Insights and Perspectives
di: Khati, Dipin, et al.
Pubblicazione: (2025)
di: Khati, Dipin, et al.
Pubblicazione: (2025)
Which Syntactic Capabilities Are Statistically Learned by Masked Language Models for Code?
di: Velasco, Alejandro, et al.
Pubblicazione: (2024)
di: Velasco, Alejandro, et al.
Pubblicazione: (2024)
Toward a Theory of Causation for Interpreting Neural Code Models
di: Palacio, David N., et al.
Pubblicazione: (2023)
di: Palacio, David N., et al.
Pubblicazione: (2023)
On Interpreting the Effectiveness of Unsupervised Software Traceability with Information Theory
di: Palacio, David N., et al.
Pubblicazione: (2024)
di: Palacio, David N., et al.
Pubblicazione: (2024)
Towards Comprehensive Benchmarking Infrastructure for LLMs In Software Engineering
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2026)
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2026)
SnipGen: A Mining Repository Framework for Evaluating LLMs for Code
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2025)
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2025)
Towards More Trustworthy Deep Code Models by Enabling Out-of-Distribution Detection
di: Yan, Yanfu, et al.
Pubblicazione: (2025)
di: Yan, Yanfu, et al.
Pubblicazione: (2025)
Measuring Emergent Capabilities of LLMs for Software Engineering: How Far Are We?
di: O'Brien, Conor, et al.
Pubblicazione: (2024)
di: O'Brien, Conor, et al.
Pubblicazione: (2024)
Towards Enabling An Artificial Self-Construction Software Life-cycle via Autopoietic Architectures
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2026)
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2026)
How Propense Are Large Language Models at Producing Code Smells? A Benchmarking Study
di: Velasco, Alejandro, et al.
Pubblicazione: (2024)
di: Velasco, Alejandro, et al.
Pubblicazione: (2024)
Toward Neurosymbolic Program Comprehension
di: Velasco, Alejandro, et al.
Pubblicazione: (2025)
di: Velasco, Alejandro, et al.
Pubblicazione: (2025)
Testing Practices, Challenges, and Developer Perspectives in Open-Source IoT Platforms
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2025)
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2025)
On the Effectiveness of LLM-as-a-judge for Code Generation and Summarization
di: Crupi, Giuseppe, et al.
Pubblicazione: (2025)
di: Crupi, Giuseppe, et al.
Pubblicazione: (2025)
How Do Semantically Equivalent Code Transformations Impact Membership Inference on LLMs for Code?
di: Yang, Hua, et al.
Pubblicazione: (2025)
di: Yang, Hua, et al.
Pubblicazione: (2025)
A Path Less Traveled: Reimagining Software Engineering Automation via a Neurosymbolic Paradigm
di: Mastropaolo, Antonio, et al.
Pubblicazione: (2025)
di: Mastropaolo, Antonio, et al.
Pubblicazione: (2025)
"False negative -- that one is going to kill you": Understanding Industry Perspectives of Static Analysis based Security Testing
di: Ami, Amit Seal, et al.
Pubblicazione: (2023)
di: Ami, Amit Seal, et al.
Pubblicazione: (2023)
Understanding Privacy Risks in Code Models Through Training Dynamics: A Causal Approach
di: Yang, Hua, et al.
Pubblicazione: (2025)
di: Yang, Hua, et al.
Pubblicazione: (2025)
On the Generalizability of Transformer Models to Code Completions of Different Lengths
di: Cooper, Nathan, et al.
Pubblicazione: (2025)
di: Cooper, Nathan, et al.
Pubblicazione: (2025)
Semantic GUI Scene Learning and Video Alignment for Detecting Duplicate Video-based Bug Reports
di: Yan, Yanfu, et al.
Pubblicazione: (2024)
di: Yan, Yanfu, et al.
Pubblicazione: (2024)
Rethinking Software Empirical Studies with Structural Causal Models
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2026)
di: Rodriguez-Cardenas, Daniel, et al.
Pubblicazione: (2026)
Mutation-based Evaluation of Cryptographic API Misuse Detectors
di: Ami, Amit Seal, et al.
Pubblicazione: (2021)
di: Ami, Amit Seal, et al.
Pubblicazione: (2021)
Challenges and Practices in Quantum Software Testing and Debugging: Insights from Practitioners
di: Zappin, Jake, et al.
Pubblicazione: (2025)
di: Zappin, Jake, et al.
Pubblicazione: (2025)
When Quantum Meets Classical: Characterizing Hybrid Quantum-Classical Issues Discussed in Developer Forums
di: Zappin, Jake, et al.
Pubblicazione: (2024)
di: Zappin, Jake, et al.
Pubblicazione: (2024)
Bridging the Quantum Divide: Aligning Academic and Industry Goals in Software Engineering
di: Zappin, Jake, et al.
Pubblicazione: (2025)
di: Zappin, Jake, et al.
Pubblicazione: (2025)
Prompting in Practice: Investigating Software Practitioners' Use of Generative AI Tools
di: Otten, Daniel, et al.
Pubblicazione: (2025)
di: Otten, Daniel, et al.
Pubblicazione: (2025)
Perspective of Software Engineering Researchers on Machine Learning Practices Regarding Research, Review, and Education
di: Mojica-Hanke, Anamaria, et al.
Pubblicazione: (2024)
di: Mojica-Hanke, Anamaria, et al.
Pubblicazione: (2024)
Towards Trustworthy LLMs for Code: A Data-Centric Synergistic Auditing Framework
di: Wang, Chong, et al.
Pubblicazione: (2024)
di: Wang, Chong, et al.
Pubblicazione: (2024)
SynConfRoute: Syntax-Aware Routing for Efficient Code Completion with Small CodeLLMs
di: Thangarajah, Kishanthan, et al.
Pubblicazione: (2026)
di: Thangarajah, Kishanthan, et al.
Pubblicazione: (2026)
"The Law Doesn't Work Like a Computer": Exploring Software Licensing Issues Faced by Legal Practitioners
di: Wintersgill, Nathan, et al.
Pubblicazione: (2024)
di: Wintersgill, Nathan, et al.
Pubblicazione: (2024)
Towards a Science of Causal Interpretability in Deep Learning for Software Engineering
di: Palacio, David N.
Pubblicazione: (2025)
di: Palacio, David N.
Pubblicazione: (2025)
SGCR: A Specification-Grounded Framework for Trustworthy LLM Code Review
di: Wang, Kai, et al.
Pubblicazione: (2025)
di: Wang, Kai, et al.
Pubblicazione: (2025)
BOMs Away! Inside the Minds of Stakeholders: A Comprehensive Study of Bills of Materials for Software Systems
di: Stalnaker, Trevor, et al.
Pubblicazione: (2023)
di: Stalnaker, Trevor, et al.
Pubblicazione: (2023)
"Don't Be Afraid, Just Learn": Insights from Industry Practitioners to Prepare Software Engineers in the Age of Generative AI
di: Otten, Daniel, et al.
Pubblicazione: (2026)
di: Otten, Daniel, et al.
Pubblicazione: (2026)
Toward Explaining Large Language Models in Software Engineering Tasks
di: Vitale, Antonio, et al.
Pubblicazione: (2025)
di: Vitale, Antonio, et al.
Pubblicazione: (2025)
GUIPilot: A Consistency-based Mobile GUI Testing Approach for Detecting Application-specific Bugs
di: Liu, Ruofan, et al.
Pubblicazione: (2025)
di: Liu, Ruofan, et al.
Pubblicazione: (2025)
Developers' Perspectives on Software Licensing: Current Practices, Challenges, and Tools
di: Wintersgill, Nathan, et al.
Pubblicazione: (2025)
di: Wintersgill, Nathan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation
di: Khati, Dipin, et al.
Pubblicazione: (2025) -
A Causal Perspective on Measuring, Explaining and Mitigating Smells in LLM-Generated Code
di: Velasco, Alejandro, et al.
Pubblicazione: (2025) -
Detecting and Correcting Hallucinations in LLM-Generated Code via Deterministic AST Analysis
di: Khati, Dipin, et al.
Pubblicazione: (2026) -
Tricky$^2$: Towards a Benchmark for Evaluating Human and LLM Error Interactions
di: Granger, Cole, et al.
Pubblicazione: (2026) -
Mapping the Trust Terrain: LLMs in Software Engineering -- Insights and Perspectives
di: Khati, Dipin, et al.
Pubblicazione: (2025)