AIRA: AI-Induced Risk Audit: A Structured Inspection Framework for AI-Generated Code
Fuente:
arXiv
Salvato in:
| Autore principale: | Parris, William M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures
di: Jahan, Sigma, et al.
Pubblicazione: (2026)
di: Jahan, Sigma, et al.
Pubblicazione: (2026)
Validating Solidity Code Defects using Symbolic and Concrete Execution powered by Large Language Models
di: Susan, Ştefan-Claudiu, et al.
Pubblicazione: (2025)
di: Susan, Ştefan-Claudiu, et al.
Pubblicazione: (2025)
The Specification as Quality Gate: Three Hypotheses on AI-Assisted Code Review
di: Zietsman, Christo
Pubblicazione: (2026)
di: Zietsman, Christo
Pubblicazione: (2026)
RepoAudit: An Autonomous LLM-Agent for Repository-Level Code Auditing
di: Guo, Jinyao, et al.
Pubblicazione: (2025)
di: Guo, Jinyao, et al.
Pubblicazione: (2025)
CodeTracer: Towards Traceable Agent States
di: Li, Han, et al.
Pubblicazione: (2026)
di: Li, Han, et al.
Pubblicazione: (2026)
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
di: Hu, Yuelin, et al.
Pubblicazione: (2026)
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
di: Ravi, Ravin, et al.
Pubblicazione: (2026)
di: Ravi, Ravin, et al.
Pubblicazione: (2026)
Towards a Probabilistic Framework for Analyzing and Improving LLM-Enabled Software
di: Baldonado, Juan Manuel, et al.
Pubblicazione: (2025)
di: Baldonado, Juan Manuel, et al.
Pubblicazione: (2025)
On the Mistaken Assumption of Interchangeable Deep Reinforcement Learning Implementations
di: Hundal, Rajdeep Singh, et al.
Pubblicazione: (2025)
di: Hundal, Rajdeep Singh, et al.
Pubblicazione: (2025)
Towards Explainable Test Case Prioritisation with Learning-to-Rank Models
di: Ramírez, Aurora, et al.
Pubblicazione: (2024)
di: Ramírez, Aurora, et al.
Pubblicazione: (2024)
Emergent Formal Verification: How an Autonomous AI Ecosystem Independently Discovered SMT-Based Safety Across Six Domains
di: Untila, Octavian
Pubblicazione: (2026)
di: Untila, Octavian
Pubblicazione: (2026)
Iterative Audit Convergence in LLM-Managed Multi-Agent Systems: A Case Study in Prompt Engineering Quality Assurance
di: Calboreanu, Elias
Pubblicazione: (2026)
di: Calboreanu, Elias
Pubblicazione: (2026)
Monitoring Agentic Systems Before They're Reliable
di: Boston, Marisa Ferrara, et al.
Pubblicazione: (2026)
di: Boston, Marisa Ferrara, et al.
Pubblicazione: (2026)
Understanding and Detecting Flaky Builds in GitHub Actions
di: Ge, Wenhao, et al.
Pubblicazione: (2026)
di: Ge, Wenhao, et al.
Pubblicazione: (2026)
Automated structural testing of LLM-based agents: methods, framework, and case studies
di: Kohl, Jens, et al.
Pubblicazione: (2026)
di: Kohl, Jens, et al.
Pubblicazione: (2026)
L2MAC: Large Language Model Automatic Computer for Extensive Code Generation
di: Holt, Samuel, et al.
Pubblicazione: (2023)
di: Holt, Samuel, et al.
Pubblicazione: (2023)
AI Bill of Materials and Beyond: Systematizing Security Assurance through the AI Risk Scanning (AIRS) Framework
di: Nathanson, Samuel, et al.
Pubblicazione: (2025)
di: Nathanson, Samuel, et al.
Pubblicazione: (2025)
A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification
di: Odmark, Joshua, et al.
Pubblicazione: (2026)
di: Odmark, Joshua, et al.
Pubblicazione: (2026)
TerraFormer: Automated Infrastructure-as-Code with LLMs Fine-Tuned via Policy-Guided Verifier Feedback
di: Jana, Prithwish, et al.
Pubblicazione: (2026)
di: Jana, Prithwish, et al.
Pubblicazione: (2026)
MFH: A Multi-faceted Heuristic Algorithm Selection Approach for Software Verification
di: Su, Jie, et al.
Pubblicazione: (2025)
di: Su, Jie, et al.
Pubblicazione: (2025)
Provable Fairness Repair for Deep Neural Networks
di: Ma, Jianan, et al.
Pubblicazione: (2026)
di: Ma, Jianan, et al.
Pubblicazione: (2026)
Constitutional Spec-Driven Development: Enforcing Security by Construction in AI-Assisted Code Generation
di: Marri, Srinivas Rao
Pubblicazione: (2026)
di: Marri, Srinivas Rao
Pubblicazione: (2026)
Comparing Human and LLM Generated Code: The Jury is Still Out!
di: Licorish, Sherlock A., et al.
Pubblicazione: (2025)
di: Licorish, Sherlock A., et al.
Pubblicazione: (2025)
Generative AI and the Transformation of Software Development Practices
di: Acharya, Vivek
Pubblicazione: (2025)
di: Acharya, Vivek
Pubblicazione: (2025)
LLMDFA: Analyzing Dataflow in Code with Large Language Models
di: Wang, Chengpeng, et al.
Pubblicazione: (2024)
di: Wang, Chengpeng, et al.
Pubblicazione: (2024)
AgentEval: DAG-Structured Step-Level Evaluation for Agentic Workflows with Error Propagation Tracking
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
Context Engineering for Multi-Agent LLM Code Assistants Using Elicit, NotebookLM, ChatGPT, and Claude Code
di: Haseeb, Muhammad
Pubblicazione: (2025)
di: Haseeb, Muhammad
Pubblicazione: (2025)
A Framework for Assessing AI Agent Decisions and Outcomes in AutoML Pipelines
di: Du, Gaoyuan, et al.
Pubblicazione: (2026)
di: Du, Gaoyuan, et al.
Pubblicazione: (2026)
Orion: Fuzzing Workflow Automation
di: Bazalii, Max, et al.
Pubblicazione: (2025)
di: Bazalii, Max, et al.
Pubblicazione: (2025)
Adaptive and AI-Augmented Security Testing: A Systematic Survey of Program Analysis, Feedback-Driven Testing, and Hybrid Learning-Based Approaches
di: Wienczkowski, Michael
Pubblicazione: (2026)
di: Wienczkowski, Michael
Pubblicazione: (2026)
Dual-Process Scaffold Reasoning for Enhancing LLM Code Debugging
di: Hsieh, Po-Chung, et al.
Pubblicazione: (2025)
di: Hsieh, Po-Chung, et al.
Pubblicazione: (2025)
Testing SSD Firmware with State Data-Aware Fuzzing: Accelerating Coverage in Nondeterministic I/O Environments
di: Yoon, Gangho, et al.
Pubblicazione: (2025)
di: Yoon, Gangho, et al.
Pubblicazione: (2025)
SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark
di: Yu, Boxi, et al.
Pubblicazione: (2026)
di: Yu, Boxi, et al.
Pubblicazione: (2026)
Combined Program Analysis Techniques: A Systematic Mapping Study
di: Braione, Pietro, et al.
Pubblicazione: (2026)
di: Braione, Pietro, et al.
Pubblicazione: (2026)
Automatically Detecting Numerical Instability in Machine Learning Applications via Soft Assertions
di: Sharmin, Shaila, et al.
Pubblicazione: (2025)
di: Sharmin, Shaila, et al.
Pubblicazione: (2025)
ATLAS: A Layered Constraint-Guided Framework for Structured Artifact Generation in LLM-Assisted MDE
di: Ma, Tong, et al.
Pubblicazione: (2025)
di: Ma, Tong, et al.
Pubblicazione: (2025)
Structural Quality Gaps in Practitioner AI Governance Prompts: An Empirical Study Using a Five-Principle Evaluation Framework
di: Zietsman, Christo
Pubblicazione: (2026)
di: Zietsman, Christo
Pubblicazione: (2026)
Automated Vulnerability Detection Using Deep Learning Technique
di: Yang, Guan-Yan, et al.
Pubblicazione: (2024)
di: Yang, Guan-Yan, et al.
Pubblicazione: (2024)
N-Version Assessment and Enhancement of Generative AI
di: Kessel, Marcus, et al.
Pubblicazione: (2024)
di: Kessel, Marcus, et al.
Pubblicazione: (2024)
VulScribeR: Exploring RAG-based Vulnerability Augmentation with LLMs
di: Daneshvar, Seyed Shayan, et al.
Pubblicazione: (2024)
di: Daneshvar, Seyed Shayan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures
di: Jahan, Sigma, et al.
Pubblicazione: (2026) -
Validating Solidity Code Defects using Symbolic and Concrete Execution powered by Large Language Models
di: Susan, Ştefan-Claudiu, et al.
Pubblicazione: (2025) -
The Specification as Quality Gate: Three Hypotheses on AI-Assisted Code Review
di: Zietsman, Christo
Pubblicazione: (2026) -
RepoAudit: An Autonomous LLM-Agent for Repository-Level Code Auditing
di: Guo, Jinyao, et al.
Pubblicazione: (2025) -
CodeTracer: Towards Traceable Agent States
di: Li, Han, et al.
Pubblicazione: (2026)