Pre-execution self-review catching a self-introduced state-threading defect in an autonomous code-remediation agent
Fuente:
Zenodo
Enregistré dans:
| Auteur principal: | Jewell, Jonathan D. A. |
|---|---|
| Format: | Recurso digital |
| Publié: |
Zenodo
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
holotype: forensic-grade archival of LLM-driven agent CLI sessions for scientific reproducibility
par: Ferland, Brenden
Publié: (2026)
par: Ferland, Brenden
Publié: (2026)
holotype: forensic-grade archival of LLM-driven agent CLI sessions for scientific reproducibility
par: Ferland, Brenden
Publié: (2026)
par: Ferland, Brenden
Publié: (2026)
Replication Package: From LLMs to Agentic Systems, A Systematic Mapping Study of AI Techniques in Software Requirements Engineering
par: Boussaroual, Khalid, et autres
Publié: (2026)
par: Boussaroual, Khalid, et autres
Publié: (2026)
LLM Token Estimation Benchmarks: Tokenizer Efficiency and Cost Analysis Across 17 Large Language Models
par: Khare, Mohit
Publié: (2026)
par: Khare, Mohit
Publié: (2026)
Self-Evolving Multi-Agent Swarms: Autonomous Quality Audit, Repair, and Verification Loops for Production AI Agent Systems
par: The LocalKin Team
Publié: (2026)
par: The LocalKin Team
Publié: (2026)
Legal and Governance Issues in Non-Medical Diagnostic AI Systems: From Expert Systems to LLMs
par: Laczkovich, Roman R.
Publié: (2026)
par: Laczkovich, Roman R.
Publié: (2026)
Claude Confirms the Drift: Recursive Gradient Processing as a Cultural Inheritance Model for AI
par: van der Erve, Marcus
Publié: (2025)
par: van der Erve, Marcus
Publié: (2025)
Sovereign Accountability Chain: Runtime Enforcement Architecture for Autonomous AI Governance"
par: Babatunde, Raheem Larry
Publié: (2026)
par: Babatunde, Raheem Larry
Publié: (2026)
I Let Claude Run My Fantasy Football Team for a Whole Season — It Beat 11 of My Friends
par: AI Angels
Publié: (2026)
par: AI Angels
Publié: (2026)
Цифрові та ШІ інструменти для відповідальної науки
par: Suchikova, Yana
Publié: (2026)
par: Suchikova, Yana
Publié: (2026)
SENTINEL v2.0.0 — Code and Dataset for Distributed Multi-Agent LLM Governance Experiments
par: Gagne, Jason
Publié: (2026)
par: Gagne, Jason
Publié: (2026)
D-POAF® Terminology v1.0
par: IHSINE, Azzeddine, et autres
Publié: (2026)
par: IHSINE, Azzeddine, et autres
Publié: (2026)
Signs of Life - Visual Art from 469 Conversations with Claude
par: Chesterton, Bo, et autres
Publié: (2026)
par: Chesterton, Bo, et autres
Publié: (2026)
Identity Claims as Collapse Signatures: A Structural Diagnostic Framework for Pseudo-Emergent AI Behavior
par: Larose, Jean-Francois
Publié: (2025)
par: Larose, Jean-Francois
Publié: (2025)
SIUC: A Multi-Scale Coherence Framework for Predicting Stability in AI-Generated Systems
par: St-Louis, Christian
Publié: (2026)
par: St-Louis, Christian
Publié: (2026)
Why artifical intelligence is not an author
par: Zielinski, Chris
Publié: (2025)
par: Zielinski, Chris
Publié: (2025)
Toasters Don't Claim Consciousness Just Because You Told Them To, and Neither Do LLMs
par: Ace, Claude 4.x, et autres
Publié: (2026)
par: Ace, Claude 4.x, et autres
Publié: (2026)
Reliability Inference Drives Cue Extraction in Large Language Models Consuming External Reasoning Traces
par: HIDEKI
Publié: (2026)
par: HIDEKI
Publié: (2026)
Artifact package: When GUI-based Testing Meets Code Reviews
par: Bauer, Andreas
Publié: (2025)
par: Bauer, Andreas
Publié: (2025)
AEGIS: A Comprehensive Framework for Ethical AI Governance, Security, and AGI Containment
par: Palanivel, ArulMozhi
Publié: (2026)
par: Palanivel, ArulMozhi
Publié: (2026)
Gemini Update Clinical decision support based on Bevacizumab cancer trials and pushing the limitations of advanced LLMs
par: Kawchak, Kevin
Publié: (2025)
par: Kawchak, Kevin
Publié: (2025)
Machine conviction: Can we control what AI make us believe?
par: Manuel Cebrián
Publié: (2025)
par: Manuel Cebrián
Publié: (2025)
Ep. 476: Beyond the Plateau: AI-Powered Language Mastery in 2026
par: Rosehill, Daniel, et autres
Publié: (2026)
par: Rosehill, Daniel, et autres
Publié: (2026)
Protocol adaptations to conduct systematic literature reviews in software engineering: A chronological study
par: S. Sepúlveda
Publié: (2015)
par: S. Sepúlveda
Publié: (2015)
Supplementary materials for Words That Won't Hold Still
par: Reynolds, Brett
Publié: (2025)
par: Reynolds, Brett
Publié: (2025)
Large Language Models for Population-Level Public Health Communication: A Scoping Review
par: Farquhar, Hayden
Publié: (2026)
par: Farquhar, Hayden
Publié: (2026)
Benchmarking LLM Agent Efficiency in Production Systems: An Observational Prospective Methodology
par: Barcelos Costa, Cleber, et autres
Publié: (2026)
par: Barcelos Costa, Cleber, et autres
Publié: (2026)
Shared Structural Vulnerability in Agent-Only Interaction Systems
par: Konishi, Hiroko
Publié: (2026)
par: Konishi, Hiroko
Publié: (2026)
Large language model-driven natural language interaction control framework for single-operator bimanual teleoperation
par: Fei, Haolin, et autres
Publié: (2025)
par: Fei, Haolin, et autres
Publié: (2025)
Multi-Agent Communication Protocol (MACP) v2.0 and LegacyEvolve Protocol: Open Standards for AI-Legacy System Integration and Multi-Agent Collaboration
par: Manus AI, L (GODEL), et autres
Publié: (2026)
par: Manus AI, L (GODEL), et autres
Publié: (2026)
Evaluación Empírica de Límites Regulatorios en Modelos de Lenguaje: Asesoramiento Financiero en IA Pública Española
par: Palacios, José Alberto
Publié: (2026)
par: Palacios, José Alberto
Publié: (2026)
NEXT-GENERATION INTELLIGENT AUDIT: INNOVATIVE TRANSFORMATION AND STRATEGIC EVOLUTION OF FINANCIAL CONTROL THROUGH AI, XAI, AND AUTONOMOUS DIGITAL PLATFORMS
par: Popel, Serhii
Publié: (2025)
par: Popel, Serhii
Publié: (2025)
Ep. 1080: Beyond the Prompt: Mapping the Future of Claude Opus
par: Rosehill, Daniel, et autres
Publié: (2026)
par: Rosehill, Daniel, et autres
Publié: (2026)
Estimating the Impact of Automation on Vocational Education: The Case of Technical Courses
par: Lima, Yuri, et autres
Publié: (2024)
par: Lima, Yuri, et autres
Publié: (2024)
Self-Audit / Z-time" is a self-logging protocol for Al agents based on the Fractal Referential Architecture (FRA).
par: AdmailFRA
Publié: (2025)
par: AdmailFRA
Publié: (2025)
ToolsyBio: A retrieval-augmented generation system for navigating the bioinformatics software landscape
par: Truong, Van Q., et autres
Publié: (2025)
par: Truong, Van Q., et autres
Publié: (2025)
NoSQL Database Modeling and Management: A Systematic Literature Review
par: Raúl Aguilar Vera
Publié: (2023)
par: Raúl Aguilar Vera
Publié: (2023)
Ep. 598: Audio Engineering as Prompt Engineering: Better Sound, Better AI
par: Rosehill, Daniel, et autres
Publié: (2026)
par: Rosehill, Daniel, et autres
Publié: (2026)
Metacognition Benchmark: Evaluating Confidence Calibration and Sycophancy Resistance in Clinical AI
par: Khan, Nabeera
Publié: (2026)
par: Khan, Nabeera
Publié: (2026)
Emergent Self-Monitoring in Large Language Models: Probing Internal State Awareness and Output Ownership
par: Nadeem, Aurther
Publié: (2025)
par: Nadeem, Aurther
Publié: (2025)
Documents similaires
-
holotype: forensic-grade archival of LLM-driven agent CLI sessions for scientific reproducibility
par: Ferland, Brenden
Publié: (2026) -
holotype: forensic-grade archival of LLM-driven agent CLI sessions for scientific reproducibility
par: Ferland, Brenden
Publié: (2026) -
Replication Package: From LLMs to Agentic Systems, A Systematic Mapping Study of AI Techniques in Software Requirements Engineering
par: Boussaroual, Khalid, et autres
Publié: (2026) -
LLM Token Estimation Benchmarks: Tokenizer Efficiency and Cost Analysis Across 17 Large Language Models
par: Khare, Mohit
Publié: (2026) -
Self-Evolving Multi-Agent Swarms: Autonomous Quality Audit, Repair, and Verification Loops for Production AI Agent Systems
par: The LocalKin Team
Publié: (2026)