Emergent Formal Verification: How an Autonomous AI Ecosystem Independently Discovered SMT-Based Safety Across Six Domains
Fuente:
arXiv
Guardado en:
| Autor principal: | Untila, Octavian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AIRA: AI-Induced Risk Audit: A Structured Inspection Framework for AI-Generated Code
por: Parris, William M.
Publicado: (2026)
por: Parris, William M.
Publicado: (2026)
Provable Fairness Repair for Deep Neural Networks
por: Ma, Jianan, et al.
Publicado: (2026)
por: Ma, Jianan, et al.
Publicado: (2026)
MFH: A Multi-faceted Heuristic Algorithm Selection Approach for Software Verification
por: Su, Jie, et al.
Publicado: (2025)
por: Su, Jie, et al.
Publicado: (2025)
A Practical Approach to Formal Methods: An Eclipse Integrated Development Environment (IDE) for Security Protocols
por: Garcia, Rémi, et al.
Publicado: (2024)
por: Garcia, Rémi, et al.
Publicado: (2024)
DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures
por: Jahan, Sigma, et al.
Publicado: (2026)
por: Jahan, Sigma, et al.
Publicado: (2026)
When Gradients Collide: Failure Modes of Multi-Objective Prompt Optimization for LLM Judges
por: Darshan, Parth, et al.
Publicado: (2026)
por: Darshan, Parth, et al.
Publicado: (2026)
Understanding and Detecting Flaky Builds in GitHub Actions
por: Ge, Wenhao, et al.
Publicado: (2026)
por: Ge, Wenhao, et al.
Publicado: (2026)
TerraFormer: Automated Infrastructure-as-Code with LLMs Fine-Tuned via Policy-Guided Verifier Feedback
por: Jana, Prithwish, et al.
Publicado: (2026)
por: Jana, Prithwish, et al.
Publicado: (2026)
RepoAudit: An Autonomous LLM-Agent for Repository-Level Code Auditing
por: Guo, Jinyao, et al.
Publicado: (2025)
por: Guo, Jinyao, et al.
Publicado: (2025)
AI-assisted JSON Schema Creation and Mapping
por: Neubauer, Felix, et al.
Publicado: (2025)
por: Neubauer, Felix, et al.
Publicado: (2025)
Automated Bug Triaging using Instruction-Tuned Large Language Models
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
LLMCup: Ranking-Enhanced Comment Updating with LLMs
por: Ge, Hua, et al.
Publicado: (2025)
por: Ge, Hua, et al.
Publicado: (2025)
A Self-Improving Architecture for Dynamic Safety in Large Language Models
por: Slater, Tyler
Publicado: (2025)
por: Slater, Tyler
Publicado: (2025)
BONSAI: A Mixed-Initiative Workspace for Human-AI Co-Development of Visual Analytics Applications
por: Spinner, Thilo, et al.
Publicado: (2026)
por: Spinner, Thilo, et al.
Publicado: (2026)
From Domain Understanding to Design Readiness: a playbook for GenAI-supported learning in Software Engineering
por: Wlodarski, Rafal
Publicado: (2026)
por: Wlodarski, Rafal
Publicado: (2026)
Can Graph-Based Microservice Performance Detection Be Used for Microservice Intrusion Detection?
por: Ma, Yunjian
Publicado: (2026)
por: Ma, Yunjian
Publicado: (2026)
AI-Assisted Engineering Should Track the Epistemic Status and Temporal Validity of Architectural Decisions
por: Gilda, Sankalp, et al.
Publicado: (2026)
por: Gilda, Sankalp, et al.
Publicado: (2026)
Secure coding for web applications: Frameworks, challenges, and the role of LLMs
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
Nidus: Externalized Reasoning for AI-Assisted Engineering
por: Gorinevski, Danil
Publicado: (2026)
por: Gorinevski, Danil
Publicado: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
por: Wang, Yuchen, et al.
Publicado: (2026)
por: Wang, Yuchen, et al.
Publicado: (2026)
Towards Continuous Assurance with Formal Verification and Assurance Cases
por: Abeywickrama, Dhaminda B., et al.
Publicado: (2025)
por: Abeywickrama, Dhaminda B., et al.
Publicado: (2025)
Knowledge Equivalence in Digital Twins of Intelligent Systems
por: Zhang, Nan, et al.
Publicado: (2022)
por: Zhang, Nan, et al.
Publicado: (2022)
On the Mistaken Assumption of Interchangeable Deep Reinforcement Learning Implementations
por: Hundal, Rajdeep Singh, et al.
Publicado: (2025)
por: Hundal, Rajdeep Singh, et al.
Publicado: (2025)
Towards Explainable Test Case Prioritisation with Learning-to-Rank Models
por: Ramírez, Aurora, et al.
Publicado: (2024)
por: Ramírez, Aurora, et al.
Publicado: (2024)
Provable Repair of Deep Neural Network Defects by Preimage Synthesis and Property Refinement
por: Ma, Jianan, et al.
Publicado: (2025)
por: Ma, Jianan, et al.
Publicado: (2025)
MOCHA: Multi-Objective Chebyshev Annealing for Agent Skill Optimization
por: Tanjim, Md Mehrab, et al.
Publicado: (2026)
por: Tanjim, Md Mehrab, et al.
Publicado: (2026)
Software Defined Vehicle Code Generation: A Few-Shot Prompting Approach
por: Nguyen, Quang-Dung, et al.
Publicado: (2025)
por: Nguyen, Quang-Dung, et al.
Publicado: (2025)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
por: Palacios, Diego Cabezas
Publicado: (2026)
por: Palacios, Diego Cabezas
Publicado: (2026)
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
por: Hu, Yuelin, et al.
Publicado: (2026)
por: Hu, Yuelin, et al.
Publicado: (2026)
Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture
por: Iscan, Mehmet
Publicado: (2026)
por: Iscan, Mehmet
Publicado: (2026)
Uncovering Bugs in Formal Explainers: A Case Study with PyXAI
por: Huang, Xuanxiang, et al.
Publicado: (2025)
por: Huang, Xuanxiang, et al.
Publicado: (2025)
A Domain-Independent Agent Architecture for Adaptive Operation in Evolving Open Worlds
por: Mohan, Shiwali, et al.
Publicado: (2023)
por: Mohan, Shiwali, et al.
Publicado: (2023)
CodeTracer: Towards Traceable Agent States
por: Li, Han, et al.
Publicado: (2026)
por: Li, Han, et al.
Publicado: (2026)
Multi-Agent Code Verification via Information Theory
por: Rajan, Shreshth
Publicado: (2025)
por: Rajan, Shreshth
Publicado: (2025)
LLMDFA: Analyzing Dataflow in Code with Large Language Models
por: Wang, Chengpeng, et al.
Publicado: (2024)
por: Wang, Chengpeng, et al.
Publicado: (2024)
LLM Agents for Generating Microservice-based Applications: how complex is your specification?
por: Yellin, Daniel M.
Publicado: (2025)
por: Yellin, Daniel M.
Publicado: (2025)
The Specification as Quality Gate: Three Hypotheses on AI-Assisted Code Review
por: Zietsman, Christo
Publicado: (2026)
por: Zietsman, Christo
Publicado: (2026)
A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification
por: Odmark, Joshua, et al.
Publicado: (2026)
por: Odmark, Joshua, et al.
Publicado: (2026)
Deriving Coding-Specific Sub-Models from LLMs using Resource-Efficient Pruning
por: Puccioni, Laura, et al.
Publicado: (2025)
por: Puccioni, Laura, et al.
Publicado: (2025)
Validating Solidity Code Defects using Symbolic and Concrete Execution powered by Large Language Models
por: Susan, Ştefan-Claudiu, et al.
Publicado: (2025)
por: Susan, Ştefan-Claudiu, et al.
Publicado: (2025)
Ejemplares similares
-
AIRA: AI-Induced Risk Audit: A Structured Inspection Framework for AI-Generated Code
por: Parris, William M.
Publicado: (2026) -
Provable Fairness Repair for Deep Neural Networks
por: Ma, Jianan, et al.
Publicado: (2026) -
MFH: A Multi-faceted Heuristic Algorithm Selection Approach for Software Verification
por: Su, Jie, et al.
Publicado: (2025) -
A Practical Approach to Formal Methods: An Eclipse Integrated Development Environment (IDE) for Security Protocols
por: Garcia, Rémi, et al.
Publicado: (2024) -
DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures
por: Jahan, Sigma, et al.
Publicado: (2026)