Fuzzing the brain: Automated stress testing for the safety of ML-driven neurostimulation
Fuente:
arXiv
Salvato in:
| Autori principali: | Downing, Mara, Peng, Matthew, Granley, Jacob, Beyeler, Michael, Bultan, Tevfik |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LLMs taking shortcuts in test generation: A study with SAP HANA and LevelDB
di: Bekmyradov, Vekil, et al.
Pubblicazione: (2026)
di: Bekmyradov, Vekil, et al.
Pubblicazione: (2026)
AgentAssay: Token-Efficient Regression Testing for Non-Deterministic AI Agent Workflows
di: Bhardwaj, Varun Pratap
Pubblicazione: (2026)
di: Bhardwaj, Varun Pratap
Pubblicazione: (2026)
Private GPTs for LLM-driven testing in software development and machine learning
di: Jagielski, Jakub, et al.
Pubblicazione: (2025)
di: Jagielski, Jakub, et al.
Pubblicazione: (2025)
Orion: Fuzzing Workflow Automation
di: Bazalii, Max, et al.
Pubblicazione: (2025)
di: Bazalii, Max, et al.
Pubblicazione: (2025)
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
di: Rehan, Tzafrir
Pubblicazione: (2026)
di: Rehan, Tzafrir
Pubblicazione: (2026)
SpecOps: A Fully Automated AI Agent Testing Framework in Real-World GUI Environments
di: Ahmed, Syed Yusuf, et al.
Pubblicazione: (2026)
di: Ahmed, Syed Yusuf, et al.
Pubblicazione: (2026)
Runtime Execution Traces Guided Automated Program Repair with Multi-Agent Debate
di: Wu, Jiaqing, et al.
Pubblicazione: (2026)
di: Wu, Jiaqing, et al.
Pubblicazione: (2026)
From Machine Learning Documentation to Requirements: Bridging Processes with Requirements Languages
di: Peng, Yi, et al.
Pubblicazione: (2025)
di: Peng, Yi, et al.
Pubblicazione: (2025)
RelRepair: Enhancing Automated Program Repair by Retrieving Relevant Code
di: Liu, Shunyu, et al.
Pubblicazione: (2025)
di: Liu, Shunyu, et al.
Pubblicazione: (2025)
xML-workFlow: an end-to-end explainable scikit-learn workflow for rapid biomedical experimentation
di: Tran, Khoa A., et al.
Pubblicazione: (2025)
di: Tran, Khoa A., et al.
Pubblicazione: (2025)
Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation
di: Karpurapu, Shanthi, et al.
Pubblicazione: (2024)
di: Karpurapu, Shanthi, et al.
Pubblicazione: (2024)
Assessing, Exploiting, and Mitigating Syntactic Robustness Failures in LLM-Based Code Generation
di: Sarker, Laboni, et al.
Pubblicazione: (2024)
di: Sarker, Laboni, et al.
Pubblicazione: (2024)
AutoBridge: Automating Smart Device Integration with Centralized Platform
di: Liu, Siyuan, et al.
Pubblicazione: (2025)
di: Liu, Siyuan, et al.
Pubblicazione: (2025)
GALA: Multimodal Graph Alignment for Bug Localization in Automated Program Repair
di: Liu, Zhuoyao, et al.
Pubblicazione: (2026)
di: Liu, Zhuoyao, et al.
Pubblicazione: (2026)
Automated structural testing of LLM-based agents: methods, framework, and case studies
di: Kohl, Jens, et al.
Pubblicazione: (2026)
di: Kohl, Jens, et al.
Pubblicazione: (2026)
Natural Language Summarization Enables Multi-Repository Bug Localization by LLMs in Microservice Architectures
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
Reasoning Provenance for Autonomous AI Agents: Structured Behavioral Analytics Beyond State Checkpoints and Execution Traces
di: Vispute, Neelmani, et al.
Pubblicazione: (2026)
di: Vispute, Neelmani, et al.
Pubblicazione: (2026)
CodeEvolve: LLM-Driven Evolutionary Optimization with Runtime-Enriched Target Selection for Multi-Language Code Enhancement
di: Borra, Ajay Krishna, et al.
Pubblicazione: (2026)
di: Borra, Ajay Krishna, et al.
Pubblicazione: (2026)
Review Beats Planning: Dual-Model Interaction Patterns for Code Synthesis
di: Miller, Jan
Pubblicazione: (2026)
di: Miller, Jan
Pubblicazione: (2026)
PICKLES: a Natural Language Framework for Requirement Specification and Model-Based Testing
di: Rodríguez, María Belén, et al.
Pubblicazione: (2026)
di: Rodríguez, María Belén, et al.
Pubblicazione: (2026)
FREYR: A Framework for Recognizing and Executing Your Requests
di: Gallotta, Roberto, et al.
Pubblicazione: (2025)
di: Gallotta, Roberto, et al.
Pubblicazione: (2025)
PathFuzzing: Worst Case Analysis by Fuzzing Symbolic-Execution Paths
di: Chen, Zimu, et al.
Pubblicazione: (2025)
di: Chen, Zimu, et al.
Pubblicazione: (2025)
RepoLaunch: Automating Build&Test Pipeline of Code Repositories on ANY Language and ANY Platform
di: Li, Kenan, et al.
Pubblicazione: (2026)
di: Li, Kenan, et al.
Pubblicazione: (2026)
Test-driven Software Experimentation with LASSO: an LLM Prompt Benchmarking Example
di: Kessel, Marcus
Pubblicazione: (2024)
di: Kessel, Marcus
Pubblicazione: (2024)
Enhancing Differential Testing With LLMs For Testing Deep Learning Libraries
di: Li, Meiziniu, et al.
Pubblicazione: (2024)
di: Li, Meiziniu, et al.
Pubblicazione: (2024)
Demystifying the Silence of Correctness Bugs in PyTorch Compiler
di: Li, Meiziniu, et al.
Pubblicazione: (2026)
di: Li, Meiziniu, et al.
Pubblicazione: (2026)
COMET: Coverage-guided Model Generation For Deep Learning Library Testing
di: Li, Meiziniu, et al.
Pubblicazione: (2022)
di: Li, Meiziniu, et al.
Pubblicazione: (2022)
Understanding and Detecting Flaky Builds in GitHub Actions
di: Ge, Wenhao, et al.
Pubblicazione: (2026)
di: Ge, Wenhao, et al.
Pubblicazione: (2026)
On Interaction Effects in Greybox Fuzzing
di: Kitsios, Konstantinos, et al.
Pubblicazione: (2025)
di: Kitsios, Konstantinos, et al.
Pubblicazione: (2025)
MeDeT: Medical Device Digital Twins Creation with Few-shot Meta-learning
di: Sartaj, Hassan, et al.
Pubblicazione: (2024)
di: Sartaj, Hassan, et al.
Pubblicazione: (2024)
Social, Legal, Ethical, Empathetic and Cultural Norm Operationalisation for AI Agents
di: Calinescu, Radu, et al.
Pubblicazione: (2026)
di: Calinescu, Radu, et al.
Pubblicazione: (2026)
Biomedical systems biology workflow orchestration and execution with PoSyMed
di: Süwer, Simon, et al.
Pubblicazione: (2026)
di: Süwer, Simon, et al.
Pubblicazione: (2026)
Refining Fuzzed Crashing Inputs for Better Fault Diagnosis
di: Kim, Kieun, et al.
Pubblicazione: (2025)
di: Kim, Kieun, et al.
Pubblicazione: (2025)
The Right Prompts for the Job: Repair Code-Review Defects with Large Language Model
di: Zhao, Zelin, et al.
Pubblicazione: (2023)
di: Zhao, Zelin, et al.
Pubblicazione: (2023)
Towards Observation Lakehouses: Living, Interactive Archives of Software Behavior
di: Kessel, Marcus
Pubblicazione: (2025)
di: Kessel, Marcus
Pubblicazione: (2025)
CIFE: Code Instruction-Following Evaluation
di: Gunnu, Sravani, et al.
Pubblicazione: (2025)
di: Gunnu, Sravani, et al.
Pubblicazione: (2025)
Addressing Data Leakage in HumanEval Using Combinatorial Test Design
di: Bradbury, Jeremy S., et al.
Pubblicazione: (2024)
di: Bradbury, Jeremy S., et al.
Pubblicazione: (2024)
Experience with GitHub Copilot for Developer Productivity at Zoominfo
di: Bakal, Gal, et al.
Pubblicazione: (2025)
di: Bakal, Gal, et al.
Pubblicazione: (2025)
Abstain and Validate: A Dual-LLM Policy for Reducing Noise in Agentic Program Repair
di: Cambronero, José, et al.
Pubblicazione: (2025)
di: Cambronero, José, et al.
Pubblicazione: (2025)
Inference-Time Intervention in Large Language Models for Reliable Requirement Verification
di: Darm, Paul, et al.
Pubblicazione: (2025)
di: Darm, Paul, et al.
Pubblicazione: (2025)
Documenti analoghi
-
LLMs taking shortcuts in test generation: A study with SAP HANA and LevelDB
di: Bekmyradov, Vekil, et al.
Pubblicazione: (2026) -
AgentAssay: Token-Efficient Regression Testing for Non-Deterministic AI Agent Workflows
di: Bhardwaj, Varun Pratap
Pubblicazione: (2026) -
Private GPTs for LLM-driven testing in software development and machine learning
di: Jagielski, Jakub, et al.
Pubblicazione: (2025) -
Orion: Fuzzing Workflow Automation
di: Bazalii, Max, et al.
Pubblicazione: (2025) -
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
di: Rehan, Tzafrir
Pubblicazione: (2026)