The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tihanyi, Norbert, Bisztray, Tamas, Jain, Ridhi, Ferrag, Mohamed Amine, Cordeiro, Lucas C., Mavroeidis, Vasileios |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How secure is AI-generated Code: A Large-Scale Comparison of Large Language Models
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2024)
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2024)
CyberMetric: A Benchmark Dataset based on Retrieval-Augmented Generation for Evaluating LLMs in Cybersecurity Knowledge
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2024)
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2024)
Securing Tomorrow's Smart Cities: Investigating Software Security in Internet of Vehicles and Deep Learning Technologies
von: Jain, Ridhi, et al.
Veröffentlicht: (2024)
von: Jain, Ridhi, et al.
Veröffentlicht: (2024)
A New Era in Software Security: Towards Self-Healing Software via Large Language Models and Formal Verification
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2023)
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2023)
Vulnerability Detection: From Formal Verification to Large Language Models and Hybrid Approaches: A Comprehensive Overview
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2025)
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2025)
CASTLE: Benchmarking Dataset for Static Code Analyzers and LLMs towards CWE Detection
von: Dubniczky, Richard A., et al.
Veröffentlicht: (2025)
von: Dubniczky, Richard A., et al.
Veröffentlicht: (2025)
The Hidden DNA of LLM-Generated JavaScript: Structural Patterns Enable High-Accuracy Authorship Attribution
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2025)
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2025)
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2024)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2024)
I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution
von: Bisztray, Tamas, et al.
Veröffentlicht: (2025)
von: Bisztray, Tamas, et al.
Veröffentlicht: (2025)
From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
Dynamic Intelligence Assessment: Benchmarking LLMs on the Road to AGI with a Focus on Model Confidence
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2024)
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2024)
SecureFalcon: Are We There Yet in Automated Software Vulnerability Detection with LLMs?
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2023)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2023)
Reasoning Beyond Limits: Advances and Open Problems for LLMs
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
The Phish, The Spam, and The Valid: Generating Feature-Rich Emails for Benchmarking LLMs
von: Toth, Rebeka, et al.
Veröffentlicht: (2025)
von: Toth, Rebeka, et al.
Veröffentlicht: (2025)
From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
You Have Been LaTeXpOsEd: A Systematic Analysis of Information Leakage in Preprint Archives Using Large Language Models
von: Dubniczky, Richard A., et al.
Veröffentlicht: (2025)
von: Dubniczky, Richard A., et al.
Veröffentlicht: (2025)
AgentDrive: An Open Benchmark Dataset for Agentic AI Reasoning with LLM-Generated Scenarios in Autonomous Systems
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2026)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2026)
UAVBench: An Open Benchmark Dataset for Autonomous and Agentic AI UAV Systems via LLM-Generated Flight Scenarios
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2025)
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response
von: Cherif, Bilel, et al.
Veröffentlicht: (2025)
von: Cherif, Bilel, et al.
Veröffentlicht: (2025)
Revolutionizing Cyber Threat Detection with Large Language Models: A privacy-preserving BERT-based Lightweight Model for IoT/IIoT Devices
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2023)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2023)
Validation of Modern JSON Schema: Formalization and Complexity
von: Attouche, Lyes, et al.
Veröffentlicht: (2023)
von: Attouche, Lyes, et al.
Veröffentlicht: (2023)
Edge Learning for 6G-enabled Internet of Things: A Comprehensive Survey of Vulnerabilities, Datasets, and Defenses
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2023)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2023)
DataLens: Enhancing Dataset Discovery via Network Topologies
von: Ollagnier, Anaïs, et al.
Veröffentlicht: (2025)
von: Ollagnier, Anaïs, et al.
Veröffentlicht: (2025)
Semantic Reverse Engineering Legacy Software Applications with ChatGPT, Gemini AI, and Claude AI
von: Mancas, Christian, et al.
Veröffentlicht: (2026)
von: Mancas, Christian, et al.
Veröffentlicht: (2026)
DataLens: ML-Oriented Interactive Tabular Data Quality Dashboard
von: Abdelaal, Mohamed, et al.
Veröffentlicht: (2025)
von: Abdelaal, Mohamed, et al.
Veröffentlicht: (2025)
6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2026)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2026)
ProMoAI: Process Modeling with Generative AI
von: Kourani, Humam, et al.
Veröffentlicht: (2024)
von: Kourani, Humam, et al.
Veröffentlicht: (2024)
How Small Can 6G Reason? Scaling Tiny Language Models for AI-Native Networks
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2026)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2026)
Auditing Yelp's Business Ranking and Review Recommendation Through the Lens of Fairness
von: Singhal, Mohit, et al.
Veröffentlicht: (2023)
von: Singhal, Mohit, et al.
Veröffentlicht: (2023)
Compliance Rating Scheme: A Data Provenance Framework for Generative AI Datasets
von: Bohacek, Matyas, et al.
Veröffentlicht: (2025)
von: Bohacek, Matyas, et al.
Veröffentlicht: (2025)
Automated Training of Learned Database Components with Generative AI
von: Davitkova, Angjela, et al.
Veröffentlicht: (2025)
von: Davitkova, Angjela, et al.
Veröffentlicht: (2025)
The Semantic Ladder: A Framework for Progressive Formalization of Natural Language Content for Knowledge Graphs and AI Systems
von: Vogt, Lars
Veröffentlicht: (2026)
von: Vogt, Lars
Veröffentlicht: (2026)
A survey of open-source data quality tools: shedding light on the materialization of data quality dimensions in practice
von: Papastergios, Vasileios, et al.
Veröffentlicht: (2024)
von: Papastergios, Vasileios, et al.
Veröffentlicht: (2024)
Stream DaQ: Stream-First Data Quality Monitoring
von: Papastergios, Vasileios, et al.
Veröffentlicht: (2025)
von: Papastergios, Vasileios, et al.
Veröffentlicht: (2025)
6G-Bench: An Open Benchmark for Semantic Communication and Network-Level Reasoning with Foundation Models in AI-Native 6G Networks
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2026)
von: Ferrag, Mohamed Amine, et al.
Veröffentlicht: (2026)
Snowpark: Performant, Secure, User-Friendly Data Engineering and AI/ML Next To Your Data
von: Baker, Brandon, et al.
Veröffentlicht: (2025)
von: Baker, Brandon, et al.
Veröffentlicht: (2025)
CERT: Finding Performance Issues in Database Systems Through the Lens of Cardinality Estimation
von: Ba, Jinsheng, et al.
Veröffentlicht: (2023)
von: Ba, Jinsheng, et al.
Veröffentlicht: (2023)
VARS-FL: Validation-Aligned Client Selection for Non-IID Federated Learning in IoT Systems
von: Lakas, Mohamed, et al.
Veröffentlicht: (2026)
von: Lakas, Mohamed, et al.
Veröffentlicht: (2026)
Automated Repair of AI Code with Large Language Models and Formal Verification
von: Charalambous, Yiannis, et al.
Veröffentlicht: (2024)
von: Charalambous, Yiannis, et al.
Veröffentlicht: (2024)
GenDFIR: Advancing Cyber Incident Timeline Analysis Through Retrieval Augmented Generation and Large Language Models
von: Loumachi, Fatma Yasmine, et al.
Veröffentlicht: (2024)
von: Loumachi, Fatma Yasmine, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
How secure is AI-generated Code: A Large-Scale Comparison of Large Language Models
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2024) -
CyberMetric: A Benchmark Dataset based on Retrieval-Augmented Generation for Evaluating LLMs in Cybersecurity Knowledge
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2024) -
Securing Tomorrow's Smart Cities: Investigating Software Security in Internet of Vehicles and Deep Learning Technologies
von: Jain, Ridhi, et al.
Veröffentlicht: (2024) -
A New Era in Software Security: Towards Self-Healing Software via Large Language Models and Formal Verification
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2023) -
Vulnerability Detection: From Formal Verification to Large Language Models and Hybrid Approaches: A Comprehensive Overview
von: Tihanyi, Norbert, et al.
Veröffentlicht: (2025)