Guardado en:
| Autor principal: | Vardanyan, Aram |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2511.19477 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Secure coding for web applications: Frameworks, challenges, and the role of LLMs
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
por: Othman, Refat
Publicado: (2026)
por: Othman, Refat
Publicado: (2026)
Continuous Discovery of Vulnerabilities in LLM Serving Systems with Fuzzing
por: Zhao, Yunze, et al.
Publicado: (2026)
por: Zhao, Yunze, et al.
Publicado: (2026)
Secure and Scalable Blockchain Voting: A Comparative Framework and the Role of Large Language Models
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
PROTEA: Offline Evaluation and Iterative Refinement for Multi-Agent LLM Workflows
por: Kawamura, Kazuki, et al.
Publicado: (2026)
por: Kawamura, Kazuki, et al.
Publicado: (2026)
Augment Engineering: A Methodology for Multi-Tool AI Orchestration Across Professional Domains
por: Calboreanu, Elias
Publicado: (2026)
por: Calboreanu, Elias
Publicado: (2026)
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
por: Ge, Yuxu
Publicado: (2026)
por: Ge, Yuxu
Publicado: (2026)
Using LLMs to Establish Implicit User Sentiment of Software Desirability
por: Weitl-Harms, Sherri, et al.
Publicado: (2024)
por: Weitl-Harms, Sherri, et al.
Publicado: (2024)
SPIRA: Building an Intelligent System for Respiratory Insufficiency Detection
por: Ferreira, Renato Cordeiro, et al.
Publicado: (2025)
por: Ferreira, Renato Cordeiro, et al.
Publicado: (2025)
Model-Driven Legacy System Modernization at Scale
por: Böhm, Tobias, et al.
Publicado: (2026)
por: Böhm, Tobias, et al.
Publicado: (2026)
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
por: Zolduoarrati, Elijah, et al.
Publicado: (2025)
por: Zolduoarrati, Elijah, et al.
Publicado: (2025)
Componentization: Decomposing Monolithic LLM Responses into Manipulable Semantic Units
por: Lingo, Ryan, et al.
Publicado: (2025)
por: Lingo, Ryan, et al.
Publicado: (2025)
AuditRepairBench: A Paired-Execution Trace Corpus for Evaluator-Channel Ranking Instability in Agent Repair
por: Hu, Yuelin, et al.
Publicado: (2026)
por: Hu, Yuelin, et al.
Publicado: (2026)
Automated Bug Triaging using Instruction-Tuned Large Language Models
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
GALA: Multimodal Graph Alignment for Bug Localization in Automated Program Repair
por: Liu, Zhuoyao, et al.
Publicado: (2026)
por: Liu, Zhuoyao, et al.
Publicado: (2026)
Toward Architecture-Aware Evaluation Metrics for LLM Agents
por: Souza, Débora, et al.
Publicado: (2026)
por: Souza, Débora, et al.
Publicado: (2026)
AgentEval: DAG-Structured Step-Level Evaluation for Agentic Workflows with Error Propagation Tracking
por: Guo, Dongxin, et al.
Publicado: (2026)
por: Guo, Dongxin, et al.
Publicado: (2026)
Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture
por: Iscan, Mehmet
Publicado: (2026)
por: Iscan, Mehmet
Publicado: (2026)
Making a Pipeline Production-Ready: Challenges and Lessons Learned in the Healthcare Domain
por: Lawand, Daniel Angelo Esteves, et al.
Publicado: (2025)
por: Lawand, Daniel Angelo Esteves, et al.
Publicado: (2025)
Towards Agentic Investigation of Security Alerts
por: Eilertsen, Even, et al.
Publicado: (2026)
por: Eilertsen, Even, et al.
Publicado: (2026)
AIRTBench: Measuring Autonomous AI Red Teaming Capabilities in Language Models
por: Dawson, Ads, et al.
Publicado: (2025)
por: Dawson, Ads, et al.
Publicado: (2025)
The Automation Advantage in AI Red Teaming
por: Mulla, Rob, et al.
Publicado: (2025)
por: Mulla, Rob, et al.
Publicado: (2025)
Security Considerations for Multi-agent Systems
por: Nguyen, Tam, et al.
Publicado: (2026)
por: Nguyen, Tam, et al.
Publicado: (2026)
Refusal Evaluation in Coding LLMs and Code Agents: A Systematic Review of Thirteen Malicious-Code Prompt Corpora (2023-2025)
por: Young, Richard J., et al.
Publicado: (2026)
por: Young, Richard J., et al.
Publicado: (2026)
AgentPulse: A Continuous Multi-Signal Framework for Evaluating AI Agents in Deployment
por: Gao, Yuxuan, et al.
Publicado: (2026)
por: Gao, Yuxuan, et al.
Publicado: (2026)
LLMs as Idiomatic Decompilers: Recovering High-Level Code from x86-64 Assembly for Dart
por: Abualazm, Raafat, et al.
Publicado: (2026)
por: Abualazm, Raafat, et al.
Publicado: (2026)
AI Bill of Materials and Beyond: Systematizing Security Assurance through the AI Risk Scanning (AIRS) Framework
por: Nathanson, Samuel, et al.
Publicado: (2025)
por: Nathanson, Samuel, et al.
Publicado: (2025)
AgentModernize: Preserving Business Logic in Legacy Modernization with Multi-Agent LLMs and Behavioral Specification Graphs
por: Ahmed, Sheikh Nazib, et al.
Publicado: (2026)
por: Ahmed, Sheikh Nazib, et al.
Publicado: (2026)
AgentSentry: Mitigating Indirect Prompt Injection in LLM Agents via Temporal Causal Diagnostics and Context Purification
por: Zhang, Tian, et al.
Publicado: (2026)
por: Zhang, Tian, et al.
Publicado: (2026)
Sample-Efficient Language Model for Hinglish Conversational AI
por: Singh, Sakshi, et al.
Publicado: (2025)
por: Singh, Sakshi, et al.
Publicado: (2025)
Towards a Probabilistic Framework for Analyzing and Improving LLM-Enabled Software
por: Baldonado, Juan Manuel, et al.
Publicado: (2025)
por: Baldonado, Juan Manuel, et al.
Publicado: (2025)
RMCBench: Benchmarking Large Language Models' Resistance to Malicious Code
por: Chen, Jiachi, et al.
Publicado: (2024)
por: Chen, Jiachi, et al.
Publicado: (2024)
VulScribeR: Exploring RAG-based Vulnerability Augmentation with LLMs
por: Daneshvar, Seyed Shayan, et al.
Publicado: (2024)
por: Daneshvar, Seyed Shayan, et al.
Publicado: (2024)
Exploring Large Language Models for Access Control Policy Synthesis and Summarization
por: Vatsa, Adarsh, et al.
Publicado: (2025)
por: Vatsa, Adarsh, et al.
Publicado: (2025)
DRS-OSS: Practical Diff Risk Scoring with LLMs
por: Sayedsalehi, Ali, et al.
Publicado: (2025)
por: Sayedsalehi, Ali, et al.
Publicado: (2025)
Streamlining Security Vulnerability Triage with Large Language Models
por: Torkamani, Mohammad Jalili, et al.
Publicado: (2025)
por: Torkamani, Mohammad Jalili, et al.
Publicado: (2025)
Creating benchmarkable components to measure the quality ofAI-enhanced developer tools
por: Paradis, Elise, et al.
Publicado: (2025)
por: Paradis, Elise, et al.
Publicado: (2025)
How much does AI impact development speed? An enterprise-based randomized controlled trial
por: Paradis, Elise, et al.
Publicado: (2024)
por: Paradis, Elise, et al.
Publicado: (2024)
Safeguarding Virtual Healthcare: A Novel Attacker-Centric Model for Data Security and Privacy
por: Herath, Suvineetha, et al.
Publicado: (2024)
por: Herath, Suvineetha, et al.
Publicado: (2024)
Validating Solidity Code Defects using Symbolic and Concrete Execution powered by Large Language Models
por: Susan, Ştefan-Claudiu, et al.
Publicado: (2025)
por: Susan, Ştefan-Claudiu, et al.
Publicado: (2025)
Ejemplares similares
-
Secure coding for web applications: Frameworks, challenges, and the role of LLMs
por: Kiashemshaki, Kiana, et al.
Publicado: (2025) -
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
por: Othman, Refat
Publicado: (2026) -
Continuous Discovery of Vulnerabilities in LLM Serving Systems with Fuzzing
por: Zhao, Yunze, et al.
Publicado: (2026) -
Secure and Scalable Blockchain Voting: A Comparative Framework and the Role of Large Language Models
por: Kiashemshaki, Kiana, et al.
Publicado: (2025) -
PROTEA: Offline Evaluation and Iterative Refinement for Multi-Agent LLM Workflows
por: Kawamura, Kazuki, et al.
Publicado: (2026)