SAFEdit: Does Multi-Agent Decomposition Resolve the Reliability Challenges of Instructed Code Editing?
Fuente:
arXiv
Guardado en:
| Autores principales: | Tarshish, Noam, Selouk, Nofar, Hodisan, Daniel, Gafniel, Bar Ezra, Elovici, Yuval, Shabtai, Asaf, Nachmani, Eliya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Provably Protecting Fine-Tuned LLMs from Training Data Extraction while Preserving Utility
por: Segal, Tom, et al.
Publicado: (2026)
por: Segal, Tom, et al.
Publicado: (2026)
DOMBA: Double Model Balancing for Access-Controlled Language Models via Minimum-Bounded Aggregation
por: Segal, Tom, et al.
Publicado: (2024)
por: Segal, Tom, et al.
Publicado: (2024)
Addressing Key Challenges of Adversarial Attacks and Defenses in the Tabular Domain: A Methodological Framework for Coherence and Consistency
por: Itzhakev, Yael, et al.
Publicado: (2024)
por: Itzhakev, Yael, et al.
Publicado: (2024)
SoK: Cybersecurity Assessment of Humanoid Ecosystem
por: Surve, Priyanka Prakash, et al.
Publicado: (2025)
por: Surve, Priyanka Prakash, et al.
Publicado: (2025)
From Tool Orchestration to Code Execution: A Study of MCP Design Choices
por: Felendler, Yuval, et al.
Publicado: (2026)
por: Felendler, Yuval, et al.
Publicado: (2026)
MIA-EPT: Membership Inference Attack via Error Prediction for Tabular Data
por: German, Eyal, et al.
Publicado: (2025)
por: German, Eyal, et al.
Publicado: (2025)
QuantAttack: Exploiting Dynamic Quantization to Attack Vision Transformers
por: Baras, Amit, et al.
Publicado: (2023)
por: Baras, Amit, et al.
Publicado: (2023)
RAPID: Robust APT Detection and Investigation Using Context-Aware Deep Learning
por: Amaru, Yonatan, et al.
Publicado: (2024)
por: Amaru, Yonatan, et al.
Publicado: (2024)
Real-World Adversarial Attacks on RF-Based Drone Detectors
por: Gazit, Omer, et al.
Publicado: (2025)
por: Gazit, Omer, et al.
Publicado: (2025)
Tag&Tab: Pretraining Data Detection in Large Language Models Using Keyword-Based Membership Inference Attack
por: Antebi, Sagiv, et al.
Publicado: (2025)
por: Antebi, Sagiv, et al.
Publicado: (2025)
CodeCloak: A Method for Evaluating and Mitigating Code Leakage by LLM Code Assistants
por: Noah, Amit Finkman, et al.
Publicado: (2024)
por: Noah, Amit Finkman, et al.
Publicado: (2024)
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
por: Ebrahimi, Amir M., et al.
Publicado: (2026)
por: Ebrahimi, Amir M., et al.
Publicado: (2026)
GPT in Sheep's Clothing: The Risk of Customized GPTs
por: Antebi, Sagiv, et al.
Publicado: (2024)
por: Antebi, Sagiv, et al.
Publicado: (2024)
LLMCloudHunter: Harnessing LLMs for Automated Extraction of Detection Rules from Cloud-Based CTI
por: Schwartz, Yuval, et al.
Publicado: (2024)
por: Schwartz, Yuval, et al.
Publicado: (2024)
InstructCoder: Instruction Tuning Large Language Models for Code Editing
por: Li, Kaixin, et al.
Publicado: (2023)
por: Li, Kaixin, et al.
Publicado: (2023)
Score Based Error Correcting Code Decoder
por: Helvits, Alon, et al.
Publicado: (2026)
por: Helvits, Alon, et al.
Publicado: (2026)
Two-Dimensional Quantization for Geometry-Aware Audio Coding
por: Shuster, Tal, et al.
Publicado: (2025)
por: Shuster, Tal, et al.
Publicado: (2025)
RuleGenie: SIEM Detection Rule Set Optimization
por: Shukla, Akansha, et al.
Publicado: (2025)
por: Shukla, Akansha, et al.
Publicado: (2025)
AgentGuardian: Learning Access Control Policies to Govern AI Agent Behavior
por: Abaev, Nadya, et al.
Publicado: (2026)
por: Abaev, Nadya, et al.
Publicado: (2026)
A Privacy Enhancing Technique to Evade Detection by Street Video Cameras Without Using Adversarial Accessories
por: Shams, Jacob, et al.
Publicado: (2025)
por: Shams, Jacob, et al.
Publicado: (2025)
Rogue Cell: Adversarial Attack and Defense in Untrusted O-RAN Setup Exploiting the Traffic Steering xApp
por: Aizikovich, Eran, et al.
Publicado: (2025)
por: Aizikovich, Eran, et al.
Publicado: (2025)
LexiMark: Robust Watermarking via Lexical Substitutions to Enhance Membership Verification of an LLM's Textual Training Data
por: German, Eyal, et al.
Publicado: (2025)
por: German, Eyal, et al.
Publicado: (2025)
GenKubeSec: LLM-Based Kubernetes Misconfiguration Detection, Localization, Reasoning, and Remediation
por: Malul, Ehud, et al.
Publicado: (2024)
por: Malul, Ehud, et al.
Publicado: (2024)
DeSparsify: Adversarial Attack Against Token Sparsification Mechanisms in Vision Transformers
por: Yehezkel, Oryan, et al.
Publicado: (2024)
por: Yehezkel, Oryan, et al.
Publicado: (2024)
ImpReSS: Implicit Recommender System for Support Conversations
por: Haller, Omri, et al.
Publicado: (2025)
por: Haller, Omri, et al.
Publicado: (2025)
Tab-MIA: A Benchmark Dataset for Membership Inference Attacks on Tabular Data in LLMs
por: German, Eyal, et al.
Publicado: (2025)
por: German, Eyal, et al.
Publicado: (2025)
Detection of Compromised Functions in a Serverless Cloud Environment
por: Lavi, Danielle, et al.
Publicado: (2024)
por: Lavi, Danielle, et al.
Publicado: (2024)
Multi-Agent Code-Orchestrated Generation for Reliable Infrastructure-as-Code
por: Khan, Rana Nameer Hussain, et al.
Publicado: (2025)
por: Khan, Rana Nameer Hussain, et al.
Publicado: (2025)
Neural Minimum Weight Perfect Matching for Quantum Error Codes
por: Peled, Yotam, et al.
Publicado: (2026)
por: Peled, Yotam, et al.
Publicado: (2026)
SecMate: Multi-Agent Adaptive Cybersecurity Troubleshooting with Tri-Context Personalization
por: Meidan, Yair, et al.
Publicado: (2026)
por: Meidan, Yair, et al.
Publicado: (2026)
ILRR: Inference-Time Steering Method for Masked Diffusion Language Models
por: Avrahami, Eden, et al.
Publicado: (2026)
por: Avrahami, Eden, et al.
Publicado: (2026)
SAQ: Stabilizer-Aware Quantum Error Correction Decoder
por: Zenati, David, et al.
Publicado: (2025)
por: Zenati, David, et al.
Publicado: (2025)
Magicoder: Empowering Code Generation with OSS-Instruct
por: Wei, Yuxiang, et al.
Publicado: (2023)
por: Wei, Yuxiang, et al.
Publicado: (2023)
UEFI Memory Forensics: A Framework for UEFI Threat Analysis
por: Segal, Kalanit Suzan, et al.
Publicado: (2025)
por: Segal, Kalanit Suzan, et al.
Publicado: (2025)
Towards an End-to-End (E2E) Adversarial Learning and Application in the Physical World
por: Biton, Dudi, et al.
Publicado: (2025)
por: Biton, Dudi, et al.
Publicado: (2025)
SHIELD: APT Detection and Intelligent Explanation Using LLM
por: Gandhi, Parth Atulbhai, et al.
Publicado: (2025)
por: Gandhi, Parth Atulbhai, et al.
Publicado: (2025)
OpenCodeInstruct: A Large-scale Instruction Tuning Dataset for Code LLMs
por: Ahmad, Wasi Uddin, et al.
Publicado: (2025)
por: Ahmad, Wasi Uddin, et al.
Publicado: (2025)
CodeAgent: Enhancing Code Generation with Tool-Integrated Agent Systems for Real-World Repo-level Coding Challenges
por: Zhang, Kechi, et al.
Publicado: (2024)
por: Zhang, Kechi, et al.
Publicado: (2024)
ATAG: AI-Agent Application Threat Assessment with Attack Graphs
por: Gandhi, Parth Atulbhai, et al.
Publicado: (2025)
por: Gandhi, Parth Atulbhai, et al.
Publicado: (2025)
EDIT-Bench: Evaluating LLM Abilities to Perform Real-World Instructed Code Edits
por: Chi, Wayne, et al.
Publicado: (2025)
por: Chi, Wayne, et al.
Publicado: (2025)
Ejemplares similares
-
Provably Protecting Fine-Tuned LLMs from Training Data Extraction while Preserving Utility
por: Segal, Tom, et al.
Publicado: (2026) -
DOMBA: Double Model Balancing for Access-Controlled Language Models via Minimum-Bounded Aggregation
por: Segal, Tom, et al.
Publicado: (2024) -
Addressing Key Challenges of Adversarial Attacks and Defenses in the Tabular Domain: A Methodological Framework for Coherence and Consistency
por: Itzhakev, Yael, et al.
Publicado: (2024) -
SoK: Cybersecurity Assessment of Humanoid Ecosystem
por: Surve, Priyanka Prakash, et al.
Publicado: (2025) -
From Tool Orchestration to Code Execution: A Study of MCP Design Choices
por: Felendler, Yuval, et al.
Publicado: (2026)