CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fu, Wenjie, Qin, Xiaoting, Zhang, Jue, Lin, Qingwei, Wutschitz, Lukas, Sim, Robert, Rajmohan, Saravan, Zhang, Dongmei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Privacy in Action: Towards Realistic Privacy Mitigation and Evaluation for LLM-Powered Agents
von: Wang, Shouju, et al.
Veröffentlicht: (2025)
von: Wang, Shouju, et al.
Veröffentlicht: (2025)
Securing LLM Agents Need Intent-to-Execution Integrity
von: Qu, Wenjie, et al.
Veröffentlicht: (2026)
von: Qu, Wenjie, et al.
Veröffentlicht: (2026)
CIMemories: A Compositional Benchmark for Contextual Integrity of Persistent Memory in LLMs
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2025)
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2025)
The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text
von: Meeus, Matthieu, et al.
Veröffentlicht: (2025)
von: Meeus, Matthieu, et al.
Veröffentlicht: (2025)
Integrating Differential Privacy and Contextual Integrity
von: Benthall, Sebastian, et al.
Veröffentlicht: (2024)
von: Benthall, Sebastian, et al.
Veröffentlicht: (2024)
Heimdallr: Characterizing and Detecting LLM-Induced Security Risks in GitHub CI Workflows
von: Ruan, Bonan, et al.
Veröffentlicht: (2026)
von: Ruan, Bonan, et al.
Veröffentlicht: (2026)
From Reasoning to Answer: Empirical, Attention-Based and Mechanistic Insights into Distilled DeepSeek R1 Models
von: Zhang, Jue, et al.
Veröffentlicht: (2025)
von: Zhang, Jue, et al.
Veröffentlicht: (2025)
Securing AI Agents with Information-Flow Control
von: Costa, Manuel, et al.
Veröffentlicht: (2025)
von: Costa, Manuel, et al.
Veröffentlicht: (2025)
Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks
von: Tan, Rongyuan, et al.
Veröffentlicht: (2026)
von: Tan, Rongyuan, et al.
Veröffentlicht: (2026)
Contextualized Privacy Defense for LLM Agents
von: Wen, Yule, et al.
Veröffentlicht: (2026)
von: Wen, Yule, et al.
Veröffentlicht: (2026)
When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents
von: Liu, Shi, et al.
Veröffentlicht: (2026)
von: Liu, Shi, et al.
Veröffentlicht: (2026)
Closed-Form Bounds for DP-SGD against Record-level Inference
von: Cherubin, Giovanni, et al.
Veröffentlicht: (2024)
von: Cherubin, Giovanni, et al.
Veröffentlicht: (2024)
MEETING DELEGATE: Benchmarking LLMs on Attending Meetings on Our Behalf
von: Hu, Lingxiang, et al.
Veröffentlicht: (2025)
von: Hu, Lingxiang, et al.
Veröffentlicht: (2025)
Do Vision-Language Models Respect Contextual Integrity in Location Disclosure?
von: Yang, Ruixin, et al.
Veröffentlicht: (2026)
von: Yang, Ruixin, et al.
Veröffentlicht: (2026)
Privacy Checklist: Privacy Violation Detection Grounding on Contextual Integrity Theory
von: Li, Haoran, et al.
Veröffentlicht: (2024)
von: Li, Haoran, et al.
Veröffentlicht: (2024)
Ghost in the Agent: Redefining Information Flow Tracking for LLM Agents
von: Cai, Yuandao, et al.
Veröffentlicht: (2026)
von: Cai, Yuandao, et al.
Veröffentlicht: (2026)
SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
von: Hu, Qi, et al.
Veröffentlicht: (2026)
von: Hu, Qi, et al.
Veröffentlicht: (2026)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
von: Zhang, Hanrong, et al.
Veröffentlicht: (2024)
von: Zhang, Hanrong, et al.
Veröffentlicht: (2024)
Profiling for Pennies: Unveiling the Privacy Iceberg of LLM Agents
von: Chen, Jiahao, et al.
Veröffentlicht: (2026)
von: Chen, Jiahao, et al.
Veröffentlicht: (2026)
An Evidence-driven Protocol for Trustworthy CI Pipelines
von: Castillo, Fernando, et al.
Veröffentlicht: (2026)
von: Castillo, Fernando, et al.
Veröffentlicht: (2026)
SEC-bench: Automated Benchmarking of LLM Agents on Real-World Software Security Tasks
von: Lee, Hwiwon, et al.
Veröffentlicht: (2025)
von: Lee, Hwiwon, et al.
Veröffentlicht: (2025)
Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agents
von: Kim, Juhee, et al.
Veröffentlicht: (2025)
von: Kim, Juhee, et al.
Veröffentlicht: (2025)
ConVerse: Benchmarking Contextual Safety in Agent-to-Agent Conversations
von: Gomaa, Amr, et al.
Veröffentlicht: (2025)
von: Gomaa, Amr, et al.
Veröffentlicht: (2025)
RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments
von: Fu, Yuchuan, et al.
Veröffentlicht: (2025)
von: Fu, Yuchuan, et al.
Veröffentlicht: (2025)
A Biosecurity Agent for Lifecycle LLM Biosecurity Alignment
von: Meng, Meiyin, et al.
Veröffentlicht: (2025)
von: Meng, Meiyin, et al.
Veröffentlicht: (2025)
LITMUS: Benchmarking Behavioral Jailbreaks of LLM Agents in Real OS Environments
von: Zhang, Chiyu, et al.
Veröffentlicht: (2026)
von: Zhang, Chiyu, et al.
Veröffentlicht: (2026)
Benchmarking LLMs and LLM-based Agents in Practical Vulnerability Detection for Code Repositories
von: Yildiz, Alperen, et al.
Veröffentlicht: (2025)
von: Yildiz, Alperen, et al.
Veröffentlicht: (2025)
TICAL: Trusted and Integrity-protected Compilation of AppLications
von: Krahn, Robert, et al.
Veröffentlicht: (2025)
von: Krahn, Robert, et al.
Veröffentlicht: (2025)
GoldCoin: Grounding Large Language Models in Privacy Laws via Contextual Integrity Theory
von: Fan, Wei, et al.
Veröffentlicht: (2024)
von: Fan, Wei, et al.
Veröffentlicht: (2024)
Teamwork Makes TEE Work: Open and Resilient Remote Attestation on Decentralized Trust
von: Zhang, Xiaolin, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaolin, et al.
Veröffentlicht: (2024)
Black-box Optimization of LLM Outputs by Asking for Directions
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs
von: Park, Sangwoo, et al.
Veröffentlicht: (2026)
von: Park, Sangwoo, et al.
Veröffentlicht: (2026)
Fast Revocable Attribute-Based Encryption with Data Integrity for Internet of Things
von: Li, Yongjiao, et al.
Veröffentlicht: (2025)
von: Li, Yongjiao, et al.
Veröffentlicht: (2025)
SoK: Runtime Integrity
von: Ammar, Mahmoud, et al.
Veröffentlicht: (2024)
von: Ammar, Mahmoud, et al.
Veröffentlicht: (2024)
Beyond Jailbreaking: Auditing Contextual Privacy in LLM Agents
von: Das, Saswat, et al.
Veröffentlicht: (2025)
von: Das, Saswat, et al.
Veröffentlicht: (2025)
Imprompter: Tricking LLM Agents into Improper Tool Use
von: Fu, Xiaohan, et al.
Veröffentlicht: (2024)
von: Fu, Xiaohan, et al.
Veröffentlicht: (2024)
When Skills Lie: Hidden-Comment Injection in LLM Agents
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
Prompt Control-Flow Integrity: A Priority-Aware Runtime Defense Against Prompt Injection in LLM Systems
von: Alam, Md Takrim Ul, et al.
Veröffentlicht: (2026)
von: Alam, Md Takrim Ul, et al.
Veröffentlicht: (2026)
Towards Privacy-Preserving LLM Inference via Covariant Obfuscation (Technical Report)
von: Lin, Yu, et al.
Veröffentlicht: (2026)
von: Lin, Yu, et al.
Veröffentlicht: (2026)
Securing UAV Communication: Authentication and Integrity
von: Ouadah, Meriem, et al.
Veröffentlicht: (2024)
von: Ouadah, Meriem, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Privacy in Action: Towards Realistic Privacy Mitigation and Evaluation for LLM-Powered Agents
von: Wang, Shouju, et al.
Veröffentlicht: (2025) -
Securing LLM Agents Need Intent-to-Execution Integrity
von: Qu, Wenjie, et al.
Veröffentlicht: (2026) -
CIMemories: A Compositional Benchmark for Contextual Integrity of Persistent Memory in LLMs
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2025) -
The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text
von: Meeus, Matthieu, et al.
Veröffentlicht: (2025) -
Integrating Differential Privacy and Contextual Integrity
von: Benthall, Sebastian, et al.
Veröffentlicht: (2024)