GoldCoin: Grounding Large Language Models in Privacy Laws via Contextual Integrity Theory
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fan, Wei, Li, Haoran, Deng, Zheye, Wang, Weiqi, Song, Yangqiu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Privacy Checklist: Privacy Violation Detection Grounding on Contextual Integrity Theory
von: Li, Haoran, et al.
Veröffentlicht: (2024)
von: Li, Haoran, et al.
Veröffentlicht: (2024)
PrivLM-Bench: A Multi-level Privacy Evaluation Benchmark for Language Models
von: Li, Haoran, et al.
Veröffentlicht: (2023)
von: Li, Haoran, et al.
Veröffentlicht: (2023)
Privacy in Large Language Models: Attacks, Defenses and Future Directions
von: Li, Haoran, et al.
Veröffentlicht: (2023)
von: Li, Haoran, et al.
Veröffentlicht: (2023)
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2023)
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2023)
Simulate and Eliminate: Revoke Backdoors for Generative Large Language Models
von: Li, Haoran, et al.
Veröffentlicht: (2024)
von: Li, Haoran, et al.
Veröffentlicht: (2024)
Federated Domain-Specific Knowledge Transfer on Large Language Models Using Synthetic Data
von: Li, Haoran, et al.
Veröffentlicht: (2024)
von: Li, Haoran, et al.
Veröffentlicht: (2024)
CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents
von: Fu, Wenjie, et al.
Veröffentlicht: (2026)
von: Fu, Wenjie, et al.
Veröffentlicht: (2026)
BaThe: Defense against the Jailbreak Attack in Multimodal Large Language Models by Treating Harmful Instruction as Backdoor Trigger
von: Chen, Yulin, et al.
Veröffentlicht: (2024)
von: Chen, Yulin, et al.
Veröffentlicht: (2024)
Token-Level Privacy in Large Language Models
von: Harel, Re'em, et al.
Veröffentlicht: (2025)
von: Harel, Re'em, et al.
Veröffentlicht: (2025)
Integrating Differential Privacy and Contextual Integrity
von: Benthall, Sebastian, et al.
Veröffentlicht: (2024)
von: Benthall, Sebastian, et al.
Veröffentlicht: (2024)
Privacy-Preserving Instructions for Aligning Large Language Models
von: Yu, Da, et al.
Veröffentlicht: (2024)
von: Yu, Da, et al.
Veröffentlicht: (2024)
Privacy-Preserving Parameter-Efficient Fine-Tuning for Large Language Model Services
von: Li, Yansong, et al.
Veröffentlicht: (2023)
von: Li, Yansong, et al.
Veröffentlicht: (2023)
Say Something Else: Rethinking Contextual Privacy as Information Sufficiency
von: Xiao, Yunze, et al.
Veröffentlicht: (2026)
von: Xiao, Yunze, et al.
Veröffentlicht: (2026)
Privacy-Preserved Neural Graph Databases
von: Hu, Qi, et al.
Veröffentlicht: (2023)
von: Hu, Qi, et al.
Veröffentlicht: (2023)
Understanding and Mitigating Over-refusal for Large Language Models via Safety Representation
von: Zhang, Junbo, et al.
Veröffentlicht: (2025)
von: Zhang, Junbo, et al.
Veröffentlicht: (2025)
ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models
von: Zhao, Yunhan, et al.
Veröffentlicht: (2026)
von: Zhao, Yunhan, et al.
Veröffentlicht: (2026)
Resource Consumption Red-Teaming for Large Vision-Language Models
von: Gao, Haoran, et al.
Veröffentlicht: (2025)
von: Gao, Haoran, et al.
Veröffentlicht: (2025)
Contextualized Privacy Defense for LLM Agents
von: Wen, Yule, et al.
Veröffentlicht: (2026)
von: Wen, Yule, et al.
Veröffentlicht: (2026)
One Trigger Token Is Enough: A Defense Strategy for Balancing Safety and Usability in Large Language Models
von: Gu, Haoran, et al.
Veröffentlicht: (2025)
von: Gu, Haoran, et al.
Veröffentlicht: (2025)
TopicAttack: An Indirect Prompt Injection Attack via Topic Transition
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
PrivacyRestore: Privacy-Preserving Inference in Large Language Models via Privacy Removal and Restoration
von: Zeng, Ziqian, et al.
Veröffentlicht: (2024)
von: Zeng, Ziqian, et al.
Veröffentlicht: (2024)
What Does the Server See? Understanding Privacy Leakage from Large Language Models in Split Inference
von: Fan, Mingyuan, et al.
Veröffentlicht: (2026)
von: Fan, Mingyuan, et al.
Veröffentlicht: (2026)
Toward Copyright Integrity and Verifiability via Multi-Bit Watermarking for Intelligent Transportation Systems
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
Do Vision-Language Models Respect Contextual Integrity in Location Disclosure?
von: Yang, Ruixin, et al.
Veröffentlicht: (2026)
von: Yang, Ruixin, et al.
Veröffentlicht: (2026)
PAPILLON: Privacy Preservation from Internet-based and Local Language Model Ensembles
von: Siyan, Li, et al.
Veröffentlicht: (2024)
von: Siyan, Li, et al.
Veröffentlicht: (2024)
Security and Privacy Challenges of Large Language Models: A Survey
von: Das, Badhan Chandra, et al.
Veröffentlicht: (2024)
von: Das, Badhan Chandra, et al.
Veröffentlicht: (2024)
Privacy Preserving In-Context-Learning Framework for Large Language Models
von: Bhusal, Bishnu, et al.
Veröffentlicht: (2025)
von: Bhusal, Bishnu, et al.
Veröffentlicht: (2025)
The Model's Language Matters: A Comparative Privacy Analysis of LLMs
von: Mishra, Abhishek K., et al.
Veröffentlicht: (2025)
von: Mishra, Abhishek K., et al.
Veröffentlicht: (2025)
ConfGuard: A Simple and Effective Backdoor Detection for Large Language Models
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
Trojan Activation Attack: Red-Teaming Large Language Models using Activation Steering for Safety-Alignment
von: Wang, Haoran, et al.
Veröffentlicht: (2023)
von: Wang, Haoran, et al.
Veröffentlicht: (2023)
Beyond Jailbreaking: Auditing Contextual Privacy in LLM Agents
von: Das, Saswat, et al.
Veröffentlicht: (2025)
von: Das, Saswat, et al.
Veröffentlicht: (2025)
EvoDefense: Co-Evolving Black-Box Defense with Large Language Models
von: Li, Yu, et al.
Veröffentlicht: (2026)
von: Li, Yu, et al.
Veröffentlicht: (2026)
Detecting RAG Extraction Attack via Dual-Path Runtime Integrity Game
von: Xie, Yuanbo, et al.
Veröffentlicht: (2026)
von: Xie, Yuanbo, et al.
Veröffentlicht: (2026)
Backdoor-Powered Prompt Injection Attacks Nullify Defense Methods
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
Graded Modal Types for Integrity and Confidentiality
von: Marshall, Danielle, et al.
Veröffentlicht: (2023)
von: Marshall, Danielle, et al.
Veröffentlicht: (2023)
PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action
von: Shao, Yijia, et al.
Veröffentlicht: (2024)
von: Shao, Yijia, et al.
Veröffentlicht: (2024)
Sanitize Your Responses: Mitigating Privacy Leakage in Large Language Models
von: Fu, Wenjie, et al.
Veröffentlicht: (2025)
von: Fu, Wenjie, et al.
Veröffentlicht: (2025)
Continuous Embedding Attacks via Clipped Inputs in Jailbreaking Large Language Models
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
von: Xu, Zihao, et al.
Veröffentlicht: (2024)
$PD^3F$: A Pluggable and Dynamic DoS-Defense Framework Against Resource Consumption Attacks Targeting Large Language Models
von: Zhang, Yuanhe, et al.
Veröffentlicht: (2025)
von: Zhang, Yuanhe, et al.
Veröffentlicht: (2025)
Synthesizing Tight Privacy and Accuracy Bounds via Weighted Model Counting
von: Oakley, Lisa, et al.
Veröffentlicht: (2024)
von: Oakley, Lisa, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Privacy Checklist: Privacy Violation Detection Grounding on Contextual Integrity Theory
von: Li, Haoran, et al.
Veröffentlicht: (2024) -
PrivLM-Bench: A Multi-level Privacy Evaluation Benchmark for Language Models
von: Li, Haoran, et al.
Veröffentlicht: (2023) -
Privacy in Large Language Models: Attacks, Defenses and Future Directions
von: Li, Haoran, et al.
Veröffentlicht: (2023) -
Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2023) -
Simulate and Eliminate: Revoke Backdoors for Generative Large Language Models
von: Li, Haoran, et al.
Veröffentlicht: (2024)