Rethinking Data Protection in the (Generative) Artificial Intelligence Era
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Yiming, Shao, Shuo, He, Yu, Guo, Junfeng, Zhang, Tianwei, Qin, Zhan, Chen, Pin-Yu, Backes, Michael, Torr, Philip, Tao, Dacheng, Ren, Kui |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SoK: Large Language Model Copyright Auditing via Fingerprinting
di: Shao, Shuo, et al.
Pubblicazione: (2025)
di: Shao, Shuo, et al.
Pubblicazione: (2025)
REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
di: Chen, Yukun, et al.
Pubblicazione: (2025)
di: Chen, Yukun, et al.
Pubblicazione: (2025)
External Data Extraction Attacks against Retrieval-Augmented Large Language Models
di: He, Yu, et al.
Pubblicazione: (2025)
di: He, Yu, et al.
Pubblicazione: (2025)
Towards Label-Only Membership Inference Attack against Pre-trained Large Language Models
di: He, Yu, et al.
Pubblicazione: (2025)
di: He, Yu, et al.
Pubblicazione: (2025)
Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack
di: Chen, Yukun, et al.
Pubblicazione: (2025)
di: Chen, Yukun, et al.
Pubblicazione: (2025)
Reading Between the Lines: Towards Reliable Black-box LLM Fingerprinting via Zeroth-order Gradient Estimation
di: Shao, Shuo, et al.
Pubblicazione: (2025)
di: Shao, Shuo, et al.
Pubblicazione: (2025)
Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution
di: Shao, Shuo, et al.
Pubblicazione: (2024)
di: Shao, Shuo, et al.
Pubblicazione: (2024)
DATABench: Evaluating Dataset Auditing in Deep Learning from an Adversarial Perspective
di: Shao, Shuo, et al.
Pubblicazione: (2025)
di: Shao, Shuo, et al.
Pubblicazione: (2025)
SoK: On the Role and Future of AIGC Watermarking in the Era of Gen-AI
di: Ren, Kui, et al.
Pubblicazione: (2024)
di: Ren, Kui, et al.
Pubblicazione: (2024)
ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
Red Teaming GPT-4V: Are GPT-4V Safe Against Uni/Multi-Modal Jailbreak Attacks?
di: Chen, Shuo, et al.
Pubblicazione: (2024)
di: Chen, Shuo, et al.
Pubblicazione: (2024)
Data Protection and Corporate Reputation Management in the Digital Era
di: Wojak, Gabriela, et al.
Pubblicazione: (2025)
di: Wojak, Gabriela, et al.
Pubblicazione: (2025)
AttriGuard: Defeating Indirect Prompt Injection in LLM Agents via Causal Attribution of Tool Invocations
di: He, Yu, et al.
Pubblicazione: (2026)
di: He, Yu, et al.
Pubblicazione: (2026)
Bag of Tricks for Subverting Reasoning-based Safety Guardrails
di: Chen, Shuo, et al.
Pubblicazione: (2025)
di: Chen, Shuo, et al.
Pubblicazione: (2025)
CoTGuard: Using Chain-of-Thought Triggering for Copyright Protection in Multi-Agent LLM Systems
di: Wen, Yan, et al.
Pubblicazione: (2025)
di: Wen, Yan, et al.
Pubblicazione: (2025)
Deep Research Brings Deeper Harm
di: Chen, Shuo, et al.
Pubblicazione: (2025)
di: Chen, Shuo, et al.
Pubblicazione: (2025)
FIT-Print: Towards False-claim-resistant Model Ownership Verification via Targeted Fingerprint
di: Shao, Shuo, et al.
Pubblicazione: (2025)
di: Shao, Shuo, et al.
Pubblicazione: (2025)
Eguard: Defending LLM Embeddings Against Inversion Attacks via Text Mutual Information Optimization
di: Liu, Tiantian, et al.
Pubblicazione: (2024)
di: Liu, Tiantian, et al.
Pubblicazione: (2024)
MAJIC: Markovian Adaptive Jailbreaking via Iterative Composition of Diverse Innovative Strategies
di: Qi, Weiwei, et al.
Pubblicazione: (2025)
di: Qi, Weiwei, et al.
Pubblicazione: (2025)
Towards Sample-specific Backdoor Attack with Clean Labels via Attribute Trigger
di: Zhu, Mingyan, et al.
Pubblicazione: (2023)
di: Zhu, Mingyan, et al.
Pubblicazione: (2023)
WMCopier: Forging Invisible Image Watermarks on Arbitrary Images
di: Dong, Ziping, et al.
Pubblicazione: (2025)
di: Dong, Ziping, et al.
Pubblicazione: (2025)
Rethinking Backdoor Detection Evaluation for Language Models
di: Yan, Jun, et al.
Pubblicazione: (2024)
di: Yan, Jun, et al.
Pubblicazione: (2024)
MIRAGE: Misleading Retrieval-Augmented Generation via Black-box and Query-agnostic Poisoning Attacks
di: Chen, Tailun, et al.
Pubblicazione: (2025)
di: Chen, Tailun, et al.
Pubblicazione: (2025)
Rethinking Robust Adversarial Concept Erasure in Diffusion Models
di: Yin, Qinghong, et al.
Pubblicazione: (2025)
di: Yin, Qinghong, et al.
Pubblicazione: (2025)
SurrogatePrompt: Bypassing the Safety Filter of Text-to-Image Models via Substitution
di: Ba, Zhongjie, et al.
Pubblicazione: (2023)
di: Ba, Zhongjie, et al.
Pubblicazione: (2023)
ICLGuard: Controlling In-Context Learning Behavior for Applicability Authorization
di: Si, Wai Man, et al.
Pubblicazione: (2024)
di: Si, Wai Man, et al.
Pubblicazione: (2024)
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations
di: Ye, Mang, et al.
Pubblicazione: (2025)
di: Ye, Mang, et al.
Pubblicazione: (2025)
AISCliteracy: Assessing Artificial Intelligence and Cybersecurity Literacy Levels and Learning Needs of Students
di: Chapagain, Devendra, et al.
Pubblicazione: (2025)
di: Chapagain, Devendra, et al.
Pubblicazione: (2025)
Proactive defense against LLM Jailbreak
di: Zhao, Weiliang, et al.
Pubblicazione: (2025)
di: Zhao, Weiliang, et al.
Pubblicazione: (2025)
CoreMark: Toward Robust and Universal Text Watermarking Technique
di: Meng, Jiale, et al.
Pubblicazione: (2025)
di: Meng, Jiale, et al.
Pubblicazione: (2025)
Shadow in the Cache: Unveiling and Mitigating Privacy Risks of KV-cache in LLM Inference
di: Luo, Zhifan, et al.
Pubblicazione: (2025)
di: Luo, Zhifan, et al.
Pubblicazione: (2025)
Membership Inference Attacks Against In-Context Learning
di: Wen, Rui, et al.
Pubblicazione: (2024)
di: Wen, Rui, et al.
Pubblicazione: (2024)
You Still See Me: How Data Protection Supports the Architecture of AI Surveillance
di: Yew, Rui-Jie, et al.
Pubblicazione: (2024)
di: Yew, Rui-Jie, et al.
Pubblicazione: (2024)
PointNCBW: Towards Dataset Ownership Verification for Point Clouds via Negative Clean-label Backdoor Watermark
di: Wei, Cheng, et al.
Pubblicazione: (2024)
di: Wei, Cheng, et al.
Pubblicazione: (2024)
FDINet: Protecting against DNN Model Extraction via Feature Distortion Index
di: Yao, Hongwei, et al.
Pubblicazione: (2023)
di: Yao, Hongwei, et al.
Pubblicazione: (2023)
SafeReview: Defending LLM-based Review Systems Against Adversarial Hidden Prompts
di: Xin, Yuan, et al.
Pubblicazione: (2026)
di: Xin, Yuan, et al.
Pubblicazione: (2026)
Is Artificial Intelligence Generated Image Detection a Solved Problem?
di: Li, Ziqiang, et al.
Pubblicazione: (2025)
di: Li, Ziqiang, et al.
Pubblicazione: (2025)
Rethinking Machine Unlearning in Image Generation Models
di: Liu, Renyang, et al.
Pubblicazione: (2025)
di: Liu, Renyang, et al.
Pubblicazione: (2025)
Inducing High Energy-Latency of Large Vision-Language Models with Verbose Images
di: Gao, Kuofeng, et al.
Pubblicazione: (2024)
di: Gao, Kuofeng, et al.
Pubblicazione: (2024)
Rethinking LLM Watermark Detection in Black-Box Settings: A Non-Intrusive Third-Party Framework
di: Wang, Zhuoshang, et al.
Pubblicazione: (2026)
di: Wang, Zhuoshang, et al.
Pubblicazione: (2026)
Documenti analoghi
-
SoK: Large Language Model Copyright Auditing via Fingerprinting
di: Shao, Shuo, et al.
Pubblicazione: (2025) -
REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
di: Chen, Yukun, et al.
Pubblicazione: (2025) -
External Data Extraction Attacks against Retrieval-Augmented Large Language Models
di: He, Yu, et al.
Pubblicazione: (2025) -
Towards Label-Only Membership Inference Attack against Pre-trained Large Language Models
di: He, Yu, et al.
Pubblicazione: (2025) -
Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack
di: Chen, Yukun, et al.
Pubblicazione: (2025)