PersonaMark: Personalized LLM watermarking for model protection and user attribution
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Yuehan, Lv, Peizhuo, Liu, Yinpeng, Ma, Yongqiang, Lu, Wei, Wang, Xiaofeng, Liu, Xiaozhong, Liu, Jiawei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models
di: Gong, Yuyang, et al.
Pubblicazione: (2025)
di: Gong, Yuyang, et al.
Pubblicazione: (2025)
Enhance Robustness of Language Models Against Variation Attack through Graph Integration
di: Xiong, Zi, et al.
Pubblicazione: (2024)
di: Xiong, Zi, et al.
Pubblicazione: (2024)
Black-Box Opinion Manipulation Attacks to Retrieval-Augmented Generation of Large Language Models
di: Chen, Zhuo, et al.
Pubblicazione: (2024)
di: Chen, Zhuo, et al.
Pubblicazione: (2024)
Proving membership in LLM pretraining data via data watermarks
di: Wei, Johnny Tian-Zheng, et al.
Pubblicazione: (2024)
di: Wei, Johnny Tian-Zheng, et al.
Pubblicazione: (2024)
Optimizing watermarks for large language models
di: Wouters, Bram
Pubblicazione: (2023)
di: Wouters, Bram
Pubblicazione: (2023)
WaterMax: breaking the LLM watermark detectability-robustness-quality trade-off
di: Giboulot, Eva, et al.
Pubblicazione: (2024)
di: Giboulot, Eva, et al.
Pubblicazione: (2024)
Adversarial Attacks on Reinforcement Learning-based Medical Questionnaire Systems: Input-level Perturbation Strategies and Medical Constraint Validation
di: Liu, Peizhuo
Pubblicazione: (2025)
di: Liu, Peizhuo
Pubblicazione: (2025)
Permute-and-Flip: An optimally stable and watermarkable decoder for LLMs
di: Zhao, Xuandong, et al.
Pubblicazione: (2024)
di: Zhao, Xuandong, et al.
Pubblicazione: (2024)
DiscourseFlip: An Oblique Discourse-Level Opinion Manipulation Attack against Black-box Retrieval-Augmented Generation
di: Gong, Yuyang, et al.
Pubblicazione: (2026)
di: Gong, Yuyang, et al.
Pubblicazione: (2026)
MarkLLM: An Open-Source Toolkit for LLM Watermarking
di: Pan, Leyi, et al.
Pubblicazione: (2024)
di: Pan, Leyi, et al.
Pubblicazione: (2024)
"Training robust watermarking model may hurt authentication!'' Exploring and Mitigating the Identity Leakage in Robust Watermarking
di: Zhang, Xinyu, et al.
Pubblicazione: (2026)
di: Zhang, Xinyu, et al.
Pubblicazione: (2026)
Personalized Attacks of Social Engineering in Multi-turn Conversations: LLM Agents for Simulation and Detection
di: Kumarage, Tharindu, et al.
Pubblicazione: (2025)
di: Kumarage, Tharindu, et al.
Pubblicazione: (2025)
LexiMark: Robust Watermarking via Lexical Substitutions to Enhance Membership Verification of an LLM's Textual Training Data
di: German, Eyal, et al.
Pubblicazione: (2025)
di: German, Eyal, et al.
Pubblicazione: (2025)
Deep Learning model integrity checking mechanism using watermarking technique
di: Hoque, Shahinul, et al.
Pubblicazione: (2023)
di: Hoque, Shahinul, et al.
Pubblicazione: (2023)
Defending LLM Watermarking Against Spoofing Attacks with Contrastive Representation Learning
di: An, Li, et al.
Pubblicazione: (2025)
di: An, Li, et al.
Pubblicazione: (2025)
Mark Your LLM: Detecting the Misuse of Open-Source Large Language Models via Watermarking
di: Xu, Yijie, et al.
Pubblicazione: (2025)
di: Xu, Yijie, et al.
Pubblicazione: (2025)
Watermarking LLM Agent Trajectories
di: Meng, Wenlong, et al.
Pubblicazione: (2026)
di: Meng, Wenlong, et al.
Pubblicazione: (2026)
Root Defence Strategies: Ensuring Safety of LLM at the Decoding Level
di: Zeng, Xinyi, et al.
Pubblicazione: (2024)
di: Zeng, Xinyi, et al.
Pubblicazione: (2024)
Making Theft Useless: Adulteration-Based Protection of Proprietary Knowledge Graphs in GraphRAG Systems
di: Wang, Weijie, et al.
Pubblicazione: (2026)
di: Wang, Weijie, et al.
Pubblicazione: (2026)
Hot-Swap MarkBoard: An Efficient Black-box Watermarking Approach for Large-scale Model Distribution
di: Zhang, Zhicheng, et al.
Pubblicazione: (2025)
di: Zhang, Zhicheng, et al.
Pubblicazione: (2025)
Geometry-Aware Localized Watermarking for Copyright Protection in Embedding-as-a-Service
di: Chen, Zhimin, et al.
Pubblicazione: (2026)
di: Chen, Zhimin, et al.
Pubblicazione: (2026)
FreqMark: Frequency-Based Watermark for Sentence-Level Detection of LLM-Generated Text
di: Xu, Zhenyu, et al.
Pubblicazione: (2024)
di: Xu, Zhenyu, et al.
Pubblicazione: (2024)
The Landscape of Prompt Injection Threats in LLM Agents: From Taxonomy to Analysis
di: Wang, Peiran, et al.
Pubblicazione: (2026)
di: Wang, Peiran, et al.
Pubblicazione: (2026)
Breaking the Ceiling: Exploring the Potential of Jailbreak Attacks through Expanding Strategy Space
di: Huang, Yao, et al.
Pubblicazione: (2025)
di: Huang, Yao, et al.
Pubblicazione: (2025)
LITMUS: Benchmarking Behavioral Jailbreaks of LLM Agents in Real OS Environments
di: Zhang, Chiyu, et al.
Pubblicazione: (2026)
di: Zhang, Chiyu, et al.
Pubblicazione: (2026)
Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption
di: Liu, Yepeng, et al.
Pubblicazione: (2025)
di: Liu, Yepeng, et al.
Pubblicazione: (2025)
EmMark: Robust Watermarks for IP Protection of Embedded Quantized Large Language Models
di: Zhang, Ruisi, et al.
Pubblicazione: (2024)
di: Zhang, Ruisi, et al.
Pubblicazione: (2024)
LLM-Enhanced Software Patch Localization
di: Yu, Jinhong, et al.
Pubblicazione: (2024)
di: Yu, Jinhong, et al.
Pubblicazione: (2024)
LLM-Powered Detection of Price Manipulation in DeFi
di: Liu, Lu, et al.
Pubblicazione: (2025)
di: Liu, Lu, et al.
Pubblicazione: (2025)
SRTJ: Self-Evolving Rule-Driven Training-Free LLM Jailbreaking
di: Li, Jindong, et al.
Pubblicazione: (2026)
di: Li, Jindong, et al.
Pubblicazione: (2026)
Counterfactual Evaluation for Blind Attack Detection in LLM-based Evaluation Systems
di: Liu, Lijia, et al.
Pubblicazione: (2025)
di: Liu, Lijia, et al.
Pubblicazione: (2025)
TimeMark: A Trustworthy Time Watermarking Framework for Exact Generation-Time Recovery from AIGC
di: Che, Shangkun, et al.
Pubblicazione: (2026)
di: Che, Shangkun, et al.
Pubblicazione: (2026)
CycleGANWM: A CycleGAN watermarking method for ownership verification
di: Lin, Dongdong, et al.
Pubblicazione: (2022)
di: Lin, Dongdong, et al.
Pubblicazione: (2022)
Overriding Safety protections of Open-source Models
di: Kumar, Sachin
Pubblicazione: (2024)
di: Kumar, Sachin
Pubblicazione: (2024)
Model-Agnostic Lifelong LLM Safety via Externalized Attack-Defense Co-Evolution
di: Zhang, Xiaozhe, et al.
Pubblicazione: (2026)
di: Zhang, Xiaozhe, et al.
Pubblicazione: (2026)
Two Birds with One Stone: Multi-Task Detection and Attribution of LLM-Generated Text
di: Rao, Zixin, et al.
Pubblicazione: (2025)
di: Rao, Zixin, et al.
Pubblicazione: (2025)
What Breaks Embodied AI Security:LLM Vulnerabilities, CPS Flaws,or Something Else?
di: Ma, Boyang, et al.
Pubblicazione: (2026)
di: Ma, Boyang, et al.
Pubblicazione: (2026)
Multi-Agent Collaboration in Incident Response with Large Language Models
di: Liu, Zefang
Pubblicazione: (2024)
di: Liu, Zefang
Pubblicazione: (2024)
AdaSteer: Your Aligned LLM is Inherently an Adaptive Jailbreak Defender
di: Zhao, Weixiang, et al.
Pubblicazione: (2025)
di: Zhao, Weixiang, et al.
Pubblicazione: (2025)
Security Attacks on LLM-based Code Completion Tools
di: Cheng, Wen, et al.
Pubblicazione: (2024)
di: Cheng, Wen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models
di: Gong, Yuyang, et al.
Pubblicazione: (2025) -
Enhance Robustness of Language Models Against Variation Attack through Graph Integration
di: Xiong, Zi, et al.
Pubblicazione: (2024) -
Black-Box Opinion Manipulation Attacks to Retrieval-Augmented Generation of Large Language Models
di: Chen, Zhuo, et al.
Pubblicazione: (2024) -
Proving membership in LLM pretraining data via data watermarks
di: Wei, Johnny Tian-Zheng, et al.
Pubblicazione: (2024) -
Optimizing watermarks for large language models
di: Wouters, Bram
Pubblicazione: (2023)