Can Watermarked LLMs be Identified by Users via Crafted Prompts?
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Aiwei, Guan, Sheng, Liu, Yiming, Pan, Leyi, Zhang, Yifei, Fang, Liancheng, Wen, Lijie, Yu, Philip S., Hu, Xuming |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Semantic Invariant Robust Watermark for Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
MarkDiffusion: An Open-Source Toolkit for Generative Watermarking of Latent Diffusion Models
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
MarkLLM: An Open-Source Toolkit for LLM Watermarking
di: Pan, Leyi, et al.
Pubblicazione: (2024)
di: Pan, Leyi, et al.
Pubblicazione: (2024)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
A Survey of Text Watermarking in the Era of Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
WaterSeeker: Pioneering Efficient Detection of Watermarked Segments in Large Documents
di: Pan, Leyi, et al.
Pubblicazione: (2024)
di: Pan, Leyi, et al.
Pubblicazione: (2024)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
On the Robustness of Document-Level Relation Extraction Models to Entity Name Variations
di: Meng, Shiao, et al.
Pubblicazione: (2024)
di: Meng, Shiao, et al.
Pubblicazione: (2024)
A Survey on Parallel Text Generation: From Parallel Decoding to Diffusion Language Models
di: Zhang, Lingzhe, et al.
Pubblicazione: (2025)
di: Zhang, Lingzhe, et al.
Pubblicazione: (2025)
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
ChatCite: LLM Agent with Human Workflow Guidance for Comparative Literature Summary
di: Li, Yutong, et al.
Pubblicazione: (2024)
di: Li, Yutong, et al.
Pubblicazione: (2024)
TIS-DPO: Token-level Importance Sampling for Direct Preference Optimization With Estimated Weights
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
di: Liu, Aiwei, et al.
Pubblicazione: (2024)
MicroRemed: Benchmarking LLMs in Microservices Remediation
di: Zhang, Lingzhe, et al.
Pubblicazione: (2025)
di: Zhang, Lingzhe, et al.
Pubblicazione: (2025)
AVEC: Bootstrapping Privacy for Local LLMs
di: Gaikwad, Madhava
Pubblicazione: (2025)
di: Gaikwad, Madhava
Pubblicazione: (2025)
Can LLMs Compute with Reasons?
di: Sandilya, Harshit, et al.
Pubblicazione: (2024)
di: Sandilya, Harshit, et al.
Pubblicazione: (2024)
Prompt Engineering and the Effectiveness of Large Language Models in Enhancing Human Productivity
di: Anam, Rizal Khoirul
Pubblicazione: (2025)
di: Anam, Rizal Khoirul
Pubblicazione: (2025)
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
di: Hill, Brennen, et al.
Pubblicazione: (2025)
di: Hill, Brennen, et al.
Pubblicazione: (2025)
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference
di: Liu, Aiwei, et al.
Pubblicazione: (2025)
di: Liu, Aiwei, et al.
Pubblicazione: (2025)
Dark LLMs: The Growing Threat of Unaligned AI Models
di: Fire, Michael, et al.
Pubblicazione: (2025)
di: Fire, Michael, et al.
Pubblicazione: (2025)
Logits of API-Protected LLMs Leak Proprietary Information
di: Finlayson, Matthew, et al.
Pubblicazione: (2024)
di: Finlayson, Matthew, et al.
Pubblicazione: (2024)
Shortcuts Arising from Contrast: Effective and Covert Clean-Label Attacks in Prompt-Based Learning
di: Xie, Xiaopeng, et al.
Pubblicazione: (2024)
di: Xie, Xiaopeng, et al.
Pubblicazione: (2024)
The Personalization Trap: How User Memory Alters Emotional Reasoning in LLMs
di: Fang, Xi, et al.
Pubblicazione: (2025)
di: Fang, Xi, et al.
Pubblicazione: (2025)
On the Validity of Traditional Vulnerability Scoring Systems for Adversarial Attacks against LLMs
di: Bahar, Atmane Ayoub Mansour, et al.
Pubblicazione: (2024)
di: Bahar, Atmane Ayoub Mansour, et al.
Pubblicazione: (2024)
How Few-shot Demonstrations Affect Prompt-based Defenses Against LLM Jailbreak Attacks
di: Wang, Yanshu, et al.
Pubblicazione: (2026)
di: Wang, Yanshu, et al.
Pubblicazione: (2026)
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
di: Dang, Kieu, et al.
Pubblicazione: (2025)
di: Dang, Kieu, et al.
Pubblicazione: (2025)
On Adversarial Examples for Text Classification by Perturbing Latent Representations
di: Sooksatra, Korn, et al.
Pubblicazione: (2024)
di: Sooksatra, Korn, et al.
Pubblicazione: (2024)
GATE: Graph-based Adaptive Tool Evolution Across Diverse Tasks
di: Luo, Jianwen, et al.
Pubblicazione: (2025)
di: Luo, Jianwen, et al.
Pubblicazione: (2025)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
di: Fang, Xi, et al.
Pubblicazione: (2024)
di: Fang, Xi, et al.
Pubblicazione: (2024)
LegalGuardian: A Privacy-Preserving Framework for Secure Integration of Large Language Models in Legal Practice
di: Demir, M. Mikail, et al.
Pubblicazione: (2025)
di: Demir, M. Mikail, et al.
Pubblicazione: (2025)
BreakFun: Jailbreaking LLMs via Schema Exploitation
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
di: Oskooei, Amirkia Rafiei, et al.
Pubblicazione: (2025)
Compression Method Matters: Benchmark-Dependent Output Dynamics in LLM Prompt Compression
di: Johnson, Warren
Pubblicazione: (2026)
di: Johnson, Warren
Pubblicazione: (2026)
Prompt Compression in Production Task Orchestration: A Pre-Registered Randomized Trial
di: Johnson, Warren, et al.
Pubblicazione: (2026)
di: Johnson, Warren, et al.
Pubblicazione: (2026)
LLMs for Legal Subsumption in German Employment Contracts
di: Wardas, Oliver, et al.
Pubblicazione: (2025)
di: Wardas, Oliver, et al.
Pubblicazione: (2025)
PropXplain: Can LLMs Enable Explainable Propaganda Detection?
di: Hasanain, Maram, et al.
Pubblicazione: (2025)
di: Hasanain, Maram, et al.
Pubblicazione: (2025)
Zero- and Few-Shot Prompting with LLMs: A Comparative Study with Fine-tuned Models for Bangla Sentiment Analysis
di: Hasan, Md. Arid, et al.
Pubblicazione: (2023)
di: Hasan, Md. Arid, et al.
Pubblicazione: (2023)
Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs
di: Balter, Samuel G., et al.
Pubblicazione: (2026)
di: Balter, Samuel G., et al.
Pubblicazione: (2026)
When Retrieval Succeeds and Fails: Rethinking Retrieval-Augmented Generation for LLMs
di: Wang, Yongjie, et al.
Pubblicazione: (2025)
di: Wang, Yongjie, et al.
Pubblicazione: (2025)
The Unlikely Duel: Evaluating Creative Writing in LLMs through a Unique Scenario
di: Gómez-Rodríguez, Carlos, et al.
Pubblicazione: (2024)
di: Gómez-Rodríguez, Carlos, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Semantic Invariant Robust Watermark for Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023) -
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
di: Pan, Leyi, et al.
Pubblicazione: (2025) -
MarkDiffusion: An Open-Source Toolkit for Generative Watermarking of Latent Diffusion Models
di: Pan, Leyi, et al.
Pubblicazione: (2025) -
MarkLLM: An Open-Source Toolkit for LLM Watermarking
di: Pan, Leyi, et al.
Pubblicazione: (2024) -
An Unforgeable Publicly Verifiable Watermark for Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023)