A Semantic Invariant Robust Watermark for Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Aiwei, Pan, Leyi, Hu, Xuming, Meng, Shiao, Wen, Lijie |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Can Watermarked LLMs be Identified by Users via Crafted Prompts?
por: Liu, Aiwei, et al.
Publicado: (2024)
por: Liu, Aiwei, et al.
Publicado: (2024)
MarkLLM: An Open-Source Toolkit for LLM Watermarking
por: Pan, Leyi, et al.
Publicado: (2024)
por: Pan, Leyi, et al.
Publicado: (2024)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
por: Liu, Aiwei, et al.
Publicado: (2023)
por: Liu, Aiwei, et al.
Publicado: (2023)
MarkDiffusion: An Open-Source Toolkit for Generative Watermarking of Latent Diffusion Models
por: Pan, Leyi, et al.
Publicado: (2025)
por: Pan, Leyi, et al.
Publicado: (2025)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
por: Pan, Leyi, et al.
Publicado: (2025)
por: Pan, Leyi, et al.
Publicado: (2025)
A Survey of Text Watermarking in the Era of Large Language Models
por: Liu, Aiwei, et al.
Publicado: (2023)
por: Liu, Aiwei, et al.
Publicado: (2023)
On the Robustness of Document-Level Relation Extraction Models to Entity Name Variations
por: Meng, Shiao, et al.
Publicado: (2024)
por: Meng, Shiao, et al.
Publicado: (2024)
WaterSeeker: Pioneering Efficient Detection of Watermarked Segments in Large Documents
por: Pan, Leyi, et al.
Publicado: (2024)
por: Pan, Leyi, et al.
Publicado: (2024)
Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models
por: Pan, Leyi, et al.
Publicado: (2025)
por: Pan, Leyi, et al.
Publicado: (2025)
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
por: Liu, Aiwei, et al.
Publicado: (2024)
por: Liu, Aiwei, et al.
Publicado: (2024)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
por: Pan, Leyi, et al.
Publicado: (2025)
por: Pan, Leyi, et al.
Publicado: (2025)
LegalGuardian: A Privacy-Preserving Framework for Secure Integration of Large Language Models in Legal Practice
por: Demir, M. Mikail, et al.
Publicado: (2025)
por: Demir, M. Mikail, et al.
Publicado: (2025)
Monotonicity as an Architectural Bias for Robust Language Models
por: Cooper, Patrick, et al.
Publicado: (2026)
por: Cooper, Patrick, et al.
Publicado: (2026)
TIS-DPO: Token-level Importance Sampling for Direct Preference Optimization With Estimated Weights
por: Liu, Aiwei, et al.
Publicado: (2024)
por: Liu, Aiwei, et al.
Publicado: (2024)
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference
por: Liu, Aiwei, et al.
Publicado: (2025)
por: Liu, Aiwei, et al.
Publicado: (2025)
ChatCite: LLM Agent with Human Workflow Guidance for Comparative Literature Summary
por: Li, Yutong, et al.
Publicado: (2024)
por: Li, Yutong, et al.
Publicado: (2024)
Logits of API-Protected LLMs Leak Proprietary Information
por: Finlayson, Matthew, et al.
Publicado: (2024)
por: Finlayson, Matthew, et al.
Publicado: (2024)
Shortcuts Arising from Contrast: Effective and Covert Clean-Label Attacks in Prompt-Based Learning
por: Xie, Xiaopeng, et al.
Publicado: (2024)
por: Xie, Xiaopeng, et al.
Publicado: (2024)
Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks
por: Merves, Tyler H., et al.
Publicado: (2026)
por: Merves, Tyler H., et al.
Publicado: (2026)
A Survey on Parallel Text Generation: From Parallel Decoding to Diffusion Language Models
por: Zhang, Lingzhe, et al.
Publicado: (2025)
por: Zhang, Lingzhe, et al.
Publicado: (2025)
On Adversarial Examples for Text Classification by Perturbing Latent Representations
por: Sooksatra, Korn, et al.
Publicado: (2024)
por: Sooksatra, Korn, et al.
Publicado: (2024)
Towards Effective and Efficient Continual Pre-training of Large Language Models
por: Chen, Jie, et al.
Publicado: (2024)
por: Chen, Jie, et al.
Publicado: (2024)
The Superalignment of Superhuman Intelligence with Large Language Models
por: Huang, Minlie, et al.
Publicado: (2024)
por: Huang, Minlie, et al.
Publicado: (2024)
Large Language Models(LLMs) on Tabular Data: Prediction, Generation, and Understanding -- A Survey
por: Fang, Xi, et al.
Publicado: (2024)
por: Fang, Xi, et al.
Publicado: (2024)
Exploring State Tracking Capabilities of Large Language Models
por: Rezaee, Kiamehr, et al.
Publicado: (2025)
por: Rezaee, Kiamehr, et al.
Publicado: (2025)
How Few-shot Demonstrations Affect Prompt-based Defenses Against LLM Jailbreak Attacks
por: Wang, Yanshu, et al.
Publicado: (2026)
por: Wang, Yanshu, et al.
Publicado: (2026)
Distilling Large Language Models for Efficient Clinical Information Extraction
por: Vedula, Karthik S., et al.
Publicado: (2024)
por: Vedula, Karthik S., et al.
Publicado: (2024)
Dark LLMs: The Growing Threat of Unaligned AI Models
por: Fire, Michael, et al.
Publicado: (2025)
por: Fire, Michael, et al.
Publicado: (2025)
Unifying Uniform and Binary-coding Quantization for Accurate Compression of Large Language Models
por: Park, Seungcheol, et al.
Publicado: (2025)
por: Park, Seungcheol, et al.
Publicado: (2025)
Accurate Sublayer Pruning for Large Language Models by Exploiting Latency and Tunability Information
por: Park, Seungcheol, et al.
Publicado: (2025)
por: Park, Seungcheol, et al.
Publicado: (2025)
Setting Standards in Turkish NLP: TR-MMLU for Large Language Model Evaluation
por: Bayram, M. Ali, et al.
Publicado: (2024)
por: Bayram, M. Ali, et al.
Publicado: (2024)
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
por: Hill, Brennen, et al.
Publicado: (2025)
por: Hill, Brennen, et al.
Publicado: (2025)
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
por: Dang, Kieu, et al.
Publicado: (2025)
por: Dang, Kieu, et al.
Publicado: (2025)
The Paradox of Poetic Intent in Back-Translation: Evaluating the Quality of Large Language Models in Chinese Translation
por: Weigang, Li, et al.
Publicado: (2025)
por: Weigang, Li, et al.
Publicado: (2025)
Unveiling Attractor Cycles in Large Language Models: A Dynamical Systems View of Successive Paraphrasing
por: Wang, Zhilin, et al.
Publicado: (2025)
por: Wang, Zhilin, et al.
Publicado: (2025)
NurValues: Real-World Nursing Values Evaluation for Large Language Models in Clinical Context
por: Yao, Ben, et al.
Publicado: (2025)
por: Yao, Ben, et al.
Publicado: (2025)
Trusted Uncertainty in Large Language Models: A Unified Framework for Confidence Calibration and Risk-Controlled Refusal
por: Oehri, Markus, et al.
Publicado: (2025)
por: Oehri, Markus, et al.
Publicado: (2025)
"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations
por: Choi, Jonghyeon, et al.
Publicado: (2025)
por: Choi, Jonghyeon, et al.
Publicado: (2025)
Generative AI Models: Opportunities and Risks for Industry and Authorities
por: Alt, Tobias, et al.
Publicado: (2024)
por: Alt, Tobias, et al.
Publicado: (2024)
Math Natural Language Inference: this should be easy!
por: de Paiva, Valeria, et al.
Publicado: (2025)
por: de Paiva, Valeria, et al.
Publicado: (2025)
Ejemplares similares
-
Can Watermarked LLMs be Identified by Users via Crafted Prompts?
por: Liu, Aiwei, et al.
Publicado: (2024) -
MarkLLM: An Open-Source Toolkit for LLM Watermarking
por: Pan, Leyi, et al.
Publicado: (2024) -
An Unforgeable Publicly Verifiable Watermark for Large Language Models
por: Liu, Aiwei, et al.
Publicado: (2023) -
MarkDiffusion: An Open-Source Toolkit for Generative Watermarking of Latent Diffusion Models
por: Pan, Leyi, et al.
Publicado: (2025) -
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
por: Pan, Leyi, et al.
Publicado: (2025)