Necessary and Sufficient Watermark for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Takezawa, Yuki, Sato, Ryoma, Bao, Han, Niwa, Kenta, Yamada, Makoto |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Parameter-free Clipped Gradient Descent Meets Polyak
von: Takezawa, Yuki, et al.
Veröffentlicht: (2024)
von: Takezawa, Yuki, et al.
Veröffentlicht: (2024)
PhiNets: Brain-inspired Non-contrastive Learning Based on Temporal Prediction Hypothesis
von: Ishikawa, Satoki, et al.
Veröffentlicht: (2024)
von: Ishikawa, Satoki, et al.
Veröffentlicht: (2024)
Even GPT-5.2 Can't Count to Five: The Case for Zero-Error Horizons in Trustworthy LLMs
von: Sato, Ryoma
Veröffentlicht: (2026)
von: Sato, Ryoma
Veröffentlicht: (2026)
Delayed Momentum Aggregation: Communication-efficient Byzantine-robust Federated Learning with Partial Participation
von: Otsuka, Kaoru, et al.
Veröffentlicht: (2025)
von: Otsuka, Kaoru, et al.
Veröffentlicht: (2025)
A Watermark for Large Language Models
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)
On the Reliability of Watermarks for Large Language Models
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)
User-Side Realization
von: Sato, Ryoma
Veröffentlicht: (2024)
von: Sato, Ryoma
Veröffentlicht: (2024)
Multi-Bit Distortion-Free Watermarking for Large Language Models
von: Boroujeny, Massieh Kordi, et al.
Veröffentlicht: (2024)
von: Boroujeny, Massieh Kordi, et al.
Veröffentlicht: (2024)
Topic-Based Watermarks for Large Language Models
von: Nemecek, Alexander, et al.
Veröffentlicht: (2024)
von: Nemecek, Alexander, et al.
Veröffentlicht: (2024)
Improving Detection of Watermarked Language Models
von: Bahri, Dara, et al.
Veröffentlicht: (2025)
von: Bahri, Dara, et al.
Veröffentlicht: (2025)
Watermarks for Embeddings-as-a-Service Large Language Models
von: Shetty, Anudeex
Veröffentlicht: (2025)
von: Shetty, Anudeex
Veröffentlicht: (2025)
$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks
von: Chowdhary, Pratim, et al.
Veröffentlicht: (2025)
von: Chowdhary, Pratim, et al.
Veröffentlicht: (2025)
TAID: Temporally Adaptive Interpolated Distillation for Efficient Knowledge Transfer in Language Models
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
von: Shing, Makoto, et al.
Veröffentlicht: (2025)
A Single Neuron Is Sufficient to Bypass Safety Alignment in Large Language Models
von: Kazemi, Hamid, et al.
Veröffentlicht: (2026)
von: Kazemi, Hamid, et al.
Veröffentlicht: (2026)
Duwak: Dual Watermarks in Large Language Models
von: Zhu, Chaoyi, et al.
Veröffentlicht: (2024)
von: Zhu, Chaoyi, et al.
Veröffentlicht: (2024)
On the Learnability of Watermarks for Language Models
von: Gu, Chenchen, et al.
Veröffentlicht: (2023)
von: Gu, Chenchen, et al.
Veröffentlicht: (2023)
Watermarking Language Models through Language Models
von: Dasgupta, Agnibh, et al.
Veröffentlicht: (2024)
von: Dasgupta, Agnibh, et al.
Veröffentlicht: (2024)
A Resilient and Accessible Distribution-Preserving Watermark for Large Language Models
von: Wu, Yihan, et al.
Veröffentlicht: (2023)
von: Wu, Yihan, et al.
Veröffentlicht: (2023)
Any-stepsize Gradient Descent for Separable Data under Fenchel-Young Losses
von: Bao, Han, et al.
Veröffentlicht: (2025)
von: Bao, Han, et al.
Veröffentlicht: (2025)
Adaptive Testing for Segmenting Watermarked Texts From Language Models
von: Li, Xingchi, et al.
Veröffentlicht: (2025)
von: Li, Xingchi, et al.
Veröffentlicht: (2025)
Debiasing Watermarks for Large Language Models via Maximal Coupling
von: Xie, Yangxinyu, et al.
Veröffentlicht: (2024)
von: Xie, Yangxinyu, et al.
Veröffentlicht: (2024)
Token-Specific Watermarking with Enhanced Detectability and Semantic Coherence for Large Language Models
von: Huo, Mingjia, et al.
Veröffentlicht: (2024)
von: Huo, Mingjia, et al.
Veröffentlicht: (2024)
Publicly-Detectable Watermarking for Language Models
von: Fairoze, Jaiden, et al.
Veröffentlicht: (2023)
von: Fairoze, Jaiden, et al.
Veröffentlicht: (2023)
Sparse-Autoencoder-Guided Internal Representation Unlearning for Large Language Models
von: Yamashita, Tomoya, et al.
Veröffentlicht: (2025)
von: Yamashita, Tomoya, et al.
Veröffentlicht: (2025)
JailNewsBench: Multi-Lingual and Regional Benchmark for Fake News Generation under Jailbreak Attacks
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2026)
von: Kaneko, Masahiro, et al.
Veröffentlicht: (2026)
Robust Detection of Watermarks for Large Language Models Under Human Edits
von: Li, Xiang, et al.
Veröffentlicht: (2024)
von: Li, Xiang, et al.
Veröffentlicht: (2024)
Concept Unlearning in Large Language Models via Self-Constructed Knowledge Triplets
von: Yamashita, Tomoya, et al.
Veröffentlicht: (2025)
von: Yamashita, Tomoya, et al.
Veröffentlicht: (2025)
Robust and Secure Code Watermarking for Large Language Models via ML/Crypto Codesign
von: Zhang, Ruisi, et al.
Veröffentlicht: (2025)
von: Zhang, Ruisi, et al.
Veröffentlicht: (2025)
Robust Distortion-free Watermarks for Language Models
von: Kuditipudi, Rohith, et al.
Veröffentlicht: (2023)
von: Kuditipudi, Rohith, et al.
Veröffentlicht: (2023)
Watermarking Language Models with Error Correcting Codes
von: Chao, Patrick, et al.
Veröffentlicht: (2024)
von: Chao, Patrick, et al.
Veröffentlicht: (2024)
A Watermark for Black-Box Language Models
von: Bahri, Dara, et al.
Veröffentlicht: (2024)
von: Bahri, Dara, et al.
Veröffentlicht: (2024)
Towards Principled Design of Mixture-of-Experts Language Models under Memory and Inference Constraints
von: Liew, Seng Pei, et al.
Veröffentlicht: (2026)
von: Liew, Seng Pei, et al.
Veröffentlicht: (2026)
On the Thinking-Language Modeling Gap in Large Language Models
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
PostMark: A Robust Blackbox Watermark for Large Language Models
von: Chang, Yapei, et al.
Veröffentlicht: (2024)
von: Chang, Yapei, et al.
Veröffentlicht: (2024)
Large Language Models on Graphs: A Comprehensive Survey
von: Jin, Bowen, et al.
Veröffentlicht: (2023)
von: Jin, Bowen, et al.
Veröffentlicht: (2023)
Watermarks in the Sand: Impossibility of Strong Watermarking for Generative Models
von: Zhang, Hanlin, et al.
Veröffentlicht: (2023)
von: Zhang, Hanlin, et al.
Veröffentlicht: (2023)
Is Random Attention Sufficient for Sequence Modeling? Disentangling Trainable Components in the Transformer
von: Dong, Yihe, et al.
Veröffentlicht: (2025)
von: Dong, Yihe, et al.
Veröffentlicht: (2025)
Robust Data Watermarking in Language Models by Injecting Fictitious Knowledge
von: Cui, Xinyue, et al.
Veröffentlicht: (2025)
von: Cui, Xinyue, et al.
Veröffentlicht: (2025)
Watermarking Makes Language Models Radioactive
von: Sander, Tom, et al.
Veröffentlicht: (2024)
von: Sander, Tom, et al.
Veröffentlicht: (2024)
SimMark: A Robust Sentence-Level Similarity-Based Watermarking Algorithm for Large Language Models
von: Dabiriaghdam, Amirhossein, et al.
Veröffentlicht: (2025)
von: Dabiriaghdam, Amirhossein, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Parameter-free Clipped Gradient Descent Meets Polyak
von: Takezawa, Yuki, et al.
Veröffentlicht: (2024) -
PhiNets: Brain-inspired Non-contrastive Learning Based on Temporal Prediction Hypothesis
von: Ishikawa, Satoki, et al.
Veröffentlicht: (2024) -
Even GPT-5.2 Can't Count to Five: The Case for Zero-Error Horizons in Trustworthy LLMs
von: Sato, Ryoma
Veröffentlicht: (2026) -
Delayed Momentum Aggregation: Communication-efficient Byzantine-robust Federated Learning with Partial Participation
von: Otsuka, Kaoru, et al.
Veröffentlicht: (2025) -
A Watermark for Large Language Models
von: Kirchenbauer, John, et al.
Veröffentlicht: (2023)