HeavyWater and SimplexWater: Distortion-Free LLM Watermarks for Low-Entropy Next-Token Predictions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tsur, Dor, Long, Carol Xuan, Verdun, Claudio Mayrink, Hsu, Hsiang, Chen, Chen-Fu, Permuter, Haim, Vithana, Sajani, Calmon, Flavio P. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Optimized Couplings for Watermarking Large Language Models
von: Tsur, Dor, et al.
Veröffentlicht: (2025)
von: Tsur, Dor, et al.
Veröffentlicht: (2025)
Multi-Group Proportional Representation for Text-to-Image Models
von: Jung, Sangwon, et al.
Veröffentlicht: (2025)
von: Jung, Sangwon, et al.
Veröffentlicht: (2025)
InfoMat: A Tool for the Analysis and Visualization Sequential Information Transfer
von: Tsur, Dor, et al.
Veröffentlicht: (2024)
von: Tsur, Dor, et al.
Veröffentlicht: (2024)
ArcMark: Distortion-Free Multi-Byte LLM Watermark via Optimal Transport
von: Gilani, Atefeh, et al.
Veröffentlicht: (2026)
von: Gilani, Atefeh, et al.
Veröffentlicht: (2026)
TREET: TRansfer Entropy Estimation via Transformers
von: Luxembourg, Omer, et al.
Veröffentlicht: (2024)
von: Luxembourg, Omer, et al.
Veröffentlicht: (2024)
Efficient Time Series Forecasting via Hyper-Complex Models and Frequency Aggregation
von: Yakir, Eyal, et al.
Veröffentlicht: (2025)
von: Yakir, Eyal, et al.
Veröffentlicht: (2025)
Multi-Group Proportional Representation in Retrieval
von: Oesterling, Alex, et al.
Veröffentlicht: (2024)
von: Oesterling, Alex, et al.
Veröffentlicht: (2024)
Temporal Sparse Autoencoders: Leveraging the Sequential Nature of Language for Interpretability
von: Bhalla, Usha, et al.
Veröffentlicht: (2025)
von: Bhalla, Usha, et al.
Veröffentlicht: (2025)
Correlated Privacy Mechanisms for Differentially Private Distributed Mean Estimation
von: Vithana, Sajani, et al.
Veröffentlicht: (2024)
von: Vithana, Sajani, et al.
Veröffentlicht: (2024)
Neural Estimation for Scaling Entropic Multimarginal Optimal Transport
von: Tsur, Dor, et al.
Veröffentlicht: (2025)
von: Tsur, Dor, et al.
Veröffentlicht: (2025)
Directed Information: Estimation, Optimization and Applications in Communications and Causality
von: Tsur, Dor, et al.
Veröffentlicht: (2026)
von: Tsur, Dor, et al.
Veröffentlicht: (2026)
Soft Best-of-n Sampling for Model Alignment
von: Verdun, Claudio Mayrink, et al.
Veröffentlicht: (2025)
von: Verdun, Claudio Mayrink, et al.
Veröffentlicht: (2025)
GradPCA: Leveraging NTK Alignment for Reliable Out-of-Distribution Detection
von: Seleznova, Mariia, et al.
Veröffentlicht: (2025)
von: Seleznova, Mariia, et al.
Veröffentlicht: (2025)
AI Alignment at Your Discretion
von: Buyl, Maarten, et al.
Veröffentlicht: (2025)
von: Buyl, Maarten, et al.
Veröffentlicht: (2025)
Task-Centric Acceleration of Small-Language Models
von: Tsur, Dor, et al.
Veröffentlicht: (2026)
von: Tsur, Dor, et al.
Veröffentlicht: (2026)
Inference-Time Reward Hacking in Large Language Models
von: Khalaf, Hadi, et al.
Veröffentlicht: (2025)
von: Khalaf, Hadi, et al.
Veröffentlicht: (2025)
More Haste, Less Speed: Weaker Single-Layer Watermark Improves Distortion-Free Watermark Ensembles
von: Chen, Ruibo, et al.
Veröffentlicht: (2026)
von: Chen, Ruibo, et al.
Veröffentlicht: (2026)
Selective Explanations
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2024)
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2024)
Invisible Entropy: Towards Safe and Efficient Low-Entropy LLM Watermarking
von: Gu, Tianle, et al.
Veröffentlicht: (2025)
von: Gu, Tianle, et al.
Veröffentlicht: (2025)
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
von: Natan, Shahar Ben, et al.
Veröffentlicht: (2026)
von: Natan, Shahar Ben, et al.
Veröffentlicht: (2026)
Robust Distortion-free Watermarks for Language Models
von: Kuditipudi, Rohith, et al.
Veröffentlicht: (2023)
von: Kuditipudi, Rohith, et al.
Veröffentlicht: (2023)
Multi-Bit Distortion-Free Watermarking for Large Language Models
von: Boroujeny, Massieh Kordi, et al.
Veröffentlicht: (2024)
von: Boroujeny, Massieh Kordi, et al.
Veröffentlicht: (2024)
Robust LLM Watermarking with Minimal Semantic Distortion for IP Protection
von: Dang, Kieu, et al.
Veröffentlicht: (2026)
von: Dang, Kieu, et al.
Veröffentlicht: (2026)
LongCat-Next: Lexicalizing Modalities as Discrete Tokens
von: Meituan LongCat Team, et al.
Veröffentlicht: (2026)
von: Meituan LongCat Team, et al.
Veröffentlicht: (2026)
Object Recognition as Next Token Prediction
von: Yue, Kaiyu, et al.
Veröffentlicht: (2023)
von: Yue, Kaiyu, et al.
Veröffentlicht: (2023)
Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generation
von: Ren, Sucheng, et al.
Veröffentlicht: (2025)
von: Ren, Sucheng, et al.
Veröffentlicht: (2025)
An Entropy-based Text Watermarking Detection Method
von: Lu, Yijian, et al.
Veröffentlicht: (2024)
von: Lu, Yijian, et al.
Veröffentlicht: (2024)
Probing Geometry of Next Token Prediction Using Cumulant Expansion of the Softmax Entropy
von: Viswanathan, Karthik, et al.
Veröffentlicht: (2025)
von: Viswanathan, Karthik, et al.
Veröffentlicht: (2025)
Measuring Progress in Dictionary Learning for Language Model Interpretability with Board Game Models
von: Karvonen, Adam, et al.
Veröffentlicht: (2024)
von: Karvonen, Adam, et al.
Veröffentlicht: (2024)
TokenPure: Watermark Removal through Tokenized Appearance and Structural Guidance
von: Yang, Pei, et al.
Veröffentlicht: (2025)
von: Yang, Pei, et al.
Veröffentlicht: (2025)
TokenTrace: Multi-Concept Attribution through Watermarked Token Recovery
von: Zhang, Li, et al.
Veröffentlicht: (2026)
von: Zhang, Li, et al.
Veröffentlicht: (2026)
From Next-Token to Next-Block: A Principled Adaptation Path for Diffusion LLMs
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
Acquired TASTE: Multimodal Stance Detection with Textual and Structural Embeddings
von: Barel, Guy, et al.
Veröffentlicht: (2024)
von: Barel, Guy, et al.
Veröffentlicht: (2024)
Aleatoric and Epistemic Discrimination: Fundamental Limits of Fairness Interventions
von: Wang, Hao, et al.
Veröffentlicht: (2023)
von: Wang, Hao, et al.
Veröffentlicht: (2023)
Diversity or Precision? A Deep Dive into Next Token Prediction
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
PMark: Towards Robust and Distortion-free Semantic-level Watermarking with Channel Constraints
von: Huo, Jiahao, et al.
Veröffentlicht: (2025)
von: Huo, Jiahao, et al.
Veröffentlicht: (2025)
WaterSearch: Exploring Seed Pooling for Improving the Quality-Detectability Trade-off in LLM Watermarking
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
von: Lin, Yukang, et al.
Veröffentlicht: (2025)
Addressing Tokenization Inconsistency in Steganography and Watermarking Based on Large Language Models
von: Yan, Ruiyi, et al.
Veröffentlicht: (2025)
von: Yan, Ruiyi, et al.
Veröffentlicht: (2025)
NoisePrints: Distortion-Free Watermarks for Authorship in Private Diffusion Models
von: Goren, Nir, et al.
Veröffentlicht: (2025)
von: Goren, Nir, et al.
Veröffentlicht: (2025)
Beyond Next-Token Alignment: Distilling Multimodal Large Language Models via Token Interactions
von: Chen, Lin, et al.
Veröffentlicht: (2026)
von: Chen, Lin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Optimized Couplings for Watermarking Large Language Models
von: Tsur, Dor, et al.
Veröffentlicht: (2025) -
Multi-Group Proportional Representation for Text-to-Image Models
von: Jung, Sangwon, et al.
Veröffentlicht: (2025) -
InfoMat: A Tool for the Analysis and Visualization Sequential Information Transfer
von: Tsur, Dor, et al.
Veröffentlicht: (2024) -
ArcMark: Distortion-Free Multi-Byte LLM Watermark via Optimal Transport
von: Gilani, Atefeh, et al.
Veröffentlicht: (2026) -
TREET: TRansfer Entropy Estimation via Transformers
von: Luxembourg, Omer, et al.
Veröffentlicht: (2024)