WaterJudge: Quality-Detection Trade-off when Watermarking Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Molenda, Piotr, Liusie, Adian, Gales, Mark J. F. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment
by: Raina, Vyas, et al.
Published: (2024)
by: Raina, Vyas, et al.
Published: (2024)
Teacher-Student Training for Debiasing: General Permutation Debiasing for Large Language Models
by: Liusie, Adian, et al.
Published: (2024)
by: Liusie, Adian, et al.
Published: (2024)
LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models
by: Liusie, Adian, et al.
Published: (2023)
by: Liusie, Adian, et al.
Published: (2023)
Finetuning LLMs for Comparative Assessment Tasks
by: Raina, Vatsal, et al.
Published: (2024)
by: Raina, Vatsal, et al.
Published: (2024)
Efficient LLM Comparative Assessment: a Product of Experts Framework for Pairwise Comparisons
by: Liusie, Adian, et al.
Published: (2024)
by: Liusie, Adian, et al.
Published: (2024)
Investigating the Emergent Audio Classification Ability of ASR Foundation Models
by: Ma, Rao, et al.
Published: (2023)
by: Ma, Rao, et al.
Published: (2023)
Zero-shot Audio Topic Reranking using Large Language Models
by: Qian, Mengjie, et al.
Published: (2023)
by: Qian, Mengjie, et al.
Published: (2023)
CrossCheckGPT: Universal Hallucination Ranking for Multimodal Foundation Models
by: Sun, Guangzhi, et al.
Published: (2024)
by: Sun, Guangzhi, et al.
Published: (2024)
WaterSearch: Exploring Seed Pooling for Improving the Quality-Detectability Trade-off in LLM Watermarking
by: Lin, Yukang, et al.
Published: (2025)
by: Lin, Yukang, et al.
Published: (2025)
Is it Possible to Modify Text to a Target Readability Level? An Initial Investigation Using Zero-Shot Large Language Models
by: Farajidizaji, Asma, et al.
Published: (2023)
by: Farajidizaji, Asma, et al.
Published: (2023)
From Trade-off to Synergy: A Versatile Symbiotic Watermarking Framework for Large Language Models
by: Wang, Yidan, et al.
Published: (2025)
by: Wang, Yidan, et al.
Published: (2025)
Blending Is All You Need: Cheaper, Better Alternative to Trillion-Parameters LLM
by: Lu, Xiaoding, et al.
Published: (2024)
by: Lu, Xiaoding, et al.
Published: (2024)
Embedding the Teacher: Distilling vLLM Preferences for Scalable Image Retrieval
by: He, Eric, et al.
Published: (2025)
by: He, Eric, et al.
Published: (2025)
ASR Error Correction using Large Language Models
by: Ma, Rao, et al.
Published: (2024)
by: Ma, Rao, et al.
Published: (2024)
The Impact of Editorial Intervention on Detecting Native Language Traces
by: Uluslu, Ahmet Yavuz, et al.
Published: (2026)
by: Uluslu, Ahmet Yavuz, et al.
Published: (2026)
Question-Based Retrieval using Atomic Units for Enterprise RAG
by: Raina, Vatsal, et al.
Published: (2024)
by: Raina, Vatsal, et al.
Published: (2024)
No Free Lunch in LLM Watermarking: Trade-offs in Watermarking Design Choices
by: Pang, Qi, et al.
Published: (2024)
by: Pang, Qi, et al.
Published: (2024)
Efficient Sample-Specific Encoder Perturbations
by: Fathullah, Yassir, et al.
Published: (2024)
by: Fathullah, Yassir, et al.
Published: (2024)
WaterPool: A Watermark Mitigating Trade-offs among Imperceptibility, Efficacy and Robustness
by: Huang, Baizhou, et al.
Published: (2024)
by: Huang, Baizhou, et al.
Published: (2024)
Controlling Whisper: Universal Acoustic Adversarial Attacks to Control Speech Foundation Models
by: Raina, Vyas, et al.
Published: (2024)
by: Raina, Vyas, et al.
Published: (2024)
Understanding the Quality-Diversity Trade-off in Diffusion Language Models
by: Buzzard, Zak
Published: (2025)
by: Buzzard, Zak
Published: (2025)
Question Difficulty Ranking for Multiple-Choice Reading Comprehension
by: Raina, Vatsal, et al.
Published: (2024)
by: Raina, Vatsal, et al.
Published: (2024)
Stylometric Watermarks for Large Language Models
by: Niess, Georg, et al.
Published: (2024)
by: Niess, Georg, et al.
Published: (2024)
Ensemble Watermarks for Large Language Models
by: Niess, Georg, et al.
Published: (2024)
by: Niess, Georg, et al.
Published: (2024)
In-Context Watermarks for Large Language Models
by: Liu, Yepeng, et al.
Published: (2025)
by: Liu, Yepeng, et al.
Published: (2025)
Linguistic Nepotism: Trading-off Quality for Language Preference in Multilingual RAG
by: Ki, Dayeon, et al.
Published: (2025)
by: Ki, Dayeon, et al.
Published: (2025)
Assessment of L2 Oral Proficiency using Speech Large Language Models
by: Ma, Rao, et al.
Published: (2025)
by: Ma, Rao, et al.
Published: (2025)
A Probability--Quality Trade-off in Aligned Language Models and its Relation to Sampling Adaptors
by: Tan, Naaman, et al.
Published: (2024)
by: Tan, Naaman, et al.
Published: (2024)
Multi-Bit Distortion-Free Watermarking for Large Language Models
by: Boroujeny, Massieh Kordi, et al.
Published: (2024)
by: Boroujeny, Massieh Kordi, et al.
Published: (2024)
Speak & Improve Corpus 2025: an L2 English Speech Corpus for Language Assessment and Feedback
by: Knill, Kate, et al.
Published: (2024)
by: Knill, Kate, et al.
Published: (2024)
Downstream Trade-offs of a Family of Text Watermarks
by: Ajith, Anirudh, et al.
Published: (2023)
by: Ajith, Anirudh, et al.
Published: (2023)
Adaptive Text Watermark for Large Language Models
by: Liu, Yepeng, et al.
Published: (2024)
by: Liu, Yepeng, et al.
Published: (2024)
Improved Unbiased Watermark for Large Language Models
by: Chen, Ruibo, et al.
Published: (2025)
by: Chen, Ruibo, et al.
Published: (2025)
Exploring Accuracy-Fairness Trade-off in Large Language Models
by: Zhang, Qingquan, et al.
Published: (2024)
by: Zhang, Qingquan, et al.
Published: (2024)
Towards Self-Referential Analytic Assessment: A Profile-Based Approach to L2 Writing Evaluation with LLMs
by: Bannò, Stefano, et al.
Published: (2026)
by: Bannò, Stefano, et al.
Published: (2026)
Grammatical Error Feedback: An Implicit Evaluation Approach
by: Bannò, Stefano, et al.
Published: (2024)
by: Bannò, Stefano, et al.
Published: (2024)
Improving Detection of Watermarked Language Models
by: Bahri, Dara, et al.
Published: (2025)
by: Bahri, Dara, et al.
Published: (2025)
D-Models and E-Models: Diversity-Stability Trade-offs in the Sampling Behavior of Large Language Models
by: Gu, Jia, et al.
Published: (2026)
by: Gu, Jia, et al.
Published: (2026)
WaterBench: Towards Holistic Evaluation of Watermarks for Large Language Models
by: Tu, Shangqing, et al.
Published: (2023)
by: Tu, Shangqing, et al.
Published: (2023)
De-mark: Watermark Removal in Large Language Models
by: Chen, Ruibo, et al.
Published: (2024)
by: Chen, Ruibo, et al.
Published: (2024)
Similar Items
-
Is LLM-as-a-Judge Robust? Investigating Universal Adversarial Attacks on Zero-shot LLM Assessment
by: Raina, Vyas, et al.
Published: (2024) -
Teacher-Student Training for Debiasing: General Permutation Debiasing for Large Language Models
by: Liusie, Adian, et al.
Published: (2024) -
LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models
by: Liusie, Adian, et al.
Published: (2023) -
Finetuning LLMs for Comparative Assessment Tasks
by: Raina, Vatsal, et al.
Published: (2024) -
Efficient LLM Comparative Assessment: a Product of Experts Framework for Pairwise Comparisons
by: Liusie, Adian, et al.
Published: (2024)