Gespeichert in:
| Hauptverfasser: | Chen, Yingquan, Li, Qianmu, Wu, Xiaocong, Li, Huifeng, Chang, Qing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2505.00977 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs
von: Mathew, Yohan, et al.
Veröffentlicht: (2024)
von: Mathew, Yohan, et al.
Veröffentlicht: (2024)
Zero-shot Generative Linguistic Steganography
von: Lin, Ke, et al.
Veröffentlicht: (2024)
von: Lin, Ke, et al.
Veröffentlicht: (2024)
TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking Agent
von: Meier, Dominik, et al.
Veröffentlicht: (2025)
von: Meier, Dominik, et al.
Veröffentlicht: (2025)
Fight Poison with Poison: Enhancing Robustness in Few-shot Machine-Generated Text Detection with Adversarial Training
von: Duan, Wenjing, et al.
Veröffentlicht: (2026)
von: Duan, Wenjing, et al.
Veröffentlicht: (2026)
Safe Text-to-Image Generation: Simply Sanitize the Prompt Embedding
von: Qiu, Huming, et al.
Veröffentlicht: (2024)
von: Qiu, Huming, et al.
Veröffentlicht: (2024)
TSCheater: Generating High-Quality Tibetan Adversarial Texts via Visual Similarity
von: Cao, Xi, et al.
Veröffentlicht: (2024)
von: Cao, Xi, et al.
Veröffentlicht: (2024)
StegoStylo: Squelching Stylometric Scrutiny through Steganographic Stitching
von: Dilworth, Robert
Veröffentlicht: (2026)
von: Dilworth, Robert
Veröffentlicht: (2026)
GradEscape: A Gradient-Based Evader Against AI-Generated Text Detectors
von: Meng, Wenlong, et al.
Veröffentlicht: (2025)
von: Meng, Wenlong, et al.
Veröffentlicht: (2025)
SeqAR: Jailbreak LLMs with Sequential Auto-Generated Characters
von: Yang, Yan, et al.
Veröffentlicht: (2024)
von: Yang, Yan, et al.
Veröffentlicht: (2024)
Early Signs of Steganographic Capabilities in Frontier LLMs
von: Zolkowski, Artur, et al.
Veröffentlicht: (2025)
von: Zolkowski, Artur, et al.
Veröffentlicht: (2025)
MGTEVAL: An Interactive Platform for Systemtic Evaluation of Machine-Generated Text Detectors
von: Li, Yuanfan, et al.
Veröffentlicht: (2026)
von: Li, Yuanfan, et al.
Veröffentlicht: (2026)
A High-Capacity and Secure Disambiguation Algorithm for Neural Linguistic Steganography
von: Feng, Yapei, et al.
Veröffentlicht: (2025)
von: Feng, Yapei, et al.
Veröffentlicht: (2025)
A General Pseudonymization Framework for Cloud-Based LLMs: Replacing Privacy Information in Controlled Text Generation
von: Hou, Shilong, et al.
Veröffentlicht: (2025)
von: Hou, Shilong, et al.
Veröffentlicht: (2025)
Iron Sharpens Iron: Defending Against Attacks in Machine-Generated Text Detection with Adversarial Training
von: Li, Yuanfan, et al.
Veröffentlicht: (2025)
von: Li, Yuanfan, et al.
Veröffentlicht: (2025)
Less is More: Sparse Watermarking in LLMs with Enhanced Text Quality
von: Hoang, Duy C., et al.
Veröffentlicht: (2024)
von: Hoang, Duy C., et al.
Veröffentlicht: (2024)
Beyond Text: Unveiling Privacy Vulnerabilities in Multi-modal Retrieval-Augmented Generation
von: Zhang, Jiankun, et al.
Veröffentlicht: (2025)
von: Zhang, Jiankun, et al.
Veröffentlicht: (2025)
LLMs for Domain Generation Algorithm Detection
von: La O, Reynier Leyva, et al.
Veröffentlicht: (2024)
von: La O, Reynier Leyva, et al.
Veröffentlicht: (2024)
Adversarial Text Generation with Dynamic Contextual Perturbation
von: Waghela, Hetvi, et al.
Veröffentlicht: (2025)
von: Waghela, Hetvi, et al.
Veröffentlicht: (2025)
GIFDL: Generated Image Fluctuation Distortion Learning for Enhancing Steganographic Security
von: Wang, Xiangkun, et al.
Veröffentlicht: (2025)
von: Wang, Xiangkun, et al.
Veröffentlicht: (2025)
Provably Secure Disambiguating Neural Linguistic Steganography
von: Qi, Yuang, et al.
Veröffentlicht: (2024)
von: Qi, Yuang, et al.
Veröffentlicht: (2024)
From Thinking to Output: Chain-of-Thought and Text Generation Characteristics in Reasoning Language Models
von: Liu, Junhao, et al.
Veröffentlicht: (2025)
von: Liu, Junhao, et al.
Veröffentlicht: (2025)
LingoLoop Attack: Trapping MLLMs via Linguistic Context and State Entrapment into Endless Loops
von: Fu, Jiyuan, et al.
Veröffentlicht: (2025)
von: Fu, Jiyuan, et al.
Veröffentlicht: (2025)
Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment
von: Wang, Jiongxiao, et al.
Veröffentlicht: (2024)
von: Wang, Jiongxiao, et al.
Veröffentlicht: (2024)
A Content-Preserving Secure Linguistic Steganography
von: Xiang, Lingyun, et al.
Veröffentlicht: (2025)
von: Xiang, Lingyun, et al.
Veröffentlicht: (2025)
Block-wise Codeword Embedding for Reliable Multi-bit Text Watermarking
von: Kim, Joeun, et al.
Veröffentlicht: (2026)
von: Kim, Joeun, et al.
Veröffentlicht: (2026)
Purified and Unified Steganographic Network
von: Li, Guobiao, et al.
Veröffentlicht: (2024)
von: Li, Guobiao, et al.
Veröffentlicht: (2024)
Text Embedding Inversion Security for Multilingual Language Models
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
Human-in-the-Loop Generation of Adversarial Texts: A Case Study on Tibetan Script
von: Cao, Xi, et al.
Veröffentlicht: (2024)
von: Cao, Xi, et al.
Veröffentlicht: (2024)
GenBreak: Red Teaming Text-to-Image Generators Using Large Language Models
von: Wang, Zilong, et al.
Veröffentlicht: (2025)
von: Wang, Zilong, et al.
Veröffentlicht: (2025)
RegionMarker: A Region-Triggered Semantic Watermarking Framework for Embedding-as-a-Service Copyright Protection
von: Yang, Shufan, et al.
Veröffentlicht: (2025)
von: Yang, Shufan, et al.
Veröffentlicht: (2025)
Simulate and Eliminate: Revoke Backdoors for Generative Large Language Models
von: Li, Haoran, et al.
Veröffentlicht: (2024)
von: Li, Haoran, et al.
Veröffentlicht: (2024)
One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs
von: Li, Linbao, et al.
Veröffentlicht: (2025)
von: Li, Linbao, et al.
Veröffentlicht: (2025)
TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts
von: Chu, Hua-Rong, et al.
Veröffentlicht: (2026)
von: Chu, Hua-Rong, et al.
Veröffentlicht: (2026)
Differentially Private Synthetic Text Generation for Retrieval-Augmented Generation (RAG)
von: Mori, Junki, et al.
Veröffentlicht: (2025)
von: Mori, Junki, et al.
Veröffentlicht: (2025)
AEIOU: A Unified Defense Framework against NSFW Prompts in Text-to-Image Models
von: Wang, Yiming, et al.
Veröffentlicht: (2024)
von: Wang, Yiming, et al.
Veröffentlicht: (2024)
Efficient Provably Secure Linguistic Steganography via Range Coding
von: Yan, Ruiyi, et al.
Veröffentlicht: (2026)
von: Yan, Ruiyi, et al.
Veröffentlicht: (2026)
Watermarking Conditional Text Generation for AI Detection: Unveiling Challenges and a Semantic-Aware Watermark Remedy
von: Fu, Yu, et al.
Veröffentlicht: (2023)
von: Fu, Yu, et al.
Veröffentlicht: (2023)
Secret-Protected Evolution for Differentially Private Synthetic Text Generation
von: Wang, Tianze, et al.
Veröffentlicht: (2025)
von: Wang, Tianze, et al.
Veröffentlicht: (2025)
Disabling Self-Correction in Retrieval-Augmented Generation via Stealthy Retriever Poisoning
von: Dai, Yanbo, et al.
Veröffentlicht: (2025)
von: Dai, Yanbo, et al.
Veröffentlicht: (2025)
GASLITEing the Retrieval: Exploring Vulnerabilities in Dense Embedding-based Search
von: Ben-Tov, Matan, et al.
Veröffentlicht: (2024)
von: Ben-Tov, Matan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs
von: Mathew, Yohan, et al.
Veröffentlicht: (2024) -
Zero-shot Generative Linguistic Steganography
von: Lin, Ke, et al.
Veröffentlicht: (2024) -
TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking Agent
von: Meier, Dominik, et al.
Veröffentlicht: (2025) -
Fight Poison with Poison: Enhancing Robustness in Few-shot Machine-Generated Text Detection with Adversarial Training
von: Duan, Wenjing, et al.
Veröffentlicht: (2026) -
Safe Text-to-Image Generation: Simply Sanitize the Prompt Embedding
von: Qiu, Huming, et al.
Veröffentlicht: (2024)