Mitigating Paraphrase Attacks on Machine-Text Detectors via Paraphrase Inversion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Soto, Rafael Rivera, Chen, Barry, Andrews, Nicholas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Language Models Optimized to Fool Detectors Still Have a Distinct Style (And How to Change It)
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2025)
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2025)
Revisiting the Robustness of Watermarking to Paraphrasing Attacks
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024)
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024)
Few-Shot Detection of Machine-Generated Text using Style Representations
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2024)
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2024)
MahaParaphrase: A Marathi Paraphrase Detection Corpus and BERT-based Models
von: Jadhav, Suramya, et al.
Veröffentlicht: (2025)
von: Jadhav, Suramya, et al.
Veröffentlicht: (2025)
Paraphrasing Adversarial Attack on LLM-as-a-Reviewer
von: Kaneko, Masahiro
Veröffentlicht: (2026)
von: Kaneko, Masahiro
Veröffentlicht: (2026)
Action Controlled Paraphrasing
von: Shi, Ning, et al.
Veröffentlicht: (2024)
von: Shi, Ning, et al.
Veröffentlicht: (2024)
VTechAGP: An Academic-to-General-Audience Text Paraphrase Dataset and Benchmark Models
von: Cheng, Ming, et al.
Veröffentlicht: (2024)
von: Cheng, Ming, et al.
Veröffentlicht: (2024)
Perturb Your Data: Paraphrase-Guided Training Data Watermarking
von: Shetty, Pranav, et al.
Veröffentlicht: (2025)
von: Shetty, Pranav, et al.
Veröffentlicht: (2025)
AliMark: Enhancing Robustness of Sentence-Level Watermarking Against Text Paraphrasing
von: Li, Yuexin, et al.
Veröffentlicht: (2026)
von: Li, Yuexin, et al.
Veröffentlicht: (2026)
Analyzing Persuasive Strategies in Meme Texts: A Fusion of Language Models with Paraphrase Enrichment
von: Nayak, Kota Shamanth Ramanath, et al.
Veröffentlicht: (2024)
von: Nayak, Kota Shamanth Ramanath, et al.
Veröffentlicht: (2024)
PRSM: A Measure to Evaluate CLIP's Robustness Against Paraphrases
von: Schlegel, Udo, et al.
Veröffentlicht: (2025)
von: Schlegel, Udo, et al.
Veröffentlicht: (2025)
AMPS: ASR with Multimodal Paraphrase Supervision
von: Gupta, Abhishek, et al.
Veröffentlicht: (2024)
von: Gupta, Abhishek, et al.
Veröffentlicht: (2024)
You Didn't Have to Say It like That: Subliminal Learning from Faithful Paraphrases
von: Gisler, Isaia, et al.
Veröffentlicht: (2026)
von: Gisler, Isaia, et al.
Veröffentlicht: (2026)
Sci-LoRA: Mixture of Scientific LoRAs for Cross-Domain Lay Paraphrasing
von: Cheng, Ming, et al.
Veröffentlicht: (2025)
von: Cheng, Ming, et al.
Veröffentlicht: (2025)
On the Impact of Language Nuances on Sentiment Analysis with Large Language Models: Paraphrasing, Sarcasm, and Emojis
von: Bhargava, Naman, et al.
Veröffentlicht: (2025)
von: Bhargava, Naman, et al.
Veröffentlicht: (2025)
Paraphrase and Aggregate with Large Language Models for Minimizing Intent Classification Errors
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
von: Yadav, Vikas, et al.
Veröffentlicht: (2024)
Paraphrase-Aligned Machine Translation
von: Chang, Ke-Ching, et al.
Veröffentlicht: (2024)
von: Chang, Ke-Ching, et al.
Veröffentlicht: (2024)
PADBen: A Comprehensive Benchmark for Evaluating AI Text Detectors Against Paraphrase Attacks
von: Zha, Yiwei, et al.
Veröffentlicht: (2025)
von: Zha, Yiwei, et al.
Veröffentlicht: (2025)
StealthRL: Reinforcement Learning Paraphrase Attacks for Multi-Detector Evasion of AI-Text Detectors
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026)
von: Ranganath, Suraj, et al.
Veröffentlicht: (2026)
Parameter Efficient Diverse Paraphrase Generation Using Sequence-Level Knowledge Distillation
von: Jayawardena, Lasal, et al.
Veröffentlicht: (2024)
von: Jayawardena, Lasal, et al.
Veröffentlicht: (2024)
WET: Overcoming Paraphrasing Vulnerabilities in Embeddings-as-a-Service with Linear Transformation Watermarks
von: Shetty, Anudeex, et al.
Veröffentlicht: (2024)
von: Shetty, Anudeex, et al.
Veröffentlicht: (2024)
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing
von: Jung, Jaehun, et al.
Veröffentlicht: (2023)
von: Jung, Jaehun, et al.
Veröffentlicht: (2023)
Adversarial Paraphrasing: A Universal Attack for Humanizing AI-Generated Text
von: Cheng, Yize, et al.
Veröffentlicht: (2025)
von: Cheng, Yize, et al.
Veröffentlicht: (2025)
RAFT: Realistic Attacks to Fool Text Detectors
von: Wang, James, et al.
Veröffentlicht: (2024)
von: Wang, James, et al.
Veröffentlicht: (2024)
Paraphrasing Attack Resilience of Various AI-Generated Text Detection Methods
von: Shportko, Andrii, et al.
Veröffentlicht: (2026)
von: Shportko, Andrii, et al.
Veröffentlicht: (2026)
Detector-Evasive LLM Paraphrasing via Constrained Policy Optimization
von: Wang, Mingyi, et al.
Veröffentlicht: (2026)
von: Wang, Mingyi, et al.
Veröffentlicht: (2026)
Your Language Model Can Secretly Write Like Humans: Contrastive Paraphrase Attacks on LLM-Generated Text Detectors
von: Fang, Hao, et al.
Veröffentlicht: (2025)
von: Fang, Hao, et al.
Veröffentlicht: (2025)
AdParaphrase: Paraphrase Dataset for Analyzing Linguistic Features toward Generating Attractive Ad Texts
von: Murakami, Soichiro, et al.
Veröffentlicht: (2025)
von: Murakami, Soichiro, et al.
Veröffentlicht: (2025)
Humanizing the Machine: Proxy Attacks to Mislead LLM Detectors
von: Wang, Tianchun, et al.
Veröffentlicht: (2024)
von: Wang, Tianchun, et al.
Veröffentlicht: (2024)
AdParaphrase v2.0: Generating Attractive Ad Texts Using a Preference-Annotated Paraphrase Dataset
von: Murakami, Soichiro, et al.
Veröffentlicht: (2025)
von: Murakami, Soichiro, et al.
Veröffentlicht: (2025)
ParaFusion: A Large-Scale LLM-Driven English Paraphrase Dataset Infused with High-Quality Lexical and Syntactic Diversity
von: Jayawardena, Lasal, et al.
Veröffentlicht: (2024)
von: Jayawardena, Lasal, et al.
Veröffentlicht: (2024)
Fine-tuning CLIP Text Encoders with Two-step Paraphrasing
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs
von: Guo, Hanxi, et al.
Veröffentlicht: (2025)
von: Guo, Hanxi, et al.
Veröffentlicht: (2025)
Zero2Text: Zero-Training Cross-Domain Inversion Attacks on Textual Embeddings
von: Kim, Doohyun, et al.
Veröffentlicht: (2026)
von: Kim, Doohyun, et al.
Veröffentlicht: (2026)
Task-Oriented Paraphrase Analytics
von: Gohsen, Marcel, et al.
Veröffentlicht: (2024)
von: Gohsen, Marcel, et al.
Veröffentlicht: (2024)
Linguistically-Controlled Paraphrase Generation
von: Elgaar, Mohamed, et al.
Veröffentlicht: (2024)
von: Elgaar, Mohamed, et al.
Veröffentlicht: (2024)
Paraphrase Types for Generation and Detection
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2023)
von: Wahle, Jan Philip, et al.
Veröffentlicht: (2023)
Improved Paraphrase Generation via Controllable Latent Diffusion
von: Zou, Wei, et al.
Veröffentlicht: (2024)
von: Zou, Wei, et al.
Veröffentlicht: (2024)
An Evaluation of Explanation Methods for Black-Box Detectors of Machine-Generated Text
von: Schoenegger, Loris, et al.
Veröffentlicht: (2024)
von: Schoenegger, Loris, et al.
Veröffentlicht: (2024)
GPT-DETOX: An In-Context Learning-Based Paraphraser for Text Detoxification
von: Pesaranghader, Ali, et al.
Veröffentlicht: (2024)
von: Pesaranghader, Ali, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Language Models Optimized to Fool Detectors Still Have a Distinct Style (And How to Change It)
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2025) -
Revisiting the Robustness of Watermarking to Paraphrasing Attacks
von: Rastogi, Saksham, et al.
Veröffentlicht: (2024) -
Few-Shot Detection of Machine-Generated Text using Style Representations
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2024) -
MahaParaphrase: A Marathi Paraphrase Detection Corpus and BERT-based Models
von: Jadhav, Suramya, et al.
Veröffentlicht: (2025) -
Paraphrasing Adversarial Attack on LLM-as-a-Reviewer
von: Kaneko, Masahiro
Veröffentlicht: (2026)