Stress-testing Machine Generated Text Detection: Shifting Language Models Writing Style to Fool Detectors
Fuente:
arXiv
Saved in:
| Main Authors: | Pedrotti, Andrea, Papucci, Michele, Ciaccio, Cristiano, Miaschi, Alessio, Puccetti, Giovanni, Dell'Orletta, Felice, Esuli, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI "News" Content Farms Are Easy to Make and Hard to Detect: A Case Study in Italian
by: Puccetti, Giovanni, et al.
Published: (2024)
by: Puccetti, Giovanni, et al.
Published: (2024)
Linguistic Knowledge Can Enhance Encoder-Decoder Models (If You Let It)
by: Miaschi, Alessio, et al.
Published: (2024)
by: Miaschi, Alessio, et al.
Published: (2024)
Linguistic Profiling of a Neural Language Model
by: Miaschi, Alessio, et al.
Published: (2020)
by: Miaschi, Alessio, et al.
Published: (2020)
Linguistically-driven Selection of Correct Arcs for Dependency Parsing
by: Felice Dell’Orletta
Published: (2013)
by: Felice Dell’Orletta
Published: (2013)
Optimizing LLMs for Italian: Reducing Token Fertility and Enhancing Efficiency Through Vocabulary Adaptation
by: Moroni, Luca, et al.
Published: (2025)
by: Moroni, Luca, et al.
Published: (2025)
Outliers Dimensions that Disrupt Transformers Are Driven by Frequency
by: Puccetti, Giovanni, et al.
Published: (2022)
by: Puccetti, Giovanni, et al.
Published: (2022)
T-FREX: A Transformer-based Feature Extraction Method from Mobile App Reviews
by: Motger, Quim, et al.
Published: (2024)
by: Motger, Quim, et al.
Published: (2024)
Leveraging Encoder-only Large Language Models for Mobile App Review Feature Extraction
by: Motger, Quim, et al.
Published: (2024)
by: Motger, Quim, et al.
Published: (2024)
Contextualized Counterspeech: Strategies for Adaptation, Personalization, and Evaluation
by: Cima, Lorenzo, et al.
Published: (2024)
by: Cima, Lorenzo, et al.
Published: (2024)
The Invalsi Benchmarks: measuring Linguistic and Mathematical understanding of Large Language Models in Italian
by: Puccetti, Giovanni, et al.
Published: (2024)
by: Puccetti, Giovanni, et al.
Published: (2024)
Fine-tuning with HED-IT: The impact of human post-editing for dialogical language models
by: Occhipinti, Daniela, et al.
Published: (2024)
by: Occhipinti, Daniela, et al.
Published: (2024)
Detection Latencies of Anomaly Detectors: An Overlooked Perspective ?
by: Puccetti, Tommaso, et al.
Published: (2024)
by: Puccetti, Tommaso, et al.
Published: (2024)
A logging and reporting tool for NVIDIA GPUs
by: Andrea Esuli
Published: (2025)
by: Andrea Esuli
Published: (2025)
Distributional Difference-in-Differences Models with Multiple Time Periods
by: Ciaccio, Andrea
Published: (2024)
by: Ciaccio, Andrea
Published: (2024)
Protein Secondary Structure Patterns in Short‐Range Cross‐Link Atlas
by: Alice Vetrano, et al.
Published: (2025)
by: Alice Vetrano, et al.
Published: (2025)
Protein Secondary Structure Patterns in Short‐Range Cross‐Link Atlas
by: Alice Vetrano, et al.
Published: (2025)
by: Alice Vetrano, et al.
Published: (2025)
Off-shell vertices in heavy particle effective theories and $B\rightarrow Dπ\ell ν$
by: Papucci, Michele, et al.
Published: (2024)
by: Papucci, Michele, et al.
Published: (2024)
Radiative Semileptonic Decays of Beautiful Hadrons
by: Cima, Federico, et al.
Published: (2025)
by: Cima, Federico, et al.
Published: (2025)
RAFT: Realistic Attacks to Fool Text Detectors
by: Wang, James, et al.
Published: (2024)
by: Wang, James, et al.
Published: (2024)
StyleFool: Fooling Video Classification Systems via Style Transfer
by: Cao, Yuxin, et al.
Published: (2022)
by: Cao, Yuxin, et al.
Published: (2022)
Prednisona, azatioprina, y N-acetilcisteína en fibrosis pulmonar idiopática (PANTHER - FPI)
by: Tulio Papucci
Published: (2012)
by: Tulio Papucci
Published: (2012)
Language Models Optimized to Fool Detectors Still Have a Distinct Style (And How to Change It)
by: Soto, Rafael Rivera, et al.
Published: (2025)
by: Soto, Rafael Rivera, et al.
Published: (2025)
Environmental Policy and Firm Performance in Europe: A Difference-in-Differences Approach with Spillovers
by: Ciaccio, Andrea, et al.
Published: (2025)
by: Ciaccio, Andrea, et al.
Published: (2025)
Learning to Quantify
by: Esuli, Andrea, et al.
Published: (2023)
by: Esuli, Andrea, et al.
Published: (2023)
REACHING OUT: EXTENDING COLLABORATION & TRAINING TO PARAEDUCATORS
by: Orletta Nguyen
Published: (2015)
by: Orletta Nguyen
Published: (2015)
On the Generalization and Adaptation Ability of Machine-Generated Text Detectors in Academic Writing
by: Liu, Yule, et al.
Published: (2024)
by: Liu, Yule, et al.
Published: (2024)
Breaking the Imitation Game: Can LLMs Fool Humans and Machines Alike?
by: Ubaid Ullah, et al.
Published: (2026)
by: Ubaid Ullah, et al.
Published: (2026)
Heat flow, log-concavity, and Lipschitz transport maps
by: Brigati, Giovanni, et al.
Published: (2024)
by: Brigati, Giovanni, et al.
Published: (2024)
LogoStyleFool: Vitiating Video Recognition Systems via Logo Style Transfer
by: Cao, Yuxin, et al.
Published: (2023)
by: Cao, Yuxin, et al.
Published: (2023)
TACO: Adversarial Camouflage Optimization on Trucks to Fool Object Detectors
by: Dimitriu, Adonisz, et al.
Published: (2024)
by: Dimitriu, Adonisz, et al.
Published: (2024)
Learning to Generate Text in Arbitrary Writing Styles
by: Khan, Aleem, et al.
Published: (2023)
by: Khan, Aleem, et al.
Published: (2023)
Extremal Dependence Concepts
by: Puccetti, Giovanni, et al.
Published: (2015)
by: Puccetti, Giovanni, et al.
Published: (2015)
LocalStyleFool: Regional Video Style Transfer Attack Using Segment Anything Model
by: Cao, Yuxin, et al.
Published: (2024)
by: Cao, Yuxin, et al.
Published: (2024)
Fooling the Watchers: Breaking AIGC Detectors via Semantic Prompt Attacks
by: Hao, Run, et al.
Published: (2025)
by: Hao, Run, et al.
Published: (2025)
Fool the Stoplight: Realistic Adversarial Patch Attacks on Traffic Light Detectors
by: Pavlitska, Svetlana, et al.
Published: (2025)
by: Pavlitska, Svetlana, et al.
Published: (2025)
Interpretable Text Classification Applied to the Detection of LLM-generated Creative Writing
by: Suvanto, Minerva, et al.
Published: (2026)
by: Suvanto, Minerva, et al.
Published: (2026)
Stumbling Blocks: Stress Testing the Robustness of Machine-Generated Text Detectors Under Attacks
by: Wang, Yichen, et al.
Published: (2024)
by: Wang, Yichen, et al.
Published: (2024)
Insights on the Gamma-Ray Bursts variability in their cosmological rest frame
by: Della Casa, Giovanni, et al.
Published: (2026)
by: Della Casa, Giovanni, et al.
Published: (2026)
All-in-one: Understanding and Generation in Multimodal Reasoning with the MAIA Benchmark
by: Testa, Davide, et al.
Published: (2025)
by: Testa, Davide, et al.
Published: (2025)
ROSpace: Intrusion Detection Dataset for a ROS2-Based Cyber-Physical System
by: Puccetti, Tommaso, et al.
Published: (2024)
by: Puccetti, Tommaso, et al.
Published: (2024)
Similar Items
-
AI "News" Content Farms Are Easy to Make and Hard to Detect: A Case Study in Italian
by: Puccetti, Giovanni, et al.
Published: (2024) -
Linguistic Knowledge Can Enhance Encoder-Decoder Models (If You Let It)
by: Miaschi, Alessio, et al.
Published: (2024) -
Linguistic Profiling of a Neural Language Model
by: Miaschi, Alessio, et al.
Published: (2020) -
Linguistically-driven Selection of Correct Arcs for Dependency Parsing
by: Felice Dell’Orletta
Published: (2013) -
Optimizing LLMs for Italian: Reducing Token Fertility and Enhancing Efficiency Through Vocabulary Adaptation
by: Moroni, Luca, et al.
Published: (2025)