From Insight to Exploit: Leveraging LLM Collaboration for Adaptive Adversarial Text Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Sultana, Najrin, Rashid, Md Rafi Ur, Gu, Kang, Mehnaz, Shagufta |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Second-Order Information Matters: Revisiting Machine Unlearning for Large Language Models
by: Gu, Kang, et al.
Published: (2024)
by: Gu, Kang, et al.
Published: (2024)
Gradient-Free Privacy Leakage in Federated Language Models through Selective Weight Tampering
by: Rashid, Md Rafi Ur, et al.
Published: (2023)
by: Rashid, Md Rafi Ur, et al.
Published: (2023)
Chain-of-Thought Driven Adversarial Scenario Extrapolation for Robust Language Models
by: Rashid, Md Rafi Ur, et al.
Published: (2025)
by: Rashid, Md Rafi Ur, et al.
Published: (2025)
Forget to Flourish: Leveraging Machine-Unlearning on Pretrained Language Models for Privacy Leakage
by: Rashid, Md Rafi Ur, et al.
Published: (2024)
by: Rashid, Md Rafi Ur, et al.
Published: (2024)
Benchmarking Robust Aggregation in Decentralized Gradient Marketplaces
by: Song, Zeyu, et al.
Published: (2025)
by: Song, Zeyu, et al.
Published: (2025)
Robustness of Vision Language Models Against Split-Image Harmful Input Attacks
by: Rashid, Md Rafi Ur, et al.
Published: (2026)
by: Rashid, Md Rafi Ur, et al.
Published: (2026)
Disparate Privacy Vulnerability: Targeted Attribute Inference Attacks and Defenses
by: Kabir, Ehsanul, et al.
Published: (2025)
by: Kabir, Ehsanul, et al.
Published: (2025)
Unsupervised Text Embedding Space Generation Using Generative Adversarial Networks for Text Synthesis
by: Lee, Jun-Min, et al.
Published: (2023)
by: Lee, Jun-Min, et al.
Published: (2023)
A Critical Look At Tokenwise Reward-Guided Text Generation
by: Rashid, Ahmad, et al.
Published: (2024)
by: Rashid, Ahmad, et al.
Published: (2024)
Towards Cost-Effective Reward Guided Text Generation
by: Rashid, Ahmad, et al.
Published: (2025)
by: Rashid, Ahmad, et al.
Published: (2025)
GNNBleed: Inference Attacks to Unveil Private Edges in Graphs with Realistic Access to GNN Models
by: Song, Zeyu, et al.
Published: (2023)
by: Song, Zeyu, et al.
Published: (2023)
Adversarial Lens: Exploiting Attention Layers to Generate Adversarial Examples for Evaluation
by: Dhole, Kaustubh
Published: (2025)
by: Dhole, Kaustubh
Published: (2025)
Adaptive Testing for Segmenting Watermarked Texts From Language Models
by: Li, Xingchi, et al.
Published: (2025)
by: Li, Xingchi, et al.
Published: (2025)
SequentialBreak: Large Language Models Can be Fooled by Embedding Jailbreak Prompts into Sequential Prompt Chains
by: Saiem, Bijoy Ahmed, et al.
Published: (2024)
by: Saiem, Bijoy Ahmed, et al.
Published: (2024)
AdaDetectGPT: Adaptive Detection of LLM-Generated Text with Statistical Guarantees
by: Zhou, Hongyi, et al.
Published: (2025)
by: Zhou, Hongyi, et al.
Published: (2025)
Decoding-Time Debiasing via Process Reward Models: From Controlled Fill-in to Open-Ended Generation
by: Khan, Muneeb Ur Raheem
Published: (2026)
by: Khan, Muneeb Ur Raheem
Published: (2026)
$\textbf{AGT$^{AO}$}$: Robust and Stabilized LLM Unlearning via Adversarial Gating Training with Adaptive Orthogonality
by: Li, Pengyu, et al.
Published: (2026)
by: Li, Pengyu, et al.
Published: (2026)
On the Detectability of LLM-Generated Text: What Exactly Is LLM-Generated Text?
by: Geng, Mingmeng, et al.
Published: (2025)
by: Geng, Mingmeng, et al.
Published: (2025)
Rep3Net: An Approach Exploiting Multimodal Representation for Molecular Bioactivity Prediction
by: Islam, Sabrina, et al.
Published: (2025)
by: Islam, Sabrina, et al.
Published: (2025)
Generative Adversarial Reasoner: Enhancing LLM Reasoning with Adversarial Reinforcement Learning
by: Liu, Qihao, et al.
Published: (2025)
by: Liu, Qihao, et al.
Published: (2025)
A Regularized LSTM Method for Detecting Fake News Articles
by: Camelia, Tanjina Sultana, et al.
Published: (2024)
by: Camelia, Tanjina Sultana, et al.
Published: (2024)
Detecting LLM-Generated Text with Performance Guarantees
by: Zhou, Hongyi, et al.
Published: (2026)
by: Zhou, Hongyi, et al.
Published: (2026)
Group-Adaptive Threshold Optimization for Robust AI-Generated Text Detection
by: Jung, Minseok, et al.
Published: (2025)
by: Jung, Minseok, et al.
Published: (2025)
Reducing and Exploiting Data Augmentation Noise through Meta Reweighting Contrastive Learning for Text Classification
by: Mou, Guanyi, et al.
Published: (2024)
by: Mou, Guanyi, et al.
Published: (2024)
Stochastic Adversarial Networks for Multi-Domain Text Classification
by: Wang, Xu, et al.
Published: (2024)
by: Wang, Xu, et al.
Published: (2024)
MedSyn: LLM-based Synthetic Medical Text Generation Framework
by: Kumichev, Gleb, et al.
Published: (2024)
by: Kumichev, Gleb, et al.
Published: (2024)
From Text to Talent: A Pipeline for Extracting Insights from Candidate Profiles
by: Frazzetto, Paolo, et al.
Published: (2025)
by: Frazzetto, Paolo, et al.
Published: (2025)
Smoothed Embeddings for Robust Language Models
by: Hase, Ryo, et al.
Published: (2025)
by: Hase, Ryo, et al.
Published: (2025)
Usable XAI: 10 Strategies Towards Exploiting Explainability in the LLM Era
by: Wu, Xuansheng, et al.
Published: (2024)
by: Wu, Xuansheng, et al.
Published: (2024)
From Text to Graph: Leveraging Graph Neural Networks for Enhanced Explainability in NLP
by: Yáñez-Romero, Fabio, et al.
Published: (2025)
by: Yáñez-Romero, Fabio, et al.
Published: (2025)
Evaluating Text Classification Robustness to Part-of-Speech Adversarial Examples
by: Samadi, Anahita, et al.
Published: (2024)
by: Samadi, Anahita, et al.
Published: (2024)
Not All LLM-Generated Data Are Equal: Rethinking Data Weighting in Text Classification
by: Kuo, Hsun-Yu, et al.
Published: (2024)
by: Kuo, Hsun-Yu, et al.
Published: (2024)
From Selection to Generation: A Survey of LLM-based Active Learning
by: Xia, Yu, et al.
Published: (2025)
by: Xia, Yu, et al.
Published: (2025)
Rethinking LLM Memorization through the Lens of Adversarial Compression
by: Schwarzschild, Avi, et al.
Published: (2024)
by: Schwarzschild, Avi, et al.
Published: (2024)
Margin Discrepancy-based Adversarial Training for Multi-Domain Text Classification
by: Wu, Yuan
Published: (2024)
by: Wu, Yuan
Published: (2024)
Model-Agnostic Sentiment Distribution Stability Analysis for Robust LLM-Generated Texts Detection
by: Li, Siyuan, et al.
Published: (2025)
by: Li, Siyuan, et al.
Published: (2025)
Knowledge Distillation in Automated Annotation: Supervised Text Classification with LLM-Generated Training Labels
by: Pangakis, Nicholas, et al.
Published: (2024)
by: Pangakis, Nicholas, et al.
Published: (2024)
Llamazip: Leveraging LLaMA for Lossless Text Compression and Training Dataset Detection
by: Dréano, Sören, et al.
Published: (2025)
by: Dréano, Sören, et al.
Published: (2025)
Amplification Effects in Test-Time Reinforcement Learning: Safety and Reasoning Vulnerabilities
by: Khattar, Vanshaj, et al.
Published: (2026)
by: Khattar, Vanshaj, et al.
Published: (2026)
Self-playing Adversarial Language Game Enhances LLM Reasoning
by: Cheng, Pengyu, et al.
Published: (2024)
by: Cheng, Pengyu, et al.
Published: (2024)
Similar Items
-
Second-Order Information Matters: Revisiting Machine Unlearning for Large Language Models
by: Gu, Kang, et al.
Published: (2024) -
Gradient-Free Privacy Leakage in Federated Language Models through Selective Weight Tampering
by: Rashid, Md Rafi Ur, et al.
Published: (2023) -
Chain-of-Thought Driven Adversarial Scenario Extrapolation for Robust Language Models
by: Rashid, Md Rafi Ur, et al.
Published: (2025) -
Forget to Flourish: Leveraging Machine-Unlearning on Pretrained Language Models for Privacy Leakage
by: Rashid, Md Rafi Ur, et al.
Published: (2024) -
Benchmarking Robust Aggregation in Decentralized Gradient Marketplaces
by: Song, Zeyu, et al.
Published: (2025)