Token-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual Information
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Zhengmian, Wu, Gang, Mitra, Saayan, Zhang, Ruiyi, Sun, Tong, Huang, Heng, Swaminathan, Viswanathan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Bayesian Approach to Harnessing the Power of LLMs in Authorship Attribution
von: Hu, Zhengmian, et al.
Veröffentlicht: (2024)
von: Hu, Zhengmian, et al.
Veröffentlicht: (2024)
A Penalty Goes a Long Way: Measuring Lexical Diversity in Synthetic Texts Under Prompt-Influenced Length Variations
von: Deshpande, Vijeta, et al.
Veröffentlicht: (2025)
von: Deshpande, Vijeta, et al.
Veröffentlicht: (2025)
Towards Optimal Multi-draft Speculative Decoding
von: Hu, Zhengmian, et al.
Veröffentlicht: (2025)
von: Hu, Zhengmian, et al.
Veröffentlicht: (2025)
A Resilient and Accessible Distribution-Preserving Watermark for Large Language Models
von: Wu, Yihan, et al.
Veröffentlicht: (2023)
von: Wu, Yihan, et al.
Veröffentlicht: (2023)
Multi-Level Contextual Token Relation Modeling for Machine-Generated Text Detection
von: Wu, Chenwang, et al.
Veröffentlicht: (2026)
von: Wu, Chenwang, et al.
Veröffentlicht: (2026)
PLPP: Prompt Learning with Perplexity Is Self-Distillation for Vision-Language Models
von: Liu, Biao, et al.
Veröffentlicht: (2024)
von: Liu, Biao, et al.
Veröffentlicht: (2024)
Perplexed by Perplexity: Perplexity-Based Data Pruning With Small Reference Models
von: Ankner, Zachary, et al.
Veröffentlicht: (2024)
von: Ankner, Zachary, et al.
Veröffentlicht: (2024)
Mitigating Forgetting in LLM Fine-Tuning via Low-Perplexity Token Learning
von: Wu, Chao-Chung, et al.
Veröffentlicht: (2025)
von: Wu, Chao-Chung, et al.
Veröffentlicht: (2025)
Perplexity Trap: PLM-Based Retrievers Overrate Low Perplexity Documents
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
Is my model perplexed for the right reason? Contrasting LLMs' Benchmark Behavior with Token-Level Perplexity
von: Prins, Zoë, et al.
Veröffentlicht: (2026)
von: Prins, Zoë, et al.
Veröffentlicht: (2026)
Addressing Tokenization Inconsistency in Steganography and Watermarking Based on Large Language Models
von: Yan, Ruiyi, et al.
Veröffentlicht: (2025)
von: Yan, Ruiyi, et al.
Veröffentlicht: (2025)
Demystifying Prompts in Language Models via Perplexity Estimation
von: Gonen, Hila, et al.
Veröffentlicht: (2022)
von: Gonen, Hila, et al.
Veröffentlicht: (2022)
GraphicBench: A Planning Benchmark for Graphic Design with Language Agents
von: Ki, Dayeon, et al.
Veröffentlicht: (2025)
von: Ki, Dayeon, et al.
Veröffentlicht: (2025)
CodeLutra: Boosting LLM Code Generation via Preference-Guided Refinement
von: Tao, Leitian, et al.
Veröffentlicht: (2024)
von: Tao, Leitian, et al.
Veröffentlicht: (2024)
Mitigating Forgetting Between Supervised and Reinforcement Learning Yields Stronger Reasoners
von: Yuan, Xiangchi, et al.
Veröffentlicht: (2025)
von: Yuan, Xiangchi, et al.
Veröffentlicht: (2025)
Inevitable Trade-off between Watermark Strength and Speculative Sampling Efficiency for Language Models
von: Hu, Zhengmian, et al.
Veröffentlicht: (2024)
von: Hu, Zhengmian, et al.
Veröffentlicht: (2024)
On the Fallacy of Global Token Perplexity in Spoken Language Model Evaluation
von: Hsu, Chan-Jan, et al.
Veröffentlicht: (2026)
von: Hsu, Chan-Jan, et al.
Veröffentlicht: (2026)
An Interactive Paradigm for Deep Research
von: Ai, Lin, et al.
Veröffentlicht: (2026)
von: Ai, Lin, et al.
Veröffentlicht: (2026)
ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
von: Xu, Xin, et al.
Veröffentlicht: (2026)
von: Xu, Xin, et al.
Veröffentlicht: (2026)
Inconsistent Tokenizations Cause Language Models to be Perplexed by Japanese Grammar
von: Gambardella, Andrew, et al.
Veröffentlicht: (2025)
von: Gambardella, Andrew, et al.
Veröffentlicht: (2025)
PromptFix: Few-shot Backdoor Removal via Adversarial Prompt Tuning
von: Zhang, Tianrong, et al.
Veröffentlicht: (2024)
von: Zhang, Tianrong, et al.
Veröffentlicht: (2024)
Few-Shot Dialogue Summarization via Skeleton-Assisted Prompt Transfer in Prompt Tuning
von: Xie, Kaige, et al.
Veröffentlicht: (2023)
von: Xie, Kaige, et al.
Veröffentlicht: (2023)
Can Perplexity Reflect Large Language Model's Ability in Long Text Understanding?
von: Hu, Yutong, et al.
Veröffentlicht: (2024)
von: Hu, Yutong, et al.
Veröffentlicht: (2024)
FrugalPrompt: Reducing Contextual Overhead in Large Language Models via Token Attribution
von: Raiyan, Syed Rifat, et al.
Veröffentlicht: (2025)
von: Raiyan, Syed Rifat, et al.
Veröffentlicht: (2025)
Rethinking Perplexity: Revealing the Impact of Input Length on Perplexity Evaluation in LLMs
von: Cheng, Letian, et al.
Veröffentlicht: (2026)
von: Cheng, Letian, et al.
Veröffentlicht: (2026)
Beyond Natural Language Perplexity: Detecting Dead Code Poisoning in Code Generation Datasets
von: Tsai, Chi-Chien, et al.
Veröffentlicht: (2025)
von: Tsai, Chi-Chien, et al.
Veröffentlicht: (2025)
RePPL: Recalibrating Perplexity by Uncertainty in Semantic Propagation and Language Generation for Explainable QA Hallucination Detection
von: Huang, Yiming, et al.
Veröffentlicht: (2025)
von: Huang, Yiming, et al.
Veröffentlicht: (2025)
Cautious Next Token Prediction
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
von: Wang, Yizhou, et al.
Veröffentlicht: (2025)
ASTPrompter: Preference-Aligned Automated Language Model Red-Teaming to Generate Low-Perplexity Unsafe Prompts
von: Hardy, Amelia F., et al.
Veröffentlicht: (2024)
von: Hardy, Amelia F., et al.
Veröffentlicht: (2024)
The Perplexity Paradox: Why Code Compresses Better Than Math in LLM Prompts
von: Johnson, Warren
Veröffentlicht: (2026)
von: Johnson, Warren
Veröffentlicht: (2026)
Mapping Overlaps in Benchmarks through Perplexity in the Wild
von: Wu, Siyang, et al.
Veröffentlicht: (2025)
von: Wu, Siyang, et al.
Veröffentlicht: (2025)
Measuring Contextual Informativeness in Child-Directed Text
von: Valentini, Maria, et al.
Veröffentlicht: (2024)
von: Valentini, Maria, et al.
Veröffentlicht: (2024)
Can Perplexity Predict Fine-tuning Performance? An Investigation of Tokenization Effects on Sequential Language Models for Nepali
von: Luitel, Nishant, et al.
Veröffentlicht: (2024)
von: Luitel, Nishant, et al.
Veröffentlicht: (2024)
Cmprsr: Abstractive Token-Level Question-Agnostic Prompt Compressor
von: Zakazov, Ivan, et al.
Veröffentlicht: (2025)
von: Zakazov, Ivan, et al.
Veröffentlicht: (2025)
Beyond Perplexity: Character Distribution Signatures and the MDTA Benchmark for AI Text Detection
von: Narayanasamy, Priyadarshan, et al.
Veröffentlicht: (2026)
von: Narayanasamy, Priyadarshan, et al.
Veröffentlicht: (2026)
Perplexity-Aware Data Scaling Law: Perplexity Landscapes Predict Performance for Continual Pre-training
von: Liu, Lei, et al.
Veröffentlicht: (2025)
von: Liu, Lei, et al.
Veröffentlicht: (2025)
Show Me What I Like: Detecting User-Specific Video Highlights Using Content-Based Multi-Head Attention
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2022)
von: Bhattacharya, Uttaran, et al.
Veröffentlicht: (2022)
Detecting Ambiguities to Guide Query Rewrite for Robust Conversations in Enterprise AI Assistants
von: Tanjim, Md Mehrab, et al.
Veröffentlicht: (2025)
von: Tanjim, Md Mehrab, et al.
Veröffentlicht: (2025)
SePer: Measure Retrieval Utility Through The Lens Of Semantic Perplexity Reduction
von: Dai, Lu, et al.
Veröffentlicht: (2025)
von: Dai, Lu, et al.
Veröffentlicht: (2025)
Probing Geometry of Next Token Prediction Using Cumulant Expansion of the Softmax Entropy
von: Viswanathan, Karthik, et al.
Veröffentlicht: (2025)
von: Viswanathan, Karthik, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Bayesian Approach to Harnessing the Power of LLMs in Authorship Attribution
von: Hu, Zhengmian, et al.
Veröffentlicht: (2024) -
A Penalty Goes a Long Way: Measuring Lexical Diversity in Synthetic Texts Under Prompt-Influenced Length Variations
von: Deshpande, Vijeta, et al.
Veröffentlicht: (2025) -
Towards Optimal Multi-draft Speculative Decoding
von: Hu, Zhengmian, et al.
Veröffentlicht: (2025) -
A Resilient and Accessible Distribution-Preserving Watermark for Large Language Models
von: Wu, Yihan, et al.
Veröffentlicht: (2023) -
Multi-Level Contextual Token Relation Modeling for Machine-Generated Text Detection
von: Wu, Chenwang, et al.
Veröffentlicht: (2026)