Opening A Pandora's Box: Things You Should Know in the Era of Custom GPTs
Fuente:
arXiv
Saved in:
| Main Authors: | Tao, Guanhong, Cheng, Siyuan, Zhang, Zhuo, Zhu, Junmin, Shen, Guangyu, Zhang, Xiangyu |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning
by: Deng, Gelei, et al.
Published: (2024)
by: Deng, Gelei, et al.
Published: (2024)
Rapid Optimization for Jailbreaking LLMs via Subconscious Exploitation and Echopraxia
by: Shen, Guangyu, et al.
Published: (2024)
by: Shen, Guangyu, et al.
Published: (2024)
UNIT: Backdoor Mitigation via Automated Neural Distribution Tightening
by: Cheng, Siyuan, et al.
Published: (2024)
by: Cheng, Siyuan, et al.
Published: (2024)
ASPIRER: Bypassing System Prompts With Permutation-based Backdoors in LLMs
by: Yan, Lu, et al.
Published: (2024)
by: Yan, Lu, et al.
Published: (2024)
From Poisoned to Aware: Fostering Backdoor Self-Awareness in LLMs
by: Shen, Guangyu, et al.
Published: (2025)
by: Shen, Guangyu, et al.
Published: (2025)
LOTUS: Evasive and Resilient Backdoor Attacks through Sub-Partitioning
by: Cheng, Siyuan, et al.
Published: (2024)
by: Cheng, Siyuan, et al.
Published: (2024)
GPT in Sheep's Clothing: The Risk of Customized GPTs
by: Antebi, Sagiv, et al.
Published: (2024)
by: Antebi, Sagiv, et al.
Published: (2024)
Assessing Prompt Injection Risks in 200+ Custom GPTs
by: Yu, Jiahao, et al.
Published: (2023)
by: Yu, Jiahao, et al.
Published: (2023)
Elijah: Eliminating Backdoors Injected in Diffusion Models via Distribution Shift
by: An, Shengwei, et al.
Published: (2023)
by: An, Shengwei, et al.
Published: (2023)
When GPT Spills the Tea: Comprehensive Assessment of Knowledge File Leakage in GPTs
by: Shen, Xinyue, et al.
Published: (2025)
by: Shen, Xinyue, et al.
Published: (2025)
CENSOR: Defense Against Gradient Inversion via Orthogonal Subspace Bayesian Sampling
by: Zhang, Kaiyuan, et al.
Published: (2025)
by: Zhang, Kaiyuan, et al.
Published: (2025)
I Know What You Sync: Covert and Side Channel Attacks on File Systems via syncfs
by: Gu, Cheng, et al.
Published: (2024)
by: Gu, Cheng, et al.
Published: (2024)
MGC: A Compiler Framework Exploiting Compositional Blindness in Aligned LLMs for Malware Generation
by: Yan, Lu, et al.
Published: (2025)
by: Yan, Lu, et al.
Published: (2025)
Everything You Wanted to Know About LLM-based Vulnerability Detection But Were Afraid to Ask
by: Li, Yue, et al.
Published: (2025)
by: Li, Yue, et al.
Published: (2025)
Temporal Logic-Based Multi-Vehicle Backdoor Attacks against Offline RL Agents in End-to-end Autonomous Driving
by: Chen, Xuan, et al.
Published: (2025)
by: Chen, Xuan, et al.
Published: (2025)
An Empirical Study on the Security Vulnerabilities of GPTs
by: Wu, Tong, et al.
Published: (2025)
by: Wu, Tong, et al.
Published: (2025)
Alleviating the Fear of Losing Alignment in LLM Fine-tuning
by: Yang, Kang, et al.
Published: (2025)
by: Yang, Kang, et al.
Published: (2025)
Gradient Shaping: Enhancing Backdoor Attack Against Reverse Engineering
by: Zhu, Rui, et al.
Published: (2023)
by: Zhu, Rui, et al.
Published: (2023)
Privacy and Security Threat for OpenAI GPTs
by: Wenying, Wei, et al.
Published: (2025)
by: Wenying, Wei, et al.
Published: (2025)
A Large-Scale Empirical Analysis of Custom GPTs' Vulnerabilities in the OpenAI Ecosystem
by: Ogundoyin, Sunday Oyinlola, et al.
Published: (2025)
by: Ogundoyin, Sunday Oyinlola, et al.
Published: (2025)
LLM Agents Should Employ Security Principles
by: Zhang, Kaiyuan, et al.
Published: (2025)
by: Zhang, Kaiyuan, et al.
Published: (2025)
A Sentence Relation-Based Approach to Sanitizing Malicious Instructions
by: Datta, Soumil, et al.
Published: (2026)
by: Datta, Soumil, et al.
Published: (2026)
Less Is More -- Until It Breaks: Security Pitfalls of Vision Token Compression in Large Vision-Language Models
by: Zhang, Xiaomei, et al.
Published: (2026)
by: Zhang, Xiaomei, et al.
Published: (2026)
Fusion is Not Enough: Single Modal Attacks on Fusion Models for 3D Object Detection
by: Cheng, Zhiyuan, et al.
Published: (2023)
by: Cheng, Zhiyuan, et al.
Published: (2023)
Dataset Ownership in the Era of Large Language Models
by: Li, Kun, et al.
Published: (2025)
by: Li, Kun, et al.
Published: (2025)
Tracking GPTs Third Party Service: Automation, Analysis, and Insights
by: Yan, Chuan, et al.
Published: (2025)
by: Yan, Chuan, et al.
Published: (2025)
I Know What You Said: Unveiling Hardware Cache Side-Channels in Local Large Language Model Inference
by: Gao, Zibo, et al.
Published: (2025)
by: Gao, Zibo, et al.
Published: (2025)
Fast Revocable Attribute-Based Encryption with Data Integrity for Internet of Things
by: Li, Yongjiao, et al.
Published: (2025)
by: Li, Yongjiao, et al.
Published: (2025)
Pandora's White-Box: Precise Training Data Detection and Extraction in Large Language Models
by: Wang, Jeffrey G., et al.
Published: (2024)
by: Wang, Jeffrey G., et al.
Published: (2024)
Post-Quantum Cryptography for Internet of Things: A Survey on Performance and Optimization
by: Liu, Tao, et al.
Published: (2024)
by: Liu, Tao, et al.
Published: (2024)
Blackbox Dataset Inference for LLM
by: Zhou, Ruikai, et al.
Published: (2025)
by: Zhou, Ruikai, et al.
Published: (2025)
HarnessAgent: Scaling Automatic Fuzzing Harness Construction with Tool-Augmented LLM Pipelines
by: Yang, Kang, et al.
Published: (2025)
by: Yang, Kang, et al.
Published: (2025)
BDFirewall: Towards Effective and Expeditiously Black-Box Backdoor Defense in MLaaS
by: Li, Ye, et al.
Published: (2025)
by: Li, Ye, et al.
Published: (2025)
How Vulnerable Is My Learned Policy? Universal Adversarial Perturbation Attacks On Modern Behavior Cloning Policies
by: Kalra, Akansha, et al.
Published: (2025)
by: Kalra, Akansha, et al.
Published: (2025)
ASTRA: Autonomous Spatial-Temporal Red-teaming for AI Software Assistants
by: Xu, Xiangzhe, et al.
Published: (2025)
by: Xu, Xiangzhe, et al.
Published: (2025)
Rethinking the Evaluation of Secure Code Generation
by: Dai, Shih-Chieh, et al.
Published: (2025)
by: Dai, Shih-Chieh, et al.
Published: (2025)
I Know What You Did Last Summer: Identifying VR User Activity Through VR Network Traffic
by: Muhaimin, Sheikh Samit, et al.
Published: (2025)
by: Muhaimin, Sheikh Samit, et al.
Published: (2025)
Instruction Backdoor Attacks Against Customized LLMs
by: Zhang, Rui, et al.
Published: (2024)
by: Zhang, Rui, et al.
Published: (2024)
AI Safeguards, Generative AI and the Pandora Box: AI Safety Measures to Protect Businesses and Personal Reputation
by: Kumar, Prasanna
Published: (2026)
by: Kumar, Prasanna
Published: (2026)
Pandora's Box in Your SSD: The Untold Dangers of NVMe
by: Wertenbroek, Rick, et al.
Published: (2024)
by: Wertenbroek, Rick, et al.
Published: (2024)
Similar Items
-
Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning
by: Deng, Gelei, et al.
Published: (2024) -
Rapid Optimization for Jailbreaking LLMs via Subconscious Exploitation and Echopraxia
by: Shen, Guangyu, et al.
Published: (2024) -
UNIT: Backdoor Mitigation via Automated Neural Distribution Tightening
by: Cheng, Siyuan, et al.
Published: (2024) -
ASPIRER: Bypassing System Prompts With Permutation-based Backdoors in LLMs
by: Yan, Lu, et al.
Published: (2024) -
From Poisoned to Aware: Fostering Backdoor Self-Awareness in LLMs
by: Shen, Guangyu, et al.
Published: (2025)