Evaluation empirique de la sécurisation et de l'alignement de ChatGPT et Gemini: analyse comparative des vulnérabilités par expérimentations de jailbreaks
Fuente:
arXiv
Saved in:
| Main Author: | Nouailles, Rafaël |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Comparative Analysis Based on DeepSeek, ChatGPT, and Google Gemini: Features, Techniques, Performance, Future Prospects
by: Rahman, Anichur, et al.
Published: (2025)
by: Rahman, Anichur, et al.
Published: (2025)
AuditGPT: Auditing Smart Contracts with ChatGPT
by: Xia, Shihao, et al.
Published: (2024)
by: Xia, Shihao, et al.
Published: (2024)
Exploring ChatGPT's Capabilities on Vulnerability Management
by: Liu, Peiyu, et al.
Published: (2023)
by: Liu, Peiyu, et al.
Published: (2023)
Exfiltration of personal information from ChatGPT via prompt injection
by: Schwartzman, Gregory
Published: (2024)
by: Schwartzman, Gregory
Published: (2024)
LLM Platform Security: Applying a Systematic Evaluation Framework to OpenAI's ChatGPT Plugins
by: Iqbal, Umar, et al.
Published: (2023)
by: Iqbal, Umar, et al.
Published: (2023)
On the Detectability of ChatGPT Content: Benchmarking, Methodology, and Evaluation through the Lens of Academic Writing
by: Liu, Zeyan, et al.
Published: (2023)
by: Liu, Zeyan, et al.
Published: (2023)
From Chatbots to PhishBots? -- Preventing Phishing scams created using ChatGPT, Google Bard and Claude
by: Roy, Sayak Saha, et al.
Published: (2023)
by: Roy, Sayak Saha, et al.
Published: (2023)
ChatGPT's Potential in Cryptography Misuse Detection: A Comparative Analysis with Static Analysis Tools
by: Firouzi, Ehsan, et al.
Published: (2024)
by: Firouzi, Ehsan, et al.
Published: (2024)
Red-Teaming Claude Opus and ChatGPT-based Security Advisors for Trusted Execution Environments
by: Mukherjee, Kunal, et al.
Published: (2026)
by: Mukherjee, Kunal, et al.
Published: (2026)
Detecting Phishing Sites Using ChatGPT
by: Koide, Takashi, et al.
Published: (2023)
by: Koide, Takashi, et al.
Published: (2023)
How Secure is Code Generated by ChatGPT?
by: Khoury, Raphaël, et al.
Published: (2023)
by: Khoury, Raphaël, et al.
Published: (2023)
Latent-space adversarial training with post-aware calibration for defending large language models against jailbreak attacks
by: Yi, Xin, et al.
Published: (2025)
by: Yi, Xin, et al.
Published: (2025)
Can ChatGPT Detect DeepFakes? A Study of Using Multimodal Large Language Models for Media Forensics
by: Jia, Shan, et al.
Published: (2024)
by: Jia, Shan, et al.
Published: (2024)
Security Analysis of ChatGPT: Threats and Privacy Risks
by: Xiang, Yushan, et al.
Published: (2025)
by: Xiang, Yushan, et al.
Published: (2025)
Digital Forensic Investigation of the ChatGPT Windows Application
by: Kankanamge, Malithi Wanniarachchi, et al.
Published: (2025)
by: Kankanamge, Malithi Wanniarachchi, et al.
Published: (2025)
Recursive language models for jailbreak detection: a procedural defense for tool-augmented agents
by: Shavit, Doron
Published: (2026)
by: Shavit, Doron
Published: (2026)
A Qualitative Study on Using ChatGPT for Software Security: Perception vs. Practicality
by: Kholoosi, M. Mehdi, et al.
Published: (2024)
by: Kholoosi, M. Mehdi, et al.
Published: (2024)
EaTVul: ChatGPT-based Evasion Attack Against Software Vulnerability Detection
by: Liu, Shigang, et al.
Published: (2024)
by: Liu, Shigang, et al.
Published: (2024)
Evaluation of ChatGPT's Smart Contract Auditing Capabilities Based on Chain of Thought
by: Du, Yuying, et al.
Published: (2024)
by: Du, Yuying, et al.
Published: (2024)
Are aligned neural networks adversarially aligned?
by: Carlini, Nicholas, et al.
Published: (2023)
by: Carlini, Nicholas, et al.
Published: (2023)
Exploring Backdoor Vulnerabilities of Chat Models
by: Hao, Yunzhuo, et al.
Published: (2024)
by: Hao, Yunzhuo, et al.
Published: (2024)
Time to Separate from StackOverflow and Match with ChatGPT for Encryption
by: Firouzi, Ehsan, et al.
Published: (2024)
by: Firouzi, Ehsan, et al.
Published: (2024)
Can ChatGPT Perform Image Splicing Detection? A Preliminary Study
by: Nath, Souradip
Published: (2025)
by: Nath, Souradip
Published: (2025)
Just another copy and paste? Comparing the security vulnerabilities of ChatGPT generated code and StackOverflow answers
by: Hamer, Sivana, et al.
Published: (2024)
by: Hamer, Sivana, et al.
Published: (2024)
ChatGPT: Excellent Paper! Accept It. Editor: Imposter Found! Review Rejected
by: Gharami, Kanchon, et al.
Published: (2025)
by: Gharami, Kanchon, et al.
Published: (2025)
DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models
by: Wang, Boxin, et al.
Published: (2023)
by: Wang, Boxin, et al.
Published: (2023)
Confused ChatGPT: Cross-App Context Poisoning via First-Party APIs
by: Wang, Chao, et al.
Published: (2026)
by: Wang, Chao, et al.
Published: (2026)
WildCode: An Empirical Analysis of Code Generated by ChatGPT
by: Khanmohammadi, Kobra, et al.
Published: (2025)
by: Khanmohammadi, Kobra, et al.
Published: (2025)
GPT-4 Jailbreaks Itself with Near-Perfect Success Using Self-Explanation
by: Ramesh, Govind, et al.
Published: (2024)
by: Ramesh, Govind, et al.
Published: (2024)
ChatGPT, is this real? The influence of generative AI on writing style in top-tier cybersecurity papers
by: Vansteenhuyse, Daan
Published: (2026)
by: Vansteenhuyse, Daan
Published: (2026)
AbuseGPT: Abuse of Generative AI ChatBots to Create Smishing Campaigns
by: Shibli, Ashfak Md, et al.
Published: (2024)
by: Shibli, Ashfak Md, et al.
Published: (2024)
Enhancing Android Malware Detection: The Influence of ChatGPT on Decision-centric Task
by: Li, Yao, et al.
Published: (2024)
by: Li, Yao, et al.
Published: (2024)
ChatGPT and Other Large Language Models for Cybersecurity of Smart Grid Applications
by: Zaboli, Aydin, et al.
Published: (2023)
by: Zaboli, Aydin, et al.
Published: (2023)
Low-Resource Languages Jailbreak GPT-4
by: Yong, Zheng-Xin, et al.
Published: (2023)
by: Yong, Zheng-Xin, et al.
Published: (2023)
Using Hallucinations to Bypass GPT4's Filter
by: Lemkin, Benjamin
Published: (2024)
by: Lemkin, Benjamin
Published: (2024)
From static to adaptive: immune memory-based jailbreak detection for large language models
by: Leng, Jun, et al.
Published: (2025)
by: Leng, Jun, et al.
Published: (2025)
Generative AI like ChatGPT in Blockchain Federated Learning: use cases, opportunities and future
by: Puppala, Sai, et al.
Published: (2024)
by: Puppala, Sai, et al.
Published: (2024)
Breaking the Prompt Wall (I): A Real-World Case Study of Attacking ChatGPT via Lightweight Prompt Injection
by: Chang, Xiangyu, et al.
Published: (2025)
by: Chang, Xiangyu, et al.
Published: (2025)
AVISE: Framework for Evaluating the Security of AI Systems
by: Lempinen, Mikko, et al.
Published: (2026)
by: Lempinen, Mikko, et al.
Published: (2026)
TeleAI-Safety: A comprehensive LLM jailbreaking benchmark towards attacks, defenses, and evaluations
by: Chen, Xiuyuan, et al.
Published: (2025)
by: Chen, Xiuyuan, et al.
Published: (2025)
Similar Items
-
Comparative Analysis Based on DeepSeek, ChatGPT, and Google Gemini: Features, Techniques, Performance, Future Prospects
by: Rahman, Anichur, et al.
Published: (2025) -
AuditGPT: Auditing Smart Contracts with ChatGPT
by: Xia, Shihao, et al.
Published: (2024) -
Exploring ChatGPT's Capabilities on Vulnerability Management
by: Liu, Peiyu, et al.
Published: (2023) -
Exfiltration of personal information from ChatGPT via prompt injection
by: Schwartzman, Gregory
Published: (2024) -
LLM Platform Security: Applying a Systematic Evaluation Framework to OpenAI's ChatGPT Plugins
by: Iqbal, Umar, et al.
Published: (2023)