When LLMs Go Online: The Emerging Threat of Web-Enabled LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Hanna, Song, Minkyoo, Na, Seung Ho, Shin, Seungwon, Lee, Kimin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Defending MoE LLMs against Harmful Fine-Tuning via Safety Routing Alignment
von: Kim, Jaehan, et al.
Veröffentlicht: (2025)
von: Kim, Jaehan, et al.
Veröffentlicht: (2025)
Obliviate: Neutralizing Task-agnostic Backdoors within the Parameter-efficient Fine-tuning Paradigm
von: Kim, Jaehan, et al.
Veröffentlicht: (2024)
von: Kim, Jaehan, et al.
Veröffentlicht: (2024)
Claim-Guided Textual Backdoor Attack for Practical Applications
von: Song, Minkyoo, et al.
Veröffentlicht: (2024)
von: Song, Minkyoo, et al.
Veröffentlicht: (2024)
PassREfinder-FL: Privacy-Preserving Credential Stuffing Risk Prediction via Graph-Based Federated Learning for Representing Password Reuse between Websites
von: Kim, Jaehan, et al.
Veröffentlicht: (2025)
von: Kim, Jaehan, et al.
Veröffentlicht: (2025)
Subgraph Reconstruction Attacks on Graph RAG Deployments with Practical Defenses
von: Song, Minkyoo, et al.
Veröffentlicht: (2026)
von: Song, Minkyoo, et al.
Veröffentlicht: (2026)
A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations
von: Zhou, Yihe, et al.
Veröffentlicht: (2025)
von: Zhou, Yihe, et al.
Veröffentlicht: (2025)
CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence
von: Alam, Md Tanvirul, et al.
Veröffentlicht: (2024)
von: Alam, Md Tanvirul, et al.
Veröffentlicht: (2024)
Proactively Detecting Threats: A Novel Approach Using LLMs
von: Chawla, Aniesh, et al.
Veröffentlicht: (2026)
von: Chawla, Aniesh, et al.
Veröffentlicht: (2026)
CyberSOCEval: Benchmarking LLMs Capabilities for Malware Analysis and Threat Intelligence Reasoning
von: Deason, Lauren, et al.
Veröffentlicht: (2025)
von: Deason, Lauren, et al.
Veröffentlicht: (2025)
AthenaBench: A Dynamic Benchmark for Evaluating LLMs in Cyber Threat Intelligence
von: Alam, Md Tanvirul, et al.
Veröffentlicht: (2025)
von: Alam, Md Tanvirul, et al.
Veröffentlicht: (2025)
WIPI: A New Web Threat for LLM-Driven Web Agents
von: Wu, Fangzhou, et al.
Veröffentlicht: (2024)
von: Wu, Fangzhou, et al.
Veröffentlicht: (2024)
When Bots Take the Bait: Exposing and Mitigating the Emerging Social Engineering Attack in Web Automation Agent
von: Wu, Xinyi, et al.
Veröffentlicht: (2026)
von: Wu, Xinyi, et al.
Veröffentlicht: (2026)
Confusion is the Final Barrier: Rethinking Jailbreak Evaluation and Investigating the Real Misuse Threat of LLMs
von: Yan, Yu, et al.
Veröffentlicht: (2025)
von: Yan, Yu, et al.
Veröffentlicht: (2025)
The Ethics of Interaction: Mitigating Security Threats in LLMs
von: Kumar, Ashutosh, et al.
Veröffentlicht: (2024)
von: Kumar, Ashutosh, et al.
Veröffentlicht: (2024)
A Systematic Evaluation of Parameter-Efficient Fine-Tuning Methods for the Security of Code LLMs
von: Lee, Kiho, et al.
Veröffentlicht: (2025)
von: Lee, Kiho, et al.
Veröffentlicht: (2025)
When LLMs Meet Cybersecurity: A Systematic Literature Review
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem
von: Beurer-Kellner, Luca, et al.
Veröffentlicht: (2026)
von: Beurer-Kellner, Luca, et al.
Veröffentlicht: (2026)
Agentic Misalignment: How LLMs Could Be Insider Threats
von: Lynch, Aengus, et al.
Veröffentlicht: (2025)
von: Lynch, Aengus, et al.
Veröffentlicht: (2025)
ThreatModeling-LLM: Automating Threat Modeling using Large Language Models for Banking System
von: Wu, Tingmin, et al.
Veröffentlicht: (2024)
von: Wu, Tingmin, et al.
Veröffentlicht: (2024)
Prompt Injection as an Emerging Threat: Evaluating the Resilience of Large Language Models
von: Ganiuly, Daniyal, et al.
Veröffentlicht: (2025)
von: Ganiuly, Daniyal, et al.
Veröffentlicht: (2025)
Security Logs to ATT&CK Insights: Leveraging LLMs for High-Level Threat Understanding and Cognitive Trait Inference
von: Hans, Soham, et al.
Veröffentlicht: (2025)
von: Hans, Soham, et al.
Veröffentlicht: (2025)
$PC^2$: Politically Controversial Content Generation via Jailbreaking Attacks on GPT-based Text-to-Image Models
von: Choi, Wonwoo, et al.
Veröffentlicht: (2026)
von: Choi, Wonwoo, et al.
Veröffentlicht: (2026)
Decoding Latent Attack Surfaces in LLMs: Prompt Injection via HTML in Web Summarization
von: Verma, Ishaan, et al.
Veröffentlicht: (2025)
von: Verma, Ishaan, et al.
Veröffentlicht: (2025)
AI-Driven Cybersecurity Threats: A Survey of Emerging Risks and Defensive Strategies
von: Erukude, Sai Teja, et al.
Veröffentlicht: (2026)
von: Erukude, Sai Teja, et al.
Veröffentlicht: (2026)
When Grammar Guides the Attack: Uncovering Control-Plane Vulnerabilities in LLMs with Structured Output
von: Zhang, Shuoming, et al.
Veröffentlicht: (2025)
von: Zhang, Shuoming, et al.
Veröffentlicht: (2025)
When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack
von: Sun, Zehan, et al.
Veröffentlicht: (2026)
von: Sun, Zehan, et al.
Veröffentlicht: (2026)
When Good Sounds Go Adversarial: Jailbreaking Audio-Language Models with Benign Inputs
von: Dingeto, Hiskias, et al.
Veröffentlicht: (2025)
von: Dingeto, Hiskias, et al.
Veröffentlicht: (2025)
Do You Trust Your Model? Emerging Malware Threats in the Deep Learning Ecosystem
von: Hitaj, Dorjan, et al.
Veröffentlicht: (2024)
von: Hitaj, Dorjan, et al.
Veröffentlicht: (2024)
Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation
von: Yu, Jiongchi, et al.
Veröffentlicht: (2025)
von: Yu, Jiongchi, et al.
Veröffentlicht: (2025)
On Technique Identification and Threat-Actor Attribution using LLMs and Embedding Models
von: Guru, Kyla, et al.
Veröffentlicht: (2025)
von: Guru, Kyla, et al.
Veröffentlicht: (2025)
PUZZLED: Jailbreaking LLMs through Word-Based Puzzles
von: Ahn, Yelim, et al.
Veröffentlicht: (2025)
von: Ahn, Yelim, et al.
Veröffentlicht: (2025)
ATLANTIS: AI-driven Threat Localization, Analysis, and Triage Intelligence System
von: Kim, Taesoo, et al.
Veröffentlicht: (2025)
von: Kim, Taesoo, et al.
Veröffentlicht: (2025)
What Really Matters in Many-Shot Attacks? An Empirical Study of Long-Context Vulnerabilities in LLMs
von: Kim, Sangyeop, et al.
Veröffentlicht: (2025)
von: Kim, Sangyeop, et al.
Veröffentlicht: (2025)
When Backdoors Go Beyond Triggers: Semantic Drift in Diffusion Models Under Encoder Attacks
von: Chen, Shenyang, et al.
Veröffentlicht: (2026)
von: Chen, Shenyang, et al.
Veröffentlicht: (2026)
Enabling Trustworthy Federated Learning via Remote Attestation for Mitigating Byzantine Threats
von: Zhang, Chaoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Chaoyu, et al.
Veröffentlicht: (2025)
Emerging Threats and Countermeasures in Neuromorphic Systems: A Survey
von: Sorrentino, Pablo, et al.
Veröffentlicht: (2026)
von: Sorrentino, Pablo, et al.
Veröffentlicht: (2026)
Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs
von: Kim, Jinhwa, et al.
Veröffentlicht: (2025)
von: Kim, Jinhwa, et al.
Veröffentlicht: (2025)
Security Threat Modeling for Emerging AI-Agent Protocols: A Comparative Analysis of MCP, A2A, Agora, and ANP
von: Anbiaee, Zeynab, et al.
Veröffentlicht: (2026)
von: Anbiaee, Zeynab, et al.
Veröffentlicht: (2026)
Privacy-Preserving LLMs Routing
von: Wu, Xidong, et al.
Veröffentlicht: (2026)
von: Wu, Xidong, et al.
Veröffentlicht: (2026)
On the Privacy of LLMs: An Ablation Study
von: Makhlouf, Karima, et al.
Veröffentlicht: (2026)
von: Makhlouf, Karima, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Defending MoE LLMs against Harmful Fine-Tuning via Safety Routing Alignment
von: Kim, Jaehan, et al.
Veröffentlicht: (2025) -
Obliviate: Neutralizing Task-agnostic Backdoors within the Parameter-efficient Fine-tuning Paradigm
von: Kim, Jaehan, et al.
Veröffentlicht: (2024) -
Claim-Guided Textual Backdoor Attack for Practical Applications
von: Song, Minkyoo, et al.
Veröffentlicht: (2024) -
PassREfinder-FL: Privacy-Preserving Credential Stuffing Risk Prediction via Graph-Based Federated Learning for Representing Password Reuse between Websites
von: Kim, Jaehan, et al.
Veröffentlicht: (2025) -
Subgraph Reconstruction Attacks on Graph RAG Deployments with Practical Defenses
von: Song, Minkyoo, et al.
Veröffentlicht: (2026)