Defending Against Social Engineering Attacks in the Age of LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ai, Lin, Kumarage, Tharindu, Bhattacharjee, Amrita, Liu, Zizhou, Hui, Zheng, Davinroy, Michael, Cook, James, Cassani, Laura, Trapeznikov, Kirill, Kirchner, Matthias, Basharat, Arslan, Hoogs, Anthony, Garland, Joshua, Liu, Huan, Hirschberg, Julia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Personalized Attacks of Social Engineering in Multi-turn Conversations: LLM Agents for Simulation and Detection
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2025)
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2025)
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2024)
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2024)
Aligning Machiavellian Agents: Behavior Steering via Test-Time Policy Shaping
von: Mujtaba, Dena, et al.
Veröffentlicht: (2025)
von: Mujtaba, Dena, et al.
Veröffentlicht: (2025)
Can LLMs Improve Multimodal Fact-Checking by Asking Relevant Questions?
von: Beigi, Alimohammad, et al.
Veröffentlicht: (2024)
von: Beigi, Alimohammad, et al.
Veröffentlicht: (2024)
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2024)
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2024)
EAGLE: A Domain Generalization Framework for AI-generated Text Detection
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2024)
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2024)
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2023)
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2023)
Can Knowledge Graphs Reduce Hallucinations in LLMs? : A Survey
von: Agrawal, Garima, et al.
Veröffentlicht: (2023)
von: Agrawal, Garima, et al.
Veröffentlicht: (2023)
Synthetic Audio Forensics Evaluation (SAFE) Challenge
von: Trapeznikov, Kirill, et al.
Veröffentlicht: (2025)
von: Trapeznikov, Kirill, et al.
Veröffentlicht: (2025)
A Survey of AI-generated Text Forensic Systems: Detection, Attribution, and Characterization
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2024)
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2024)
Mindful-RAG: A Study of Points of Failure in Retrieval Augmented Generation
von: Agrawal, Garima, et al.
Veröffentlicht: (2024)
von: Agrawal, Garima, et al.
Veröffentlicht: (2024)
QASE Enhanced PLMs: Improved Control in Text Generation for MRC
von: Ai, Lin, et al.
Veröffentlicht: (2024)
von: Ai, Lin, et al.
Veröffentlicht: (2024)
Enhancing Pre-Trained Generative Language Models with Question Attended Span Extraction on Machine Reading Comprehension
von: Ai, Lin, et al.
Veröffentlicht: (2024)
von: Ai, Lin, et al.
Veröffentlicht: (2024)
Adversarial Tuning: Defending Against Jailbreak Attacks for LLMs
von: Liu, Fan, et al.
Veröffentlicht: (2024)
von: Liu, Fan, et al.
Veröffentlicht: (2024)
Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
von: Keluskar, Aryan, et al.
Veröffentlicht: (2024)
von: Keluskar, Aryan, et al.
Veröffentlicht: (2024)
Cross-Platform Hate Speech Detection with Weakly Supervised Causal Disentanglement
von: Sheth, Paras, et al.
Veröffentlicht: (2024)
von: Sheth, Paras, et al.
Veröffentlicht: (2024)
Steerable Pluralism: Pluralistic Alignment via Few-Shot Comparative Regression
von: Adams, Jadie, et al.
Veröffentlicht: (2025)
von: Adams, Jadie, et al.
Veröffentlicht: (2025)
Advancing Reliable Synthetic Video Detection: Insights from the SAFE Challenge
von: Trapeznikov, Kirill, et al.
Veröffentlicht: (2026)
von: Trapeznikov, Kirill, et al.
Veröffentlicht: (2026)
RedditESS: A Mental Health Social Support Interaction Dataset -- Understanding Effective Social Support to Refine AI-Driven Support Tools
von: Alghamdi, Zeyad, et al.
Veröffentlicht: (2025)
von: Alghamdi, Zeyad, et al.
Veröffentlicht: (2025)
ALIGN: Prompt-based Attribute Alignment for Reliable, Responsible, and Personalized LLM-based Decision-Making
von: Ravichandran, Bharadwaj, et al.
Veröffentlicht: (2025)
von: Ravichandran, Bharadwaj, et al.
Veröffentlicht: (2025)
A Review of Incorporating Psychological Theories in LLMs
von: Liu, Zizhou, et al.
Veröffentlicht: (2025)
von: Liu, Zizhou, et al.
Veröffentlicht: (2025)
Time Traveling to Defend Against Adversarial Example Attacks in Image Classification
von: Etim, Anthony, et al.
Veröffentlicht: (2024)
von: Etim, Anthony, et al.
Veröffentlicht: (2024)
SecureGaze: Defending Gaze Estimation Against Backdoor Attacks
von: Du, Lingyu, et al.
Veröffentlicht: (2025)
von: Du, Lingyu, et al.
Veröffentlicht: (2025)
Defending LLM Watermarking Against Spoofing Attacks with Contrastive Representation Learning
von: An, Li, et al.
Veröffentlicht: (2025)
von: An, Li, et al.
Veröffentlicht: (2025)
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
von: Nirmal, Ayushi, et al.
Veröffentlicht: (2024)
von: Nirmal, Ayushi, et al.
Veröffentlicht: (2024)
Adversarial Text Purification: A Large Language Model Approach for Defense
von: Moraffah, Raha, et al.
Veröffentlicht: (2024)
von: Moraffah, Raha, et al.
Veröffentlicht: (2024)
DETAM: Defending LLMs Against Jailbreak Attacks via Targeted Attention Modification
von: Li, Yu, et al.
Veröffentlicht: (2025)
von: Li, Yu, et al.
Veröffentlicht: (2025)
No Free Lunch for Defending Against Prefilling Attack by In-Context Learning
von: Xue, Zhiyu, et al.
Veröffentlicht: (2024)
von: Xue, Zhiyu, et al.
Veröffentlicht: (2024)
Attention is All You Need to Defend Against Indirect Prompt Injection Attacks in LLMs
von: Zhong, Yinan, et al.
Veröffentlicht: (2025)
von: Zhong, Yinan, et al.
Veröffentlicht: (2025)
ResumeFlow: An LLM-facilitated Pipeline for Personalized Resume Generation and Refinement
von: Zinjad, Saurabh Bhausaheb, et al.
Veröffentlicht: (2024)
von: Zinjad, Saurabh Bhausaheb, et al.
Veröffentlicht: (2024)
Defending LVLMs Against Vision Attacks through Partial-Perception Supervision
von: Zhou, Qi, et al.
Veröffentlicht: (2024)
von: Zhou, Qi, et al.
Veröffentlicht: (2024)
Ontology-Aware RAG for Improved Question-Answering in Cybersecurity Education
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2024)
von: Zhao, Chengshuai, et al.
Veröffentlicht: (2024)
PropaInsight: Toward Deeper Understanding of Propaganda in Terms of Techniques, Appeals, and Intent
von: Liu, Jiateng, et al.
Veröffentlicht: (2024)
von: Liu, Jiateng, et al.
Veröffentlicht: (2024)
Defending Against Indirect Prompt Injection Attacks With Spotlighting
von: Hines, Keegan, et al.
Veröffentlicht: (2024)
von: Hines, Keegan, et al.
Veröffentlicht: (2024)
Defending Against Frequency-Based Attacks with Diffusion Models
von: Amerehi, Fatemeh, et al.
Veröffentlicht: (2025)
von: Amerehi, Fatemeh, et al.
Veröffentlicht: (2025)
Defending Against Poisoning Attacks in Federated Learning with Blockchain
von: Dong, Nanqing, et al.
Veröffentlicht: (2023)
von: Dong, Nanqing, et al.
Veröffentlicht: (2023)
The Role of Information Incompleteness in Defending Against Stealth Attacks
von: Sun, Ke, et al.
Veröffentlicht: (2025)
von: Sun, Ke, et al.
Veröffentlicht: (2025)
Defending Against Data Reconstruction Attacks in Federated Learning: An Information Theory Approach
von: Tan, Qi, et al.
Veröffentlicht: (2024)
von: Tan, Qi, et al.
Veröffentlicht: (2024)
Kaleidoscopic Teaming in Multi Agent Simulations
von: Mehrabi, Ninareh, et al.
Veröffentlicht: (2025)
von: Mehrabi, Ninareh, et al.
Veröffentlicht: (2025)
LLMs Can Defend Themselves Against Jailbreaking in a Practical Manner: A Vision Paper
von: Wu, Daoyuan, et al.
Veröffentlicht: (2024)
von: Wu, Daoyuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Personalized Attacks of Social Engineering in Multi-turn Conversations: LLM Agents for Simulation and Detection
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2025) -
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2024) -
Aligning Machiavellian Agents: Behavior Steering via Test-Time Policy Shaping
von: Mujtaba, Dena, et al.
Veröffentlicht: (2025) -
Can LLMs Improve Multimodal Fact-Checking by Asking Relevant Questions?
von: Beigi, Alimohammad, et al.
Veröffentlicht: (2024) -
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
von: Bhattacharjee, Amrita, et al.
Veröffentlicht: (2024)