EdgeAIGuard: Agentic LLMs for Minor Protection in Digital Spaces
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mujtaba, Ghulam, Khowaja, Sunder Ali, Dev, Kapal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Pathway to Secure and Trustworthy ZSM for LLMs: Attacks, Defense, and Opportunities
von: Khowaja, Sunder Ali, et al.
Veröffentlicht: (2024)
von: Khowaja, Sunder Ali, et al.
Veröffentlicht: (2024)
ChatGPT Needs SPADE (Sustainability, PrivAcy, Digital divide, and Ethics) Evaluation: A Review
von: Khowaja, Sunder Ali, et al.
Veröffentlicht: (2023)
von: Khowaja, Sunder Ali, et al.
Veröffentlicht: (2023)
Agentic LLMs as Powerful Deanonymizers: Re-identification of Participants in the Anthropic Interviewer Dataset
von: Li, Tianshi
Veröffentlicht: (2026)
von: Li, Tianshi
Veröffentlicht: (2026)
RedTeamLLM: an Agentic AI framework for offensive security
von: Challita, Brian, et al.
Veröffentlicht: (2025)
von: Challita, Brian, et al.
Veröffentlicht: (2025)
RLCP: A Reinforcement Learning-based Copyright Protection Method for Text-to-Image Diffusion Model
von: Shi, Zhuan, et al.
Veröffentlicht: (2024)
von: Shi, Zhuan, et al.
Veröffentlicht: (2024)
Global Challenge for Safe and Secure LLMs Track 1
von: Jia, Xiaojun, et al.
Veröffentlicht: (2024)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2024)
Gender-Based Heterogeneity in Youth Privacy-Protective Behavior for Smart Voice Assistants: Evidence from Multigroup PLS-SEM
von: Campbell, Molly, et al.
Veröffentlicht: (2026)
von: Campbell, Molly, et al.
Veröffentlicht: (2026)
Fingerprinting and Tracing Shadows: The Development and Impact of Browser Fingerprinting on Digital Privacy
von: Lawall, Alexander
Veröffentlicht: (2024)
von: Lawall, Alexander
Veröffentlicht: (2024)
On Technique Identification and Threat-Actor Attribution using LLMs and Embedding Models
von: Guru, Kyla, et al.
Veröffentlicht: (2025)
von: Guru, Kyla, et al.
Veröffentlicht: (2025)
DSSmoothing: Toward Certified Dataset Ownership Verification for Pre-trained Language Models via Dual-Space Smoothing
von: Qiao, Ting, et al.
Veröffentlicht: (2025)
von: Qiao, Ting, et al.
Veröffentlicht: (2025)
Framework, Standards, Applications and Best practices of Responsible AI : A Comprehensive Survey
von: Gadekallu, Thippa Reddy, et al.
Veröffentlicht: (2025)
von: Gadekallu, Thippa Reddy, et al.
Veröffentlicht: (2025)
Secure and Trustworthy Artificial Intelligence-Extended Reality (AI-XR) for Metaverses
von: Qayyum, Adnan, et al.
Veröffentlicht: (2022)
von: Qayyum, Adnan, et al.
Veröffentlicht: (2022)
Medical Malice: A Dataset for Context-Aware Safety in Healthcare LLMs
von: D'addario, Andrew Maranhão Ventura
Veröffentlicht: (2025)
von: D'addario, Andrew Maranhão Ventura
Veröffentlicht: (2025)
A Survey on Responsible LLMs: Inherent Risk, Malicious Use, and Mitigation Strategy
von: Wang, Huandong, et al.
Veröffentlicht: (2025)
von: Wang, Huandong, et al.
Veröffentlicht: (2025)
Can LLMs Infer Conversational Agent Users' Personality Traits from Chat History?
von: Cögendez, Derya, et al.
Veröffentlicht: (2026)
von: Cögendez, Derya, et al.
Veröffentlicht: (2026)
The Earth is Flat because...: Investigating LLMs' Belief towards Misinformation via Persuasive Conversation
von: Xu, Rongwu, et al.
Veröffentlicht: (2023)
von: Xu, Rongwu, et al.
Veröffentlicht: (2023)
SEAL: An Open, Auditable, and Fair Data Generation Framework for AI-Native 6G Networks
von: Khowaja, Sunder Ali, et al.
Veröffentlicht: (2026)
von: Khowaja, Sunder Ali, et al.
Veröffentlicht: (2026)
Advanced Architectures Integrated with Agentic AI for Next-Generation Wireless Networks
von: Dev, Kapal, et al.
Veröffentlicht: (2025)
von: Dev, Kapal, et al.
Veröffentlicht: (2025)
Leveraging Digital Twin Technologies for Public Space Protection and Vulnerability Assessment
von: Stefanidou, Artemis, et al.
Veröffentlicht: (2024)
von: Stefanidou, Artemis, et al.
Veröffentlicht: (2024)
Enforcing Cybersecurity Constraints for LLM-driven Robot Agents for Online Transactions
von: Shah, Shraddha Pradipbhai, et al.
Veröffentlicht: (2025)
von: Shah, Shraddha Pradipbhai, et al.
Veröffentlicht: (2025)
Ask What Your Country Can Do For You: Towards a Public Red Teaming Model
von: Kennedy, Wm. Matthew, et al.
Veröffentlicht: (2025)
von: Kennedy, Wm. Matthew, et al.
Veröffentlicht: (2025)
The End Of Universal Lifelong Identifiers: Identity Systems For The AI Era
von: Palakodety, Shriphani
Veröffentlicht: (2025)
von: Palakodety, Shriphani
Veröffentlicht: (2025)
AI Safety vs. AI Security: Demystifying the Distinction and Boundaries
von: Lin, Zhiqiang, et al.
Veröffentlicht: (2025)
von: Lin, Zhiqiang, et al.
Veröffentlicht: (2025)
Comparing AI Agents to Cybersecurity Professionals in Real-World Penetration Testing
von: Lin, Justin W., et al.
Veröffentlicht: (2025)
von: Lin, Justin W., et al.
Veröffentlicht: (2025)
Synthetic Data and Health Privacy
von: Abgrall, Gwénolé, et al.
Veröffentlicht: (2025)
von: Abgrall, Gwénolé, et al.
Veröffentlicht: (2025)
Playing Devil's Advocate: Unmasking Toxicity and Vulnerabilities in Large Vision-Language Models
von: Erol, Abdulkadir, et al.
Veröffentlicht: (2025)
von: Erol, Abdulkadir, et al.
Veröffentlicht: (2025)
Position: Ensuring mutual privacy is necessary for effective external evaluation of proprietary AI systems
von: Bucknall, Ben, et al.
Veröffentlicht: (2025)
von: Bucknall, Ben, et al.
Veröffentlicht: (2025)
The Pitfalls of "Security by Obscurity" And What They Mean for Transparent AI
von: Hall, Peter, et al.
Veröffentlicht: (2025)
von: Hall, Peter, et al.
Veröffentlicht: (2025)
A software security review on Uganda's Mobile Money Services: Dr. Jim Spire's tweets sentiment analysis
von: Wilberforce, Nsengiyumva
Veröffentlicht: (2025)
von: Wilberforce, Nsengiyumva
Veröffentlicht: (2025)
User Privacy and Large Language Models: An Analysis of Frontier Developers' Privacy Policies
von: King, Jennifer, et al.
Veröffentlicht: (2025)
von: King, Jennifer, et al.
Veröffentlicht: (2025)
AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models
von: Barrett, Anthony M., et al.
Veröffentlicht: (2025)
von: Barrett, Anthony M., et al.
Veröffentlicht: (2025)
Recommender Systems for Democracy: Toward Adversarial Robustness in Voting Advice Applications
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
von: Berdoz, Frédéric, et al.
Veröffentlicht: (2025)
Naming is framing: How cybersecurity's language problems are repeating in AI governance
von: Potter, Lianne
Veröffentlicht: (2025)
von: Potter, Lianne
Veröffentlicht: (2025)
The Impact of AI on the Cyber Offense-Defense Balance and the Character of Cyber Conflict
von: Lohn, Andrew J.
Veröffentlicht: (2025)
von: Lohn, Andrew J.
Veröffentlicht: (2025)
BadScientist: Can a Research Agent Write Convincing but Unsound Papers that Fool LLM Reviewers?
von: Jiang, Fengqing, et al.
Veröffentlicht: (2025)
von: Jiang, Fengqing, et al.
Veröffentlicht: (2025)
Love, Lies, and Language Models: Investigating AI's Role in Romance-Baiting Scams
von: Gressel, Gilad, et al.
Veröffentlicht: (2025)
von: Gressel, Gilad, et al.
Veröffentlicht: (2025)
Private, Verifiable, and Auditable AI Systems
von: South, Tobin
Veröffentlicht: (2025)
von: South, Tobin
Veröffentlicht: (2025)
Trust and Dependability in Blockchain & AI Based MedIoT Applications: Research Challenges and Future Directions
von: Solaiman, Ellis, et al.
Veröffentlicht: (2025)
von: Solaiman, Ellis, et al.
Veröffentlicht: (2025)
Optimal Allocation of Privacy Budget on Hierarchical Data Release
von: Ko, Joonhyuk, et al.
Veröffentlicht: (2025)
von: Ko, Joonhyuk, et al.
Veröffentlicht: (2025)
Can AI Models be Jailbroken to Phish Elderly Victims? An End-to-End Evaluation
von: Heiding, Fred, et al.
Veröffentlicht: (2025)
von: Heiding, Fred, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Pathway to Secure and Trustworthy ZSM for LLMs: Attacks, Defense, and Opportunities
von: Khowaja, Sunder Ali, et al.
Veröffentlicht: (2024) -
ChatGPT Needs SPADE (Sustainability, PrivAcy, Digital divide, and Ethics) Evaluation: A Review
von: Khowaja, Sunder Ali, et al.
Veröffentlicht: (2023) -
Agentic LLMs as Powerful Deanonymizers: Re-identification of Participants in the Anthropic Interviewer Dataset
von: Li, Tianshi
Veröffentlicht: (2026) -
RedTeamLLM: an Agentic AI framework for offensive security
von: Challita, Brian, et al.
Veröffentlicht: (2025) -
RLCP: A Reinforcement Learning-based Copyright Protection Method for Text-to-Image Diffusion Model
von: Shi, Zhuan, et al.
Veröffentlicht: (2024)