Is On-Device AI Broken and Exploitable? Assessing the Trust and Ethics in Small Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Nakka, Kalyan, Dani, Jimmy, Saxena, Nitesh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LiteLMGuard: Seamless and Lightweight On-Device Prompt Filtering for Safeguarding Small Language Models against Quantization-induced Risks and Vulnerabilities
by: Nakka, Kalyan, et al.
Published: (2025)
by: Nakka, Kalyan, et al.
Published: (2025)
A Machine Learning-Based Framework for Assessing Cryptographic Indistinguishability of Lightweight Block Ciphers
by: Dani, Jimmy, et al.
Published: (2024)
by: Dani, Jimmy, et al.
Published: (2024)
BitBypass: A New Direction in Jailbreaking Aligned Large Language Models with Bitstream Camouflage
by: Nakka, Kalyan, et al.
Published: (2025)
by: Nakka, Kalyan, et al.
Published: (2025)
When AI Defeats Password Deception! A Deep Learning Framework to Distinguish Passwords and Honeywords
by: Dani, Jimmy, et al.
Published: (2024)
by: Dani, Jimmy, et al.
Published: (2024)
The First Early Evidence of the Use of Browser Fingerprinting for Online Tracking
by: Liu, Zengrui, et al.
Published: (2024)
by: Liu, Zengrui, et al.
Published: (2024)
The System Prompt Is the Attack Surface: How LLM Agent Configuration Shapes Security and Creates Exploitable Vulnerabilities
by: Litvak, Ron
Published: (2026)
by: Litvak, Ron
Published: (2026)
Robust and Verifiable MPC with Applications to Linear Machine Learning Inference
by: Wang, Tzu-Shen, et al.
Published: (2025)
by: Wang, Tzu-Shen, et al.
Published: (2025)
A Comprehensive Study of Exploitable Patterns in Smart Contracts: From Vulnerability to Defense
by: Ding, Yuchen, et al.
Published: (2025)
by: Ding, Yuchen, et al.
Published: (2025)
SME-TEAM: Leveraging Trust and Ethics for Secure and Responsible Use of AI and LLMs in SMEs
by: Sarker, Iqbal H., et al.
Published: (2025)
by: Sarker, Iqbal H., et al.
Published: (2025)
Building Trustworthy Multimodal AI: A Review of Fairness, Transparency, and Ethics in Vision-Language Tasks
by: Saleh, Mohammad, et al.
Published: (2025)
by: Saleh, Mohammad, et al.
Published: (2025)
MemTrust: A Zero-Trust Architecture for Unified AI Memory System
by: Zhou, Xing, et al.
Published: (2026)
by: Zhou, Xing, et al.
Published: (2026)
Memory-Efficient and Secure DNN Inference on TrustZone-enabled Consumer IoT Devices
by: Xie, Xueshuo, et al.
Published: (2024)
by: Xie, Xueshuo, et al.
Published: (2024)
Unvalidated Trust: Cross-Stage Vulnerabilities in Large Language Model Architectures
by: Schwarz, Dominik
Published: (2025)
by: Schwarz, Dominik
Published: (2025)
The End of Trust: How Agentic AI Breaks Security Assumptions
by: Zafar, Osama, et al.
Published: (2026)
by: Zafar, Osama, et al.
Published: (2026)
Synthetic Trust Attacks: Modeling How Generative AI Manipulates Human Decisions in Social Engineering Fraud
by: Ashraf, Muhammad Tahir
Published: (2026)
by: Ashraf, Muhammad Tahir
Published: (2026)
LLM-PBE: Assessing Data Privacy in Large Language Models
by: Li, Qinbin, et al.
Published: (2024)
by: Li, Qinbin, et al.
Published: (2024)
MAIF: Enforcing AI Trust and Provenance with an Artifact-Centric Agentic Paradigm
by: Narajala, Vineeth Sai, et al.
Published: (2025)
by: Narajala, Vineeth Sai, et al.
Published: (2025)
Caging the Agents: A Zero Trust Security Architecture for Autonomous AI in Healthcare
by: Maiti, Saikat
Published: (2026)
by: Maiti, Saikat
Published: (2026)
AgentTrust: Runtime Safety Evaluation and Interception for AI Agent Tool Use
by: Yang, Chenglin
Published: (2026)
by: Yang, Chenglin
Published: (2026)
Toward a Unified Security Framework for AI Agents: Trust, Risk, and Liability
by: Mo, Jiayun, et al.
Published: (2025)
by: Mo, Jiayun, et al.
Published: (2025)
Win-k: Improved Membership Inference Attacks on Small Language Models
by: Arkhmammadova, Roya, et al.
Published: (2025)
by: Arkhmammadova, Roya, et al.
Published: (2025)
Towards Small Language Models for Security Query Generation in SOC Workflows
by: Muzammil, Saleha, et al.
Published: (2025)
by: Muzammil, Saleha, et al.
Published: (2025)
Assessing Spear-Phishing Website Generation in Large Language Model Coding Agents
by: Malloy, Tailia, et al.
Published: (2026)
by: Malloy, Tailia, et al.
Published: (2026)
A Unified Framework for Human AI Collaboration in Security Operations Centers with Trusted Autonomy
by: Mohsin, Ahmad, et al.
Published: (2025)
by: Mohsin, Ahmad, et al.
Published: (2025)
Can You Trust Your Copilot? A Privacy Scorecard for AI Coding Assistants
by: AL-Maamari, Amir
Published: (2025)
by: AL-Maamari, Amir
Published: (2025)
Trusted AI Agents in the Cloud
by: Bodea, Teofil, et al.
Published: (2025)
by: Bodea, Teofil, et al.
Published: (2025)
Broken by Default: A Formal Verification Study of Security Vulnerabilities in AI-Generated Code
by: Blain, Dominik, et al.
Published: (2026)
by: Blain, Dominik, et al.
Published: (2026)
Small Language Models for Phishing Website Detection: Cost, Performance, and Privacy Trade-Offs
by: Goldenits, Georg, et al.
Published: (2025)
by: Goldenits, Georg, et al.
Published: (2025)
Towards Privacy-Preserving and Personalized Smart Homes via Tailored Small Language Models
by: Huang, Xinyu, et al.
Published: (2025)
by: Huang, Xinyu, et al.
Published: (2025)
Fine-Tuning Small Language Models for Solution-Oriented Windows Event Log Analysis
by: Akhtar, Siraaj, et al.
Published: (2026)
by: Akhtar, Siraaj, et al.
Published: (2026)
Can Small Language Models Reliably Resist Jailbreak Attacks? A Comprehensive Evaluation
by: Zhang, Wenhui, et al.
Published: (2025)
by: Zhang, Wenhui, et al.
Published: (2025)
Securing Generative AI in Healthcare: A Zero-Trust Architecture Powered by Confidential Computing on Google Cloud
by: Amanna, Adaobi, et al.
Published: (2025)
by: Amanna, Adaobi, et al.
Published: (2025)
DP-FedLoRA: Privacy-Enhanced Federated Fine-Tuning for On-Device Large Language Models
by: Xu, Honghui, et al.
Published: (2025)
by: Xu, Honghui, et al.
Published: (2025)
On-Device Watermarking: A Socio-Technical Imperative For Authenticity In The Age of Generative AI
by: Kherraz, Houssam
Published: (2025)
by: Kherraz, Houssam
Published: (2025)
GUARD-SLM: Token Activation-Based Defense Against Jailbreak Attacks for Small Language Models
by: Mia, Md Jueal, et al.
Published: (2026)
by: Mia, Md Jueal, et al.
Published: (2026)
Case Study: Fine-tuning Small Language Models for Accurate and Private CWE Detection in Python Code
by: Bappy, Md. Azizul Hakim, et al.
Published: (2025)
by: Bappy, Md. Azizul Hakim, et al.
Published: (2025)
Winning at All Cost: A Small Environment for Eliciting Specification Gaming Behaviors in Large Language Models
by: Malmqvist, Lars
Published: (2025)
by: Malmqvist, Lars
Published: (2025)
Revolutionizing Cyber Threat Detection with Large Language Models: A privacy-preserving BERT-based Lightweight Model for IoT/IIoT Devices
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
Securing GenAI Multi-Agent Systems Against Tool Squatting: A Zero Trust Registry-Based Approach
by: Narajala, Vineeth Sai, et al.
Published: (2025)
by: Narajala, Vineeth Sai, et al.
Published: (2025)
Do You Trust Your Model? Emerging Malware Threats in the Deep Learning Ecosystem
by: Hitaj, Dorjan, et al.
Published: (2024)
by: Hitaj, Dorjan, et al.
Published: (2024)
Similar Items
-
LiteLMGuard: Seamless and Lightweight On-Device Prompt Filtering for Safeguarding Small Language Models against Quantization-induced Risks and Vulnerabilities
by: Nakka, Kalyan, et al.
Published: (2025) -
A Machine Learning-Based Framework for Assessing Cryptographic Indistinguishability of Lightweight Block Ciphers
by: Dani, Jimmy, et al.
Published: (2024) -
BitBypass: A New Direction in Jailbreaking Aligned Large Language Models with Bitstream Camouflage
by: Nakka, Kalyan, et al.
Published: (2025) -
When AI Defeats Password Deception! A Deep Learning Framework to Distinguish Passwords and Honeywords
by: Dani, Jimmy, et al.
Published: (2024) -
The First Early Evidence of the Use of Browser Fingerprinting for Online Tracking
by: Liu, Zengrui, et al.
Published: (2024)