Guardado en:
| Autores principales: | Shahariar, G M, Nazi, Zabir Al, Bhuiyan, Md Olid Hasan, Shi, Zhouxing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2601.05739 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Are Vision Language Models Cross-Cultural Theory of Mind Reasoners?
por: Nazi, Zabir Al, et al.
Publicado: (2025)
por: Nazi, Zabir Al, et al.
Publicado: (2025)
PII-Bench: Evaluating Query-Aware Privacy Protection Systems
por: Shen, Hao, et al.
Publicado: (2025)
por: Shen, Hao, et al.
Publicado: (2025)
An Evaluation of Chat Safety Moderations in Roblox
por: Kaushik, Priya, et al.
Publicado: (2026)
por: Kaushik, Priya, et al.
Publicado: (2026)
CAPID: Context-Aware PII Detection for Question-Answering Systems
por: Ponomarenko, Mariia, et al.
Publicado: (2026)
por: Ponomarenko, Mariia, et al.
Publicado: (2026)
Evaluating Granularity in Markov Chain-Based Trust Models for Vehicular Ad Hoc Networks (VANETs)
por: Shahariar, Rezvi
Publicado: (2026)
por: Shahariar, Rezvi
Publicado: (2026)
VisualLeakBench: Auditing the Fragility of Large Vision-Language Models against PII Leakage and Social Engineering
por: Wang, Youting, et al.
Publicado: (2026)
por: Wang, Youting, et al.
Publicado: (2026)
PIIGuard: Mitigating PII Harvesting under Adversarial Sanitization
por: Liu, Mingshuo, et al.
Publicado: (2026)
por: Liu, Mingshuo, et al.
Publicado: (2026)
PATCH: Mitigating PII Leakage in Language Models with Privacy-Aware Targeted Circuit PatcHing
por: Hughes, Anthony, et al.
Publicado: (2025)
por: Hughes, Anthony, et al.
Publicado: (2025)
PII-Compass: Guiding LLM training data extraction prompts towards the target PII via grounding
por: Nakka, Krishna Kanth, et al.
Publicado: (2024)
por: Nakka, Krishna Kanth, et al.
Publicado: (2024)
Comparative Analysis Based on DeepSeek, ChatGPT, and Google Gemini: Features, Techniques, Performance, Future Prospects
por: Rahman, Anichur, et al.
Publicado: (2025)
por: Rahman, Anichur, et al.
Publicado: (2025)
SecureGate: Learning When to Reveal PII Safely via Token-Gated Dual-Adapters for Federated LLMs
por: Shaaban, Mohamed, et al.
Publicado: (2026)
por: Shaaban, Mohamed, et al.
Publicado: (2026)
TRIDENT -- A Three-Tier Privacy-Preserving Propaganda Detection Model in Mobile Networks using Transformers, Adversarial Learning, and Differential Privacy
por: Emran, Al Nahian Bin, et al.
Publicado: (2025)
por: Emran, Al Nahian Bin, et al.
Publicado: (2025)
"I Strongly Suspect This Website Is a Scam": Benchmarking PII Leakage and Detection without Defense in Autonomous Web Agents
por: Roy, Soham, et al.
Publicado: (2026)
por: Roy, Soham, et al.
Publicado: (2026)
Privacy-Preserving Federated Vision Transformer Learning Leveraging Lightweight Homomorphic Encryption in Medical AI
por: Amin, Al, et al.
Publicado: (2025)
por: Amin, Al, et al.
Publicado: (2025)
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
por: Tong, Haibo, et al.
Publicado: (2026)
por: Tong, Haibo, et al.
Publicado: (2026)
FedPoisonTTP: A Threat Model and Poisoning Attack for Federated Test-Time Personalization
por: Iftee, Md Akil Raihan, et al.
Publicado: (2025)
por: Iftee, Md Akil Raihan, et al.
Publicado: (2025)
What's Privacy Good for? Measuring Privacy as a Shield from Harms due to Personal Data Use
por: Gajavalli, Sri Harsha, et al.
Publicado: (2025)
por: Gajavalli, Sri Harsha, et al.
Publicado: (2025)
DePrompt: Desensitization and Evaluation of Personal Identifiable Information in Large Language Model Prompts
por: Sun, Xiongtao, et al.
Publicado: (2024)
por: Sun, Xiongtao, et al.
Publicado: (2024)
ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models
por: Zhao, Yunhan, et al.
Publicado: (2026)
por: Zhao, Yunhan, et al.
Publicado: (2026)
A Case Study on the Impact of Anonymization Along the RAG Pipeline
por: Bodea, Andreea-Elena, et al.
Publicado: (2026)
por: Bodea, Andreea-Elena, et al.
Publicado: (2026)
Assessing the influence of cybersecurity threats and risks on the adoption and growth of digital banking: a systematic literature review
por: Waliullah, Md., et al.
Publicado: (2025)
por: Waliullah, Md., et al.
Publicado: (2025)
PII Jailbreaking in LLMs via Activation Steering Reveals Personal Information Leakage
por: Nakka, Krishna Kanth, et al.
Publicado: (2025)
por: Nakka, Krishna Kanth, et al.
Publicado: (2025)
Large language models in healthcare and medical domain: A review
por: Nazi, Zabir Al, et al.
Publicado: (2023)
por: Nazi, Zabir Al, et al.
Publicado: (2023)
A Hardware-Anchored Privacy Middleware for PII Sharing Across Heterogeneous Embedded Consumer Devices
por: Sabbineni, Aditya, et al.
Publicado: (2026)
por: Sabbineni, Aditya, et al.
Publicado: (2026)
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations
por: Ye, Mang, et al.
Publicado: (2025)
por: Ye, Mang, et al.
Publicado: (2025)
GuidedBench: Measuring and Mitigating the Evaluation Discrepancies of In-the-wild LLM Jailbreak Methods
por: Huang, Ruixuan, et al.
Publicado: (2025)
por: Huang, Ruixuan, et al.
Publicado: (2025)
The Need Of Trustworthy Announcements To Achieve Driving Comfort
por: Shahariar, Rezvi, et al.
Publicado: (2024)
por: Shahariar, Rezvi, et al.
Publicado: (2024)
A trust management framework for vehicular ad hoc networks
por: Shahariar, Rezvi, et al.
Publicado: (2024)
por: Shahariar, Rezvi, et al.
Publicado: (2024)
A fuzzy reward and punishment scheme for vehicular ad hoc networks
por: Shahariar, Rezvi, et al.
Publicado: (2024)
por: Shahariar, Rezvi, et al.
Publicado: (2024)
A Survey of Security Threats and Trust Management in Vehicular Ad Hoc Networks
por: Shahariar, Rezvi, et al.
Publicado: (2026)
por: Shahariar, Rezvi, et al.
Publicado: (2026)
Adversarial Attacks on Parts of Speech: An Empirical Study in Text-to-Image Generation
por: Shahariar, G M, et al.
Publicado: (2024)
por: Shahariar, G M, et al.
Publicado: (2024)
UnPII: Unlearning Personally Identifiable Information with Quantifiable Exposure Risk
por: Jeon, Intae, et al.
Publicado: (2026)
por: Jeon, Intae, et al.
Publicado: (2026)
PrivLM-Bench: A Multi-level Privacy Evaluation Benchmark for Language Models
por: Li, Haoran, et al.
Publicado: (2023)
por: Li, Haoran, et al.
Publicado: (2023)
HarmLevelBench: Evaluating Harm-Level Compliance and the Impact of Quantization on Model Alignment
por: Belkhiter, Yannis, et al.
Publicado: (2024)
por: Belkhiter, Yannis, et al.
Publicado: (2024)
Reconstruction of Personally Identifiable Information from Supervised Finetuned Models
por: Furukawa, Sae, et al.
Publicado: (2026)
por: Furukawa, Sae, et al.
Publicado: (2026)
T2VSafetyBench: Evaluating the Safety of Text-to-Video Generative Models
por: Miao, Yibo, et al.
Publicado: (2024)
por: Miao, Yibo, et al.
Publicado: (2024)
WebPII: Benchmarking Visual PII Detection for Computer-Use Agents
por: Zhao, Nathan
Publicado: (2026)
por: Zhao, Nathan
Publicado: (2026)
SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces
por: Jin, Chang, et al.
Publicado: (2026)
por: Jin, Chang, et al.
Publicado: (2026)
PandaGuard: Systematic Evaluation of LLM Safety against Jailbreaking Attacks
por: Shen, Guobin, et al.
Publicado: (2025)
por: Shen, Guobin, et al.
Publicado: (2025)
S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models
por: Yuan, Xiaohan, et al.
Publicado: (2024)
por: Yuan, Xiaohan, et al.
Publicado: (2024)
Ejemplares similares
-
Are Vision Language Models Cross-Cultural Theory of Mind Reasoners?
por: Nazi, Zabir Al, et al.
Publicado: (2025) -
PII-Bench: Evaluating Query-Aware Privacy Protection Systems
por: Shen, Hao, et al.
Publicado: (2025) -
An Evaluation of Chat Safety Moderations in Roblox
por: Kaushik, Priya, et al.
Publicado: (2026) -
CAPID: Context-Aware PII Detection for Question-Answering Systems
por: Ponomarenko, Mariia, et al.
Publicado: (2026) -
Evaluating Granularity in Markov Chain-Based Trust Models for Vehicular Ad Hoc Networks (VANETs)
por: Shahariar, Rezvi
Publicado: (2026)