Saved in:
| Main Authors: | Wang, Yanting, Jia, Jinyuan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.30034 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FCert: Certifiably Robust Few-Shot Classification in the Era of Foundation Models
by: Wang, Yanting, et al.
Published: (2024)
by: Wang, Yanting, et al.
Published: (2024)
AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption
by: Wang, Yanting, et al.
Published: (2025)
by: Wang, Yanting, et al.
Published: (2025)
TracLLM: A Generic Framework for Attributing Long Context LLMs
by: Wang, Yanting, et al.
Published: (2025)
by: Wang, Yanting, et al.
Published: (2025)
TrojanDec: Data-free Detection of Trojan Inputs in Self-supervised Learning
by: Liu, Yupei, et al.
Published: (2025)
by: Liu, Yupei, et al.
Published: (2025)
AgentWatcher: A Rule-based Prompt Injection Monitor
by: Wang, Yanting, et al.
Published: (2026)
by: Wang, Yanting, et al.
Published: (2026)
FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption
by: Wang, Yanting, et al.
Published: (2026)
by: Wang, Yanting, et al.
Published: (2026)
Distributed Backdoor Attacks on Federated Graph Learning and Certified Defenses
by: Yang, Yuxin, et al.
Published: (2024)
by: Yang, Yuxin, et al.
Published: (2024)
TASO: Jailbreak LLMs via Alternative Template and Suffix Optimization
by: Wang, Yanting, et al.
Published: (2025)
by: Wang, Yanting, et al.
Published: (2025)
PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses
by: Yin, Chenlong, et al.
Published: (2026)
by: Yin, Chenlong, et al.
Published: (2026)
Certifiably Robust Image Watermark
by: Jiang, Zhengyuan, et al.
Published: (2024)
by: Jiang, Zhengyuan, et al.
Published: (2024)
SecInfer: Preventing Prompt Injection via Inference-time Scaling
by: Liu, Yupei, et al.
Published: (2025)
by: Liu, Yupei, et al.
Published: (2025)
MMCert: Provable Defense against Adversarial Attacks to Multi-modal Models
by: Wang, Yanting, et al.
Published: (2024)
by: Wang, Yanting, et al.
Published: (2024)
UniC-RAG: Universal Knowledge Corruption Attacks to Retrieval-Augmented Generation
by: Geng, Runpeng, et al.
Published: (2025)
by: Geng, Runpeng, et al.
Published: (2025)
PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
by: Zou, Wei, et al.
Published: (2025)
by: Zou, Wei, et al.
Published: (2025)
Provably Robust Explainable Graph Neural Networks against Graph Perturbation Attacks
by: Li, Jiate, et al.
Published: (2025)
by: Li, Jiate, et al.
Published: (2025)
PIArena: A Platform for Prompt Injection Evaluation
by: Geng, Runpeng, et al.
Published: (2026)
by: Geng, Runpeng, et al.
Published: (2026)
RobustMask: Certified Robustness against Adversarial Neural Ranking Attack via Randomized Masking
by: Liu, Jiawei, et al.
Published: (2025)
by: Liu, Jiawei, et al.
Published: (2025)
Ideal Attribution and Faithful Watermarks for Language Models
by: Song, Min Jae, et al.
Published: (2025)
by: Song, Min Jae, et al.
Published: (2025)
Adaptive Randomized Smoothing: Certified Adversarial Robustness for Multi-Step Defences
by: Lyu, Saiyue, et al.
Published: (2024)
by: Lyu, Saiyue, et al.
Published: (2024)
Explainable Threat Attribution for IoT Networks Using Conditional SHAP and Flow Behavior Modelling
by: Ozechi, Samuel, et al.
Published: (2026)
by: Ozechi, Samuel, et al.
Published: (2026)
FedGMark: Certifiably Robust Watermarking for Federated Graph Learning
by: Yang, Yuxin, et al.
Published: (2024)
by: Yang, Yuxin, et al.
Published: (2024)
Registered Attribute-Based Encryption with Publicly Verifiable Certified Deletion, Everlasting Security, and More
by: Murshid, Shayeef, et al.
Published: (2026)
by: Murshid, Shayeef, et al.
Published: (2026)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
by: Geng, Runpeng, et al.
Published: (2025)
by: Geng, Runpeng, et al.
Published: (2025)
Interpretable Ensemble Learning for Network Traffic Anomaly Detection: A SHAP-based Explainable AI Framework for Embedded Systems Security
by: Shao, Wanru
Published: (2026)
by: Shao, Wanru
Published: (2026)
A Robust Certified Machine Unlearning Method Under Distribution Shift
by: Guo, Jinduo, et al.
Published: (2026)
by: Guo, Jinduo, et al.
Published: (2026)
Boosting Certified Robustness for Time Series Classification with Efficient Self-Ensemble
by: Dong, Chang, et al.
Published: (2024)
by: Dong, Chang, et al.
Published: (2024)
A Certified Robust Watermark For Large Language Models
by: Feng, Xianheng, et al.
Published: (2024)
by: Feng, Xianheng, et al.
Published: (2024)
Evaluating LLM-based Personal Information Extraction and Countermeasures
by: Liu, Yupei, et al.
Published: (2024)
by: Liu, Yupei, et al.
Published: (2024)
Towards Strong Certified Defense with Universal Asymmetric Randomization
by: Hong, Hanbin, et al.
Published: (2025)
by: Hong, Hanbin, et al.
Published: (2025)
Certified PEFTSmoothing: Parameter-Efficient Fine-Tuning with Randomized Smoothing
by: Fu, Chengyan, et al.
Published: (2024)
by: Fu, Chengyan, et al.
Published: (2024)
Certified Adversarial Robustness of Machine Learning-based Malware Detectors via (De)Randomized Smoothing
by: Gibert, Daniel, et al.
Published: (2024)
by: Gibert, Daniel, et al.
Published: (2024)
AdaptDel: Adaptable Deletion Rate Randomized Smoothing for Certified Robustness
by: Huang, Zhuoqun, et al.
Published: (2025)
by: Huang, Zhuoqun, et al.
Published: (2025)
TRUSTDESC: Preventing Tool Poisoning in LLM Applications via Trusted Description Generation
by: Ye, Hengkai, et al.
Published: (2026)
by: Ye, Hengkai, et al.
Published: (2026)
Provably Robust Multi-bit Watermarking for AI-generated Text
by: Qu, Wenjie, et al.
Published: (2024)
by: Qu, Wenjie, et al.
Published: (2024)
Getting a-Round Guarantees: Floating-Point Attacks on Certified Robustness
by: Jin, Jiankai, et al.
Published: (2022)
by: Jin, Jiankai, et al.
Published: (2022)
TrojFM: Resource-efficient Backdoor Attacks against Very Large Foundation Models
by: Nie, Yuzhou., et al.
Published: (2024)
by: Nie, Yuzhou., et al.
Published: (2024)
On the Equivalence between Classical Position Verification and Certified Randomness
by: Kaleoglu, Fatih, et al.
Published: (2024)
by: Kaleoglu, Fatih, et al.
Published: (2024)
On Using Certified Training towards Empirical Robustness
by: De Palma, Alessandro, et al.
Published: (2024)
by: De Palma, Alessandro, et al.
Published: (2024)
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
by: Zou, Wei, et al.
Published: (2024)
by: Zou, Wei, et al.
Published: (2024)
Certified Causal Attribution for Real-Time Attack Forensics in 6G Network Slicing
by: Quan, Minh K., et al.
Published: (2026)
by: Quan, Minh K., et al.
Published: (2026)
Similar Items
-
FCert: Certifiably Robust Few-Shot Classification in the Era of Foundation Models
by: Wang, Yanting, et al.
Published: (2024) -
AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption
by: Wang, Yanting, et al.
Published: (2025) -
TracLLM: A Generic Framework for Attributing Long Context LLMs
by: Wang, Yanting, et al.
Published: (2025) -
TrojanDec: Data-free Detection of Trojan Inputs in Self-supervised Learning
by: Liu, Yupei, et al.
Published: (2025) -
AgentWatcher: A Rule-based Prompt Injection Monitor
by: Wang, Yanting, et al.
Published: (2026)