FCert: Certifiably Robust Few-Shot Classification in the Era of Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yanting, Zou, Wei, Jia, Jinyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EnsembleSHAP: Faithful and Certifiably Robust Attribution for Random Subspace Method
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
AgentWatcher: A Rule-based Prompt Injection Monitor
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
MMCert: Provable Defense against Adversarial Attacks to Multi-modal Models
von: Wang, Yanting, et al.
Veröffentlicht: (2024)
von: Wang, Yanting, et al.
Veröffentlicht: (2024)
TracLLM: A Generic Framework for Attributing Long Context LLMs
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
TrojanDec: Data-free Detection of Trojan Inputs in Self-supervised Learning
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
Distributed Backdoor Attacks on Federated Graph Learning and Certified Defenses
von: Yang, Yuxin, et al.
Veröffentlicht: (2024)
von: Yang, Yuxin, et al.
Veröffentlicht: (2024)
PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
von: Zou, Wei, et al.
Veröffentlicht: (2025)
von: Zou, Wei, et al.
Veröffentlicht: (2025)
TASO: Jailbreak LLMs via Alternative Template and Suffix Optimization
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses
von: Yin, Chenlong, et al.
Veröffentlicht: (2026)
von: Yin, Chenlong, et al.
Veröffentlicht: (2026)
Certifiably Robust Image Watermark
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2024)
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2024)
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
von: Zou, Wei, et al.
Veröffentlicht: (2024)
von: Zou, Wei, et al.
Veröffentlicht: (2024)
SecInfer: Preventing Prompt Injection via Inference-time Scaling
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
TrojFM: Resource-efficient Backdoor Attacks against Very Large Foundation Models
von: Nie, Yuzhou., et al.
Veröffentlicht: (2024)
von: Nie, Yuzhou., et al.
Veröffentlicht: (2024)
UniC-RAG: Universal Knowledge Corruption Attacks to Retrieval-Augmented Generation
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
Provably Robust Multi-bit Watermarking for AI-generated Text
von: Qu, Wenjie, et al.
Veröffentlicht: (2024)
von: Qu, Wenjie, et al.
Veröffentlicht: (2024)
Provably Robust Explainable Graph Neural Networks against Graph Perturbation Attacks
von: Li, Jiate, et al.
Veröffentlicht: (2025)
von: Li, Jiate, et al.
Veröffentlicht: (2025)
A Certified Robust Watermark For Large Language Models
von: Feng, Xianheng, et al.
Veröffentlicht: (2024)
von: Feng, Xianheng, et al.
Veröffentlicht: (2024)
CleanBase: Detecting Malicious Documents in RAG Knowledge Databases
von: Jin, Weifei, et al.
Veröffentlicht: (2026)
von: Jin, Weifei, et al.
Veröffentlicht: (2026)
A Classification-by-Retrieval Framework for Few-Shot Anomaly Detection to Detect API Injection Attacks
von: Aharon, Udi, et al.
Veröffentlicht: (2024)
von: Aharon, Udi, et al.
Veröffentlicht: (2024)
FedGMark: Certifiably Robust Watermarking for Federated Graph Learning
von: Yang, Yuxin, et al.
Veröffentlicht: (2024)
von: Yang, Yuxin, et al.
Veröffentlicht: (2024)
PCEvolve: Private Contrastive Evolution for Synthetic Dataset Generation via Few-Shot Private Data and Generative APIs
von: Zhang, Jianqing, et al.
Veröffentlicht: (2025)
von: Zhang, Jianqing, et al.
Veröffentlicht: (2025)
Robust Privacy: Inference-Time Privacy through Certified Robustness
von: Jin, Jiankai, et al.
Veröffentlicht: (2026)
von: Jin, Jiankai, et al.
Veröffentlicht: (2026)
PIArena: A Platform for Prompt Injection Evaluation
von: Geng, Runpeng, et al.
Veröffentlicht: (2026)
von: Geng, Runpeng, et al.
Veröffentlicht: (2026)
Evaluating LLM-based Personal Information Extraction and Countermeasures
von: Liu, Yupei, et al.
Veröffentlicht: (2024)
von: Liu, Yupei, et al.
Veröffentlicht: (2024)
MalMixer: Few-Shot Malware Classification with Retrieval-Augmented Semi-Supervised Learning
von: Li, Jiliang, et al.
Veröffentlicht: (2024)
von: Li, Jiliang, et al.
Veröffentlicht: (2024)
Boosting Certified Robustness for Time Series Classification with Efficient Self-Ensemble
von: Dong, Chang, et al.
Veröffentlicht: (2024)
von: Dong, Chang, et al.
Veröffentlicht: (2024)
CERT-ED: Certifiably Robust Text Classification for Edit Distance
von: Huang, Zhuoqun, et al.
Veröffentlicht: (2024)
von: Huang, Zhuoqun, et al.
Veröffentlicht: (2024)
Getting a-Round Guarantees: Floating-Point Attacks on Certified Robustness
von: Jin, Jiankai, et al.
Veröffentlicht: (2022)
von: Jin, Jiankai, et al.
Veröffentlicht: (2022)
RobustMask: Certified Robustness against Adversarial Neural Ranking Attack via Randomized Masking
von: Liu, Jiawei, et al.
Veröffentlicht: (2025)
von: Liu, Jiawei, et al.
Veröffentlicht: (2025)
Risks of Practicing Large Language Models in Smart Grid: Threat Modeling and Validation
von: Li, Jiangnan, et al.
Veröffentlicht: (2024)
von: Li, Jiangnan, et al.
Veröffentlicht: (2024)
Position: Certified Robustness Does Not (Yet) Imply Model Security
von: Cullen, Andrew C., et al.
Veröffentlicht: (2025)
von: Cullen, Andrew C., et al.
Veröffentlicht: (2025)
TRUSTDESC: Preventing Tool Poisoning in LLM Applications via Trusted Description Generation
von: Ye, Hengkai, et al.
Veröffentlicht: (2026)
von: Ye, Hengkai, et al.
Veröffentlicht: (2026)
Few-Shot Learning-Based Cyber Incident Detection with Augmented Context Intelligence
von: Zuo, Fei, et al.
Veröffentlicht: (2025)
von: Zuo, Fei, et al.
Veröffentlicht: (2025)
On Using Certified Training towards Empirical Robustness
von: De Palma, Alessandro, et al.
Veröffentlicht: (2024)
von: De Palma, Alessandro, et al.
Veröffentlicht: (2024)
Distributed Intrusion Detection in Dynamic Networks of UAVs using Few-Shot Federated Learning
von: Ceviz, Ozlem, et al.
Veröffentlicht: (2025)
von: Ceviz, Ozlem, et al.
Veröffentlicht: (2025)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
Certified Causal Defense with Generalizable Robustness
von: Qiao, Yiran, et al.
Veröffentlicht: (2024)
von: Qiao, Yiran, et al.
Veröffentlicht: (2024)
Dataset Ownership in the Era of Large Language Models
von: Li, Kun, et al.
Veröffentlicht: (2025)
von: Li, Kun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EnsembleSHAP: Faithful and Certifiably Robust Attribution for Random Subspace Method
von: Wang, Yanting, et al.
Veröffentlicht: (2026) -
AgentWatcher: A Rule-based Prompt Injection Monitor
von: Wang, Yanting, et al.
Veröffentlicht: (2026) -
MMCert: Provable Defense against Adversarial Attacks to Multi-modal Models
von: Wang, Yanting, et al.
Veröffentlicht: (2024) -
TracLLM: A Generic Framework for Attributing Long Context LLMs
von: Wang, Yanting, et al.
Veröffentlicht: (2025) -
FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption
von: Wang, Yanting, et al.
Veröffentlicht: (2026)