FCert: Certifiably Robust Few-Shot Classification in the Era of Foundation Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Yanting, Zou, Wei, Jia, Jinyuan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
EnsembleSHAP: Faithful and Certifiably Robust Attribution for Random Subspace Method
por: Wang, Yanting, et al.
Publicado: (2026)
por: Wang, Yanting, et al.
Publicado: (2026)
AgentWatcher: A Rule-based Prompt Injection Monitor
por: Wang, Yanting, et al.
Publicado: (2026)
por: Wang, Yanting, et al.
Publicado: (2026)
MMCert: Provable Defense against Adversarial Attacks to Multi-modal Models
por: Wang, Yanting, et al.
Publicado: (2024)
por: Wang, Yanting, et al.
Publicado: (2024)
TracLLM: A Generic Framework for Attributing Long Context LLMs
por: Wang, Yanting, et al.
Publicado: (2025)
por: Wang, Yanting, et al.
Publicado: (2025)
FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption
por: Wang, Yanting, et al.
Publicado: (2026)
por: Wang, Yanting, et al.
Publicado: (2026)
TrojanDec: Data-free Detection of Trojan Inputs in Self-supervised Learning
por: Liu, Yupei, et al.
Publicado: (2025)
por: Liu, Yupei, et al.
Publicado: (2025)
Distributed Backdoor Attacks on Federated Graph Learning and Certified Defenses
por: Yang, Yuxin, et al.
Publicado: (2024)
por: Yang, Yuxin, et al.
Publicado: (2024)
PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
por: Zou, Wei, et al.
Publicado: (2025)
por: Zou, Wei, et al.
Publicado: (2025)
TASO: Jailbreak LLMs via Alternative Template and Suffix Optimization
por: Wang, Yanting, et al.
Publicado: (2025)
por: Wang, Yanting, et al.
Publicado: (2025)
AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption
por: Wang, Yanting, et al.
Publicado: (2025)
por: Wang, Yanting, et al.
Publicado: (2025)
PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses
por: Yin, Chenlong, et al.
Publicado: (2026)
por: Yin, Chenlong, et al.
Publicado: (2026)
Certifiably Robust Image Watermark
por: Jiang, Zhengyuan, et al.
Publicado: (2024)
por: Jiang, Zhengyuan, et al.
Publicado: (2024)
PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models
por: Zou, Wei, et al.
Publicado: (2024)
por: Zou, Wei, et al.
Publicado: (2024)
SecInfer: Preventing Prompt Injection via Inference-time Scaling
por: Liu, Yupei, et al.
Publicado: (2025)
por: Liu, Yupei, et al.
Publicado: (2025)
TrojFM: Resource-efficient Backdoor Attacks against Very Large Foundation Models
por: Nie, Yuzhou., et al.
Publicado: (2024)
por: Nie, Yuzhou., et al.
Publicado: (2024)
UniC-RAG: Universal Knowledge Corruption Attacks to Retrieval-Augmented Generation
por: Geng, Runpeng, et al.
Publicado: (2025)
por: Geng, Runpeng, et al.
Publicado: (2025)
Provably Robust Multi-bit Watermarking for AI-generated Text
por: Qu, Wenjie, et al.
Publicado: (2024)
por: Qu, Wenjie, et al.
Publicado: (2024)
Provably Robust Explainable Graph Neural Networks against Graph Perturbation Attacks
por: Li, Jiate, et al.
Publicado: (2025)
por: Li, Jiate, et al.
Publicado: (2025)
A Certified Robust Watermark For Large Language Models
por: Feng, Xianheng, et al.
Publicado: (2024)
por: Feng, Xianheng, et al.
Publicado: (2024)
CleanBase: Detecting Malicious Documents in RAG Knowledge Databases
por: Jin, Weifei, et al.
Publicado: (2026)
por: Jin, Weifei, et al.
Publicado: (2026)
A Classification-by-Retrieval Framework for Few-Shot Anomaly Detection to Detect API Injection Attacks
por: Aharon, Udi, et al.
Publicado: (2024)
por: Aharon, Udi, et al.
Publicado: (2024)
FedGMark: Certifiably Robust Watermarking for Federated Graph Learning
por: Yang, Yuxin, et al.
Publicado: (2024)
por: Yang, Yuxin, et al.
Publicado: (2024)
PCEvolve: Private Contrastive Evolution for Synthetic Dataset Generation via Few-Shot Private Data and Generative APIs
por: Zhang, Jianqing, et al.
Publicado: (2025)
por: Zhang, Jianqing, et al.
Publicado: (2025)
Robust Privacy: Inference-Time Privacy through Certified Robustness
por: Jin, Jiankai, et al.
Publicado: (2026)
por: Jin, Jiankai, et al.
Publicado: (2026)
PIArena: A Platform for Prompt Injection Evaluation
por: Geng, Runpeng, et al.
Publicado: (2026)
por: Geng, Runpeng, et al.
Publicado: (2026)
Evaluating LLM-based Personal Information Extraction and Countermeasures
por: Liu, Yupei, et al.
Publicado: (2024)
por: Liu, Yupei, et al.
Publicado: (2024)
MalMixer: Few-Shot Malware Classification with Retrieval-Augmented Semi-Supervised Learning
por: Li, Jiliang, et al.
Publicado: (2024)
por: Li, Jiliang, et al.
Publicado: (2024)
Boosting Certified Robustness for Time Series Classification with Efficient Self-Ensemble
por: Dong, Chang, et al.
Publicado: (2024)
por: Dong, Chang, et al.
Publicado: (2024)
CERT-ED: Certifiably Robust Text Classification for Edit Distance
por: Huang, Zhuoqun, et al.
Publicado: (2024)
por: Huang, Zhuoqun, et al.
Publicado: (2024)
Getting a-Round Guarantees: Floating-Point Attacks on Certified Robustness
por: Jin, Jiankai, et al.
Publicado: (2022)
por: Jin, Jiankai, et al.
Publicado: (2022)
RobustMask: Certified Robustness against Adversarial Neural Ranking Attack via Randomized Masking
por: Liu, Jiawei, et al.
Publicado: (2025)
por: Liu, Jiawei, et al.
Publicado: (2025)
Risks of Practicing Large Language Models in Smart Grid: Threat Modeling and Validation
por: Li, Jiangnan, et al.
Publicado: (2024)
por: Li, Jiangnan, et al.
Publicado: (2024)
Position: Certified Robustness Does Not (Yet) Imply Model Security
por: Cullen, Andrew C., et al.
Publicado: (2025)
por: Cullen, Andrew C., et al.
Publicado: (2025)
TRUSTDESC: Preventing Tool Poisoning in LLM Applications via Trusted Description Generation
por: Ye, Hengkai, et al.
Publicado: (2026)
por: Ye, Hengkai, et al.
Publicado: (2026)
Few-Shot Learning-Based Cyber Incident Detection with Augmented Context Intelligence
por: Zuo, Fei, et al.
Publicado: (2025)
por: Zuo, Fei, et al.
Publicado: (2025)
On Using Certified Training towards Empirical Robustness
por: De Palma, Alessandro, et al.
Publicado: (2024)
por: De Palma, Alessandro, et al.
Publicado: (2024)
Distributed Intrusion Detection in Dynamic Networks of UAVs using Few-Shot Federated Learning
por: Ceviz, Ozlem, et al.
Publicado: (2025)
por: Ceviz, Ozlem, et al.
Publicado: (2025)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
por: Geng, Runpeng, et al.
Publicado: (2025)
por: Geng, Runpeng, et al.
Publicado: (2025)
Certified Causal Defense with Generalizable Robustness
por: Qiao, Yiran, et al.
Publicado: (2024)
por: Qiao, Yiran, et al.
Publicado: (2024)
Dataset Ownership in the Era of Large Language Models
por: Li, Kun, et al.
Publicado: (2025)
por: Li, Kun, et al.
Publicado: (2025)
Ejemplares similares
-
EnsembleSHAP: Faithful and Certifiably Robust Attribution for Random Subspace Method
por: Wang, Yanting, et al.
Publicado: (2026) -
AgentWatcher: A Rule-based Prompt Injection Monitor
por: Wang, Yanting, et al.
Publicado: (2026) -
MMCert: Provable Defense against Adversarial Attacks to Multi-modal Models
por: Wang, Yanting, et al.
Publicado: (2024) -
TracLLM: A Generic Framework for Attributing Long Context LLMs
por: Wang, Yanting, et al.
Publicado: (2025) -
FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption
por: Wang, Yanting, et al.
Publicado: (2026)