GREAT Score: Global Robustness Evaluation of Adversarial Perturbation using Generative Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Zaitang, Chen, Pin-Yu, Ho, Tsung-Yi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Retention Score: Quantifying Jailbreak Risks for Vision Language Models
por: Li, Zaitang, et al.
Publicado: (2024)
por: Li, Zaitang, et al.
Publicado: (2024)
Defining and Evaluating Physical Safety for Large Language Models
por: Tang, Yung-Chen, et al.
Publicado: (2024)
por: Tang, Yung-Chen, et al.
Publicado: (2024)
NaNa and MiGu: Semantic Data Augmentation Techniques to Enhance Protein Classification in Graph Neural Networks
por: Lan, Yi-Shan, et al.
Publicado: (2024)
por: Lan, Yi-Shan, et al.
Publicado: (2024)
Gradient Cuff: Detecting Jailbreak Attacks on Large Language Models by Exploring Refusal Loss Landscapes
por: Hu, Xiaomeng, et al.
Publicado: (2024)
por: Hu, Xiaomeng, et al.
Publicado: (2024)
Data-Driven Lipschitz Continuity: A Cost-Effective Approach to Improve Adversarial Robustness
por: Chen, Erh-Chung, et al.
Publicado: (2024)
por: Chen, Erh-Chung, et al.
Publicado: (2024)
How Worst-Case Are Adversarial Attacks? Linking Adversarial and Perturbation Robustness
por: Rossolini, Giulio
Publicado: (2026)
por: Rossolini, Giulio
Publicado: (2026)
Computational Safety for Generative AI: A Signal Processing Perspective
por: Chen, Pin-Yu
Publicado: (2025)
por: Chen, Pin-Yu
Publicado: (2025)
Towards Reliable Evaluation of Adversarial Robustness for Spiking Neural Networks
por: Wang, Jihang, et al.
Publicado: (2025)
por: Wang, Jihang, et al.
Publicado: (2025)
Generating Universal Adversarial Perturbations for Quantum Classifiers
por: Anil, Gautham, et al.
Publicado: (2024)
por: Anil, Gautham, et al.
Publicado: (2024)
KCLNet: Electrically Equivalence-Oriented Graph Representation Learning for Analog Circuits
por: Xu, Peng, et al.
Publicado: (2026)
por: Xu, Peng, et al.
Publicado: (2026)
Certified Robustness against Sparse Adversarial Perturbations via Data Localization
por: Pal, Ambar, et al.
Publicado: (2024)
por: Pal, Ambar, et al.
Publicado: (2024)
Perturbing the Phase: Analyzing Adversarial Robustness of Complex-Valued Neural Networks
por: Eilers, Florian, et al.
Publicado: (2026)
por: Eilers, Florian, et al.
Publicado: (2026)
Robust Deep Reinforcement Learning with Adaptive Adversarial Perturbations in Action Space
por: Liu, Qianmei, et al.
Publicado: (2024)
por: Liu, Qianmei, et al.
Publicado: (2024)
Bridging Models to Defend: A Population-Based Strategy for Robust Adversarial Defense
por: Wang, Ren, et al.
Publicado: (2023)
por: Wang, Ren, et al.
Publicado: (2023)
Optimization-Free Universal Watermark Forgery with Regenerative Diffusion Models
por: Zhu, Chaoyi, et al.
Publicado: (2025)
por: Zhu, Chaoyi, et al.
Publicado: (2025)
CoP: Agentic Red-teaming for Large Language Models using Composition of Principles
por: Xiong, Chen, et al.
Publicado: (2025)
por: Xiong, Chen, et al.
Publicado: (2025)
Unifying Adversarial Perturbation for Graph Neural Networks
por: Yang, Jinluan, et al.
Publicado: (2025)
por: Yang, Jinluan, et al.
Publicado: (2025)
Neural Clamping: Joint Input Perturbation and Temperature Scaling for Neural Network Calibration
por: Tang, Yung-Chen, et al.
Publicado: (2022)
por: Tang, Yung-Chen, et al.
Publicado: (2022)
Sim2Act: Robust Simulation-to-Decision Learning via Adversarial Calibration and Group-Relative Perturbation
por: Cao, Hongyu, et al.
Publicado: (2026)
por: Cao, Hongyu, et al.
Publicado: (2026)
PermLLM: Learnable Channel Permutation for N:M Sparse Large Language Models
por: Zou, Lancheng, et al.
Publicado: (2025)
por: Zou, Lancheng, et al.
Publicado: (2025)
Assessing Robustness via Score-Based Adversarial Image Generation
por: Kollovieh, Marcel, et al.
Publicado: (2023)
por: Kollovieh, Marcel, et al.
Publicado: (2023)
Adversarial Robustness Overestimation and Instability in TRADES
por: Li, Jonathan Weiping, et al.
Publicado: (2024)
por: Li, Jonathan Weiping, et al.
Publicado: (2024)
Explanation-Guided Adversarial Training for Robust and Interpretable Models
por: Chen, Chao, et al.
Publicado: (2026)
por: Chen, Chao, et al.
Publicado: (2026)
A Novel Perturb-ability Score to Mitigate Evasion Adversarial Attacks on Flow-Based ML-NIDS
por: elShehaby, Mohamed, et al.
Publicado: (2024)
por: elShehaby, Mohamed, et al.
Publicado: (2024)
Linking Robustness and Generalization: A k* Distribution Analysis of Concept Clustering in Latent Space for Vision Models
por: Kotyan, Shashank, et al.
Publicado: (2024)
por: Kotyan, Shashank, et al.
Publicado: (2024)
Robustness-enhanced Uplift Modeling with Adversarial Feature Desensitization
por: Sun, Zexu, et al.
Publicado: (2023)
por: Sun, Zexu, et al.
Publicado: (2023)
Diffusion Guided Adversarial State Perturbations in Reinforcement Learning
por: Sun, Xiaolin, et al.
Publicado: (2025)
por: Sun, Xiaolin, et al.
Publicado: (2025)
Exploring Adversarial Robustness of Deep State Space Models
por: Qi, Biqing, et al.
Publicado: (2024)
por: Qi, Biqing, et al.
Publicado: (2024)
Elijah: Eliminating Backdoors Injected in Diffusion Models via Distribution Shift
por: An, Shengwei, et al.
Publicado: (2023)
por: An, Shengwei, et al.
Publicado: (2023)
GF-Score: Certified Class-Conditional Robustness Evaluation with Fairness Guarantees
por: Shah, Arya, et al.
Publicado: (2026)
por: Shah, Arya, et al.
Publicado: (2026)
GREAT-EER: Graph Edge Attention Network for Emergency Evacuation Responses
por: Lischka, Attila, et al.
Publicado: (2026)
por: Lischka, Attila, et al.
Publicado: (2026)
A GREAT Architecture for Edge-Based Graph Problems Like TSP
por: Lischka, Attila, et al.
Publicado: (2024)
por: Lischka, Attila, et al.
Publicado: (2024)
Improved Diffusion-based Generative Model with Better Adversarial Robustness
por: Wang, Zekun, et al.
Publicado: (2025)
por: Wang, Zekun, et al.
Publicado: (2025)
Evaluating Adversarial Robustness of Concept Representations in Sparse Autoencoders
por: Li, Aaron J., et al.
Publicado: (2025)
por: Li, Aaron J., et al.
Publicado: (2025)
Zero-Sacrifice Persistent-Robustness Adversarial Defense for Pre-Trained Encoders
por: Lei, Zhuxin, et al.
Publicado: (2026)
por: Lei, Zhuxin, et al.
Publicado: (2026)
Diversity Boosts AI-Generated Text Detection
por: Basani, Advik Raj, et al.
Publicado: (2025)
por: Basani, Advik Raj, et al.
Publicado: (2025)
Adversarial Preference Learning for Robust LLM Alignment
por: Wang, Yuanfu, et al.
Publicado: (2025)
por: Wang, Yuanfu, et al.
Publicado: (2025)
Robust Model-Based Reinforcement Learning with an Adversarial Auxiliary Model
por: Herremans, Siemen, et al.
Publicado: (2024)
por: Herremans, Siemen, et al.
Publicado: (2024)
Ignition Phase : Standard Training for Fast Adversarial Robustness
por: Yu-Hang, Wang, et al.
Publicado: (2025)
por: Yu-Hang, Wang, et al.
Publicado: (2025)
Mitigating Adversarial Perturbations for Deep Reinforcement Learning via Vector Quantization
por: Luu, Tung M., et al.
Publicado: (2024)
por: Luu, Tung M., et al.
Publicado: (2024)
Ejemplares similares
-
Retention Score: Quantifying Jailbreak Risks for Vision Language Models
por: Li, Zaitang, et al.
Publicado: (2024) -
Defining and Evaluating Physical Safety for Large Language Models
por: Tang, Yung-Chen, et al.
Publicado: (2024) -
NaNa and MiGu: Semantic Data Augmentation Techniques to Enhance Protein Classification in Graph Neural Networks
por: Lan, Yi-Shan, et al.
Publicado: (2024) -
Gradient Cuff: Detecting Jailbreak Attacks on Large Language Models by Exploring Refusal Loss Landscapes
por: Hu, Xiaomeng, et al.
Publicado: (2024) -
Data-Driven Lipschitz Continuity: A Cost-Effective Approach to Improve Adversarial Robustness
por: Chen, Erh-Chung, et al.
Publicado: (2024)