Structured Gradient-based Interpretations via Norm-Regularized Adversarial Training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gong, Shizhan, Dou, Qi, Farnia, Farzan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Super-pixel-based Approach to the Stable Interpretation of Neural Networks
von: Gong, Shizhan, et al.
Veröffentlicht: (2024)
von: Gong, Shizhan, et al.
Veröffentlicht: (2024)
Kernel-based Unsupervised Embedding Alignment for Enhanced Visual Representation in Vision-language Models
von: Gong, Shizhan, et al.
Veröffentlicht: (2025)
von: Gong, Shizhan, et al.
Veröffentlicht: (2025)
Sparse Domain Transfer via Elastic Net Regularization
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
Interpretation of Neural Networks is Susceptible to Universal Adversarial Perturbations
von: Oskouie, Haniyeh Ehsani, et al.
Veröffentlicht: (2022)
von: Oskouie, Haniyeh Ehsani, et al.
Veröffentlicht: (2022)
Concepts from Representations: Post-hoc Concept Bottleneck Models via Sparse Decomposition of Visual Representations
von: Gong, Shizhan, et al.
Veröffentlicht: (2026)
von: Gong, Shizhan, et al.
Veröffentlicht: (2026)
Gaussian Smoothing in Saliency Maps: The Stability-Fidelity Trade-Off in Neural Network Interpretability
von: Ye, Zhuorui, et al.
Veröffentlicht: (2024)
von: Ye, Zhuorui, et al.
Veröffentlicht: (2024)
Saliency-R1: Enforcing Interpretable and Faithful Vision-language Reasoning via Saliency-map Alignment Reward
von: Gong, Shizhan, et al.
Veröffentlicht: (2026)
von: Gong, Shizhan, et al.
Veröffentlicht: (2026)
An Interpretable Evaluation of Entropy-based Novelty of Generative Models
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
Scendi Score: Prompt-Aware Diversity Evaluation via Schur Complement of CLIP Embeddings
von: Ospanov, Azim, et al.
Veröffentlicht: (2024)
von: Ospanov, Azim, et al.
Veröffentlicht: (2024)
Boosting Adversarial Training via Fisher-Rao Norm-based Regularization
von: Yin, Xiangyu, et al.
Veröffentlicht: (2024)
von: Yin, Xiangyu, et al.
Veröffentlicht: (2024)
Training-Free Distribution Adaptation for Diffusion Models via Maximum Mean Discrepancy Guidance
von: Sani, Matina Mahdizadeh, et al.
Veröffentlicht: (2026)
von: Sani, Matina Mahdizadeh, et al.
Veröffentlicht: (2026)
When Exploration Comes for Free with Mixture-Greedy: Do we need UCB in Diversity-Aware Multi-Armed Bandits?
von: Nia, Bahar Dibaei, et al.
Veröffentlicht: (2026)
von: Nia, Bahar Dibaei, et al.
Veröffentlicht: (2026)
PromptSplit: Revealing Prompt-Level Disagreement in Generative Models
von: Lotfian, Mehdi, et al.
Veröffentlicht: (2026)
von: Lotfian, Mehdi, et al.
Veröffentlicht: (2026)
SPARKE: Scalable Prompt-Aware Diversity and Novelty Guidance in Diffusion Models via RKE Score
von: Jalali, Mohammad, et al.
Veröffentlicht: (2025)
von: Jalali, Mohammad, et al.
Veröffentlicht: (2025)
Defense Against Adversarial Attacks on No-Reference Image Quality Models with Gradient Norm Regularization
von: Liu, Yujia, et al.
Veröffentlicht: (2024)
von: Liu, Yujia, et al.
Veröffentlicht: (2024)
Exposing Diversity Bias in Deep Generative Models: Statistical Origins and Correction of Diversity Error
von: Farnia, Farzan, et al.
Veröffentlicht: (2026)
von: Farnia, Farzan, et al.
Veröffentlicht: (2026)
Unveiling Differences in Generative Models: A Scalable Differential Clustering Approach
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
ACCORD: Alleviating Concept Coupling through Dependence Regularization for Text-to-Image Diffusion Personalization
von: Liu, Shizhan, et al.
Veröffentlicht: (2025)
von: Liu, Shizhan, et al.
Veröffentlicht: (2025)
Physical Adversarial Camouflage through Gradient Calibration and Regularization
von: Liang, Jiawei, et al.
Veröffentlicht: (2025)
von: Liang, Jiawei, et al.
Veröffentlicht: (2025)
3DSAM-adapter: Holistic adaptation of SAM from 2D to 3D for promptable tumor segmentation
von: Gong, Shizhan, et al.
Veröffentlicht: (2023)
von: Gong, Shizhan, et al.
Veröffentlicht: (2023)
Towards an Explainable Comparison and Alignment of Feature Embeddings
von: Jalali, Mohammad, et al.
Veröffentlicht: (2025)
von: Jalali, Mohammad, et al.
Veröffentlicht: (2025)
Conditional Vendi Score: An Information-Theoretic Approach to Diversity Evaluation of Prompt-based Generative Models
von: Jalali, Mohammad, et al.
Veröffentlicht: (2024)
von: Jalali, Mohammad, et al.
Veröffentlicht: (2024)
DiffPattern-Flex: Efficient Layout Pattern Generation via Discrete Diffusion
von: Wang, Zixiao, et al.
Veröffentlicht: (2025)
von: Wang, Zixiao, et al.
Veröffentlicht: (2025)
Boosting Adversarial Transferability via Commonality-Oriented Gradient Optimization
von: Gao, Yanting, et al.
Veröffentlicht: (2025)
von: Gao, Yanting, et al.
Veröffentlicht: (2025)
Training-Free and Interpretable Hateful Video Detection via Multi-stage Adversarial Reasoning
von: Yang, Shuonan, et al.
Veröffentlicht: (2026)
von: Yang, Shuonan, et al.
Veröffentlicht: (2026)
Releasing Inequality Phenomenon in $\ell_{\infty}$-norm Adversarial Training via Input Gradient Distillation
von: Chen, Junxi, et al.
Veröffentlicht: (2023)
von: Chen, Junxi, et al.
Veröffentlicht: (2023)
SORA: Free Second-Order Attacks in Fast Adversarial Training
von: Teymourian, Mazdak, et al.
Veröffentlicht: (2026)
von: Teymourian, Mazdak, et al.
Veröffentlicht: (2026)
pFedFair: Towards Optimal Group Fairness-Accuracy Trade-off in Heterogeneous Federated Learning
von: Lei, Haoyu, et al.
Veröffentlicht: (2025)
von: Lei, Haoyu, et al.
Veröffentlicht: (2025)
Low-Rank Tensor Recovery via Variational Schatten-p Quasi-Norm and Jacobian Regularization
von: Cheng, Zhengyun, et al.
Veröffentlicht: (2025)
von: Cheng, Zhengyun, et al.
Veröffentlicht: (2025)
GReAT: A Graph Regularized Adversarial Training Method
von: Bayram, Samet, et al.
Veröffentlicht: (2023)
von: Bayram, Samet, et al.
Veröffentlicht: (2023)
Nexar Dashcam Collision Prediction Dataset and Challenge
von: Moura, Daniel C., et al.
Veröffentlicht: (2025)
von: Moura, Daniel C., et al.
Veröffentlicht: (2025)
RoMA: Robust Malware Attribution via Byte-level Adversarial Training with Global Perturbations and Adversarial Consistency Regularization
von: Sun, Yuxia, et al.
Veröffentlicht: (2025)
von: Sun, Yuxia, et al.
Veröffentlicht: (2025)
Fast Image-based Neural Relighting with Translucency-Reflection Modeling
von: Zhu, Shizhan, et al.
Veröffentlicht: (2023)
von: Zhu, Shizhan, et al.
Veröffentlicht: (2023)
Contextual Interaction via Primitive-based Adversarial Training For Compositional Zero-shot Learning
von: Li, Suyi, et al.
Veröffentlicht: (2024)
von: Li, Suyi, et al.
Veröffentlicht: (2024)
MultiEdit: Advancing Instruction-based Image Editing on Diverse and Challenging Tasks
von: Li, Mingsong, et al.
Veröffentlicht: (2025)
von: Li, Mingsong, et al.
Veröffentlicht: (2025)
SARA: Structural and Adversarial Representation Alignment for Training-efficient Diffusion Models
von: Chen, Hesen, et al.
Veröffentlicht: (2025)
von: Chen, Hesen, et al.
Veröffentlicht: (2025)
Adaptive Discriminative Regularization for Visual Classification
von: Zhao, Qingsong, et al.
Veröffentlicht: (2022)
von: Zhao, Qingsong, et al.
Veröffentlicht: (2022)
LTD: Low Temperature Distillation for Gradient Masking-free Adversarial Training
von: Chen, Erh-Chung, et al.
Veröffentlicht: (2021)
von: Chen, Erh-Chung, et al.
Veröffentlicht: (2021)
VoxNeuS: Enhancing Voxel-Based Neural Surface Reconstruction via Gradient Interpolation
von: Liu, Sidun, et al.
Veröffentlicht: (2024)
von: Liu, Sidun, et al.
Veröffentlicht: (2024)
Improving Adversarial Robustness of Attribution via Implicit Regularization
von: Mehrpanah, Amir, et al.
Veröffentlicht: (2026)
von: Mehrpanah, Amir, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Super-pixel-based Approach to the Stable Interpretation of Neural Networks
von: Gong, Shizhan, et al.
Veröffentlicht: (2024) -
Kernel-based Unsupervised Embedding Alignment for Enhanced Visual Representation in Vision-language Models
von: Gong, Shizhan, et al.
Veröffentlicht: (2025) -
Sparse Domain Transfer via Elastic Net Regularization
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024) -
Interpretation of Neural Networks is Susceptible to Universal Adversarial Perturbations
von: Oskouie, Haniyeh Ehsani, et al.
Veröffentlicht: (2022) -
Concepts from Representations: Post-hoc Concept Bottleneck Models via Sparse Decomposition of Visual Representations
von: Gong, Shizhan, et al.
Veröffentlicht: (2026)