Regulating Model Reliance on Non-Robust Features by Smoothing Input Marginal Density
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Peiyu, Akhtar, Naveed, Shah, Mubarak, Mian, Ajmal |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Backdoor-based Explainable AI Benchmark for High Fidelity Evaluation of Attributions
von: Yang, Peiyu, et al.
Veröffentlicht: (2024)
von: Yang, Peiyu, et al.
Veröffentlicht: (2024)
Attribution-Guided Model Rectification of Unreliable Neural Network Behaviors
von: Yang, Peiyu, et al.
Veröffentlicht: (2026)
von: Yang, Peiyu, et al.
Veröffentlicht: (2026)
On Transfer-based Universal Attacks in Pure Black-box Setting
von: Jalwana, Mohammad A. A. K., et al.
Veröffentlicht: (2025)
von: Jalwana, Mohammad A. A. K., et al.
Veröffentlicht: (2025)
Safety Without Semantic Disruptions: Editing-free Safe Image Generation via Context-preserving Dual Latent Reconstruction
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
Manipulating and Mitigating Generative Model Biases without Retraining
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
Exploring Bias in over 100 Text-to-Image Generative Models
von: Vice, Jordan, et al.
Veröffentlicht: (2025)
von: Vice, Jordan, et al.
Veröffentlicht: (2025)
FastBO: Fast HPO and NAS with Adaptive Fidelity Identification
von: Jiang, Jiantong, et al.
Veröffentlicht: (2024)
von: Jiang, Jiantong, et al.
Veröffentlicht: (2024)
FiLoRA: Focus-and-Ignore LoRA for Controllable Feature Reliance
von: Chung, Hyunsuk, et al.
Veröffentlicht: (2026)
von: Chung, Hyunsuk, et al.
Veröffentlicht: (2026)
On the Fairness, Diversity and Reliability of Text-to-Image Generative Models
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
von: Vice, Jordan, et al.
Veröffentlicht: (2024)
Intriguing Properties of Input-dependent Randomized Smoothing
von: Súkeník, Peter, et al.
Veröffentlicht: (2021)
von: Súkeník, Peter, et al.
Veröffentlicht: (2021)
Multistream Network for LiDAR and Camera-based 3D Object Detection in Outdoor Scenes
von: Ibrahim, Muhammad, et al.
Veröffentlicht: (2025)
von: Ibrahim, Muhammad, et al.
Veröffentlicht: (2025)
Causal Reinforcement Learning for Complex Card Games: A Magic The Gathering Benchmark
von: Cunha, Cristiano da Costa, et al.
Veröffentlicht: (2026)
von: Cunha, Cristiano da Costa, et al.
Veröffentlicht: (2026)
M-CELS: Counterfactual Explanation for Multivariate Time Series Data Guided by Learned Saliency Maps
von: Li, Peiyu, et al.
Veröffentlicht: (2024)
von: Li, Peiyu, et al.
Veröffentlicht: (2024)
Certified Robustness for Deep Equilibrium Models via Serialized Random Smoothing
von: Gao, Weizhi, et al.
Veröffentlicht: (2024)
von: Gao, Weizhi, et al.
Veröffentlicht: (2024)
Skip Mamba Diffusion for Monocular 3D Semantic Scene Completion
von: Liang, Li, et al.
Veröffentlicht: (2025)
von: Liang, Li, et al.
Veröffentlicht: (2025)
Fast-PGM: Fast Probabilistic Graphical Model Learning and Inference
von: Jiang, Jiantong, et al.
Veröffentlicht: (2024)
von: Jiang, Jiantong, et al.
Veröffentlicht: (2024)
Margin-aware Fuzzy Rough Feature Selection: Bridging Uncertainty Characterization and Pattern Classification
von: Xu, Suping, et al.
Veröffentlicht: (2025)
von: Xu, Suping, et al.
Veröffentlicht: (2025)
Efficient and Private Marginal Reconstruction with Local Non-Negativity
von: Mullins, Brett, et al.
Veröffentlicht: (2024)
von: Mullins, Brett, et al.
Veröffentlicht: (2024)
Generative Marginalization Models
von: Liu, Sulin, et al.
Veröffentlicht: (2023)
von: Liu, Sulin, et al.
Veröffentlicht: (2023)
Reasoning Stabilization Point: A Training-Time Signal for Stable Evidence and Shortcut Reliance
von: Dhayalkar, Sahil Rajesh
Veröffentlicht: (2026)
von: Dhayalkar, Sahil Rajesh
Veröffentlicht: (2026)
From What to How: Attributing CLIP's Latent Components Reveals Unexpected Semantic Reliance
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2025)
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2025)
Certified Robustness via Dynamic Margin Maximization and Improved Lipschitz Regularization
von: Fazlyab, Mahyar, et al.
Veröffentlicht: (2023)
von: Fazlyab, Mahyar, et al.
Veröffentlicht: (2023)
Certifying Language Model Robustness with Fuzzed Randomized Smoothing: An Efficient Defense Against Backdoor Attacks
von: He, Bowei, et al.
Veröffentlicht: (2025)
von: He, Bowei, et al.
Veröffentlicht: (2025)
Breaking the Barrier: Enhanced Utility and Robustness in Smoothed DRL Agents
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
von: Sun, Chung-En, et al.
Veröffentlicht: (2024)
Certified Adversarial Robustness via Partition-based Randomized Smoothing
von: Goli, Hossein, et al.
Veröffentlicht: (2024)
von: Goli, Hossein, et al.
Veröffentlicht: (2024)
Beyond Interpretability: The Gains of Feature Monosemanticity on Model Robustness
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
Robustness-enhanced Uplift Modeling with Adversarial Feature Desensitization
von: Sun, Zexu, et al.
Veröffentlicht: (2023)
von: Sun, Zexu, et al.
Veröffentlicht: (2023)
Diffusion Models in Vision: A Survey
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2022)
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2022)
Reinforcement Learning Enabled Peer-to-Peer Energy Trading for Dairy Farms
von: Shah, Mian Ibad Ali, et al.
Veröffentlicht: (2024)
von: Shah, Mian Ibad Ali, et al.
Veröffentlicht: (2024)
Physics-integrated generative modeling using attentive planar normalizing flow based variational autoencoder
von: Akhtar, Sheikh Waqas
Veröffentlicht: (2024)
von: Akhtar, Sheikh Waqas
Veröffentlicht: (2024)
On Tuning Neural ODE for Stability, Consistency and Faster Convergence
von: Akhtar, Sheikh Waqas
Veröffentlicht: (2023)
von: Akhtar, Sheikh Waqas
Veröffentlicht: (2023)
Federated Smoothing Proximal Gradient for Quantile Regression with Non-Convex Penalties
von: Mirzaeifard, Reza, et al.
Veröffentlicht: (2024)
von: Mirzaeifard, Reza, et al.
Veröffentlicht: (2024)
Toward Robust Signed Graph Learning through Joint Input-Target Denoising
von: Wu, Junran, et al.
Veröffentlicht: (2025)
von: Wu, Junran, et al.
Veröffentlicht: (2025)
Reconcile Certified Robustness and Accuracy for DNN-based Smoothed Majority Vote Classifier
von: Jin, Gaojie, et al.
Veröffentlicht: (2025)
von: Jin, Gaojie, et al.
Veröffentlicht: (2025)
CluCERT: Certifying LLM Robustness via Clustering-Guided Denoising Smoothing
von: Wang, Zixia, et al.
Veröffentlicht: (2025)
von: Wang, Zixia, et al.
Veröffentlicht: (2025)
Imitation Learning Inputting Image Feature to Each Layer of Neural Network
von: Yamane, Koki, et al.
Veröffentlicht: (2024)
von: Yamane, Koki, et al.
Veröffentlicht: (2024)
Non-Smooth Weakly-Convex Finite-sum Coupled Compositional Optimization
von: Hu, Quanqi, et al.
Veröffentlicht: (2023)
von: Hu, Quanqi, et al.
Veröffentlicht: (2023)
Explainable and Interpretable Forecasts on Non-Smooth Multivariate Time Series for Responsible Gameplay
von: Jagirdar, Hussain, et al.
Veröffentlicht: (2025)
von: Jagirdar, Hussain, et al.
Veröffentlicht: (2025)
TangledFeatures: Robust Feature Selection in Highly Correlated Spaces
von: Sunny, Allen Daniel
Veröffentlicht: (2025)
von: Sunny, Allen Daniel
Veröffentlicht: (2025)
eMargin: Revisiting Contrastive Learning with Margin-Based Separation
von: Shamba, Abdul-Kazeem, et al.
Veröffentlicht: (2025)
von: Shamba, Abdul-Kazeem, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Backdoor-based Explainable AI Benchmark for High Fidelity Evaluation of Attributions
von: Yang, Peiyu, et al.
Veröffentlicht: (2024) -
Attribution-Guided Model Rectification of Unreliable Neural Network Behaviors
von: Yang, Peiyu, et al.
Veröffentlicht: (2026) -
On Transfer-based Universal Attacks in Pure Black-box Setting
von: Jalwana, Mohammad A. A. K., et al.
Veröffentlicht: (2025) -
Safety Without Semantic Disruptions: Editing-free Safe Image Generation via Context-preserving Dual Latent Reconstruction
von: Vice, Jordan, et al.
Veröffentlicht: (2024) -
Manipulating and Mitigating Generative Model Biases without Retraining
von: Vice, Jordan, et al.
Veröffentlicht: (2024)