Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models
Fuente:
arXiv
Saved in:
| Main Authors: | Devynck, Tom, Faye, Bilal, Bouchaffra, Djamel, Lazaar, Nadjib, Azzag, Hanane, Lebbah, Mustapha |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptative Context Normalization: A Boost for Deep Learning in Image Processing
by: Faye, Bilal, et al.
Published: (2024)
by: Faye, Bilal, et al.
Published: (2024)
Context Normalization Layer with Applications
by: Faye, Bilal, et al.
Published: (2023)
by: Faye, Bilal, et al.
Published: (2023)
Enhancing Neural Network Representations with Prior Knowledge-Based Normalization
by: Faye, Bilal, et al.
Published: (2024)
by: Faye, Bilal, et al.
Published: (2024)
Lightweight Cross-Modal Representation Learning
by: Faye, Bilal, et al.
Published: (2024)
by: Faye, Bilal, et al.
Published: (2024)
OneEncoder: A Lightweight Framework for Progressive Alignment of Modalities
by: Faye, Bilal, et al.
Published: (2024)
by: Faye, Bilal, et al.
Published: (2024)
Lightweight Modular Parameter-Efficient Tuning for Open-Vocabulary Object Detection
by: Faye, Bilal, et al.
Published: (2024)
by: Faye, Bilal, et al.
Published: (2024)
Game Theory Meets Statistical Mechanics in Deep Learning Design
by: Bouchaffra, Djamel, et al.
Published: (2024)
by: Bouchaffra, Djamel, et al.
Published: (2024)
Coalition Free Energy and Adaptive Precision in Multi-Agent Cooperation
by: Bouchaffra, Djamel, et al.
Published: (2026)
by: Bouchaffra, Djamel, et al.
Published: (2026)
A Collective Variational Principle Unifying Bayesian Inference, Game Theory, and Thermodynamics
by: Bouchaffra, Djamel, et al.
Published: (2026)
by: Bouchaffra, Djamel, et al.
Published: (2026)
Prototype-Guided Diffusion: Visual Conditioning without External Memory
by: Faye, Bilal, et al.
Published: (2025)
by: Faye, Bilal, et al.
Published: (2025)
Supervised Batch Normalization
by: Faye, Bilal, et al.
Published: (2024)
by: Faye, Bilal, et al.
Published: (2024)
Value-Free Policy Optimization via Reward Partitioning
by: Faye, Bilal, et al.
Published: (2025)
by: Faye, Bilal, et al.
Published: (2025)
NeuroGame Transformer: Gibbs-Inspired Attention Driven by Game Theory and Statistical Physics
by: Bouchaffra, Djamel, et al.
Published: (2026)
by: Bouchaffra, Djamel, et al.
Published: (2026)
Unsupervised Adaptive Normalization
by: Faye, Bilal, et al.
Published: (2024)
by: Faye, Bilal, et al.
Published: (2024)
Adaptive Head Budgeting for Efficient Multi-Head Attention
by: Faye, Bilal, et al.
Published: (2026)
by: Faye, Bilal, et al.
Published: (2026)
MB-ORES: A Multi-Branch Object Reasoner for Visual Grounding in Remote Sensing
by: Radouane, Karim, et al.
Published: (2025)
by: Radouane, Karim, et al.
Published: (2025)
Trustworthy Automated Driving through Qualitative Scene Understanding and Explanations
by: Belmecheri, Nassim, et al.
Published: (2024)
by: Belmecheri, Nassim, et al.
Published: (2024)
Explainable Scene Understanding with Qualitative Representations and Graph Neural Networks
by: Belmecheri, Nassim, et al.
Published: (2025)
by: Belmecheri, Nassim, et al.
Published: (2025)
A Game Theoretic Free Energy Analysis of Higher Order Synergy in Attention Heads of Large Language Models
by: Bouchaffra, Djamel
Published: (2026)
by: Bouchaffra, Djamel
Published: (2026)
Distributed MCMC inference for Bayesian Non-Parametric Latent Block Model
by: Khoufache, Reda, et al.
Published: (2024)
by: Khoufache, Reda, et al.
Published: (2024)
VM-BeautyNet: A Synergistic Ensemble of Vision Transformer and Mamba for Facial Beauty Prediction
by: Boukhari, Djamel Eddine
Published: (2025)
by: Boukhari, Djamel Eddine
Published: (2025)
Enhancing Robustness of Vision-Language Models through Orthogonality Learning and Self-Regularization
by: Li, Jinlong, et al.
Published: (2024)
by: Li, Jinlong, et al.
Published: (2024)
FairViT-GAN: A Hybrid Vision Transformer with Adversarial Debiasing for Fair and Explainable Facial Beauty Prediction
by: Boukhari, Djamel Eddine
Published: (2025)
by: Boukhari, Djamel Eddine
Published: (2025)
SynergyNet: Fusing Generative Priors and State-Space Models for Facial Beauty Prediction
by: Boukhari, Djamel Eddine
Published: (2025)
by: Boukhari, Djamel Eddine
Published: (2025)
SVR-GS: Spatially Variant Regularization for Probabilistic Masks in 3D Gaussian Splatting
by: Taghipour, Ashkan, et al.
Published: (2025)
by: Taghipour, Ashkan, et al.
Published: (2025)
Group Orthogonalization Regularization For Vision Models Adaptation and Robustness
by: Kurtz, Yoav, et al.
Published: (2023)
by: Kurtz, Yoav, et al.
Published: (2023)
Analysing the Robustness of Vision-Language-Models to Common Corruptions
by: Usama, Muhammad, et al.
Published: (2025)
by: Usama, Muhammad, et al.
Published: (2025)
Multi-Modal Interpretability for Enhanced Localization in Vision-Language Models
by: Imran, Muhammad, et al.
Published: (2025)
by: Imran, Muhammad, et al.
Published: (2025)
Masked Depth Modeling for Spatial Perception
by: Tan, Bin, et al.
Published: (2026)
by: Tan, Bin, et al.
Published: (2026)
Vision-Language Modeling with Regularized Spatial Transformer Networks for All Weather Crosswind Landing of Aircraft
by: Pal, Debabrata, et al.
Published: (2024)
by: Pal, Debabrata, et al.
Published: (2024)
SATA: Spatial Autocorrelation Token Analysis for Enhancing the Robustness of Vision Transformers
by: Nikzad, Nick, et al.
Published: (2024)
by: Nikzad, Nick, et al.
Published: (2024)
Mask Consistency Regularization in Object Removal
by: Yuan, Hua, et al.
Published: (2025)
by: Yuan, Hua, et al.
Published: (2025)
Mamba-CNN: A Hybrid Architecture for Efficient and Accurate Facial Beauty Prediction
by: Boukhari, Djamel Eddine
Published: (2025)
by: Boukhari, Djamel Eddine
Published: (2025)
Scale-interaction transformer: a hybrid cnn-transformer model for facial beauty prediction
by: Boukhari, Djamel Eddine
Published: (2025)
by: Boukhari, Djamel Eddine
Published: (2025)
Frequency-Guided Masking for Enhanced Vision Self-Supervised Learning
by: Monsefi, Amin Karimi, et al.
Published: (2024)
by: Monsefi, Amin Karimi, et al.
Published: (2024)
Quality Text, Robust Vision: The Role of Language in Enhancing Visual Robustness of Vision-Language Models
by: Waseda, Futa, et al.
Published: (2025)
by: Waseda, Futa, et al.
Published: (2025)
Regularization by denoising: Bayesian model and Langevin-within-split Gibbs sampling
by: Faye, Elhadji C., et al.
Published: (2024)
by: Faye, Elhadji C., et al.
Published: (2024)
Towards Robust Vision Transformer via Masked Adaptive Ensemble
by: Lin, Fudong, et al.
Published: (2024)
by: Lin, Fudong, et al.
Published: (2024)
Unmasking Dementia Detection by Masking Input Gradients: A JSM Approach to Model Interpretability and Precision
by: Mustafa, Yasmine, et al.
Published: (2024)
by: Mustafa, Yasmine, et al.
Published: (2024)
Advancing Vision Transformer with Enhanced Spatial Priors
by: Fan, Qihang, et al.
Published: (2026)
by: Fan, Qihang, et al.
Published: (2026)
Similar Items
-
Adaptative Context Normalization: A Boost for Deep Learning in Image Processing
by: Faye, Bilal, et al.
Published: (2024) -
Context Normalization Layer with Applications
by: Faye, Bilal, et al.
Published: (2023) -
Enhancing Neural Network Representations with Prior Knowledge-Based Normalization
by: Faye, Bilal, et al.
Published: (2024) -
Lightweight Cross-Modal Representation Learning
by: Faye, Bilal, et al.
Published: (2024) -
OneEncoder: A Lightweight Framework for Progressive Alignment of Modalities
by: Faye, Bilal, et al.
Published: (2024)