Hierarchical Invariance for Robust and Interpretable Vision Tasks at Larger Scales
Fuente:
arXiv
Saved in:
| Main Authors: | Qi, Shuren, Zhang, Yushu, Wang, Chao, Xia, Zhihua, Cao, Xiaochun, Weng, Jian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Representing Noisy Image Without Denoising
by: Qi, Shuren, et al.
Published: (2023)
by: Qi, Shuren, et al.
Published: (2023)
Spatial-Frequency Discriminability for Revealing Adversarial Perturbations
by: Wang, Chao, et al.
Published: (2023)
by: Wang, Chao, et al.
Published: (2023)
Less is More: Fewer Interpretable Region via Submodular Subset Selection
by: Chen, Ruoyu, et al.
Published: (2024)
by: Chen, Ruoyu, et al.
Published: (2024)
Hierarchical Multi-Graphs Learning for Robust Group Re-Identification
by: Liu, Ruiqi, et al.
Published: (2024)
by: Liu, Ruiqi, et al.
Published: (2024)
Interpreting Neurons in Deep Vision Networks with Language Models
by: Bai, Nicholas, et al.
Published: (2024)
by: Bai, Nicholas, et al.
Published: (2024)
SGW-based Multi-Task Learning in Vision Tasks
by: Zhang, Ruiyuan, et al.
Published: (2024)
by: Zhang, Ruiyuan, et al.
Published: (2024)
Less is More: Efficient Black-box Attribution via Minimal Interpretable Subset Selection
by: Chen, Ruoyu, et al.
Published: (2025)
by: Chen, Ruoyu, et al.
Published: (2025)
RViDeformer: Efficient Raw Video Denoising Transformer with a Larger Benchmark Dataset
by: Yue, Huanjing, et al.
Published: (2023)
by: Yue, Huanjing, et al.
Published: (2023)
HyperPriv-EPN: Hypergraph Learning with Privileged Knowledge for Ependymoma Prognosis
by: Yu, Shuren Gabriel, et al.
Published: (2026)
by: Yu, Shuren Gabriel, et al.
Published: (2026)
Counterfactual Image Editing
by: Pan, Yushu, et al.
Published: (2024)
by: Pan, Yushu, et al.
Published: (2024)
Intriguing Frequency Interpretation of Adversarial Robustness for CNNs and ViTs
by: Chen, Lu, et al.
Published: (2025)
by: Chen, Lu, et al.
Published: (2025)
Hierarchically Robust Zero-shot Vision-language Models
by: Dong, Junhao, et al.
Published: (2026)
by: Dong, Junhao, et al.
Published: (2026)
Naturally Computed Scale Invariance in the Residual Stream of ResNet18
by: Longon, André
Published: (2025)
by: Longon, André
Published: (2025)
Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging
by: Shen, Li, et al.
Published: (2024)
by: Shen, Li, et al.
Published: (2024)
DirMixE: Harnessing Test Agnostic Long-tail Recognition with Hierarchical Label Vartiations
by: Yang, Zhiyong, et al.
Published: (2024)
by: Yang, Zhiyong, et al.
Published: (2024)
SkillDiffuser: Interpretable Hierarchical Planning via Skill Abstractions in Diffusion-Based Task Execution
by: Liang, Zhixuan, et al.
Published: (2023)
by: Liang, Zhixuan, et al.
Published: (2023)
Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models
by: Devynck, Tom, et al.
Published: (2026)
by: Devynck, Tom, et al.
Published: (2026)
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks
by: Ding, Yuhe, et al.
Published: (2024)
by: Ding, Yuhe, et al.
Published: (2024)
Meta Invariance Defense Towards Generalizable Robustness to Unknown Adversarial Attacks
by: Zhang, Lei, et al.
Published: (2024)
by: Zhang, Lei, et al.
Published: (2024)
Unveil Inversion and Invariance in Flow Transformer for Versatile Image Editing
by: Xu, Pengcheng, et al.
Published: (2024)
by: Xu, Pengcheng, et al.
Published: (2024)
Calibrated and Robust Foundation Models for Vision-Language and Medical Image Tasks Under Distribution Shift
by: Khan, Behraj, et al.
Published: (2025)
by: Khan, Behraj, et al.
Published: (2025)
Hierarchical Concept Embedding & Pursuit for Interpretable Image Classification
by: Nguyen, Nghia, et al.
Published: (2026)
by: Nguyen, Nghia, et al.
Published: (2026)
Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders
by: Jiang, Yitong, et al.
Published: (2026)
by: Jiang, Yitong, et al.
Published: (2026)
Accelerating Augmentation Invariance Pretraining
by: Lin, Jinhong, et al.
Published: (2024)
by: Lin, Jinhong, et al.
Published: (2024)
Land Surface Temperature Super-Resolution with a Scale-Invariance-Free Neural Approach: Application to MODIS
by: Ait-Bachir, Romuald, et al.
Published: (2025)
by: Ait-Bachir, Romuald, et al.
Published: (2025)
SpecFormer: Guarding Vision Transformer Robustness via Maximum Singular Value Penalization
by: Hu, Xixu, et al.
Published: (2024)
by: Hu, Xixu, et al.
Published: (2024)
Synthesizer Based Efficient Self-Attention for Vision Tasks
by: Zhu, Guangyang, et al.
Published: (2022)
by: Zhu, Guangyang, et al.
Published: (2022)
Beyond Top Activations: Efficient and Reliable Crowdsourced Evaluation of Automated Interpretability
by: Oikarinen, Tuomas, et al.
Published: (2025)
by: Oikarinen, Tuomas, et al.
Published: (2025)
Lotus: learning-based online thermal and latency variation management for two-stage detectors on edge devices
by: Gong, Yifan, et al.
Published: (2024)
by: Gong, Yifan, et al.
Published: (2024)
CI-CBM: Class-Incremental Concept Bottleneck Model for Interpretable Continual Learning
by: Javadi, Amirhosein, et al.
Published: (2026)
by: Javadi, Amirhosein, et al.
Published: (2026)
Approximate Borderline Sampling using Granular-Ball for Classification Tasks
by: Xie, Qin, et al.
Published: (2025)
by: Xie, Qin, et al.
Published: (2025)
Uncertainty-Supervised Interpretable and Robust Evidential Segmentation
by: Li, Yuzhu, et al.
Published: (2025)
by: Li, Yuzhu, et al.
Published: (2025)
A General Framework for Robust G-Invariance in G-Equivariant Networks
by: Sanborn, Sophia, et al.
Published: (2023)
by: Sanborn, Sophia, et al.
Published: (2023)
Ask, Pose, Unite: Scaling Data Acquisition for Close Interactions with Vision Language Models
by: Bravo-Sánchez, Laura, et al.
Published: (2024)
by: Bravo-Sánchez, Laura, et al.
Published: (2024)
VLG-CBM: Training Concept Bottleneck Models with Vision-Language Guidance
by: Srivastava, Divyansh, et al.
Published: (2024)
by: Srivastava, Divyansh, et al.
Published: (2024)
Interpretable and Steerable Concept Bottleneck Sparse Autoencoders
by: Kulkarni, Akshay, et al.
Published: (2025)
by: Kulkarni, Akshay, et al.
Published: (2025)
Learning to Transform for Generalizable Instance-wise Invariance
by: Singhal, Utkarsh, et al.
Published: (2023)
by: Singhal, Utkarsh, et al.
Published: (2023)
Background Invariance Testing According to Semantic Proximity
by: Liao, Zukang, et al.
Published: (2022)
by: Liao, Zukang, et al.
Published: (2022)
Learning Conditional Invariances through Non-Commutativity
by: Chaudhuri, Abhra, et al.
Published: (2024)
by: Chaudhuri, Abhra, et al.
Published: (2024)
Tailored Transformation Invariance for Industrial Anomaly Detection
by: Schönfeld, Mariette, et al.
Published: (2025)
by: Schönfeld, Mariette, et al.
Published: (2025)
Similar Items
-
Representing Noisy Image Without Denoising
by: Qi, Shuren, et al.
Published: (2023) -
Spatial-Frequency Discriminability for Revealing Adversarial Perturbations
by: Wang, Chao, et al.
Published: (2023) -
Less is More: Fewer Interpretable Region via Submodular Subset Selection
by: Chen, Ruoyu, et al.
Published: (2024) -
Hierarchical Multi-Graphs Learning for Robust Group Re-Identification
by: Liu, Ruiqi, et al.
Published: (2024) -
Interpreting Neurons in Deep Vision Networks with Language Models
by: Bai, Nicholas, et al.
Published: (2024)