Sensitive Image Classification by Vision Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | He, Hanxian, Wilson, Campbell, Nguyen, Thanh Thi, Dalins, Janis |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Knowledge Distillation for Classification of Hand Images using Explainable Vision Transformers
by: Nguyen, Thanh Thi, et al.
Published: (2024)
by: Nguyen, Thanh Thi, et al.
Published: (2024)
Object Detection Approaches to Identifying Hand Images with High Forensic Values
by: Nguyen, Thanh Thi, et al.
Published: (2024)
by: Nguyen, Thanh Thi, et al.
Published: (2024)
Large Language Models for Detection of Life-Threatening Texts
by: Nguyen, Thanh Thi, et al.
Published: (2025)
by: Nguyen, Thanh Thi, et al.
Published: (2025)
Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization
by: Nguyen, Thanh Thi, et al.
Published: (2025)
by: Nguyen, Thanh Thi, et al.
Published: (2025)
GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification
by: Quang, Ngoc Bui Lam, et al.
Published: (2025)
by: Quang, Ngoc Bui Lam, et al.
Published: (2025)
Hierarchical Vision Transformer Enhanced by Graph Convolutional Network for Image Classification
by: Jiao, Haibin
Published: (2026)
by: Jiao, Haibin
Published: (2026)
Brain Tumor Segmentation in MRI Images with 3D U-Net and Contextual Transformer
by: Nguyen, Thien-Qua T., et al.
Published: (2024)
by: Nguyen, Thien-Qua T., et al.
Published: (2024)
Siamese Transformer Networks for Few-shot Image Classification
by: Jiang, Weihao, et al.
Published: (2024)
by: Jiang, Weihao, et al.
Published: (2024)
AC-MAMBASEG: An adaptive convolution and Mamba-based architecture for enhanced skin lesion segmentation
by: Nguyen, Viet-Thanh, et al.
Published: (2024)
by: Nguyen, Viet-Thanh, et al.
Published: (2024)
Probing the Efficacy of Federated Parameter-Efficient Fine-Tuning of Vision Transformers for Medical Image Classification
by: Alkhunaizi, Naif, et al.
Published: (2024)
by: Alkhunaizi, Naif, et al.
Published: (2024)
Stitching Gaps: Fusing Situated Perceptual Knowledge with Vision Transformers for High-Level Image Classification
by: Pandiani, Delfina Sol Martinez, et al.
Published: (2024)
by: Pandiani, Delfina Sol Martinez, et al.
Published: (2024)
Evaluating Deep Learning Models for African Wildlife Image Classification: From DenseNet to Vision Transformers
by: Aliyu, Lukman Jibril, et al.
Published: (2025)
by: Aliyu, Lukman Jibril, et al.
Published: (2025)
Contrastive Integrated Gradients: A Feature Attribution-Based Method for Explaining Whole Slide Image Classification
by: Vu, Anh Mai, et al.
Published: (2025)
by: Vu, Anh Mai, et al.
Published: (2025)
LangXAI: Integrating Large Vision Models for Generating Textual Explanations to Enhance Explainability in Visual Perception Tasks
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
Salient Mask-Guided Vision Transformer for Fine-Grained Classification
by: Demidov, Dmitry, et al.
Published: (2023)
by: Demidov, Dmitry, et al.
Published: (2023)
BRAIN: Bias-Mitigation Continual Learning Approach to Vision-Brain Understanding
by: Nguyen, Xuan-Bac, et al.
Published: (2025)
by: Nguyen, Xuan-Bac, et al.
Published: (2025)
HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image Classification
by: Ouyang, Shuyi, et al.
Published: (2024)
by: Ouyang, Shuyi, et al.
Published: (2024)
Examining Monitoring System: Detecting Abnormal Behavior In Online Examinations
by: Ngo, Dinh An, et al.
Published: (2024)
by: Ngo, Dinh An, et al.
Published: (2024)
SimGraph: A Unified Framework for Scene Graph-Based Image Generation and Editing
by: Vo, Thanh-Nhan, et al.
Published: (2026)
by: Vo, Thanh-Nhan, et al.
Published: (2026)
IoT Botnet Detection: Application of Vision Transformer to Classification of Network Flow Traffic
by: Wasswa, Hassan, et al.
Published: (2025)
by: Wasswa, Hassan, et al.
Published: (2025)
A Data-Centric Vision Transformer Baseline for SAR Sea Ice Classification
by: Mike-Ewewie, David, et al.
Published: (2026)
by: Mike-Ewewie, David, et al.
Published: (2026)
Zoom-shot: Fast and Efficient Unsupervised Zero-Shot Transfer of CLIP to Vision Encoders with Multimodal Loss
by: Shipard, Jordan, et al.
Published: (2024)
by: Shipard, Jordan, et al.
Published: (2024)
Disentangling Visual Transformers: Patch-level Interpretability for Image Classification
by: Jeanneret, Guillaume, et al.
Published: (2025)
by: Jeanneret, Guillaume, et al.
Published: (2025)
SPARO: Selective Attention for Robust and Compositional Transformer Encodings for Vision
by: Vani, Ankit, et al.
Published: (2024)
by: Vani, Ankit, et al.
Published: (2024)
Intersectional Fairness in Vision-Language Models for Medical Image Disease Classification
by: Zhang, Yupeng, et al.
Published: (2025)
by: Zhang, Yupeng, et al.
Published: (2025)
STER-VLM: Spatio-Temporal With Enhanced Reference Vision-Language Models
by: Nguyen-Nhu, Tinh-Anh, et al.
Published: (2025)
by: Nguyen-Nhu, Tinh-Anh, et al.
Published: (2025)
Cross-Task Multi-Branch Vision Transformer for Facial Expression and Mask Wearing Classification
by: Zhu, Armando, et al.
Published: (2024)
by: Zhu, Armando, et al.
Published: (2024)
Surformer v1: Transformer-Based Surface Classification Using Tactile and Vision Features
by: Kansana, Manish, et al.
Published: (2025)
by: Kansana, Manish, et al.
Published: (2025)
Can Biases in ImageNet Models Explain Generalization?
by: Gavrikov, Paul, et al.
Published: (2024)
by: Gavrikov, Paul, et al.
Published: (2024)
A Simple Interpretable Transformer for Fine-Grained Image Classification and Analysis
by: Paul, Dipanjyoti, et al.
Published: (2023)
by: Paul, Dipanjyoti, et al.
Published: (2023)
Bridging the Training-Deployment Gap: Gated Encoding and Multi-Scale Refinement for Efficient Quantization-Aware Image Enhancement
by: To-Thanh, Dat, et al.
Published: (2026)
by: To-Thanh, Dat, et al.
Published: (2026)
Uncovering Critical Features for Deepfake Detection through the Lottery Ticket Hypothesis
by: Amin, Lisan Al, et al.
Published: (2025)
by: Amin, Lisan Al, et al.
Published: (2025)
LightX3ECG: A Lightweight and eXplainable Deep Learning System for 3-lead Electrocardiogram Classification
by: Le, Khiem H., et al.
Published: (2022)
by: Le, Khiem H., et al.
Published: (2022)
Multimodal Contextualized Support for Enhancing Video Retrieval System
by: Nguyen-Le, Quoc-Bao, et al.
Published: (2024)
by: Nguyen-Le, Quoc-Bao, et al.
Published: (2024)
Tiny-ViT: A Compact Vision Transformer for Efficient and Explainable Potato Leaf Disease Classification
by: Mia, Shakil, et al.
Published: (2026)
by: Mia, Shakil, et al.
Published: (2026)
EfficientFSL: Enhancing Few-Shot Classification via Query-Only Tuning in Vision Transformers
by: Liao, Wenwen, et al.
Published: (2026)
by: Liao, Wenwen, et al.
Published: (2026)
KTVIC: A Vietnamese Image Captioning Dataset on the Life Domain
by: Pham, Anh-Cuong, et al.
Published: (2024)
by: Pham, Anh-Cuong, et al.
Published: (2024)
Texture Image Synthesis Using Spatial GAN Based on Vision Transformers
by: Salari, Elahe, et al.
Published: (2025)
by: Salari, Elahe, et al.
Published: (2025)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
From Specialist to Generalist: Unlocking SAM's Learning Potential on Unlabeled Medical Images
by: Vu, Vi, et al.
Published: (2026)
by: Vu, Vi, et al.
Published: (2026)
Similar Items
-
Adaptive Knowledge Distillation for Classification of Hand Images using Explainable Vision Transformers
by: Nguyen, Thanh Thi, et al.
Published: (2024) -
Object Detection Approaches to Identifying Hand Images with High Forensic Values
by: Nguyen, Thanh Thi, et al.
Published: (2024) -
Large Language Models for Detection of Life-Threatening Texts
by: Nguyen, Thanh Thi, et al.
Published: (2025) -
Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization
by: Nguyen, Thanh Thi, et al.
Published: (2025) -
GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification
by: Quang, Ngoc Bui Lam, et al.
Published: (2025)