Inspecting Explainability of Transformer Models with Additional Statistical Information
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen, Hoang C., Lee, Haeil, Kim, Junmo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Effects of Mixed Sample Data Augmentation are Class Dependent
von: Lee, Haeil, et al.
Veröffentlicht: (2023)
von: Lee, Haeil, et al.
Veröffentlicht: (2023)
Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis
von: Lee, Haeil, et al.
Veröffentlicht: (2024)
von: Lee, Haeil, et al.
Veröffentlicht: (2024)
Test-Time Mixup Augmentation for Data and Class-Specific Uncertainty Estimation in Deep Learning Image Classification
von: Lee, Hansang, et al.
Veröffentlicht: (2022)
von: Lee, Hansang, et al.
Veröffentlicht: (2022)
Do Vision Models Encode Object-Level Semantic Relatedness? A Cognitive Psychology-Inspired Benchmark
von: Lee, Hansang, et al.
Veröffentlicht: (2017)
von: Lee, Hansang, et al.
Veröffentlicht: (2017)
Noisy Label Classification using Label Noise Selection with Test-Time Augmentation Cross-Entropy and NoiseMix Learning
von: Lee, Hansang, et al.
Veröffentlicht: (2022)
von: Lee, Hansang, et al.
Veröffentlicht: (2022)
IWP: Token Pruning as Implicit Weight Pruning in Large Vision Language Models
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2026)
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2026)
Frequency-Aware Token Reduction for Efficient Vision Transformer
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025)
von: Lee, Dong-Jae, et al.
Veröffentlicht: (2025)
VLM's Eye Examination: Instruct and Inspect Visual Competency of Vision Language Models
von: Hyeon-Woo, Nam, et al.
Veröffentlicht: (2024)
von: Hyeon-Woo, Nam, et al.
Veröffentlicht: (2024)
Cross-Axis Feature Fusion with Joint-Wise Motion Difference Prediction for Text-Based 3D Human Motion Editing
von: Han, Gyojin, et al.
Veröffentlicht: (2026)
von: Han, Gyojin, et al.
Veröffentlicht: (2026)
AH-OCDA: Amplitude-based Curriculum Learning and Hopfield Segmentation Model for Open Compound Domain Adaptation
von: Choi, Jaehyun, et al.
Veröffentlicht: (2024)
von: Choi, Jaehyun, et al.
Veröffentlicht: (2024)
Pygmalion Effect in Vision: Image-to-Clay Translation for Reflective Geometry Reconstruction
von: Lee, Gayoung, et al.
Veröffentlicht: (2025)
von: Lee, Gayoung, et al.
Veröffentlicht: (2025)
Learning Question-Aware Keyframe Selection with Synthetic Supervision for Video Question Answering
von: Kwon, Minchan, et al.
Veröffentlicht: (2026)
von: Kwon, Minchan, et al.
Veröffentlicht: (2026)
Self-supervised Transformation Learning for Equivariant Representations
von: Yu, Jaemyung, et al.
Veröffentlicht: (2025)
von: Yu, Jaemyung, et al.
Veröffentlicht: (2025)
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
von: Lee, Dongyeun, et al.
Veröffentlicht: (2025)
von: Lee, Dongyeun, et al.
Veröffentlicht: (2025)
MATE: Meet At The Embedding -- Connecting Images with Long Texts
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
Video Diffusion Models Excel at Tracking Similar-Looking Objects Without Supervision
von: Zhang, Chenshuang, et al.
Veröffentlicht: (2025)
von: Zhang, Chenshuang, et al.
Veröffentlicht: (2025)
SFLD: Reducing the content bias for AI-generated Image Detection
von: Gye, Seoyeon, et al.
Veröffentlicht: (2025)
von: Gye, Seoyeon, et al.
Veröffentlicht: (2025)
IMSE: Intrinsic Mixture of Spectral Experts Fine-tuning for Test-Time Adaptation
von: Baek, Sunghyun, et al.
Veröffentlicht: (2026)
von: Baek, Sunghyun, et al.
Veröffentlicht: (2026)
Refining Visual Artifacts in Diffusion Models via Explainable AI-based Flaw Activation Maps
von: Lee, Seoyeon, et al.
Veröffentlicht: (2025)
von: Lee, Seoyeon, et al.
Veröffentlicht: (2025)
ARGOS: Who, Where, and When in Agentic Multi-Camera Person Search
von: Kim, Myungchul, et al.
Veröffentlicht: (2026)
von: Kim, Myungchul, et al.
Veröffentlicht: (2026)
Transferring Visual Explainability of Self-Explaining Models to Prediction-Only Models without Additional Training
von: Yoshikawa, Yuya, et al.
Veröffentlicht: (2025)
von: Yoshikawa, Yuya, et al.
Veröffentlicht: (2025)
Text-to-image Diffusion Models in Generative AI: A Survey
von: Zhang, Chenshuang, et al.
Veröffentlicht: (2023)
von: Zhang, Chenshuang, et al.
Veröffentlicht: (2023)
Enhancing the Fairness and Performance of Edge Cameras with Explainable AI
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
InfoDisent: Explainability of Image Classification Models by Information Disentanglement
von: Struski, Łukasz, et al.
Veröffentlicht: (2024)
von: Struski, Łukasz, et al.
Veröffentlicht: (2024)
Rethinking Top Probability from Multi-view for Distracted Driver Behaviour Localization
von: Nguyen, Quang Vinh, et al.
Veröffentlicht: (2024)
von: Nguyen, Quang Vinh, et al.
Veröffentlicht: (2024)
DAM: Domain-Aware Module for Multi-Domain Dataset Condensation
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025)
von: Choi, Jaehyun, et al.
Veröffentlicht: (2025)
Brain Stroke Detection and Classification Using CT Imaging with Transformer Models and Explainable AI
von: Qari, Shomukh, et al.
Veröffentlicht: (2025)
von: Qari, Shomukh, et al.
Veröffentlicht: (2025)
LangXAI: Integrating Large Vision Models for Generating Textual Explanations to Enhance Explainability in Visual Perception Tasks
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality
von: Oh, Youngtaek, et al.
Veröffentlicht: (2024)
von: Oh, Youngtaek, et al.
Veröffentlicht: (2024)
Adaptive Knowledge Distillation for Classification of Hand Images using Explainable Vision Transformers
von: Nguyen, Thanh Thi, et al.
Veröffentlicht: (2024)
von: Nguyen, Thanh Thi, et al.
Veröffentlicht: (2024)
Exploring the Spectrum of Visio-Linguistic Compositionality and Recognition
von: Oh, Youngtaek, et al.
Veröffentlicht: (2024)
von: Oh, Youngtaek, et al.
Veröffentlicht: (2024)
Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
IPTQ-ViT: Post-Training Quantization of Non-linear Functions for Integer-only Vision Transformers
von: Kim, Gihwan, et al.
Veröffentlicht: (2025)
von: Kim, Gihwan, et al.
Veröffentlicht: (2025)
Mixed Non-linear Quantization for Vision Transformers
von: Kim, Gihwan, et al.
Veröffentlicht: (2024)
von: Kim, Gihwan, et al.
Veröffentlicht: (2024)
Knowledge-Guided Textual Reasoning for Explainable Video Anomaly Detection via LLMs
von: Lee, Hari
Veröffentlicht: (2025)
von: Lee, Hari
Veröffentlicht: (2025)
Explainable Parkinsons Disease Gait Recognition Using Multimodal RGB-D Fusion and Large Language Models
von: Alnaasan, Manar, et al.
Veröffentlicht: (2025)
von: Alnaasan, Manar, et al.
Veröffentlicht: (2025)
SwiftBrush v2: Make Your One-step Diffusion Model Better Than Its Teacher
von: Dao, Trung, et al.
Veröffentlicht: (2024)
von: Dao, Trung, et al.
Veröffentlicht: (2024)
Multi-Head Explainer: A General Framework to Improve Explainability in CNNs and Transformers
von: Sun, Bohang, et al.
Veröffentlicht: (2025)
von: Sun, Bohang, et al.
Veröffentlicht: (2025)
From Pixels to Explanations: Interpretable Diabetic Retinopathy Grading with CNN-Transformer Ensembles, Visual Explainability and Vision-Language Models
von: Khokhar, Pir Bakhsh, et al.
Veröffentlicht: (2026)
von: Khokhar, Pir Bakhsh, et al.
Veröffentlicht: (2026)
ImageNet-D: Benchmarking Neural Network Robustness on Diffusion Synthetic Object
von: Zhang, Chenshuang, et al.
Veröffentlicht: (2024)
von: Zhang, Chenshuang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
The Effects of Mixed Sample Data Augmentation are Class Dependent
von: Lee, Haeil, et al.
Veröffentlicht: (2023) -
Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis
von: Lee, Haeil, et al.
Veröffentlicht: (2024) -
Test-Time Mixup Augmentation for Data and Class-Specific Uncertainty Estimation in Deep Learning Image Classification
von: Lee, Hansang, et al.
Veröffentlicht: (2022) -
Do Vision Models Encode Object-Level Semantic Relatedness? A Cognitive Psychology-Inspired Benchmark
von: Lee, Hansang, et al.
Veröffentlicht: (2017) -
Noisy Label Classification using Label Noise Selection with Test-Time Augmentation Cross-Entropy and NoiseMix Learning
von: Lee, Hansang, et al.
Veröffentlicht: (2022)