Efficient and Concise Explanations for Object Detection with Gaussian-Class Activation Mapping Explainer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen, Quoc Khanh, Nguyen, Truong Thanh Hung, Nguyen, Vo Thanh Khang, Truong, Van Binh, Phan, Tuong, Cao, Hung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing the Fairness and Performance of Edge Cameras with Explainable AI
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
LangXAI: Integrating Large Vision Models for Generating Textual Explanations to Enhance Explainability in Visual Perception Tasks
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
ODExAI: A Comprehensive Object Detection Explainable AI Evaluation
von: Nguyen, Loc Phuc Truong, et al.
Veröffentlicht: (2025)
von: Nguyen, Loc Phuc Truong, et al.
Veröffentlicht: (2025)
XEdgeAI: A Human-centered Industrial Inspection Framework with Data-centric Explainable Edge AI Approach
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
Multimedia Verification Through Multi-Agent Deep Research Multimodal Large Language Models
von: Le, Huy Hoan, et al.
Veröffentlicht: (2025)
von: Le, Huy Hoan, et al.
Veröffentlicht: (2025)
Multi-Perspective Data Augmentation for Few-shot Object Detection
von: Vu, Anh-Khoa Nguyen, et al.
Veröffentlicht: (2025)
von: Vu, Anh-Khoa Nguyen, et al.
Veröffentlicht: (2025)
Examining Monitoring System: Detecting Abnormal Behavior In Online Examinations
von: Ngo, Dinh An, et al.
Veröffentlicht: (2024)
von: Ngo, Dinh An, et al.
Veröffentlicht: (2024)
WaveletGaussian: Wavelet-domain Diffusion for Sparse-view 3D Gaussian Object Reconstruction
von: Nguyen, Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Hung, et al.
Veröffentlicht: (2025)
From Coarse to Fine: Learnable Discrete Wavelet Transforms for Efficient 3D Gaussian Splatting
von: Nguyen, Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Hung, et al.
Veröffentlicht: (2025)
XAI-Enhanced Semantic Segmentation Models for Visual Quality Inspection
von: Clement, Tobias, et al.
Veröffentlicht: (2024)
von: Clement, Tobias, et al.
Veröffentlicht: (2024)
Learnable Multi-level Discrete Wavelet Transforms for 3D Gaussian Splatting Frequency Modulation
von: Nguyen, Hung, et al.
Veröffentlicht: (2026)
von: Nguyen, Hung, et al.
Veröffentlicht: (2026)
LGCA: Enhancing Semantic Representation via Progressive Expansion
von: Cao, Thanh Hieu, et al.
Veröffentlicht: (2025)
von: Cao, Thanh Hieu, et al.
Veröffentlicht: (2025)
DWTGS: Rethinking Frequency Regularization for Sparse-view 3D Gaussian Splatting
von: Nguyen, Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Hung, et al.
Veröffentlicht: (2025)
Hierarchical Neural Collapse Detection Transformer for Class Incremental Object Detection
von: Pham, Duc Thanh, et al.
Veröffentlicht: (2025)
von: Pham, Duc Thanh, et al.
Veröffentlicht: (2025)
Self-supervised Video Object Segmentation with Distillation Learning of Deformable Attention
von: Truong, Quang-Trung, et al.
Veröffentlicht: (2024)
von: Truong, Quang-Trung, et al.
Veröffentlicht: (2024)
Variational Quantum Rainbow Deep Q-Network for Optimizing Resource Allocation Problem
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2025)
A Two-Stage, Object-Centric Deep Learning Framework for Robust Exam Cheating Detection
von: Le, Van-Truong, et al.
Veröffentlicht: (2026)
von: Le, Van-Truong, et al.
Veröffentlicht: (2026)
Contestable Multi-Agent Debate with Arena-based Argumentative Computation for Multimedia Verification
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2026)
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2026)
CT to PET Translation: A Large-scale Dataset and Domain-Knowledge-Guided Diffusion Approach
von: Nguyen, Dac Thai, et al.
Veröffentlicht: (2024)
von: Nguyen, Dac Thai, et al.
Veröffentlicht: (2024)
DWTNeRF: Boosting Few-shot Neural Radiance Fields via Discrete Wavelet Transform
von: Nguyen, Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Hung, et al.
Veröffentlicht: (2025)
Adaptive Cache Enhancement for Test-Time Adaptation of Vision-Language Models
von: Nguyen, Khanh-Binh, et al.
Veröffentlicht: (2025)
von: Nguyen, Khanh-Binh, et al.
Veröffentlicht: (2025)
Motion2Meaning: A Clinician-Centered Framework for Contestable LLM in Parkinson's Disease Gait Interpretation
von: Nguyen, Loc Phuc Truong, et al.
Veröffentlicht: (2025)
von: Nguyen, Loc Phuc Truong, et al.
Veröffentlicht: (2025)
Multi-Dimensional Model Integrity and Responsibility Assessment Index and Scoring Framework
von: Nguyen, Phuc Truong Loc, et al.
Veröffentlicht: (2026)
von: Nguyen, Phuc Truong Loc, et al.
Veröffentlicht: (2026)
ViCocktail: Automated Multi-Modal Data Collection for Vietnamese Audio-Visual Speech Recognition
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thai-Binh, et al.
Veröffentlicht: (2025)
Enhancing YOLOv11n for Reliable Child Detection in Noisy Surveillance Footage
von: Tran, Khanh Linh, et al.
Veröffentlicht: (2026)
von: Tran, Khanh Linh, et al.
Veröffentlicht: (2026)
AC-MAMBASEG: An adaptive convolution and Mamba-based architecture for enhanced skin lesion segmentation
von: Nguyen, Viet-Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Viet-Thanh, et al.
Veröffentlicht: (2024)
Privacy-Preserving Multi-Stage Fall Detection Framework with Semi-supervised Federated Learning and Robotic Vision Confirmation
von: Azghadi, Seyed Alireza Rahimi, et al.
Veröffentlicht: (2025)
von: Azghadi, Seyed Alireza Rahimi, et al.
Veröffentlicht: (2025)
Multi-view Action Recognition via Directed Gromov-Wasserstein Discrepancy
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2024)
von: Nguyen, Hoang-Quan, et al.
Veröffentlicht: (2024)
Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation
von: Nguyen, Huu Tien, et al.
Veröffentlicht: (2025)
von: Nguyen, Huu Tien, et al.
Veröffentlicht: (2025)
ESRPCB: an Edge guided Super-Resolution model and Ensemble learning for tiny Printed Circuit Board Defect detection
von: HoangVan, Xiem, et al.
Veröffentlicht: (2025)
von: HoangVan, Xiem, et al.
Veröffentlicht: (2025)
LP-OVOD: Open-Vocabulary Object Detection by Linear Probing
von: Pham, Chau, et al.
Veröffentlicht: (2023)
von: Pham, Chau, et al.
Veröffentlicht: (2023)
A Deep-Learning Framework for Land-Sliding Classification from Remote Sensing Image
von: Tang, Hieu, et al.
Veröffentlicht: (2025)
von: Tang, Hieu, et al.
Veröffentlicht: (2025)
MACeIP: A Multimodal Ambient Context-enriched Intelligence Platform in Smart Cities
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024)
IQBench: How "Smart'' Are Vision-Language Models? A Study with Human IQ Tests
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2025)
von: Pham, Tan-Hanh, et al.
Veröffentlicht: (2025)
SEMT: Static-Expansion-Mesh Transformer Network Architecture for Remote Sensing Image Captioning
von: Truong, Khang, et al.
Veröffentlicht: (2025)
von: Truong, Khang, et al.
Veröffentlicht: (2025)
Semi-Supervised Semantic Segmentation using Redesigned Self-Training for White Blood Cells
von: Luu, Vinh Quoc, et al.
Veröffentlicht: (2024)
von: Luu, Vinh Quoc, et al.
Veröffentlicht: (2024)
Towards Robust and Fair Vision Learning in Open-World Environments
von: Truong, Thanh-Dat
Veröffentlicht: (2024)
von: Truong, Thanh-Dat
Veröffentlicht: (2024)
DiFlow-TTS: Compact and Low-Latency Zero-Shot Text-to-Speech with Factorized Discrete Flow Matching
von: Nguyen, Ngoc-Son, et al.
Veröffentlicht: (2025)
von: Nguyen, Ngoc-Son, et al.
Veröffentlicht: (2025)
Multimodal Contextualized Support for Enhancing Video Retrieval System
von: Nguyen-Le, Quoc-Bao, et al.
Veröffentlicht: (2024)
von: Nguyen-Le, Quoc-Bao, et al.
Veröffentlicht: (2024)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
von: Tran, Quoc-Khang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Enhancing the Fairness and Performance of Edge Cameras with Explainable AI
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024) -
LangXAI: Integrating Large Vision Models for Generating Textual Explanations to Enhance Explainability in Visual Perception Tasks
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024) -
ODExAI: A Comprehensive Object Detection Explainable AI Evaluation
von: Nguyen, Loc Phuc Truong, et al.
Veröffentlicht: (2025) -
XEdgeAI: A Human-centered Industrial Inspection Framework with Data-centric Explainable Edge AI Approach
von: Nguyen, Truong Thanh Hung, et al.
Veröffentlicht: (2024) -
Multimedia Verification Through Multi-Agent Deep Research Multimodal Large Language Models
von: Le, Huy Hoan, et al.
Veröffentlicht: (2025)