Zero-Shot Anomaly Detection in Battery Thermal Images Using Visual Question Answering with Prior Knowledge
Fuente:
arXiv
Saved in:
| Main Authors: | Astrid, Marcella, Shabayek, Abdelrahman, Aouada, Djamila |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Training Free Zero-Shot Visual Anomaly Localization via Diffusion Inversion
by: Hicsonmez, Samet, et al.
Published: (2026)
by: Hicsonmez, Samet, et al.
Published: (2026)
VLMDiff: Leveraging Vision-Language Models for Multi-Class Anomaly Detection with Diffusion
by: Hicsonmez, Samet, et al.
Published: (2025)
by: Hicsonmez, Samet, et al.
Published: (2025)
Detecting Audio-Visual Deepfakes with Fine-Grained Inconsistencies
by: Astrid, Marcella, et al.
Published: (2024)
by: Astrid, Marcella, et al.
Published: (2024)
Motion Aware ViT-based Framework for Monocular 6-DoF Spacecraft Pose Estimation
by: Sosa, Jose, et al.
Published: (2025)
by: Sosa, Jose, et al.
Published: (2025)
Audio-Visual Deepfake Detection With Local Temporal Inconsistencies
by: Astrid, Marcella, et al.
Published: (2025)
by: Astrid, Marcella, et al.
Published: (2025)
Removing Geometric Bias in One-Class Anomaly Detection with Adaptive Feature Perturbation
by: Hermary, Romain, et al.
Published: (2025)
by: Hermary, Romain, et al.
Published: (2025)
FakeFormer: Efficient Vulnerability-Driven Transformers for Generalisable Deepfake Detection
by: Nguyen, Dat, et al.
Published: (2024)
by: Nguyen, Dat, et al.
Published: (2024)
Domain Adaptive Object Detection for Space Applications with Real-Time Constraints
by: Hicsonmez, Samet, et al.
Published: (2025)
by: Hicsonmez, Samet, et al.
Published: (2025)
ASTER: Latent Pseudo-Anomaly Generation for Unsupervised Time-Series Anomaly Detection
by: Hermary, Romain, et al.
Published: (2026)
by: Hermary, Romain, et al.
Published: (2026)
Exploiting Autoencoder's Weakness to Generate Pseudo Anomalies
by: Astrid, Marcella, et al.
Published: (2024)
by: Astrid, Marcella, et al.
Published: (2024)
Statistics-aware Audio-visual Deepfake Detector
by: Astrid, Marcella, et al.
Published: (2024)
by: Astrid, Marcella, et al.
Published: (2024)
Vulnerability-Aware Spatio-Temporal Learning for Generalizable Deepfake Video Detection
by: Nguyen, Dat, et al.
Published: (2025)
by: Nguyen, Dat, et al.
Published: (2025)
LAA-X: Unified Localized Artifact Attention for Quality-Agnostic and Generalizable Face Forgery Detection
by: Nguyen, Dat, et al.
Published: (2026)
by: Nguyen, Dat, et al.
Published: (2026)
A Hitchhikers Guide to Fine-Grained Face Forgery Detection Using Common Sense Reasoning
by: Foteinopoulou, Niki Maria, et al.
Published: (2024)
by: Foteinopoulou, Niki Maria, et al.
Published: (2024)
Annotation Free Spacecraft Detection and Segmentation using Vision Language Models
by: Hicsonmez, Samet, et al.
Published: (2026)
by: Hicsonmez, Samet, et al.
Published: (2026)
Question-Instructed Visual Descriptions for Zero-Shot Video Question Answering
by: Romero, David, et al.
Published: (2024)
by: Romero, David, et al.
Published: (2024)
Hybrid Attention for Robust RGB-T Pedestrian Detection in Real-World Conditions
by: Rathinam, Arunkumar, et al.
Published: (2024)
by: Rathinam, Arunkumar, et al.
Published: (2024)
LAA-Net: Localized Artifact Attention Network for Quality-Agnostic and Generalizable Deepfake Detection
by: Nguyen, Dat, et al.
Published: (2024)
by: Nguyen, Dat, et al.
Published: (2024)
Bridging the Synthetic-Real Gap: Supervised Domain Adaptation for Robust Spacecraft 6-DoF Pose Estimation
by: Singh, Inder Pal, et al.
Published: (2025)
by: Singh, Inder Pal, et al.
Published: (2025)
TALON: Token-Aligned Lightweight Adapters for 6-DoF Spacecraft Pose Estimation
by: Ali, Abid, et al.
Published: (2026)
by: Ali, Abid, et al.
Published: (2026)
When Unsupervised Domain Adaptation meets One-class Anomaly Detection: Addressing the Two-fold Unsupervised Curse by Leveraging Anomaly Scarcity
by: Mejri, Nesryne, et al.
Published: (2025)
by: Mejri, Nesryne, et al.
Published: (2025)
Knowledge Detection by Relevant Question and Image Attributes in Visual Question Answering
by: Ahir, Param, et al.
Published: (2023)
by: Ahir, Param, et al.
Published: (2023)
Domain Adaptation for Multi-label Image Classification: a Discriminator-free Approach
by: Singh, Inder Pal, et al.
Published: (2025)
by: Singh, Inder Pal, et al.
Published: (2025)
Investigating Prompting Techniques for Zero- and Few-Shot Visual Question Answering
by: Awal, Rabiul, et al.
Published: (2023)
by: Awal, Rabiul, et al.
Published: (2023)
Multi-label Image Classification using Adaptive Graph Convolutional Networks: from a Single Domain to Multiple Domains
by: Singh, Indel Pal, et al.
Published: (2023)
by: Singh, Indel Pal, et al.
Published: (2023)
Overcoming Language Priors for Visual Question Answering Based on Knowledge Distillation
by: Peng, Daowan, et al.
Published: (2025)
by: Peng, Daowan, et al.
Published: (2025)
Zero-Shot Image Anomaly Detection Using Generative Foundation Models
by: Abdi, Lemar, et al.
Published: (2025)
by: Abdi, Lemar, et al.
Published: (2025)
Few-Shot Image Classification and Segmentation as Visual Question Answering Using Vision-Language Models
by: Meng, Tian, et al.
Published: (2024)
by: Meng, Tian, et al.
Published: (2024)
Uncertainty-Aware Knowledge Distillation for Compact and Efficient 6DoF Pose Estimation
by: Ousalah, Nassim Ali, et al.
Published: (2025)
by: Ousalah, Nassim Ali, et al.
Published: (2025)
Combining Knowledge Graph and LLMs for Enhanced Zero-shot Visual Question Answering
by: Tao, Qian, et al.
Published: (2025)
by: Tao, Qian, et al.
Published: (2025)
MultiMAE Meets Earth Observation: Pre-training Multi-modal Multi-task Masked Autoencoders for Earth Observation Tasks
by: Sosa, Jose, et al.
Published: (2025)
by: Sosa, Jose, et al.
Published: (2025)
Enabling Training-Free Text-Based Remote Sensing Segmentation
by: Sosa, Jose, et al.
Published: (2026)
by: Sosa, Jose, et al.
Published: (2026)
Dual-Image Enhanced CLIP for Zero-Shot Anomaly Detection
by: Zhang, Zhaoxiang, et al.
Published: (2024)
by: Zhang, Zhaoxiang, et al.
Published: (2024)
Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering
by: Cocchi, Federico, et al.
Published: (2024)
by: Cocchi, Federico, et al.
Published: (2024)
Constricting Normal Latent Space for Anomaly Detection with Normal-only Training Data
by: Astrid, Marcella, et al.
Published: (2024)
by: Astrid, Marcella, et al.
Published: (2024)
Object Retrieval for Visual Question Answering with Outside Knowledge
by: Kan, Shichao, et al.
Published: (2024)
by: Kan, Shichao, et al.
Published: (2024)
Stabilizing Adversarially Learned One-Class Novelty Detection Using Pseudo Anomalies
by: Zaheer, Muhammad Zaigham, et al.
Published: (2022)
by: Zaheer, Muhammad Zaigham, et al.
Published: (2022)
Breaking the Visual Shortcuts in Multimodal Knowledge-Based Visual Question Answering
by: Lee, Dosung, et al.
Published: (2025)
by: Lee, Dosung, et al.
Published: (2025)
VisualAD: Language-Free Zero-Shot Anomaly Detection via Vision Transformer
by: Hou, Yanning, et al.
Published: (2026)
by: Hou, Yanning, et al.
Published: (2026)
VMAD: Visual-enhanced Multimodal Large Language Model for Zero-Shot Anomaly Detection
by: Deng, Huilin, et al.
Published: (2024)
by: Deng, Huilin, et al.
Published: (2024)
Similar Items
-
Training Free Zero-Shot Visual Anomaly Localization via Diffusion Inversion
by: Hicsonmez, Samet, et al.
Published: (2026) -
VLMDiff: Leveraging Vision-Language Models for Multi-Class Anomaly Detection with Diffusion
by: Hicsonmez, Samet, et al.
Published: (2025) -
Detecting Audio-Visual Deepfakes with Fine-Grained Inconsistencies
by: Astrid, Marcella, et al.
Published: (2024) -
Motion Aware ViT-based Framework for Monocular 6-DoF Spacecraft Pose Estimation
by: Sosa, Jose, et al.
Published: (2025) -
Audio-Visual Deepfake Detection With Local Temporal Inconsistencies
by: Astrid, Marcella, et al.
Published: (2025)