Out of Context: Reliability in Multimodal Anomaly Detection Requires Contextual Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Wilkinghoff, Kevin, Madan, Neelu, Valverde, Juan Miguel, Nasrollahi, Kamal, Ionescu, Radu Tudor, Wisniewski, Rafal, Moeslund, Thomas B., Wang, Wenwu, Tan, Zheng-Hua |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CL-MAE: Curriculum-Learned Masked Autoencoders
by: Madan, Neelu, et al.
Published: (2023)
by: Madan, Neelu, et al.
Published: (2023)
SlotMatch: Distilling Object-Centric Representations for Unsupervised Video Segmentation
by: Grigore, Diana-Nicoleta, et al.
Published: (2025)
by: Grigore, Diana-Nicoleta, et al.
Published: (2025)
A Hyperbolic Perspective on Hierarchical Structure in Object-Centric Scene Representations
by: Madan, Neelu, et al.
Published: (2026)
by: Madan, Neelu, et al.
Published: (2026)
Video Anomaly Detection with Contours -- A Study
by: Siemon, Mia, et al.
Published: (2025)
by: Siemon, Mia, et al.
Published: (2025)
Bounding Boxes and Probabilistic Graphical Models: Video Anomaly Detection Simplified
by: Siemon, Mia, et al.
Published: (2024)
by: Siemon, Mia, et al.
Published: (2024)
PrismVAU: Prompt-Refined Inference System for Multimodal Video Anomaly Understanding
by: Erregue, Iñaki, et al.
Published: (2026)
by: Erregue, Iñaki, et al.
Published: (2026)
Only Whats Necessary: Pareto Optimal Data Minimization for Privacy Preserving Video Anomaly Detection
by: Aslam, Nazia, et al.
Published: (2026)
by: Aslam, Nazia, et al.
Published: (2026)
Machine Unlearning in Hyperbolic vs. Euclidean Multimodal Contrastive Learning: Adapting Alignment Calibration to MERU
by: Vidal, Àlex Pujol, et al.
Published: (2025)
by: Vidal, Àlex Pujol, et al.
Published: (2025)
Foundation Models for Video Understanding: A Survey
by: Madan, Neelu, et al.
Published: (2024)
by: Madan, Neelu, et al.
Published: (2024)
DSpAST: Disentangled Representations for Spatial Audio Reasoning with Large Language Models
by: Wilkinghoff, Kevin, et al.
Published: (2025)
by: Wilkinghoff, Kevin, et al.
Published: (2025)
A Novel Cartography-Based Curriculum Learning Method Applied on RoNLI: The First Romanian Natural Language Inference Corpus
by: Poesina, Eduard, et al.
Published: (2024)
by: Poesina, Eduard, et al.
Published: (2024)
Machine Unlearning in the Era of Quantum Machine Learning: An Empirical Study
by: Crivoi, Carla, et al.
Published: (2025)
by: Crivoi, Carla, et al.
Published: (2025)
Every Character Counts: From Vulnerability to Defense in Phishing Detection
by: Chiper, Maria, et al.
Published: (2025)
by: Chiper, Maria, et al.
Published: (2025)
Towards Few-Call Model Stealing via Active Self-Paced Knowledge Distillation and Diffusion-Based Image Generation
by: Hondru, Vlad, et al.
Published: (2023)
by: Hondru, Vlad, et al.
Published: (2023)
UnibucLLM: Harnessing LLMs for Automated Prediction of Item Difficulty and Response Time for Multiple-Choice Questions
by: Rogoz, Ana-Cristina, et al.
Published: (2024)
by: Rogoz, Ana-Cristina, et al.
Published: (2024)
AdaProj: Adaptively Scaled Angular Margin Subspace Projections for Anomalous Sound Detection with Auxiliary Classification Tasks
by: Wilkinghoff, Kevin
Published: (2024)
by: Wilkinghoff, Kevin
Published: (2024)
How Much Does Machine Identity Matter in Anomalous Sound Detection at Test Time?
by: Wilkinghoff, Kevin, et al.
Published: (2026)
by: Wilkinghoff, Kevin, et al.
Published: (2026)
Temporal Pooling Strategies for Training-Free Anomalous Sound Detection with Self-Supervised Audio Embeddings
by: Wilkinghoff, Kevin, et al.
Published: (2026)
by: Wilkinghoff, Kevin, et al.
Published: (2026)
CLewR: Curriculum Learning with Restarts for Machine Translation Preference Learning
by: Dragomir, Alexandra, et al.
Published: (2026)
by: Dragomir, Alexandra, et al.
Published: (2026)
Multi-Level Feature Distillation of Joint Teachers Trained on Distinct Image Datasets
by: Iordache, Adrian, et al.
Published: (2024)
by: Iordache, Adrian, et al.
Published: (2024)
Curriculum Multi-Task Self-Supervision Improves Lightweight Architectures for Onboard Satellite Hyperspectral Image Segmentation
by: Carlesso, Hugo, et al.
Published: (2025)
by: Carlesso, Hugo, et al.
Published: (2025)
MTL-MAD: Multi-Task Learners are Effective Medical Anomaly Detectors
by: Bercean, Bogdan Alexandru, et al.
Published: (2026)
by: Bercean, Bogdan Alexandru, et al.
Published: (2026)
Large Multimodal Models for Low-Resource Languages: A Survey
by: Lupascu, Marian, et al.
Published: (2025)
by: Lupascu, Marian, et al.
Published: (2025)
Quantization-Based Score Calibration for Few-Shot Keyword Spotting with Dynamic Time Warping in Noisy Environments
by: Wilkinghoff, Kevin, et al.
Published: (2025)
by: Wilkinghoff, Kevin, et al.
Published: (2025)
VQPP: Video Query Performance Prediction Benchmark
by: Lutu, Adrian Catalin, et al.
Published: (2026)
by: Lutu, Adrian Catalin, et al.
Published: (2026)
PRNU-Bench: A Novel Benchmark and Model for PRNU-Based Camera Identification
by: Croitoru, Florinel Alin, et al.
Published: (2025)
by: Croitoru, Florinel Alin, et al.
Published: (2025)
Task-Informed Anti-Curriculum by Masking Improves Downstream Performance on Text
by: Jarca, Andrei, et al.
Published: (2025)
by: Jarca, Andrei, et al.
Published: (2025)
RoDia: A New Dataset for Romanian Dialect Identification from Speech
by: Rotaru, Codrut, et al.
Published: (2023)
by: Rotaru, Codrut, et al.
Published: (2023)
CBM: Curriculum by Masking
by: Jarca, Andrei, et al.
Published: (2024)
by: Jarca, Andrei, et al.
Published: (2024)
Cascaded Cross-Modal Transformer for Audio-Textual Classification
by: Ristea, Nicolae-Catalin, et al.
Published: (2024)
by: Ristea, Nicolae-Catalin, et al.
Published: (2024)
Moment matching based reduced closed-loop design to achieve asymptotic performance
by: Ionescu, Tudor C.
Published: (2024)
by: Ionescu, Tudor C.
Published: (2024)
Balancing Privacy and Action Performance: A Penalty-Driven Approach to Image Anonymization
by: Aslam, Nazia, et al.
Published: (2025)
by: Aslam, Nazia, et al.
Published: (2025)
Rip Current Segmentation: A Novel Benchmark and YOLOv8 Baseline Results
by: Dumitriu, Andrei, et al.
Published: (2025)
by: Dumitriu, Andrei, et al.
Published: (2025)
PoPreRo: A New Dataset for Popularity Prediction of Romanian Reddit Posts
by: Rogoz, Ana-Cristina, et al.
Published: (2024)
by: Rogoz, Ana-Cristina, et al.
Published: (2024)
Learning Using Generated Privileged Information by Text-to-Image Diffusion Models
by: Menadil, Rafael-Edy, et al.
Published: (2023)
by: Menadil, Rafael-Edy, et al.
Published: (2023)
ExDDV: A New Dataset for Explainable Deepfake Detection in Video
by: Hondru, Vlad, et al.
Published: (2025)
by: Hondru, Vlad, et al.
Published: (2025)
SegMate: Asymmetric Attention-Based Lightweight Architecture for Efficient Multi-Organ Segmentation
by: Bunea, Andrei-Alexandru, et al.
Published: (2026)
by: Bunea, Andrei-Alexandru, et al.
Published: (2026)
Visual Context-Aware Person Fall Detection
by: Nagaj, Aleksander, et al.
Published: (2024)
by: Nagaj, Aleksander, et al.
Published: (2024)
SOVABench: A Vehicle Surveillance Action Retrieval Benchmark for Multimodal Large Language Models
by: Rabasseda, Oriol, et al.
Published: (2026)
by: Rabasseda, Oriol, et al.
Published: (2026)
Verifying Machine Unlearning with Explainable AI
by: Vidal, Àlex Pujol, et al.
Published: (2024)
by: Vidal, Àlex Pujol, et al.
Published: (2024)
Similar Items
-
CL-MAE: Curriculum-Learned Masked Autoencoders
by: Madan, Neelu, et al.
Published: (2023) -
SlotMatch: Distilling Object-Centric Representations for Unsupervised Video Segmentation
by: Grigore, Diana-Nicoleta, et al.
Published: (2025) -
A Hyperbolic Perspective on Hierarchical Structure in Object-Centric Scene Representations
by: Madan, Neelu, et al.
Published: (2026) -
Video Anomaly Detection with Contours -- A Study
by: Siemon, Mia, et al.
Published: (2025) -
Bounding Boxes and Probabilistic Graphical Models: Video Anomaly Detection Simplified
by: Siemon, Mia, et al.
Published: (2024)