Visual Error Patterns in Multi-Modal AI: A Statistical Approach
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Wang, Ching-Yi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Promoting AI Equity in Science: Generalized Domain Prompt Learning for Accessible VLM Research
von: Cao, Qinglong, et al.
Veröffentlicht: (2024)
von: Cao, Qinglong, et al.
Veröffentlicht: (2024)
BiDepth: A Bidirectional-Depth Neural Network for Spatio-Temporal Prediction
von: Ehsani, Sina, et al.
Veröffentlicht: (2025)
von: Ehsani, Sina, et al.
Veröffentlicht: (2025)
Provable Contrastive Continual Learning
von: Wen, Yichen, et al.
Veröffentlicht: (2024)
von: Wen, Yichen, et al.
Veröffentlicht: (2024)
Massimo: Public Queue Monitoring and Management using Mass-Spring Model
von: Kumar, Abhijeet, et al.
Veröffentlicht: (2024)
von: Kumar, Abhijeet, et al.
Veröffentlicht: (2024)
KM-GPT: An Automated Pipeline for Reconstructing Individual Patient Data from Kaplan-Meier Plots
von: Zhao, Yao, et al.
Veröffentlicht: (2025)
von: Zhao, Yao, et al.
Veröffentlicht: (2025)
AlphaEarth Satellite Embeddings for Modelling Climate Sensitive Diseases Towards Global Health Resilience
von: Nazir, Usman, et al.
Veröffentlicht: (2026)
von: Nazir, Usman, et al.
Veröffentlicht: (2026)
TinyBayes: Closed-Form Bayesian Inference via Jacobi Prior for Real-Time Image Classification on Edge Devices
von: Sardar, Shouvik, et al.
Veröffentlicht: (2026)
von: Sardar, Shouvik, et al.
Veröffentlicht: (2026)
Style-Based Neural Architectures for Real-Time Weather Classification
von: Ouattara, Hamed, et al.
Veröffentlicht: (2026)
von: Ouattara, Hamed, et al.
Veröffentlicht: (2026)
Deep Learning Approaches with Explainable AI for Differentiating Alzheimer Disease and Mild Cognitive Impairment
von: Mostafa, Fahad, et al.
Veröffentlicht: (2025)
von: Mostafa, Fahad, et al.
Veröffentlicht: (2025)
A Large-Scale Benchmark of Cross-Modal Learning for Histology and Gene Expression in Spatial Transcriptomics
von: Gindra, Rushin H., et al.
Veröffentlicht: (2025)
von: Gindra, Rushin H., et al.
Veröffentlicht: (2025)
Understanding Learning with Sliced-Wasserstein Requires Rethinking Informative Slices
von: Tran, Huy, et al.
Veröffentlicht: (2024)
von: Tran, Huy, et al.
Veröffentlicht: (2024)
Diffusion-Based Cross-Modal Feature Extraction for Multi-Label Classification
von: Lan, Tian, et al.
Veröffentlicht: (2025)
von: Lan, Tian, et al.
Veröffentlicht: (2025)
TASTE: A Designer-Annotated Multi-Dimensional Preference Dataset for AI-Generated Graphic Design
von: Zhu, Haonan, et al.
Veröffentlicht: (2026)
von: Zhu, Haonan, et al.
Veröffentlicht: (2026)
Visual Explanations of Image-Text Representations via Multi-Modal Information Bottleneck Attribution
von: Wang, Ying, et al.
Veröffentlicht: (2023)
von: Wang, Ying, et al.
Veröffentlicht: (2023)
AI-Assisted Decision-Making for Clinical Assessment of Auto-Segmented Contour Quality
von: Wang, Biling, et al.
Veröffentlicht: (2025)
von: Wang, Biling, et al.
Veröffentlicht: (2025)
Statistical Edge Detection And UDF Learning For Shape Representation
von: Foy, Virgile, et al.
Veröffentlicht: (2024)
von: Foy, Virgile, et al.
Veröffentlicht: (2024)
ColonScopeX: Leveraging Explainable Expert Systems with Multimodal Data for Improved Early Diagnosis of Colorectal Cancer
von: Sikora, Natalia, et al.
Veröffentlicht: (2025)
von: Sikora, Natalia, et al.
Veröffentlicht: (2025)
Real-Time Localization and Bimodal Point Pattern Analysis of Palms Using UAV Imagery
von: Cui, Kangning, et al.
Veröffentlicht: (2024)
von: Cui, Kangning, et al.
Veröffentlicht: (2024)
FedDiff: Diffusion Model Driven Federated Learning for Multi-Modal and Multi-Clients
von: Li, DaiXun, et al.
Veröffentlicht: (2023)
von: Li, DaiXun, et al.
Veröffentlicht: (2023)
OpenViewer: Openness-Aware Multi-View Learning
von: Du, Shide, et al.
Veröffentlicht: (2024)
von: Du, Shide, et al.
Veröffentlicht: (2024)
Multimodal Whole Slide Foundation Model for Pathology
von: Ding, Tong, et al.
Veröffentlicht: (2024)
von: Ding, Tong, et al.
Veröffentlicht: (2024)
Multi-Modal Video Feature Extraction for Popularity Prediction
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
Jailbreaking Vision-Language Models Through the Visual Modality
von: Azulay, Aharon, et al.
Veröffentlicht: (2026)
von: Azulay, Aharon, et al.
Veröffentlicht: (2026)
Visual Spatial Learning: Single-Field Spatial Interpolation Using Convolutional Neural Networks
von: Tinoco, Daniel, et al.
Veröffentlicht: (2026)
von: Tinoco, Daniel, et al.
Veröffentlicht: (2026)
Human and AI Perceptual Differences in Image Classification Errors
von: Liu, Minghao, et al.
Veröffentlicht: (2023)
von: Liu, Minghao, et al.
Veröffentlicht: (2023)
Visual Modality Prompt for Adapting Vision-Language Object Detectors
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings
von: Li, Zijie, et al.
Veröffentlicht: (2026)
von: Li, Zijie, et al.
Veröffentlicht: (2026)
Learnable Expansion of Graph Operators for Multi-Modal Feature Fusion
von: Ding, Dexuan, et al.
Veröffentlicht: (2024)
von: Ding, Dexuan, et al.
Veröffentlicht: (2024)
A Survey on Cache Methods in Diffusion Models: Toward Efficient Multi-Modal Generation
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
von: Liu, Jiacheng, et al.
Veröffentlicht: (2025)
Multi-Modal Adapter for Vision-Language Models
von: Seputis, Dominykas, et al.
Veröffentlicht: (2024)
von: Seputis, Dominykas, et al.
Veröffentlicht: (2024)
VIFO: Visual Feature Empowered Multivariate Time Series Forecasting with Cross-Modal Fusion
von: Wang, Yanlong, et al.
Veröffentlicht: (2025)
von: Wang, Yanlong, et al.
Veröffentlicht: (2025)
End-to-End Multi-Modal Diffusion Mamba
von: Lu, Chunhao, et al.
Veröffentlicht: (2025)
von: Lu, Chunhao, et al.
Veröffentlicht: (2025)
Multi-Modal Multi-Agent Reinforcement Learning for Radiology Report Generation
von: Baba, Kaito, et al.
Veröffentlicht: (2026)
von: Baba, Kaito, et al.
Veröffentlicht: (2026)
StitchFusion: Weaving Any Visual Modalities to Enhance Multimodal Semantic Segmentation
von: Li, Bingyu, et al.
Veröffentlicht: (2024)
von: Li, Bingyu, et al.
Veröffentlicht: (2024)
MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering
von: Li, Xu, et al.
Veröffentlicht: (2025)
von: Li, Xu, et al.
Veröffentlicht: (2025)
ProJudge: A Multi-Modal Multi-Discipline Benchmark and Instruction-Tuning Dataset for MLLM-based Process Judges
von: Ai, Jiaxin, et al.
Veröffentlicht: (2025)
von: Ai, Jiaxin, et al.
Veröffentlicht: (2025)
Detecting Dataset Bias in Medical AI: A Generalized and Modality-Agnostic Auditing Framework
von: Drenkow, Nathan, et al.
Veröffentlicht: (2025)
von: Drenkow, Nathan, et al.
Veröffentlicht: (2025)
Visual Knowledge in the Big Model Era: Retrospect and Prospect
von: Wang, Wenguan, et al.
Veröffentlicht: (2024)
von: Wang, Wenguan, et al.
Veröffentlicht: (2024)
MultiOOD: Scaling Out-of-Distribution Detection for Multiple Modalities
von: Dong, Hao, et al.
Veröffentlicht: (2024)
von: Dong, Hao, et al.
Veröffentlicht: (2024)
Learning Multi-Manifold Embedding for Out-Of-Distribution Detection
von: Li, Jeng-Lin, et al.
Veröffentlicht: (2024)
von: Li, Jeng-Lin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Promoting AI Equity in Science: Generalized Domain Prompt Learning for Accessible VLM Research
von: Cao, Qinglong, et al.
Veröffentlicht: (2024) -
BiDepth: A Bidirectional-Depth Neural Network for Spatio-Temporal Prediction
von: Ehsani, Sina, et al.
Veröffentlicht: (2025) -
Provable Contrastive Continual Learning
von: Wen, Yichen, et al.
Veröffentlicht: (2024) -
Massimo: Public Queue Monitoring and Management using Mass-Spring Model
von: Kumar, Abhijeet, et al.
Veröffentlicht: (2024) -
KM-GPT: An Automated Pipeline for Reconstructing Individual Patient Data from Kaplan-Meier Plots
von: Zhao, Yao, et al.
Veröffentlicht: (2025)