Explaining Black-box Model Predictions via Two-level Nested Feature Attributions with Consistency Property
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yoshikawa, Yuya, Kimura, Masanari, Shimizu, Ryotaro, Saito, Yuki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Transferring Visual Explainability of Self-Explaining Models to Prediction-Only Models without Additional Training
von: Yoshikawa, Yuya, et al.
Veröffentlicht: (2025)
von: Yoshikawa, Yuya, et al.
Veröffentlicht: (2025)
A Fashion Item Recommendation Model in Hyperbolic Space
von: Shimizu, Ryotaro, et al.
Veröffentlicht: (2024)
von: Shimizu, Ryotaro, et al.
Veröffentlicht: (2024)
Masked Language Prompting for Generative Data Augmentation in Few-shot Fashion Style Recognition
von: Hirakawa, Yuki, et al.
Veröffentlicht: (2025)
von: Hirakawa, Yuki, et al.
Veröffentlicht: (2025)
Explaining Caption-Image Interactions in CLIP Models with Second-Order Attributions
von: Möller, Lucas, et al.
Veröffentlicht: (2024)
von: Möller, Lucas, et al.
Veröffentlicht: (2024)
Automated Black-box Prompt Engineering for Personalized Text-to-Image Generation
von: He, Yutong, et al.
Veröffentlicht: (2024)
von: He, Yutong, et al.
Veröffentlicht: (2024)
Reference-Free Image Quality Assessment for Virtual Try-On via Human Feedback
von: Hirakawa, Yuki, et al.
Veröffentlicht: (2026)
von: Hirakawa, Yuki, et al.
Veröffentlicht: (2026)
An Empirical Analysis of GPT-4V's Performance on Fashion Aesthetic Evaluation
von: Hirakawa, Yuki, et al.
Veröffentlicht: (2024)
von: Hirakawa, Yuki, et al.
Veröffentlicht: (2024)
Multi-modal, Multi-task, Multi-criteria Automatic Evaluation with Vision Language Models
von: Ohi, Masanari, et al.
Veröffentlicht: (2024)
von: Ohi, Masanari, et al.
Veröffentlicht: (2024)
BSPA: Exploring Black-box Stealthy Prompt Attacks against Image Generators
von: Tian, Yu, et al.
Veröffentlicht: (2024)
von: Tian, Yu, et al.
Veröffentlicht: (2024)
Attributed Synthetic Data Generation for Zero-shot Domain-specific Image Classification
von: Wang, Shijian, et al.
Veröffentlicht: (2025)
von: Wang, Shijian, et al.
Veröffentlicht: (2025)
VideoAVE: A Multi-Attribute Video-to-Text Attribute Value Extraction Dataset and Benchmark Models
von: Cheng, Ming, et al.
Veröffentlicht: (2025)
von: Cheng, Ming, et al.
Veröffentlicht: (2025)
Fashionability-Enhancing Outfit Image Editing with Conditional Diffusion Models
von: Qin, Qice, et al.
Veröffentlicht: (2024)
von: Qin, Qice, et al.
Veröffentlicht: (2024)
Prediction Exposes Your Face: Black-box Model Inversion via Prediction Alignment
von: Liu, Yufan, et al.
Veröffentlicht: (2024)
von: Liu, Yufan, et al.
Veröffentlicht: (2024)
Decompose and Compare Consistency: Measuring VLMs' Answer Reliability via Task-Decomposition Consistency Comparison
von: Yang, Qian, et al.
Veröffentlicht: (2024)
von: Yang, Qian, et al.
Veröffentlicht: (2024)
Little Data, Big Impact: Privacy-Aware Visual Language Models via Minimal Tuning
von: Samson, Laurens, et al.
Veröffentlicht: (2024)
von: Samson, Laurens, et al.
Veröffentlicht: (2024)
MolSight: Molecular Property Prediction with Images
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2026)
von: Baranwal, Aaditya, et al.
Veröffentlicht: (2026)
Explanation-based Training with Differentiable Insertion/Deletion Metric-aware Regularizers
von: Yoshikawa, Yuya, et al.
Veröffentlicht: (2023)
von: Yoshikawa, Yuya, et al.
Veröffentlicht: (2023)
Improving Text-to-Image Consistency via Automatic Prompt Optimization
von: Mañas, Oscar, et al.
Veröffentlicht: (2024)
von: Mañas, Oscar, et al.
Veröffentlicht: (2024)
CAVE: Detecting and Explaining Commonsense Anomalies in Visual Environments
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025)
von: Bhagwatkar, Rishika, et al.
Veröffentlicht: (2025)
Do Vision Encoders Truly Explain Object Hallucination?: Mitigating Object Hallucination via Simple Fine-Grained CLIPScore
von: Oh, Hongseok, et al.
Veröffentlicht: (2025)
von: Oh, Hongseok, et al.
Veröffentlicht: (2025)
GLEAM: Learning to Match and Explain in Cross-View Geo-Localization
von: Lu, Xudong, et al.
Veröffentlicht: (2025)
von: Lu, Xudong, et al.
Veröffentlicht: (2025)
Attribution Analysis Meets Model Editing: Advancing Knowledge Correction in Vision Language Models with VisEdit
von: Chen, Qizhou, et al.
Veröffentlicht: (2024)
von: Chen, Qizhou, et al.
Veröffentlicht: (2024)
Linguistic Binding in Diffusion Models: Enhancing Attribute Correspondence through Attention Map Alignment
von: Rassin, Royi, et al.
Veröffentlicht: (2023)
von: Rassin, Royi, et al.
Veröffentlicht: (2023)
From Heads to Neurons: Causal Attribution and Steering in Multi-Task Vision-Language Models
von: Wang, Qidong, et al.
Veröffentlicht: (2026)
von: Wang, Qidong, et al.
Veröffentlicht: (2026)
Evaluating Fairness in Large Vision-Language Models Across Diverse Demographic Attributes and Prompts
von: Wu, Xuyang, et al.
Veröffentlicht: (2024)
von: Wu, Xuyang, et al.
Veröffentlicht: (2024)
MM-Interleaved: Interleaved Image-Text Generative Modeling via Multi-modal Feature Synchronizer
von: Tian, Changyao, et al.
Veröffentlicht: (2024)
von: Tian, Changyao, et al.
Veröffentlicht: (2024)
Generative Visual Commonsense Answering and Explaining with Generative Scene Graph Constructing
von: Yuan, Fan, et al.
Veröffentlicht: (2025)
von: Yuan, Fan, et al.
Veröffentlicht: (2025)
VinaBench: Benchmark for Faithful and Consistent Visual Narratives
von: Gao, Silin, et al.
Veröffentlicht: (2025)
von: Gao, Silin, et al.
Veröffentlicht: (2025)
Look & Mark: Leveraging Radiologist Eye Fixations and Bounding boxes in Multimodal Large Language Models for Chest X-ray Report Generation
von: Kim, Yunsoo, et al.
Veröffentlicht: (2025)
von: Kim, Yunsoo, et al.
Veröffentlicht: (2025)
Less is More: Efficient Black-box Attribution via Minimal Interpretable Subset Selection
von: Chen, Ruoyu, et al.
Veröffentlicht: (2025)
von: Chen, Ruoyu, et al.
Veröffentlicht: (2025)
MatFormer: Nested Transformer for Elastic Inference
von: Devvrit, et al.
Veröffentlicht: (2023)
von: Devvrit, et al.
Veröffentlicht: (2023)
Learning How To Ask: Cycle-Consistency Refines Prompts in Multimodal Foundation Models
von: Diesendruck, Maurice, et al.
Veröffentlicht: (2024)
von: Diesendruck, Maurice, et al.
Veröffentlicht: (2024)
CAIRe: Cultural Attribution of Images by Retrieval-Augmented Evaluation
von: Yayavaram, Arnav, et al.
Veröffentlicht: (2025)
von: Yayavaram, Arnav, et al.
Veröffentlicht: (2025)
Insight-A: Attribution-aware for Multimodal Misinformation Detection
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
von: Wu, Junjie, et al.
Veröffentlicht: (2025)
$\left|\,\circlearrowright\,\boxed{\text{BUS}}\,\right|$: A Large and Diverse Multimodal Benchmark for evaluating the ability of Vision-Language Models to understand Rebus Puzzles
von: Das, Trishanu, et al.
Veröffentlicht: (2025)
von: Das, Trishanu, et al.
Veröffentlicht: (2025)
ICON: Improving Inter-Report Consistency in Radiology Report Generation via Lesion-aware Mixup Augmentation
von: Hou, Wenjun, et al.
Veröffentlicht: (2024)
von: Hou, Wenjun, et al.
Veröffentlicht: (2024)
Unblocking Fine-Grained Evaluation of Detailed Captions: An Explaining AutoRater and Critic-and-Revise Pipeline
von: Gordon, Brian, et al.
Veröffentlicht: (2025)
von: Gordon, Brian, et al.
Veröffentlicht: (2025)
See, Explain, and Intervene: A Few-Shot Multimodal Agent Framework for Hateful Meme Moderation
von: Rizwan, Naquee, et al.
Veröffentlicht: (2026)
von: Rizwan, Naquee, et al.
Veröffentlicht: (2026)
Hierarchical Document Parsing via Large Margin Feature Matching and Heuristics
von: Kiet, Duong Anh
Veröffentlicht: (2025)
von: Kiet, Duong Anh
Veröffentlicht: (2025)
Intriguing Properties of Large Language and Vision Models
von: Lee, Young-Jun, et al.
Veröffentlicht: (2024)
von: Lee, Young-Jun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Transferring Visual Explainability of Self-Explaining Models to Prediction-Only Models without Additional Training
von: Yoshikawa, Yuya, et al.
Veröffentlicht: (2025) -
A Fashion Item Recommendation Model in Hyperbolic Space
von: Shimizu, Ryotaro, et al.
Veröffentlicht: (2024) -
Masked Language Prompting for Generative Data Augmentation in Few-shot Fashion Style Recognition
von: Hirakawa, Yuki, et al.
Veröffentlicht: (2025) -
Explaining Caption-Image Interactions in CLIP Models with Second-Order Attributions
von: Möller, Lucas, et al.
Veröffentlicht: (2024) -
Automated Black-box Prompt Engineering for Personalized Text-to-Image Generation
von: He, Yutong, et al.
Veröffentlicht: (2024)