What You See is What You Classify: Black Box Attributions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Stalder, Steven, Perraudin, Nathanaël, Achanta, Radhakrishna, Perez-Cruz, Fernando, Volpi, Michele |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
What You See is What You GAN: Rendering Every Pixel for High-Fidelity Geometry in 3D GANs
von: Trevithick, Alex, et al.
Veröffentlicht: (2024)
von: Trevithick, Alex, et al.
Veröffentlicht: (2024)
What You See is What You Ask: Evaluating Audio Descriptions
von: Kala, Divy, et al.
Veröffentlicht: (2025)
von: Kala, Divy, et al.
Veröffentlicht: (2025)
Tell What You Hear From What You See -- Video to Audio Generation Through Text
von: Liu, Xiulong, et al.
Veröffentlicht: (2024)
von: Liu, Xiulong, et al.
Veröffentlicht: (2024)
See What You Are Told: Visual Attention Sink in Large Multimodal Models
von: Kang, Seil, et al.
Veröffentlicht: (2025)
von: Kang, Seil, et al.
Veröffentlicht: (2025)
Not All Attention Heads Are What You Need: Refining CLIP's Image Representation with Attention Ablation
von: Lin, Feng, et al.
Veröffentlicht: (2025)
von: Lin, Feng, et al.
Veröffentlicht: (2025)
Revisit What You See: Revealing Visual Semantics in Vision Tokens to Guide LVLM Decoding
von: Cho, Beomsik, et al.
Veröffentlicht: (2025)
von: Cho, Beomsik, et al.
Veröffentlicht: (2025)
Decom--CAM: Tell Me What You See, In Details! Feature-Level Interpretation via Decomposition Class Activation Map
von: Yang, Yuguang, et al.
Veröffentlicht: (2023)
von: Yang, Yuguang, et al.
Veröffentlicht: (2023)
RewardFlow: Generate Images by Optimizing What You Reward
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
von: Susladkar, Onkar, et al.
Veröffentlicht: (2026)
Looking Beyond What You See: An Empirical Analysis on Subgroup Intersectional Fairness for Multi-label Chest X-ray Classification Using Social Determinants of Racial Health Inequities
von: Moukheiber, Dana, et al.
Veröffentlicht: (2024)
von: Moukheiber, Dana, et al.
Veröffentlicht: (2024)
Anatomy Might Be All You Need: Forecasting What to Do During Surgery
von: Sarwin, Gary, et al.
Veröffentlicht: (2025)
von: Sarwin, Gary, et al.
Veröffentlicht: (2025)
Multiple Different Black Box Explanations for Image Classifiers
von: Chockler, Hana, et al.
Veröffentlicht: (2023)
von: Chockler, Hana, et al.
Veröffentlicht: (2023)
What Matters to You? Towards Visual Representation Alignment for Robot Learning
von: Tian, Ran, et al.
Veröffentlicht: (2023)
von: Tian, Ran, et al.
Veröffentlicht: (2023)
Who Can See Through You? Adversarial Shielding Against VLM-Based Attribute Inference Attacks
von: Fan, Yucheng, et al.
Veröffentlicht: (2025)
von: Fan, Yucheng, et al.
Veröffentlicht: (2025)
Can You Trust What You See? Alpha Channel No-Box Attacks on Video Object Detection
von: Yi, Ariana, et al.
Veröffentlicht: (2025)
von: Yi, Ariana, et al.
Veröffentlicht: (2025)
Motivation is Something You Need
von: Acheli, Mehdi, et al.
Veröffentlicht: (2026)
von: Acheli, Mehdi, et al.
Veröffentlicht: (2026)
Count What You Want: Exemplar Identification and Few-shot Counting of Human Actions in the Wild
von: Huang, Yifeng, et al.
Veröffentlicht: (2023)
von: Huang, Yifeng, et al.
Veröffentlicht: (2023)
Biased Binary Attribute Classifiers Ignore the Majority Classes
von: Zhang, Xinyi, et al.
Veröffentlicht: (2024)
von: Zhang, Xinyi, et al.
Veröffentlicht: (2024)
Be the Change You Want to See: Revisiting Remote Sensing Change Detection Practices
von: Rolih, Blaž, et al.
Veröffentlicht: (2025)
von: Rolih, Blaž, et al.
Veröffentlicht: (2025)
KAN You See It? KANs and Sentinel for Effective and Explainable Crop Field Segmentation
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
LensWalk: Agentic Video Understanding by Planning How You See in Videos
von: Li, Keliang, et al.
Veröffentlicht: (2026)
von: Li, Keliang, et al.
Veröffentlicht: (2026)
SemVideo: Reconstructs What You Watch from Brain Activity via Hierarchical Semantic Guidance
von: Yang, Minghan, et al.
Veröffentlicht: (2026)
von: Yang, Minghan, et al.
Veröffentlicht: (2026)
What You See is (Usually) What You Get: Multimodal Prototype Networks that Abstain from Expensive Modalities
von: Bahng, Muchang, et al.
Veröffentlicht: (2025)
von: Bahng, Muchang, et al.
Veröffentlicht: (2025)
What are You Looking at? Modality Contribution in Multimodal Medical Deep Learning
von: Gapp, Christian, et al.
Veröffentlicht: (2025)
von: Gapp, Christian, et al.
Veröffentlicht: (2025)
See Before You Code: Learning Visual Priors for Spatially Aware Educational Animation Generation
von: Li, Yuejia, et al.
Veröffentlicht: (2026)
von: Li, Yuejia, et al.
Veröffentlicht: (2026)
Do You Keep an Eye on What I Ask? Mitigating Multimodal Hallucination via Attention-Guided Ensemble Decoding
von: Cho, Yeongjae, et al.
Veröffentlicht: (2025)
von: Cho, Yeongjae, et al.
Veröffentlicht: (2025)
Self-supervised learning unveils change in urban housing from street-level images
von: Stalder, Steven, et al.
Veröffentlicht: (2023)
von: Stalder, Steven, et al.
Veröffentlicht: (2023)
What Helps---and What Hurts: Bidirectional Explanations for Vision Transformers
von: Su, Qin, et al.
Veröffentlicht: (2026)
von: Su, Qin, et al.
Veröffentlicht: (2026)
Three Creates All: You Only Sample 3 Steps
von: Cai, Yuren, et al.
Veröffentlicht: (2026)
von: Cai, Yuren, et al.
Veröffentlicht: (2026)
May the Forgetting Be with You: Alternate Replay for Learning with Noisy Labels
von: Millunzi, Monica, et al.
Veröffentlicht: (2024)
von: Millunzi, Monica, et al.
Veröffentlicht: (2024)
Fake it till You Make it: Reward Modeling as Discriminative Prediction
von: Liu, Runtao, et al.
Veröffentlicht: (2025)
von: Liu, Runtao, et al.
Veröffentlicht: (2025)
Is Hyperbolic Space All You Need for Medical Anomaly Detection?
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
von: Gonzalez-Jimenez, Alvaro, et al.
Veröffentlicht: (2025)
With Limited Data for Multimodal Alignment, Let the STRUCTURE Guide You
von: Gröger, Fabian, et al.
Veröffentlicht: (2025)
von: Gröger, Fabian, et al.
Veröffentlicht: (2025)
Aligning Text to Image in Diffusion Models is Easier Than You Think
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2025)
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2025)
Camouflaged Image Synthesis Is All You Need to Boost Camouflaged Detection
von: Zhang, Haichao, et al.
Veröffentlicht: (2023)
von: Zhang, Haichao, et al.
Veröffentlicht: (2023)
Aligning Logits Generatively for Principled Black-Box Knowledge Distillation
von: Ma, Jing, et al.
Veröffentlicht: (2022)
von: Ma, Jing, et al.
Veröffentlicht: (2022)
Prompting the Unseen: Detecting Hidden Backdoors in Black-Box Models
von: Huang, Zi-Xuan, et al.
Veröffentlicht: (2024)
von: Huang, Zi-Xuan, et al.
Veröffentlicht: (2024)
Where Do You Go? Pedestrian Trajectory Prediction using Scene Features
von: Rezaei, Mohammad Ali, et al.
Veröffentlicht: (2025)
von: Rezaei, Mohammad Ali, et al.
Veröffentlicht: (2025)
You Only Submit One Image to Find the Most Suitable Generative Model
von: Zhou, Zhi, et al.
Veröffentlicht: (2024)
von: Zhou, Zhi, et al.
Veröffentlicht: (2024)
Why Are You Wrong? Counterfactual Explanations for Language Grounding with 3D Objects
von: Preintner, Tobias, et al.
Veröffentlicht: (2025)
von: Preintner, Tobias, et al.
Veröffentlicht: (2025)
ADBA:Approximation Decision Boundary Approach for Black-Box Adversarial Attacks
von: Wang, Feiyang, et al.
Veröffentlicht: (2024)
von: Wang, Feiyang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
What You See is What You GAN: Rendering Every Pixel for High-Fidelity Geometry in 3D GANs
von: Trevithick, Alex, et al.
Veröffentlicht: (2024) -
What You See is What You Ask: Evaluating Audio Descriptions
von: Kala, Divy, et al.
Veröffentlicht: (2025) -
Tell What You Hear From What You See -- Video to Audio Generation Through Text
von: Liu, Xiulong, et al.
Veröffentlicht: (2024) -
See What You Are Told: Visual Attention Sink in Large Multimodal Models
von: Kang, Seil, et al.
Veröffentlicht: (2025) -
Not All Attention Heads Are What You Need: Refining CLIP's Image Representation with Attention Ablation
von: Lin, Feng, et al.
Veröffentlicht: (2025)