Fool Me Once? Contrasting Textual and Visual Explanations in a Clinical Decision-Support Setting
Fuente:
arXiv
Salvato in:
| Autori principali: | Kayser, Maxime, Menzat, Bayar, Emde, Cornelius, Bercean, Bogdan, Novak, Alex, Espinosa, Abdala, Papiez, Bartlomiej W., Gaube, Susanne, Lukasiewicz, Thomas, Camburu, Oana-Maria |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Benchmarking Predictive Coding Networks -- Made Simple
di: Pinchetti, Luca, et al.
Pubblicazione: (2024)
di: Pinchetti, Luca, et al.
Pubblicazione: (2024)
Shh, don't say that! Domain Certification in LLMs
di: Emde, Cornelius, et al.
Pubblicazione: (2025)
di: Emde, Cornelius, et al.
Pubblicazione: (2025)
Using Natural Language Explanations to Improve Robustness of In-context Learning
di: He, Xuanli, et al.
Pubblicazione: (2023)
di: He, Xuanli, et al.
Pubblicazione: (2023)
SPARSEFIT: Few-shot Prompting with Sparse Fine-tuning for Jointly Generating Predictions and Natural Language Explanations
di: Solano, Jesus, et al.
Pubblicazione: (2023)
di: Solano, Jesus, et al.
Pubblicazione: (2023)
Fool Me, Fool Me: User Attitudes Toward LLM Falsehoods
di: Nirman, Diana Bar-Or, et al.
Pubblicazione: (2024)
di: Nirman, Diana Bar-Or, et al.
Pubblicazione: (2024)
iLLuMinaTE: An LLM-XAI Framework Leveraging Social Science Explanation Theories Towards Actionable Student Performance Feedback
di: Swamy, Vinitra, et al.
Pubblicazione: (2024)
di: Swamy, Vinitra, et al.
Pubblicazione: (2024)
The Probabilities Also Matter: A More Faithful Metric for Faithfulness of Free-Text Explanations in Large Language Models
di: Siegel, Noah Y., et al.
Pubblicazione: (2024)
di: Siegel, Noah Y., et al.
Pubblicazione: (2024)
Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations
di: Siegel, Noah Y., et al.
Pubblicazione: (2025)
di: Siegel, Noah Y., et al.
Pubblicazione: (2025)
Identifying Linear Relational Concepts in Large Language Models
di: Chanin, David, et al.
Pubblicazione: (2023)
di: Chanin, David, et al.
Pubblicazione: (2023)
Exploring the Effectiveness of Deep Features from Domain-Specific Foundation Models in Retinal Image Synthesis
di: Skorniewska, Zuzanna, et al.
Pubblicazione: (2025)
di: Skorniewska, Zuzanna, et al.
Pubblicazione: (2025)
Prototype Transformer: Towards Language Model Architectures Interpretable by Design
di: Yordanov, Yordan, et al.
Pubblicazione: (2026)
di: Yordanov, Yordan, et al.
Pubblicazione: (2026)
SpineFM: Leveraging Foundation Models for Automatic Spine X-ray Segmentation
di: Simons, Samuel J., et al.
Pubblicazione: (2024)
di: Simons, Samuel J., et al.
Pubblicazione: (2024)
Rethinking Foundation Models for Medical Image Classification through a Benchmark Study on MedMNIST
di: Wu, Fuping, et al.
Pubblicazione: (2025)
di: Wu, Fuping, et al.
Pubblicazione: (2025)
AnnoCaseLaw: A Richly-Annotated Dataset For Benchmarking Explainable Legal Judgment Prediction
di: Sesodia, Magnus, et al.
Pubblicazione: (2025)
di: Sesodia, Magnus, et al.
Pubblicazione: (2025)
Interpretable Rheumatoid Arthritis Scoring via Anatomy-aware Multiple Instance Learning
di: Bo, Zhiyan, et al.
Pubblicazione: (2025)
di: Bo, Zhiyan, et al.
Pubblicazione: (2025)
Fooling the Textual Fooler via Randomizing Latent Representations
di: Hoang, Duy C., et al.
Pubblicazione: (2023)
di: Hoang, Duy C., et al.
Pubblicazione: (2023)
A Stable, Fast, and Fully Automatic Learning Algorithm for Predictive Coding Networks
di: Salvatori, Tommaso, et al.
Pubblicazione: (2022)
di: Salvatori, Tommaso, et al.
Pubblicazione: (2022)
Paired Diffusion: Generation of related, synthetic PET-CT-Segmentation scans using Linked Denoising Diffusion Probabilistic Models
di: Bradbury, Rowan, et al.
Pubblicazione: (2024)
di: Bradbury, Rowan, et al.
Pubblicazione: (2024)
Deep Learning Models to Automate the Scoring of Hand Radiographs for Rheumatoid Arthritis
di: Bo, Zhiyan, et al.
Pubblicazione: (2024)
di: Bo, Zhiyan, et al.
Pubblicazione: (2024)
Atomic Inference for NLI with Generated Facts as Atoms
di: Stacey, Joe, et al.
Pubblicazione: (2023)
di: Stacey, Joe, et al.
Pubblicazione: (2023)
Privacy Collapse: Benign Fine-Tuning Can Break Contextual Privacy in Language Models
di: Goel, Anmol, et al.
Pubblicazione: (2026)
di: Goel, Anmol, et al.
Pubblicazione: (2026)
Prediction of recurrence free survival of head and neck cancer using PET/CT radiomics and clinical information
di: Furukawa, Mona, et al.
Pubblicazione: (2024)
di: Furukawa, Mona, et al.
Pubblicazione: (2024)
On Biases in a UK Biobank-based Retinal Image Classification Model
di: Alloula, Anissa, et al.
Pubblicazione: (2024)
di: Alloula, Anissa, et al.
Pubblicazione: (2024)
Don't be Fooled: The Misinformation Effect of Explanations in Human-AI Collaboration
di: Spitzer, Philipp, et al.
Pubblicazione: (2024)
di: Spitzer, Philipp, et al.
Pubblicazione: (2024)
Spinal Osteophyte Detection via Robust Patch Extraction on minimally annotated X-rays
di: Kundu, Soumya Snigdha, et al.
Pubblicazione: (2024)
di: Kundu, Soumya Snigdha, et al.
Pubblicazione: (2024)
KneeXNeT: An Ensemble-Based Approach for Knee Radiographic Evaluation
di: Srikijkasemwat, Nicharee, et al.
Pubblicazione: (2024)
di: Srikijkasemwat, Nicharee, et al.
Pubblicazione: (2024)
Multimodal Deformable Image Registration for Long-COVID Analysis Based on Progressive Alignment and Multi-perspective Loss
di: Li, Jiahua, et al.
Pubblicazione: (2024)
di: Li, Jiahua, et al.
Pubblicazione: (2024)
Fooling Contrastive Language-Image Pre-trained Models with CLIPMasterPrints
di: Freiberger, Matthias, et al.
Pubblicazione: (2023)
di: Freiberger, Matthias, et al.
Pubblicazione: (2023)
Graph-Guided Textual Explanation Generation Framework
di: Yuan, Shuzhou, et al.
Pubblicazione: (2024)
di: Yuan, Shuzhou, et al.
Pubblicazione: (2024)
Tell Me What's Next: Textual Foresight for Generic UI Representations
di: Burns, Andrea, et al.
Pubblicazione: (2024)
di: Burns, Andrea, et al.
Pubblicazione: (2024)
Visual or Textual: Effects of Explanation Format and Personal Characteristics on the Perception of Explanations in an Educational Recommender System
di: Ain, Qurat Ul, et al.
Pubblicazione: (2026)
di: Ain, Qurat Ul, et al.
Pubblicazione: (2026)
Everything, Everywhere, All at Once: Is Mechanistic Interpretability Identifiable?
di: Méloux, Maxime, et al.
Pubblicazione: (2025)
di: Méloux, Maxime, et al.
Pubblicazione: (2025)
Rethinking Visual Counterfactual Explanations Through Region Constraint
di: Sobieski, Bartlomiej, et al.
Pubblicazione: (2024)
di: Sobieski, Bartlomiej, et al.
Pubblicazione: (2024)
Too Big to Fool: Resisting Deception in Language Models
di: Samsami, Mohammad Reza, et al.
Pubblicazione: (2024)
di: Samsami, Mohammad Reza, et al.
Pubblicazione: (2024)
Integrating Personalized Parsons Problems with Multi-Level Textual Explanations to Scaffold Code Writing
di: Hou, Xinying, et al.
Pubblicazione: (2024)
di: Hou, Xinying, et al.
Pubblicazione: (2024)
Emerging Semantic Segmentation from Positive and Negative Coarse Label Learning
di: Zhang, Le, et al.
Pubblicazione: (2025)
di: Zhang, Le, et al.
Pubblicazione: (2025)
When to Extract ReID Features: A Selective Approach for Improved Multiple Object Tracking
di: Bayar, Emirhan, et al.
Pubblicazione: (2024)
di: Bayar, Emirhan, et al.
Pubblicazione: (2024)
WBCBench 2026: A Challenge for Robust White Blood Cell Classification Under Class Imbalance
di: Tian, Xin, et al.
Pubblicazione: (2026)
di: Tian, Xin, et al.
Pubblicazione: (2026)
Recursive Deformable Image Registration Network with Mutual Attention
di: Zheng, Jian-Qing, et al.
Pubblicazione: (2022)
di: Zheng, Jian-Qing, et al.
Pubblicazione: (2022)
FeatureFool: Zero-Query Fooling of Video Models via Feature Map
di: Tang, Duoxun, et al.
Pubblicazione: (2025)
di: Tang, Duoxun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Benchmarking Predictive Coding Networks -- Made Simple
di: Pinchetti, Luca, et al.
Pubblicazione: (2024) -
Shh, don't say that! Domain Certification in LLMs
di: Emde, Cornelius, et al.
Pubblicazione: (2025) -
Using Natural Language Explanations to Improve Robustness of In-context Learning
di: He, Xuanli, et al.
Pubblicazione: (2023) -
SPARSEFIT: Few-shot Prompting with Sparse Fine-tuning for Jointly Generating Predictions and Natural Language Explanations
di: Solano, Jesus, et al.
Pubblicazione: (2023) -
Fool Me, Fool Me: User Attitudes Toward LLM Falsehoods
di: Nirman, Diana Bar-Or, et al.
Pubblicazione: (2024)