Position: Explain to Question not to Justify
Fuente:
arXiv
Saved in:
| Main Authors: | Biecek, Przemyslaw, Samek, Wojciech |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Model Science: getting serious about verification, explanation and control of AI systems
by: Biecek, Przemyslaw, et al.
Published: (2025)
by: Biecek, Przemyslaw, et al.
Published: (2025)
Adversarial attacks and defenses in explainable artificial intelligence: A survey
by: Baniecki, Hubert, et al.
Published: (2023)
by: Baniecki, Hubert, et al.
Published: (2023)
Strategic Dishonesty Can Undermine AI Safety Evaluations of Frontier LLMs
by: Panfilov, Alexander, et al.
Published: (2025)
by: Panfilov, Alexander, et al.
Published: (2025)
An AI Architecture with the Capability to Classify and Explain Hardware Trojans
by: Whitten, Paul, et al.
Published: (2024)
by: Whitten, Paul, et al.
Published: (2024)
Uncovering, Explaining, and Mitigating the Superficial Safety of Backdoor Defense
by: Min, Rui, et al.
Published: (2024)
by: Min, Rui, et al.
Published: (2024)
SHIELD: Secure Hypernetworks for Incremental Expansion Learning Defense
by: Krukowski, Patryk, et al.
Published: (2025)
by: Krukowski, Patryk, et al.
Published: (2025)
Position: Retire the "Positive Backdoor" Label -- Secret Alignment Requires Strict and Systematic Evaluation
by: Li, Jianwei, et al.
Published: (2026)
by: Li, Jianwei, et al.
Published: (2026)
Position: Do Not Explain Vision Models Without Context
by: Tomaszewska, Paulina, et al.
Published: (2024)
by: Tomaszewska, Paulina, et al.
Published: (2024)
Position: Challenges and Opportunities for Differential Privacy in the U.S. Federal Government
by: Khanna, Amol, et al.
Published: (2024)
by: Khanna, Amol, et al.
Published: (2024)
Position: AI Security Policy Should Target Systems, Not Models
by: Riegler, Michael A., et al.
Published: (2026)
by: Riegler, Michael A., et al.
Published: (2026)
A Privacy Preserving System for Movie Recommendations Using Federated Learning
by: Neumann, David, et al.
Published: (2023)
by: Neumann, David, et al.
Published: (2023)
Position: Privacy Is Not Just Memorization!
by: Mireshghallah, Niloofar, et al.
Published: (2025)
by: Mireshghallah, Niloofar, et al.
Published: (2025)
There Are No Silly Questions: Evaluation of Offline LLM Capabilities from a Turkish Perspective
by: Yilmaz, Edibe, et al.
Published: (2026)
by: Yilmaz, Edibe, et al.
Published: (2026)
Trojan Horse Hunt in Time Series Forecasting for Space Operations
by: Kotowski, Krzysztof, et al.
Published: (2025)
by: Kotowski, Krzysztof, et al.
Published: (2025)
Fake or Real: The Impostor Hunt in Texts for Space Operations
by: Kaczmarek, Agata, et al.
Published: (2025)
by: Kaczmarek, Agata, et al.
Published: (2025)
Prompt Injection Attacks on Large Language Models in Oncology
by: Clusmann, Jan, et al.
Published: (2024)
by: Clusmann, Jan, et al.
Published: (2024)
Unveiling the Threat of Fraud Gangs to Graph Neural Networks: Multi-Target Graph Injection Attacks Against GNN-Based Fraud Detectors
by: Choi, Jinhyeok, et al.
Published: (2024)
by: Choi, Jinhyeok, et al.
Published: (2024)
PBP: Post-training Backdoor Purification for Malware Classifiers
by: Nguyen, Dung Thuy, et al.
Published: (2024)
by: Nguyen, Dung Thuy, et al.
Published: (2024)
Manipulating hidden-Markov-model inferences by corrupting batch data
by: Caballero, William N., et al.
Published: (2024)
by: Caballero, William N., et al.
Published: (2024)
Generative Models are Self-Watermarked: Declaring Model Authentication through Re-Generation
by: Desu, Aditya, et al.
Published: (2024)
by: Desu, Aditya, et al.
Published: (2024)
On the (In)feasibility of ML Backdoor Detection as an Hypothesis Testing Problem
by: Pichler, Georg, et al.
Published: (2024)
by: Pichler, Georg, et al.
Published: (2024)
Free Lunch for Federated Remote Sensing Target Fine-Grained Classification: A Parameter-Efficient Framework
by: Chen, Shengchao, et al.
Published: (2024)
by: Chen, Shengchao, et al.
Published: (2024)
Linearizing Models for Efficient yet Robust Private Inference
by: Sarkar, Sreetama, et al.
Published: (2024)
by: Sarkar, Sreetama, et al.
Published: (2024)
I can't see it but I can Fine-tune it: On Encrypted Fine-tuning of Transformers using Fully Homomorphic Encryption
by: Panzade, Prajwal, et al.
Published: (2024)
by: Panzade, Prajwal, et al.
Published: (2024)
Improving Clean Accuracy via a Tangent-Space Perspective on Adversarial Training
by: Yi, Bongsoo, et al.
Published: (2024)
by: Yi, Bongsoo, et al.
Published: (2024)
On TinyML and Cybersecurity: Electric Vehicle Charging Infrastructure Use Case
by: Dehrouyeh, Fatemeh, et al.
Published: (2024)
by: Dehrouyeh, Fatemeh, et al.
Published: (2024)
Theoretical Analysis of Privacy Leakage in Trustworthy Federated Learning: A Perspective from Linear Algebra and Optimization Theory
by: Zhang, Xiaojin, et al.
Published: (2024)
by: Zhang, Xiaojin, et al.
Published: (2024)
Knowledge-Informed Auto-Penetration Testing Based on Reinforcement Learning with Reward Machine
by: Li, Yuanliang, et al.
Published: (2024)
by: Li, Yuanliang, et al.
Published: (2024)
TrojFM: Resource-efficient Backdoor Attacks against Very Large Foundation Models
by: Nie, Yuzhou., et al.
Published: (2024)
by: Nie, Yuzhou., et al.
Published: (2024)
Rethinking Pruning for Backdoor Mitigation: An Optimization Perspective
by: Li, Nan, et al.
Published: (2024)
by: Li, Nan, et al.
Published: (2024)
Fooling SHAP with Output Shuffling Attacks
by: Yuan, Jun, et al.
Published: (2024)
by: Yuan, Jun, et al.
Published: (2024)
Perturbation Towards Easy Samples Improves Targeted Adversarial Transferability
by: Gao, Junqi, et al.
Published: (2024)
by: Gao, Junqi, et al.
Published: (2024)
Automated Creation of Source Code Variants of a Cryptographic Hash Function Implementation Using Generative Pre-Trained Transformer Models
by: Pelofske, Elijah, et al.
Published: (2024)
by: Pelofske, Elijah, et al.
Published: (2024)
C-RADAR: A Centralized Deep Learning System for Intrusion Detection in Software Defined Networks
by: Mustafa, Osama, et al.
Published: (2024)
by: Mustafa, Osama, et al.
Published: (2024)
Center-Based Relaxed Learning Against Membership Inference Attacks
by: Fang, Xingli, et al.
Published: (2024)
by: Fang, Xingli, et al.
Published: (2024)
Towards Characterizing Cyber Networks with Large Language Models
by: Hartsock, Alaric, et al.
Published: (2024)
by: Hartsock, Alaric, et al.
Published: (2024)
Data Poisoning Attacks on Off-Policy Policy Evaluation Methods
by: Lobo, Elita, et al.
Published: (2024)
by: Lobo, Elita, et al.
Published: (2024)
Investigating the Impact of Quantization on Adversarial Robustness
by: Li, Qun, et al.
Published: (2024)
by: Li, Qun, et al.
Published: (2024)
Comprehensive evaluation of Mal-API-2019 dataset by machine learning in malware detection
by: Li, Zhenglin, et al.
Published: (2024)
by: Li, Zhenglin, et al.
Published: (2024)
A Survey on Adversarial Robustness of LiDAR-based Machine Learning Perception in Autonomous Vehicles
by: Kim, Junae, et al.
Published: (2024)
by: Kim, Junae, et al.
Published: (2024)
Similar Items
-
Model Science: getting serious about verification, explanation and control of AI systems
by: Biecek, Przemyslaw, et al.
Published: (2025) -
Adversarial attacks and defenses in explainable artificial intelligence: A survey
by: Baniecki, Hubert, et al.
Published: (2023) -
Strategic Dishonesty Can Undermine AI Safety Evaluations of Frontier LLMs
by: Panfilov, Alexander, et al.
Published: (2025) -
An AI Architecture with the Capability to Classify and Explain Hardware Trojans
by: Whitten, Paul, et al.
Published: (2024) -
Uncovering, Explaining, and Mitigating the Superficial Safety of Backdoor Defense
by: Min, Rui, et al.
Published: (2024)