Guarantee Regions for Local Explanations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Havasi, Marton, Parbhoo, Sonali, Doshi-Velez, Finale |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Integrating Personal Knowledge into Test-Time Predictions
von: Lage, Isaac, et al.
Veröffentlicht: (2024)
von: Lage, Isaac, et al.
Veröffentlicht: (2024)
Diverse Concept Proposals for Concept Bottleneck Models
von: Brown, Katrina, et al.
Veröffentlicht: (2024)
von: Brown, Katrina, et al.
Veröffentlicht: (2024)
What Makes a Good Explanation?: A Harmonized View of Properties of Explanations
von: Chen, Zixi, et al.
Veröffentlicht: (2022)
von: Chen, Zixi, et al.
Veröffentlicht: (2022)
Decision-Point Guided Safe Policy Improvement
von: Sharma, Abhishek, et al.
Veröffentlicht: (2024)
von: Sharma, Abhishek, et al.
Veröffentlicht: (2024)
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
von: Benac, Leo, et al.
Veröffentlicht: (2024)
von: Benac, Leo, et al.
Veröffentlicht: (2024)
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
von: Sharma, Abhishek, et al.
Veröffentlicht: (2023)
von: Sharma, Abhishek, et al.
Veröffentlicht: (2023)
Semi-parametric Expert Bayesian Network Learning with Gaussian Processes and Horseshoe Priors
von: Weng, Yidou, et al.
Veröffentlicht: (2024)
von: Weng, Yidou, et al.
Veröffentlicht: (2024)
Strategically Linked Decisions in Long-Term Planning and Reinforcement Learning
von: Hüyük, Alihan, et al.
Veröffentlicht: (2025)
von: Hüyük, Alihan, et al.
Veröffentlicht: (2025)
Transparent Trade-offs between Properties of Explanations
von: Tadesse, Hiwot Belay, et al.
Veröffentlicht: (2024)
von: Tadesse, Hiwot Belay, et al.
Veröffentlicht: (2024)
Tree-Based Leakage Inspection and Control in Concept Bottleneck Models
von: Ragkousis, Angelos, et al.
Veröffentlicht: (2024)
von: Ragkousis, Angelos, et al.
Veröffentlicht: (2024)
Towards Model-Agnostic Posterior Approximation for Fast and Accurate Variational Autoencoders
von: Yacoby, Yaniv, et al.
Veröffentlicht: (2024)
von: Yacoby, Yaniv, et al.
Veröffentlicht: (2024)
On the Effective Horizon of Inverse Reinforcement Learning
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
von: Xu, Yiqing, et al.
Veröffentlicht: (2023)
Connecting Federated ADMM to Bayes
von: Swaroop, Siddharth, et al.
Veröffentlicht: (2025)
von: Swaroop, Siddharth, et al.
Veröffentlicht: (2025)
Quantifying Potential Observation Missingness in Inverse Reinforcement Learning
von: Benac, Leo, et al.
Veröffentlicht: (2026)
von: Benac, Leo, et al.
Veröffentlicht: (2026)
Do regularization methods for shortcut mitigation work as intended?
von: Hong, Haoyang, et al.
Veröffentlicht: (2025)
von: Hong, Haoyang, et al.
Veröffentlicht: (2025)
Inverse Reinforcement Learning with Multiple Planning Horizons
von: Yao, Jiayu, et al.
Veröffentlicht: (2024)
von: Yao, Jiayu, et al.
Veröffentlicht: (2024)
Concept-driven Off Policy Evaluation
von: Majumdar, Ritam, et al.
Veröffentlicht: (2024)
von: Majumdar, Ritam, et al.
Veröffentlicht: (2024)
Feature Importance Depends on Properties of the Data: Towards Choosing the Correct Explanations for Your Data and Decision Trees based Models
von: Ayad, Célia Wafa, et al.
Veröffentlicht: (2025)
von: Ayad, Célia Wafa, et al.
Veröffentlicht: (2025)
A Benchmark for Multi-Party Negotiation Games from Real Negotiation Data
von: Benac, Leo, et al.
Veröffentlicht: (2026)
von: Benac, Leo, et al.
Veröffentlicht: (2026)
Federated ADMM from Bayesian Duality
von: Möllenhoff, Thomas, et al.
Veröffentlicht: (2025)
von: Möllenhoff, Thomas, et al.
Veröffentlicht: (2025)
Causal Bayesian Optimization with Unknown Graphs
von: Durand, Jean, et al.
Veröffentlicht: (2025)
von: Durand, Jean, et al.
Veröffentlicht: (2025)
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
von: Nofshin, Eura, et al.
Veröffentlicht: (2024)
von: Nofshin, Eura, et al.
Veröffentlicht: (2024)
A Sim2Real Approach for Identifying Task-Relevant Properties in Interpretable Machine Learning
von: Nofshin, Eura, et al.
Veröffentlicht: (2024)
von: Nofshin, Eura, et al.
Veröffentlicht: (2024)
Understanding the Relationship between Prompts and Response Uncertainty in Large Language Models
von: Zhang, Ze Yu, et al.
Veröffentlicht: (2024)
von: Zhang, Ze Yu, et al.
Veröffentlicht: (2024)
Understanding and Mitigating Tokenization Bias in Language Models
von: Phan, Buu, et al.
Veröffentlicht: (2024)
von: Phan, Buu, et al.
Veröffentlicht: (2024)
Edit Flows: Flow Matching with Edit Operations
von: Havasi, Marton, et al.
Veröffentlicht: (2025)
von: Havasi, Marton, et al.
Veröffentlicht: (2025)
Improving ARDS Diagnosis Through Context-Aware Concept Bottleneck Models
von: Narain, Anish, et al.
Veröffentlicht: (2025)
von: Narain, Anish, et al.
Veröffentlicht: (2025)
Learning from Failures: Understanding LLM Alignment through Failure-Aware Inverse RL
von: Patel, Nyal, et al.
Veröffentlicht: (2025)
von: Patel, Nyal, et al.
Veröffentlicht: (2025)
The Alignment Auditor: A Bayesian Framework for Verifying and Refining LLM Objectives
von: Bou, Matthieu, et al.
Veröffentlicht: (2025)
von: Bou, Matthieu, et al.
Veröffentlicht: (2025)
Non-Stationary Latent Auto-Regressive Bandits
von: Trella, Anna L., et al.
Veröffentlicht: (2024)
von: Trella, Anna L., et al.
Veröffentlicht: (2024)
Exact Byte-Level Probabilities from Tokenized Language Models for FIM-Tasks and Model Ensembles
von: Phan, Buu, et al.
Veröffentlicht: (2024)
von: Phan, Buu, et al.
Veröffentlicht: (2024)
Monitoring Fidelity of Online Reinforcement Learning Algorithms in Clinical Trials
von: Trella, Anna L., et al.
Veröffentlicht: (2024)
von: Trella, Anna L., et al.
Veröffentlicht: (2024)
CONFEX: Uncertainty-Aware Counterfactual Explanations with Conformal Guarantees
von: Bilkhoo, Aman, et al.
Veröffentlicht: (2025)
von: Bilkhoo, Aman, et al.
Veröffentlicht: (2025)
Regional Explanations: Bridging Local and Global Variable Importance
von: Amoukou, Salim I., et al.
Veröffentlicht: (2026)
von: Amoukou, Salim I., et al.
Veröffentlicht: (2026)
Rigorous Probabilistic Guarantees for Robust Counterfactual Explanations
von: Marzari, Luca, et al.
Veröffentlicht: (2024)
von: Marzari, Luca, et al.
Veröffentlicht: (2024)
Pruning the Path to Optimal Care: Identifying Systematically Suboptimal Medical Decision-Making with Inverse Reinforcement Learning
von: Bovenzi, Inko, et al.
Veröffentlicht: (2024)
von: Bovenzi, Inko, et al.
Veröffentlicht: (2024)
Counterfactual Explanations with Probabilistic Guarantees on their Robustness to Model Change
von: Stępka, Ignacy, et al.
Veröffentlicht: (2024)
von: Stępka, Ignacy, et al.
Veröffentlicht: (2024)
Set Block Decoding is a Language Model Inference Accelerator
von: Gat, Itai, et al.
Veröffentlicht: (2025)
von: Gat, Itai, et al.
Veröffentlicht: (2025)
Shaping AI's Impact on Billions of Lives
von: Cuéllar, Mariano-Florentino, et al.
Veröffentlicht: (2024)
von: Cuéllar, Mariano-Florentino, et al.
Veröffentlicht: (2024)
Guaranteed Optimal Compositional Explanations for Neurons
von: La Rosa, Biagio, et al.
Veröffentlicht: (2025)
von: La Rosa, Biagio, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards Integrating Personal Knowledge into Test-Time Predictions
von: Lage, Isaac, et al.
Veröffentlicht: (2024) -
Diverse Concept Proposals for Concept Bottleneck Models
von: Brown, Katrina, et al.
Veröffentlicht: (2024) -
What Makes a Good Explanation?: A Harmonized View of Properties of Explanations
von: Chen, Zixi, et al.
Veröffentlicht: (2022) -
Decision-Point Guided Safe Policy Improvement
von: Sharma, Abhishek, et al.
Veröffentlicht: (2024) -
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
von: Benac, Leo, et al.
Veröffentlicht: (2024)