Trustworthy AI Must Account for Interactions
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Cresswell, Jesse C. |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Knowledge Distillation Must Account for What It Loses
par: Wang, Wenshuo
Publié: (2026)
par: Wang, Wenshuo
Publié: (2026)
A Geometric View of Data Complexity: Efficient Local Intrinsic Dimension Estimation with Diffusion Models
par: Kamkari, Hamidreza, et autres
Publié: (2024)
par: Kamkari, Hamidreza, et autres
Publié: (2024)
Trustworthy Actionable Perturbations
par: Friedbaum, Jesse, et autres
Publié: (2024)
par: Friedbaum, Jesse, et autres
Publié: (2024)
Deep Generative Models through the Lens of the Manifold Hypothesis: A Survey and New Connections
par: Loaiza-Ganem, Gabriel, et autres
Publié: (2024)
par: Loaiza-Ganem, Gabriel, et autres
Publié: (2024)
Explainable AI Systems Must Be Contestable: Here's How to Make It Happen
par: Moreira, Catarina, et autres
Publié: (2025)
par: Moreira, Catarina, et autres
Publié: (2025)
Inconsistencies In Consistency Models: Better ODE Solving Does Not Imply Better Samples
par: Vouitsis, Noël, et autres
Publié: (2024)
par: Vouitsis, Noël, et autres
Publié: (2024)
The Cell Must Go On: Agar.io for Continual Reinforcement Learning
par: Mohamed, Mohamed A., et autres
Publié: (2025)
par: Mohamed, Mohamed A., et autres
Publié: (2025)
Toward Maturity-Based Certification of Embodied AI: Quantifying Trustworthiness Through Measurement Mechanisms
par: Darling, Michael C., et autres
Publié: (2026)
par: Darling, Michael C., et autres
Publié: (2026)
Co-Investigator AI: The Rise of Agentic AI for Smarter, Trustworthy AML Compliance Narratives
par: Naik, Prathamesh Vasudeo, et autres
Publié: (2025)
par: Naik, Prathamesh Vasudeo, et autres
Publié: (2025)
TabDPT: Scaling Tabular Foundation Models on Real Data
par: Ma, Junwei, et autres
Publié: (2024)
par: Ma, Junwei, et autres
Publié: (2024)
Position: A Theory of Deep Learning Must Include Compositional Sparsity
par: Danhofer, David A., et autres
Publié: (2025)
par: Danhofer, David A., et autres
Publié: (2025)
Grounding Generative Planners in Verifiable Logic: A Hybrid Architecture for Trustworthy Embodied AI
par: Wu, Feiyu, et autres
Publié: (2026)
par: Wu, Feiyu, et autres
Publié: (2026)
Whatever Remains Must Be True: Filtering Drives Reasoning in LLMs, Shaping Diversity
par: Kruszewski, Germán, et autres
Publié: (2025)
par: Kruszewski, Germán, et autres
Publié: (2025)
Simulating the Unseen: Crash Prediction Must Learn from What Did Not Happen
par: Li, Zihao, et autres
Publié: (2025)
par: Li, Zihao, et autres
Publié: (2025)
Standardization Trends on Safety and Trustworthiness Technology for Advanced AI
par: Jeon, Jonghong
Publié: (2024)
par: Jeon, Jonghong
Publié: (2024)
Assessing Trustworthiness of AI Training Dataset using Subjective Logic -- A Use Case on Bias
par: Ouattara, Koffi Ismael, et autres
Publié: (2025)
par: Ouattara, Koffi Ismael, et autres
Publié: (2025)
Parent-Guided Adaptive Reliability (PGAR): A Behavioural Meta-Learning Framework for Stable and Trustworthy AI
par: Rankawat, Anshum
Publié: (2026)
par: Rankawat, Anshum
Publié: (2026)
Towards Trustworthy Keylogger detection: A Comprehensive Analysis of Ensemble Techniques and Feature Selections through Explainable AI
par: Mahmud, Monirul Islam
Publié: (2025)
par: Mahmud, Monirul Islam
Publié: (2025)
Textual Bayes: Quantifying Prompt Uncertainty in LLM-Based Systems
par: Ross, Brendan Leigh, et autres
Publié: (2025)
par: Ross, Brendan Leigh, et autres
Publié: (2025)
Towards Trustworthy Vital Sign Forecasting: Leveraging Uncertainty for Prediction Intervals
par: Wang, Li Rong, et autres
Publié: (2025)
par: Wang, Li Rong, et autres
Publié: (2025)
Mechanistic Interpretability of Fine-Tuned Vision Transformers on Distorted Images: Decoding Attention Head Behavior for Transparent and Trustworthy AI
par: Bahador, Nooshin
Publié: (2025)
par: Bahador, Nooshin
Publié: (2025)
Data Heterogeneity Modeling for Trustworthy Machine Learning
par: Liu, Jiashuo, et autres
Publié: (2025)
par: Liu, Jiashuo, et autres
Publié: (2025)
LLINBO: Trustworthy LLM-in-the-Loop Bayesian Optimization
par: Chang, Chih-Yu, et autres
Publié: (2025)
par: Chang, Chih-Yu, et autres
Publié: (2025)
Actionable Interpretability Must Be Defined in Terms of Symmetries
par: Barbiero, Pietro, et autres
Publié: (2026)
par: Barbiero, Pietro, et autres
Publié: (2026)
Energy per Successful Goal: Goal-Level Energy Accounting for Agentic AI Systems
par: Panigrahy, Deepak, et autres
Publié: (2026)
par: Panigrahy, Deepak, et autres
Publié: (2026)
Faithful and Stable Neuron Explanations for Trustworthy Mechanistic Interpretability
par: Yan, Ge, et autres
Publié: (2025)
par: Yan, Ge, et autres
Publié: (2025)
Trustworthy GNNs with LLMs: A Systematic Review and Taxonomy
par: Xue, Ruizhan, et autres
Publié: (2025)
par: Xue, Ruizhan, et autres
Publié: (2025)
Trustworthy Graph Neural Networks: Aspects, Methods and Trends
par: Zhang, He, et autres
Publié: (2022)
par: Zhang, He, et autres
Publié: (2022)
Generalizing Trust: Weak-to-Strong Trustworthiness in Language Models
par: Pawelczyk, Martin, et autres
Publié: (2024)
par: Pawelczyk, Martin, et autres
Publié: (2024)
A Geometric Explanation of the Likelihood OOD Detection Paradox
par: Kamkari, Hamidreza, et autres
Publié: (2024)
par: Kamkari, Hamidreza, et autres
Publié: (2024)
High Noise Scheduling is a Must
par: Gokmen, Mahmut S., et autres
Publié: (2024)
par: Gokmen, Mahmut S., et autres
Publié: (2024)
Gradients Must Earn Their Influence: Unifying SFT with Generalized Entropic Objectives
par: Wang, Zecheng, et autres
Publié: (2026)
par: Wang, Zecheng, et autres
Publié: (2026)
Large Language Models Must Be Taught to Know What They Don't Know
par: Kapoor, Sanyam, et autres
Publié: (2024)
par: Kapoor, Sanyam, et autres
Publié: (2024)
Exploration of the Rashomon Set Assists Trustworthy Explanations for Medical Data
par: Kobylińska, Katarzyna, et autres
Publié: (2023)
par: Kobylińska, Katarzyna, et autres
Publié: (2023)
SpecTM: Spectral Targeted Masking for Trustworthy Foundation Models
par: Imtiaz, Syed Usama, et autres
Publié: (2026)
par: Imtiaz, Syed Usama, et autres
Publié: (2026)
On the Utility of Accounting for Human Beliefs about AI Intention in Human-AI Collaboration
par: Yu, Guanghui, et autres
Publié: (2024)
par: Yu, Guanghui, et autres
Publié: (2024)
Towards Trustworthy AI: Secure Deepfake Detection using CNNs and Zero-Knowledge Proofs
par: Islam, H M Mohaimanul, et autres
Publié: (2025)
par: Islam, H M Mohaimanul, et autres
Publié: (2025)
Trustworthy AI: Safety, Bias, and Privacy -- A Survey
par: Fang, Xingli, et autres
Publié: (2025)
par: Fang, Xingli, et autres
Publié: (2025)
A Unified Memory Perspective for Probabilistic Trustworthy AI
par: Zhao, Xueji, et autres
Publié: (2026)
par: Zhao, Xueji, et autres
Publié: (2026)
Improving Health Professionals' Onboarding with AI and XAI for Trustworthy Human-AI Collaborative Decision Making
par: Lee, Min Hun, et autres
Publié: (2024)
par: Lee, Min Hun, et autres
Publié: (2024)
Documents similaires
-
Knowledge Distillation Must Account for What It Loses
par: Wang, Wenshuo
Publié: (2026) -
A Geometric View of Data Complexity: Efficient Local Intrinsic Dimension Estimation with Diffusion Models
par: Kamkari, Hamidreza, et autres
Publié: (2024) -
Trustworthy Actionable Perturbations
par: Friedbaum, Jesse, et autres
Publié: (2024) -
Deep Generative Models through the Lens of the Manifold Hypothesis: A Survey and New Connections
par: Loaiza-Ganem, Gabriel, et autres
Publié: (2024) -
Explainable AI Systems Must Be Contestable: Here's How to Make It Happen
par: Moreira, Catarina, et autres
Publié: (2025)