Trustworthy AI Suffers from Invariance Conflicts and Causality is The Solution
Fuente:
arXiv
Salvato in:
| Autori principali: | Binkyte, Ruta, Sheth, Ivaxi, Jin, Zhijing, Havaei, Mohammad, Schölkopf, Bernhard, Fritz, Mario |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Causality Is Key to Understand and Balance Multiple Goals in Trustworthy ML and Foundation Models
di: Binkyte, Ruta, et al.
Pubblicazione: (2025)
di: Binkyte, Ruta, et al.
Pubblicazione: (2025)
Safety Must Precede the Deployment of Open-Ended AI
di: Sheth, Ivaxi, et al.
Pubblicazione: (2025)
di: Sheth, Ivaxi, et al.
Pubblicazione: (2025)
LLM4GRN: Discovering Causal Gene Regulatory Networks with LLMs -- Evaluation through Synthetic Data Generation
di: Afonja, Tejumade, et al.
Pubblicazione: (2024)
di: Afonja, Tejumade, et al.
Pubblicazione: (2024)
IV Co-Scientist: Multi-Agent LLM Framework for Causal Instrumental Variable Discovery
di: Sheth, Ivaxi, et al.
Pubblicazione: (2026)
di: Sheth, Ivaxi, et al.
Pubblicazione: (2026)
Justice in Judgment: Unveiling (Hidden) Bias in LLM-assisted Peer Reviews
di: Vasu, Sai Suresh Macharla, et al.
Pubblicazione: (2025)
di: Vasu, Sai Suresh Macharla, et al.
Pubblicazione: (2025)
Causal Responsibility Attribution for Human-AI Collaboration
di: Qi, Yahang, et al.
Pubblicazione: (2024)
di: Qi, Yahang, et al.
Pubblicazione: (2024)
Interactional Fairness in LLM Multi-Agent Systems: An Evaluation Framework
di: Binkyte, Ruta
Pubblicazione: (2025)
di: Binkyte, Ruta
Pubblicazione: (2025)
ProtocolLLM: RTL Benchmark for SystemVerilog Generation of Communication Protocols
di: Sheth, Arnav, et al.
Pubblicazione: (2025)
di: Sheth, Arnav, et al.
Pubblicazione: (2025)
Causality can systematically address the monsters under the bench(marks)
di: Leeb, Felix, et al.
Pubblicazione: (2025)
di: Leeb, Felix, et al.
Pubblicazione: (2025)
On the Need and Applicability of Causality for Fairness: A Unified Framework for AI Auditing and Legal Analysis
di: Binkyte, Ruta, et al.
Pubblicazione: (2022)
di: Binkyte, Ruta, et al.
Pubblicazione: (2022)
Inspectable AI for Science: A Research Object Approach to Generative AI Governance
di: Binkyte, Ruta, et al.
Pubblicazione: (2026)
di: Binkyte, Ruta, et al.
Pubblicazione: (2026)
Funny or Persuasive, but Not Both: Evaluating Fine-Grained Multi-Concept Control in LLMs
di: Labroo, Arya, et al.
Pubblicazione: (2026)
di: Labroo, Arya, et al.
Pubblicazione: (2026)
BaBE: Enhancing Fairness via Estimation of Latent Explaining Variables
di: Binkyte, Ruta, et al.
Pubblicazione: (2023)
di: Binkyte, Ruta, et al.
Pubblicazione: (2023)
Stargazer: A Scalable Model-Fitting Benchmark Environment for AI Agents under Astrophysical Constraints
di: Liu, Xinge, et al.
Pubblicazione: (2026)
di: Liu, Xinge, et al.
Pubblicazione: (2026)
Improving Large Language Model Safety with Contrastive Representation Learning
di: Simko, Samuel, et al.
Pubblicazione: (2025)
di: Simko, Samuel, et al.
Pubblicazione: (2025)
CausalCite: A Causal Formulation of Paper Citations
di: Kumar, Ishan, et al.
Pubblicazione: (2023)
di: Kumar, Ishan, et al.
Pubblicazione: (2023)
Causal Discovery Under Local Privacy
di: Binkytė, Rūta, et al.
Pubblicazione: (2023)
di: Binkytė, Rūta, et al.
Pubblicazione: (2023)
CausalGraph2LLM: Evaluating LLMs for Causal Queries
di: Sheth, Ivaxi, et al.
Pubblicazione: (2024)
di: Sheth, Ivaxi, et al.
Pubblicazione: (2024)
Context-Aware Reasoning On Parametric Knowledge for Inferring Causal Variables
di: Sheth, Ivaxi, et al.
Pubblicazione: (2024)
di: Sheth, Ivaxi, et al.
Pubblicazione: (2024)
Causality for Natural Language Processing
di: Jin, Zhijing
Pubblicazione: (2025)
di: Jin, Zhijing
Pubblicazione: (2025)
Survey on AI Ethics: A Socio-technical Perspective
di: Mbiazi, Dave, et al.
Pubblicazione: (2023)
di: Mbiazi, Dave, et al.
Pubblicazione: (2023)
Corrupted by Reasoning: Reasoning Language Models Become Free-Riders in Public Goods Games
di: Piedrahita, David Guzman, et al.
Pubblicazione: (2025)
di: Piedrahita, David Guzman, et al.
Pubblicazione: (2025)
A Measure-Theoretic Axiomatisation of Causality
di: Park, Junhyung, et al.
Pubblicazione: (2023)
di: Park, Junhyung, et al.
Pubblicazione: (2023)
Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
di: Jenny, David F., et al.
Pubblicazione: (2023)
di: Jenny, David F., et al.
Pubblicazione: (2023)
Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI
di: Huang, Xuanqiang Angelo, et al.
Pubblicazione: (2026)
di: Huang, Xuanqiang Angelo, et al.
Pubblicazione: (2026)
Quriosity: Analyzing Human Questioning Behavior and Causal Inquiry through Curiosity-Driven Queries
di: Ceraolo, Roberto, et al.
Pubblicazione: (2024)
di: Ceraolo, Roberto, et al.
Pubblicazione: (2024)
Computational Arbitrage in AI Model Markets
di: Olmedo, Ricardo, et al.
Pubblicazione: (2026)
di: Olmedo, Ricardo, et al.
Pubblicazione: (2026)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
di: Backmann, Steffen, et al.
Pubblicazione: (2025)
di: Backmann, Steffen, et al.
Pubblicazione: (2025)
Position: Cracking the Code of Cascading Disparity Towards Marginalized Communities
di: Farnadi, Golnoosh, et al.
Pubblicazione: (2024)
di: Farnadi, Golnoosh, et al.
Pubblicazione: (2024)
Can Theoretical Physics Research Benefit from Language Agents?
di: Lu, Sirui, et al.
Pubblicazione: (2025)
di: Lu, Sirui, et al.
Pubblicazione: (2025)
CLadder: Assessing Causal Reasoning in Language Models
di: Jin, Zhijing, et al.
Pubblicazione: (2023)
di: Jin, Zhijing, et al.
Pubblicazione: (2023)
Deep Backtracking Counterfactuals for Causally Compliant Explanations
di: Kladny, Klaus-Rudolf, et al.
Pubblicazione: (2023)
di: Kladny, Klaus-Rudolf, et al.
Pubblicazione: (2023)
Identifiable Exchangeable Mechanisms for Causal Structure and Representation Learning
di: Reizinger, Patrik, et al.
Pubblicazione: (2024)
di: Reizinger, Patrik, et al.
Pubblicazione: (2024)
Causal vs. Anticausal merging of predictors
di: Mejia, Sergio Hernan Garrido, et al.
Pubblicazione: (2025)
di: Mejia, Sergio Hernan Garrido, et al.
Pubblicazione: (2025)
Can Large Language Models Infer Causation from Correlation?
di: Jin, Zhijing, et al.
Pubblicazione: (2023)
di: Jin, Zhijing, et al.
Pubblicazione: (2023)
PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?
di: Pulipaka, Sidharth, et al.
Pubblicazione: (2026)
di: Pulipaka, Sidharth, et al.
Pubblicazione: (2026)
Uncovering Hidden Correctness in LLM Causal Reasoning via Symbolic Verification
di: He, Paul, et al.
Pubblicazione: (2026)
di: He, Paul, et al.
Pubblicazione: (2026)
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies
di: Kekić, Armin, et al.
Pubblicazione: (2025)
di: Kekić, Armin, et al.
Pubblicazione: (2025)
Hidden in Memory: Sleeper Memory Poisoning in LLM Agents
di: Pulipaka, Sidharth, et al.
Pubblicazione: (2026)
di: Pulipaka, Sidharth, et al.
Pubblicazione: (2026)
HyperCausalLP: Causal Link Prediction using Hyper-Relational Knowledge Graph
di: Jaimini, Utkarshani, et al.
Pubblicazione: (2024)
di: Jaimini, Utkarshani, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Causality Is Key to Understand and Balance Multiple Goals in Trustworthy ML and Foundation Models
di: Binkyte, Ruta, et al.
Pubblicazione: (2025) -
Safety Must Precede the Deployment of Open-Ended AI
di: Sheth, Ivaxi, et al.
Pubblicazione: (2025) -
LLM4GRN: Discovering Causal Gene Regulatory Networks with LLMs -- Evaluation through Synthetic Data Generation
di: Afonja, Tejumade, et al.
Pubblicazione: (2024) -
IV Co-Scientist: Multi-Agent LLM Framework for Causal Instrumental Variable Discovery
di: Sheth, Ivaxi, et al.
Pubblicazione: (2026) -
Justice in Judgment: Unveiling (Hidden) Bias in LLM-assisted Peer Reviews
di: Vasu, Sai Suresh Macharla, et al.
Pubblicazione: (2025)