Architecting Trust in Artificial Epistemic Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Marchal, Nahema, Chan, Stephanie, Franklin, Matija, Revel, Manon, Keeling, Geoff, Fischli, Roberta, Chandra, Bilva, Gabriel, Iason |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data
von: Marchal, Nahema, et al.
Veröffentlicht: (2024)
von: Marchal, Nahema, et al.
Veröffentlicht: (2024)
We Need a New Ethics for a World of AI Agents
von: Gabriel, Iason, et al.
Veröffentlicht: (2025)
von: Gabriel, Iason, et al.
Veröffentlicht: (2025)
Virtual Agent Economies
von: Tomasev, Nenad, et al.
Veröffentlicht: (2025)
von: Tomasev, Nenad, et al.
Veröffentlicht: (2025)
Model-Free RL Agents Demonstrate System 1-Like Intentionality
von: Ashton, Hal, et al.
Veröffentlicht: (2025)
von: Ashton, Hal, et al.
Veröffentlicht: (2025)
AI-Enhanced Deliberative Democracy and the Future of the Collective Will
von: Revel, Manon, et al.
Veröffentlicht: (2025)
von: Revel, Manon, et al.
Veröffentlicht: (2025)
On the attribution of confidence to large language models
von: Keeling, Geoff, et al.
Veröffentlicht: (2024)
von: Keeling, Geoff, et al.
Veröffentlicht: (2024)
A Mechanism-Based Approach to Mitigating Harms from Persuasive Generative AI
von: El-Sayed, Seliem, et al.
Veröffentlicht: (2024)
von: El-Sayed, Seliem, et al.
Veröffentlicht: (2024)
Characterizing AI Agents for Alignment and Governance
von: Kasirzadeh, Atoosa, et al.
Veröffentlicht: (2025)
von: Kasirzadeh, Atoosa, et al.
Veröffentlicht: (2025)
Intelligent AI Delegation
von: Tomašev, Nenad, et al.
Veröffentlicht: (2026)
von: Tomašev, Nenad, et al.
Veröffentlicht: (2026)
Deflating Deflationism: A Critical Perspective on Debunking Arguments Against LLM Mentality
von: Grzankowski, Alex, et al.
Veröffentlicht: (2025)
von: Grzankowski, Alex, et al.
Veröffentlicht: (2025)
Know When to Trust the Skill: Delayed Appraisal and Epistemic Vigilance for Single-Agent LLMs
von: Unlu, Eren
Veröffentlicht: (2026)
von: Unlu, Eren
Veröffentlicht: (2026)
SEAL: Systematic Error Analysis for Value ALignment
von: Revel, Manon, et al.
Veröffentlicht: (2024)
von: Revel, Manon, et al.
Veröffentlicht: (2024)
Theory of Mind and Self-Attributions of Mentality are Dissociable in LLMs
von: Kim, Junsol, et al.
Veröffentlicht: (2026)
von: Kim, Junsol, et al.
Veröffentlicht: (2026)
Resource Rational Contractualism Should Guide AI Alignment
von: Levine, Sydney, et al.
Veröffentlicht: (2025)
von: Levine, Sydney, et al.
Veröffentlicht: (2025)
Beyond Preferences in AI Alignment
von: Zhi-Xuan, Tan, et al.
Veröffentlicht: (2024)
von: Zhi-Xuan, Tan, et al.
Veröffentlicht: (2024)
AI Governance through Markets
von: Tomei, Philip Moreira, et al.
Veröffentlicht: (2025)
von: Tomei, Philip Moreira, et al.
Veröffentlicht: (2025)
(Unfair) Norms in Fairness Research: A Meta-Analysis
von: Chien, Jennifer, et al.
Veröffentlicht: (2024)
von: Chien, Jennifer, et al.
Veröffentlicht: (2024)
Distributional AGI Safety
von: Tomašev, Nenad, et al.
Veröffentlicht: (2025)
von: Tomašev, Nenad, et al.
Veröffentlicht: (2025)
Epistemic Context Learning: Building Trust the Right Way in LLM-Based Multi-Agent Systems
von: Zhou, Ruiwen, et al.
Veröffentlicht: (2026)
von: Zhou, Ruiwen, et al.
Veröffentlicht: (2026)
Correctness, Artificial Intelligence, and the Epistemic Value of Mathematical Proof
von: Weatherall, James Owen, et al.
Veröffentlicht: (2026)
von: Weatherall, James Owen, et al.
Veröffentlicht: (2026)
Architecting Clinical Collaboration: Multi-Agent Reasoning Systems for Multimodal Medical VQA
von: Thakrar, Karishma, et al.
Veröffentlicht: (2025)
von: Thakrar, Karishma, et al.
Veröffentlicht: (2025)
Should agentic conversational AI change how we think about ethics? Characterising an interactional ethics centred on respect
von: Alberts, Lize, et al.
Veröffentlicht: (2024)
von: Alberts, Lize, et al.
Veröffentlicht: (2024)
Architecting AgentOS: From Token-Level Context to Emergent System-Level Intelligence
von: Li, ChengYou, et al.
Veröffentlicht: (2026)
von: Li, ChengYou, et al.
Veröffentlicht: (2026)
Positive Alignment: Artificial Intelligence for Human Flourishing
von: Laukkonen, Ruben, et al.
Veröffentlicht: (2026)
von: Laukkonen, Ruben, et al.
Veröffentlicht: (2026)
TodoEvolve: Learning to Architect Agent Planning Systems
von: Liu, Jiaxi, et al.
Veröffentlicht: (2026)
von: Liu, Jiaxi, et al.
Veröffentlicht: (2026)
An Epistemic Perspective on Agent Awareness
von: Naumov, Pavel, et al.
Veröffentlicht: (2025)
von: Naumov, Pavel, et al.
Veröffentlicht: (2025)
Modeling Epistemic Uncertainty in Social Perception via Rashomon Set Agents
von: Yang, Jinming, et al.
Veröffentlicht: (2026)
von: Yang, Jinming, et al.
Veröffentlicht: (2026)
From Aleatoric to Epistemic: Exploring Uncertainty Quantification Techniques in Artificial Intelligence
von: Wang, Tianyang, et al.
Veröffentlicht: (2025)
von: Wang, Tianyang, et al.
Veröffentlicht: (2025)
DynaTrust: Defending Multi-Agent Systems Against Sleeper Agents via Dynamic Trust Graphs
von: Li, Yu, et al.
Veröffentlicht: (2026)
von: Li, Yu, et al.
Veröffentlicht: (2026)
Agentic Inequality
von: Sharp, Matthew, et al.
Veröffentlicht: (2025)
von: Sharp, Matthew, et al.
Veröffentlicht: (2025)
Epistemic Artificial Intelligence is Essential for Machine Learning Models to Truly 'Know When They Do Not Know'
von: Manchingal, Shireen Kudukkil, et al.
Veröffentlicht: (2025)
von: Manchingal, Shireen Kudukkil, et al.
Veröffentlicht: (2025)
Proof-of-Guardrail in AI Agents and What (Not) to Trust from It
von: Jin, Xisen, et al.
Veröffentlicht: (2026)
von: Jin, Xisen, et al.
Veröffentlicht: (2026)
Architecture Without Architects: How AI Coding Agents Shape Software Architecture
von: Konrad, Phongsakon Mark, et al.
Veröffentlicht: (2026)
von: Konrad, Phongsakon Mark, et al.
Veröffentlicht: (2026)
CodeMem: Architecting Reproducible Agents via Dynamic MCP and Procedural Memory
von: Gaurav, Nishant, et al.
Veröffentlicht: (2025)
von: Gaurav, Nishant, et al.
Veröffentlicht: (2025)
Epistemic Filtering and Collective Hallucination: A Jury Theorem for Confidence-Calibrated Agents
von: Karge, Jonas
Veröffentlicht: (2026)
von: Karge, Jonas
Veröffentlicht: (2026)
Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary
von: Wang, Hongru, et al.
Veröffentlicht: (2025)
von: Wang, Hongru, et al.
Veröffentlicht: (2025)
How will advanced AI systems impact democracy?
von: Summerfield, Christopher, et al.
Veröffentlicht: (2024)
von: Summerfield, Christopher, et al.
Veröffentlicht: (2024)
Imitation Learning in the Deep Learning Era: A Novel Taxonomy and Recent Advances
von: Chrysomallis, Iason, et al.
Veröffentlicht: (2025)
von: Chrysomallis, Iason, et al.
Veröffentlicht: (2025)
Defense Against the Dark Prompts: Mitigating Best-of-N Jailbreaking with Prompt Evaluation
von: Armstrong, Stuart, et al.
Veröffentlicht: (2025)
von: Armstrong, Stuart, et al.
Veröffentlicht: (2025)
Architecting Agentic Communities using Design Patterns
von: Milosevic, Zoran, et al.
Veröffentlicht: (2026)
von: Milosevic, Zoran, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data
von: Marchal, Nahema, et al.
Veröffentlicht: (2024) -
We Need a New Ethics for a World of AI Agents
von: Gabriel, Iason, et al.
Veröffentlicht: (2025) -
Virtual Agent Economies
von: Tomasev, Nenad, et al.
Veröffentlicht: (2025) -
Model-Free RL Agents Demonstrate System 1-Like Intentionality
von: Ashton, Hal, et al.
Veröffentlicht: (2025) -
AI-Enhanced Deliberative Democracy and the Future of the Collective Will
von: Revel, Manon, et al.
Veröffentlicht: (2025)