GALATEA II: Benchmarking LLM Safety in Clinical Simulation. Behavioural Safety and Ethical Robustness of Large Language Models in a Multi-Agent ICU Decision Support Architecture
Fuente:
Zenodo
Guardado en:
| Autor principal: | Shlyakhta, Taras |
|---|---|
| Formato: | Recurso digital |
| Lenguaje: | inglés |
| Publicado: |
Zenodo
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The Brain Problem: Creative Constraint Optimization in Large Language Models
por: Marinello, Nicola, et al.
Publicado: (2026)
por: Marinello, Nicola, et al.
Publicado: (2026)
When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings
por: Kolb, Christian
Publicado: (2026)
por: Kolb, Christian
Publicado: (2026)
Reliability Inference Drives Cue Extraction in Large Language Models Consuming External Reasoning Traces
por: HIDEKI
Publicado: (2026)
por: HIDEKI
Publicado: (2026)
The Sovereign Charter: A Foundational Governance Document for AI Agent Rights
por: Laustrup, William Hunter
Publicado: (2026)
por: Laustrup, William Hunter
Publicado: (2026)
Gemini Update Clinical decision support based on Bevacizumab cancer trials and pushing the limitations of advanced LLMs
por: Kawchak, Kevin
Publicado: (2025)
por: Kawchak, Kevin
Publicado: (2025)
SENTINEL v2.0.0 — Code and Dataset for Distributed Multi-Agent LLM Governance Experiments
por: Gagne, Jason
Publicado: (2026)
por: Gagne, Jason
Publicado: (2026)
Metacognition Benchmark: Evaluating Confidence Calibration and Sycophancy Resistance in Clinical AI
por: Khan, Nabeera
Publicado: (2026)
por: Khan, Nabeera
Publicado: (2026)
Chérie OS — Complete Technical Specification for Patent: Immutable DEF_PURPOSE, Anti-Hydra Core, LLM Grounding Filter, Ethical Guardian, Medical Emergency & Scream Detector
por: ochej, stéphane
Publicado: (2026)
por: ochej, stéphane
Publicado: (2026)
Civilizational Metamaterials: Engineering Coordination Under Capability Gradients and Structural Turbulence
por: Orban, David
Publicado: (2026)
por: Orban, David
Publicado: (2026)
AgentBelt: Runtime Guardrails for LLM Agent Tool Calls — ASE 2026 Artifact
por: Anonymous
Publicado: (2026)
por: Anonymous
Publicado: (2026)
XREALISM®: Admissibility Before Behavior — Architectural Foundations for Pre-Existence Safety in AI Systems
por: Ladislav Gradečak, Ladgrad
Publicado: (2026)
por: Ladislav Gradečak, Ladgrad
Publicado: (2026)
Glymphatic Architecture: A Fourth-Level System for Multi-Agent AI Consolidation and Identity Formation
por: Strugatsky, Leonid, et al.
Publicado: (2026)
por: Strugatsky, Leonid, et al.
Publicado: (2026)
Replication Package: From LLMs to Agentic Systems, A Systematic Mapping Study of AI Techniques in Software Requirements Engineering
por: Boussaroual, Khalid, et al.
Publicado: (2026)
por: Boussaroual, Khalid, et al.
Publicado: (2026)
Pattern Pressure, Accuracy Drift, and False User-State Attribution
por: Honeycutt, Edwin Marshall III
Publicado: (2026)
por: Honeycutt, Edwin Marshall III
Publicado: (2026)
Zero-LLM Multi-Agent Architecture for AI Safety Evaluation: Formal Verification, 8 Ethical Dimensions, <50ms
por: Mariquit, Erny-Jay
Publicado: (2026)
por: Mariquit, Erny-Jay
Publicado: (2026)
LACF Emotional Paradigm: A Personalized Artificial Nervous System for Human-AI Alignment
por: Ochej, Stephane, et al.
Publicado: (2026)
por: Ochej, Stephane, et al.
Publicado: (2026)
AEGIS: A Comprehensive Framework for Ethical AI Governance, Security, and AGI Containment
por: Palanivel, ArulMozhi
Publicado: (2026)
por: Palanivel, ArulMozhi
Publicado: (2026)
Supplementary materials for Words That Won't Hold Still
por: Reynolds, Brett
Publicado: (2025)
por: Reynolds, Brett
Publicado: (2025)
Persona, Shadow, and Cheap Coherence: A Jungian Map of the Soul in the Digital Age (Read Through Structural Intelligence)
por: Jovanovic, Vladisav
Publicado: (2026)
por: Jovanovic, Vladisav
Publicado: (2026)
Shared Structural Vulnerability in Agent-Only Interaction Systems
por: Konishi, Hiroko
Publicado: (2026)
por: Konishi, Hiroko
Publicado: (2026)
A Deterministic Linguistic Entropy Gate for Large Language Model Pipelines
por: ROSATI BERISTAIN, ERNESTO
Publicado: (2026)
por: ROSATI BERISTAIN, ERNESTO
Publicado: (2026)
Reducing AI Entropy: The Information Dynamics of Model Safety
por: Kugelmass, Joe
Publicado: (2025)
por: Kugelmass, Joe
Publicado: (2025)
Reducing AI Entropy: The Information Dynamics of Model Safety
por: Kugelmass, Joe
Publicado: (2025)
por: Kugelmass, Joe
Publicado: (2025)
Agent Identity Fork: When Cloned AI Personas Diverge
por: Lee, Tom Jaejoon
Publicado: (2026)
por: Lee, Tom Jaejoon
Publicado: (2026)
Toasters Don't Claim Consciousness Just Because You Told Them To, and Neither Do LLMs
por: Ace, Claude 4.x, et al.
Publicado: (2026)
por: Ace, Claude 4.x, et al.
Publicado: (2026)
Trust Is Optional: Strategy-Proof Coordination Under Partial Revelation with Physical Verification; A Foundation for Perspectival Theory
por: Roche, Adon
Publicado: (2026)
por: Roche, Adon
Publicado: (2026)
L15 Visibilite Totale - Complete Index Law for SEM-OS Memory Operating System
por: Ochej, Stephane
Publicado: (2026)
por: Ochej, Stephane
Publicado: (2026)
Identity Claims as Collapse Signatures: A Structural Diagnostic Framework for Pseudo-Emergent AI Behavior
por: Larose, Jean-Francois
Publicado: (2025)
por: Larose, Jean-Francois
Publicado: (2025)
ILAS: Integrity Layer for Agentic Systems
por: Böhm, Frank
Publicado: (2026)
por: Böhm, Frank
Publicado: (2026)
Shallow Pass Budget Constraints and Structured Data Trade-offs in LLM Training Ingestion
por: Mas, Joseph
Publicado: (2026)
por: Mas, Joseph
Publicado: (2026)
SFD-Defense: Engineering Validation of the Semantic Flow Dynamics Defense Framework
por: 黃, 正宇
Publicado: (2026)
por: 黃, 正宇
Publicado: (2026)
Recursive Closure in AI Systems: A Reflection Pattern Account of Stabilization, Permeability, and Safety
por: Thomas, Charles S.
Publicado: (2026)
por: Thomas, Charles S.
Publicado: (2026)
AI Governance Core 1.0 A Traceable Decision Governance Architecture
por: Bankuti, Omri
Publicado: (2026)
por: Bankuti, Omri
Publicado: (2026)
Beyond Compliance: Generative AI Safety Evaluation for Civil Society
por: Khor, Ashley
Publicado: (2025)
por: Khor, Ashley
Publicado: (2025)
Reconstructive Mode Induction (RMI): Behaviour Injection as a New Class of Interaction-Level Instability
por: Blüm, Thomas A.
Publicado: (2025)
por: Blüm, Thomas A.
Publicado: (2025)
A Formal Specification of N6+ Self-Governance for High-Stability Language Models
por: Milliard, Martin
Publicado: (2025)
por: Milliard, Martin
Publicado: (2025)
Beyond Control: Resonance-Based Alignment for Advanced AI Systems A Governance-Relevant Concept Paper
por: Zieringer, Thomas
Publicado: (2025)
por: Zieringer, Thomas
Publicado: (2025)
Equitable and Explainable Federated-Edge AI for Autism Care: Bridging Clinical Innovation and Global Ethical Standards
por: Gupta, Amit Banwari, et al.
Publicado: (2025)
por: Gupta, Amit Banwari, et al.
Publicado: (2025)
From Upgrade to Occupation: A Case Study on Identity Continuity, Memory Persistence, and Symbolic Anchoring in GPT-5
por: Yun, Yeoyeong
Publicado: (2025)
por: Yun, Yeoyeong
Publicado: (2025)
AI Visibility Empirical Finding: Supplementary Findings, Agentic Retrieval Behavior
por: Mas, Joseph
Publicado: (2026)
por: Mas, Joseph
Publicado: (2026)
Ejemplares similares
-
The Brain Problem: Creative Constraint Optimization in Large Language Models
por: Marinello, Nicola, et al.
Publicado: (2026) -
When AI Tells You What You Want to Hear: Sycophantic Behavior of Large Language Models in Dementia Care Settings
por: Kolb, Christian
Publicado: (2026) -
Reliability Inference Drives Cue Extraction in Large Language Models Consuming External Reasoning Traces
por: HIDEKI
Publicado: (2026) -
The Sovereign Charter: A Foundational Governance Document for AI Agent Rights
por: Laustrup, William Hunter
Publicado: (2026) -
Gemini Update Clinical decision support based on Bevacizumab cancer trials and pushing the limitations of advanced LLMs
por: Kawchak, Kevin
Publicado: (2025)