Dialogical Reasoning Across AI Architectures: A Multi-Model Framework for Testing AI Alignment Strategies
Fuente:
arXiv
Guardado en:
| Autor principal: | Cox, Gray |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Penetration Testing of Agentic AI: A Comparative Security Analysis Across Models and Frameworks
por: Nguyen, Viet K., et al.
Publicado: (2025)
por: Nguyen, Viet K., et al.
Publicado: (2025)
AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems
por: Badagi, Chitra, et al.
Publicado: (2026)
por: Badagi, Chitra, et al.
Publicado: (2026)
Towards Dialogues for Joint Human-AI Reasoning and Value Alignment
por: Bezou-Vrakatseli, Elfia, et al.
Publicado: (2024)
por: Bezou-Vrakatseli, Elfia, et al.
Publicado: (2024)
Rules, Cases, and Reasoning: Positivist Legal Theory as a Framework for Pluralistic AI Alignment
por: Caputo, Nicholas A.
Publicado: (2024)
por: Caputo, Nicholas A.
Publicado: (2024)
Script-Based Dialog Policy Planning for LLM-Powered Conversational Agents: A Basic Architecture for an "AI Therapist"
por: Wasenmüller, Robert, et al.
Publicado: (2024)
por: Wasenmüller, Robert, et al.
Publicado: (2024)
Privacy Ethics Alignment in AI: A Stakeholder-Centric Framework for Ethical AI
por: Barthwal, Ankur, et al.
Publicado: (2025)
por: Barthwal, Ankur, et al.
Publicado: (2025)
Do AI Models Perform Human-like Abstract Reasoning Across Modalities?
por: Beger, Claas, et al.
Publicado: (2025)
por: Beger, Claas, et al.
Publicado: (2025)
To Mask or to Mirror: Human-AI Alignment in Collective Reasoning
por: Qian, Crystal, et al.
Publicado: (2025)
por: Qian, Crystal, et al.
Publicado: (2025)
Bridging AI and Clinical Reasoning: Abductive Explanations for Alignment on Critical Symptoms
por: Sonna, Belona, et al.
Publicado: (2026)
por: Sonna, Belona, et al.
Publicado: (2026)
Agentic AI Frameworks: Architectures, Protocols, and Design Challenges
por: Derouiche, Hana, et al.
Publicado: (2025)
por: Derouiche, Hana, et al.
Publicado: (2025)
Modular Speaker Architecture: A Framework for Sustaining Responsibility and Contextual Integrity in Multi-Agent AI Communication
por: Toh, Khe-Han, et al.
Publicado: (2025)
por: Toh, Khe-Han, et al.
Publicado: (2025)
AI-Compass: A Comprehensive and Effective Multi-module Testing Tool for AI Systems
por: Zhu, Zhiyu, et al.
Publicado: (2024)
por: Zhu, Zhiyu, et al.
Publicado: (2024)
HySafe-AI: Hybrid Safety Architectural Analysis Framework for AI Systems: A Case Study
por: Pitale, Mandar, et al.
Publicado: (2025)
por: Pitale, Mandar, et al.
Publicado: (2025)
The Value of Gen-AI Conversations: A bottom-up Framework for AI Value Alignment
por: Motnikar, Lenart, et al.
Publicado: (2025)
por: Motnikar, Lenart, et al.
Publicado: (2025)
Breaking Up with Normatively Monolithic Agency with GRACE: A Reason-Based Neuro-Symbolic Architecture for Safe and Ethical AI Alignment
por: Jahn, Felix, et al.
Publicado: (2026)
por: Jahn, Felix, et al.
Publicado: (2026)
Match Point AI: A Novel AI Framework for Evaluating Data-Driven Tennis Strategies
por: Nübel, Carlo, et al.
Publicado: (2024)
por: Nübel, Carlo, et al.
Publicado: (2024)
Conversation AI Dialog for Medicare powered by Finetuning and Retrieval Augmented Generation
por: Agrawal, Atharva Mangeshkumar, et al.
Publicado: (2025)
por: Agrawal, Atharva Mangeshkumar, et al.
Publicado: (2025)
A Context Alignment Pre-processor for Enhancing the Coherence of Human-LLM Dialog
por: Wei, Ding
Publicado: (2026)
por: Wei, Ding
Publicado: (2026)
Architectural Constraints Alignment in AI-assisted, Platform-based Service Development
por: Irion, Julius, et al.
Publicado: (2026)
por: Irion, Julius, et al.
Publicado: (2026)
Data Poisoning Vulnerabilities Across Healthcare AI Architectures: A Security Threat Analysis
por: Abtahi, Farhad, et al.
Publicado: (2025)
por: Abtahi, Farhad, et al.
Publicado: (2025)
AI and Human Oversight: A Risk-Based Framework for Alignment
por: Kandikatla, Laxmiraju, et al.
Publicado: (2025)
por: Kandikatla, Laxmiraju, et al.
Publicado: (2025)
Real-Time AI Service Economy: A Framework for Agentic Computing Across the Continuum
por: Lovén, Lauri, et al.
Publicado: (2026)
por: Lovén, Lauri, et al.
Publicado: (2026)
AI Alignment: A Comprehensive Survey
por: Ji, Jiaming, et al.
Publicado: (2023)
por: Ji, Jiaming, et al.
Publicado: (2023)
Strategies of Code-switching in Human-Machine Dialogs
por: Geckt, Dean, et al.
Publicado: (2025)
por: Geckt, Dean, et al.
Publicado: (2025)
System 2 Reasoning for Human-AI Alignment: Generality and Adaptivity via ARC-AGI
por: Kim, Sejin, et al.
Publicado: (2024)
por: Kim, Sejin, et al.
Publicado: (2024)
Beyond Preferences in AI Alignment
por: Zhi-Xuan, Tan, et al.
Publicado: (2024)
por: Zhi-Xuan, Tan, et al.
Publicado: (2024)
Evaluating AI Alignment in LLMs: Output Analysis of Value Priorities Across 75 Models with Human Benchmarking
por: Lau, Gabriel Rongyang, et al.
Publicado: (2025)
por: Lau, Gabriel Rongyang, et al.
Publicado: (2025)
Moral Anchor System: A Predictive Framework for AI Value Alignment and Drift Prevention
por: Ravindran, Santhosh Kumar
Publicado: (2025)
por: Ravindran, Santhosh Kumar
Publicado: (2025)
Multi-level Value Alignment in Agentic AI Systems: Survey and Perspectives
por: Zeng, Wei, et al.
Publicado: (2025)
por: Zeng, Wei, et al.
Publicado: (2025)
MAGMA: A Multi-Graph based Agentic Memory Architecture for AI Agents
por: Jiang, Dongming, et al.
Publicado: (2026)
por: Jiang, Dongming, et al.
Publicado: (2026)
The Impact of AI on Educational Assessment: A Framework for Constructive Alignment
por: Stokkink, Patrick
Publicado: (2025)
por: Stokkink, Patrick
Publicado: (2025)
A Revealed Preference Framework for AI Alignment
por: Suleymanov, Elchin
Publicado: (2026)
por: Suleymanov, Elchin
Publicado: (2026)
Maia-2: A Unified Model for Human-AI Alignment in Chess
por: Tang, Zhenwei, et al.
Publicado: (2024)
por: Tang, Zhenwei, et al.
Publicado: (2024)
TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools
por: Gao, Shanghua, et al.
Publicado: (2025)
por: Gao, Shanghua, et al.
Publicado: (2025)
CERN for AI: A Theoretical Framework for Autonomous Simulation-Based Artificial Intelligence Testing and Alignment
por: Bojic, Ljubisa, et al.
Publicado: (2023)
por: Bojic, Ljubisa, et al.
Publicado: (2023)
Pen-Strategist: A Reasoning Framework for Penetration Testing Strategy Formation and Analysis
por: Ginige, Yasod, et al.
Publicado: (2026)
por: Ginige, Yasod, et al.
Publicado: (2026)
Analysis Of Linguistic Stereotypes in Single and Multi-Agent Generative AI Architectures
por: Ullasci, Martina, et al.
Publicado: (2026)
por: Ullasci, Martina, et al.
Publicado: (2026)
Rationalize: Shared Semantic Reasoning for Human-AI Alignment
por: Dasgupta, Aritra, et al.
Publicado: (2026)
por: Dasgupta, Aritra, et al.
Publicado: (2026)
AI Alignment Strategies from a Risk Perspective: Independent Safety Mechanisms or Shared Failures?
por: Dung, Leonard, et al.
Publicado: (2025)
por: Dung, Leonard, et al.
Publicado: (2025)
DiffusionDialog: A Diffusion Model for Diverse Dialog Generation with Latent Space
por: Xiang, Jianxiang, et al.
Publicado: (2024)
por: Xiang, Jianxiang, et al.
Publicado: (2024)
Ejemplares similares
-
Penetration Testing of Agentic AI: A Comparative Security Analysis Across Models and Frameworks
por: Nguyen, Viet K., et al.
Publicado: (2025) -
AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems
por: Badagi, Chitra, et al.
Publicado: (2026) -
Towards Dialogues for Joint Human-AI Reasoning and Value Alignment
por: Bezou-Vrakatseli, Elfia, et al.
Publicado: (2024) -
Rules, Cases, and Reasoning: Positivist Legal Theory as a Framework for Pluralistic AI Alignment
por: Caputo, Nicholas A.
Publicado: (2024) -
Script-Based Dialog Policy Planning for LLM-Powered Conversational Agents: A Basic Architecture for an "AI Therapist"
por: Wasenmüller, Robert, et al.
Publicado: (2024)