CERN for AI: A Theoretical Framework for Autonomous Simulation-Based Artificial Intelligence Testing and Alignment
Fuente:
arXiv
Guardado en:
| Autores principales: | Bojic, Ljubisa, Cinelli, Matteo, Culibrk, Dubravko, Delibasic, Boris |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AI of the People, by the People, for the People: A Social Choice Approach to Collective Control of Artificial Intelligence
por: Bachmann, Paul Anton, et al.
Publicado: (2026)
por: Bachmann, Paul Anton, et al.
Publicado: (2026)
A Game-Theoretic Negotiation Framework for Cross-Cultural Consensus in LLMs
por: Zhang, Guoxi, et al.
Publicado: (2025)
por: Zhang, Guoxi, et al.
Publicado: (2025)
Auction-Based Regulation for Artificial Intelligence
por: Bornstein, Marco, et al.
Publicado: (2024)
por: Bornstein, Marco, et al.
Publicado: (2024)
AI Arms and Influence: Frontier Models Exhibit Sophisticated Reasoning in Simulated Nuclear Crises
por: Payne, Kenneth
Publicado: (2026)
por: Payne, Kenneth
Publicado: (2026)
Alignment as Institutional Design: From Behavioral Correction to Transaction Structure in Intelligent Systems
por: Chai, Rui
Publicado: (2026)
por: Chai, Rui
Publicado: (2026)
Competition and Diversity in Generative AI
por: Raghavan, Manish
Publicado: (2024)
por: Raghavan, Manish
Publicado: (2024)
Path Dependence under Adaptive AI Delegation
por: Huang, Lingxiao, et al.
Publicado: (2026)
por: Huang, Lingxiao, et al.
Publicado: (2026)
Modeling the Economic Impacts of AI Openness Regulation
por: Qiu, Tori, et al.
Publicado: (2025)
por: Qiu, Tori, et al.
Publicado: (2025)
Does GPT-4 surpass human performance in linguistic pragmatics?
por: Bojic, Ljubisa, et al.
Publicado: (2023)
por: Bojic, Ljubisa, et al.
Publicado: (2023)
Designing Algorithmic Delegates: The Role of Indistinguishability in Human-AI Handoff
por: Greenwood, Sophie, et al.
Publicado: (2025)
por: Greenwood, Sophie, et al.
Publicado: (2025)
Game Theory Meets LLM and Agentic AI: Reimagining Cybersecurity for the Age of Intelligent Threats
por: Zhu, Quanyan
Publicado: (2025)
por: Zhu, Quanyan
Publicado: (2025)
Representative Social Choice: From Learning Theory to AI Alignment
por: Qiu, Tianyi
Publicado: (2024)
por: Qiu, Tianyi
Publicado: (2024)
From Outcome-Based to Language-Based Preferences
por: Capraro, Valerio, et al.
Publicado: (2022)
por: Capraro, Valerio, et al.
Publicado: (2022)
Test-Time Compute Games
por: Velasco, Ander Artola, et al.
Publicado: (2026)
por: Velasco, Ander Artola, et al.
Publicado: (2026)
AgentSociety: Incentivizing Agentic Social Intelligence
por: Kesari, Aditya Vema Reddy, et al.
Publicado: (2026)
por: Kesari, Aditya Vema Reddy, et al.
Publicado: (2026)
A Learning Framework for Distribution-Based Game-Theoretic Solution Concepts
por: Jha, Tushant, et al.
Publicado: (2019)
por: Jha, Tushant, et al.
Publicado: (2019)
Integrated Design and Governance of Agentic AI Systems through Adaptive Information Modulation
por: Chen, Qiliang, et al.
Publicado: (2024)
por: Chen, Qiliang, et al.
Publicado: (2024)
Prompting Fairness: Artificial Intelligence as Game Players
por: Henry, Jazmia
Publicado: (2024)
por: Henry, Jazmia
Publicado: (2024)
AI's assigned gender affects human-AI cooperation
por: Bazazi, Sepideh, et al.
Publicado: (2024)
por: Bazazi, Sepideh, et al.
Publicado: (2024)
Delegation and Verification Under AI
por: Huang, Lingxiao, et al.
Publicado: (2026)
por: Huang, Lingxiao, et al.
Publicado: (2026)
Using deep reinforcement learning to promote sustainable human behaviour on a common pool resource problem
por: Koster, Raphael, et al.
Publicado: (2024)
por: Koster, Raphael, et al.
Publicado: (2024)
Designing DSIC Mechanisms for Data Sharing in the Era of Large Language Models
por: Ayyoubzadeh, Seyed Moein, et al.
Publicado: (2025)
por: Ayyoubzadeh, Seyed Moein, et al.
Publicado: (2025)
Multi-Agent Strategic Games with LLMs
por: Chupilkin, Maxim
Publicado: (2026)
por: Chupilkin, Maxim
Publicado: (2026)
Towards Strategic Persuasion with Language Models
por: Cheng, Zirui, et al.
Publicado: (2025)
por: Cheng, Zirui, et al.
Publicado: (2025)
If It's Nice, Do It Twice: We Should Try Iterative Corpus Curation
por: Young, Robin
Publicado: (2025)
por: Young, Robin
Publicado: (2025)
Interpretable Risk Mitigation in LLM Agent Systems
por: Chojnacki, Jan
Publicado: (2025)
por: Chojnacki, Jan
Publicado: (2025)
Assessing Group Fairness with Social Welfare Optimization
por: Chen, Violet, et al.
Publicado: (2024)
por: Chen, Violet, et al.
Publicado: (2024)
The Backfiring Effect of Weak AI Safety Regulation
por: Laufer, Benjamin, et al.
Publicado: (2025)
por: Laufer, Benjamin, et al.
Publicado: (2025)
The Fair Game: Auditing & Debiasing AI Algorithms Over Time
por: Basu, Debabrota, et al.
Publicado: (2025)
por: Basu, Debabrota, et al.
Publicado: (2025)
Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?
por: Fontana, Nicoló, et al.
Publicado: (2024)
por: Fontana, Nicoló, et al.
Publicado: (2024)
Do LLMs trust AI regulation? Emerging behaviour of game-theoretic LLM agents
por: Buscemi, Alessio, et al.
Publicado: (2025)
por: Buscemi, Alessio, et al.
Publicado: (2025)
A Revealed Preference Framework for AI Alignment
por: Suleymanov, Elchin
Publicado: (2026)
por: Suleymanov, Elchin
Publicado: (2026)
The Dual Impact of Virtual Reality: Examining the Addictive Potential and Therapeutic Applications of Immersive Media in the Metaverse
por: Bojic, Ljubisa, et al.
Publicado: (2024)
por: Bojic, Ljubisa, et al.
Publicado: (2024)
Incentives, Equilibria, and the Limits of Healthcare AI: A Game-Theoretic Perspective
por: Ercole, Ari
Publicado: (2026)
por: Ercole, Ari
Publicado: (2026)
Asymptotic Universal Alignment: A New Alignment Framework via Test-Time Scaling
por: Cai, Yang, et al.
Publicado: (2026)
por: Cai, Yang, et al.
Publicado: (2026)
An AI Theory of Mind Will Enhance Our Collective Intelligence
por: Harré, Michael S., et al.
Publicado: (2024)
por: Harré, Michael S., et al.
Publicado: (2024)
AI Cap-and-Trade: Efficiency Incentives for Accessibility and Sustainability
por: Bornstein, Marco, et al.
Publicado: (2026)
por: Bornstein, Marco, et al.
Publicado: (2026)
AI Testing Should Account for Sophisticated Strategic Behaviour
por: Kovarik, Vojtech, et al.
Publicado: (2025)
por: Kovarik, Vojtech, et al.
Publicado: (2025)
ADAPT: A Game-Theoretic and Neuro-Symbolic Framework for Automated Distributed Adaptive Penetration Testing
por: Lei, Haozhe, et al.
Publicado: (2024)
por: Lei, Haozhe, et al.
Publicado: (2024)
Towards Recommender Systems LLMs Playground (RecSysLLMsP): Exploring Polarization and Engagement in Simulated Social Networks
por: Bojic, Ljubisa, et al.
Publicado: (2025)
por: Bojic, Ljubisa, et al.
Publicado: (2025)
Ejemplares similares
-
AI of the People, by the People, for the People: A Social Choice Approach to Collective Control of Artificial Intelligence
por: Bachmann, Paul Anton, et al.
Publicado: (2026) -
A Game-Theoretic Negotiation Framework for Cross-Cultural Consensus in LLMs
por: Zhang, Guoxi, et al.
Publicado: (2025) -
Auction-Based Regulation for Artificial Intelligence
por: Bornstein, Marco, et al.
Publicado: (2024) -
AI Arms and Influence: Frontier Models Exhibit Sophisticated Reasoning in Simulated Nuclear Crises
por: Payne, Kenneth
Publicado: (2026) -
Alignment as Institutional Design: From Behavioral Correction to Transaction Structure in Intelligent Systems
por: Chai, Rui
Publicado: (2026)