Quantifying stability of non-power-seeking in artificial agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gunter, Evan Ryan, Liokumovich, Yevgeny, Krakovna, Victoria |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Gram: Assessing sabotage propensities via automated alignment auditing
von: Lindner, David, et al.
Veröffentlicht: (2026)
von: Lindner, David, et al.
Veröffentlicht: (2026)
The Smale conjecture and min-max theory
von: Ketover, Daniel, et al.
Veröffentlicht: (2023)
von: Ketover, Daniel, et al.
Veröffentlicht: (2023)
Parametric inequalities and Weyl law for the volume spectrum
von: Guth, Larry, et al.
Veröffentlicht: (2022)
von: Guth, Larry, et al.
Veröffentlicht: (2022)
Will artificial agents pursue power by default?
von: Tarsney, Christian
Veröffentlicht: (2025)
von: Tarsney, Christian
Veröffentlicht: (2025)
Limitations of Agents Simulated by Predictive Models
von: Douglas, Raymond, et al.
Veröffentlicht: (2024)
von: Douglas, Raymond, et al.
Veröffentlicht: (2024)
Length of a closed geodesic in 3-manifolds of positive scalar curvature
von: Liokumovich, Yevgeny, et al.
Veröffentlicht: (2025)
von: Liokumovich, Yevgeny, et al.
Veröffentlicht: (2025)
On the existence of minimal Heegaard surfaces
von: Ketover, Daniel, et al.
Veröffentlicht: (2019)
von: Ketover, Daniel, et al.
Veröffentlicht: (2019)
Supervising the search process produces reliable and generalizable information-seeking agents
von: Xiong, Guangzhi, et al.
Veröffentlicht: (2025)
von: Xiong, Guangzhi, et al.
Veröffentlicht: (2025)
Artificial intelligence is algorithmic mimicry: why artificial "agents" are not (and won't be) proper agents
von: Jaeger, Johannes
Veröffentlicht: (2023)
von: Jaeger, Johannes
Veröffentlicht: (2023)
MedClarify: An information-seeking AI agent for medical diagnosis with case-specific follow-up questions
von: Wong, Hui Min, et al.
Veröffentlicht: (2026)
von: Wong, Hui Min, et al.
Veröffentlicht: (2026)
Quantifying construct validity in large language model evaluations
von: Kearns, Ryan Othniel
Veröffentlicht: (2026)
von: Kearns, Ryan Othniel
Veröffentlicht: (2026)
Quantifying artificial intelligence through algorithmic generalization
von: Ito, Takuya, et al.
Veröffentlicht: (2024)
von: Ito, Takuya, et al.
Veröffentlicht: (2024)
Quantifying non deterministic drift in large language models
von: Nicholson, Claire
Veröffentlicht: (2026)
von: Nicholson, Claire
Veröffentlicht: (2026)
Quantifying Harm
von: Beckers, Sander, et al.
Veröffentlicht: (2022)
von: Beckers, Sander, et al.
Veröffentlicht: (2022)
A Note on the Complexity of the Satisfiability Problem for Graded Modal Logics
von: Kazakov, Yevgeny, et al.
Veröffentlicht: (2009)
von: Kazakov, Yevgeny, et al.
Veröffentlicht: (2009)
Neuromorphic dreaming: A pathway to efficient learning in artificial agents
von: Blakowski, Ingo, et al.
Veröffentlicht: (2024)
von: Blakowski, Ingo, et al.
Veröffentlicht: (2024)
LLM-powered Multi-agent Framework for Goal-oriented Learning in Intelligent Tutoring System
von: Wang, Tianfu, et al.
Veröffentlicht: (2025)
von: Wang, Tianfu, et al.
Veröffentlicht: (2025)
Critic as Lyapunov function (CALF): a model-free, stability-ensuring agent
von: Osinenko, Pavel, et al.
Veröffentlicht: (2024)
von: Osinenko, Pavel, et al.
Veröffentlicht: (2024)
MaD Physics: Evaluating information seeking under constraints in physical environments
von: Jain, Moksh, et al.
Veröffentlicht: (2026)
von: Jain, Moksh, et al.
Veröffentlicht: (2026)
The Philosophic Turn for AI Agents: Replacing centralized digital rhetoric with decentralized truth-seeking
von: Koralus, Philipp
Veröffentlicht: (2025)
von: Koralus, Philipp
Veröffentlicht: (2025)
The impact of behavioral diversity in multi-agent reinforcement learning
von: Bettini, Matteo, et al.
Veröffentlicht: (2024)
von: Bettini, Matteo, et al.
Veröffentlicht: (2024)
Fault Detection for agents on power grid topology optimization: A Comprehensive analysis
von: Lehna, Malte, et al.
Veröffentlicht: (2024)
von: Lehna, Malte, et al.
Veröffentlicht: (2024)
Learning to Ground Existentially Quantified Goals
von: Funkquist, Martin, et al.
Veröffentlicht: (2024)
von: Funkquist, Martin, et al.
Veröffentlicht: (2024)
Improving ensemble extreme precipitation forecasts using generative artificial intelligence
von: Sha, Yingkai, et al.
Veröffentlicht: (2024)
von: Sha, Yingkai, et al.
Veröffentlicht: (2024)
AIstorian lets AI be a historian: A KG-powered multi-agent system for accurate biography generation
von: Li, Fengyu, et al.
Veröffentlicht: (2025)
von: Li, Fengyu, et al.
Veröffentlicht: (2025)
Efficient Data Generation for Source-grounded Information-seeking Dialogs: A Use Case for Meeting Transcripts
von: Golany, Lotem, et al.
Veröffentlicht: (2024)
von: Golany, Lotem, et al.
Veröffentlicht: (2024)
Quantifying Model Uniqueness in Heterogeneous AI Ecosystems
von: You, Lei
Veröffentlicht: (2026)
von: You, Lei
Veröffentlicht: (2026)
Quantifying and Optimizing Simplicity via Polynomial Representations
von: Zhang, Tianren, et al.
Veröffentlicht: (2026)
von: Zhang, Tianren, et al.
Veröffentlicht: (2026)
Simulated patient systems powered by large language model-based AI agents offer potential for transforming medical education
von: Yu, Huizi, et al.
Veröffentlicht: (2024)
von: Yu, Huizi, et al.
Veröffentlicht: (2024)
PathFound: An Agentic Multimodal Model Activating Evidence-seeking Pathological Diagnosis
von: Hua, Shengyi, et al.
Veröffentlicht: (2025)
von: Hua, Shengyi, et al.
Veröffentlicht: (2025)
KwaiAgents: Generalized Information-seeking Agent System with Large Language Models
von: Pan, Haojie, et al.
Veröffentlicht: (2023)
von: Pan, Haojie, et al.
Veröffentlicht: (2023)
Grounding Natural Language for Multi-agent Decision-Making with Multi-agentic LLMs
von: Huh, Dom, et al.
Veröffentlicht: (2025)
von: Huh, Dom, et al.
Veröffentlicht: (2025)
Configurable multi-agent framework for scalable and realistic testing of llm-based agents
von: Wang, Sai, et al.
Veröffentlicht: (2025)
von: Wang, Sai, et al.
Veröffentlicht: (2025)
Towards Intelligent Geospatial Data Discovery: a knowledge graph-driven multi-agent framework powered by large language models
von: Liu, Ruixiang, et al.
Veröffentlicht: (2026)
von: Liu, Ruixiang, et al.
Veröffentlicht: (2026)
Grade Score: Quantifying LLM Performance in Option Selection
von: Iourovitski, Dmitri
Veröffentlicht: (2024)
von: Iourovitski, Dmitri
Veröffentlicht: (2024)
Quantifying Divergence for Human-AI Collaboration and Cognitive Trust
von: Kural, Müge, et al.
Veröffentlicht: (2023)
von: Kural, Müge, et al.
Veröffentlicht: (2023)
ReEfBench: Quantifying the Reasoning Efficiency of LLMs
von: Fu, Zhizhang, et al.
Veröffentlicht: (2026)
von: Fu, Zhizhang, et al.
Veröffentlicht: (2026)
Stochasticity in Agentic Evaluations: Quantifying Inconsistency with Intraclass Correlation
von: Mustahsan, Zairah, et al.
Veröffentlicht: (2025)
von: Mustahsan, Zairah, et al.
Veröffentlicht: (2025)
Contrastive explanations of BDI agents
von: Winikoff, Michael
Veröffentlicht: (2026)
von: Winikoff, Michael
Veröffentlicht: (2026)
Archive-based Single-Objective Evolutionary Algorithms for Submodular Optimization
von: Neumann, Frank, et al.
Veröffentlicht: (2024)
von: Neumann, Frank, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Gram: Assessing sabotage propensities via automated alignment auditing
von: Lindner, David, et al.
Veröffentlicht: (2026) -
The Smale conjecture and min-max theory
von: Ketover, Daniel, et al.
Veröffentlicht: (2023) -
Parametric inequalities and Weyl law for the volume spectrum
von: Guth, Larry, et al.
Veröffentlicht: (2022) -
Will artificial agents pursue power by default?
von: Tarsney, Christian
Veröffentlicht: (2025) -
Limitations of Agents Simulated by Predictive Models
von: Douglas, Raymond, et al.
Veröffentlicht: (2024)