Towards a Theory of AI Personhood
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Ward, Francis Rhys |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Password-Activated Shutdown Protocols for Misaligned Frontier Agents
von: Williams, Kai, et al.
Veröffentlicht: (2025)
von: Williams, Kai, et al.
Veröffentlicht: (2025)
The Elicitation Game: Evaluating Capability Elicitation Techniques
von: Hofstätter, Felix, et al.
Veröffentlicht: (2025)
von: Hofstätter, Felix, et al.
Veröffentlicht: (2025)
AI Sandbagging: Language Models can Strategically Underperform on Evaluations
von: van der Weij, Teun, et al.
Veröffentlicht: (2024)
von: van der Weij, Teun, et al.
Veröffentlicht: (2024)
Continuous-Time Analysis of Adaptive Optimization and Normalization
von: Gould, Rhys, et al.
Veröffentlicht: (2024)
von: Gould, Rhys, et al.
Veröffentlicht: (2024)
Large language models as uncertainty-calibrated optimizers for experimental discovery
von: Ranković, Bojana, et al.
Veröffentlicht: (2025)
von: Ranković, Bojana, et al.
Veröffentlicht: (2025)
Towards a Learning Theory of Representation Alignment
von: Insulla, Francesco, et al.
Veröffentlicht: (2025)
von: Insulla, Francesco, et al.
Veröffentlicht: (2025)
Towards a Medical AI Scientist
von: Wu, Hongtao, et al.
Veröffentlicht: (2026)
von: Wu, Hongtao, et al.
Veröffentlicht: (2026)
Towards a More Complete Theory of Function Preserving Transforms
von: Painter, Michael
Veröffentlicht: (2024)
von: Painter, Michael
Veröffentlicht: (2024)
A Pragmatic View of AI Personhood
von: Leibo, Joel Z., et al.
Veröffentlicht: (2025)
von: Leibo, Joel Z., et al.
Veröffentlicht: (2025)
Towards a Formal Creativity Theory: Preliminary results in Novelty and Transformativeness
von: Santo, Luís Espírito, et al.
Veröffentlicht: (2024)
von: Santo, Luís Espírito, et al.
Veröffentlicht: (2024)
Interpretable Representations in Explainable AI: From Theory to Practice
von: Sokol, Kacper, et al.
Veröffentlicht: (2020)
von: Sokol, Kacper, et al.
Veröffentlicht: (2020)
ML-Master: Towards AI-for-AI via Integration of Exploration and Reasoning
von: Liu, Zexi, et al.
Veröffentlicht: (2025)
von: Liu, Zexi, et al.
Veröffentlicht: (2025)
Dispelling the Mirage of Progress in Offline MARL through Standardised Baselines and Evaluation
von: Formanek, Claude, et al.
Veröffentlicht: (2024)
von: Formanek, Claude, et al.
Veröffentlicht: (2024)
Towards Effective Theory of LLMs: A Representation Learning Approach
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2026)
von: Ustaomeroglu, Muhammed, et al.
Veröffentlicht: (2026)
Toward an Evaluation Science for Generative AI Systems
von: Weidinger, Laura, et al.
Veröffentlicht: (2025)
von: Weidinger, Laura, et al.
Veröffentlicht: (2025)
Towards Unified Attribution in Explainable AI, Data-Centric AI, and Mechanistic Interpretability
von: Zhang, Shichang, et al.
Veröffentlicht: (2025)
von: Zhang, Shichang, et al.
Veröffentlicht: (2025)
Towards a perturbation-based explanation for medical AI as differentiable programs
von: Abe, Takeshi, et al.
Veröffentlicht: (2025)
von: Abe, Takeshi, et al.
Veröffentlicht: (2025)
INSPIRE-GNN: Intelligent Sensor Placement to Improve Sparse Bicycling Network Prediction via Reinforcement Learning Boosted Graph Neural Networks
von: Gupta, Mohit, et al.
Veröffentlicht: (2025)
von: Gupta, Mohit, et al.
Veröffentlicht: (2025)
Learning the Preferences of a Learning Agent
von: Sadek, Karim Abdel, et al.
Veröffentlicht: (2026)
von: Sadek, Karim Abdel, et al.
Veröffentlicht: (2026)
Diversifying AI: Towards Creative Chess with AlphaZero
von: Zahavy, Tom, et al.
Veröffentlicht: (2023)
von: Zahavy, Tom, et al.
Veröffentlicht: (2023)
Towards certifiable AI in aviation: landscape, challenges, and opportunities
von: Bello, Hymalai, et al.
Veröffentlicht: (2024)
von: Bello, Hymalai, et al.
Veröffentlicht: (2024)
Tutorial on the Probabilistic Unification of Estimation Theory, Machine Learning, and Generative AI
von: Elmusrati, Mohammed
Veröffentlicht: (2025)
von: Elmusrati, Mohammed
Veröffentlicht: (2025)
An AI Architecture with the Capability to Classify and Explain Hardware Trojans
von: Whitten, Paul, et al.
Veröffentlicht: (2024)
von: Whitten, Paul, et al.
Veröffentlicht: (2024)
Towards Machine Theory of Mind with Large Language Model-Augmented Inverse Planning
von: Gelpí, Rebekah A., et al.
Veröffentlicht: (2025)
von: Gelpí, Rebekah A., et al.
Veröffentlicht: (2025)
xEEGNet: Towards Explainable AI in EEG Dementia Classification
von: Zanola, Andrea, et al.
Veröffentlicht: (2025)
von: Zanola, Andrea, et al.
Veröffentlicht: (2025)
Towards deployment-centric multimodal AI beyond vision and language
von: Liu, Xianyuan, et al.
Veröffentlicht: (2025)
von: Liu, Xianyuan, et al.
Veröffentlicht: (2025)
Curie: Toward Rigorous and Automated Scientific Experimentation with AI Agents
von: Kon, Patrick Tser Jern, et al.
Veröffentlicht: (2025)
von: Kon, Patrick Tser Jern, et al.
Veröffentlicht: (2025)
Towards Scientific Discovery with Generative AI: Progress, Opportunities, and Challenges
von: Reddy, Chandan K, et al.
Veröffentlicht: (2024)
von: Reddy, Chandan K, et al.
Veröffentlicht: (2024)
SEAL: Towards Safe Autonomous Driving via Skill-Enabled Adversary Learning for Closed-Loop Scenario Generation
von: Stoler, Benjamin, et al.
Veröffentlicht: (2024)
von: Stoler, Benjamin, et al.
Veröffentlicht: (2024)
Sheaf-Theoretic Transport and Obstruction for Detecting Scientific Theory Shift in AI Agents
von: Olivieri, David N., et al.
Veröffentlicht: (2026)
von: Olivieri, David N., et al.
Veröffentlicht: (2026)
Rank-1 LoRAs Encode Interpretable Reasoning Signals
von: Ward, Jake, et al.
Veröffentlicht: (2025)
von: Ward, Jake, et al.
Veröffentlicht: (2025)
Correcting Human Labels for Rater Effects in AI Evaluation: An Item Response Theory Approach
von: Casabianca, Jodi M., et al.
Veröffentlicht: (2026)
von: Casabianca, Jodi M., et al.
Veröffentlicht: (2026)
The Galerkin method beats Graph-Based Approaches for Spectral Algorithms
von: Cabannes, Vivien, et al.
Veröffentlicht: (2023)
von: Cabannes, Vivien, et al.
Veröffentlicht: (2023)
Are Large Language Models Useful for Time Series Data Analysis?
von: Tang, Francis, et al.
Veröffentlicht: (2024)
von: Tang, Francis, et al.
Veröffentlicht: (2024)
Interpretable AI for Time-Series: Multi-Model Heatmap Fusion with Global Attention and NLP-Generated Explanations
von: Francis, Jiztom Kavalakkatt, et al.
Veröffentlicht: (2025)
von: Francis, Jiztom Kavalakkatt, et al.
Veröffentlicht: (2025)
Variance-Bounded Evaluation of Entity-Centric AI Systems Without Ground Truth: Theory and Measurement
von: Ding, Kaihua
Veröffentlicht: (2025)
von: Ding, Kaihua
Veröffentlicht: (2025)
Towards a Science of AI Agent Reliability
von: Rabanser, Stephan, et al.
Veröffentlicht: (2026)
von: Rabanser, Stephan, et al.
Veröffentlicht: (2026)
Neural Nonmyopic Bayesian Optimization in Dynamic Cost Settings
von: Truong, Sang T., et al.
Veröffentlicht: (2026)
von: Truong, Sang T., et al.
Veröffentlicht: (2026)
Position: AI Evaluations Should be Grounded on a Theory of Capability
von: Jo, Nathanael, et al.
Veröffentlicht: (2025)
von: Jo, Nathanael, et al.
Veröffentlicht: (2025)
Towards Green AI in Fine-tuning Large Language Models via Adaptive Backpropagation
von: Huang, Kai, et al.
Veröffentlicht: (2023)
von: Huang, Kai, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Password-Activated Shutdown Protocols for Misaligned Frontier Agents
von: Williams, Kai, et al.
Veröffentlicht: (2025) -
The Elicitation Game: Evaluating Capability Elicitation Techniques
von: Hofstätter, Felix, et al.
Veröffentlicht: (2025) -
AI Sandbagging: Language Models can Strategically Underperform on Evaluations
von: van der Weij, Teun, et al.
Veröffentlicht: (2024) -
Continuous-Time Analysis of Adaptive Optimization and Normalization
von: Gould, Rhys, et al.
Veröffentlicht: (2024) -
Large language models as uncertainty-calibrated optimizers for experimental discovery
von: Ranković, Bojana, et al.
Veröffentlicht: (2025)