A computational framework for human values
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Osman, Nardine, d'Inverno, Mark |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Modelling Human Values for AI Reasoning
par: Osman, Nardine, et autres
Publié: (2024)
par: Osman, Nardine, et autres
Publié: (2024)
Value-Aware Multiagent Systems
par: Osman, Nardine
Publié: (2025)
par: Osman, Nardine
Publié: (2025)
Instilling Organisational Values in Firefighters through Simulation-Based Training
par: Osman, Nardine, et autres
Publié: (2025)
par: Osman, Nardine, et autres
Publié: (2025)
A Hormetic Approach to the Value-Loading Problem: Preventing the Paperclip Apocalypse?
par: Henry, Nathan I. N., et autres
Publié: (2024)
par: Henry, Nathan I. N., et autres
Publié: (2024)
A Modular Cognitive Architecture for Assisted Reasoning: The Nemosine Framework
par: Melo, Edervaldo
Publié: (2025)
par: Melo, Edervaldo
Publié: (2025)
Interpretation, Learning, and Empathy as One Constraint: A Residual-Adequacy Architecture with Accountable Abstention
par: Amornbunchornvej, Chainarong
Publié: (2026)
par: Amornbunchornvej, Chainarong
Publié: (2026)
Taxonomy to Regulation: A (Geo)Political Taxonomy for AI Risks and Regulatory Measures in the EU AI Act
par: Arda, Sinan
Publié: (2024)
par: Arda, Sinan
Publié: (2024)
A Theoretical Framework for Adaptive Utility-Weighted Benchmarking
par: Waggoner, Philip
Publié: (2026)
par: Waggoner, Philip
Publié: (2026)
A Taxonomy of Omnicidal Futures Involving Artificial Intelligence
par: Critch, Andrew, et autres
Publié: (2025)
par: Critch, Andrew, et autres
Publié: (2025)
The Specification Trap: Why Static Value Alignment Alone Is Insufficient for Robust Alignment
par: Spizzirri, Austin
Publié: (2025)
par: Spizzirri, Austin
Publié: (2025)
Feature Relevancy, Necessity and Usefulness: Complexity and Algorithms
par: Capdevielle, Tomás, et autres
Publié: (2025)
par: Capdevielle, Tomás, et autres
Publié: (2025)
From Language Models to Practical Self-Improving Computer Agents
par: Sheng, Alex
Publié: (2024)
par: Sheng, Alex
Publié: (2024)
Synthetic emotions and consciousness: exploring architectural boundaries
par: Borotschnig, Hermann
Publié: (2025)
par: Borotschnig, Hermann
Publié: (2025)
Return of the Schema: Building Complete Datasets for Machine Learning and Reasoning on Knowledge Graphs
par: Diliso, Ivan, et autres
Publié: (2026)
par: Diliso, Ivan, et autres
Publié: (2026)
Unlocking the Potential of Metaverse in Innovative and Immersive Digital Health
par: Ebrahimzadeh, Fatemeh, et autres
Publié: (2024)
par: Ebrahimzadeh, Fatemeh, et autres
Publié: (2024)
Compressible Softmax-Attended Language under Incompressible Attention
par: Lee, Wonsuk
Publié: (2026)
par: Lee, Wonsuk
Publié: (2026)
Dynamic Observation Policies in Observation Cost-Sensitive Reinforcement Learning
par: Bellinger, Colin, et autres
Publié: (2023)
par: Bellinger, Colin, et autres
Publié: (2023)
Intelligence as Computation
par: Brock, Oliver
Publié: (2024)
par: Brock, Oliver
Publié: (2024)
Slipstream: Trajectory-Grounded Compaction Validation for Long-Horizon Agents
par: Chen, Zhuofu, et autres
Publié: (2026)
par: Chen, Zhuofu, et autres
Publié: (2026)
Decentralizing Coordination in Open Vehicle Fleets for Scalable and Dynamic Task Allocation
par: Lujak, Marin, et autres
Publié: (2024)
par: Lujak, Marin, et autres
Publié: (2024)
Cognition is All You Need -- The Next Layer of AI Above Large Language Models
par: Spivack, Nova, et autres
Publié: (2024)
par: Spivack, Nova, et autres
Publié: (2024)
The Station: An Open-World Environment for AI-Driven Discovery
par: Chung, Stephen, et autres
Publié: (2025)
par: Chung, Stephen, et autres
Publié: (2025)
Quantifying Behavioral Dissimilarity Between Mathematical Expressions
par: Mežnar, Sebastian, et autres
Publié: (2024)
par: Mežnar, Sebastian, et autres
Publié: (2024)
Charting the Future of Scholarly Knowledge with AI: A Community Perspective
par: Jiomekong, Azanzi, et autres
Publié: (2025)
par: Jiomekong, Azanzi, et autres
Publié: (2025)
Heckerthoughts
par: Heckerman, David
Publié: (2023)
par: Heckerman, David
Publié: (2023)
On the Invariants of Softmax Attention
par: Lee, Wonsuk
Publié: (2026)
par: Lee, Wonsuk
Publié: (2026)
Achieving Distributive Justice in Federated Learning via Uncertainty Quantification
par: Carey, Alycia, et autres
Publié: (2025)
par: Carey, Alycia, et autres
Publié: (2025)
ATEX-CF: Attack-Informed Counterfactual Explanations for Graph Neural Networks
par: Zhang, Yu, et autres
Publié: (2026)
par: Zhang, Yu, et autres
Publié: (2026)
Attack Selection Reduces Safety in Concentrated AI Control Settings against Trusted Monitoring
par: Schaeffer, Joachim, et autres
Publié: (2026)
par: Schaeffer, Joachim, et autres
Publié: (2026)
Personality-Driven Decision-Making in LLM-Based Autonomous Agents
par: Newsham, Lewis, et autres
Publié: (2025)
par: Newsham, Lewis, et autres
Publié: (2025)
An Automatic Text Classification Method Based on Hierarchical Taxonomies, Neural Networks and Document Embedding: The NETHIC Tool
par: Lomasto, Luigi, et autres
Publié: (2026)
par: Lomasto, Luigi, et autres
Publié: (2026)
Generative AI Usage of University Students: Navigating Between Education and Business
par: Walke, Fabian, et autres
Publié: (2026)
par: Walke, Fabian, et autres
Publié: (2026)
VACoDe: Visual Augmented Contrastive Decoding
par: Kim, Sihyeon, et autres
Publié: (2024)
par: Kim, Sihyeon, et autres
Publié: (2024)
A Mixed User-Centered Approach to Enable Augmented Intelligence in Intelligent Tutoring Systems: The Case of MathAIde app
par: Guerino, Guilherme, et autres
Publié: (2025)
par: Guerino, Guilherme, et autres
Publié: (2025)
On Privacy Leakage in Tabular Diffusion Models: Influential Factors, Attacker Knowledge, and Metrics
par: Shafieinejad, Masoumeh, et autres
Publié: (2026)
par: Shafieinejad, Masoumeh, et autres
Publié: (2026)
Data and AI governance: Promoting equity, ethics, and fairness in large language models
par: Abhishek, Alok, et autres
Publié: (2025)
par: Abhishek, Alok, et autres
Publié: (2025)
SHARP: Social Harm Analysis via Risk Profiles for Measuring Inequities in Large Language Models
par: Abhishek, Alok, et autres
Publié: (2026)
par: Abhishek, Alok, et autres
Publié: (2026)
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
par: Abhishek, Alok, et autres
Publié: (2025)
par: Abhishek, Alok, et autres
Publié: (2025)
Reasoning Promotes Robustness in Theory of Mind Tasks
par: de Haan, Ian B., et autres
Publié: (2026)
par: de Haan, Ian B., et autres
Publié: (2026)
Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems
par: Berman, Shmuel, et autres
Publié: (2024)
par: Berman, Shmuel, et autres
Publié: (2024)
Documents similaires
-
Modelling Human Values for AI Reasoning
par: Osman, Nardine, et autres
Publié: (2024) -
Value-Aware Multiagent Systems
par: Osman, Nardine
Publié: (2025) -
Instilling Organisational Values in Firefighters through Simulation-Based Training
par: Osman, Nardine, et autres
Publié: (2025) -
A Hormetic Approach to the Value-Loading Problem: Preventing the Paperclip Apocalypse?
par: Henry, Nathan I. N., et autres
Publié: (2024) -
A Modular Cognitive Architecture for Assisted Reasoning: The Nemosine Framework
par: Melo, Edervaldo
Publié: (2025)