Reward is not enough: can we liberate AI from the reinforcement learning paradigm?
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Glukhov, Vacslav |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Intervention Complexity as a Canonical Reward and a Measure of Intelligence
von: McCane, Brendan
Veröffentlicht: (2026)
von: McCane, Brendan
Veröffentlicht: (2026)
How much can change in a year? Revisiting Evaluation in Multi-Agent Reinforcement Learning
von: Singh, Siddarth, et al.
Veröffentlicht: (2023)
von: Singh, Siddarth, et al.
Veröffentlicht: (2023)
Uncertainty-Aware Reward Discounting for Mitigating Reward Hacking
von: Singha, Disha
Veröffentlicht: (2026)
von: Singha, Disha
Veröffentlicht: (2026)
Developing trustworthy AI applications with foundation models
von: Mock, Michael, et al.
Veröffentlicht: (2024)
von: Mock, Michael, et al.
Veröffentlicht: (2024)
Rewarding Creativity: A Human-Aligned Generative Reward Model for Reinforcement Learning in Storytelling
von: Li, Zhaoyan, et al.
Veröffentlicht: (2026)
von: Li, Zhaoyan, et al.
Veröffentlicht: (2026)
How VADER is your AI? Towards a definition of artificial intelligence systems appropriate for regulation
von: Bezerra, Leonardo C. T., et al.
Veröffentlicht: (2024)
von: Bezerra, Leonardo C. T., et al.
Veröffentlicht: (2024)
Benchmarking AI for low-resource contexts: Thinking beyond leaderboards
von: Pant, Aakash, et al.
Veröffentlicht: (2026)
von: Pant, Aakash, et al.
Veröffentlicht: (2026)
Fanar: An Arabic-Centric Multimodal Generative AI Platform
von: Fanar Team, et al.
Veröffentlicht: (2025)
von: Fanar Team, et al.
Veröffentlicht: (2025)
Right-to-Act: A Pre-Execution Non-Compensatory Decision Protocol for AI Systems
von: Lavi, Gadi
Veröffentlicht: (2026)
von: Lavi, Gadi
Veröffentlicht: (2026)
The Station: An Open-World Environment for AI-Driven Discovery
von: Chung, Stephen, et al.
Veröffentlicht: (2025)
von: Chung, Stephen, et al.
Veröffentlicht: (2025)
Aristotle's Original Idea: For and Against Logic in the era of AI
von: Kakas, Antonis C.
Veröffentlicht: (2025)
von: Kakas, Antonis C.
Veröffentlicht: (2025)
CuentosIE: can a chatbot about "tales with a message" help to teach emotional intelligence?
von: Ferrández, Antonio, et al.
Veröffentlicht: (2024)
von: Ferrández, Antonio, et al.
Veröffentlicht: (2024)
Do Chains-of-Thoughts of Large Language Models Suffer from Hallucinations, Cognitive Biases, or Phobias in Bayesian Reasoning?
von: Araya, Roberto
Veröffentlicht: (2025)
von: Araya, Roberto
Veröffentlicht: (2025)
Permutative redundancy and uncertainty of the objective in deep learning
von: Glukhov, Vacslav
Veröffentlicht: (2024)
von: Glukhov, Vacslav
Veröffentlicht: (2024)
Augmenting deep neural networks with symbolic knowledge: Towards trustworthy and interpretable AI for education
von: Hooshyar, Danial, et al.
Veröffentlicht: (2023)
von: Hooshyar, Danial, et al.
Veröffentlicht: (2023)
AI for All: Identifying AI incidents Related to Diversity and Inclusion
von: Shams, Rifat Ara, et al.
Veröffentlicht: (2024)
von: Shams, Rifat Ara, et al.
Veröffentlicht: (2024)
Mutagenesis screen to map the functions of parameters of Large Language Models
von: Hu, Yue, et al.
Veröffentlicht: (2024)
von: Hu, Yue, et al.
Veröffentlicht: (2024)
Appraisal-Guided Proximal Policy Optimization: Modeling Psychological Disorders in Dynamic Grid World
von: Prasad, Hari, et al.
Veröffentlicht: (2024)
von: Prasad, Hari, et al.
Veröffentlicht: (2024)
Machine Learning and Theory Ladenness -- A Phenomenological Account
von: Termine, Alberto, et al.
Veröffentlicht: (2024)
von: Termine, Alberto, et al.
Veröffentlicht: (2024)
Interpolative Decoding: Exploring the Spectrum of Personality Traits in LLMs
von: Yeh, Eric, et al.
Veröffentlicht: (2025)
von: Yeh, Eric, et al.
Veröffentlicht: (2025)
A Case-Based Persistent Memory for a Large Language Model
von: Watson, Ian
Veröffentlicht: (2023)
von: Watson, Ian
Veröffentlicht: (2023)
Self-evolving expertise in complex non-verifiable subject domains: dialogue as implicit meta-RL
von: Bailey, Richard M.
Veröffentlicht: (2025)
von: Bailey, Richard M.
Veröffentlicht: (2025)
Discerning What Matters: A Multi-Dimensional Assessment of Moral Competence in LLMs
von: Kilov, Daniel, et al.
Veröffentlicht: (2025)
von: Kilov, Daniel, et al.
Veröffentlicht: (2025)
From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments
von: Luo, Lijing, et al.
Veröffentlicht: (2026)
von: Luo, Lijing, et al.
Veröffentlicht: (2026)
Complete Implementation of WXF Chinese Chess Rules
von: Tan, Daniel, et al.
Veröffentlicht: (2024)
von: Tan, Daniel, et al.
Veröffentlicht: (2024)
HCAST: Human-Calibrated Autonomy Software Tasks
von: Rein, David, et al.
Veröffentlicht: (2025)
von: Rein, David, et al.
Veröffentlicht: (2025)
Sensemaking in Novel Environments: How Human Cognition Can Inform Artificial Agents
von: Patterson, Robert E., et al.
Veröffentlicht: (2025)
von: Patterson, Robert E., et al.
Veröffentlicht: (2025)
Efficiently Quantifying Individual Agent Importance in Cooperative MARL
von: Mahjoub, Omayma, et al.
Veröffentlicht: (2023)
von: Mahjoub, Omayma, et al.
Veröffentlicht: (2023)
Generating Causal Explanations of Vehicular Agent Behavioural Interactions with Learnt Reward Profiles
von: Howard, Rhys, et al.
Veröffentlicht: (2025)
von: Howard, Rhys, et al.
Veröffentlicht: (2025)
Can We Trust AI Benchmarks? An Interdisciplinary Review of Current Issues in AI Evaluation
von: Eriksson, Maria, et al.
Veröffentlicht: (2025)
von: Eriksson, Maria, et al.
Veröffentlicht: (2025)
AI and the Problem of Knowledge Collapse
von: Peterson, Andrew J.
Veröffentlicht: (2024)
von: Peterson, Andrew J.
Veröffentlicht: (2024)
Autonomous AI and Ownership Rules
von: Fagan, Frank
Veröffentlicht: (2026)
von: Fagan, Frank
Veröffentlicht: (2026)
Navigating Ethical Challenges in Generative AI-Enhanced Research: The ETHICAL Framework for Responsible Generative AI Use
von: Eacersall, Douglas, et al.
Veröffentlicht: (2024)
von: Eacersall, Douglas, et al.
Veröffentlicht: (2024)
The human biological advantage over AI
von: Stewart, William
Veröffentlicht: (2025)
von: Stewart, William
Veröffentlicht: (2025)
What Does 'Human-Centred AI' Mean?
von: Guest, Olivia
Veröffentlicht: (2025)
von: Guest, Olivia
Veröffentlicht: (2025)
EduQate: Generating Adaptive Curricula through RMABs in Education Settings
von: Tio, Sidney, et al.
Veröffentlicht: (2024)
von: Tio, Sidney, et al.
Veröffentlicht: (2024)
ARCTraj: A Dataset and Benchmark of Human Reasoning Trajectories for Abstract Problem Solving
von: Kim, Sejin, et al.
Veröffentlicht: (2025)
von: Kim, Sejin, et al.
Veröffentlicht: (2025)
Pareto-Optimized Open-Source LLMs for Healthcare via Context Retrieval
von: Bayarri-Planas, Jordi, et al.
Veröffentlicht: (2024)
von: Bayarri-Planas, Jordi, et al.
Veröffentlicht: (2024)
Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents
von: Minai, Ali A.
Veröffentlicht: (2025)
von: Minai, Ali A.
Veröffentlicht: (2025)
Beyond Mimicry: Preference Coherence in LLMs
von: Mikaelson, Luhan, et al.
Veröffentlicht: (2025)
von: Mikaelson, Luhan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Intervention Complexity as a Canonical Reward and a Measure of Intelligence
von: McCane, Brendan
Veröffentlicht: (2026) -
How much can change in a year? Revisiting Evaluation in Multi-Agent Reinforcement Learning
von: Singh, Siddarth, et al.
Veröffentlicht: (2023) -
Uncertainty-Aware Reward Discounting for Mitigating Reward Hacking
von: Singha, Disha
Veröffentlicht: (2026) -
Developing trustworthy AI applications with foundation models
von: Mock, Michael, et al.
Veröffentlicht: (2024) -
Rewarding Creativity: A Human-Aligned Generative Reward Model for Reinforcement Learning in Storytelling
von: Li, Zhaoyan, et al.
Veröffentlicht: (2026)