Intervention Complexity as a Canonical Reward and a Measure of Intelligence
Fuente:
arXiv
Guardado en:
| Autor principal: | McCane, Brendan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Uncertainty-Aware Reward Discounting for Mitigating Reward Hacking
por: Singha, Disha
Publicado: (2026)
por: Singha, Disha
Publicado: (2026)
ARCTraj: A Dataset and Benchmark of Human Reasoning Trajectories for Abstract Problem Solving
por: Kim, Sejin, et al.
Publicado: (2025)
por: Kim, Sejin, et al.
Publicado: (2025)
Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents
por: Minai, Ali A.
Publicado: (2025)
por: Minai, Ali A.
Publicado: (2025)
Beyond Mimicry: Preference Coherence in LLMs
por: Mikaelson, Luhan, et al.
Publicado: (2025)
por: Mikaelson, Luhan, et al.
Publicado: (2025)
Generating Causal Explanations of Vehicular Agent Behavioural Interactions with Learnt Reward Profiles
por: Howard, Rhys, et al.
Publicado: (2025)
por: Howard, Rhys, et al.
Publicado: (2025)
An Axiomatic Approach to General Intelligence: SANC(E3) -- Self-organizing Active Network of Concepts with Energy E3
por: Kwon, Daesuk, et al.
Publicado: (2026)
por: Kwon, Daesuk, et al.
Publicado: (2026)
Augmenting deep neural networks with symbolic knowledge: Towards trustworthy and interpretable AI for education
por: Hooshyar, Danial, et al.
Publicado: (2023)
por: Hooshyar, Danial, et al.
Publicado: (2023)
Graceful task adaptation with a bi-hemispheric RL agent
por: Nicholas, Grant, et al.
Publicado: (2024)
por: Nicholas, Grant, et al.
Publicado: (2024)
Interpretability-Guided Bi-objective Optimization: Aligning Accuracy and Explainability
por: Fouladi, Kasra, et al.
Publicado: (2026)
por: Fouladi, Kasra, et al.
Publicado: (2026)
Revisiting Parameter-Based Knowledge Editing in Large Language Models: Theoretical Limits and Empirical Evidence
por: Ren, Wanying, et al.
Publicado: (2026)
por: Ren, Wanying, et al.
Publicado: (2026)
The Parameters of Educability
por: Valiant, Leslie G.
Publicado: (2024)
por: Valiant, Leslie G.
Publicado: (2024)
Training Language Models to Win Debates with Self-Play Improves Judge Accuracy
por: Arnesen, Samuel, et al.
Publicado: (2024)
por: Arnesen, Samuel, et al.
Publicado: (2024)
Reasoning Beyond the Obvious: Evaluating Divergent and Convergent Thinking in LLMs for Financial Scenarios
por: Bok, Zhuang Qiang, et al.
Publicado: (2025)
por: Bok, Zhuang Qiang, et al.
Publicado: (2025)
Simulation-Driven Railway Delay Prediction: An Imitation Learning Approach
por: Elliker, Clément, et al.
Publicado: (2025)
por: Elliker, Clément, et al.
Publicado: (2025)
A Domain-Agnostic Neurosymbolic Approach for Big Social Data Analysis: Evaluating Mental Health Sentiment on Social Media during COVID-19
por: Khandelwal, Vedant, et al.
Publicado: (2024)
por: Khandelwal, Vedant, et al.
Publicado: (2024)
The Information-Theoretic Imperative: Compression and the Epistemic Foundations of Intelligence
por: Dittrich, Christian, et al.
Publicado: (2025)
por: Dittrich, Christian, et al.
Publicado: (2025)
The ARC of Progress towards AGI: A Living Survey of Abstraction and Reasoning
por: Vahdati, Sahar, et al.
Publicado: (2026)
por: Vahdati, Sahar, et al.
Publicado: (2026)
Deep Reinforcement Learning for Adverse Garage Scenario Generation
por: Li, Kai
Publicado: (2024)
por: Li, Kai
Publicado: (2024)
The Drill-Down and Fabricate Test (DDFT): A Protocol for Measuring Epistemic Robustness in Language Models
por: Baxi, Rahul
Publicado: (2025)
por: Baxi, Rahul
Publicado: (2025)
Super-additive Cooperation in Language Model Agents
por: Tonini, Filippo, et al.
Publicado: (2025)
por: Tonini, Filippo, et al.
Publicado: (2025)
Scalable APT Malware Classification via Parallel Feature Extraction and GPU-Accelerated Learning
por: Subedar, Noah, et al.
Publicado: (2025)
por: Subedar, Noah, et al.
Publicado: (2025)
LaPro-DTA: Latent Dual-View Drug Representations and Salient Protein Feature Extraction for Generalizable Drug--Target Affinity Prediction
por: Dun, Zihan, et al.
Publicado: (2026)
por: Dun, Zihan, et al.
Publicado: (2026)
Emergence of Self-Awareness in Artificial Systems: A Minimalist Three-Layer Approach to Artificial Consciousness
por: Iida, Kurando
Publicado: (2025)
por: Iida, Kurando
Publicado: (2025)
Matryoshka Policy Gradient for Entropy-Regularized RL: Convergence and Global Optimality
por: Ged, François, et al.
Publicado: (2023)
por: Ged, François, et al.
Publicado: (2023)
Vector Symbolic Architectures answer Jackendoff's challenges for cognitive neuroscience
por: Gayler, Ross W.
Publicado: (2004)
por: Gayler, Ross W.
Publicado: (2004)
Perturbation Dose Responses in Recursive LLM Loops: Raw Switching, Stochastic Floors, and Persistent Escape under Append, Replace, and Dialog Updates
por: Kaplanski, Pawel
Publicado: (2026)
por: Kaplanski, Pawel
Publicado: (2026)
Incentives or Ontology? A Structural Rebuttal to OpenAI's Hallucination Thesis
por: Ackermann, Richard, et al.
Publicado: (2025)
por: Ackermann, Richard, et al.
Publicado: (2025)
Behavioural vs. Representational Systematicity in End-to-End Models: An Opinionated Survey
por: Vegner, Ivan, et al.
Publicado: (2025)
por: Vegner, Ivan, et al.
Publicado: (2025)
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
por: Sarkar, Nilesh, et al.
Publicado: (2026)
por: Sarkar, Nilesh, et al.
Publicado: (2026)
Feel-Good Thompson Sampling for Contextual Bandits: a Markov Chain Monte Carlo Showdown
por: Anand, Emile, et al.
Publicado: (2025)
por: Anand, Emile, et al.
Publicado: (2025)
Prompt Readiness Levels (PRL): a maturity scale and scoring framework for production grade prompt assets
por: Guinard, Sebastien
Publicado: (2026)
por: Guinard, Sebastien
Publicado: (2026)
Choosing DAG Models Using Markov and Minimal Edge Count in the Absence of Ground Truth
por: Ramsey, Joseph D., et al.
Publicado: (2024)
por: Ramsey, Joseph D., et al.
Publicado: (2024)
Reward Machines for Deep RL in Noisy and Uncertain Environments
por: Li, Andrew C., et al.
Publicado: (2024)
por: Li, Andrew C., et al.
Publicado: (2024)
Learning Can Converge Stably to the Wrong Belief under Latent Reliability
por: Zhang, Zhipeng, et al.
Publicado: (2026)
por: Zhang, Zhipeng, et al.
Publicado: (2026)
Implicit Counterfactual Data Augmentation for Robust Learning
por: Zhou, Xiaoling, et al.
Publicado: (2023)
por: Zhou, Xiaoling, et al.
Publicado: (2023)
STAR : Bridging Statistical and Agentic Reasoning for Large Model Performance Prediction
por: Wang, Xiaoxiao, et al.
Publicado: (2026)
por: Wang, Xiaoxiao, et al.
Publicado: (2026)
FF-INT8: Efficient Forward-Forward DNN Training on Edge Devices with INT8 Precision
por: Ma, Jingxiao, et al.
Publicado: (2025)
por: Ma, Jingxiao, et al.
Publicado: (2025)
Simulation-Based Counterfactual Causal Discovery on Real World Driver Behaviour
por: Howard, Rhys, et al.
Publicado: (2023)
por: Howard, Rhys, et al.
Publicado: (2023)
How Metacognitive Architectures Remember Their Own Thoughts: A Systematic Review
por: Nolte, Robin, et al.
Publicado: (2025)
por: Nolte, Robin, et al.
Publicado: (2025)
MODP: Multi Objective Directional Prompting
por: Nema, Aashutosh, et al.
Publicado: (2025)
por: Nema, Aashutosh, et al.
Publicado: (2025)
Ejemplares similares
-
Uncertainty-Aware Reward Discounting for Mitigating Reward Hacking
por: Singha, Disha
Publicado: (2026) -
ARCTraj: A Dataset and Benchmark of Human Reasoning Trajectories for Abstract Problem Solving
por: Kim, Sejin, et al.
Publicado: (2025) -
Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents
por: Minai, Ali A.
Publicado: (2025) -
Beyond Mimicry: Preference Coherence in LLMs
por: Mikaelson, Luhan, et al.
Publicado: (2025) -
Generating Causal Explanations of Vehicular Agent Behavioural Interactions with Learnt Reward Profiles
por: Howard, Rhys, et al.
Publicado: (2025)