Semantic Reward Collapse and the Preservation of Epistemic Integrity in Adaptive AI Systems
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Parris, William |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ProfileXAI: User-Adaptive Explainable AI
von: Corrales, Gilber A., et al.
Veröffentlicht: (2025)
von: Corrales, Gilber A., et al.
Veröffentlicht: (2025)
SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
From Features to Graphs: Exploring Graph Structures and Pairwise Interactions via GNNs
von: Yamchote, Phaphontee, et al.
Veröffentlicht: (2025)
von: Yamchote, Phaphontee, et al.
Veröffentlicht: (2025)
Mapping representations in Reinforcement Learning via Semantic Alignment for Zero-Shot Stitching
von: Ricciardi, Antonio Pio, et al.
Veröffentlicht: (2025)
von: Ricciardi, Antonio Pio, et al.
Veröffentlicht: (2025)
HyperMask: Adaptive Hypernetwork-based Masks for Continual Learning
von: Książek, Kamil, et al.
Veröffentlicht: (2023)
von: Książek, Kamil, et al.
Veröffentlicht: (2023)
NeuronSpark: A Spiking Neural Network Language Model with Selective State Space Dynamics
von: Tang, Zhengzheng
Veröffentlicht: (2026)
von: Tang, Zhengzheng
Veröffentlicht: (2026)
Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design
von: Pepe, Alberto, et al.
Veröffentlicht: (2026)
von: Pepe, Alberto, et al.
Veröffentlicht: (2026)
On the Origin of Algorithmic Progress in AI
von: Gundlach, Hans, et al.
Veröffentlicht: (2025)
von: Gundlach, Hans, et al.
Veröffentlicht: (2025)
Harnessing non-adversarial robustness in large language models
von: Zhou, Qinghua, et al.
Veröffentlicht: (2026)
von: Zhou, Qinghua, et al.
Veröffentlicht: (2026)
Lost or Hidden? A Concept-Level Forgetting in Supervised Continual Learning
von: Filus, Katarzyna, et al.
Veröffentlicht: (2026)
von: Filus, Katarzyna, et al.
Veröffentlicht: (2026)
A Machine Learning Framework for Turbofan Health Estimation via Inverse Problem Formulation
von: Leyli-Abadi, Milad, et al.
Veröffentlicht: (2026)
von: Leyli-Abadi, Milad, et al.
Veröffentlicht: (2026)
Simple Network Graph Comparative Learning
von: Yu, Qiang, et al.
Veröffentlicht: (2026)
von: Yu, Qiang, et al.
Veröffentlicht: (2026)
Reliable Unlearning Harmful Information in LLMs with Metamorphosis Representation Projection
von: Wu, Chengcan, et al.
Veröffentlicht: (2025)
von: Wu, Chengcan, et al.
Veröffentlicht: (2025)
HGTUL: A Hypergraph-based Model For Trajectory User Linking
von: Chang, Fengjie, et al.
Veröffentlicht: (2025)
von: Chang, Fengjie, et al.
Veröffentlicht: (2025)
Is ReLU Adversarially Robust?
von: Sooksatra, Korn, et al.
Veröffentlicht: (2024)
von: Sooksatra, Korn, et al.
Veröffentlicht: (2024)
Upside Down Reinforcement Learning with Policy Generators
von: Di Ventura, Jacopo, et al.
Veröffentlicht: (2025)
von: Di Ventura, Jacopo, et al.
Veröffentlicht: (2025)
Multi-Level Fusion Graph Neural Network for Molecule Property Prediction
von: Liu, XiaYu, et al.
Veröffentlicht: (2025)
von: Liu, XiaYu, et al.
Veröffentlicht: (2025)
CGLearn: Consistent Gradient-Based Learning for Out-of-Distribution Generalization
von: Chowdhury, Jawad, et al.
Veröffentlicht: (2024)
von: Chowdhury, Jawad, et al.
Veröffentlicht: (2024)
Concept Prerequisite Relation Prediction by Using Permutation-Equivariant Directed Graph Neural Networks
von: Qu, Xiran, et al.
Veröffentlicht: (2023)
von: Qu, Xiran, et al.
Veröffentlicht: (2023)
ProactBench: Beyond What The User Asked For
von: Harfi, Sepehr, et al.
Veröffentlicht: (2026)
von: Harfi, Sepehr, et al.
Veröffentlicht: (2026)
Thinking Machines: Mathematical Reasoning in the Age of LLMs
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
Neural Concept Verifier: Scaling Prover-Verifier Games via Concept Encodings
von: Turan, Berkant, et al.
Veröffentlicht: (2025)
von: Turan, Berkant, et al.
Veröffentlicht: (2025)
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
AIPsy-Affect: A Keyword-Free Clinical Stimulus Battery for Mechanistic Interpretability of Emotion in Language Models
von: Keeman, Michael
Veröffentlicht: (2026)
von: Keeman, Michael
Veröffentlicht: (2026)
The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth
von: Henry, James
Veröffentlicht: (2026)
von: Henry, James
Veröffentlicht: (2026)
A Practical Guide to Streaming Continual Learning
von: Cossu, Andrea, et al.
Veröffentlicht: (2026)
von: Cossu, Andrea, et al.
Veröffentlicht: (2026)
On Privacy Leakage in Tabular Diffusion Models: Influential Factors, Attacker Knowledge, and Metrics
von: Shafieinejad, Masoumeh, et al.
Veröffentlicht: (2026)
von: Shafieinejad, Masoumeh, et al.
Veröffentlicht: (2026)
Don't Look Back in Anger: MAGIC Net for Streaming Continual Learning with Temporal Dependence
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference
von: Mathew, Aby Mammen
Veröffentlicht: (2026)
von: Mathew, Aby Mammen
Veröffentlicht: (2026)
cPNN: Continuous Progressive Neural Networks for Evolving Streaming Time Series
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
Geometric Evolution Maps: Extracting Stable Concept Probes from Transformer Residual Streams
von: Henry, James
Veröffentlicht: (2026)
von: Henry, James
Veröffentlicht: (2026)
BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025)
Decentralized Time Series Classification with ROCKET Features
von: Casella, Bruno, et al.
Veröffentlicht: (2025)
von: Casella, Bruno, et al.
Veröffentlicht: (2025)
Matryoshka Policy Gradient for Entropy-Regularized RL: Convergence and Global Optimality
von: Ged, François, et al.
Veröffentlicht: (2023)
von: Ged, François, et al.
Veröffentlicht: (2023)
Combining AI Control Systems and Human Decision Support via Robustness and Criticality
von: Woods, Walt, et al.
Veröffentlicht: (2024)
von: Woods, Walt, et al.
Veröffentlicht: (2024)
Streaming Continual Learning for Unified Adaptive Intelligence in Dynamic Environments
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
Batch Matrix-form Equations and Implementation of Multilayer Perceptrons
von: Wesselink, Wieger, et al.
Veröffentlicht: (2025)
von: Wesselink, Wieger, et al.
Veröffentlicht: (2025)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
von: Yang, Yibo
Veröffentlicht: (2025)
von: Yang, Yibo
Veröffentlicht: (2025)
A Constraint-Preserving Neural Network Approach for Solving Mean-Field Games Equilibrium
von: Liu, Jinwei, et al.
Veröffentlicht: (2025)
von: Liu, Jinwei, et al.
Veröffentlicht: (2025)
How to Correctly do Semantic Backpropagation on Language-based Agentic Systems
von: Wang, Wenyi, et al.
Veröffentlicht: (2024)
von: Wang, Wenyi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ProfileXAI: User-Adaptive Explainable AI
von: Corrales, Gilber A., et al.
Veröffentlicht: (2025) -
SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
von: Zhou, Xueyang, et al.
Veröffentlicht: (2025) -
From Features to Graphs: Exploring Graph Structures and Pairwise Interactions via GNNs
von: Yamchote, Phaphontee, et al.
Veröffentlicht: (2025) -
Mapping representations in Reinforcement Learning via Semantic Alignment for Zero-Shot Stitching
von: Ricciardi, Antonio Pio, et al.
Veröffentlicht: (2025) -
HyperMask: Adaptive Hypernetwork-based Masks for Continual Learning
von: Książek, Kamil, et al.
Veröffentlicht: (2023)