Appraisal-Guided Proximal Policy Optimization: Modeling Psychological Disorders in Dynamic Grid World
Fuente:
arXiv
Saved in:
| Main Authors: | Prasad, Hari, Jacob, Chinnu, P, Imthias Ahamed T. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Interpretability-Guided Bi-objective Optimization: Aligning Accuracy and Explainability
by: Fouladi, Kasra, et al.
Published: (2026)
by: Fouladi, Kasra, et al.
Published: (2026)
Pareto-Optimized Open-Source LLMs for Healthcare via Context Retrieval
by: Bayarri-Planas, Jordi, et al.
Published: (2024)
by: Bayarri-Planas, Jordi, et al.
Published: (2024)
The Station: An Open-World Environment for AI-Driven Discovery
by: Chung, Stephen, et al.
Published: (2025)
by: Chung, Stephen, et al.
Published: (2025)
Mutagenesis screen to map the functions of parameters of Large Language Models
by: Hu, Yue, et al.
Published: (2024)
by: Hu, Yue, et al.
Published: (2024)
A Case-Based Persistent Memory for a Large Language Model
by: Watson, Ian
Published: (2023)
by: Watson, Ian
Published: (2023)
Do Chains-of-Thoughts of Large Language Models Suffer from Hallucinations, Cognitive Biases, or Phobias in Bayesian Reasoning?
by: Araya, Roberto
Published: (2025)
by: Araya, Roberto
Published: (2025)
A Taxonomy of Omnicidal Futures Involving Artificial Intelligence
by: Critch, Andrew, et al.
Published: (2025)
by: Critch, Andrew, et al.
Published: (2025)
Enhancing Multi-Agent Collaboration with Attention-Based Actor-Critic Policies
by: Belinchon, Hugo Garrido-Lestache, et al.
Published: (2025)
by: Belinchon, Hugo Garrido-Lestache, et al.
Published: (2025)
Machine Learning and Theory Ladenness -- A Phenomenological Account
by: Termine, Alberto, et al.
Published: (2024)
by: Termine, Alberto, et al.
Published: (2024)
Complete Implementation of WXF Chinese Chess Rules
by: Tan, Daniel, et al.
Published: (2024)
by: Tan, Daniel, et al.
Published: (2024)
Developing trustworthy AI applications with foundation models
by: Mock, Michael, et al.
Published: (2024)
by: Mock, Michael, et al.
Published: (2024)
How VADER is your AI? Towards a definition of artificial intelligence systems appropriate for regulation
by: Bezerra, Leonardo C. T., et al.
Published: (2024)
by: Bezerra, Leonardo C. T., et al.
Published: (2024)
Interpolative Decoding: Exploring the Spectrum of Personality Traits in LLMs
by: Yeh, Eric, et al.
Published: (2025)
by: Yeh, Eric, et al.
Published: (2025)
Self-evolving expertise in complex non-verifiable subject domains: dialogue as implicit meta-RL
by: Bailey, Richard M.
Published: (2025)
by: Bailey, Richard M.
Published: (2025)
Reward is not enough: can we liberate AI from the reinforcement learning paradigm?
by: Glukhov, Vacslav
Published: (2022)
by: Glukhov, Vacslav
Published: (2022)
Discerning What Matters: A Multi-Dimensional Assessment of Moral Competence in LLMs
by: Kilov, Daniel, et al.
Published: (2025)
by: Kilov, Daniel, et al.
Published: (2025)
From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments
by: Luo, Lijing, et al.
Published: (2026)
by: Luo, Lijing, et al.
Published: (2026)
HCAST: Human-Calibrated Autonomy Software Tasks
by: Rein, David, et al.
Published: (2025)
by: Rein, David, et al.
Published: (2025)
Sensemaking in Novel Environments: How Human Cognition Can Inform Artificial Agents
by: Patterson, Robert E., et al.
Published: (2025)
by: Patterson, Robert E., et al.
Published: (2025)
How much can change in a year? Revisiting Evaluation in Multi-Agent Reinforcement Learning
by: Singh, Siddarth, et al.
Published: (2023)
by: Singh, Siddarth, et al.
Published: (2023)
Efficiently Quantifying Individual Agent Importance in Cooperative MARL
by: Mahjoub, Omayma, et al.
Published: (2023)
by: Mahjoub, Omayma, et al.
Published: (2023)
EduQate: Generating Adaptive Curricula through RMABs in Education Settings
by: Tio, Sidney, et al.
Published: (2024)
by: Tio, Sidney, et al.
Published: (2024)
Intervention Complexity as a Canonical Reward and a Measure of Intelligence
by: McCane, Brendan
Published: (2026)
by: McCane, Brendan
Published: (2026)
Benchmarking AI for low-resource contexts: Thinking beyond leaderboards
by: Pant, Aakash, et al.
Published: (2026)
by: Pant, Aakash, et al.
Published: (2026)
Right-to-Act: A Pre-Execution Non-Compensatory Decision Protocol for AI Systems
by: Lavi, Gadi
Published: (2026)
by: Lavi, Gadi
Published: (2026)
ARCTraj: A Dataset and Benchmark of Human Reasoning Trajectories for Abstract Problem Solving
by: Kim, Sejin, et al.
Published: (2025)
by: Kim, Sejin, et al.
Published: (2025)
Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents
by: Minai, Ali A.
Published: (2025)
by: Minai, Ali A.
Published: (2025)
Beyond Mimicry: Preference Coherence in LLMs
by: Mikaelson, Luhan, et al.
Published: (2025)
by: Mikaelson, Luhan, et al.
Published: (2025)
Dynamic Observation Policies in Observation Cost-Sensitive Reinforcement Learning
by: Bellinger, Colin, et al.
Published: (2023)
by: Bellinger, Colin, et al.
Published: (2023)
How Data Quality Affects Machine Learning Models for Credit Risk Assessment
by: Maurino, Andrea
Published: (2025)
by: Maurino, Andrea
Published: (2025)
Learning Actionable World Models for Industrial Process Control
by: Yan, Peng, et al.
Published: (2025)
by: Yan, Peng, et al.
Published: (2025)
Fanar: An Arabic-Centric Multimodal Generative AI Platform
by: Fanar Team, et al.
Published: (2025)
by: Fanar Team, et al.
Published: (2025)
ConciseRL: Conciseness-Guided Reinforcement Learning for Efficient Reasoning Models
by: Dumitru, Razvan-Gabriel, et al.
Published: (2025)
by: Dumitru, Razvan-Gabriel, et al.
Published: (2025)
Decentralizing Coordination in Open Vehicle Fleets for Scalable and Dynamic Task Allocation
by: Lujak, Marin, et al.
Published: (2024)
by: Lujak, Marin, et al.
Published: (2024)
Not All Explanations are Created Equal: Investigating the Pitfalls of Current XAI Evaluation
by: Shymanski, Joe, et al.
Published: (2025)
by: Shymanski, Joe, et al.
Published: (2025)
Modeling Emotions and Ethics with Large Language Models
by: Chang, Edward Y.
Published: (2024)
by: Chang, Edward Y.
Published: (2024)
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
by: Venkatasubramanian, Venkat, et al.
Published: (2024)
by: Venkatasubramanian, Venkat, et al.
Published: (2024)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
by: Khanna, Danush, et al.
Published: (2025)
by: Khanna, Danush, et al.
Published: (2025)
Aristotle's Original Idea: For and Against Logic in the era of AI
by: Kakas, Antonis C.
Published: (2025)
by: Kakas, Antonis C.
Published: (2025)
Reducing Selection Bias in Large Language Models
by: Eicher, J. E., et al.
Published: (2024)
by: Eicher, J. E., et al.
Published: (2024)
Similar Items
-
Interpretability-Guided Bi-objective Optimization: Aligning Accuracy and Explainability
by: Fouladi, Kasra, et al.
Published: (2026) -
Pareto-Optimized Open-Source LLMs for Healthcare via Context Retrieval
by: Bayarri-Planas, Jordi, et al.
Published: (2024) -
The Station: An Open-World Environment for AI-Driven Discovery
by: Chung, Stephen, et al.
Published: (2025) -
Mutagenesis screen to map the functions of parameters of Large Language Models
by: Hu, Yue, et al.
Published: (2024) -
A Case-Based Persistent Memory for a Large Language Model
by: Watson, Ian
Published: (2023)