Objective Decoupling in Social Reinforcement Learning: Recovering Ground Truth from Sycophantic Majorities
Fuente:
arXiv
Saved in:
| Main Authors: | Ghasemi, Majid, Crowley, Mark |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Graph-Enhanced Deep Reinforcement Learning for Multi-Objective Unrelated Parallel Machine Scheduling
by: Soykan, Bulent, et al.
Published: (2026)
by: Soykan, Bulent, et al.
Published: (2026)
Rethinking Agentic Reinforcement Learning In Large Language Models
by: Cui, Fangming, et al.
Published: (2026)
by: Cui, Fangming, et al.
Published: (2026)
Hybrid Action Based Reinforcement Learning for Multi-Objective Compatible Autonomous Driving
by: Jin, Guizhe, et al.
Published: (2025)
by: Jin, Guizhe, et al.
Published: (2025)
Quantum Reinforcement Learning with Transformers for the Capacitated Vehicle Routing Problem
by: Andrés, Eva
Published: (2026)
by: Andrés, Eva
Published: (2026)
TruthTensor: Evaluating LLMs through Human Imitation on Prediction Market under Drift and Holistic Reasoning
by: Shahabi, Shirin, et al.
Published: (2026)
by: Shahabi, Shirin, et al.
Published: (2026)
Auditing an Automatic Grading Model with deep Reinforcement Learning
by: Condor, Aubrey, et al.
Published: (2024)
by: Condor, Aubrey, et al.
Published: (2024)
Neurosymbolic Graph Enrichment for Grounded World Models
by: De Giorgis, Stefano, et al.
Published: (2024)
by: De Giorgis, Stefano, et al.
Published: (2024)
Reinforcement Learning for an Efficient and Effective Malware Investigation during Cyber Incident Response
by: Dunsin, Dipo, et al.
Published: (2024)
by: Dunsin, Dipo, et al.
Published: (2024)
Symbolic-AI-Fusion Deep Learning (SAIF-DL): Encoding Knowledge into Training with Answer Set Programming Loss Penalties by a Novel Loss Function Approach
by: Machot, Fadi Al, et al.
Published: (2024)
by: Machot, Fadi Al, et al.
Published: (2024)
HGT-Scheduler: Deep Reinforcement Learning for the Job Shop Scheduling Problem via Heterogeneous Graph Transformers
by: Soykan, Bulent
Published: (2026)
by: Soykan, Bulent
Published: (2026)
Coordinating Ride-Pooling with Public Transit using Reward-Guided Conservative Q-Learning: An Offline Training and Online Fine-Tuning Reinforcement Learning Framework
by: Hu, Yulong, et al.
Published: (2025)
by: Hu, Yulong, et al.
Published: (2025)
One Step is Enough: Multi-Agent Reinforcement Learning based on One-Step Policy Optimization for Order Dispatch on Ride-Sharing Platforms
by: Zhao, Zijian, et al.
Published: (2025)
by: Zhao, Zijian, et al.
Published: (2025)
Relational Norms for Human-AI Cooperation
by: Earp, Brian D., et al.
Published: (2025)
by: Earp, Brian D., et al.
Published: (2025)
Toward Virtuous Reinforcement Learning: A Critique and Roadmap
by: Ghasemi, Majid, et al.
Published: (2025)
by: Ghasemi, Majid, et al.
Published: (2025)
MAPS: Multi-Fidelity AI-Augmented Photonic Simulation and Inverse Design Infrastructure
by: Ma, Pingchuan, et al.
Published: (2025)
by: Ma, Pingchuan, et al.
Published: (2025)
MM-tau-p$^2$: Persona-Adaptive Prompting for Robust Multi-Modal Agent Evaluation in Dual-Control Settings
by: Purwar, Anupam, et al.
Published: (2026)
by: Purwar, Anupam, et al.
Published: (2026)
Mind the Gap: How Elicitation Protocols Shape the Stated-Revealed Preference Gap in Language Models
by: Mahajan, Pranav, et al.
Published: (2026)
by: Mahajan, Pranav, et al.
Published: (2026)
Traffic and weather driven hybrid digital twin for bridge monitoring
by: Balijepalli, Phani Raja Bharath, et al.
Published: (2026)
by: Balijepalli, Phani Raja Bharath, et al.
Published: (2026)
mcp-proto-okn: Natural-language access to open scientific knowledge graphs through the Model Context Protocol
by: Rose, Peter W., et al.
Published: (2026)
by: Rose, Peter W., et al.
Published: (2026)
VERA-MH: Validation of Ethical and Responsible AI in Mental Health
by: Belli, Luca, et al.
Published: (2026)
by: Belli, Luca, et al.
Published: (2026)
An Onto-Relational-Sophic Framework for Governing Synthetic Minds
by: Ning, Huansheng, et al.
Published: (2026)
by: Ning, Huansheng, et al.
Published: (2026)
UrbanMoE: A Sparse Multi-Modal Mixture-of-Experts Framework for Multi-Task Urban Region Profiling
by: Liu, Pingping, et al.
Published: (2026)
by: Liu, Pingping, et al.
Published: (2026)
Developing AI Agents with Simulated Data: Why, what, and how?
by: Liu, Xiaoran, et al.
Published: (2026)
by: Liu, Xiaoran, et al.
Published: (2026)
Enhancing Software Quality Assurance with an Adaptive Differential Evolution based Quantum Variational Autoencoder-Transformer Model
by: Barma, Seshu Babu, et al.
Published: (2025)
by: Barma, Seshu Babu, et al.
Published: (2025)
Cognitive maps are generative programs
by: Kryven, Marta, et al.
Published: (2025)
by: Kryven, Marta, et al.
Published: (2025)
Optimizing Package Delivery with Quantum Annealers: Addressing Time-Windows and Simultaneous Pickup and Delivery
by: Osaba, Eneko, et al.
Published: (2025)
by: Osaba, Eneko, et al.
Published: (2025)
Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization
by: Giabbanelli, Philippe J., et al.
Published: (2025)
by: Giabbanelli, Philippe J., et al.
Published: (2025)
An Approach to Checking Correctness for Agentic Systems
by: Sheffler, Thomas J
Published: (2025)
by: Sheffler, Thomas J
Published: (2025)
OpenMENA: An Open-Source Memristor Interfacing and Compute Board for Neuromorphic Edge-AI Applications
by: Safa, Ali, et al.
Published: (2025)
by: Safa, Ali, et al.
Published: (2025)
LSDTs: LLM-Augmented Semantic Digital Twins for Adaptive Knowledge-Intensive Infrastructure Planning
by: Li, Naiyi, et al.
Published: (2025)
by: Li, Naiyi, et al.
Published: (2025)
Harnessing AI Agents to Advance Research on Refugee Child Mental Health
by: Shrivastava, Aditya, et al.
Published: (2025)
by: Shrivastava, Aditya, et al.
Published: (2025)
Artificial Intelligence for Atmospheric Sciences: A Research Roadmap
by: Zaidan, Martha Arbayani, et al.
Published: (2025)
by: Zaidan, Martha Arbayani, et al.
Published: (2025)
Dendritic Computing with Multi-Gate Ferroelectric Field-Effect Transistors
by: Islam, A N M Nafiul, et al.
Published: (2025)
by: Islam, A N M Nafiul, et al.
Published: (2025)
A Distributed Emulation Environment for In-Memory Computing Systems
by: Bougioukou, Eleni, et al.
Published: (2025)
by: Bougioukou, Eleni, et al.
Published: (2025)
Towards Intelligent Transportation with Pedestrians and Vehicles In-the-Loop: A Surveillance Video-Assisted Federated Digital Twin Framework
by: Li, Xiaolong, et al.
Published: (2025)
by: Li, Xiaolong, et al.
Published: (2025)
From Questions to Insightful Answers: Building an Informed Chatbot for University Resources
by: Neupane, Subash, et al.
Published: (2024)
by: Neupane, Subash, et al.
Published: (2024)
Virtuous Machines: Towards Artificial General Science
by: Wehr, Gabrielle, et al.
Published: (2025)
by: Wehr, Gabrielle, et al.
Published: (2025)
Trustworthy Orchestration Artificial Intelligence by the Ten Criteria with Control-Plane Governance
by: Kang, Byeong Ho, et al.
Published: (2025)
by: Kang, Byeong Ho, et al.
Published: (2025)
The Maximum Coverage Model and Recommendation System for UAV Vertiports Location Planning
by: Hua, Chunliang, et al.
Published: (2025)
by: Hua, Chunliang, et al.
Published: (2025)
Building Trustworthy AI: Transparent AI Systems via Large Language Models, Ontologies, and Logical Reasoning (TranspNet)
by: Machot, Fadi Al, et al.
Published: (2024)
by: Machot, Fadi Al, et al.
Published: (2024)
Similar Items
-
Graph-Enhanced Deep Reinforcement Learning for Multi-Objective Unrelated Parallel Machine Scheduling
by: Soykan, Bulent, et al.
Published: (2026) -
Rethinking Agentic Reinforcement Learning In Large Language Models
by: Cui, Fangming, et al.
Published: (2026) -
Hybrid Action Based Reinforcement Learning for Multi-Objective Compatible Autonomous Driving
by: Jin, Guizhe, et al.
Published: (2025) -
Quantum Reinforcement Learning with Transformers for the Capacitated Vehicle Routing Problem
by: Andrés, Eva
Published: (2026) -
TruthTensor: Evaluating LLMs through Human Imitation on Prediction Market under Drift and Holistic Reasoning
by: Shahabi, Shirin, et al.
Published: (2026)