TruthTensor: Evaluating LLMs through Human Imitation on Prediction Market under Drift and Holistic Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Shahabi, Shirin, Graham, Spencer, Isah, Haruna |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MI9: An Integrated Runtime Governance Framework for Agentic AI
por: Wang, Charles L., et al.
Publicado: (2025)
por: Wang, Charles L., et al.
Publicado: (2025)
Autonomous Traffic Signal Optimization Using Digital Twin and Agentic AI for Real-Time Decision-Making
por: Jan, Salman, et al.
Publicado: (2026)
por: Jan, Salman, et al.
Publicado: (2026)
One Step is Enough: Multi-Agent Reinforcement Learning based on One-Step Policy Optimization for Order Dispatch on Ride-Sharing Platforms
por: Zhao, Zijian, et al.
Publicado: (2025)
por: Zhao, Zijian, et al.
Publicado: (2025)
Towards a HIPAA Compliant Agentic AI System in Healthcare
por: Neupane, Subash, et al.
Publicado: (2025)
por: Neupane, Subash, et al.
Publicado: (2025)
A Gossip-Enhanced Communication Substrate for Agentic AI: Toward Decentralized Coordination in Large-Scale Multi-Agent Systems
por: Khan, Nafiul I., et al.
Publicado: (2025)
por: Khan, Nafiul I., et al.
Publicado: (2025)
Modular Autonomous Vehicle in Heterogeneous Traffic Flow: Modeling, Simulation, and Implication
por: Ye, Lanhang, et al.
Publicado: (2024)
por: Ye, Lanhang, et al.
Publicado: (2024)
AGENTSAFE: A Unified Framework for Ethical Assurance and Governance in Agentic AI
por: Khan, Rafflesia, et al.
Publicado: (2025)
por: Khan, Rafflesia, et al.
Publicado: (2025)
Q-RESTORE: Quantum-Driven Framework for Resilient and Equitable Transportation Network Restoration
por: Udekwe, Daniel, et al.
Publicado: (2025)
por: Udekwe, Daniel, et al.
Publicado: (2025)
Content Caching-Assisted Vehicular Edge Computing Using Multi-Agent Graph Attention Reinforcement Learning
por: Shen, Jinjin, et al.
Publicado: (2024)
por: Shen, Jinjin, et al.
Publicado: (2024)
Fusion Intelligence: Confluence of Natural and Artificial Intelligence for Enhanced Problem-Solving Efficiency
por: Kalavakonda, Rohan Reddy, et al.
Publicado: (2024)
por: Kalavakonda, Rohan Reddy, et al.
Publicado: (2024)
KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning
por: Shek, Chak Lam, et al.
Publicado: (2026)
por: Shek, Chak Lam, et al.
Publicado: (2026)
Urban Air Mobility as a System of Systems: An LLM-Enhanced Holonic Approach
por: Sadik, Ahmed R., et al.
Publicado: (2025)
por: Sadik, Ahmed R., et al.
Publicado: (2025)
Scale-Plan: Scalable Language-Enabled Task Planning for Heterogeneous Multi-Robot Teams
por: Gupta, Piyush, et al.
Publicado: (2026)
por: Gupta, Piyush, et al.
Publicado: (2026)
Incorporating Large Language Models into Production Systems for Enhanced Task Automation and Flexibility
por: Xia, Yuchen, et al.
Publicado: (2024)
por: Xia, Yuchen, et al.
Publicado: (2024)
LLM experiments with simulation: Large Language Model Multi-Agent System for Simulation Model Parametrization in Digital Twins
por: Xia, Yuchen, et al.
Publicado: (2024)
por: Xia, Yuchen, et al.
Publicado: (2024)
Beyond the model: Key differentiators in large language models and multi-agent services
por: Goyal, Muskaan, et al.
Publicado: (2025)
por: Goyal, Muskaan, et al.
Publicado: (2025)
BMG-Q: Localized Bipartite Match Graph Attention Q-Learning for Ride-Pooling Order Dispatch
por: Hu, Yulong, et al.
Publicado: (2025)
por: Hu, Yulong, et al.
Publicado: (2025)
On the Ethical Considerations of Generative Agents
por: Diamond, N'yoma, et al.
Publicado: (2024)
por: Diamond, N'yoma, et al.
Publicado: (2024)
SentinelAI: A Multi-Agent Framework for Structuring and Linking NG9-1-1 Emergency Incident Data
por: Ho, Kliment, et al.
Publicado: (2026)
por: Ho, Kliment, et al.
Publicado: (2026)
LCGuard: Latent Communication Guard for Safe KV Sharing in Multi-Agent Systems
por: Asif, Sadia, et al.
Publicado: (2026)
por: Asif, Sadia, et al.
Publicado: (2026)
ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents
por: Lu, Yuxing, et al.
Publicado: (2026)
por: Lu, Yuxing, et al.
Publicado: (2026)
Demystifying AI Agents: The Final Generation of Intelligence
por: McNamara, Kevin J, et al.
Publicado: (2025)
por: McNamara, Kevin J, et al.
Publicado: (2025)
Multi-agent Systems for Misinformation Lifecycle : Detection, Correction And Source Identification
por: Gautam, Aditya
Publicado: (2025)
por: Gautam, Aditya
Publicado: (2025)
Synergistic Simulations: Multi-Agent Problem Solving with Large Language Models
por: Sprigler, Asher, et al.
Publicado: (2024)
por: Sprigler, Asher, et al.
Publicado: (2024)
Scaling Multiagent Systems with Process Rewards
por: Li, Ed, et al.
Publicado: (2026)
por: Li, Ed, et al.
Publicado: (2026)
LLM-Ehnanced Holonic Architecture for Ad-Hoc Scalable SoS
por: Ashfaq, Muhammad, et al.
Publicado: (2025)
por: Ashfaq, Muhammad, et al.
Publicado: (2025)
Agentic Witnessing: Pragmatic and Scalable TEE-Enabled Privacy-Preserving Auditing
por: Rowstron, Antony
Publicado: (2026)
por: Rowstron, Antony
Publicado: (2026)
Training-Free Agentic AI: Probabilistic Control and Coordination in Multi-Agent LLM Systems
por: Hosseini, Mohammad Parsa, et al.
Publicado: (2026)
por: Hosseini, Mohammad Parsa, et al.
Publicado: (2026)
SciDataCopilot: An Agentic Data Preparation Framework for AGI-driven Scientific Discovery
por: Rao, Jiyong, et al.
Publicado: (2026)
por: Rao, Jiyong, et al.
Publicado: (2026)
RAGAR, Your Falsehood Radar: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language Models
por: Khaliq, M. Abdul, et al.
Publicado: (2024)
por: Khaliq, M. Abdul, et al.
Publicado: (2024)
Visual Reasoning and Multi-Agent Approach in Multimodal Large Language Models (MLLMs): Solving TSP and mTSP Combinatorial Challenges
por: Elhenawy, Mohammed, et al.
Publicado: (2024)
por: Elhenawy, Mohammed, et al.
Publicado: (2024)
Interacting Large Language Model Agents. Interpretable Models and Social Learning
por: Jain, Adit, et al.
Publicado: (2024)
por: Jain, Adit, et al.
Publicado: (2024)
The Algorithmic State Architecture (ASA): An Integrated Framework for AI-Enabled Government
por: Engin, Zeynep, et al.
Publicado: (2025)
por: Engin, Zeynep, et al.
Publicado: (2025)
From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents
por: Padhi, Trilok, et al.
Publicado: (2026)
por: Padhi, Trilok, et al.
Publicado: (2026)
Exploring Agentic Artificial Intelligence Systems: Towards a Typological Framework
por: Wissuchek, Christopher, et al.
Publicado: (2025)
por: Wissuchek, Christopher, et al.
Publicado: (2025)
Adaptive Traffic-Following Scheme for Orderly Distributed Control of Multi-Vehicle Systems
por: Jain, Anahita, et al.
Publicado: (2025)
por: Jain, Anahita, et al.
Publicado: (2025)
A Replica for our Democracies? On Using Digital Twins to Enhance Deliberative Democracy
por: Novelli, Claudio, et al.
Publicado: (2025)
por: Novelli, Claudio, et al.
Publicado: (2025)
Quantum Computing and Neuromorphic Computing for Safe, Reliable, and explainable Multi-Agent Reinforcement Learning: Optimal Control in Autonomous Robotics
por: Taghavi, Mazyar, et al.
Publicado: (2024)
por: Taghavi, Mazyar, et al.
Publicado: (2024)
Market share maximizing strategies of CAV fleet operators may cause chaos in our cities
por: Jamróz, Grzegorz, et al.
Publicado: (2025)
por: Jamróz, Grzegorz, et al.
Publicado: (2025)
Mitigating "Epistemic Debt" in Generative AI-Scaffolded Novice Programming using Metacognitive Scripts
por: Sankaranarayanan, Sreecharan
Publicado: (2026)
por: Sankaranarayanan, Sreecharan
Publicado: (2026)
Ejemplares similares
-
MI9: An Integrated Runtime Governance Framework for Agentic AI
por: Wang, Charles L., et al.
Publicado: (2025) -
Autonomous Traffic Signal Optimization Using Digital Twin and Agentic AI for Real-Time Decision-Making
por: Jan, Salman, et al.
Publicado: (2026) -
One Step is Enough: Multi-Agent Reinforcement Learning based on One-Step Policy Optimization for Order Dispatch on Ride-Sharing Platforms
por: Zhao, Zijian, et al.
Publicado: (2025) -
Towards a HIPAA Compliant Agentic AI System in Healthcare
por: Neupane, Subash, et al.
Publicado: (2025) -
A Gossip-Enhanced Communication Substrate for Agentic AI: Toward Decentralized Coordination in Large-Scale Multi-Agent Systems
por: Khan, Nafiul I., et al.
Publicado: (2025)