Saved in:
| Main Authors: | Shi, Jerick, Zhang, Terry Jingcheng, Jin, Zhijing, Conitzer, Vincent |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.04782 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Hallucination to Scheming: A Unified Taxonomy and Benchmark Analysis for LLM Deception
by: Shi, Jerick, et al.
Published: (2026)
by: Shi, Jerick, et al.
Published: (2026)
Cheap Talk
by: St. Pierre, Joshua
Published: (2025)
by: St. Pierre, Joshua
Published: (2025)
CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas
by: Tewolde, Emanuel, et al.
Published: (2026)
by: Tewolde, Emanuel, et al.
Published: (2026)
Predictive Power of LLMs in Financial Markets
by: Shi, Jerick, et al.
Published: (2024)
by: Shi, Jerick, et al.
Published: (2024)
Now, Later, and Lasting: Ten Priorities for AI Research, Policy, and Practice
by: Horvitz, Eric, et al.
Published: (2024)
by: Horvitz, Eric, et al.
Published: (2024)
Causality for Natural Language Processing
by: Jin, Zhijing
Published: (2025)
by: Jin, Zhijing
Published: (2025)
On the Pros and Cons of Active Learning for Moral Preference Elicitation
by: Keswani, Vijay, et al.
Published: (2024)
by: Keswani, Vijay, et al.
Published: (2024)
Market-Dependent Communication in Multi-Agent Alpha Generation
by: Shi, Jerick, et al.
Published: (2025)
by: Shi, Jerick, et al.
Published: (2025)
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
by: Cobben, Pepijn, et al.
Published: (2026)
by: Cobben, Pepijn, et al.
Published: (2026)
Assessing Large Language Models' ability to predict how humans balance self-interest and the interest of others
by: Capraro, Valerio, et al.
Published: (2023)
by: Capraro, Valerio, et al.
Published: (2023)
Evaluating the Promise and Pitfalls of LLMs in Hiring Decisions
by: Anzenberg, Eitan, et al.
Published: (2025)
by: Anzenberg, Eitan, et al.
Published: (2025)
The Promises and Perils of using LLMs for Effective Public Services
by: Moon, Erina Seh-Young, et al.
Published: (2026)
by: Moon, Erina Seh-Young, et al.
Published: (2026)
Algorithmic Cheap Talk
by: Babichenko, Yakov, et al.
Published: (2023)
by: Babichenko, Yakov, et al.
Published: (2023)
Can AI Model the Complexities of Human Moral Decision-Making? A Qualitative Study of Kidney Allocation Decisions
by: Keswani, Vijay, et al.
Published: (2025)
by: Keswani, Vijay, et al.
Published: (2025)
Cheap Expertise: Mapping and Challenging Industry Perspectives in the Expert Data Gig Economy
by: Wolfe, Robert, et al.
Published: (2026)
by: Wolfe, Robert, et al.
Published: (2026)
The promise and perils of AI in medicine
by: Sparrow, Robert, et al.
Published: (2025)
by: Sparrow, Robert, et al.
Published: (2025)
Cheap Talk in Bilateral Trade
by: Tucker-Foltz, Jamie, et al.
Published: (2026)
by: Tucker-Foltz, Jamie, et al.
Published: (2026)
Equitable Access to Justice: Logical LLMs Show Promise
by: Kant, Manuj, et al.
Published: (2024)
by: Kant, Manuj, et al.
Published: (2024)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
by: Choi, Younwoo, et al.
Published: (2025)
by: Choi, Younwoo, et al.
Published: (2025)
On The Stability of Moral Preferences: A Problem with Computational Elicitation Methods
by: Boerstler, Kyle, et al.
Published: (2024)
by: Boerstler, Kyle, et al.
Published: (2024)
Moral Change or Noise? On Problems of Aligning AI With Temporally Unstable Human Feedback
by: Keswani, Vijay, et al.
Published: (2025)
by: Keswani, Vijay, et al.
Published: (2025)
The Complexity of Computing Robust Mediated Equilibria in Ordinal Games
by: Conitzer, Vincent
Published: (2024)
by: Conitzer, Vincent
Published: (2024)
Exploring Consciousness in LLMs: A Systematic Survey of Theories, Implementations, and Frontier Risks
by: Chen, Sirui, et al.
Published: (2025)
by: Chen, Sirui, et al.
Published: (2025)
Gender inequality and self-publication patterns among scientific editors
by: Liu, Fengyuan, et al.
Published: (2022)
by: Liu, Fengyuan, et al.
Published: (2022)
Frontier AI systems have surpassed the self-replicating red line
by: Pan, Xudong, et al.
Published: (2024)
by: Pan, Xudong, et al.
Published: (2024)
Are LLMs Court-Ready? Evaluating Frontier Models on Indian Legal Reasoning
by: Juvekar, Kush, et al.
Published: (2025)
by: Juvekar, Kush, et al.
Published: (2025)
Can LLMs Talk 'Sex'? Exploring How AI Models Handle Intimate Conversations
by: Lai, Huiqian
Published: (2025)
by: Lai, Huiqian
Published: (2025)
An evidence-based and critical analysis of the Fediverse decentralization promises
by: Xavier, Henrique S.
Published: (2024)
by: Xavier, Henrique S.
Published: (2024)
WHBench: Evaluating Frontier LLMs with Expert-in-the-Loop Validation on Women's Health Topics
by: Maurya, Sneha, et al.
Published: (2026)
by: Maurya, Sneha, et al.
Published: (2026)
Talking the Talk Does Not Entail Walking the Walk: On the Limits of Large Language Models in Lexical Entailment Recognition
by: Greco, Candida M., et al.
Published: (2024)
by: Greco, Candida M., et al.
Published: (2024)
When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas
by: Backmann, Steffen, et al.
Published: (2025)
by: Backmann, Steffen, et al.
Published: (2025)
Estimating household contact matrices structure from easily collectable metadata
by: Dall'Amico, Lorenzo, et al.
Published: (2022)
by: Dall'Amico, Lorenzo, et al.
Published: (2022)
Breaking the ICE: Exploring promises and challenges of benchmarks for Inference Carbon & Energy estimation for LLMs
by: Sikand, Samarth, et al.
Published: (2025)
by: Sikand, Samarth, et al.
Published: (2025)
Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
by: Jenny, David F., et al.
Published: (2023)
by: Jenny, David F., et al.
Published: (2023)
A Randomized Controlled Trial on Anonymizing Reviewers to Each Other in Peer Review Discussions
by: Rastogi, Charvi, et al.
Published: (2024)
by: Rastogi, Charvi, et al.
Published: (2024)
Frontier AI's Impact on the Cybersecurity Landscape
by: Potter, Yujin, et al.
Published: (2025)
by: Potter, Yujin, et al.
Published: (2025)
Preserving Historical Truth: Detecting Historical Revisionism in Large Language Models
by: Ortu, Francesco, et al.
Published: (2026)
by: Ortu, Francesco, et al.
Published: (2026)
Evaluating the Critical Risks of Amazon's Nova Premier under the Frontier Model Safety Framework
by: Krishna, Satyapriya, et al.
Published: (2025)
by: Krishna, Satyapriya, et al.
Published: (2025)
The California Report on Frontier AI Policy
by: Bommasani, Rishi, et al.
Published: (2025)
by: Bommasani, Rishi, et al.
Published: (2025)
Tie-breaking in self interest cumulative subtraction games
by: Bhagat, Anjali, et al.
Published: (2025)
by: Bhagat, Anjali, et al.
Published: (2025)
Similar Items
-
From Hallucination to Scheming: A Unified Taxonomy and Benchmark Analysis for LLM Deception
by: Shi, Jerick, et al.
Published: (2026) -
Cheap Talk
by: St. Pierre, Joshua
Published: (2025) -
CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas
by: Tewolde, Emanuel, et al.
Published: (2026) -
Predictive Power of LLMs in Financial Markets
by: Shi, Jerick, et al.
Published: (2024) -
Now, Later, and Lasting: Ten Priorities for AI Research, Policy, and Practice
by: Horvitz, Eric, et al.
Published: (2024)