Emergent LLM behaviors are observationally equivalent to data leakage
Fuente:
arXiv
Salvato in:
| Autori principali: | Barrie, Christopher, Törnberg, Petter |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Reply to "Emergent LLM behaviors are observationally equivalent to data leakage"
di: Ashery, Ariel Flint, et al.
Pubblicazione: (2025)
di: Ashery, Ariel Flint, et al.
Pubblicazione: (2025)
STRIDE: A Tool-Assisted LLM Agent Framework for Strategic and Interactive Decision-Making
di: Li, Chuanhao, et al.
Pubblicazione: (2024)
di: Li, Chuanhao, et al.
Pubblicazione: (2024)
Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas
di: Gallego, Víctor
Pubblicazione: (2026)
di: Gallego, Víctor
Pubblicazione: (2026)
How memory can affect collective and cooperative behaviors in an LLM-Based Social Particle Swarm
di: Hishiki, Taisei, et al.
Pubblicazione: (2026)
di: Hishiki, Taisei, et al.
Pubblicazione: (2026)
Ad Insertion in LLM-Generated Responses
di: Xu, Shengwei, et al.
Pubblicazione: (2026)
di: Xu, Shengwei, et al.
Pubblicazione: (2026)
Incentive-Aligned Multi-Source LLM Summaries
di: Jiang, Yanchen, et al.
Pubblicazione: (2025)
di: Jiang, Yanchen, et al.
Pubblicazione: (2025)
Prompt Stability Scoring for Text Annotation with Large Language Models
di: Barrie, Christopher, et al.
Pubblicazione: (2024)
di: Barrie, Christopher, et al.
Pubblicazione: (2024)
The Language of Bargaining: Linguistic Effects in LLM Negotiations
di: Sinha, Stuti, et al.
Pubblicazione: (2026)
di: Sinha, Stuti, et al.
Pubblicazione: (2026)
Framing the Game: How Context Shapes LLM Decision-Making
di: Robinson, Isaac, et al.
Pubblicazione: (2025)
di: Robinson, Isaac, et al.
Pubblicazione: (2025)
Fairshare Data Pricing via Data Valuation for Large Language Models
di: Zhang, Luyang, et al.
Pubblicazione: (2025)
di: Zhang, Luyang, et al.
Pubblicazione: (2025)
Make an Offer They Can't Refuse: Grounding Bayesian Persuasion in Real-World Dialogues without Pre-Commitment
di: He, Buwei, et al.
Pubblicazione: (2025)
di: He, Buwei, et al.
Pubblicazione: (2025)
Beyond Game Theory Optimal: Profit-Maximizing Poker Agents for No-Limit Holdem
di: Yi, SeungHyun, et al.
Pubblicazione: (2025)
di: Yi, SeungHyun, et al.
Pubblicazione: (2025)
Evolving Diverse Red-team Language Models in Multi-round Multi-agent Games
di: Ma, Chengdong, et al.
Pubblicazione: (2023)
di: Ma, Chengdong, et al.
Pubblicazione: (2023)
Measuring Bargaining Abilities of LLMs: A Benchmark and A Buyer-Enhancement Method
di: Xia, Tian, et al.
Pubblicazione: (2024)
di: Xia, Tian, et al.
Pubblicazione: (2024)
Verification Required: The Impact of Information Credibility on AI Persuasion
di: Mahmud, Saaduddin, et al.
Pubblicazione: (2026)
di: Mahmud, Saaduddin, et al.
Pubblicazione: (2026)
Beyond Arrow's Impossibility: Fairness as an Emergent Property of Multi-Agent Collaboration
di: Chaki, Sayan Kumar, et al.
Pubblicazione: (2026)
di: Chaki, Sayan Kumar, et al.
Pubblicazione: (2026)
EconEvals: Benchmarks and Litmus Tests for Economic Decision-Making by LLM Agents
di: Fish, Sara, et al.
Pubblicazione: (2025)
di: Fish, Sara, et al.
Pubblicazione: (2025)
ALYMPICS: LLM Agents Meet Game Theory -- Exploring Strategic Decision-Making with AI Agents
di: Mao, Shaoguang, et al.
Pubblicazione: (2023)
di: Mao, Shaoguang, et al.
Pubblicazione: (2023)
Seven kinds of equivalent models for generalized coalition logics
di: Chen, Zixuan, et al.
Pubblicazione: (2025)
di: Chen, Zixuan, et al.
Pubblicazione: (2025)
Strategic Collusion of LLM Agents: Market Division in Multi-Commodity Competitions
di: Lin, Ryan Y., et al.
Pubblicazione: (2024)
di: Lin, Ryan Y., et al.
Pubblicazione: (2024)
Modeling reputation-based behavioral biases in school choice
di: Kleinberg, Jon, et al.
Pubblicazione: (2024)
di: Kleinberg, Jon, et al.
Pubblicazione: (2024)
Does bilevel optimization result in more competitive racing behavior?
di: Cinar, Andrew, et al.
Pubblicazione: (2024)
di: Cinar, Andrew, et al.
Pubblicazione: (2024)
The Memory Curse: How Expanded Recall Erodes Cooperative Intent in LLM Agents
di: Liu, Jiayuan, et al.
Pubblicazione: (2026)
di: Liu, Jiayuan, et al.
Pubblicazione: (2026)
PrefBench: Evaluating Zero-Shot LLM Agents in Hidden-Preference Personalized Pricing Negotiations
di: Lei, Yingjie
Pubblicazione: (2026)
di: Lei, Yingjie
Pubblicazione: (2026)
Re-evaluating Open-ended Evaluation of Large Language Models
di: Liu, Siqi, et al.
Pubblicazione: (2025)
di: Liu, Siqi, et al.
Pubblicazione: (2025)
BEDA: Belief Estimation as Probabilistic Constraints for Performing Strategic Dialogue Acts
di: Li, Hengli, et al.
Pubblicazione: (2025)
di: Li, Hengli, et al.
Pubblicazione: (2025)
AI-Generated Compromises for Coalition Formation: Modeling, Simulation, and a Textual Case Study
di: Briman, Eyal, et al.
Pubblicazione: (2025)
di: Briman, Eyal, et al.
Pubblicazione: (2025)
Generating Fair Consensus Statements with Social Choice on Token-Level MDPs
di: Blair, Carter, et al.
Pubblicazione: (2025)
di: Blair, Carter, et al.
Pubblicazione: (2025)
Language Self-Play For Data-Free Training
di: Kuba, Jakub Grudzien, et al.
Pubblicazione: (2025)
di: Kuba, Jakub Grudzien, et al.
Pubblicazione: (2025)
Personality Modeling for Persuasion of Misinformation using AI Agent
di: Lou, Qianmin, et al.
Pubblicazione: (2025)
di: Lou, Qianmin, et al.
Pubblicazione: (2025)
PokerBench: Training Large Language Models to become Professional Poker Players
di: Zhuang, Richard, et al.
Pubblicazione: (2025)
di: Zhuang, Richard, et al.
Pubblicazione: (2025)
Strategic Intelligence in Large Language Models: Evidence from evolutionary Game Theory
di: Payne, Kenneth, et al.
Pubblicazione: (2025)
di: Payne, Kenneth, et al.
Pubblicazione: (2025)
ShortageSim: Simulating Drug Shortages under Information Asymmetry
di: Cui, Mingxuan, et al.
Pubblicazione: (2025)
di: Cui, Mingxuan, et al.
Pubblicazione: (2025)
Incentivizing Inclusive Contributions in Model Sharing Markets
di: Zhang, Enpei, et al.
Pubblicazione: (2025)
di: Zhang, Enpei, et al.
Pubblicazione: (2025)
Dynamic Coalition Structure Detection in Natural Language-based Interactions
di: Kulkarni, Abhishek N., et al.
Pubblicazione: (2025)
di: Kulkarni, Abhishek N., et al.
Pubblicazione: (2025)
Online Learning and Equilibrium Computation with Ranking Feedback
di: Liu, Mingyang, et al.
Pubblicazione: (2026)
di: Liu, Mingyang, et al.
Pubblicazione: (2026)
Why Do LLMs Struggle in Strategic Play? Broken Links Between Observations, Beliefs, and Actions
di: Sobotka, Jan, et al.
Pubblicazione: (2026)
di: Sobotka, Jan, et al.
Pubblicazione: (2026)
How Well Can LLMs Negotiate? NegotiationArena Platform and Analysis
di: Bianchi, Federico, et al.
Pubblicazione: (2024)
di: Bianchi, Federico, et al.
Pubblicazione: (2024)
Are Large Language Models Strategic Decision Makers? A Study of Performance and Bias in Two-Player Non-Zero-Sum Games
di: Herr, Nathan, et al.
Pubblicazione: (2024)
di: Herr, Nathan, et al.
Pubblicazione: (2024)
Eliciting Informative Text Evaluations with Large Language Models
di: Lu, Yuxuan, et al.
Pubblicazione: (2024)
di: Lu, Yuxuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Reply to "Emergent LLM behaviors are observationally equivalent to data leakage"
di: Ashery, Ariel Flint, et al.
Pubblicazione: (2025) -
STRIDE: A Tool-Assisted LLM Agent Framework for Strategic and Interactive Decision-Making
di: Li, Chuanhao, et al.
Pubblicazione: (2024) -
Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas
di: Gallego, Víctor
Pubblicazione: (2026) -
How memory can affect collective and cooperative behaviors in an LLM-Based Social Particle Swarm
di: Hishiki, Taisei, et al.
Pubblicazione: (2026) -
Ad Insertion in LLM-Generated Responses
di: Xu, Shengwei, et al.
Pubblicazione: (2026)