Should I Trust You? Detecting Deception in Negotiations using Counterfactual RL
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wongkamjan, Wichayaporn, Wang, Yanze, Gu, Feng, Peskoff, Denis, Kummerfeld, Jonathan K., May, Jonathan, Boyd-Graber, Jordan Lee |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Personalized Help for Optimizing Low-Skilled Users' Strategy
par: Gu, Feng, et autres
Publié: (2024)
par: Gu, Feng, et autres
Publié: (2024)
More Victories, Less Cooperation: Assessing Cicero's Diplomacy Play
par: Wongkamjan, Wichayaporn, et autres
Publié: (2024)
par: Wongkamjan, Wichayaporn, et autres
Publié: (2024)
What if Red Can Talk? Dynamic Dialogue Generation Using Large Language Models
par: Nananukul, Navapat, et autres
Publié: (2024)
par: Nananukul, Navapat, et autres
Publié: (2024)
COPlanner: Plan to Roll Out Conservatively but to Explore Optimistically for Model-Based RL
par: Wang, Xiyao, et autres
Publié: (2023)
par: Wang, Xiyao, et autres
Publié: (2023)
Your Students Don't Use LLMs Like You Wish They Did
par: Kobler, Sebastian, et autres
Publié: (2026)
par: Kobler, Sebastian, et autres
Publié: (2026)
Simple and Effective Baselines for Code Summarisation Evaluation
par: Robinson, Jade, et autres
Publié: (2025)
par: Robinson, Jade, et autres
Publié: (2025)
AI-Resilient Interfaces
par: Glassman, Elena L., et autres
Publié: (2024)
par: Glassman, Elena L., et autres
Publié: (2024)
Which of These Best Describes Multiple Choice Evaluation with LLMs? A) Forced B) Flawed C) Fixable D) All of the Above
par: Balepur, Nishant, et autres
Publié: (2025)
par: Balepur, Nishant, et autres
Publié: (2025)
KARL: Knowledge-Aware Retrieval and Representations aid Retention and Learning in Students
par: Shu, Matthew, et autres
Publié: (2024)
par: Shu, Matthew, et autres
Publié: (2024)
DiscoTrace: Representing and Comparing Answering Strategies of Humans and LLMs in Information-Seeking Question Answering
par: Srikanth, Neha, et autres
Publié: (2026)
par: Srikanth, Neha, et autres
Publié: (2026)
Labeled Interactive Topic Models
par: Seelman, Kyle, et autres
Publié: (2023)
par: Seelman, Kyle, et autres
Publié: (2023)
Reverse Question Answering: Can an LLM Write a Question so Hard (or Bad) that it Can't Answer?
par: Balepur, Nishant, et autres
Publié: (2024)
par: Balepur, Nishant, et autres
Publié: (2024)
Seamless Deception: Larger Language Models Are Better Knowledge Concealers
par: Ashok, Dhananjay, et autres
Publié: (2026)
par: Ashok, Dhananjay, et autres
Publié: (2026)
Large Language Models Are Effective Human Annotation Assistants, But Not Good Independent Annotators
par: Gu, Feng, et autres
Publié: (2025)
par: Gu, Feng, et autres
Publié: (2025)
How the Advent of Ubiquitous Large Language Models both Stymie and Turbocharge Dynamic Adversarial Question Generation
par: Sung, Yoo Yeon, et autres
Publié: (2024)
par: Sung, Yoo Yeon, et autres
Publié: (2024)
The Rise of AI-Generated Content in Wikipedia
par: Brooks, Creston, et autres
Publié: (2024)
par: Brooks, Creston, et autres
Publié: (2024)
AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?
par: Gor, Maharshi, et autres
Publié: (2026)
par: Gor, Maharshi, et autres
Publié: (2026)
NAVIG: Natural Language-guided Analysis with Vision Language Models for Image Geo-localization
par: Zhang, Zheyuan, et autres
Publié: (2025)
par: Zhang, Zheyuan, et autres
Publié: (2025)
Student ID Cards: What You Should Know About Them
par: Hoffman, Jonathan, et autres
Publié: (1973)
par: Hoffman, Jonathan, et autres
Publié: (1973)
An AI-Resilient Text Rendering Technique for Reading and Skimming Documents
par: Gu, Ziwei, et autres
Publié: (2024)
par: Gu, Ziwei, et autres
Publié: (2024)
Whose Boat Does it Float? Improving Personalization in Preference Tuning via Inferred User Personas
par: Balepur, Nishant, et autres
Publié: (2025)
par: Balepur, Nishant, et autres
Publié: (2025)
An Empirical Analysis of Static Analysis Methods for Detection and Mitigation of Code Library Hallucinations
par: Miranda-Pena, Clarissa, et autres
Publié: (2026)
par: Miranda-Pena, Clarissa, et autres
Publié: (2026)
Aligning AI Research with the Needs of Clinical Coding Workflows: Eight Recommendations Based on US Data Analysis and Critical Review
par: Gan, Yidong, et autres
Publié: (2024)
par: Gan, Yidong, et autres
Publié: (2024)
SMART-Editor: A Multi-Agent Framework for Human-Like Design Editing with Structural Integrity
par: Mondal, Ishani, et autres
Publié: (2025)
par: Mondal, Ishani, et autres
Publié: (2025)
PEDANTS: Cheap but Effective and Interpretable Answer Equivalence
par: Li, Zongxia, et autres
Publié: (2024)
par: Li, Zongxia, et autres
Publié: (2024)
Out-of-Distribution Adaptation in Offline RL: Counterfactual Reasoning via Causal Normalizing Flows
par: Cho, Minjae, et autres
Publié: (2024)
par: Cho, Minjae, et autres
Publié: (2024)
GNOME: Generating Negotiations through Open-Domain Mapping of Exchanges
par: Deshpande, Darshan, et autres
Publié: (2024)
par: Deshpande, Darshan, et autres
Publié: (2024)
ProxAnn: Use-Oriented Evaluations of Topic Models and Document Clustering
par: Hoyle, Alexander, et autres
Publié: (2025)
par: Hoyle, Alexander, et autres
Publié: (2025)
Do great minds think alike? Investigating Human-AI Complementarity in Question Answering with CAIMIRA
par: Gor, Maharshi, et autres
Publié: (2024)
par: Gor, Maharshi, et autres
Publié: (2024)
Negotiating Science‐in‐the‐Wild
par: Douglas Allchin, et autres
Publié: (2026)
par: Douglas Allchin, et autres
Publié: (2026)
Supporting Sensemaking of Large Language Model Outputs at Scale
par: Gero, Katy Ilonka, et autres
Publié: (2024)
par: Gero, Katy Ilonka, et autres
Publié: (2024)
Question Negotiation at Orem Public Library.
par: Buckley, Jonathan, et autres
Publié: (1990)
par: Buckley, Jonathan, et autres
Publié: (1990)
Is your benchmark truly adversarial? AdvScore: Evaluating Human-Grounded Adversarialness
par: Sung, Yoo Yeon, et autres
Publié: (2024)
par: Sung, Yoo Yeon, et autres
Publié: (2024)
GRACE: A Granular Benchmark for Evaluating Model Calibration against Human Calibration
par: Sung, Yoo Yeon, et autres
Publié: (2025)
par: Sung, Yoo Yeon, et autres
Publié: (2025)
Closer People Hurt You More: How Social Distance Modulates Deception‐Triggered Trust Decline and Trust Repair
par: Ziying Li, et autres
Publié: (2025)
par: Ziying Li, et autres
Publié: (2025)
CFMatch: Aligning Automated Answer Equivalence Evaluation with Expert Judgments For Open-Domain Question Answering
par: Li, Zongxia, et autres
Publié: (2024)
par: Li, Zongxia, et autres
Publié: (2024)
A Mechanistic Understanding of Alignment Algorithms: A Case Study on DPO and Toxicity
par: Lee, Andrew, et autres
Publié: (2024)
par: Lee, Andrew, et autres
Publié: (2024)
SQLucid: Grounding Natural Language Database Queries with Interactive Explanations
par: Tian, Yuan, et autres
Publié: (2024)
par: Tian, Yuan, et autres
Publié: (2024)
Should metaphysics be (re)conceived as metalinguistic negotiation?
par: Jonathan Knowles
Publié: (2025)
par: Jonathan Knowles
Publié: (2025)
Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users
par: Balepur, Nishant, et autres
Publié: (2026)
par: Balepur, Nishant, et autres
Publié: (2026)
Documents similaires
-
Personalized Help for Optimizing Low-Skilled Users' Strategy
par: Gu, Feng, et autres
Publié: (2024) -
More Victories, Less Cooperation: Assessing Cicero's Diplomacy Play
par: Wongkamjan, Wichayaporn, et autres
Publié: (2024) -
What if Red Can Talk? Dynamic Dialogue Generation Using Large Language Models
par: Nananukul, Navapat, et autres
Publié: (2024) -
COPlanner: Plan to Roll Out Conservatively but to Explore Optimistically for Model-Based RL
par: Wang, Xiyao, et autres
Publié: (2023) -
Your Students Don't Use LLMs Like You Wish They Did
par: Kobler, Sebastian, et autres
Publié: (2026)