ProMediate: A Socio-cognitive framework for evaluating proactive agents in multi-party negotiation
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Ziyi, Sarrafzadeh, Bahar, Zhou, Pei, Yang, Longqi, Zhao, Jieyu, Sharma, Ashish |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Conversational User-AI Intervention: A Study on Prompt Rewriting for Improved LLM Response Generation
by: Sarkar, Rupak, et al.
Published: (2025)
by: Sarkar, Rupak, et al.
Published: (2025)
MARS-SQL: A multi-agent reinforcement learning framework for Text-to-SQL
by: Yang, Haolin, et al.
Published: (2025)
by: Yang, Haolin, et al.
Published: (2025)
Diverse And Private Synthetic Datasets Generation for RAG evaluation: A multi-agent framework
by: Driouich, Ilias, et al.
Published: (2025)
by: Driouich, Ilias, et al.
Published: (2025)
Evaluating LLM-Simulated Conversations in Modeling Inconsistent and Uncollaborative Behaviors in Human Social Interaction
by: Kamoi, Ryo, et al.
Published: (2026)
by: Kamoi, Ryo, et al.
Published: (2026)
Self-Contradictory Reasoning Evaluation and Detection
by: Liu, Ziyi, et al.
Published: (2023)
by: Liu, Ziyi, et al.
Published: (2023)
Teaching Language Models To Gather Information Proactively
by: Huang, Tenghao, et al.
Published: (2025)
by: Huang, Tenghao, et al.
Published: (2025)
Pearl: Personalizing Large Language Model Writing Assistants with Generation-Calibrated Retrievers
by: Mysore, Sheshera, et al.
Published: (2023)
by: Mysore, Sheshera, et al.
Published: (2023)
CBEval: A framework for evaluating and interpreting cognitive biases in LLMs
by: Shaikh, Ammar, et al.
Published: (2024)
by: Shaikh, Ammar, et al.
Published: (2024)
WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback
by: Shi, Taiwei, et al.
Published: (2024)
by: Shi, Taiwei, et al.
Published: (2024)
WiseMind: a knowledge-guided multi-agent framework for accurate and empathetic psychiatric diagnosis
by: Wu, Yuqi, et al.
Published: (2025)
by: Wu, Yuqi, et al.
Published: (2025)
Can LLMs Grasp Implicit Cultural Values? Benchmarking LLMs' Cultural Intelligence with CQ-Bench
by: Liu, Ziyi, et al.
Published: (2025)
by: Liu, Ziyi, et al.
Published: (2025)
One Model, All Roles: Multi-Turn, Multi-Agent Self-Play Reinforcement Learning for Conversational Social Intelligence
by: Jiang, Bowen, et al.
Published: (2026)
by: Jiang, Bowen, et al.
Published: (2026)
GenTool: Enhancing Tool Generalization in Language Models through Zero-to-One and Weak-to-Strong Simulation
by: He, Jie, et al.
Published: (2025)
by: He, Jie, et al.
Published: (2025)
RAM-SD: Retrieval-Augmented Multi-agent framework for Sarcasm Detection
by: Zhou, Ziyang, et al.
Published: (2026)
by: Zhou, Ziyang, et al.
Published: (2026)
AI Sees Your Location, But With A Bias Toward The Wealthy World
by: Huang, Jingyuan, et al.
Published: (2025)
by: Huang, Jingyuan, et al.
Published: (2025)
Beyond Output Critique: Self-Correction via Task Distillation
by: Rahmani, Hossein A., et al.
Published: (2026)
by: Rahmani, Hossein A., et al.
Published: (2026)
Prototypical Human-AI Collaboration Behaviors from LLM-Assisted Writing in the Wild
by: Mysore, Sheshera, et al.
Published: (2025)
by: Mysore, Sheshera, et al.
Published: (2025)
A process algebraic framework for multi-agent dynamic epistemic systems
by: Aldini, Alessandro
Published: (2024)
by: Aldini, Alessandro
Published: (2024)
Deliberate Planning in Language Models with Symbolic Representation
by: Xiong, Siheng, et al.
Published: (2025)
by: Xiong, Siheng, et al.
Published: (2025)
Discovering Knowledge Deficiencies of Language Models on Massive Knowledge Base
by: Song, Linxin, et al.
Published: (2025)
by: Song, Linxin, et al.
Published: (2025)
The Hallucination Tax of Reinforcement Finetuning
by: Song, Linxin, et al.
Published: (2025)
by: Song, Linxin, et al.
Published: (2025)
Can Third-parties Read Our Emotions?
by: Li, Jiayi, et al.
Published: (2025)
by: Li, Jiayi, et al.
Published: (2025)
Leveraging Large Language Models for Collective Decision-Making
by: Papachristou, Marios, et al.
Published: (2023)
by: Papachristou, Marios, et al.
Published: (2023)
m&m's: A Benchmark to Evaluate Tool-Use for multi-step multi-modal Tasks
by: Ma, Zixian, et al.
Published: (2024)
by: Ma, Zixian, et al.
Published: (2024)
Friends-MMC: A Dataset for Multi-modal Multi-party Conversation Understanding
by: Wang, Yueqian, et al.
Published: (2024)
by: Wang, Yueqian, et al.
Published: (2024)
Images Speak Louder than Words: Understanding and Mitigating Bias in Vision-Language Model from a Causal Mediation Perspective
by: Weng, Zhaotian, et al.
Published: (2024)
by: Weng, Zhaotian, et al.
Published: (2024)
Reshaping MOFs text mining with a dynamic multi-agents framework of large language model
by: Lin, Zuhong, et al.
Published: (2025)
by: Lin, Zuhong, et al.
Published: (2025)
Talking with Oompa Loompas: A novel framework for evaluating linguistic acquisition of LLM agents
by: Swain, Sankalp Tattwadarshi, et al.
Published: (2025)
by: Swain, Sankalp Tattwadarshi, et al.
Published: (2025)
Struct-Bench: A Benchmark for Differentially Private Structured Text Generation
by: Wang, Shuaiqi, et al.
Published: (2025)
by: Wang, Shuaiqi, et al.
Published: (2025)
AKReF: An argumentative knowledge representation framework for structured argumentation
by: Bhattacharjee, Debarati, et al.
Published: (2025)
by: Bhattacharjee, Debarati, et al.
Published: (2025)
Safer-Instruct: Aligning Language Models with Automated Preference Data
by: Shi, Taiwei, et al.
Published: (2023)
by: Shi, Taiwei, et al.
Published: (2023)
LiTransProQA: an LLM-based Literary Translation evaluation metric with Professional Question Answering
by: Zhang, Ran, et al.
Published: (2025)
by: Zhang, Ran, et al.
Published: (2025)
Evaluating Style-Personalized Text Generation: Challenges and Directions
by: Jangra, Anubhav, et al.
Published: (2025)
by: Jangra, Anubhav, et al.
Published: (2025)
Analyzing Uncertainty of LLM-as-a-Judge: Interval Evaluations with Conformal Prediction
by: Sheng, Huanxin, et al.
Published: (2025)
by: Sheng, Huanxin, et al.
Published: (2025)
Efficient Reinforcement Finetuning via Adaptive Curriculum Learning
by: Shi, Taiwei, et al.
Published: (2025)
by: Shi, Taiwei, et al.
Published: (2025)
Video-Based Reward Modeling for Computer-Use Agents
by: Song, Linxin, et al.
Published: (2026)
by: Song, Linxin, et al.
Published: (2026)
A unified foundational framework for knowledge injection and evaluation of Large Language Models in Combustion Science
by: Yang, Zonglin, et al.
Published: (2026)
by: Yang, Zonglin, et al.
Published: (2026)
Multi-party Response Generation with Relation Disentanglement
by: Dai, Tianhao, et al.
Published: (2024)
by: Dai, Tianhao, et al.
Published: (2024)
Correcting misinformation on social media with a large language model
by: Zhou, Xinyi, et al.
Published: (2024)
by: Zhou, Xinyi, et al.
Published: (2024)
RadEval: A framework for radiology text evaluation
by: Xu, Justin, et al.
Published: (2025)
by: Xu, Justin, et al.
Published: (2025)
Similar Items
-
Conversational User-AI Intervention: A Study on Prompt Rewriting for Improved LLM Response Generation
by: Sarkar, Rupak, et al.
Published: (2025) -
MARS-SQL: A multi-agent reinforcement learning framework for Text-to-SQL
by: Yang, Haolin, et al.
Published: (2025) -
Diverse And Private Synthetic Datasets Generation for RAG evaluation: A multi-agent framework
by: Driouich, Ilias, et al.
Published: (2025) -
Evaluating LLM-Simulated Conversations in Modeling Inconsistent and Uncollaborative Behaviors in Human Social Interaction
by: Kamoi, Ryo, et al.
Published: (2026) -
Self-Contradictory Reasoning Evaluation and Detection
by: Liu, Ziyi, et al.
Published: (2023)