Don't lie to your friends: Learning what you know from collaborative self-play
Fuente:
arXiv
Saved in:
| Main Authors: | Eisenstein, Jacob, Aghajani, Reza, Fisch, Adam, Dua, Dheeru, Huot, Fantine, Lapata, Mirella, Zayats, Vicky, Berant, Jonathan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Steerable Clarification Policies with Collaborative Self-play
by: Berant, Jonathan, et al.
Published: (2025)
by: Berant, Jonathan, et al.
Published: (2025)
MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games
by: Eisenstein, Jacob, et al.
Published: (2026)
by: Eisenstein, Jacob, et al.
Published: (2026)
Help Me Write a Story: Evaluating LLMs' Ability to Generate Writing Feedback
by: Rashkin, Hannah, et al.
Published: (2025)
by: Rashkin, Hannah, et al.
Published: (2025)
Evaluating LLMs for Targeted Concept Simplification for Domain-Specific Texts
by: Asthana, Sumit, et al.
Published: (2024)
by: Asthana, Sumit, et al.
Published: (2024)
Robust Preference Optimization through Reward Model Distillation
by: Fisch, Adam, et al.
Published: (2024)
by: Fisch, Adam, et al.
Published: (2024)
Low-Rank Adaptation for Multilingual Summarization: An Empirical Study
by: Whitehouse, Chenxi, et al.
Published: (2023)
by: Whitehouse, Chenxi, et al.
Published: (2023)
DOLOMITES: Domain-Specific Long-Form Methodical Tasks
by: Malaviya, Chaitanya, et al.
Published: (2024)
by: Malaviya, Chaitanya, et al.
Published: (2024)
Agents' Room: Narrative Generation through Multi-step Collaboration
by: Huot, Fantine, et al.
Published: (2024)
by: Huot, Fantine, et al.
Published: (2024)
Learning to Plan and Generate Text with Citations
by: Fierro, Constanza, et al.
Published: (2024)
by: Fierro, Constanza, et al.
Published: (2024)
Cost-Optimal Active AI Model Evaluation
by: Angelopoulos, Anastasios N., et al.
Published: (2025)
by: Angelopoulos, Anastasios N., et al.
Published: (2025)
$μ$PLAN: Summarizing using a Content Plan as Cross-Lingual Bridge
by: Huot, Fantine, et al.
Published: (2023)
by: Huot, Fantine, et al.
Published: (2023)
Plantain: Plan-Answer Interleaved Reasoning
by: Liang, Anthony, et al.
Published: (2025)
by: Liang, Anthony, et al.
Published: (2025)
Don't drop your samples! Coherence-aware training benefits Conditional diffusion
by: Dufour, Nicolas, et al.
Published: (2024)
by: Dufour, Nicolas, et al.
Published: (2024)
Think Before you Write: QA-Guided Reasoning for Character Descriptions in Books
by: Papoudakis, Argyrios, et al.
Published: (2026)
by: Papoudakis, Argyrios, et al.
Published: (2026)
Europe. We know what's best for your old age, why can't you see it?
Published: (2002)
Published: (2002)
Do you know what q-means?
by: Cornelissen, Arjan, et al.
Published: (2023)
by: Cornelissen, Arjan, et al.
Published: (2023)
It’s not what you know but who you know: Heterogeneous peer effects at a Colombian University
by: Ana María Díaz
Published: (2018)
by: Ana María Díaz
Published: (2018)
Essential information : what to know before you go
Published: (2005)
Published: (2005)
ABCs of personal investments: what you need to know
by: Ruth Downs
Published: (2025)
by: Ruth Downs
Published: (2025)
K*-Means: A Parameter-free Clustering Algorithm
by: Mahon, Louis, et al.
Published: (2025)
by: Mahon, Louis, et al.
Published: (2025)
Learning to Reason for Long-Form Story Generation
by: Gurung, Alexander, et al.
Published: (2025)
by: Gurung, Alexander, et al.
Published: (2025)
Multimodal Latent Reasoning via Predictive Embeddings
by: Adhikari, Ashutosh, et al.
Published: (2026)
by: Adhikari, Ashutosh, et al.
Published: (2026)
Reasoning about Intent for Ambiguous Requests
by: Saparina, Irina, et al.
Published: (2025)
by: Saparina, Irina, et al.
Published: (2025)
A Modular Approach for Multimodal Summarization of TV Shows
by: Mahon, Louis, et al.
Published: (2024)
by: Mahon, Louis, et al.
Published: (2024)
ScreenWriter: Automatic Screenplay Generation and Movie Summarisation
by: Mahon, Louis, et al.
Published: (2024)
by: Mahon, Louis, et al.
Published: (2024)
Debating for Better Reasoning: An Unsupervised Multimodal Approach
by: Adhikari, Ashutosh, et al.
Published: (2025)
by: Adhikari, Ashutosh, et al.
Published: (2025)
AMBROSIA: A Benchmark for Parsing Ambiguous Questions into Database Queries
by: Saparina, Irina, et al.
Published: (2024)
by: Saparina, Irina, et al.
Published: (2024)
CHIRON: Rich Character Representations in Long-Form Narratives
by: Gurung, Alexander, et al.
Published: (2024)
by: Gurung, Alexander, et al.
Published: (2024)
Integrating Large Language Models with Graph-based Reasoning for Conversational Question Answering
by: Jain, Parag, et al.
Published: (2024)
by: Jain, Parag, et al.
Published: (2024)
Disambiguate First, Parse Later: Generating Interpretations for Ambiguity Resolution in Semantic Parsing
by: Saparina, Irina, et al.
Published: (2025)
by: Saparina, Irina, et al.
Published: (2025)
Parameter-free Video Segmentation for Vision and Language Understanding
by: Mahon, Louis, et al.
Published: (2025)
by: Mahon, Louis, et al.
Published: (2025)
Context-Aware Hierarchical Merging for Long Document Summarization
by: Ou, Litu, et al.
Published: (2025)
by: Ou, Litu, et al.
Published: (2025)
Improving Generalization in Semantic Parsing by Increasing Natural Language Variation
by: Saparina, Irina, et al.
Published: (2024)
by: Saparina, Irina, et al.
Published: (2024)
Don't Forget your Inverse DDIM for Image Editing
by: Gomez-Trenado, Guillermo, et al.
Published: (2025)
by: Gomez-Trenado, Guillermo, et al.
Published: (2025)
What do your logits know? (The answer may surprise you!)
by: Fedzechkina, Masha, et al.
Published: (2026)
by: Fedzechkina, Masha, et al.
Published: (2026)
Uncertainty Quantification in Retrieval Augmented Question Answering
by: Perez-Beltrachini, Laura, et al.
Published: (2025)
by: Perez-Beltrachini, Laura, et al.
Published: (2025)
Multi‐cohort evaluation of “Don't know” responders to self‐report oral health questions: Implications for etiologic research
by: Julia C. Bond, et al.
Published: (2025)
by: Julia C. Bond, et al.
Published: (2025)
Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning
by: Setlur, Amrith, et al.
Published: (2024)
by: Setlur, Amrith, et al.
Published: (2024)
Don't trust your eyes: on the (un)reliability of feature visualizations
by: Geirhos, Robert, et al.
Published: (2023)
by: Geirhos, Robert, et al.
Published: (2023)
What you think you know shapes what you see: How epistemology shapes curriculum, teaching, and learning
by: Cayla R. Teal, et al.
Published: (2024)
by: Cayla R. Teal, et al.
Published: (2024)
Similar Items
-
Learning Steerable Clarification Policies with Collaborative Self-play
by: Berant, Jonathan, et al.
Published: (2025) -
MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games
by: Eisenstein, Jacob, et al.
Published: (2026) -
Help Me Write a Story: Evaluating LLMs' Ability to Generate Writing Feedback
by: Rashkin, Hannah, et al.
Published: (2025) -
Evaluating LLMs for Targeted Concept Simplification for Domain-Specific Texts
by: Asthana, Sumit, et al.
Published: (2024) -
Robust Preference Optimization through Reward Model Distillation
by: Fisch, Adam, et al.
Published: (2024)