The Collaboration Gap
Fuente:
arXiv
Saved in:
| Main Authors: | Davidson, Tim R., Fourney, Adam, Amershi, Saleema, West, Robert, Horvitz, Eric, Kamar, Ece |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Challenges in Human-Agent Communication
by: Bansal, Gagan, et al.
Published: (2024)
by: Bansal, Gagan, et al.
Published: (2024)
Tandem Training for Language Models
by: West, Robert, et al.
Published: (2025)
by: West, Robert, et al.
Published: (2025)
AutoGen Studio: A No-Code Developer Tool for Building and Debugging Multi-Agent Systems
by: Dibia, Victor, et al.
Published: (2024)
by: Dibia, Victor, et al.
Published: (2024)
Accurate Measures of Vaccination and Concerns of Vaccine Holdouts from Web Search Logs
by: Chang, Serina, et al.
Published: (2023)
by: Chang, Serina, et al.
Published: (2023)
Interactive Debugging and Steering of Multi-Agent AI Systems
by: Epperson, Will, et al.
Published: (2025)
by: Epperson, Will, et al.
Published: (2025)
Overseeing Agents Without Constant Oversight: Challenges and Opportunities
by: Grunde-McLaughlin, Madeleine, et al.
Published: (2026)
by: Grunde-McLaughlin, Madeleine, et al.
Published: (2026)
"Is This It?": Towards Ecologically Valid Benchmarks for Situated Collaboration
by: Bohus, Dan, et al.
Published: (2024)
by: Bohus, Dan, et al.
Published: (2024)
JSONSchemaBench: A Rigorous Benchmark of Structured Outputs for Language Models
by: Geng, Saibo, et al.
Published: (2025)
by: Geng, Saibo, et al.
Published: (2025)
Self-Recognition in Language Models
by: Davidson, Tim R., et al.
Published: (2024)
by: Davidson, Tim R., et al.
Published: (2024)
Magentic-UI: Towards Human-in-the-loop Agentic Systems
by: Mozannar, Hussein, et al.
Published: (2025)
by: Mozannar, Hussein, et al.
Published: (2025)
Navigating Rifts in Human-LLM Grounding: Study and Benchmark
by: Shaikh, Omar, et al.
Published: (2025)
by: Shaikh, Omar, et al.
Published: (2025)
Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks
by: Fourney, Adam, et al.
Published: (2024)
by: Fourney, Adam, et al.
Published: (2024)
Evaluating Language Model Agency through Negotiations
by: Davidson, Tim R., et al.
Published: (2024)
by: Davidson, Tim R., et al.
Published: (2024)
Getting Serious about Humor: Crafting Humor Datasets with Unfunny Large Language Models
by: Horvitz, Zachary, et al.
Published: (2024)
by: Horvitz, Zachary, et al.
Published: (2024)
Attention Satisfies: A Constraint-Satisfaction Lens on Factual Errors of Language Models
by: Yuksekgonul, Mert, et al.
Published: (2023)
by: Yuksekgonul, Mert, et al.
Published: (2023)
Towards better Human-Agent Alignment: Assessing Task Utility in LLM-Powered Applications
by: Arabzadeh, Negar, et al.
Published: (2024)
by: Arabzadeh, Negar, et al.
Published: (2024)
Improving Instruction-Following in Language Models through Activation Steering
by: Stolfo, Alessandro, et al.
Published: (2024)
by: Stolfo, Alessandro, et al.
Published: (2024)
ParaGuide: Guided Diffusion Paraphrasers for Plug-and-Play Textual Style Transfer
by: Horvitz, Zachary, et al.
Published: (2023)
by: Horvitz, Zachary, et al.
Published: (2023)
Creating General User Models from Computer Use
by: Shaikh, Omar, et al.
Published: (2025)
by: Shaikh, Omar, et al.
Published: (2025)
Challenging the Machine: Contestability in Government AI Systems
by: Landau, Susan, et al.
Published: (2024)
by: Landau, Susan, et al.
Published: (2024)
Knowledge-Infused Legal Wisdom: Navigating LLM Consultation through the Lens of Diagnostics and Positive-Unlabeled Reinforcement Learning
by: Wu, Yang, et al.
Published: (2024)
by: Wu, Yang, et al.
Published: (2024)
Reasoning-Driven Synthetic Data Generation and Evaluation
by: Davidson, Tim R., et al.
Published: (2026)
by: Davidson, Tim R., et al.
Published: (2026)
Response: Emergent analogical reasoning in large language models
by: Hodel, Damian, et al.
Published: (2023)
by: Hodel, Damian, et al.
Published: (2023)
Separating Tongue from Thought: Activation Patching Reveals Language-Agnostic Concept Representations in Transformers
by: Dumas, Clément, et al.
Published: (2024)
by: Dumas, Clément, et al.
Published: (2024)
Tracing Persona Vectors Through LLM Pretraining
by: Moskvoretskii, Viktor, et al.
Published: (2026)
by: Moskvoretskii, Viktor, et al.
Published: (2026)
Activation Scaling for Steering and Interpreting Language Models
by: Stoehr, Niklas, et al.
Published: (2024)
by: Stoehr, Niklas, et al.
Published: (2024)
Describing Images $\textit{Fast and Slow}$: Quantifying and Predicting the Variation in Human Signals during Visuo-Linguistic Processes
by: Takmaz, Ece, et al.
Published: (2024)
by: Takmaz, Ece, et al.
Published: (2024)
Grammar-Constrained Decoding for Structured NLP Tasks without Finetuning
by: Geng, Saibo, et al.
Published: (2023)
by: Geng, Saibo, et al.
Published: (2023)
Narrow Finetuning Leaves Clearly Readable Traces in Activation Differences
by: Minder, Julian, et al.
Published: (2025)
by: Minder, Julian, et al.
Published: (2025)
Controllable Context Sensitivity and the Knob Behind It
by: Minder, Julian, et al.
Published: (2024)
by: Minder, Julian, et al.
Published: (2024)
Completion $\neq$ Collaboration: Scaling Collaborative Effort with Agents
by: Shen, Shannon Zejiang, et al.
Published: (2025)
by: Shen, Shannon Zejiang, et al.
Published: (2025)
Mind the Language Gap: Automated and Augmented Evaluation of Bias in LLMs for High- and Low-Resource Languages
by: Buscemi, Alessio, et al.
Published: (2025)
by: Buscemi, Alessio, et al.
Published: (2025)
Agentic AI: The Era of Semantic Decoding
by: Peyrard, Maxime, et al.
Published: (2024)
by: Peyrard, Maxime, et al.
Published: (2024)
Mitigating the Linguistic Gap with Phonemic Representations for Robust Cross-lingual Transfer
by: Jung, Haeji, et al.
Published: (2024)
by: Jung, Haeji, et al.
Published: (2024)
Beyond Prompts: Dynamic Conversational Benchmarking of Large Language Models
by: Castillo-Bolado, David, et al.
Published: (2024)
by: Castillo-Bolado, David, et al.
Published: (2024)
Handling Ontology Gaps in Semantic Parsing
by: Bacciu, Andrea, et al.
Published: (2024)
by: Bacciu, Andrea, et al.
Published: (2024)
Collaborative Causal Sensemaking: Closing the Complementarity Gap in Human-AI Decision Support
by: Jain, Raunak
Published: (2025)
by: Jain, Raunak
Published: (2025)
Investigating the interaction of linguistic and mathematical reasoning in language models using multilingual number puzzles
by: Bhattacharya, Antara Raaghavi, et al.
Published: (2025)
by: Bhattacharya, Antara Raaghavi, et al.
Published: (2025)
Localized Cultural Knowledge is Conserved and Controllable in Large Language Models
by: Veselovsky, Veniamin, et al.
Published: (2025)
by: Veselovsky, Veniamin, et al.
Published: (2025)
GRAD: Generative Retrieval-Aligned Demonstration Sampler for Efficient Few-Shot Reasoning
by: Gabouj, Oussama, et al.
Published: (2025)
by: Gabouj, Oussama, et al.
Published: (2025)
Similar Items
-
Challenges in Human-Agent Communication
by: Bansal, Gagan, et al.
Published: (2024) -
Tandem Training for Language Models
by: West, Robert, et al.
Published: (2025) -
AutoGen Studio: A No-Code Developer Tool for Building and Debugging Multi-Agent Systems
by: Dibia, Victor, et al.
Published: (2024) -
Accurate Measures of Vaccination and Concerns of Vaccine Holdouts from Web Search Logs
by: Chang, Serina, et al.
Published: (2023) -
Interactive Debugging and Steering of Multi-Agent AI Systems
by: Epperson, Will, et al.
Published: (2025)