Characterizing Language Use in a Collaborative Situated Game
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tomlin, Nicholas, Zhou, Naitian, Fleisig, Eve, Chen, Liangyuan, Wright, Téa, Vinh, Lauren, Ma, Laura X., Eisape, Seun, French, Ellie, Du, Tingting, Zhang, Tianjiao, Koller, Alexander, Suhr, Alane |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Long Chain-of-Thought Reasoning Across Languages
von: Barua, Josh, et al.
Veröffentlicht: (2025)
von: Barua, Josh, et al.
Veröffentlicht: (2025)
Ghostbuster: Detecting Text Ghostwritten by Large Language Models
von: Verma, Vivek, et al.
Veröffentlicht: (2023)
von: Verma, Vivek, et al.
Veröffentlicht: (2023)
TULIP: Towards Unified Language-Image Pretraining
von: Tang, Zineng, et al.
Veröffentlicht: (2025)
von: Tang, Zineng, et al.
Veröffentlicht: (2025)
Autonomous Evaluation and Refinement of Digital Agents
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
ScribbleEdit: Synthetic Data for Image Editing with Scribbles and Text
von: Ji, Anya, et al.
Veröffentlicht: (2026)
von: Ji, Anya, et al.
Veröffentlicht: (2026)
Mapping Social Choice Theory to RLHF
von: Dai, Jessica, et al.
Veröffentlicht: (2024)
von: Dai, Jessica, et al.
Veröffentlicht: (2024)
Evaluating Model Perception of Color Illusions in Photorealistic Scenes
von: Mao, Lingjun, et al.
Veröffentlicht: (2024)
von: Mao, Lingjun, et al.
Veröffentlicht: (2024)
Grounding Language in Multi-Perspective Referential Communication
von: Tang, Zineng, et al.
Veröffentlicht: (2024)
von: Tang, Zineng, et al.
Veröffentlicht: (2024)
When the Majority is Wrong: Modeling Annotator Disagreement for Subjective Tasks
von: Fleisig, Eve, et al.
Veröffentlicht: (2023)
von: Fleisig, Eve, et al.
Veröffentlicht: (2023)
Using Language Models to Disambiguate Lexical Choices in Translation
von: Barua, Josh, et al.
Veröffentlicht: (2024)
von: Barua, Josh, et al.
Veröffentlicht: (2024)
Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
von: Sclar, Melanie, et al.
Veröffentlicht: (2023)
von: Sclar, Melanie, et al.
Veröffentlicht: (2023)
UNcommonsense Reasoning: Abductive Reasoning about Uncommon Situations
von: Zhao, Wenting, et al.
Veröffentlicht: (2023)
von: Zhao, Wenting, et al.
Veröffentlicht: (2023)
Collaborative Problem-Solving in an Optimization Game
von: Jeknic, Isidora, et al.
Veröffentlicht: (2025)
von: Jeknic, Isidora, et al.
Veröffentlicht: (2025)
A Dialogue Game for Eliciting Balanced Collaboration
von: Jeknić, Isidora, et al.
Veröffentlicht: (2024)
von: Jeknić, Isidora, et al.
Veröffentlicht: (2024)
Efficacy of Language Model Self-Play in Non-Zero-Sum Games
von: Liao, Austen, et al.
Veröffentlicht: (2024)
von: Liao, Austen, et al.
Veröffentlicht: (2024)
First Tragedy, then Parse: History Repeats Itself in the New Era of Large Language Models
von: Saphra, Naomi, et al.
Veröffentlicht: (2023)
von: Saphra, Naomi, et al.
Veröffentlicht: (2023)
Balancing Quality and Variation: Spam Filtering Distorts Data Label Distributions
von: Fleisig, Eve, et al.
Veröffentlicht: (2025)
von: Fleisig, Eve, et al.
Veröffentlicht: (2025)
Accurate and Data-Efficient Toxicity Prediction when Annotators Disagree
von: Jaggi, Harbani, et al.
Veröffentlicht: (2024)
von: Jaggi, Harbani, et al.
Veröffentlicht: (2024)
ABBEL: LLM Agents Acting through Belief Bottlenecks Expressed in Language
von: Lidayan, Aly, et al.
Veröffentlicht: (2025)
von: Lidayan, Aly, et al.
Veröffentlicht: (2025)
Reforming collateral laws to expand access to finance / Heywood Fleisig, Mehnaz Safavian, Nuria de la Peña
von: Fleisig, Heywood
Veröffentlicht: (2006)
von: Fleisig, Heywood
Veröffentlicht: (2006)
The Perspectivist Paradigm Shift: Assumptions and Challenges of Capturing Human Labels
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
Decision-Oriented Dialogue for Human-AI Collaboration
von: Lin, Jessy, et al.
Veröffentlicht: (2023)
von: Lin, Jessy, et al.
Veröffentlicht: (2023)
A Knapsack by Any Other Name: Presentation impacts LLM performance on NP-hard problems
von: Duchnowski, Alex, et al.
Veröffentlicht: (2025)
von: Duchnowski, Alex, et al.
Veröffentlicht: (2025)
Community‐Led Forest Management and Conservation: Insights From Nigeria's Obudu and Mambilla Mountain Forests
von: Seun Bamidele
Veröffentlicht: (2025)
von: Seun Bamidele
Veröffentlicht: (2025)
Standard Language Ideology in AI-Generated Language
von: Smith, Genevieve, et al.
Veröffentlicht: (2024)
von: Smith, Genevieve, et al.
Veröffentlicht: (2024)
Measuring General Intelligence with Generated Games
von: Verma, Vivek, et al.
Veröffentlicht: (2025)
von: Verma, Vivek, et al.
Veröffentlicht: (2025)
A Retrospective Study of the Relationship Between Blood Transfusion and 30-Day Postoperative Outcomes in Patients Undergoing Isolated Off-Pump Coronary Artery Bypass Grafting
von: Liangyuan Lu
Veröffentlicht: (2022)
von: Liangyuan Lu
Veröffentlicht: (2022)
OpenAIRE Graph Community Call - January 15 2025
von: Brunschweiger, Alane
Veröffentlicht: (2025)
von: Brunschweiger, Alane
Veröffentlicht: (2025)
Once More, With Feeling: Measuring Emotion of Acting Performances in Contemporary American Film
von: Zhou, Naitian, et al.
Veröffentlicht: (2024)
von: Zhou, Naitian, et al.
Veröffentlicht: (2024)
Model‐Free Approximate Dynamic Programming for Stochastic Zero‐Sum Games: Algorithm Design and Analysis
von: Liangyuan Guo, et al.
Veröffentlicht: (2025)
von: Liangyuan Guo, et al.
Veröffentlicht: (2025)
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
von: Fleisig, Eve, et al.
Veröffentlicht: (2024)
DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning
von: Bai, Hao, et al.
Veröffentlicht: (2024)
von: Bai, Hao, et al.
Veröffentlicht: (2024)
Training Software Engineering Agents and Verifiers with SWE-Gym
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
Characterization of Multiphase Systems in Process Engineering by the Investigation of Their Thermophysical Properties
von: Koller, Thomas Manfred
Veröffentlicht: (2025)
von: Koller, Thomas Manfred
Veröffentlicht: (2025)
Evolução do programa nacional de segurança do paciente: uma análise dos dados públicos disponibilizados pela Agência Nacional de Vigilância Sanitária
von: Alane Martins Andrade
Veröffentlicht: (2020)
von: Alane Martins Andrade
Veröffentlicht: (2020)
Editorial da Edição Especial - 30 anos do Curso de Ciências Atuariais da UFC
von: Alane Siqueira Rocha
Veröffentlicht: (2023)
von: Alane Siqueira Rocha
Veröffentlicht: (2023)
Is your benchmark truly adversarial? AdvScore: Evaluating Human-Grounded Adversarialness
von: Sung, Yoo Yeon, et al.
Veröffentlicht: (2024)
von: Sung, Yoo Yeon, et al.
Veröffentlicht: (2024)
GRACE: A Granular Benchmark for Evaluating Model Calibration against Human Calibration
von: Sung, Yoo Yeon, et al.
Veröffentlicht: (2025)
von: Sung, Yoo Yeon, et al.
Veröffentlicht: (2025)
Disputed Sovereignty and Forced Displacement: Rethinking the Bakassi Peninsula Crisis Through International Law and Human Security
von: Gbeke Adenuga, et al.
Veröffentlicht: (2025)
von: Gbeke Adenuga, et al.
Veröffentlicht: (2025)
A Dynamic Model of Performative Human-ML Collaboration: Theory and Empirical Evidence
von: Sühr, Tom, et al.
Veröffentlicht: (2024)
von: Sühr, Tom, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Long Chain-of-Thought Reasoning Across Languages
von: Barua, Josh, et al.
Veröffentlicht: (2025) -
Ghostbuster: Detecting Text Ghostwritten by Large Language Models
von: Verma, Vivek, et al.
Veröffentlicht: (2023) -
TULIP: Towards Unified Language-Image Pretraining
von: Tang, Zineng, et al.
Veröffentlicht: (2025) -
Autonomous Evaluation and Refinement of Digital Agents
von: Pan, Jiayi, et al.
Veröffentlicht: (2024) -
ScribbleEdit: Synthetic Data for Image Editing with Scribbles and Text
von: Ji, Anya, et al.
Veröffentlicht: (2026)