Salvato in:
| Autori principali: | Jeknić, Isidora, Schlangen, David, Koller, Alexander |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2406.08202 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Collaborative Problem-Solving in an Optimization Game
di: Jeknic, Isidora, et al.
Pubblicazione: (2025)
di: Jeknic, Isidora, et al.
Pubblicazione: (2025)
Can Visual Dialogue Models Do Scorekeeping? Exploring How Dialogue Representations Incrementally Encode Shared Knowledge
di: Madureira, Brielen, et al.
Pubblicazione: (2022)
di: Madureira, Brielen, et al.
Pubblicazione: (2022)
Learning Communication Policies for Different Follower Behaviors in a Collaborative Reference Game
di: Sadler, Philipp, et al.
Pubblicazione: (2024)
di: Sadler, Philipp, et al.
Pubblicazione: (2024)
Multi-Turn Multi-Agent Dialogue for Collaborative Reconstruction Improves VLM Performance on Spatial Reasoning, But Only Barely
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2026)
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2026)
A Third Paradigm for LLM Evaluation: Dialogue Game-Based Evaluation using clembench
di: Schlangen, David, et al.
Pubblicazione: (2025)
di: Schlangen, David, et al.
Pubblicazione: (2025)
Sharing the Cost of Success: A Game for Evaluating and Learning Collaborative Multi-Agent Instruction Giving and Following Policies
di: Sadler, Philipp, et al.
Pubblicazione: (2024)
di: Sadler, Philipp, et al.
Pubblicazione: (2024)
The Image Reconstruction Game: Drawing Common Ground Through Iterative Multimodal Dialogue
di: Hakimov, Sherzod, et al.
Pubblicazione: (2026)
di: Hakimov, Sherzod, et al.
Pubblicazione: (2026)
Triangulating LLM Progress through Benchmarks, Games, and Cognitive Tests
di: Momentè, Filippo, et al.
Pubblicazione: (2025)
di: Momentè, Filippo, et al.
Pubblicazione: (2025)
clem:todd: A Framework for the Systematic Benchmarking of LLM-Based Task-Oriented Dialogue System Realisations
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2025)
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2025)
Prior Lessons of Incremental Dialogue and Robot Action Management for the Age of Language Models
di: Kennington, Casey, et al.
Pubblicazione: (2025)
di: Kennington, Casey, et al.
Pubblicazione: (2025)
Ad-hoc Concept Forming in the Game Codenames as a Means for Evaluating Large Language Models
di: Hakimov, Sherzod, et al.
Pubblicazione: (2025)
di: Hakimov, Sherzod, et al.
Pubblicazione: (2025)
LLMs as Function Approximators: Terminology, Taxonomy, and Questions for Evaluation
di: Schlangen, David
Pubblicazione: (2024)
di: Schlangen, David
Pubblicazione: (2024)
Taking Action Towards Graceful Interaction: The Effects of Performing Actions on Modelling Policies for Instruction Clarification Requests
di: Madureira, Brielen, et al.
Pubblicazione: (2024)
di: Madureira, Brielen, et al.
Pubblicazione: (2024)
It Couldn't Help But Overhear: On the Limits of Modelling Meta-Communicative Grounding Acts with Supervised Learning
di: Madureira, Brielen, et al.
Pubblicazione: (2024)
di: Madureira, Brielen, et al.
Pubblicazione: (2024)
Could the Road to Grounded, Neuro-symbolic AI be Paved with Words-as-Classifiers?
di: Kennington, Casey, et al.
Pubblicazione: (2025)
di: Kennington, Casey, et al.
Pubblicazione: (2025)
Incremental Processing in the Age of Non-Incremental Encoders: An Empirical Assessment of Bidirectional Models for Incremental NLU
di: Madureira, Brielen, et al.
Pubblicazione: (2020)
di: Madureira, Brielen, et al.
Pubblicazione: (2020)
How Many Parameters Does it Take to Change a Light Bulb? Evaluating Performance in Self-Play of Conversational Games as a Function of Model Characteristics
di: Bhavsar, Nidhir, et al.
Pubblicazione: (2024)
di: Bhavsar, Nidhir, et al.
Pubblicazione: (2024)
Characterizing Language Use in a Collaborative Situated Game
di: Tomlin, Nicholas, et al.
Pubblicazione: (2025)
di: Tomlin, Nicholas, et al.
Pubblicazione: (2025)
Retrieval-Augmented Code Generation for Situated Action Generation: A Case Study on Minecraft
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2024)
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2024)
A Survey on Complex Tasks for Goal-Directed Interactive Agents
di: Hartmann, Mareike, et al.
Pubblicazione: (2024)
di: Hartmann, Mareike, et al.
Pubblicazione: (2024)
When Only Time Will Tell: Interpreting How Transformers Process Local Ambiguities Through the Lens of Restart-Incrementality
di: Madureira, Brielen, et al.
Pubblicazione: (2024)
di: Madureira, Brielen, et al.
Pubblicazione: (2024)
The Unreasonable Ineffectiveness of Nucleus Sampling on Mitigating Text Memorization
di: Borec, Luka, et al.
Pubblicazione: (2024)
di: Borec, Luka, et al.
Pubblicazione: (2024)
Towards Incremental Transformers: An Empirical Analysis of Transformer Models for Incremental NLU
di: Kahardipraja, Patrick, et al.
Pubblicazione: (2021)
di: Kahardipraja, Patrick, et al.
Pubblicazione: (2021)
Plant in Cupboard, Orange on Rably, Inat Aphone. Benchmarking Incremental Learning of Situation and Language Model using a Text-Simulated Situated Environment
di: Jordan, Jonathan, et al.
Pubblicazione: (2025)
di: Jordan, Jonathan, et al.
Pubblicazione: (2025)
From Templates to Natural Language: Generalization Challenges in Instruction-Tuned LLMs for Spatial Reasoning
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2025)
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2025)
Simple and effective data augmentation for compositional generalization
di: Yao, Yuekun, et al.
Pubblicazione: (2024)
di: Yao, Yuekun, et al.
Pubblicazione: (2024)
Fine-grained Controllable Text Generation through In-context Learning with Feedback
di: Thillainathan, Sarubi, et al.
Pubblicazione: (2024)
di: Thillainathan, Sarubi, et al.
Pubblicazione: (2024)
Predicting generalization performance with correctness discriminators
di: Yao, Yuekun, et al.
Pubblicazione: (2023)
di: Yao, Yuekun, et al.
Pubblicazione: (2023)
SafeTy Reasoning Elicitation Alignment for Multi-Turn Dialogues
di: Kuo, Martin, et al.
Pubblicazione: (2025)
di: Kuo, Martin, et al.
Pubblicazione: (2025)
Towards No-Code Programming of Cobots: Experiments with Code Synthesis by Large Code Models for Conversational Programming
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2024)
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2024)
Representations of Fact, Fiction and Forecast in Large Language Models: Epistemics and Attitudes
di: Li, Meng, et al.
Pubblicazione: (2025)
di: Li, Meng, et al.
Pubblicazione: (2025)
Using Game Play to Investigate Multimodal and Conversational Grounding in Large Multimodal Models
di: Hakimov, Sherzod, et al.
Pubblicazione: (2024)
di: Hakimov, Sherzod, et al.
Pubblicazione: (2024)
Synthetic Dialogue Generation for Interactive Conversational Elicitation & Recommendation (ICER)
di: Ryu, Moonkyung, et al.
Pubblicazione: (2025)
di: Ryu, Moonkyung, et al.
Pubblicazione: (2025)
Deflanderization for Game Dialogue: Balancing Character Authenticity with Task Execution in LLM-based NPCs
di: Buakhaw, Pasin, et al.
Pubblicazione: (2025)
di: Buakhaw, Pasin, et al.
Pubblicazione: (2025)
Strengthening Structural Inductive Biases by Pre-training to Perform Syntactic Transformations
di: Lindemann, Matthias, et al.
Pubblicazione: (2024)
di: Lindemann, Matthias, et al.
Pubblicazione: (2024)
SIP: Injecting a Structural Inductive Bias into a Seq2Seq Model by Simulation
di: Lindemann, Matthias, et al.
Pubblicazione: (2023)
di: Lindemann, Matthias, et al.
Pubblicazione: (2023)
LLMs syntactically adapt their language use to their conversational partner
di: Kandra, Florian, et al.
Pubblicazione: (2025)
di: Kandra, Florian, et al.
Pubblicazione: (2025)
Direct Neural Machine Translation with Task-level Mixture of Experts models
di: Tourni, Isidora Chara, et al.
Pubblicazione: (2023)
di: Tourni, Isidora Chara, et al.
Pubblicazione: (2023)
Towards Negotiative Dialogue for the Talkamatic Dialogue Manager
di: Larsson, Staffan, et al.
Pubblicazione: (2024)
di: Larsson, Staffan, et al.
Pubblicazione: (2024)
A Knapsack by Any Other Name: Presentation impacts LLM performance on NP-hard problems
di: Duchnowski, Alex, et al.
Pubblicazione: (2025)
di: Duchnowski, Alex, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Collaborative Problem-Solving in an Optimization Game
di: Jeknic, Isidora, et al.
Pubblicazione: (2025) -
Can Visual Dialogue Models Do Scorekeeping? Exploring How Dialogue Representations Incrementally Encode Shared Knowledge
di: Madureira, Brielen, et al.
Pubblicazione: (2022) -
Learning Communication Policies for Different Follower Behaviors in a Collaborative Reference Game
di: Sadler, Philipp, et al.
Pubblicazione: (2024) -
Multi-Turn Multi-Agent Dialogue for Collaborative Reconstruction Improves VLM Performance on Spatial Reasoning, But Only Barely
di: Kranti, Chalamalasetti, et al.
Pubblicazione: (2026) -
A Third Paradigm for LLM Evaluation: Dialogue Game-Based Evaluation using clembench
di: Schlangen, David, et al.
Pubblicazione: (2025)