When Only Time Will Tell: Interpreting How Transformers Process Local Ambiguities Through the Lens of Restart-Incrementality
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Madureira, Brielen, Kahardipraja, Patrick, Schlangen, David |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Incremental Transformers: An Empirical Analysis of Transformer Models for Incremental NLU
von: Kahardipraja, Patrick, et al.
Veröffentlicht: (2021)
von: Kahardipraja, Patrick, et al.
Veröffentlicht: (2021)
Incremental Processing in the Age of Non-Incremental Encoders: An Empirical Assessment of Bidirectional Models for Incremental NLU
von: Madureira, Brielen, et al.
Veröffentlicht: (2020)
von: Madureira, Brielen, et al.
Veröffentlicht: (2020)
Can Visual Dialogue Models Do Scorekeeping? Exploring How Dialogue Representations Incrementally Encode Shared Knowledge
von: Madureira, Brielen, et al.
Veröffentlicht: (2022)
von: Madureira, Brielen, et al.
Veröffentlicht: (2022)
Taking Action Towards Graceful Interaction: The Effects of Performing Actions on Modelling Policies for Instruction Clarification Requests
von: Madureira, Brielen, et al.
Veröffentlicht: (2024)
von: Madureira, Brielen, et al.
Veröffentlicht: (2024)
It Couldn't Help But Overhear: On the Limits of Modelling Meta-Communicative Grounding Acts with Supervised Learning
von: Madureira, Brielen, et al.
Veröffentlicht: (2024)
von: Madureira, Brielen, et al.
Veröffentlicht: (2024)
clembench-2024: A Challenging, Dynamic, Complementary, Multilingual Benchmark and Underlying Flexible Framework for LLMs as Multi-Action Agents
von: Beyer, Anne, et al.
Veröffentlicht: (2024)
von: Beyer, Anne, et al.
Veröffentlicht: (2024)
How Loud Rumbles Hit Newsstands: A Data Analysis of Coverage and Spatial Bias in German News about Landslides Around the World
von: Madureira, Brielen, et al.
Veröffentlicht: (2026)
von: Madureira, Brielen, et al.
Veröffentlicht: (2026)
The Newsworthiness of Brazilian Distress: A Peak Analysis on Time Series of International Media Attention to Disasters in Brazil
von: Madureira, Brielen, et al.
Veröffentlicht: (2026)
von: Madureira, Brielen, et al.
Veröffentlicht: (2026)
Geolocating News about Extreme Climate Events: A Comparative Analysis of Off-the-Shelf Tools for Toponym Identification in German
von: Madureira, Brielen, et al.
Veröffentlicht: (2026)
von: Madureira, Brielen, et al.
Veröffentlicht: (2026)
Retrieving Floods without Floodlights: Topic Models as Binary Classifiers for Extreme Climate Events in German News
von: Madureira, Brielen, et al.
Veröffentlicht: (2026)
von: Madureira, Brielen, et al.
Veröffentlicht: (2026)
Prior Lessons of Incremental Dialogue and Robot Action Management for the Age of Language Models
von: Kennington, Casey, et al.
Veröffentlicht: (2025)
von: Kennington, Casey, et al.
Veröffentlicht: (2025)
Multi-Turn Multi-Agent Dialogue for Collaborative Reconstruction Improves VLM Performance on Spatial Reasoning, But Only Barely
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2026)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2026)
A Close Look at Decomposition-based XAI-Methods for Transformer Language Models
von: Arras, Leila, et al.
Veröffentlicht: (2025)
von: Arras, Leila, et al.
Veröffentlicht: (2025)
Plant in Cupboard, Orange on Rably, Inat Aphone. Benchmarking Incremental Learning of Situation and Language Model using a Text-Simulated Situated Environment
von: Jordan, Jonathan, et al.
Veröffentlicht: (2025)
von: Jordan, Jonathan, et al.
Veröffentlicht: (2025)
LLMs as Function Approximators: Terminology, Taxonomy, and Questions for Evaluation
von: Schlangen, David
Veröffentlicht: (2024)
von: Schlangen, David
Veröffentlicht: (2024)
The Atlas of In-Context Learning: How Attention Heads Shape In-Context Retrieval Augmentation
von: Kahardipraja, Patrick, et al.
Veröffentlicht: (2025)
von: Kahardipraja, Patrick, et al.
Veröffentlicht: (2025)
Correct-Detect: Balancing Performance and Ambiguity Through the Lens of Coreference Resolution in LLMs
von: Shore, Amber, et al.
Veröffentlicht: (2025)
von: Shore, Amber, et al.
Veröffentlicht: (2025)
Could the Road to Grounded, Neuro-symbolic AI be Paved with Words-as-Classifiers?
von: Kennington, Casey, et al.
Veröffentlicht: (2025)
von: Kennington, Casey, et al.
Veröffentlicht: (2025)
DecoderLens: Layerwise Interpretation of Encoder-Decoder Transformers
von: Langedijk, Anna, et al.
Veröffentlicht: (2023)
von: Langedijk, Anna, et al.
Veröffentlicht: (2023)
Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
von: Lu, Taiming, et al.
Veröffentlicht: (2024)
von: Lu, Taiming, et al.
Veröffentlicht: (2024)
LLMs Can Plan Only If We Tell Them
von: Sel, Bilgehan, et al.
Veröffentlicht: (2025)
von: Sel, Bilgehan, et al.
Veröffentlicht: (2025)
Only One Relation Possible? Modeling the Ambiguity in Event Temporal Relation Extraction
von: Hu, Yutong, et al.
Veröffentlicht: (2024)
von: Hu, Yutong, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Code Generation for Situated Action Generation: A Case Study on Minecraft
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2024)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2024)
A Dialogue Game for Eliciting Balanced Collaboration
von: Jeknić, Isidora, et al.
Veröffentlicht: (2024)
von: Jeknić, Isidora, et al.
Veröffentlicht: (2024)
Learning Communication Policies for Different Follower Behaviors in a Collaborative Reference Game
von: Sadler, Philipp, et al.
Veröffentlicht: (2024)
von: Sadler, Philipp, et al.
Veröffentlicht: (2024)
The Unreasonable Ineffectiveness of Nucleus Sampling on Mitigating Text Memorization
von: Borec, Luka, et al.
Veröffentlicht: (2024)
von: Borec, Luka, et al.
Veröffentlicht: (2024)
Ad-hoc Concept Forming in the Game Codenames as a Means for Evaluating Large Language Models
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2025)
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2025)
clem:todd: A Framework for the Systematic Benchmarking of LLM-Based Task-Oriented Dialogue System Realisations
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2025)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2025)
From Templates to Natural Language: Generalization Challenges in Instruction-Tuned LLMs for Spatial Reasoning
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2025)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2025)
How Many Parameters Does it Take to Change a Light Bulb? Evaluating Performance in Self-Play of Conversational Games as a Function of Model Characteristics
von: Bhavsar, Nidhir, et al.
Veröffentlicht: (2024)
von: Bhavsar, Nidhir, et al.
Veröffentlicht: (2024)
The Image Reconstruction Game: Drawing Common Ground Through Iterative Multimodal Dialogue
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2026)
von: Hakimov, Sherzod, et al.
Veröffentlicht: (2026)
Incremental Sentence Processing Mechanisms in Autoregressive Transformer Language Models
von: Hanna, Michael, et al.
Veröffentlicht: (2024)
von: Hanna, Michael, et al.
Veröffentlicht: (2024)
How Powerful are Decoder-Only Transformer Neural Models?
von: Roberts, Jesse
Veröffentlicht: (2023)
von: Roberts, Jesse
Veröffentlicht: (2023)
Towards No-Code Programming of Cobots: Experiments with Code Synthesis by Large Code Models for Conversational Programming
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2024)
von: Kranti, Chalamalasetti, et al.
Veröffentlicht: (2024)
Sharing the Cost of Success: A Game for Evaluating and Learning Collaborative Multi-Agent Instruction Giving and Following Policies
von: Sadler, Philipp, et al.
Veröffentlicht: (2024)
von: Sadler, Philipp, et al.
Veröffentlicht: (2024)
Representations of Fact, Fiction and Forecast in Large Language Models: Epistemics and Attitudes
von: Li, Meng, et al.
Veröffentlicht: (2025)
von: Li, Meng, et al.
Veröffentlicht: (2025)
Interpreting Transformers Through Attention Head Intervention
von: Kadem, Mason, et al.
Veröffentlicht: (2026)
von: Kadem, Mason, et al.
Veröffentlicht: (2026)
DistillLens: Symmetric Knowledge Distillation Through Logit Lens
von: Dhakal, Manish, et al.
Veröffentlicht: (2026)
von: Dhakal, Manish, et al.
Veröffentlicht: (2026)
When Smiley Turns Hostile: Interpreting How Emojis Trigger LLMs' Toxicity
von: Cui, Shiyao, et al.
Veröffentlicht: (2025)
von: Cui, Shiyao, et al.
Veröffentlicht: (2025)
MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference
von: Park, Jeonghyun, et al.
Veröffentlicht: (2025)
von: Park, Jeonghyun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards Incremental Transformers: An Empirical Analysis of Transformer Models for Incremental NLU
von: Kahardipraja, Patrick, et al.
Veröffentlicht: (2021) -
Incremental Processing in the Age of Non-Incremental Encoders: An Empirical Assessment of Bidirectional Models for Incremental NLU
von: Madureira, Brielen, et al.
Veröffentlicht: (2020) -
Can Visual Dialogue Models Do Scorekeeping? Exploring How Dialogue Representations Incrementally Encode Shared Knowledge
von: Madureira, Brielen, et al.
Veröffentlicht: (2022) -
Taking Action Towards Graceful Interaction: The Effects of Performing Actions on Modelling Policies for Instruction Clarification Requests
von: Madureira, Brielen, et al.
Veröffentlicht: (2024) -
It Couldn't Help But Overhear: On the Limits of Modelling Meta-Communicative Grounding Acts with Supervised Learning
von: Madureira, Brielen, et al.
Veröffentlicht: (2024)