Detecting Referring Expressions in Visually Grounded Dialogue with Autoregressive Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Willemsen, Bram, Skantze, Gabriel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Referring Expression Generation in Visually Grounded Dialogue with Discourse-aware Comprehension Guiding
von: Willemsen, Bram, et al.
Veröffentlicht: (2024)
von: Willemsen, Bram, et al.
Veröffentlicht: (2024)
Joint Learning of Context and Feedback Embeddings in Spoken Dialogue
von: Qian, Livia, et al.
Veröffentlicht: (2024)
von: Qian, Livia, et al.
Veröffentlicht: (2024)
Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning
von: Qian, Livia, et al.
Veröffentlicht: (2026)
von: Qian, Livia, et al.
Veröffentlicht: (2026)
An Analysis of User Behaviors for Objectively Evaluating Spoken Dialogue Systems
von: Inoue, Koji, et al.
Veröffentlicht: (2024)
von: Inoue, Koji, et al.
Veröffentlicht: (2024)
CoT Referring: Improving Referring Expression Tasks with Grounded Reasoning
von: Dong, Qihua, et al.
Veröffentlicht: (2025)
von: Dong, Qihua, et al.
Veröffentlicht: (2025)
Referring Expressions as a Lens into Spatial Language Grounding in Vision-Language Models
von: Tumu, Akshar, et al.
Veröffentlicht: (2025)
von: Tumu, Akshar, et al.
Veröffentlicht: (2025)
Exploring Spatial Language Grounding Through Referring Expressions
von: Tumu, Akshar, et al.
Veröffentlicht: (2025)
von: Tumu, Akshar, et al.
Veröffentlicht: (2025)
Representation of perceived prosodic similarity of conversational feedback
von: Qian, Livia, et al.
Veröffentlicht: (2025)
von: Qian, Livia, et al.
Veröffentlicht: (2025)
Personalized Topic Selection Model for Topic-Grounded Dialogue
von: Fan, Shixuan, et al.
Veröffentlicht: (2024)
von: Fan, Shixuan, et al.
Veröffentlicht: (2024)
DRAGON: A Dialogue-Based Robot for Assistive Navigation with Visual Language Grounding
von: Liu, Shuijing, et al.
Veröffentlicht: (2023)
von: Liu, Shuijing, et al.
Veröffentlicht: (2023)
Reference-free Hallucination Detection for Large Vision-Language Models
von: Li, Qing, et al.
Veröffentlicht: (2024)
von: Li, Qing, et al.
Veröffentlicht: (2024)
SEADialogues: A Multilingual Culturally Grounded Multi-turn Dialogue Dataset on Southeast Asian Languages
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
GPT-SW3: An Autoregressive Language Model for the Nordic Languages
von: Ekgren, Ariel, et al.
Veröffentlicht: (2023)
von: Ekgren, Ariel, et al.
Veröffentlicht: (2023)
Improving the Robustness of Knowledge-Grounded Dialogue via Contrastive Learning
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
Differences in Text Generated by Diffusion and Autoregressive Language Models
von: Zhang, Zeyang, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyang, et al.
Veröffentlicht: (2026)
RecycleGPT: An Autoregressive Language Model with Recyclable Module
von: Jiang, Yufan, et al.
Veröffentlicht: (2023)
von: Jiang, Yufan, et al.
Veröffentlicht: (2023)
Ref-Adv: Exploring MLLM Visual Reasoning in Referring Expression Tasks
von: Dong, Qihua, et al.
Veröffentlicht: (2026)
von: Dong, Qihua, et al.
Veröffentlicht: (2026)
Red Teaming Language Models for Processing Contradictory Dialogues
von: Wen, Xiaofei, et al.
Veröffentlicht: (2024)
von: Wen, Xiaofei, et al.
Veröffentlicht: (2024)
Continuous Autoregressive Language Models
von: Shao, Chenze, et al.
Veröffentlicht: (2025)
von: Shao, Chenze, et al.
Veröffentlicht: (2025)
Sequence-Level Certainty Reduces Hallucination In Knowledge-Grounded Dialogue Generation
von: Wan, Yixin, et al.
Veröffentlicht: (2023)
von: Wan, Yixin, et al.
Veröffentlicht: (2023)
LLMs and Cultural Values: the Impact of Prompt Language and Explicit Cultural Framing
von: Bulté, Bram, et al.
Veröffentlicht: (2025)
von: Bulté, Bram, et al.
Veröffentlicht: (2025)
From Bytes to Ideas: Language Modeling with Autoregressive U-Nets
von: Videau, Mathurin, et al.
Veröffentlicht: (2025)
von: Videau, Mathurin, et al.
Veröffentlicht: (2025)
Beyond Single-User Dialogue: Assessing Multi-User Dialogue State Tracking Capabilities of Large Language Models
von: Song, Sangmin, et al.
Veröffentlicht: (2025)
von: Song, Sangmin, et al.
Veröffentlicht: (2025)
Chronological Thinking in Full-Duplex Spoken Dialogue Language Models
von: Wu, Donghang, et al.
Veröffentlicht: (2025)
von: Wu, Donghang, et al.
Veröffentlicht: (2025)
Empathy by Design: Aligning Large Language Models for Healthcare Dialogue
von: Umucu, Emre, et al.
Veröffentlicht: (2025)
von: Umucu, Emre, et al.
Veröffentlicht: (2025)
Empirical Analysis of Dialogue Relation Extraction with Large Language Models
von: Li, Guozheng, et al.
Veröffentlicht: (2024)
von: Li, Guozheng, et al.
Veröffentlicht: (2024)
DFlow: Diverse Dialogue Flow Simulation with Large Language Models
von: Du, Wanyu, et al.
Veröffentlicht: (2024)
von: Du, Wanyu, et al.
Veröffentlicht: (2024)
Circuit Complexity Bounds for Visual Autoregressive Model
von: Ke, Yekun, et al.
Veröffentlicht: (2025)
von: Ke, Yekun, et al.
Veröffentlicht: (2025)
Transcoders Trace Visual Grounding and Hallucinations in Vision-Language Models
von: Damianos, Dimitrios, et al.
Veröffentlicht: (2026)
von: Damianos, Dimitrios, et al.
Veröffentlicht: (2026)
Lexicon-Level Contrastive Visual-Grounding Improves Language Modeling
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2024)
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2024)
Grounded Misunderstandings in Asymmetric Dialogue: A Perspectivist Annotation Scheme for MapTask
von: Li, Nan, et al.
Veröffentlicht: (2025)
von: Li, Nan, et al.
Veröffentlicht: (2025)
Simulated Annealing Enhances Theory-of-Mind Reasoning in Autoregressive Language Models
von: Hu, Xucong, et al.
Veröffentlicht: (2026)
von: Hu, Xucong, et al.
Veröffentlicht: (2026)
STRIDE-ED: A Strategy-Grounded Stepwise Reasoning Framework for Empathetic Dialogue Systems
von: Ji, Hongru, et al.
Veröffentlicht: (2026)
von: Ji, Hongru, et al.
Veröffentlicht: (2026)
SAM4MLLM: Enhance Multi-Modal Large Language Model for Referring Expression Segmentation
von: Chen, Yi-Chia, et al.
Veröffentlicht: (2024)
von: Chen, Yi-Chia, et al.
Veröffentlicht: (2024)
Why Diffusion Language Models Struggle with Truly Parallel (Non-Autoregressive) Decoding?
von: Li, Pengxiang, et al.
Veröffentlicht: (2026)
von: Li, Pengxiang, et al.
Veröffentlicht: (2026)
A Survey of the Evolution of Language Model-Based Dialogue Systems: Data, Task and Models
von: Wang, Hongru, et al.
Veröffentlicht: (2023)
von: Wang, Hongru, et al.
Veröffentlicht: (2023)
Hexa: Self-Improving for Knowledge-Grounded Dialogue System
von: Jo, Daejin, et al.
Veröffentlicht: (2023)
von: Jo, Daejin, et al.
Veröffentlicht: (2023)
Context Dependence and Reliability in Autoregressive Language Models
von: Sengupta, Poushali, et al.
Veröffentlicht: (2026)
von: Sengupta, Poushali, et al.
Veröffentlicht: (2026)
From Medical Records to Diagnostic Dialogues: A Clinical-Grounded Approach and Dataset for Psychiatric Comorbidity
von: Wan, Tianxi, et al.
Veröffentlicht: (2025)
von: Wan, Tianxi, et al.
Veröffentlicht: (2025)
Longitudinal Abuse and Sentiment Analysis of Hollywood Movie Dialogues using Language Models
von: Chandra, Rohitash, et al.
Veröffentlicht: (2025)
von: Chandra, Rohitash, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Referring Expression Generation in Visually Grounded Dialogue with Discourse-aware Comprehension Guiding
von: Willemsen, Bram, et al.
Veröffentlicht: (2024) -
Joint Learning of Context and Feedback Embeddings in Spoken Dialogue
von: Qian, Livia, et al.
Veröffentlicht: (2024) -
Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning
von: Qian, Livia, et al.
Veröffentlicht: (2026) -
An Analysis of User Behaviors for Objectively Evaluating Spoken Dialogue Systems
von: Inoue, Koji, et al.
Veröffentlicht: (2024) -
CoT Referring: Improving Referring Expression Tasks with Grounded Reasoning
von: Dong, Qihua, et al.
Veröffentlicht: (2025)