Multimodal Conversation Structure Understanding
Fuente:
arXiv
Salvato in:
| Autori principali: | Chang, Kent K., Cramer, Mackenzie Hanh, Ho, Anna, Nguyen, Ti Ti, Yuan, Yilin, Bamman, David |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Subversive Characters and Stereotyping Readers: Characterizing Queer Relationalities with Dialogue-Based Relation Extraction
di: Chang, Kent K., et al.
Pubblicazione: (2024)
di: Chang, Kent K., et al.
Pubblicazione: (2024)
Conversational Orientation Reasoning: Egocentric-to-Allocentric Navigation with Multimodal Chain-of-Thought
di: Huang, Yu Ti
Pubblicazione: (2025)
di: Huang, Yu Ti
Pubblicazione: (2025)
On Classification with Large Language Models in Cultural Analytics
di: Bamman, David, et al.
Pubblicazione: (2024)
di: Bamman, David, et al.
Pubblicazione: (2024)
The Impact of LoRA Adapters on LLMs for Clinical Text Classification Under Computational and Data Constraints
di: Le, Thanh-Dung, et al.
Pubblicazione: (2024)
di: Le, Thanh-Dung, et al.
Pubblicazione: (2024)
Once More, With Feeling: Measuring Emotion of Acting Performances in Contemporary American Film
di: Zhou, Naitian, et al.
Pubblicazione: (2024)
di: Zhou, Naitian, et al.
Pubblicazione: (2024)
Improving Multilingual Social Media Insights: Aspect-based Comment Analysis
di: Zhang, Longyin, et al.
Pubblicazione: (2025)
di: Zhang, Longyin, et al.
Pubblicazione: (2025)
Culture is Not Trivia: Sociocultural Theory for Cultural NLP
di: Zhou, Naitian, et al.
Pubblicazione: (2025)
di: Zhou, Naitian, et al.
Pubblicazione: (2025)
Advancing Singlish Understanding: Bridging the Gap with Datasets and Multimodal Models
di: Wang, Bin, et al.
Pubblicazione: (2025)
di: Wang, Bin, et al.
Pubblicazione: (2025)
PRIMA: Boosting Animal Mesh Recovery with Biological Priors and Test-Time Adaptation
di: Yu, Xiaohang, et al.
Pubblicazione: (2026)
di: Yu, Xiaohang, et al.
Pubblicazione: (2026)
FMPose3D: monocular 3D pose estimation via flow matching
di: Wang, Ti, et al.
Pubblicazione: (2026)
di: Wang, Ti, et al.
Pubblicazione: (2026)
IFEval-Audio: Benchmarking Instruction-Following Capability in Audio-based Large Language Models
di: Gao, Yiming, et al.
Pubblicazione: (2025)
di: Gao, Yiming, et al.
Pubblicazione: (2025)
High-Precision Intelligent Reflecting Surfaces-assisted Positioning Service in 5G Networks with Flexible Numerology
di: Nguyen, Ti Ti, et al.
Pubblicazione: (2024)
di: Nguyen, Ti Ti, et al.
Pubblicazione: (2024)
MERaLiON-TextLLM: Cross-Lingual Understanding of Large Language Models in Chinese, Indonesian, Malay, and Singlish
di: Huang, Xin, et al.
Pubblicazione: (2024)
di: Huang, Xin, et al.
Pubblicazione: (2024)
Direct Preference Optimization for English-Mandarin Code-Switching Speech Recognition in Audio LLMs
di: Quang, Trung Nguyen, et al.
Pubblicazione: (2026)
di: Quang, Trung Nguyen, et al.
Pubblicazione: (2026)
Contextual Paralinguistic Data Creation for Multi-Modal Speech-LLM: Data Condensation and Spoken QA Generation
di: Wang, Qiongqiong, et al.
Pubblicazione: (2025)
di: Wang, Qiongqiong, et al.
Pubblicazione: (2025)
Tell, Don't Show: Leveraging Language Models' Abstractive Retellings to Model Literary Themes
di: Lucy, Li, et al.
Pubblicazione: (2025)
di: Lucy, Li, et al.
Pubblicazione: (2025)
Can GRPO Boost Complex Multimodal Table Understanding?
di: Kang, Xiaoqiang, et al.
Pubblicazione: (2025)
di: Kang, Xiaoqiang, et al.
Pubblicazione: (2025)
A Survey of Ontology Expansion for Conversational Understanding
di: Liang, Jinggui, et al.
Pubblicazione: (2024)
di: Liang, Jinggui, et al.
Pubblicazione: (2024)
Query Understanding in LLM-based Conversational Information Seeking
di: Yuan, Yifei, et al.
Pubblicazione: (2025)
di: Yuan, Yifei, et al.
Pubblicazione: (2025)
Structured Attention Matters to Multimodal LLMs in Document Understanding
di: Liu, Chang, et al.
Pubblicazione: (2025)
di: Liu, Chang, et al.
Pubblicazione: (2025)
Emotion and Intent Joint Understanding in Multimodal Conversation: A Benchmarking Dataset
di: Liu, Rui, et al.
Pubblicazione: (2024)
di: Liu, Rui, et al.
Pubblicazione: (2024)
AboutMe: Using Self-Descriptions in Webpages to Document the Effects of English Pretraining Data Filters
di: Lucy, Li, et al.
Pubblicazione: (2024)
di: Lucy, Li, et al.
Pubblicazione: (2024)
Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models
di: Wang, Qiongqiong, et al.
Pubblicazione: (2025)
di: Wang, Qiongqiong, et al.
Pubblicazione: (2025)
CCL-XCoT: An Efficient Cross-Lingual Knowledge Transfer Method for Mitigating Hallucination Generation
di: Zheng, Weihua, et al.
Pubblicazione: (2025)
di: Zheng, Weihua, et al.
Pubblicazione: (2025)
Structured Prompting and LLM Ensembling for Multimodal Conversational Aspect-based Sentiment Analysis
di: Gao, Zhiqiang, et al.
Pubblicazione: (2025)
di: Gao, Zhiqiang, et al.
Pubblicazione: (2025)
MMRC: A Large-Scale Benchmark for Understanding Multimodal Large Language Model in Real-World Conversation
di: Xue, Haochen, et al.
Pubblicazione: (2025)
di: Xue, Haochen, et al.
Pubblicazione: (2025)
SeaEval for Multilingual Foundation Models: From Cross-Lingual Alignment to Cultural Reasoning
di: Wang, Bin, et al.
Pubblicazione: (2023)
di: Wang, Bin, et al.
Pubblicazione: (2023)
From Simulation to Strategy: Automating Personalized Interaction Planning for Conversational Agents
di: Chang, Wen-Yu, et al.
Pubblicazione: (2025)
di: Chang, Wen-Yu, et al.
Pubblicazione: (2025)
Empathy Through Multimodality in Conversational Interfaces
di: Abbasian, Mahyar, et al.
Pubblicazione: (2024)
di: Abbasian, Mahyar, et al.
Pubblicazione: (2024)
Asking Multimodal Clarifying Questions in Mixed-Initiative Conversational Search
di: Yuan, Yifei, et al.
Pubblicazione: (2024)
di: Yuan, Yifei, et al.
Pubblicazione: (2024)
Credit C-GPT: A Domain-Specialized Large Language Model for Conversational Understanding in Vietnamese Debt Collection
di: Hong, Nhung Nguyen Thi, et al.
Pubblicazione: (2026)
di: Hong, Nhung Nguyen Thi, et al.
Pubblicazione: (2026)
Multimodal Transformer Models for Turn-taking Prediction: Effects on Conversational Dynamics of Human-Agent Interaction during Cooperative Gameplay
di: Bae, Young-Ho, et al.
Pubblicazione: (2025)
di: Bae, Young-Ho, et al.
Pubblicazione: (2025)
Benchmarking Contextual Understanding for In-Car Conversational Systems
di: Habicht, Philipp, et al.
Pubblicazione: (2025)
di: Habicht, Philipp, et al.
Pubblicazione: (2025)
Retcon -- a Prompt-Based Technique for Precise Control of LLMs in Conversations
di: Kogan, David, et al.
Pubblicazione: (2026)
di: Kogan, David, et al.
Pubblicazione: (2026)
MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models
di: He, Yingxu, et al.
Pubblicazione: (2024)
di: He, Yingxu, et al.
Pubblicazione: (2024)
Using Game Play to Investigate Multimodal and Conversational Grounding in Large Multimodal Models
di: Hakimov, Sherzod, et al.
Pubblicazione: (2024)
di: Hakimov, Sherzod, et al.
Pubblicazione: (2024)
Conversation Understanding using Relational Temporal Graph Neural Networks with Auxiliary Cross-Modality Interaction
di: Nguyen, Cam-Van Thi, et al.
Pubblicazione: (2023)
di: Nguyen, Cam-Van Thi, et al.
Pubblicazione: (2023)
NEU-ESC: A Comprehensive Vietnamese dataset for Educational Sentiment analysis and topic Classification toward multitask learning
di: Mai, Phan Quoc Hung, et al.
Pubblicazione: (2025)
di: Mai, Phan Quoc Hung, et al.
Pubblicazione: (2025)
MTP: A Dataset for Multi-Modal Turning Points in Casual Conversations
di: Ho, Gia-Bao Dinh, et al.
Pubblicazione: (2024)
di: Ho, Gia-Bao Dinh, et al.
Pubblicazione: (2024)
TriageSim: A Conversational Emergency Triage Simulation Framework from Structured Electronic Health Records
di: Srirag, Dipankar, et al.
Pubblicazione: (2026)
di: Srirag, Dipankar, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Subversive Characters and Stereotyping Readers: Characterizing Queer Relationalities with Dialogue-Based Relation Extraction
di: Chang, Kent K., et al.
Pubblicazione: (2024) -
Conversational Orientation Reasoning: Egocentric-to-Allocentric Navigation with Multimodal Chain-of-Thought
di: Huang, Yu Ti
Pubblicazione: (2025) -
On Classification with Large Language Models in Cultural Analytics
di: Bamman, David, et al.
Pubblicazione: (2024) -
The Impact of LoRA Adapters on LLMs for Clinical Text Classification Under Computational and Data Constraints
di: Le, Thanh-Dung, et al.
Pubblicazione: (2024) -
Once More, With Feeling: Measuring Emotion of Acting Performances in Contemporary American Film
di: Zhou, Naitian, et al.
Pubblicazione: (2024)