Multimodal Fusion with LLMs for Engagement Prediction in Natural Conversation
Fuente:
arXiv
Salvato in:
| Autori principali: | Ma, Cheng Charles, Joo, Kevin Hyekang, Vail, Alexandria K., Bhattacharya, Sunreeta, García, Álvaro Fernández, Baker-Matsuoka, Kailana, Mathew, Sheryl, Holt, Lori L., De la Torre, Fernando |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ADAPTS: Agentic Decomposition for Automated Protocol-agnostic Tracking of Symptoms
di: Vail, Alexandria K., et al.
Pubblicazione: (2026)
di: Vail, Alexandria K., et al.
Pubblicazione: (2026)
Effect of Performance Feedback Timing on Motor Learning for a Surgical Training Task
di: Gale, Mary Kate, et al.
Pubblicazione: (2025)
di: Gale, Mary Kate, et al.
Pubblicazione: (2025)
Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection
di: Lee, Sanghoon, et al.
Pubblicazione: (2026)
di: Lee, Sanghoon, et al.
Pubblicazione: (2026)
3DPillars: Pillar-based two-stage 3D object detection
di: Noh, Jongyoun, et al.
Pubblicazione: (2025)
di: Noh, Jongyoun, et al.
Pubblicazione: (2025)
LLAMADRS: Evaluating Open-Source LLMs on Real Clinical Interviews--To Reason or Not to Reason?
di: Kebe, Gaoussou Youssouf, et al.
Pubblicazione: (2025)
di: Kebe, Gaoussou Youssouf, et al.
Pubblicazione: (2025)
FESTA: Functionally Equivalent Sampling for Trust Assessment of Multimodal LLMs
di: Bhattacharya, Debarpan, et al.
Pubblicazione: (2025)
di: Bhattacharya, Debarpan, et al.
Pubblicazione: (2025)
Retrieval-Augmented Generation for Natural Language Art Provenance Searches in the Getty Provenance Index
di: Henrickson, Mathew
Pubblicazione: (2025)
di: Henrickson, Mathew
Pubblicazione: (2025)
Can Large Language Models Robustly Perform Natural Language Inference for Japanese Comparatives?
di: Mikami, Yosuke, et al.
Pubblicazione: (2025)
di: Mikami, Yosuke, et al.
Pubblicazione: (2025)
Seeing the Big Picture: Evaluating Multimodal LLMs' Ability to Interpret and Grade Handwritten Student Work
di: Henkel, Owen, et al.
Pubblicazione: (2025)
di: Henkel, Owen, et al.
Pubblicazione: (2025)
Counterfactual Reward Model Training for Bias Mitigation in Multimodal Reinforcement Learning
di: Mathew, Sheryl, et al.
Pubblicazione: (2025)
di: Mathew, Sheryl, et al.
Pubblicazione: (2025)
Animating Language Practice: Engagement with Stylized Conversational Agents in Japanese Learning
di: Rackauckas, Zackary, et al.
Pubblicazione: (2025)
di: Rackauckas, Zackary, et al.
Pubblicazione: (2025)
Iteratively Prompting Multimodal LLMs to Reproduce Natural and AI-Generated Images
di: Naseh, Ali, et al.
Pubblicazione: (2024)
di: Naseh, Ali, et al.
Pubblicazione: (2024)
Reasoning LLMs for User-Aware Multimodal Conversational Agents
di: Rahimi, Hamed, et al.
Pubblicazione: (2025)
di: Rahimi, Hamed, et al.
Pubblicazione: (2025)
Natural Language Processing and Multimodal Stock Price Prediction
di: Taylor, Kevin, et al.
Pubblicazione: (2024)
di: Taylor, Kevin, et al.
Pubblicazione: (2024)
Talk Less, Interact Better: Evaluating In-context Conversational Adaptation in Multimodal LLMs
di: Hua, Yilun, et al.
Pubblicazione: (2024)
di: Hua, Yilun, et al.
Pubblicazione: (2024)
Cross-modal Context Fusion and Adaptive Graph Convolutional Network for Multimodal Conversational Emotion Recognition
di: Feng, Junwei, et al.
Pubblicazione: (2025)
di: Feng, Junwei, et al.
Pubblicazione: (2025)
Investigating Conversational Agents to Support Secondary School Students Learning CSP
di: Frazier, Matthew, et al.
Pubblicazione: (2026)
di: Frazier, Matthew, et al.
Pubblicazione: (2026)
Multimodal Fusion with Semi-Supervised Learning Minimizes Annotation Quantity for Modeling Videoconference Conversation Experience
di: Chang, Andrew, et al.
Pubblicazione: (2025)
di: Chang, Andrew, et al.
Pubblicazione: (2025)
Leveraging Interesting Facts to Enhance User Engagement with Conversational Interfaces
di: Vedula, Nikhita, et al.
Pubblicazione: (2024)
di: Vedula, Nikhita, et al.
Pubblicazione: (2024)
IntrEx: A Dataset for Modeling Engagement in Educational Conversations
di: Tan, Xingwei, et al.
Pubblicazione: (2025)
di: Tan, Xingwei, et al.
Pubblicazione: (2025)
Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs
di: Yu, Zeping, et al.
Pubblicazione: (2025)
di: Yu, Zeping, et al.
Pubblicazione: (2025)
Enhancing Perception Capabilities of Multimodal LLMs with Training-Free Fusion
di: Chen, Zhuokun, et al.
Pubblicazione: (2024)
di: Chen, Zhuokun, et al.
Pubblicazione: (2024)
Text-Guided Layer Fusion Mitigates Hallucination in Multimodal LLMs
di: Lin, Chenchen, et al.
Pubblicazione: (2026)
di: Lin, Chenchen, et al.
Pubblicazione: (2026)
Understanding Nature Engagement Experiences of Blind People
di: Tang, Mengjie, et al.
Pubblicazione: (2026)
di: Tang, Mengjie, et al.
Pubblicazione: (2026)
Understanding the role of FFNs in driving multilingual behaviour in LLMs
di: Bhattacharya, Sunit, et al.
Pubblicazione: (2024)
di: Bhattacharya, Sunit, et al.
Pubblicazione: (2024)
Engagement-Optimized Care: When LLMs become Mental Health Infrastructure
di: Vecchione, Briana, et al.
Pubblicazione: (2026)
di: Vecchione, Briana, et al.
Pubblicazione: (2026)
LLM-I: LLMs are Naturally Interleaved Multimodal Creators
di: Guo, Zirun, et al.
Pubblicazione: (2025)
di: Guo, Zirun, et al.
Pubblicazione: (2025)
DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimation
di: Li, Jia, et al.
Pubblicazione: (2024)
di: Li, Jia, et al.
Pubblicazione: (2024)
Multimodal Shannon Game with Images
di: Zouhar, Vilém, et al.
Pubblicazione: (2023)
di: Zouhar, Vilém, et al.
Pubblicazione: (2023)
ERIT Lightweight Multimodal Dataset for Elderly Emotion Recognition and Multimodal Fusion Evaluation
di: Frieske, Rita, et al.
Pubblicazione: (2024)
di: Frieske, Rita, et al.
Pubblicazione: (2024)
Enhancing Online Learning by Integrating Biosensors and Multimodal Learning Analytics for Detecting and Predicting Student Behavior: A Review
di: Becerra, Alvaro, et al.
Pubblicazione: (2025)
di: Becerra, Alvaro, et al.
Pubblicazione: (2025)
LLMs Get Lost In Multi-Turn Conversation
di: Laban, Philippe, et al.
Pubblicazione: (2025)
di: Laban, Philippe, et al.
Pubblicazione: (2025)
Survival of the Fittest: Evolutionary Adaptation of Policies for Environmental Shifts
di: Paul, Sheryl, et al.
Pubblicazione: (2024)
di: Paul, Sheryl, et al.
Pubblicazione: (2024)
Bridging the Digital Divide in Postsecondary Education: Technology Access for Youth with Disabilities. Information Brief.
di: Burgstahler, Sheryl
Pubblicazione: (2002)
di: Burgstahler, Sheryl
Pubblicazione: (2002)
Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis
di: Bhowmik, Shimanto, et al.
Pubblicazione: (2025)
di: Bhowmik, Shimanto, et al.
Pubblicazione: (2025)
Multimodal Conversation Structure Understanding
di: Chang, Kent K., et al.
Pubblicazione: (2025)
di: Chang, Kent K., et al.
Pubblicazione: (2025)
Optimizing Dataflow Systems for Scalable Interactive Visualization
di: Yang, Junran, et al.
Pubblicazione: (2024)
di: Yang, Junran, et al.
Pubblicazione: (2024)
SpatialGeo:Boosting Spatial Reasoning in Multimodal LLMs via Geometry-Semantics Fusion
di: Guo, Jiajie, et al.
Pubblicazione: (2025)
di: Guo, Jiajie, et al.
Pubblicazione: (2025)
Multi-Layer Visual Feature Fusion in Multimodal LLMs: Methods, Analysis, and Best Practices
di: Lin, Junyan, et al.
Pubblicazione: (2025)
di: Lin, Junyan, et al.
Pubblicazione: (2025)
Centering Emotion Hotspots: Multimodal Local-Global Fusion and Cross-Modal Alignment for Emotion Recognition in Conversations
di: Liu, Yu, et al.
Pubblicazione: (2025)
di: Liu, Yu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
ADAPTS: Agentic Decomposition for Automated Protocol-agnostic Tracking of Symptoms
di: Vail, Alexandria K., et al.
Pubblicazione: (2026) -
Effect of Performance Feedback Timing on Motor Learning for a Surgical Training Task
di: Gale, Mary Kate, et al.
Pubblicazione: (2025) -
Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection
di: Lee, Sanghoon, et al.
Pubblicazione: (2026) -
3DPillars: Pillar-based two-stage 3D object detection
di: Noh, Jongyoun, et al.
Pubblicazione: (2025) -
LLAMADRS: Evaluating Open-Source LLMs on Real Clinical Interviews--To Reason or Not to Reason?
di: Kebe, Gaoussou Youssouf, et al.
Pubblicazione: (2025)