Multimodal Methods for Analyzing Learning and Training Environments: A Systematic Literature Review
Fuente:
arXiv
Saved in:
| Main Authors: | Cohn, Clayton, Davalos, Eduardo, Vatral, Caleb, Fonteles, Joyce Horn, Wang, Hanchen David, Coursey, Austin, Rayala, Surya, S, Ashwin T, Ma, Meiyi, Biswas, Gautam |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A multimodal approach to support teacher, researcher and AI collaboration in STEM +C learning environments
by: Clayton Cohn, et al.
Published: (2024)
by: Clayton Cohn, et al.
Published: (2024)
A Theory of Adaptive Scaffolding for LLM-Based Pedagogical Agents
by: Cohn, Clayton, et al.
Published: (2025)
by: Cohn, Clayton, et al.
Published: (2025)
Analyzing Recursiveness in Multimodal Generative Artificial Intelligence: Stability or Divergence?
by: Conde, Javier, et al.
Published: (2024)
by: Conde, Javier, et al.
Published: (2024)
OpenVNA: A Framework for Analyzing the Behavior of Multimodal Language Understanding System under Noisy Scenarios
by: Yuan, Ziqi, et al.
Published: (2024)
by: Yuan, Ziqi, et al.
Published: (2024)
Personalizing Student-Agent Interactions Using Log-Contextualized Retrieval-Augmented Generation (RAG)
by: Cohn, Clayton, et al.
Published: (2025)
by: Cohn, Clayton, et al.
Published: (2025)
FLARE: Full-Modality Long-Video Audiovisual Retrieval Benchmark with User-Simulated Queries
by: You, Qijie, et al.
Published: (2026)
by: You, Qijie, et al.
Published: (2026)
BEAGLE: Behavior-Enforced Agent for Grounded Learner Emulation
by: Wang, Hanchen David, et al.
Published: (2026)
by: Wang, Hanchen David, et al.
Published: (2026)
CoTAL: Human-in-the-Loop Prompt Engineering for Generalizable Formative Assessment Scoring
by: Cohn, Clayton, et al.
Published: (2025)
by: Cohn, Clayton, et al.
Published: (2025)
Video-Based Performance Evaluation for ECR Drills in Synthetic Training Environments
by: Rayala, Surya, et al.
Published: (2025)
by: Rayala, Surya, et al.
Published: (2025)
Loss-tolerant neural video codec aware congestion control for real time video communication
by: Xia, Zhengxu, et al.
Published: (2024)
by: Xia, Zhengxu, et al.
Published: (2024)
3D Gaze Tracking for Studying Collaborative Interactions in Mixed-Reality Environments
by: Davalos, Eduardo, et al.
Published: (2024)
by: Davalos, Eduardo, et al.
Published: (2024)
AI-Assisted Competency Assessment from Egocentric Video in Simulation-Based Nursing Education
by: Wang, Hanchen David, et al.
Published: (2026)
by: Wang, Hanchen David, et al.
Published: (2026)
On the Design of Safe Continual RL Methods for Control of Nonlinear Systems
by: Coursey, Austin, et al.
Published: (2025)
by: Coursey, Austin, et al.
Published: (2025)
Multimodal Classification and Out-of-distribution Detection for Multimodal Intent Understanding
by: Zhang, Hanlei, et al.
Published: (2024)
by: Zhang, Hanlei, et al.
Published: (2024)
Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions
by: Wang, Tianshi, et al.
Published: (2023)
by: Wang, Tianshi, et al.
Published: (2023)
GalleryGPT: Analyzing Paintings with Large Multimodal Models
by: Bin, Yi, et al.
Published: (2024)
by: Bin, Yi, et al.
Published: (2024)
Dark Side of Modalities: Reinforced Multimodal Distillation for Multimodal Knowledge Graph Reasoning
by: Zhao, Yu, et al.
Published: (2025)
by: Zhao, Yu, et al.
Published: (2025)
MInD: Improving Multimodal Sentiment Analysis via Multimodal Information Disentanglement
by: Dai, Weichen, et al.
Published: (2024)
by: Dai, Weichen, et al.
Published: (2024)
Hyperbolic Multimodal Generative Representation Learning for Generalized Zero-Shot Multimodal Information Extraction
by: Zhou, Baohang, et al.
Published: (2026)
by: Zhou, Baohang, et al.
Published: (2026)
Multimodal Graph-Based Variational Mixture of Experts Network for Zero-Shot Multimodal Information Extraction
by: Zhou, Baohang, et al.
Published: (2025)
by: Zhou, Baohang, et al.
Published: (2025)
MM-InstructEval: Zero-Shot Evaluation of (Multimodal) Large Language Models on Multimodal Reasoning Tasks
by: Yang, Xiaocui, et al.
Published: (2024)
by: Yang, Xiaocui, et al.
Published: (2024)
Subjective and Objective Quality Assessment Methods of Stereoscopic Videos with Visibility Affecting Distortions
by: Biswas, Sria, et al.
Published: (2024)
by: Biswas, Sria, et al.
Published: (2024)
From Multimodal Signals to Adaptive XR Experiences for De-escalation Training
by: Nierula, Birgit, et al.
Published: (2026)
by: Nierula, Birgit, et al.
Published: (2026)
Exploring the Role of Audio in Multimodal Misinformation Detection
by: Liu, Moyang, et al.
Published: (2024)
by: Liu, Moyang, et al.
Published: (2024)
Towards Multimodal Emotional Support Conversation Systems
by: Chu, Yuqi, et al.
Published: (2024)
by: Chu, Yuqi, et al.
Published: (2024)
Multimodal Emotion Recognition with Large Language Models
by: Zhang, Hongrui, et al.
Published: (2026)
by: Zhang, Hongrui, et al.
Published: (2026)
An Efficient Digital Watermarking Technique for Small Scale devices
by: Talathi, Kaushik, et al.
Published: (2025)
by: Talathi, Kaushik, et al.
Published: (2025)
Contrastive Knowledge Distillation for Robust Multimodal Sentiment Analysis
by: Sang, Zhongyi, et al.
Published: (2024)
by: Sang, Zhongyi, et al.
Published: (2024)
Multimodal LLM-based Query Paraphrasing for Video Search
by: Wu, Jiaxin, et al.
Published: (2024)
by: Wu, Jiaxin, et al.
Published: (2024)
From Natural Alignment to Conditional Controllability in Multimodal Dialogue
by: Jin, Zeyu, et al.
Published: (2026)
by: Jin, Zeyu, et al.
Published: (2026)
Training Data Efficiency in Multimodal Process Reward Models
by: Li, Jinyuan, et al.
Published: (2026)
by: Li, Jinyuan, et al.
Published: (2026)
A Multimodal Framework for Explainable Evaluation of Soft Skills in Educational Environments
by: Guerrero-Sosa, Jared D. T., et al.
Published: (2025)
by: Guerrero-Sosa, Jared D. T., et al.
Published: (2025)
Subjective Quality Assessment of Dynamic 3D Meshes in Virtual Reality Environment
by: Nguyen, Duc V., et al.
Published: (2026)
by: Nguyen, Duc V., et al.
Published: (2026)
Virbo: Multimodal Multilingual Avatar Video Generation in Digital Marketing
by: Zhang, Juan, et al.
Published: (2024)
by: Zhang, Juan, et al.
Published: (2024)
Angle-Optimized Partial Disentanglement for Multimodal Emotion Recognition in Conversation
by: Che, Xinyi, et al.
Published: (2025)
by: Che, Xinyi, et al.
Published: (2025)
Learning Shared Sentiment Prototypes for Adaptive Multimodal Sentiment Analysis
by: Su, Chen, et al.
Published: (2026)
by: Su, Chen, et al.
Published: (2026)
Retrieval Augmented Verification for Zero-Shot Detection of Multimodal Disinformation
by: Dey, Arka Ujjal, et al.
Published: (2024)
by: Dey, Arka Ujjal, et al.
Published: (2024)
Differential Mental Disorder Detection with Psychology-Inspired Multimodal Stimuli
by: Zhou, Zhiyuan, et al.
Published: (2026)
by: Zhou, Zhiyuan, et al.
Published: (2026)
EmpathyEar: An Open-source Avatar Multimodal Empathetic Chatbot
by: Fei, Hao, et al.
Published: (2024)
by: Fei, Hao, et al.
Published: (2024)
Multimodal Semantic Communication for Generative Audio-Driven Video Conferencing
by: Tong, Haonan, et al.
Published: (2024)
by: Tong, Haonan, et al.
Published: (2024)
Similar Items
-
A multimodal approach to support teacher, researcher and AI collaboration in STEM +C learning environments
by: Clayton Cohn, et al.
Published: (2024) -
A Theory of Adaptive Scaffolding for LLM-Based Pedagogical Agents
by: Cohn, Clayton, et al.
Published: (2025) -
Analyzing Recursiveness in Multimodal Generative Artificial Intelligence: Stability or Divergence?
by: Conde, Javier, et al.
Published: (2024) -
OpenVNA: A Framework for Analyzing the Behavior of Multimodal Language Understanding System under Noisy Scenarios
by: Yuan, Ziqi, et al.
Published: (2024) -
Personalizing Student-Agent Interactions Using Log-Contextualized Retrieval-Augmented Generation (RAG)
by: Cohn, Clayton, et al.
Published: (2025)