Supporting Experts with a Multimodal Machine-Learning-Based Tool for Human Behavior Analysis of Conversational Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Arakawa, Riku, Maeda, Kiyosu, Yakura, Hiromu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Coaching Copilot: Blended Form of an LLM-Powered Chatbot and a Human Coach to Effectively Support Self-Reflection for Leadership Growth
by: Arakawa, Riku, et al.
Published: (2024)
by: Arakawa, Riku, et al.
Published: (2024)
PrISM-Observer: Intervention Agent to Help Users Perform Everyday Procedures Sensed using a Smartwatch
by: Arakawa, Riku, et al.
Published: (2024)
by: Arakawa, Riku, et al.
Published: (2024)
Evaluating Large Language Models' Ability Using a Psychiatric Screening Tool Based on Metaphor and Sarcasm Scenarios
by: Yakura, Hiromu
Published: (2023)
by: Yakura, Hiromu
Published: (2023)
HiFiGaze: Improving Eye Tracking Accuracy Using Screen Content Knowledge
by: Kim, Taejun, et al.
Published: (2026)
by: Kim, Taejun, et al.
Published: (2026)
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
by: Ozdel, Suleyman, et al.
Published: (2024)
by: Ozdel, Suleyman, et al.
Published: (2024)
Biometrics and Behavior Analysis for Detecting Distractions in e-Learning
by: Becerra, Álvaro, et al.
Published: (2024)
by: Becerra, Álvaro, et al.
Published: (2024)
Continuous Human Action Recognition for Human-Machine Interaction: A Review
by: Gammulle, Harshala, et al.
Published: (2022)
by: Gammulle, Harshala, et al.
Published: (2022)
Surgment: Segmentation-enabled Semantic Search and Creation of Visual Question and Feedback to Support Video-Based Surgery Learning
by: Wang, Jingying, et al.
Published: (2024)
by: Wang, Jingying, et al.
Published: (2024)
Multimodal Ambivalence/Hesitancy Recognition in Videos for Personalized Digital Health Interventions
by: González-González, Manuela, et al.
Published: (2026)
by: González-González, Manuela, et al.
Published: (2026)
Accessible, At-Home Detection of Parkinson's Disease via Multi-task Video Analysis
by: Islam, Md Saiful, et al.
Published: (2024)
by: Islam, Md Saiful, et al.
Published: (2024)
ICo3D: An Interactive Conversational 3D Virtual Human
by: Shaw, Richard, et al.
Published: (2026)
by: Shaw, Richard, et al.
Published: (2026)
VRMN-bD: A Multi-modal Natural Behavior Dataset of Immersive Human Fear Responses in VR Stand-up Interactive Games
by: Zhang, He, et al.
Published: (2024)
by: Zhang, He, et al.
Published: (2024)
A Deep Learning-based Multimodal Depth-Aware Dynamic Hand Gesture Recognition System
by: Mahmud, Hasan, et al.
Published: (2021)
by: Mahmud, Hasan, et al.
Published: (2021)
Improve accessibility for Low Vision and Blind people using Machine Learning and Computer Vision
by: Shukurov, Jasur
Published: (2024)
by: Shukurov, Jasur
Published: (2024)
Machine Learning-Based Jamun Leaf Disease Detection: A Comprehensive Review
by: Bhowmik, Auvick Chandra, et al.
Published: (2023)
by: Bhowmik, Auvick Chandra, et al.
Published: (2023)
VideoMap: Supporting Video Editing Exploration, Brainstorming, and Prototyping in the Latent Space
by: Lin, David Chuan-En, et al.
Published: (2022)
by: Lin, David Chuan-En, et al.
Published: (2022)
Benchmarking Early Agitation Prediction in Community-Dwelling People with Dementia Using Multimodal Sensors and Machine Learning
by: Abedi, Ali, et al.
Published: (2025)
by: Abedi, Ali, et al.
Published: (2025)
VAAD: Visual Attention Analysis Dashboard applied to e-Learning
by: Navarro, Miriam, et al.
Published: (2024)
by: Navarro, Miriam, et al.
Published: (2024)
VideoMix: Aggregating How-To Videos for Task-Oriented Learning
by: Yang, Saelyne, et al.
Published: (2025)
by: Yang, Saelyne, et al.
Published: (2025)
IVISIT: An Interactive Visual Simulation Tool for system simulation, visualization, optimization, and parameter management
by: Knoblauch, Andreas
Published: (2024)
by: Knoblauch, Andreas
Published: (2024)
On the Interpretability of Part-Prototype Based Classifiers: A Human Centric Analysis
by: Davoodi, Omid, et al.
Published: (2023)
by: Davoodi, Omid, et al.
Published: (2023)
Toward Human-Robot Teaming: Learning Handover Behaviors from 3D Scenes
by: Wu, Yuekun, et al.
Published: (2025)
by: Wu, Yuekun, et al.
Published: (2025)
Behavioral Engagement in VR-Based Sign Language Learning: Visual Attention as a Predictor of Performance and Temporal Dynamics
by: Traini, Davide, et al.
Published: (2026)
by: Traini, Davide, et al.
Published: (2026)
A Comparison of Human and Machine Learning Errors in Face Recognition
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
Hybrid 3D Human Pose Estimation with Monocular Video and Sparse IMUs
by: Bao, Yiming, et al.
Published: (2024)
by: Bao, Yiming, et al.
Published: (2024)
Acoustic Field Video for Multimodal Scene Understanding
by: Kim, Daehwa, et al.
Published: (2026)
by: Kim, Daehwa, et al.
Published: (2026)
EduGage: Methods and Dataset for Sensor-Based Momentary Assessment of Engagement in Self-Guided Video Learning
by: Leng, Zikang, et al.
Published: (2026)
by: Leng, Zikang, et al.
Published: (2026)
A Review of Driver Gaze Estimation and Application in Gaze Behavior Understanding
by: Sharma, Pavan Kumar, et al.
Published: (2023)
by: Sharma, Pavan Kumar, et al.
Published: (2023)
Privacy-Preserving Empathy Detection in Video Interactions
by: Hasan, Md Rakibul, et al.
Published: (2025)
by: Hasan, Md Rakibul, et al.
Published: (2025)
Deep Learning Based Approach to Enhanced Recognition of Emotions and Behavioral Patterns of Autistic Children
by: R, Nelaka K. A., et al.
Published: (2025)
by: R, Nelaka K. A., et al.
Published: (2025)
L-WISE: Boosting Human Visual Category Learning Through Model-Based Image Selection and Enhancement
by: Talbot, Morgan B., et al.
Published: (2024)
by: Talbot, Morgan B., et al.
Published: (2024)
Random Channel Ablation for Robust Hand Gesture Classification with Multimodal Biosignals
by: Bimbraw, Keshav, et al.
Published: (2024)
by: Bimbraw, Keshav, et al.
Published: (2024)
A Rigorous Behavior Assessment of CNNs Using a Data-Domain Sampling Regime
by: Jiang, Shuning, et al.
Published: (2025)
by: Jiang, Shuning, et al.
Published: (2025)
VFA: Vision Frequency Analysis of Foundation Models and Human
by: Darvishi-Bayazi, Mohammad-Javad, et al.
Published: (2024)
by: Darvishi-Bayazi, Mohammad-Javad, et al.
Published: (2024)
An Evaluation of Hybrid Annotation Workflows on High-Ambiguity Spatiotemporal Video Footage
by: Gutiérrez, Juan, et al.
Published: (2025)
by: Gutiérrez, Juan, et al.
Published: (2025)
User Experience Estimation in Human-Robot Interaction Via Multi-Instance Learning of Multimodal Social Signals
by: Miyoshi, Ryo, et al.
Published: (2025)
by: Miyoshi, Ryo, et al.
Published: (2025)
ADAS-TO: A Large-Scale Multimodal Naturalistic Dataset and Empirical Characterization of Human Takeovers during ADAS Engagement
by: Wang, Yuhang, et al.
Published: (2026)
by: Wang, Yuhang, et al.
Published: (2026)
Machine Vision-Based Surgical Lighting System:Design and Implementation
by: Gharghabi, Amir, et al.
Published: (2025)
by: Gharghabi, Amir, et al.
Published: (2025)
Evaluating the Evaluators: Towards Human-aligned Metrics for Missing Markers Reconstruction
by: Kucherenko, Taras, et al.
Published: (2024)
by: Kucherenko, Taras, et al.
Published: (2024)
Human Motion Synthesis_ A Diffusion Approach for Motion Stitching and In-Betweening
by: Adewole, Michael, et al.
Published: (2024)
by: Adewole, Michael, et al.
Published: (2024)
Similar Items
-
Coaching Copilot: Blended Form of an LLM-Powered Chatbot and a Human Coach to Effectively Support Self-Reflection for Leadership Growth
by: Arakawa, Riku, et al.
Published: (2024) -
PrISM-Observer: Intervention Agent to Help Users Perform Everyday Procedures Sensed using a Smartwatch
by: Arakawa, Riku, et al.
Published: (2024) -
Evaluating Large Language Models' Ability Using a Psychiatric Screening Tool Based on Metaphor and Sarcasm Scenarios
by: Yakura, Hiromu
Published: (2023) -
HiFiGaze: Improving Eye Tracking Accuracy Using Screen Content Knowledge
by: Kim, Taejun, et al.
Published: (2026) -
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
by: Ozdel, Suleyman, et al.
Published: (2024)