Three-Class Emotion Classification for Audiovisual Scenes Based on Ensemble Learning Scheme
Fuente:
arXiv
Saved in:
| Main Authors: | Xiong, Xiangrui, Zhou, Zhou, Nong, Guocai, Deng, Junlin, Wu, Ning |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LightBeam: An Accurate and Memory-Efficient CTC Decoder for Speech Neuroprostheses
by: Feghhi, Ebrahim, et al.
Published: (2026)
by: Feghhi, Ebrahim, et al.
Published: (2026)
A Semantically Consistent Dataset for Data-Efficient Query-Based Universal Sound Separation
by: Li, Kai, et al.
Published: (2026)
by: Li, Kai, et al.
Published: (2026)
EarResp-ANS : Audio-Based On-Device Respiration Rate Estimation on Earphones with Adaptive Noise Suppression
by: Küttner, Michael, et al.
Published: (2026)
by: Küttner, Michael, et al.
Published: (2026)
Bridging the Gap between Micro-scale Traffic Simulation and 4D Digital Cityscapes
by: Jiao, Longxiang, et al.
Published: (2026)
by: Jiao, Longxiang, et al.
Published: (2026)
Towards LLM-Empowered Fine-Grained Speech Descriptors for Explainable Emotion Recognition
by: Chen, Youjun, et al.
Published: (2025)
by: Chen, Youjun, et al.
Published: (2025)
FireRedChat: A Pluggable, Full-Duplex Voice Interaction System with Cascaded and Semi-Cascaded Implementations
by: Chen, Junjie, et al.
Published: (2025)
by: Chen, Junjie, et al.
Published: (2025)
EmoAugNet: A Signal-Augmented Hybrid CNN-LSTM Framework for Speech Emotion Recognition
by: Paul, Durjoy Chandra, et al.
Published: (2025)
by: Paul, Durjoy Chandra, et al.
Published: (2025)
A Theory-Based Explainable Deep Learning Architecture for Music Emotion
by: Fong, Hortense, et al.
Published: (2024)
by: Fong, Hortense, et al.
Published: (2024)
Improved Dysarthric Speech to Text Conversion via TTS Personalization
by: Mihajlik, Péter, et al.
Published: (2025)
by: Mihajlik, Péter, et al.
Published: (2025)
SpeechCompass: Enhancing Mobile Captioning with Diarization and Directional Guidance via Multi-Microphone Localization
by: Dementyev, Artem, et al.
Published: (2025)
by: Dementyev, Artem, et al.
Published: (2025)
Effect of Avatar Head Movement on Communication Behaviour, Experience of Presence and Conversation Success in Triadic Conversations
by: Kothe, Angelika, et al.
Published: (2025)
by: Kothe, Angelika, et al.
Published: (2025)
AuthGlass: Benchmarking Voice Liveness Detection and Authentication on Smart Glasses via Comprehensive Acoustic Features
by: Xu, Weiye, et al.
Published: (2025)
by: Xu, Weiye, et al.
Published: (2025)
Erie: A Declarative Grammar for Data Sonification
by: Kim, Hyeok, et al.
Published: (2024)
by: Kim, Hyeok, et al.
Published: (2024)
Accessible Fine-grained Data Representation via Spatial Audio
by: Liu, Can, et al.
Published: (2026)
by: Liu, Can, et al.
Published: (2026)
FlueBricks: A Construction Kit of Flute-like Instruments for Acoustic Reasoning
by: Chen, Bo-Yu, et al.
Published: (2026)
by: Chen, Bo-Yu, et al.
Published: (2026)
Opening the Design Space: Two Years of Performance with Intelligent Musical Instruments
by: Martin, Charles Patrick
Published: (2026)
by: Martin, Charles Patrick
Published: (2026)
Sound-Based Recognition of Touch Gestures and Emotions for Enhanced Human-Robot Interaction
by: Hou, Yuanbo, et al.
Published: (2024)
by: Hou, Yuanbo, et al.
Published: (2024)
Improving Multimodal Emotion Recognition by Leveraging Acoustic Adaptation and Visual Alignment
by: Zhao, Zhixian, et al.
Published: (2024)
by: Zhao, Zhixian, et al.
Published: (2024)
Revisiting Your Memory: Reconstruction of Affect-Contextualized Memory via EEG-guided Audiovisual Generation
by: Kwon, Joonwoo, et al.
Published: (2024)
by: Kwon, Joonwoo, et al.
Published: (2024)
Advancing User-Voice Interaction: Exploring Emotion-Aware Voice Assistants Through a Role-Swapping Approach
by: Ma, Yong, et al.
Published: (2025)
by: Ma, Yong, et al.
Published: (2025)
Are Expressions for Music Emotions the Same Across Cultures?
by: Celen, Elif, et al.
Published: (2025)
by: Celen, Elif, et al.
Published: (2025)
SCDiar: a streaming diarization system based on speaker change detection and speech recognition
by: Zheng, Naijun, et al.
Published: (2025)
by: Zheng, Naijun, et al.
Published: (2025)
Lla-VAP: LSTM Ensemble of Llama and VAP for Turn-Taking Prediction
by: Jeon, Hyunbae, et al.
Published: (2024)
by: Jeon, Hyunbae, et al.
Published: (2024)
LoopLens: Supporting Search as Creation in Loop-Based Music Composition
by: Long, Sheng, et al.
Published: (2026)
by: Long, Sheng, et al.
Published: (2026)
Sound Clouds: Exploring ambient intelligence in public spaces to elicit deep human experience of awe, wonder, and beauty
by: Zhang, Chengzhi, et al.
Published: (2025)
by: Zhang, Chengzhi, et al.
Published: (2025)
TalkSketch: Multimodal Generative AI for Real-time Sketch Ideation with Speech
by: Shi, Weiyan, et al.
Published: (2025)
by: Shi, Weiyan, et al.
Published: (2025)
Semi-Automatic Flute Robot and Its Acoustic Sensing
by: Kuriyama, Hikari, et al.
Published: (2026)
by: Kuriyama, Hikari, et al.
Published: (2026)
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers
by: Mishra, Ruchik, et al.
Published: (2024)
by: Mishra, Ruchik, et al.
Published: (2024)
An Intelligent AI glasses System with Multi-Agent Architecture for Real-Time Voice Processing and Task Execution
by: Chen, Sheng-Kai, et al.
Published: (2026)
by: Chen, Sheng-Kai, et al.
Published: (2026)
AVE Speech: A Comprehensive Multi-Modal Dataset for Speech Recognition Integrating Audio, Visual, and Electromyographic Signals
by: Zhou, Dongliang, et al.
Published: (2025)
by: Zhou, Dongliang, et al.
Published: (2025)
Beyond Acoustic Emotion Recognition: Multimodal Pathos Analysis in Political Speech Using LLM-Based and Acoustic Emotion Models
by: Dietrich, Juergen
Published: (2026)
by: Dietrich, Juergen
Published: (2026)
ListenNet: A Lightweight Spatio-Temporal Enhancement Nested Network for Auditory Attention Detection
by: Fan, Cunhang, et al.
Published: (2025)
by: Fan, Cunhang, et al.
Published: (2025)
MHANet: Multi-scale Hybrid Attention Network for Auditory Attention Detection
by: Li, Lu, et al.
Published: (2025)
by: Li, Lu, et al.
Published: (2025)
Exploring Gender Bias in Alzheimer's Disease Detection: Insights from Mandarin and Greek Speech Perception
by: He, Liu, et al.
Published: (2025)
by: He, Liu, et al.
Published: (2025)
Multilingual and Continuous Backchannel Prediction: A Cross-lingual Study
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
Evaluating the Impact of AI-Powered Audiovisual Personalization on Learner Emotion, Focus, and Learning Outcomes
by: Wang, George Xi, et al.
Published: (2025)
by: Wang, George Xi, et al.
Published: (2025)
Optimizing Multilingual Text-To-Speech with Accents & Emotions
by: Pawar, Pranav, et al.
Published: (2025)
by: Pawar, Pranav, et al.
Published: (2025)
Human Feedback Driven Dynamic Speech Emotion Recognition
by: Fedorov, Ilya, et al.
Published: (2025)
by: Fedorov, Ilya, et al.
Published: (2025)
RespEar: Earable-Based Robust Respiratory Rate Monitoring
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
Cervical Auscultation Machine Learning for Dysphagia Assessment
by: Chia, An An, et al.
Published: (2024)
by: Chia, An An, et al.
Published: (2024)
Similar Items
-
LightBeam: An Accurate and Memory-Efficient CTC Decoder for Speech Neuroprostheses
by: Feghhi, Ebrahim, et al.
Published: (2026) -
A Semantically Consistent Dataset for Data-Efficient Query-Based Universal Sound Separation
by: Li, Kai, et al.
Published: (2026) -
EarResp-ANS : Audio-Based On-Device Respiration Rate Estimation on Earphones with Adaptive Noise Suppression
by: Küttner, Michael, et al.
Published: (2026) -
Bridging the Gap between Micro-scale Traffic Simulation and 4D Digital Cityscapes
by: Jiao, Longxiang, et al.
Published: (2026) -
Towards LLM-Empowered Fine-Grained Speech Descriptors for Explainable Emotion Recognition
by: Chen, Youjun, et al.
Published: (2025)