Privileged Contrastive Pretraining for Multimodal Affect Modelling
Fuente:
arXiv
Salvato in:
| Autori principali: | Pinitas, Kosmas, Makantasis, Konstantinos, Yannakakis, Georgios N. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Across-Game Engagement Modelling via Few-Shot Learning
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
GameVibe: A Multimodal Affective Game Corpus
di: Barthet, Matthew, et al.
Pubblicazione: (2024)
di: Barthet, Matthew, et al.
Pubblicazione: (2024)
To Fuse or to Drop? Dual-Path Learning for Resolving Modality Conflicts in Multimodal Emotion Recognition
di: Yu, Yangchen, et al.
Pubblicazione: (2026)
di: Yu, Yangchen, et al.
Pubblicazione: (2026)
Multimodal Dialog Systems with Dual Knowledge-enhanced Generative Pretrained Language Model
di: Chen, Xiaolin, et al.
Pubblicazione: (2022)
di: Chen, Xiaolin, et al.
Pubblicazione: (2022)
Chain-of-Modality: Learning Manipulation Programs from Multimodal Human Videos with Vision-Language-Models
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
Multimodal Infusion Tuning for Large Models
di: Sun, Hao, et al.
Pubblicazione: (2024)
di: Sun, Hao, et al.
Pubblicazione: (2024)
Body Ownership Affects the Processing of Sensorimotor Contingencies in Virtual Reality
di: Center, Evan G., et al.
Pubblicazione: (2025)
di: Center, Evan G., et al.
Pubblicazione: (2025)
Language-Guided Multimodal Texture Authoring via Generative Models
di: Qian, Wanli, et al.
Pubblicazione: (2026)
di: Qian, Wanli, et al.
Pubblicazione: (2026)
AffectMachine-Pop: A controllable expert system for real-time pop music generation
di: Agres, Kat R., et al.
Pubblicazione: (2025)
di: Agres, Kat R., et al.
Pubblicazione: (2025)
From Perception to Cognition: How Latency Affects Interaction Fluency and Social Presence in VR Conferencing
di: Song, Jiarun, et al.
Pubblicazione: (2026)
di: Song, Jiarun, et al.
Pubblicazione: (2026)
Multimodal Fusion with Semi-Supervised Learning Minimizes Annotation Quantity for Modeling Videoconference Conversation Experience
di: Chang, Andrew, et al.
Pubblicazione: (2025)
di: Chang, Andrew, et al.
Pubblicazione: (2025)
MindCine: Multimodal EEG-to-Video Reconstruction with Large-Scale Pretrained Models
di: Zhou, Tian-Yi, et al.
Pubblicazione: (2026)
di: Zhou, Tian-Yi, et al.
Pubblicazione: (2026)
Explainable Multimodal Emotion Recognition
di: Lian, Zheng, et al.
Pubblicazione: (2023)
di: Lian, Zheng, et al.
Pubblicazione: (2023)
Human-Machine Ritual: Synergic Performance through Real-Time Motion Recognition
di: Cai, Zhuodi, et al.
Pubblicazione: (2025)
di: Cai, Zhuodi, et al.
Pubblicazione: (2025)
LibEMER: A novel benchmark and algorithms library for EEG-based Multimodal Emotion Recognition
di: Liu, Zejun, et al.
Pubblicazione: (2025)
di: Liu, Zejun, et al.
Pubblicazione: (2025)
From Multimodal Signals to Adaptive XR Experiences for De-escalation Training
di: Nierula, Birgit, et al.
Pubblicazione: (2026)
di: Nierula, Birgit, et al.
Pubblicazione: (2026)
Crafting Dynamic Virtual Activities with Advanced Multimodal Models
di: Li, Changyang, et al.
Pubblicazione: (2024)
di: Li, Changyang, et al.
Pubblicazione: (2024)
Towards Aligning Multimodal LLMs with Human Experts: A Focus on Parent-Child Interaction
di: Shi, Weiyan, et al.
Pubblicazione: (2025)
di: Shi, Weiyan, et al.
Pubblicazione: (2025)
When Drawing Is Not Enough: Exploring Spontaneous Speech with Sketch for Intent Alignment in Multimodal LLMs
di: Shi, Weiyan, et al.
Pubblicazione: (2026)
di: Shi, Weiyan, et al.
Pubblicazione: (2026)
MULTI-CASE: A Transformer-based Ethics-aware Multimodal Investigative Intelligence Framework
di: Fischer, Maximilian T., et al.
Pubblicazione: (2024)
di: Fischer, Maximilian T., et al.
Pubblicazione: (2024)
Dynamic Multimodal Expression Generation for LLM-Driven Pedagogical Agents: From User Experience Perspective
di: Wan, Ninghao, et al.
Pubblicazione: (2026)
di: Wan, Ninghao, et al.
Pubblicazione: (2026)
Memento: Augmenting Personalized Memory via Practical Multimodal Wearable Sensing in Visual Search and Wayfinding Navigation
di: Ghosh, Indrajeet, et al.
Pubblicazione: (2025)
di: Ghosh, Indrajeet, et al.
Pubblicazione: (2025)
Multimodal Digital Sensing of Early-Life Laying Hens: A Pilot Study Integrating Thermal, Acoustic, Optical-Flow and Environmental Data
di: Dhaliwal, Yashan, et al.
Pubblicazione: (2026)
di: Dhaliwal, Yashan, et al.
Pubblicazione: (2026)
Talking Slide Avatars: Open-Source Multimodal Communication Approach for Teaching
di: Wu, Xinxing
Pubblicazione: (2026)
di: Wu, Xinxing
Pubblicazione: (2026)
BioArtlas: Computational Clustering of Multi-Dimensional Complexity in Bioart
di: Bae, Joonhyung
Pubblicazione: (2025)
di: Bae, Joonhyung
Pubblicazione: (2025)
See, Think, Act: Online Shopper Behavior Simulation with VLM Agents
di: Zhang, Yimeng, et al.
Pubblicazione: (2025)
di: Zhang, Yimeng, et al.
Pubblicazione: (2025)
Towards Interactive Multimodal Representation of ML Functions for Human Understanding of ML
di: Wang, Bokang, et al.
Pubblicazione: (2026)
di: Wang, Bokang, et al.
Pubblicazione: (2026)
TalkSketch: Multimodal Generative AI for Real-time Sketch Ideation with Speech
di: Shi, Weiyan, et al.
Pubblicazione: (2025)
di: Shi, Weiyan, et al.
Pubblicazione: (2025)
3D Modelling to Address Pandemic Challenges: A Project-Based Learning Methodology
di: Rocha, Tânia, et al.
Pubblicazione: (2024)
di: Rocha, Tânia, et al.
Pubblicazione: (2024)
On Parallelism in Music and Language: A Perspective from Symbol Emergence Systems based on Probabilistic Generative Models
di: Taniguchi, Tadahiro
Pubblicazione: (2025)
di: Taniguchi, Tadahiro
Pubblicazione: (2025)
Human-Machine Collaboration-Guided Space Design: Combination of Machine Learning Models and Humanistic Design Concepts
di: Yang, Yuxuan
Pubblicazione: (2025)
di: Yang, Yuxuan
Pubblicazione: (2025)
Coordinated 2D-3D Visualization of Volumetric Medical Data in XR with Multimodal Interactions
di: Liu, Qixuan, et al.
Pubblicazione: (2025)
di: Liu, Qixuan, et al.
Pubblicazione: (2025)
Resource-Constrained Affect Modelling via Variance Regularisation Pruning
di: Pinitas, Kosmas, et al.
Pubblicazione: (2026)
di: Pinitas, Kosmas, et al.
Pubblicazione: (2026)
MultiVox: A Benchmark for Evaluating Voice Assistants for Multimodal Interactions
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2025)
di: Selvakumar, Ramaneswaran, et al.
Pubblicazione: (2025)
Enhanced Web User Interface Design Via Cross-Device Responsiveness Assessment Using An Improved HCI-INTEGRATED DL Schemes
di: Balasubramanian, Shrinivass Arunachalam
Pubblicazione: (2025)
di: Balasubramanian, Shrinivass Arunachalam
Pubblicazione: (2025)
Emotion-Driven Personalized Recommendation for AI-Generated Content Using Multi-Modal Sentiment and Intent Analysis
di: Hu, Zheqi, et al.
Pubblicazione: (2025)
di: Hu, Zheqi, et al.
Pubblicazione: (2025)
Workflow-Based Evaluation of Music Generation Systems
di: Dadman, Shayan, et al.
Pubblicazione: (2025)
di: Dadman, Shayan, et al.
Pubblicazione: (2025)
Evaluating gesture generation in a large-scale open challenge: The GENEA Challenge 2022
di: Kucherenko, Taras, et al.
Pubblicazione: (2023)
di: Kucherenko, Taras, et al.
Pubblicazione: (2023)
AI TrackMate: Finally, Someone Who Will Give Your Music More Than Just "Sounds Great!"
di: Jiang, Yi-Lin, et al.
Pubblicazione: (2024)
di: Jiang, Yi-Lin, et al.
Pubblicazione: (2024)
Closing the Loop: Unified 3D Scene Generation and Immersive Interaction via LLM-RL Coupling
di: Vo, Anh H., et al.
Pubblicazione: (2026)
di: Vo, Anh H., et al.
Pubblicazione: (2026)
Documenti analoghi
-
Across-Game Engagement Modelling via Few-Shot Learning
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024) -
GameVibe: A Multimodal Affective Game Corpus
di: Barthet, Matthew, et al.
Pubblicazione: (2024) -
To Fuse or to Drop? Dual-Path Learning for Resolving Modality Conflicts in Multimodal Emotion Recognition
di: Yu, Yangchen, et al.
Pubblicazione: (2026) -
Multimodal Dialog Systems with Dual Knowledge-enhanced Generative Pretrained Language Model
di: Chen, Xiaolin, et al.
Pubblicazione: (2022) -
Chain-of-Modality: Learning Manipulation Programs from Multimodal Human Videos with Vision-Language-Models
di: Wang, Chen, et al.
Pubblicazione: (2025)