Guardado en:
| Autores principales: | Zhou, Emily, Ma, Marcus, Avramidis, Kleanthis, Toth, Gabor Mihaly, Narayanan, Shrikanth |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2604.22016 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Encoding Emotion Through Self-Supervised Eye Movement Reconstruction
por: Ma, Marcus, et al.
Publicado: (2026)
por: Ma, Marcus, et al.
Publicado: (2026)
Smiling Regulates Emotion During Traumatic Recollection
por: Ma, Marcus, et al.
Publicado: (2026)
por: Ma, Marcus, et al.
Publicado: (2026)
Emotion-Aligned Contrastive Learning Between Images and Music
por: Stewart, Shanti, et al.
Publicado: (2023)
por: Stewart, Shanti, et al.
Publicado: (2023)
Early Detection of Coffee Leaf Rust Through Convolutional Neural Networks Trained on Low-Resolution Images
por: Cabrera, Angelly, et al.
Publicado: (2024)
por: Cabrera, Angelly, et al.
Publicado: (2024)
Knowledge-guided EEG Representation Learning
por: Kommineni, Aditya, et al.
Publicado: (2024)
por: Kommineni, Aditya, et al.
Publicado: (2024)
An Emotion Recognition Framework via Cross-modal Alignment of EEG and Eye Movement Data
por: Wang, Jianlu, et al.
Publicado: (2025)
por: Wang, Jianlu, et al.
Publicado: (2025)
VoxEmo: Benchmarking Speech Emotion Recognition with Speech LLMs
por: Zhang, Hezhao, et al.
Publicado: (2026)
por: Zhang, Hezhao, et al.
Publicado: (2026)
Look, Listen and Segment: Towards Weakly Supervised Audio-visual Semantic Segmentation
por: Li, Chengzhi, et al.
Publicado: (2026)
por: Li, Chengzhi, et al.
Publicado: (2026)
Evaluating Atypical Gaze Patterns through Vision Models: The Case of Cortical Visual Impairment
por: Avramidis, Kleanthis, et al.
Publicado: (2024)
por: Avramidis, Kleanthis, et al.
Publicado: (2024)
Deep Learning Characterizes Depression and Suicidal Ideation from Eye Movements
por: Avramidis, Kleanthis, et al.
Publicado: (2025)
por: Avramidis, Kleanthis, et al.
Publicado: (2025)
Informed Bootstrap Augmentation Improves EEG Decoding
por: Jeong, Woojae, et al.
Publicado: (2025)
por: Jeong, Woojae, et al.
Publicado: (2025)
Movement- and Traffic-based User Identification in Commercial Virtual Reality Applications: Threats and Opportunities
por: Baldoni, Sara, et al.
Publicado: (2025)
por: Baldoni, Sara, et al.
Publicado: (2025)
Listening to the Unspoken: Exploring "365" Aspects of Multimodal Interview Performance Assessment
por: Li, Jia, et al.
Publicado: (2025)
por: Li, Jia, et al.
Publicado: (2025)
Remember Past, Anticipate Future: Learning Continual Multimodal Misinformation Detectors
por: Wang, Bing, et al.
Publicado: (2025)
por: Wang, Bing, et al.
Publicado: (2025)
Archiving Body Movements: Collective Generation of Chinese Calligraphy
por: Zhou, Aven Le, et al.
Publicado: (2023)
por: Zhou, Aven Le, et al.
Publicado: (2023)
VoxCare: Studying Natural Communication Behaviors of Hospital Caregivers through Wearable Sensing of Egocentric Audio
por: Feng, Tiantian, et al.
Publicado: (2026)
por: Feng, Tiantian, et al.
Publicado: (2026)
Editing on the Generative Manifold: A Theoretical and Empirical Study of General Diffusion-Based Image Editing Trade-offs
por: Hu, Yi, et al.
Publicado: (2026)
por: Hu, Yi, et al.
Publicado: (2026)
DanceCamera3D: 3D Camera Movement Synthesis with Music and Dance
por: Wang, Zixuan, et al.
Publicado: (2024)
por: Wang, Zixuan, et al.
Publicado: (2024)
Conformer-based Ultrasound-to-Speech Conversion
por: Ibrahimov, Ibrahim, et al.
Publicado: (2025)
por: Ibrahimov, Ibrahim, et al.
Publicado: (2025)
Listen, Look, Drive: Coupling Audio Instructions for User-aware VLA-based Autonomous Driving
por: Guo, Ziang, et al.
Publicado: (2026)
por: Guo, Ziang, et al.
Publicado: (2026)
Characterizing Multimedia Information Environment through Multi-modal Clustering of YouTube Videos
por: Yousefi, Niloofar, et al.
Publicado: (2024)
por: Yousefi, Niloofar, et al.
Publicado: (2024)
Modular Conversational Agents for Surveys and Interviews
por: Yu, Jiangbo, et al.
Publicado: (2024)
por: Yu, Jiangbo, et al.
Publicado: (2024)
Through Their Eyes: Fixation-aligned Tuning for Personalized User Emulation
por: Huang, Lingfeng, et al.
Publicado: (2026)
por: Huang, Lingfeng, et al.
Publicado: (2026)
Toward Fully-End-to-End Listened Speech Decoding from EEG Signals
por: Lee, Jihwan, et al.
Publicado: (2024)
por: Lee, Jihwan, et al.
Publicado: (2024)
Looking Backward: Streaming Video-to-Video Translation with Feature Banks
por: Liang, Feng, et al.
Publicado: (2024)
por: Liang, Feng, et al.
Publicado: (2024)
Unravelling the Power of Single-Pass Look-Ahead in Modern Codecs for Optimized Transcoding Deployment
por: Vibhoothi, Vibhoothi, et al.
Publicado: (2024)
por: Vibhoothi, Vibhoothi, et al.
Publicado: (2024)
SimInterview: Transforming Business Education through Large Language Model-Based Simulated Multilingual Interview Training System
por: Nguyen, Truong Thanh Hung, et al.
Publicado: (2025)
por: Nguyen, Truong Thanh Hung, et al.
Publicado: (2025)
Applying LLM-Powered Virtual Humans to Child Interviews in Child-Centered Design
por: Li, Linshi, et al.
Publicado: (2025)
por: Li, Linshi, et al.
Publicado: (2025)
Look, Compare and Draw: Differential Query Transformer for Automatic Oil Painting
por: Liu, Lingyu, et al.
Publicado: (2026)
por: Liu, Lingyu, et al.
Publicado: (2026)
Unraveling Instance Associations: A Closer Look for Audio-Visual Segmentation
por: Chen, Yuanhong, et al.
Publicado: (2023)
por: Chen, Yuanhong, et al.
Publicado: (2023)
State-Anchored Complete-View Distillation for Robust Conversational Multimodal Emotion Recognition
por: Pan, Zhaoyan, et al.
Publicado: (2026)
por: Pan, Zhaoyan, et al.
Publicado: (2026)
EyeNexus: Adaptive Gaze-Driven Quality and Bitrate Streaming for Seamless VR Cloud Gaming Experiences
por: Wu, Ze, et al.
Publicado: (2025)
por: Wu, Ze, et al.
Publicado: (2025)
Can Video Diffusion Models Predict Past Frames? Bidirectional Cycle Consistency for Reversible Interpolation
por: Liu, Lingyu, et al.
Publicado: (2026)
por: Liu, Lingyu, et al.
Publicado: (2026)
Panonut360: A Head and Eye Tracking Dataset for Panoramic Video
por: Xu, Yutong, et al.
Publicado: (2024)
por: Xu, Yutong, et al.
Publicado: (2024)
FineBadminton: A Multi-Level Dataset for Fine-Grained Badminton Video Understanding
por: He, Xusheng, et al.
Publicado: (2025)
por: He, Xusheng, et al.
Publicado: (2025)
Augment Before Copy-Paste: Data and Memory Efficiency-Oriented Instance Segmentation Framework for Sport-scenes
por: Hsu, Chih-Chung, et al.
Publicado: (2024)
por: Hsu, Chih-Chung, et al.
Publicado: (2024)
Real-time 3D Light-field Viewing with Eye-tracking on Conventional Displays
por: Pham, Trung Hieu, et al.
Publicado: (2025)
por: Pham, Trung Hieu, et al.
Publicado: (2025)
Subjective Evaluation of Frame Rate in Bitrate-Constrained Live Streaming
por: He, Jiaqi, et al.
Publicado: (2026)
por: He, Jiaqi, et al.
Publicado: (2026)
Relationship Analysis of Image-Text Pair in SNS Posts
por: Nabeoka, Takuto, et al.
Publicado: (2025)
por: Nabeoka, Takuto, et al.
Publicado: (2025)
Promisedland: An XR Narrative Attraction Integrating Diorama-to-Virtual Workflow and Elemental Storytelling
por: Wang, Xianghan, et al.
Publicado: (2025)
por: Wang, Xianghan, et al.
Publicado: (2025)
Ejemplares similares
-
Encoding Emotion Through Self-Supervised Eye Movement Reconstruction
por: Ma, Marcus, et al.
Publicado: (2026) -
Smiling Regulates Emotion During Traumatic Recollection
por: Ma, Marcus, et al.
Publicado: (2026) -
Emotion-Aligned Contrastive Learning Between Images and Music
por: Stewart, Shanti, et al.
Publicado: (2023) -
Early Detection of Coffee Leaf Rust Through Convolutional Neural Networks Trained on Low-Resolution Images
por: Cabrera, Angelly, et al.
Publicado: (2024) -
Knowledge-guided EEG Representation Learning
por: Kommineni, Aditya, et al.
Publicado: (2024)