Panonut360: A Head and Eye Tracking Dataset for Panoramic Video
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Yutong, Du, Junhao, Wang, Jiahe, Ning, Yuwei, Cao, Sihan Zhou Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024)
von: He, Xu, et al.
Veröffentlicht: (2024)
Editing Physiological Signals in Videos Using Latent Representations
von: Zhou, Tianwen, et al.
Veröffentlicht: (2025)
von: Zhou, Tianwen, et al.
Veröffentlicht: (2025)
EyeNavGS: A 6-DoF Navigation Dataset and Record-n-Replay Software for Real-World 3DGS Scenes in VR
von: Ding, Zihao, et al.
Veröffentlicht: (2025)
von: Ding, Zihao, et al.
Veröffentlicht: (2025)
SVFAP: Self-supervised Video Facial Affect Perceiver
von: Sun, Licai, et al.
Veröffentlicht: (2023)
von: Sun, Licai, et al.
Veröffentlicht: (2023)
MindCine: Multimodal EEG-to-Video Reconstruction with Large-Scale Pretrained Models
von: Zhou, Tian-Yi, et al.
Veröffentlicht: (2026)
von: Zhou, Tian-Yi, et al.
Veröffentlicht: (2026)
VideoMap: Supporting Video Editing Exploration, Brainstorming, and Prototyping in the Latent Space
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
MindCross: Fast New Subject Adaptation with Limited Data for Cross-subject Video Reconstruction from Brain Signals
von: Liu, Xuan-Hao, et al.
Veröffentlicht: (2025)
von: Liu, Xuan-Hao, et al.
Veröffentlicht: (2025)
BATON: A Multimodal Benchmark for Bidirectional Automation Transition Observation in Naturalistic Driving
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
Videogenic: Identifying Highlight Moments in Videos with Professional Photographs as a Prior
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
von: Lin, David Chuan-En, et al.
Veröffentlicht: (2022)
Seeing, Hearing, and Knowing Together: Multimodal Strategies in Deepfake Videos Detection
von: Chen, Chen, et al.
Veröffentlicht: (2026)
von: Chen, Chen, et al.
Veröffentlicht: (2026)
AttentionBender: Manipulating Cross-Attention in Video Diffusion Transformers as a Creative Probe
von: Cole, Adam, et al.
Veröffentlicht: (2026)
von: Cole, Adam, et al.
Veröffentlicht: (2026)
G3R: Generating Rich and Fine-grained mmWave Radar Data from 2D Videos for Generalized Gesture Recognition
von: Deng, Kaikai, et al.
Veröffentlicht: (2024)
von: Deng, Kaikai, et al.
Veröffentlicht: (2024)
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
von: Rakesh, Vineet Kumar, et al.
Veröffentlicht: (2025)
von: Rakesh, Vineet Kumar, et al.
Veröffentlicht: (2025)
FastPerson: Enhancing Video Learning through Effective Video Summarization that Preserves Linguistic and Visual Contexts
von: Kawamura, Kazuki, et al.
Veröffentlicht: (2024)
von: Kawamura, Kazuki, et al.
Veröffentlicht: (2024)
Plug-and-Play Clarifier: A Zero-Shot Multimodal Framework for Egocentric Intent Disambiguation
von: Yang, Sicheng, et al.
Veröffentlicht: (2025)
von: Yang, Sicheng, et al.
Veröffentlicht: (2025)
Color When It Counts: Grayscale-Guided Online Triggering for Always-On Streaming Video Sensing
von: Cai, Weitong, et al.
Veröffentlicht: (2026)
von: Cai, Weitong, et al.
Veröffentlicht: (2026)
ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion Model
von: Cheng, Luo, et al.
Veröffentlicht: (2025)
von: Cheng, Luo, et al.
Veröffentlicht: (2025)
SentiAvatar: Towards Expressive and Interactive Digital Humans
von: Jin, Chuhao, et al.
Veröffentlicht: (2026)
von: Jin, Chuhao, et al.
Veröffentlicht: (2026)
ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions
von: Luo, Cheng, et al.
Veröffentlicht: (2023)
von: Luo, Cheng, et al.
Veröffentlicht: (2023)
Across-Game Engagement Modelling via Few-Shot Learning
von: Pinitas, Kosmas, et al.
Veröffentlicht: (2024)
von: Pinitas, Kosmas, et al.
Veröffentlicht: (2024)
Emotion Based Prediction in the Context of Optimized Trajectory Planning for Immersive Learning
von: Sungheetha, Akey, et al.
Veröffentlicht: (2023)
von: Sungheetha, Akey, et al.
Veröffentlicht: (2023)
MS2Mesh-XR: Multi-modal Sketch-to-Mesh Generation in XR Environments
von: Tong, Yuqi, et al.
Veröffentlicht: (2024)
von: Tong, Yuqi, et al.
Veröffentlicht: (2024)
Generative Timelines for Instructed Visual Assembly
von: Pardo, Alejandro, et al.
Veröffentlicht: (2024)
von: Pardo, Alejandro, et al.
Veröffentlicht: (2024)
Shu Dao: A Calligraphy Score Framework Linking Calligraphy, Music, and Performance
von: Huang, Lican
Veröffentlicht: (2026)
von: Huang, Lican
Veröffentlicht: (2026)
Secure & Personalized Music-to-Video Generation via CHARCHA
von: Agarwal, Mehul, et al.
Veröffentlicht: (2025)
von: Agarwal, Mehul, et al.
Veröffentlicht: (2025)
LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs
von: Sun, Boyuan, et al.
Veröffentlicht: (2025)
von: Sun, Boyuan, et al.
Veröffentlicht: (2025)
DiffMesh: A Motion-aware Diffusion Framework for Human Mesh Recovery from Videos
von: Zheng, Ce, et al.
Veröffentlicht: (2023)
von: Zheng, Ce, et al.
Veröffentlicht: (2023)
Learning High-Quality Navigation and Zooming on Omnidirectional Images in Virtual Reality
von: Cao, Zidong, et al.
Veröffentlicht: (2024)
von: Cao, Zidong, et al.
Veröffentlicht: (2024)
"I Can See Forever!": Evaluating Real-time VideoLLMs for Assisting Individuals with Visual Impairments
von: Zhang, Ziyi, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyi, et al.
Veröffentlicht: (2025)
SkinGEN: an Explainable Dermatology Diagnosis-to-Generation Framework with Interactive Vision-Language Models
von: Lin, Bo, et al.
Veröffentlicht: (2024)
von: Lin, Bo, et al.
Veröffentlicht: (2024)
Focus360: Guiding User Attention in Immersive Videos for VR
von: Silva, Paulo Vitor S., et al.
Veröffentlicht: (2026)
von: Silva, Paulo Vitor S., et al.
Veröffentlicht: (2026)
Unveiling the Visual Rhetoric of Persuasive Cartography: A Case Study of the Design of Octopus Maps
von: Lin, Daocheng, et al.
Veröffentlicht: (2025)
von: Lin, Daocheng, et al.
Veröffentlicht: (2025)
MV-Crafter: An Intelligent System for Music-guided Video Generation
von: Chen, Chuer, et al.
Veröffentlicht: (2025)
von: Chen, Chuer, et al.
Veröffentlicht: (2025)
Code2Video: A Code-centric Paradigm for Educational Video Generation
von: Chen, Yanzhe, et al.
Veröffentlicht: (2025)
von: Chen, Yanzhe, et al.
Veröffentlicht: (2025)
WebXR, A-Frame and Networked-Aframe as a Basis for an Open Metaverse: A Conceptual Architecture
von: Macario, Giuseppe
Veröffentlicht: (2024)
von: Macario, Giuseppe
Veröffentlicht: (2024)
Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation
von: Ki, Taekyung, et al.
Veröffentlicht: (2026)
von: Ki, Taekyung, et al.
Veröffentlicht: (2026)
adder-viz: Real-Time Visualization Software for Transcoding Event Video
von: Freeman, Andrew C., et al.
Veröffentlicht: (2025)
von: Freeman, Andrew C., et al.
Veröffentlicht: (2025)
MIMOSA: Human-AI Co-Creation of Computational Spatial Audio Effects on Videos
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
SPICA: Interactive Video Content Exploration through Augmented Audio Descriptions for Blind or Low-Vision Viewers
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
von: Ning, Zheng, et al.
Veröffentlicht: (2024)
SMPLest-X: Ultimate Scaling for Expressive Human Pose and Shape Estimation
von: Yin, Wanqi, et al.
Veröffentlicht: (2025)
von: Yin, Wanqi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
von: He, Xu, et al.
Veröffentlicht: (2024) -
Editing Physiological Signals in Videos Using Latent Representations
von: Zhou, Tianwen, et al.
Veröffentlicht: (2025) -
EyeNavGS: A 6-DoF Navigation Dataset and Record-n-Replay Software for Real-World 3DGS Scenes in VR
von: Ding, Zihao, et al.
Veröffentlicht: (2025) -
SVFAP: Self-supervised Video Facial Affect Perceiver
von: Sun, Licai, et al.
Veröffentlicht: (2023) -
MindCine: Multimodal EEG-to-Video Reconstruction with Large-Scale Pretrained Models
von: Zhou, Tian-Yi, et al.
Veröffentlicht: (2026)