Listen to the Unexpected: Self-Supervised Surprise Detection for Efficient Viewport Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Khah, Arman Nik, Prakash, Ravi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Meaning over Motion: A Semantic-First Approach to 360° Viewport Prediction
von: Khah, Arman Nik, et al.
Veröffentlicht: (2026)
von: Khah, Arman Nik, et al.
Veröffentlicht: (2026)
Score Distillation Sampling for Audio: Source Separation, Synthesis, and Beyond
von: Richter-Powell, Jessie, et al.
Veröffentlicht: (2025)
von: Richter-Powell, Jessie, et al.
Veröffentlicht: (2025)
Automatic Album Sequencing
von: Herrmann, Vincent, et al.
Veröffentlicht: (2024)
von: Herrmann, Vincent, et al.
Veröffentlicht: (2024)
Adaptable Symbolic Music Infilling with MIDI-RWKV
von: Zhou-Zheng, Christian, et al.
Veröffentlicht: (2025)
von: Zhou-Zheng, Christian, et al.
Veröffentlicht: (2025)
Music2P: A Multi-Modal AI-Driven Tool for Simplifying Album Cover Design
von: Choi, Joong Ho, et al.
Veröffentlicht: (2024)
von: Choi, Joong Ho, et al.
Veröffentlicht: (2024)
DFingerNet: Noise-Adaptive Speech Enhancement for Hearing Aids
von: Tsangko, Iosif, et al.
Veröffentlicht: (2025)
von: Tsangko, Iosif, et al.
Veröffentlicht: (2025)
PromptReverb: Multimodal Room Impulse Response Generation Through Latent Rectified Flow Matching
von: Vosoughi, Ali, et al.
Veröffentlicht: (2025)
von: Vosoughi, Ali, et al.
Veröffentlicht: (2025)
BemaGANv2: Discriminator Combination Strategies for GAN-based Vocoders in Long-Term Audio Generation
von: Park, Taesoo, et al.
Veröffentlicht: (2025)
von: Park, Taesoo, et al.
Veröffentlicht: (2025)
Information-Theoretic Quality Metric of Low-Dimensional Embeddings
von: Gutiérrez-Bernal, Sebastián, et al.
Veröffentlicht: (2025)
von: Gutiérrez-Bernal, Sebastián, et al.
Veröffentlicht: (2025)
FlightSense: An End-to-End MLOps Platform for Real-Time Flight Delay Prediction via Rotation-Chain Propagation Features and Agentic Conversational AI
von: Shelke, Aditi J., et al.
Veröffentlicht: (2026)
von: Shelke, Aditi J., et al.
Veröffentlicht: (2026)
A Framework for Multimodal Medical Image Interaction
von: Schütz, Laura, et al.
Veröffentlicht: (2024)
von: Schütz, Laura, et al.
Veröffentlicht: (2024)
Benchmarking Sub-Genre Classification For Mainstage Dance Music
von: Shu, Hongzhi, et al.
Veröffentlicht: (2024)
von: Shu, Hongzhi, et al.
Veröffentlicht: (2024)
MuQ-Eval: An Open-Source Per-Sample Quality Metric for AI Music Generation Evaluation
von: Zhu, Di, et al.
Veröffentlicht: (2026)
von: Zhu, Di, et al.
Veröffentlicht: (2026)
Augmented Assembly: Object Recognition and Hand Tracking for Adaptive Assembly Instructions in Augmented Reality
von: Kyaw, Alexander Htet, et al.
Veröffentlicht: (2025)
von: Kyaw, Alexander Htet, et al.
Veröffentlicht: (2025)
DGTEN: A Robust Deep Gaussian based Graph Neural Network for Dynamic Trust Evaluation with Uncertainty-Quantification Support
von: Usman, Muhammad, et al.
Veröffentlicht: (2025)
von: Usman, Muhammad, et al.
Veröffentlicht: (2025)
Quantum-Enhanced Analysis and Grading of Vocal Performance
von: Agarwal, Rohan
Veröffentlicht: (2025)
von: Agarwal, Rohan
Veröffentlicht: (2025)
Beyond Seeing Is Believing: On Crowdsourced Detection of Audiovisual Deepfakes
von: Soprano, Michael, et al.
Veröffentlicht: (2026)
von: Soprano, Michael, et al.
Veröffentlicht: (2026)
Self++: Co-Determined Agency for Human--AI Symbiosis in Extended Reality
von: Piumsomboon, Thammathip
Veröffentlicht: (2026)
von: Piumsomboon, Thammathip
Veröffentlicht: (2026)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
The Interaction Fidelity Model: A Taxonomy to Distinguish the Aspects of Fidelity in Virtual Reality
von: Bonfert, Michael, et al.
Veröffentlicht: (2024)
von: Bonfert, Michael, et al.
Veröffentlicht: (2024)
SonoHaptics: An Audio-Haptic Cursor for Gaze-Based Object Selection in XR
von: Cho, Hyunsung, et al.
Veröffentlicht: (2024)
von: Cho, Hyunsung, et al.
Veröffentlicht: (2024)
MoXaRt: Audio-Visual Object-Guided Sound Interaction for XR
von: Xu, Tianyu, et al.
Veröffentlicht: (2026)
von: Xu, Tianyu, et al.
Veröffentlicht: (2026)
Audio Foundation Models Outperform Symbolic Representations for Piano Performance Evaluation
von: Dhiman, Jai
Veröffentlicht: (2026)
von: Dhiman, Jai
Veröffentlicht: (2026)
APEX: Large-scale Multi-task Aesthetic-Informed Popularity Prediction for AI-Generated Music
von: Husain, Jaavid Aktar, et al.
Veröffentlicht: (2026)
von: Husain, Jaavid Aktar, et al.
Veröffentlicht: (2026)
Audience Amplified: Virtual Audiences in Asynchronously Performed AR Theater
von: Kim, You-Jin, et al.
Veröffentlicht: (2025)
von: Kim, You-Jin, et al.
Veröffentlicht: (2025)
Transformers Meet Relational Databases
von: Peleška, Jakub, et al.
Veröffentlicht: (2024)
von: Peleška, Jakub, et al.
Veröffentlicht: (2024)
LEFT: Learnable Fusion of Tri-view Tokens for Unsupervised Time Series Anomaly Detection
von: Wang, Dezheng, et al.
Veröffentlicht: (2026)
von: Wang, Dezheng, et al.
Veröffentlicht: (2026)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
von: Bachar, Or, et al.
Veröffentlicht: (2026)
von: Bachar, Or, et al.
Veröffentlicht: (2026)
Generative AI for Video Translation: A Scalable Architecture for Multilingual Video Conferencing
von: Oskooei, Amirkia Rafiei, et al.
Veröffentlicht: (2025)
von: Oskooei, Amirkia Rafiei, et al.
Veröffentlicht: (2025)
Refining music sample identification with a self-supervised graph neural network
von: Bhattacharjee, Aditya, et al.
Veröffentlicht: (2025)
von: Bhattacharjee, Aditya, et al.
Veröffentlicht: (2025)
Predicting Traffic Accident Severity with Deep Neural Networks
von: Bibb, Meghan, et al.
Veröffentlicht: (2025)
von: Bibb, Meghan, et al.
Veröffentlicht: (2025)
Enhancing Classification with Semi-Supervised Deep Learning Using Distance-Based Sample Weights
von: Abedinia, Aydin, et al.
Veröffentlicht: (2025)
von: Abedinia, Aydin, et al.
Veröffentlicht: (2025)
Prevailing Research Areas for Music AI in the Era of Foundation Models
von: Wei, Megan, et al.
Veröffentlicht: (2024)
von: Wei, Megan, et al.
Veröffentlicht: (2024)
Evaluating Keyframe Layouts for Visual Known-Item Search in Homogeneous Collections
von: Jäckl, Bastian, et al.
Veröffentlicht: (2025)
von: Jäckl, Bastian, et al.
Veröffentlicht: (2025)
Geo2Sound: A Scalable Geo-Aligned Framework for Soundscape Generation from Satellite Imagery
von: Wu, Kunlin, et al.
Veröffentlicht: (2026)
von: Wu, Kunlin, et al.
Veröffentlicht: (2026)
LakeMLB: Data Lake Machine Learning Benchmark
von: Pan, Feiyu, et al.
Veröffentlicht: (2026)
von: Pan, Feiyu, et al.
Veröffentlicht: (2026)
CASE: Contrastive Activation for Saliency Estimation
von: Williamson, Dane, et al.
Veröffentlicht: (2025)
von: Williamson, Dane, et al.
Veröffentlicht: (2025)
Machine Learning Framework for Audio-Based Content Evaluation using MFCC, Chroma, Spectral Contrast, and Temporal Feature Engineering
von: Aristorenas, Aris J.
Veröffentlicht: (2024)
von: Aristorenas, Aris J.
Veröffentlicht: (2024)
Joint Estimation of Piano Dynamics and Metrical Structure with a Multi-task Multi-Scale Network
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
FaceValue: Exploring Real-Time Self-View Overlays to Prompt Meaning-Oriented Self-Awareness in Remote Meetings
von: Park, Gun Woo Warren, et al.
Veröffentlicht: (2026)
von: Park, Gun Woo Warren, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Meaning over Motion: A Semantic-First Approach to 360° Viewport Prediction
von: Khah, Arman Nik, et al.
Veröffentlicht: (2026) -
Score Distillation Sampling for Audio: Source Separation, Synthesis, and Beyond
von: Richter-Powell, Jessie, et al.
Veröffentlicht: (2025) -
Automatic Album Sequencing
von: Herrmann, Vincent, et al.
Veröffentlicht: (2024) -
Adaptable Symbolic Music Infilling with MIDI-RWKV
von: Zhou-Zheng, Christian, et al.
Veröffentlicht: (2025) -
Music2P: A Multi-Modal AI-Driven Tool for Simplifying Album Cover Design
von: Choi, Joong Ho, et al.
Veröffentlicht: (2024)