Gespeichert in:
| Hauptverfasser: | Bishay, Mina, Page, Graham, Emad, Waleed, Mavadati, Mohammad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2504.06237 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PainNet: Statistical Relation Network with Episode-Based Training for Pain Estimation
von: Bishay, Mina, et al.
Veröffentlicht: (2025)
von: Bishay, Mina, et al.
Veröffentlicht: (2025)
Radiance Field Learners As UAV First-Person Viewers
von: Yan, Liqi, et al.
Veröffentlicht: (2024)
von: Yan, Liqi, et al.
Veröffentlicht: (2024)
AdsQA: Towards Advertisement Video Understanding
von: Long, Xinwei, et al.
Veröffentlicht: (2025)
von: Long, Xinwei, et al.
Veröffentlicht: (2025)
VideoAds for Fast-Paced Video Understanding
von: Zhang, Zheyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Zheyuan, et al.
Veröffentlicht: (2025)
The Face of Persuasion: Analyzing Bias and Generating Culture-Aware Ads
von: Aghazadeh, Aysan, et al.
Veröffentlicht: (2025)
von: Aghazadeh, Aysan, et al.
Veröffentlicht: (2025)
OpenViewer: Openness-Aware Multi-View Learning
von: Du, Shide, et al.
Veröffentlicht: (2024)
von: Du, Shide, et al.
Veröffentlicht: (2024)
ManzaiSet: A Multimodal Dataset of Viewer Responses to Japanese Manzai Comedy
von: Kawamura, Kazuki, et al.
Veröffentlicht: (2025)
von: Kawamura, Kazuki, et al.
Veröffentlicht: (2025)
ViASNet: A Video Ad Saliency Network for Predicting Dynamic Saliency and Viewer Engagement
von: Ye, Jianping, et al.
Veröffentlicht: (2026)
von: Ye, Jianping, et al.
Veröffentlicht: (2026)
InnoAds-Composer: Efficient Condition Composition for E-Commerce Poster Generation
von: Qin, Yuxin, et al.
Veröffentlicht: (2026)
von: Qin, Yuxin, et al.
Veröffentlicht: (2026)
SplatBus: A Gaussian Splatting Viewer Framework via GPU Interprocess Communication
von: Xu, Yinghan, et al.
Veröffentlicht: (2026)
von: Xu, Yinghan, et al.
Veröffentlicht: (2026)
Visual Affect Analysis: Predicting Emotions of Image Viewers with Vision-Language Models
von: Nowicki, Filip, et al.
Veröffentlicht: (2026)
von: Nowicki, Filip, et al.
Veröffentlicht: (2026)
Class-Incremental Learning for Honey Botanical Origin Classification with Hyperspectral Images: A Study with Continual Backpropagation
von: Zhang, Guyang, et al.
Veröffentlicht: (2025)
von: Zhang, Guyang, et al.
Veröffentlicht: (2025)
On filter design in deep convolutional neural network
von: Hirani, Gaurav, et al.
Veröffentlicht: (2024)
von: Hirani, Gaurav, et al.
Veröffentlicht: (2024)
NextAds: Towards Next-generation Personalized Video Advertising
von: Xu, Yiyan, et al.
Veröffentlicht: (2026)
von: Xu, Yiyan, et al.
Veröffentlicht: (2026)
FLAASH: Flow-Attention Adaptive Semantic Hierarchical Fusion for Multi-Modal Tobacco Content Analysis
von: Chappa, Naga VS Raviteja, et al.
Veröffentlicht: (2024)
von: Chappa, Naga VS Raviteja, et al.
Veröffentlicht: (2024)
DASS: Differentiable Architecture Search for Sparse neural networks
von: Mousavi, Hamid, et al.
Veröffentlicht: (2022)
von: Mousavi, Hamid, et al.
Veröffentlicht: (2022)
IBO: Inpainting-Based Occlusion to Enhance Explainable Artificial Intelligence Evaluation in Histopathology
von: Afshar, Pardis, et al.
Veröffentlicht: (2024)
von: Afshar, Pardis, et al.
Veröffentlicht: (2024)
Long-Term Ad Memorability: Understanding & Generating Memorable Ads
von: SI, Harini, et al.
Veröffentlicht: (2023)
von: SI, Harini, et al.
Veröffentlicht: (2023)
CalibRefine: Deep Learning-Based Online Automatic Targetless LiDAR-Camera Calibration with Iterative and Attention-Driven Post-Refinement
von: Cheng, Lei, et al.
Veröffentlicht: (2025)
von: Cheng, Lei, et al.
Veröffentlicht: (2025)
Noise is an Efficient Learner for Zero-Shot Vision-Language Models
von: Imam, Raza, et al.
Veröffentlicht: (2025)
von: Imam, Raza, et al.
Veröffentlicht: (2025)
Camera Calibration through Geometric Constraints from Rotation and Projection Matrices
von: Waleed, Muhammad, et al.
Veröffentlicht: (2024)
von: Waleed, Muhammad, et al.
Veröffentlicht: (2024)
ExIQA: Explainable Image Quality Assessment Using Distortion Attributes
von: Ranjbar, Sepehr Kazemi, et al.
Veröffentlicht: (2024)
von: Ranjbar, Sepehr Kazemi, et al.
Veröffentlicht: (2024)
Static Key Attention in Vision
von: Hu, Zizhao, et al.
Veröffentlicht: (2024)
von: Hu, Zizhao, et al.
Veröffentlicht: (2024)
Partial Convolution Meets Visual Attention
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
von: Huang, Haiduo, et al.
Veröffentlicht: (2025)
Saliency-guided Emotion Modeling: Predicting Viewer Reactions from Video Stimuli
von: Yaragoppa, Akhila, et al.
Veröffentlicht: (2025)
von: Yaragoppa, Akhila, et al.
Veröffentlicht: (2025)
PEAVS: Perceptual Evaluation of Audio-Visual Synchrony Grounded in Viewers' Opinion Scores
von: Goncalves, Lucas, et al.
Veröffentlicht: (2024)
von: Goncalves, Lucas, et al.
Veröffentlicht: (2024)
Transformers Meet Hyperspectral Imaging: A Comprehensive Study of Models, Challenges and Open Problems
von: Zhang, Guyang, et al.
Veröffentlicht: (2025)
von: Zhang, Guyang, et al.
Veröffentlicht: (2025)
GAZE: Grounded Agentic Zero-shot Evaluation with Viewer-Level Tools and Literature Retrieval on Rare Brain MRI
von: Alim, Duaa, et al.
Veröffentlicht: (2026)
von: Alim, Duaa, et al.
Veröffentlicht: (2026)
SoGAR: Self-supervised Spatiotemporal Attention-based Social Group Activity Recognition
von: Chappa, Naga VS Raviteja, et al.
Veröffentlicht: (2023)
von: Chappa, Naga VS Raviteja, et al.
Veröffentlicht: (2023)
Smart Camera Parking System With Auto Parking Spot Detection
von: Nguyen, Tuan T., et al.
Veröffentlicht: (2024)
von: Nguyen, Tuan T., et al.
Veröffentlicht: (2024)
Online Continual Domain Adaptation for Semantic Image Segmentation Using Internal Representations
von: Stan, Serban, et al.
Veröffentlicht: (2024)
von: Stan, Serban, et al.
Veröffentlicht: (2024)
DiffusionWorldViewer: Exposing and Broadening the Worldview Reflected by Generative Text-to-Image Models
von: De Simone, Zoe, et al.
Veröffentlicht: (2023)
von: De Simone, Zoe, et al.
Veröffentlicht: (2023)
Online Micro-gesture Recognition Using Data Augmentation and Spatial-Temporal Attention
von: Liu, Pengyu, et al.
Veröffentlicht: (2025)
von: Liu, Pengyu, et al.
Veröffentlicht: (2025)
Robust Multiview Multimodal Driver Monitoring System Using Masked Multi-Head Self-Attention
von: Ma, Yiming, et al.
Veröffentlicht: (2023)
von: Ma, Yiming, et al.
Veröffentlicht: (2023)
Progressive Conditioned Scale-Shift Recalibration of Self-Attention for Online Test-time Adaptation
von: Tang, Yushun, et al.
Veröffentlicht: (2025)
von: Tang, Yushun, et al.
Veröffentlicht: (2025)
Taylor-Series Expanded Kolmogorov-Arnold Network for Medical Imaging Classification
von: Fatema, Kaniz, et al.
Veröffentlicht: (2025)
von: Fatema, Kaniz, et al.
Veröffentlicht: (2025)
Hierarchical Vector Quantization for Unsupervised Action Segmentation
von: Spurio, Federico, et al.
Veröffentlicht: (2024)
von: Spurio, Federico, et al.
Veröffentlicht: (2024)
CASP: Compression of Large Multimodal Models Based on Attention Sparsity
von: Gholami, Mohsen, et al.
Veröffentlicht: (2025)
von: Gholami, Mohsen, et al.
Veröffentlicht: (2025)
Attention-based U-Net Method for Autonomous Lane Detection
von: Tangestanizadeh, Mohammadhamed, et al.
Veröffentlicht: (2024)
von: Tangestanizadeh, Mohammadhamed, et al.
Veröffentlicht: (2024)
Non-Contact Health Monitoring During Daily Personal Care Routines
von: Ma, Xulin, et al.
Veröffentlicht: (2025)
von: Ma, Xulin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PainNet: Statistical Relation Network with Episode-Based Training for Pain Estimation
von: Bishay, Mina, et al.
Veröffentlicht: (2025) -
Radiance Field Learners As UAV First-Person Viewers
von: Yan, Liqi, et al.
Veröffentlicht: (2024) -
AdsQA: Towards Advertisement Video Understanding
von: Long, Xinwei, et al.
Veröffentlicht: (2025) -
VideoAds for Fast-Paced Video Understanding
von: Zhang, Zheyuan, et al.
Veröffentlicht: (2025) -
The Face of Persuasion: Analyzing Bias and Generating Culture-Aware Ads
von: Aghazadeh, Aysan, et al.
Veröffentlicht: (2025)