Video-Based MPAA Rating Prediction: An Attention-Driven Hybrid Architecture Using Contrastive Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Neogi, Dipta, Chowdhury, Nourash Azmine, Kabir, Muhammad Rafsan, Khan, Mohammad Ashrafuzzaman |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Balancing Accuracy and Efficiency: CNN Fusion Models for Diabetic Retinopathy Screening
by: Islam, Md Rafid, et al.
Published: (2025)
by: Islam, Md Rafid, et al.
Published: (2025)
BanglaEmbed: Efficient Sentence Embedding Models for a Low-Resource Language Using Cross-Lingual Distillation Techniques
by: Kabir, Muhammad Rafsan, et al.
Published: (2024)
by: Kabir, Muhammad Rafsan, et al.
Published: (2024)
Supervised Contrastive Frame Aggregation for Video Representation Learning
by: Chowdhury, Shaif, et al.
Published: (2025)
by: Chowdhury, Shaif, et al.
Published: (2025)
Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual Text-to-Video Retrieval
by: Dipta, Shubhashis Roy, et al.
Published: (2025)
by: Dipta, Shubhashis Roy, et al.
Published: (2025)
Breaking the Geometric Bottleneck: Contrastive Expansion in Asymmetric Cross-Modal Distillation
by: Thayani, Kabir
Published: (2026)
by: Thayani, Kabir
Published: (2026)
MHAFF: Multi-Head Attention Feature Fusion of CNN and Transformer for Cattle Identification
by: Dulal, Rabin, et al.
Published: (2025)
by: Dulal, Rabin, et al.
Published: (2025)
TriFusion-AE: Language-Guided Depth and LiDAR Fusion for Robust Point Cloud Processing
by: Neogi, Susmit
Published: (2025)
by: Neogi, Susmit
Published: (2025)
Deep Learning for Early Alzheimer Disease Detection with MRI Scans
by: Rafsan, Mohammad, et al.
Published: (2025)
by: Rafsan, Mohammad, et al.
Published: (2025)
Attention-Guided Dual-Stream Learning for Group Engagement Recognition: Fusing Transformer-Encoded Motion Dynamics with Scene Context via Adaptive Gating
by: Chowdhury, Saniah Kayenat, et al.
Published: (2026)
by: Chowdhury, Saniah Kayenat, et al.
Published: (2026)
Beyond Real Weights: Hypercomplex Representations for Stable Quantization
by: Ahad, Jawad Ibn, et al.
Published: (2025)
by: Ahad, Jawad Ibn, et al.
Published: (2025)
NURBGen: High-Fidelity Text-to-CAD Generation through LLM-Driven NURBS Modeling
by: Usama, Muhammad, et al.
Published: (2025)
by: Usama, Muhammad, et al.
Published: (2025)
SRVP: Strong Recollection Video Prediction Model Using Attention-Based Spatiotemporal Correlation Fusion
by: Kim, Yuseon, et al.
Published: (2025)
by: Kim, Yuseon, et al.
Published: (2025)
VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding
by: Maaz, Muhammad, et al.
Published: (2024)
by: Maaz, Muhammad, et al.
Published: (2024)
VC-Inspector: Advancing Reference-free Evaluation of Video Captions with Factual Analysis
by: Dipta, Shubhashis Roy, et al.
Published: (2025)
by: Dipta, Shubhashis Roy, et al.
Published: (2025)
An Attention-Guided Deep Learning Approach for Classifying 39 Skin Lesion Types
by: Hanum, Sauda Adiv, et al.
Published: (2025)
by: Hanum, Sauda Adiv, et al.
Published: (2025)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
by: Baliah, Sanoojan, et al.
Published: (2026)
by: Baliah, Sanoojan, et al.
Published: (2026)
Adapt, But Don't Forget: Fine-Tuning and Contrastive Routing for Lane Detection under Distribution Shift
by: Khan, Mohammed Abdul Hafeez, et al.
Published: (2025)
by: Khan, Mohammed Abdul Hafeez, et al.
Published: (2025)
Attention Based Simple Primitives for Open World Compositional Zero-Shot Learning
by: Munir, Ans, et al.
Published: (2024)
by: Munir, Ans, et al.
Published: (2024)
Point Cloud Understanding via Attention-Driven Contrastive Learning
by: Wang, Yi, et al.
Published: (2024)
by: Wang, Yi, et al.
Published: (2024)
DFCon: Attention-Driven Supervised Contrastive Learning for Robust Deepfake Detection
by: Shanto, MD Sadik Hossain, et al.
Published: (2025)
by: Shanto, MD Sadik Hossain, et al.
Published: (2025)
Beyond Uniform Query Distribution: Key-Driven Grouped Query Attention
by: Khan, Zohaib, et al.
Published: (2024)
by: Khan, Zohaib, et al.
Published: (2024)
Contrastive-SDXL: Annotation-Preserving Night-Time Augmentation for Pedestrian Detection
by: George, Franky, et al.
Published: (2026)
by: George, Franky, et al.
Published: (2026)
Chronic Obstructive Pulmonary Disease Prediction Using Deep Convolutional Network
by: Alve, Shahran Rahman, et al.
Published: (2024)
by: Alve, Shahran Rahman, et al.
Published: (2024)
Video Representation Learning with Joint-Embedding Predictive Architectures
by: Drozdov, Katrina, et al.
Published: (2024)
by: Drozdov, Katrina, et al.
Published: (2024)
LMFLOSS: A Hybrid Loss For Imbalanced Medical Image Classification
by: Sadi, Abu Adnan, et al.
Published: (2022)
by: Sadi, Abu Adnan, et al.
Published: (2022)
Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
by: Maaz, Muhammad, et al.
Published: (2023)
by: Maaz, Muhammad, et al.
Published: (2023)
Echo-Path: Pathology-Conditioned Echo Video Generation
by: Muhammad, Kabir Hamzah, et al.
Published: (2025)
by: Muhammad, Kabir Hamzah, et al.
Published: (2025)
Joint Flow And Feature Refinement Using Attention For Video Restoration
by: Merugu, Ranjith, et al.
Published: (2025)
by: Merugu, Ranjith, et al.
Published: (2025)
Attention-Based Ensemble Learning for Crop Classification Using Landsat 8-9 Fusion
by: Ramzan, Zeeshan, et al.
Published: (2025)
by: Ramzan, Zeeshan, et al.
Published: (2025)
How Good is my Video LMM? Complex Video Reasoning and Robustness Evaluation Suite for Video-LMMs
by: Khattak, Muhammad Uzair, et al.
Published: (2024)
by: Khattak, Muhammad Uzair, et al.
Published: (2024)
Exploring Personalized Federated Learning Architectures for Violence Detection in Surveillance Videos
by: Kassir, Mohammad, et al.
Published: (2025)
by: Kassir, Mohammad, et al.
Published: (2025)
WoundFormer: Multi-Scale Spatial Feature Fusion for Multi-Class Wound Tissue Segmentation
by: Kabir, Muhammad Ashad, et al.
Published: (2026)
by: Kabir, Muhammad Ashad, et al.
Published: (2026)
Unmasking Deep Fakes: Leveraging Deep Learning for Video Authenticity Detection
by: Hasan, Mahmudul, et al.
Published: (2025)
by: Hasan, Mahmudul, et al.
Published: (2025)
VEDIT: Latent Prediction Architecture For Procedural Video Representation Learning
by: Lin, Han, et al.
Published: (2024)
by: Lin, Han, et al.
Published: (2024)
ReHyAt: Recurrent Hybrid Attention for Video Diffusion Transformers
by: Ghafoorian, Mohsen, et al.
Published: (2026)
by: Ghafoorian, Mohsen, et al.
Published: (2026)
SurvRNC: Learning Ordered Representations for Survival Prediction using Rank-N-Contrast
by: Saeed, Numan, et al.
Published: (2024)
by: Saeed, Numan, et al.
Published: (2024)
FPGA-Based Hardware Architecture for Contrast Maximization in Event-Based Vision
by: Filipkowski, Michal, et al.
Published: (2026)
by: Filipkowski, Michal, et al.
Published: (2026)
Squeezed-Eff-Net: Edge-Computed Boost of Tomography Based Brain Tumor Classification leveraging Hybrid Neural Network Architecture
by: Chowdhury, Md. Srabon, et al.
Published: (2025)
by: Chowdhury, Md. Srabon, et al.
Published: (2025)
Efficient Video Object Segmentation via Modulated Cross-Attention Memory
by: Shaker, Abdelrahman, et al.
Published: (2024)
by: Shaker, Abdelrahman, et al.
Published: (2024)
SparseContrast: Dynamic Sparse Attention for Efficient and Accurate Contrastive Learning in Medical Imaging
by: Prasad, Paarth, et al.
Published: (2026)
by: Prasad, Paarth, et al.
Published: (2026)
Similar Items
-
Balancing Accuracy and Efficiency: CNN Fusion Models for Diabetic Retinopathy Screening
by: Islam, Md Rafid, et al.
Published: (2025) -
BanglaEmbed: Efficient Sentence Embedding Models for a Low-Resource Language Using Cross-Lingual Distillation Techniques
by: Kabir, Muhammad Rafsan, et al.
Published: (2024) -
Supervised Contrastive Frame Aggregation for Video Representation Learning
by: Chowdhury, Shaif, et al.
Published: (2025) -
Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual Text-to-Video Retrieval
by: Dipta, Shubhashis Roy, et al.
Published: (2025) -
Breaking the Geometric Bottleneck: Contrastive Expansion in Asymmetric Cross-Modal Distillation
by: Thayani, Kabir
Published: (2026)