Demystifying Visual Features of Movie Posters for Multi-Label Genre Identification
Fuente:
arXiv
Saved in:
| Main Authors: | Nareti, Utsav Kumar, Adak, Chandranath, Chattopadhyay, Soumi |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unraveling Movie Genres through Cross-Attention Fusion of Bi-Modal Synergy of Poster
by: Nareti, Utsav Kumar, et al.
Published: (2024)
by: Nareti, Utsav Kumar, et al.
Published: (2024)
Detecting Severity of Diabetic Retinopathy from Fundus Images: A Transformer Network-based Review
by: Karkera, Tejas, et al.
Published: (2023)
by: Karkera, Tejas, et al.
Published: (2023)
Adaptive Data-Resilient Multi-Modal Hierarchical Multi-Label Book Genre Identification
by: Nareti, Utsav Kumar, et al.
Published: (2025)
by: Nareti, Utsav Kumar, et al.
Published: (2025)
ProtoSiTex: Learning Semi-Interpretable Prototypes for Multi-label Text Classification
by: Nareti, Utsav Kumar, et al.
Published: (2025)
by: Nareti, Utsav Kumar, et al.
Published: (2025)
Blurb-Refined Inference from Crowdsourced Book Reviews using Hierarchical Genre Mining with Dual-Path Graph Convolutions
by: Kumar, Suraj, et al.
Published: (2025)
by: Kumar, Suraj, et al.
Published: (2025)
Impact of Visual Context on Noisy Multimodal NMT: An Empirical Study for English to Indian Languages
by: Gain, Baban, et al.
Published: (2023)
by: Gain, Baban, et al.
Published: (2023)
Movie Trailer Genre Classification Using Multimodal Pretrained Features
by: Sulun, Serkan, et al.
Published: (2024)
by: Sulun, Serkan, et al.
Published: (2024)
Demystifying the Visual Quality Paradox in Multimodal Large Language Models
by: Xing, Shuo, et al.
Published: (2025)
by: Xing, Shuo, et al.
Published: (2025)
MovieCORE: COgnitive REasoning in Movies
by: Faure, Gueter Josmy, et al.
Published: (2025)
by: Faure, Gueter Josmy, et al.
Published: (2025)
MovieTeller: Tool-augmented Movie Synopsis with ID Consistent Progressive Abstraction
by: Li, Yizhi, et al.
Published: (2026)
by: Li, Yizhi, et al.
Published: (2026)
Demystifying Video Reasoning
by: Wang, Ruisi, et al.
Published: (2026)
by: Wang, Ruisi, et al.
Published: (2026)
Seeing is Believing (and Predicting): Context-Aware Multi-Human Behavior Prediction with Vision Language Models
by: Panchal, Utsav, et al.
Published: (2025)
by: Panchal, Utsav, et al.
Published: (2025)
Demystifying Foreground-Background Memorization in Diffusion Models
by: Di, Jimmy Z., et al.
Published: (2025)
by: Di, Jimmy Z., et al.
Published: (2025)
Movie Gen: SWOT Analysis of Meta's Generative AI Foundation Model for Transforming Media Generation, Advertising, and Entertainment Industries
by: Ehtesham, Abul, et al.
Published: (2024)
by: Ehtesham, Abul, et al.
Published: (2024)
Diffusion-Based Cross-Modal Feature Extraction for Multi-Label Classification
by: Lan, Tian, et al.
Published: (2025)
by: Lan, Tian, et al.
Published: (2025)
PosterSum: A Multimodal Benchmark for Scientific Poster Summarization
by: Saxena, Rohit, et al.
Published: (2025)
by: Saxena, Rohit, et al.
Published: (2025)
Demystifying KAN for Vision Tasks: The RepKAN Approach
by: Cheon, Minjong
Published: (2026)
by: Cheon, Minjong
Published: (2026)
Feature Identification for Hierarchical Contrastive Learning
by: Ott, Julius, et al.
Published: (2025)
by: Ott, Julius, et al.
Published: (2025)
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents
by: Yang, Rui, et al.
Published: (2026)
by: Yang, Rui, et al.
Published: (2026)
Reliable or Deceptive? Investigating Gated Features for Smooth Visual Explanations in CNNs
by: Mitra, Soham, et al.
Published: (2024)
by: Mitra, Soham, et al.
Published: (2024)
Actional Atomic-Concept Learning for Demystifying Vision-Language Navigation
by: Lin, Bingqian, et al.
Published: (2023)
by: Lin, Bingqian, et al.
Published: (2023)
Multi-Label Contrastive Learning for Abstract Visual Reasoning
by: Małkiński, Mikołaj, et al.
Published: (2020)
by: Małkiński, Mikołaj, et al.
Published: (2020)
No Labels, No Problem: Training Visual Reasoners with Multimodal Verifiers
by: Marsili, Damiano, et al.
Published: (2025)
by: Marsili, Damiano, et al.
Published: (2025)
Demystifying the Potential of ChatGPT-4 Vision for Construction Progress Monitoring
by: Ersoz, Ahmet Bahaddin
Published: (2024)
by: Ersoz, Ahmet Bahaddin
Published: (2024)
ScreenWriter: Automatic Screenplay Generation and Movie Summarisation
by: Mahon, Louis, et al.
Published: (2024)
by: Mahon, Louis, et al.
Published: (2024)
Poster: Camera Tampering Detection for Outdoor IoT Systems
by: Attarha, Shadi, et al.
Published: (2026)
by: Attarha, Shadi, et al.
Published: (2026)
Label Sharing Incremental Learning Framework for Independent Multi-Label Segmentation Tasks
by: Anand, Deepa, et al.
Published: (2024)
by: Anand, Deepa, et al.
Published: (2024)
Binding Visual Features Point by Point
by: Haputhanthri, Udith, et al.
Published: (2026)
by: Haputhanthri, Udith, et al.
Published: (2026)
HELM: Hierarchical and Explicit Label Modeling with Graph Learning for Multi-Label Image Classification
by: Stoimchev, Marjan, et al.
Published: (2026)
by: Stoimchev, Marjan, et al.
Published: (2026)
IDA-VLM: Towards Movie Understanding via ID-Aware Large Vision-Language Model
by: Ji, Yatai, et al.
Published: (2024)
by: Ji, Yatai, et al.
Published: (2024)
L3A: Label-Augmented Analytic Adaptation for Multi-Label Class Incremental Learning
by: Zhang, Xiang, et al.
Published: (2025)
by: Zhang, Xiang, et al.
Published: (2025)
VisNet: Efficient Person Re-Identification via Alpha-Divergence Loss, Feature Fusion and Dynamic Multi-Task Learning
by: Ijaz, Anns, et al.
Published: (2026)
by: Ijaz, Anns, et al.
Published: (2026)
Non-planar Object Detection and Identification by Features Matching and Triangulation Growth
by: Leveni, Filippo
Published: (2025)
by: Leveni, Filippo
Published: (2025)
Hard to See, Hard to Label: Generative and Symbolic Acquisition for Subtle Visual Phenomena
by: Prasad, Renjith, et al.
Published: (2026)
by: Prasad, Renjith, et al.
Published: (2026)
Generalizable Object Re-Identification via Visual In-Context Prompting
by: Huang, Zhizhong, et al.
Published: (2025)
by: Huang, Zhizhong, et al.
Published: (2025)
Multi-Label Classification Framework for Hurricane Damage Assessment
by: Liu, Zhangding, et al.
Published: (2025)
by: Liu, Zhangding, et al.
Published: (2025)
Relieving Universal Label Noise for Unsupervised Visible-Infrared Person Re-Identification by Inferring from Neighbors
by: Teng, Xiao, et al.
Published: (2024)
by: Teng, Xiao, et al.
Published: (2024)
Adaptive Disentangled Representation Learning for Incomplete Multi-View Multi-Label Classification
by: Li, Quanjiang, et al.
Published: (2026)
by: Li, Quanjiang, et al.
Published: (2026)
AttZoom: Attention Zoom for Better Visual Features
by: DeAlcala, Daniel, et al.
Published: (2025)
by: DeAlcala, Daniel, et al.
Published: (2025)
Understanding Visual Feature Reliance through the Lens of Complexity
by: Fel, Thomas, et al.
Published: (2024)
by: Fel, Thomas, et al.
Published: (2024)
Similar Items
-
Unraveling Movie Genres through Cross-Attention Fusion of Bi-Modal Synergy of Poster
by: Nareti, Utsav Kumar, et al.
Published: (2024) -
Detecting Severity of Diabetic Retinopathy from Fundus Images: A Transformer Network-based Review
by: Karkera, Tejas, et al.
Published: (2023) -
Adaptive Data-Resilient Multi-Modal Hierarchical Multi-Label Book Genre Identification
by: Nareti, Utsav Kumar, et al.
Published: (2025) -
ProtoSiTex: Learning Semi-Interpretable Prototypes for Multi-label Text Classification
by: Nareti, Utsav Kumar, et al.
Published: (2025) -
Blurb-Refined Inference from Crowdsourced Book Reviews using Hierarchical Genre Mining with Dual-Path Graph Convolutions
by: Kumar, Suraj, et al.
Published: (2025)