Visual Content Detection in Educational Videos with Transfer Learning and Dataset Enrichment
Fuente:
arXiv
Saved in:
| Main Authors: | Biswas, Dipayan, Shah, Shishir, Subhlok, Jaspal |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lecture Video Visual Objects (LVVO) Dataset: A Benchmark for Visual Object Detection in Educational Videos
by: Biswas, Dipayan, et al.
Published: (2025)
by: Biswas, Dipayan, et al.
Published: (2025)
Attention-based Shape and Gait Representations Learning for Video-based Cloth-Changing Person Re-Identification
by: Nguyen, Vuong D., et al.
Published: (2024)
by: Nguyen, Vuong D., et al.
Published: (2024)
GeoStack: A Framework for Quasi-Abelian Knowledge Composition in VLMs
by: Mantini, Pranav, et al.
Published: (2026)
by: Mantini, Pranav, et al.
Published: (2026)
Generated Contents Enrichment
by: Naseri, Mahdi, et al.
Published: (2024)
by: Naseri, Mahdi, et al.
Published: (2024)
CCPA: Long-term Person Re-Identification via Contrastive Clothing and Pose Augmentation
by: Nguyen, Vuong D., et al.
Published: (2024)
by: Nguyen, Vuong D., et al.
Published: (2024)
Leveraging Pre-Trained Visual Models for AI-Generated Video Detection
by: Veeramachaneni, Keerthi, et al.
Published: (2025)
by: Veeramachaneni, Keerthi, et al.
Published: (2025)
TeleStyle: Content-Preserving Style Transfer in Images and Videos
by: Zhang, Shiwen, et al.
Published: (2026)
by: Zhang, Shiwen, et al.
Published: (2026)
CAE-AV: Improving Audio-Visual Learning via Cross-modal Interactive Enrichment
by: Hu, Yunzuo, et al.
Published: (2026)
by: Hu, Yunzuo, et al.
Published: (2026)
Tuning-free Visual Effect Transfer across Videos
by: Jones, Maxwell, et al.
Published: (2026)
by: Jones, Maxwell, et al.
Published: (2026)
MotionAdapter: Video Motion Transfer via Content-Aware Attention Customization
by: Zhang, Zhexin, et al.
Published: (2026)
by: Zhang, Zhexin, et al.
Published: (2026)
SAILS: Segment Anything with Incrementally Learned Semantics for Task-Invariant and Training-Free Continual Learning
by: Muralidhara, Shishir, et al.
Published: (2026)
by: Muralidhara, Shishir, et al.
Published: (2026)
SegVG: Transferring Object Bounding Box to Segmentation for Visual Grounding
by: Kang, Weitai, et al.
Published: (2024)
by: Kang, Weitai, et al.
Published: (2024)
Unsupervised Modality-Transferable Video Highlight Detection with Representation Activation Sequence Learning
by: Li, Tingtian, et al.
Published: (2024)
by: Li, Tingtian, et al.
Published: (2024)
Deep Learning for Sports Video Event Detection: Tasks, Datasets, Methods, and Challenges
by: Xu, Hao, et al.
Published: (2025)
by: Xu, Hao, et al.
Published: (2025)
Data Quality Aware Approaches for Addressing Model Drift of Semantic Segmentation Models
by: Mirza, Samiha, et al.
Published: (2024)
by: Mirza, Samiha, et al.
Published: (2024)
Unsupervised Video Highlight Detection by Learning from Audio and Visual Recurrence
by: Islam, Zahidul, et al.
Published: (2024)
by: Islam, Zahidul, et al.
Published: (2024)
Video Repurposing from User Generated Content: A Large-scale Dataset and Benchmark
by: Wu, Yongliang, et al.
Published: (2024)
by: Wu, Yongliang, et al.
Published: (2024)
CoVA: Text-Guided Composed Video Retrieval for Audio-Visual Content
by: Han, Gyuwon, et al.
Published: (2026)
by: Han, Gyuwon, et al.
Published: (2026)
Interpretable Deep Transfer Learning for Breast Ultrasound Cancer Detection: A Multi-Dataset Study
by: Abbadi, Mohammad, et al.
Published: (2025)
by: Abbadi, Mohammad, et al.
Published: (2025)
Fall Detection from Indoor Videos using MediaPipe and Handcrafted Feature
by: Ahmed, Fatima, et al.
Published: (2025)
by: Ahmed, Fatima, et al.
Published: (2025)
BiCLIP: Domain Canonicalization via Structured Geometric Transformation
by: Mantini, Pranav, et al.
Published: (2026)
by: Mantini, Pranav, et al.
Published: (2026)
Inversion-Free Video Style Transfer with Trajectory Reset Attention Control and Content-Style Bridging
by: Lin, Jiang, et al.
Published: (2025)
by: Lin, Jiang, et al.
Published: (2025)
Enhancing Multimodal Large Language Models with Multi-instance Visual Prompt Generator for Visual Representation Enrichment
by: Zhong, Wenliang, et al.
Published: (2024)
by: Zhong, Wenliang, et al.
Published: (2024)
Bridging Annotation Gaps: Transferring Labels to Align Object Detection Datasets
by: Kennerley, Mikhail, et al.
Published: (2025)
by: Kennerley, Mikhail, et al.
Published: (2025)
Learning Skill-Attributes for Transferable Assessment in Video
by: Ashutosh, Kumar, et al.
Published: (2025)
by: Ashutosh, Kumar, et al.
Published: (2025)
ContextFlow: Training-Free Video Object Editing via Adaptive Context Enrichment
by: Chen, Yiyang, et al.
Published: (2025)
by: Chen, Yiyang, et al.
Published: (2025)
MV-Adapter: Multimodal Video Transfer Learning for Video Text Retrieval
by: Jin, Xiaojie, et al.
Published: (2023)
by: Jin, Xiaojie, et al.
Published: (2023)
Learning Transferable Temporal Primitives for Video Reasoning via Synthetic Videos
by: Jiang, Songtao, et al.
Published: (2026)
by: Jiang, Songtao, et al.
Published: (2026)
MoTE: Reconciling Generalization with Specialization for Visual-Language to Video Knowledge Transfer
by: Zhu, Minghao, et al.
Published: (2024)
by: Zhu, Minghao, et al.
Published: (2024)
BVI-Artefact: An Artefact Detection Benchmark Dataset for Streamed Videos
by: Feng, Chen, et al.
Published: (2023)
by: Feng, Chen, et al.
Published: (2023)
Context-aware Video Anomaly Detection in Long-Term Datasets
by: Yang, Zhengye, et al.
Published: (2024)
by: Yang, Zhengye, et al.
Published: (2024)
XS-VID: An Extremely Small Video Object Detection Dataset
by: Guo, Jiahao, et al.
Published: (2024)
by: Guo, Jiahao, et al.
Published: (2024)
Class Prototypes based Contrastive Learning for Classifying Multi-Label and Fine-Grained Educational Videos
by: Gupta, Rohit, et al.
Published: (2025)
by: Gupta, Rohit, et al.
Published: (2025)
Simple Visual Artifact Detection in Sora-Generated Videos
by: Sugiyama, Misora, et al.
Published: (2025)
by: Sugiyama, Misora, et al.
Published: (2025)
VideoWorld 2: Learning Transferable Knowledge from Real-world Videos
by: Ren, Zhongwei, et al.
Published: (2026)
by: Ren, Zhongwei, et al.
Published: (2026)
Modality-Incremental Learning with Disjoint Relevance Mapping Networks for Image-based Semantic Segmentation
by: Hegde, Niharika, et al.
Published: (2024)
by: Hegde, Niharika, et al.
Published: (2024)
Efficient Transfer Learning for Video-language Foundation Models
by: Chen, Haoxing, et al.
Published: (2024)
by: Chen, Haoxing, et al.
Published: (2024)
Dual-Perspective Knowledge Enrichment for Semi-Supervised 3D Object Detection
by: Han, Yucheng, et al.
Published: (2024)
by: Han, Yucheng, et al.
Published: (2024)
FADE: A Dataset for Detecting Falling Objects around Buildings in Video
by: Tu, Zhigang, et al.
Published: (2024)
by: Tu, Zhigang, et al.
Published: (2024)
Visualizing Celebrity Dynamics in Video Content: A Proposed Approach Using Face Recognition Timestamp Data
by: Demir, Doğanay, et al.
Published: (2025)
by: Demir, Doğanay, et al.
Published: (2025)
Similar Items
-
Lecture Video Visual Objects (LVVO) Dataset: A Benchmark for Visual Object Detection in Educational Videos
by: Biswas, Dipayan, et al.
Published: (2025) -
Attention-based Shape and Gait Representations Learning for Video-based Cloth-Changing Person Re-Identification
by: Nguyen, Vuong D., et al.
Published: (2024) -
GeoStack: A Framework for Quasi-Abelian Knowledge Composition in VLMs
by: Mantini, Pranav, et al.
Published: (2026) -
Generated Contents Enrichment
by: Naseri, Mahdi, et al.
Published: (2024) -
CCPA: Long-term Person Re-Identification via Contrastive Clothing and Pose Augmentation
by: Nguyen, Vuong D., et al.
Published: (2024)