VVitCutLER: Towards Unsupervised Object Detection and Segmentation in Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Zhijing, Hashmi, Khurram Azeem, Stricker, Didier, Afzal, Muhammad Zeshan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Boxes: Mask-Guided Spatio-Temporal Feature Aggregation for Video Object Detection
by: Hashmi, Khurram Azeem, et al.
Published: (2024)
by: Hashmi, Khurram Azeem, et al.
Published: (2024)
Sparse Semi-DETR: Sparse Learnable Queries for Semi-Supervised Object Detection
by: Shehzadi, Tahira, et al.
Published: (2024)
by: Shehzadi, Tahira, et al.
Published: (2024)
UnSupDLA: Towards Unsupervised Document Layout Analysis
by: Sheikh, Talha Uddin, et al.
Published: (2024)
by: Sheikh, Talha Uddin, et al.
Published: (2024)
Semi-Supervised Object Detection: A Survey on Progress from CNN to Transformer
by: Shehzadi, Tahira, et al.
Published: (2024)
by: Shehzadi, Tahira, et al.
Published: (2024)
End-to-End Semi-Supervised approach with Modulated Object Queries for Table Detection in Documents
by: Ehsan, Iqraa, et al.
Published: (2024)
by: Ehsan, Iqraa, et al.
Published: (2024)
Towards End-to-End Semi-Supervised Table Detection with Semantic Aligned Matching Transformer
by: Shehzadi, Tahira, et al.
Published: (2024)
by: Shehzadi, Tahira, et al.
Published: (2024)
A Hybrid Approach for Document Layout Analysis in Document images
by: Shehzadi, Tahira, et al.
Published: (2024)
by: Shehzadi, Tahira, et al.
Published: (2024)
SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion
by: Nazir, Danish, et al.
Published: (2022)
by: Nazir, Danish, et al.
Published: (2022)
SituationalLLM: Proactive language models with scene awareness for dynamic, contextual task guidance
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
NURBGen: High-Fidelity Text-to-CAD Generation through LLM-Driven NURBS Modeling
by: Usama, Muhammad, et al.
Published: (2025)
by: Usama, Muhammad, et al.
Published: (2025)
Estimating Human Poses Across Datasets: A Unified Skeleton and Multi-Teacher Distillation Approach
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
Enhanced Bank Check Security: Introducing a Novel Dataset and Transformer-Based Approach for Detection and Verification
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
ReConText3D: Replay-based Continual Text-to-3D Generation
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
Shape2.5D: A Dataset of Texture-less Surfaces for Depth and Normals Estimation
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
CICA: Content-Injected Contrastive Alignment for Zero-Shot Document Image Classification
by: Sinha, Sankalp, et al.
Published: (2024)
by: Sinha, Sankalp, et al.
Published: (2024)
Classroom-Inspired Multi-Mentor Distillation with Adaptive Learning Strategies
by: Sarode, Shalini, et al.
Published: (2024)
by: Sarode, Shalini, et al.
Published: (2024)
Human Pose Descriptions and Subject-Focused Attention for Improved Zero-Shot Transfer in Human-Centric Classification Tasks
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
Text2CAD: Generating Sequential CAD Models from Beginner-to-Expert Level Text Prompts
by: Khan, Mohammad Sadil, et al.
Published: (2024)
by: Khan, Mohammad Sadil, et al.
Published: (2024)
Object-Centric 2D Gaussian Splatting: Background Removal and Occlusion-Aware Pruning for Compact Object Models
by: Rogge, Marcel, et al.
Published: (2025)
by: Rogge, Marcel, et al.
Published: (2025)
DreamCAD: Scaling Multi-modal CAD Generation using Differentiable Parametric Surfaces
by: Khan, Mohammad Sadil, et al.
Published: (2026)
by: Khan, Mohammad Sadil, et al.
Published: (2026)
Sensor Generalization for Adaptive Sensing in Event-based Object Detection via Joint Distribution Training
by: Saha, Aheli, et al.
Published: (2026)
by: Saha, Aheli, et al.
Published: (2026)
Towards Unconstrained 2D Pose Estimation of the Human Spine
by: Khan, Muhammad Saif Ullah, et al.
Published: (2025)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2025)
Small Object Detection with YOLO: A Performance Analysis Across Model Versions and Hardware
by: Tariq, Muhammad Fasih, et al.
Published: (2025)
by: Tariq, Muhammad Fasih, et al.
Published: (2025)
Jointly Learning Spatial, Angular, and Temporal Information for Enhanced Lane Detection
by: Alam, Muhammad Zeshan
Published: (2024)
by: Alam, Muhammad Zeshan
Published: (2024)
Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation
by: Zhuge, Yunzhi, et al.
Published: (2025)
by: Zhuge, Yunzhi, et al.
Published: (2025)
MARVEL-40M+: Multi-Level Visual Elaboration for High-Fidelity Text-to-3D Content Creation
by: Sinha, Sankalp, et al.
Published: (2024)
by: Sinha, Sankalp, et al.
Published: (2024)
DriverGaze360: OmniDirectional Driver Attention with Object-Level Guidance
by: Govil, Shreedhar, et al.
Published: (2025)
by: Govil, Shreedhar, et al.
Published: (2025)
Domain-Incremental Semantic Segmentation for Autonomous Driving under Adverse Driving Conditions
by: Muralidhara, Shishir, et al.
Published: (2025)
by: Muralidhara, Shishir, et al.
Published: (2025)
SIMSPINE: A Biomechanics-Aware Simulation Framework for 3D Spine Motion Annotation and Benchmarking
by: Khan, Muhammad Saif Ullah, et al.
Published: (2026)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2026)
PoseAdapt: Sustainable Human Pose Estimation via Continual Learning Benchmarks and Toolkit
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
MILE: Mixture of Incremental LoRA Experts for Continual Semantic Segmentation across Domains and Modalities
by: Muralidhara, Shishir, et al.
Published: (2026)
by: Muralidhara, Shishir, et al.
Published: (2026)
PanoSAMic: Panoramic Image Segmentation from SAM Feature Encoding and Dual View Fusion
by: Chamseddine, Mahdi, et al.
Published: (2026)
by: Chamseddine, Mahdi, et al.
Published: (2026)
SAILS: Segment Anything with Incrementally Learned Semantics for Task-Invariant and Training-Free Continual Learning
by: Muralidhara, Shishir, et al.
Published: (2026)
by: Muralidhara, Shishir, et al.
Published: (2026)
Unsupervised Segmentation by Diffusing, Walking and Cutting
by: Ivanova, Daniela, et al.
Published: (2024)
by: Ivanova, Daniela, et al.
Published: (2024)
Beyond Averages: Open-Vocabulary 3D Scene Understanding with Gaussian Splatting and Bag of Embeddings
by: Arafa, Abdalla, et al.
Published: (2025)
by: Arafa, Abdalla, et al.
Published: (2025)
FlowCut: Unsupervised Video Instance Segmentation via Temporal Mask Matching
by: Sari, Alp Eren, et al.
Published: (2025)
by: Sari, Alp Eren, et al.
Published: (2025)
Dual Prototype Attention for Unsupervised Video Object Segmentation
by: Cho, Suhwan, et al.
Published: (2022)
by: Cho, Suhwan, et al.
Published: (2022)
Guided Slot Attention for Unsupervised Video Object Segmentation
by: Lee, Minhyeok, et al.
Published: (2023)
by: Lee, Minhyeok, et al.
Published: (2023)
Modality-Incremental Learning with Disjoint Relevance Mapping Networks for Image-based Semantic Segmentation
by: Hegde, Niharika, et al.
Published: (2024)
by: Hegde, Niharika, et al.
Published: (2024)
Enhancing Object Discovery for Unsupervised Instance Segmentation and Object Detection
by: Feng, Xingyu, et al.
Published: (2025)
by: Feng, Xingyu, et al.
Published: (2025)
Similar Items
-
Beyond Boxes: Mask-Guided Spatio-Temporal Feature Aggregation for Video Object Detection
by: Hashmi, Khurram Azeem, et al.
Published: (2024) -
Sparse Semi-DETR: Sparse Learnable Queries for Semi-Supervised Object Detection
by: Shehzadi, Tahira, et al.
Published: (2024) -
UnSupDLA: Towards Unsupervised Document Layout Analysis
by: Sheikh, Talha Uddin, et al.
Published: (2024) -
Semi-Supervised Object Detection: A Survey on Progress from CNN to Transformer
by: Shehzadi, Tahira, et al.
Published: (2024) -
End-to-End Semi-Supervised approach with Modulated Object Queries for Table Detection in Documents
by: Ehsan, Iqraa, et al.
Published: (2024)