Semi-supervised Active Learning for Video Action Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Singh, Ayush, Rana, Aayush J, Kumar, Akash, Vyas, Shruti, Rawat, Yogesh Singh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Stable Mean Teacher for Semi-supervised Video Action Detection
por: Kumar, Akash, et al.
Publicado: (2024)
por: Kumar, Akash, et al.
Publicado: (2024)
OmViD: Omni-supervised active learning for video action detection
por: Rana, Aayush, et al.
Publicado: (2025)
por: Rana, Aayush, et al.
Publicado: (2025)
MolVision: Molecular Property Prediction with Vision Language Models
por: Adak, Deepan, et al.
Publicado: (2025)
por: Adak, Deepan, et al.
Publicado: (2025)
iSafetyBench: A video-language benchmark for safety in industrial environment
por: Abdullah, Raiyaan, et al.
Publicado: (2025)
por: Abdullah, Raiyaan, et al.
Publicado: (2025)
Advancing Automatic Photovoltaic Defect Detection using Semi-Supervised Semantic Segmentation of Electroluminescence Images
por: Jha, Abhishek, et al.
Publicado: (2024)
por: Jha, Abhishek, et al.
Publicado: (2024)
Contextual Self-paced Learning for Weakly Supervised Spatio-Temporal Video Grounding
por: Kumar, Akash, et al.
Publicado: (2025)
por: Kumar, Akash, et al.
Publicado: (2025)
Scaling Open-Vocabulary Action Detection
por: Sia, Zhen Hao, et al.
Publicado: (2025)
por: Sia, Zhen Hao, et al.
Publicado: (2025)
On Occlusions in Video Action Detection: Benchmark Datasets And Training Recipes
por: Modi, Rajat, et al.
Publicado: (2024)
por: Modi, Rajat, et al.
Publicado: (2024)
LR0.FM: Low-Res Benchmark and Improving Robustness for Zero-Shot Classification in Foundation Models
por: Pathak, Priyank, et al.
Publicado: (2025)
por: Pathak, Priyank, et al.
Publicado: (2025)
A Large-Scale Analysis on Contextual Self-Supervised Video Representation Learning
por: Kumar, Akash, et al.
Publicado: (2025)
por: Kumar, Akash, et al.
Publicado: (2025)
MolSight: Molecular Property Prediction with Images
por: Baranwal, Aaditya, et al.
Publicado: (2026)
por: Baranwal, Aaditya, et al.
Publicado: (2026)
STPro: Spatial and Temporal Progressive Learning for Weakly Supervised Spatio-Temporal Grounding
por: Garg, Aaryan, et al.
Publicado: (2025)
por: Garg, Aaryan, et al.
Publicado: (2025)
StreamReady: Learning What to Answer and When in Long Streaming Videos
por: Azad, Shehreen, et al.
Publicado: (2026)
por: Azad, Shehreen, et al.
Publicado: (2026)
Re:Verse -- Can Your VLM Read a Manga?
por: Baranwal, Aaditya, et al.
Publicado: (2025)
por: Baranwal, Aaditya, et al.
Publicado: (2025)
HierarQ: Task-Aware Hierarchical Q-Former for Enhanced Video Understanding
por: Azad, Shehreen, et al.
Publicado: (2025)
por: Azad, Shehreen, et al.
Publicado: (2025)
Activity-Biometrics: Person Identification from Daily Activities
por: Azad, Shehreen, et al.
Publicado: (2024)
por: Azad, Shehreen, et al.
Publicado: (2024)
EZ-CLIP: Efficient Zeroshot Video Action Recognition
por: Ahmad, Shahzad, et al.
Publicado: (2023)
por: Ahmad, Shahzad, et al.
Publicado: (2023)
Asynchronous Perception Machine For Efficient Test-Time-Training
por: Modi, Rajat, et al.
Publicado: (2024)
por: Modi, Rajat, et al.
Publicado: (2024)
CoSPlan: Corrective Sequential Planning via Scene Graph Incremental Updates
por: Grover, Shresth, et al.
Publicado: (2025)
por: Grover, Shresth, et al.
Publicado: (2025)
RobustGait: Robustness Analysis for Appearance Based Gait Recognition
por: Sayera, Reeshoon, et al.
Publicado: (2025)
por: Sayera, Reeshoon, et al.
Publicado: (2025)
Semi-supervised Open-World Object Detection
por: Mullappilly, Sahal Shaji, et al.
Publicado: (2024)
por: Mullappilly, Sahal Shaji, et al.
Publicado: (2024)
Dual Guidance Semi-Supervised Action Detection
por: Singh, Ankit, et al.
Publicado: (2025)
por: Singh, Ankit, et al.
Publicado: (2025)
VISTA: Video Interaction Spatio-Temporal Analysis Benchmark
por: Aparcedo, Alejandro, et al.
Publicado: (2026)
por: Aparcedo, Alejandro, et al.
Publicado: (2026)
Bridging Foundation Models and ASTM Metallurgical Standards for Automated Grain Size Estimation from Microscopy Images
por: Mueez, Abdul, et al.
Publicado: (2026)
por: Mueez, Abdul, et al.
Publicado: (2026)
ProDiG: Progressive Diffusion-Guided Gaussian Splatting for Aerial to Ground Reconstruction
por: Mitra, Sirshapan, et al.
Publicado: (2026)
por: Mitra, Sirshapan, et al.
Publicado: (2026)
Colors See Colors Ignore: Clothes Changing ReID with Color Disentanglement
por: Pathak, Priyank, et al.
Publicado: (2025)
por: Pathak, Priyank, et al.
Publicado: (2025)
DIFFER: Disentangling Identity Features via Semantic Cues for Clothes-Changing Person Re-ID
por: Liang, Xin, et al.
Publicado: (2025)
por: Liang, Xin, et al.
Publicado: (2025)
DisenQ: Disentangling Q-Former for Activity-Biometrics
por: Azad, Shehreen, et al.
Publicado: (2025)
por: Azad, Shehreen, et al.
Publicado: (2025)
Coarse Attribute Prediction with Task Agnostic Distillation for Real World Clothes Changing ReID
por: Pathak, Priyank, et al.
Publicado: (2025)
por: Pathak, Priyank, et al.
Publicado: (2025)
ALFred: An Active Learning Framework for Real-world Semi-supervised Anomaly Detection with Adaptive Thresholds
por: Yao, Shanle, et al.
Publicado: (2025)
por: Yao, Shanle, et al.
Publicado: (2025)
Language-Guided Temporal Token Pruning for Efficient VideoLLM Processing
por: Kumar, Yogesh
Publicado: (2025)
por: Kumar, Yogesh
Publicado: (2025)
Streamlining Video Analysis for Efficient Violence Detection
por: Pathak, Gourang, et al.
Publicado: (2024)
por: Pathak, Gourang, et al.
Publicado: (2024)
SITAR: Semi-supervised Image Transformer for Action Recognition
por: Iqbal, Owais, et al.
Publicado: (2024)
por: Iqbal, Owais, et al.
Publicado: (2024)
ACTRESS: Active Retraining for Semi-supervised Visual Grounding
por: Kang, Weitai, et al.
Publicado: (2024)
por: Kang, Weitai, et al.
Publicado: (2024)
Collaboratively Self-supervised Video Representation Learning for Action Recognition
por: Zhang, Jie, et al.
Publicado: (2024)
por: Zhang, Jie, et al.
Publicado: (2024)
Foundation Models for Video Understanding: A Survey
por: Madan, Neelu, et al.
Publicado: (2024)
por: Madan, Neelu, et al.
Publicado: (2024)
GaitCrafter: Diffusion Model for Biometric Preserving Gait Synthesis
por: Mitra, Sirshapan, et al.
Publicado: (2025)
por: Mitra, Sirshapan, et al.
Publicado: (2025)
Navigating Hallucinations for Reasoning of Unintentional Activities
por: Grover, Shresth, et al.
Publicado: (2024)
por: Grover, Shresth, et al.
Publicado: (2024)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
por: Kumar, Yogesh, et al.
Publicado: (2025)
por: Kumar, Yogesh, et al.
Publicado: (2025)
VideoLLM Benchmarks and Evaluation: A Survey
por: Kumar, Yogesh
Publicado: (2025)
por: Kumar, Yogesh
Publicado: (2025)
Ejemplares similares
-
Stable Mean Teacher for Semi-supervised Video Action Detection
por: Kumar, Akash, et al.
Publicado: (2024) -
OmViD: Omni-supervised active learning for video action detection
por: Rana, Aayush, et al.
Publicado: (2025) -
MolVision: Molecular Property Prediction with Vision Language Models
por: Adak, Deepan, et al.
Publicado: (2025) -
iSafetyBench: A video-language benchmark for safety in industrial environment
por: Abdullah, Raiyaan, et al.
Publicado: (2025) -
Advancing Automatic Photovoltaic Defect Detection using Semi-Supervised Semantic Segmentation of Electroluminescence Images
por: Jha, Abhishek, et al.
Publicado: (2024)