MaskVD: Region Masking for Efficient Video Object Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Sarkar, Sreetama, Datta, Gourav, Kundu, Souvik, Zheng, Kai, Bhattacharyya, Chirayata, Beerel, Peter A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Block Selective Reprogramming for On-device Training of Vision Transformers
por: Sarkar, Sreetama, et al.
Publicado: (2024)
por: Sarkar, Sreetama, et al.
Publicado: (2024)
Region Masking to Accelerate Video Processing on Neuromorphic Hardware
por: Sarkar, Sreetama, et al.
Publicado: (2025)
por: Sarkar, Sreetama, et al.
Publicado: (2025)
Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression
por: Sarkar, Sreetama, et al.
Publicado: (2025)
por: Sarkar, Sreetama, et al.
Publicado: (2025)
RedVTP: Training-Free Acceleration of Diffusion Vision-Language Models Inference via Masked Token-Guided Visual Token Pruning
por: Xu, Jingqi, et al.
Publicado: (2025)
por: Xu, Jingqi, et al.
Publicado: (2025)
Linearizing Models for Efficient yet Robust Private Inference
por: Sarkar, Sreetama, et al.
Publicado: (2024)
por: Sarkar, Sreetama, et al.
Publicado: (2024)
Energy-Efficient & Real-Time Computer Vision with Intelligent Skipping via Reconfigurable CMOS Image Sensors
por: Kaiser, Md Abdullah-Al, et al.
Publicado: (2024)
por: Kaiser, Md Abdullah-Al, et al.
Publicado: (2024)
EVEREST: Efficient Masked Video Autoencoder by Removing Redundant Spatiotemporal Tokens
por: Hwang, Sunil, et al.
Publicado: (2022)
por: Hwang, Sunil, et al.
Publicado: (2022)
Self-Distilled Masked Auto-Encoders are Efficient Video Anomaly Detectors
por: Ristea, Nicolae-Catalin, et al.
Publicado: (2023)
por: Ristea, Nicolae-Catalin, et al.
Publicado: (2023)
E-MD3C: Taming Masked Diffusion Transformers for Efficient Zero-Shot Object Customization
por: Pham, Trung X., et al.
Publicado: (2025)
por: Pham, Trung X., et al.
Publicado: (2025)
General and Efficient Visual Goal-Conditioned Reinforcement Learning using Object-Agnostic Masks
por: Shahriar, Fahim, et al.
Publicado: (2025)
por: Shahriar, Fahim, et al.
Publicado: (2025)
Two-Step Data Augmentation for Masked Face Detection and Recognition: Turning Fake Masks to Real
por: Yang, Yan, et al.
Publicado: (2025)
por: Yang, Yan, et al.
Publicado: (2025)
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
por: Yariv, Guy, et al.
Publicado: (2025)
por: Yariv, Guy, et al.
Publicado: (2025)
PEEKABOO: Interactive Video Generation via Masked-Diffusion
por: Jain, Yash, et al.
Publicado: (2023)
por: Jain, Yash, et al.
Publicado: (2023)
Effective and Efficient Masked Image Generation Models
por: You, Zebin, et al.
Publicado: (2025)
por: You, Zebin, et al.
Publicado: (2025)
HIVTP: A Training-Free Method to Improve VLMs Efficiency via Hierarchical Visual Token Pruning Using Middle-Layer-Based Importance Score
por: Xu, Jingqi, et al.
Publicado: (2025)
por: Xu, Jingqi, et al.
Publicado: (2025)
Mask and Restore: Blind Backdoor Defense at Test Time with Masked Autoencoder
por: Sun, Tao, et al.
Publicado: (2023)
por: Sun, Tao, et al.
Publicado: (2023)
MaskAttn-SDXL: Controllable Region-Level Text-To-Image Generation
por: Chang, Yu, et al.
Publicado: (2025)
por: Chang, Yu, et al.
Publicado: (2025)
Masked Vector Quantization
por: Nguyen, David D., et al.
Publicado: (2023)
por: Nguyen, David D., et al.
Publicado: (2023)
SupMAE: Supervised Masked Autoencoders Are Efficient Vision Learners
por: Liang, Feng, et al.
Publicado: (2022)
por: Liang, Feng, et al.
Publicado: (2022)
The Role of Masking for Efficient Supervised Knowledge Distillation of Vision Transformers
por: Son, Seungwoo, et al.
Publicado: (2023)
por: Son, Seungwoo, et al.
Publicado: (2023)
IterMask3D: Unsupervised Anomaly Detection and Segmentation with Test-Time Iterative Mask Refinement in 3D Brain MR
por: Liang, Ziyun, et al.
Publicado: (2025)
por: Liang, Ziyun, et al.
Publicado: (2025)
Pack and Detect: Fast Object Detection in Videos Using Region-of-Interest Packing
por: Kumar, Athindran Ramesh, et al.
Publicado: (2018)
por: Kumar, Athindran Ramesh, et al.
Publicado: (2018)
Downstream Task Guided Masking Learning in Masked Autoencoders Using Multi-Level Optimization
por: Guo, Han, et al.
Publicado: (2024)
por: Guo, Han, et al.
Publicado: (2024)
Click2Mask: Local Editing with Dynamic Mask Generation
por: Regev, Omer, et al.
Publicado: (2024)
por: Regev, Omer, et al.
Publicado: (2024)
Combating Noisy Labels via Dynamic Connection Masking
por: Zhang, Xinlei, et al.
Publicado: (2025)
por: Zhang, Xinlei, et al.
Publicado: (2025)
Unsupervised Video Domain Adaptation with Masked Pre-Training and Collaborative Self-Training
por: Reddy, Arun, et al.
Publicado: (2023)
por: Reddy, Arun, et al.
Publicado: (2023)
MIMIC: Masked Image Modeling with Image Correspondences
por: Marathe, Kalyani, et al.
Publicado: (2023)
por: Marathe, Kalyani, et al.
Publicado: (2023)
MaskMedPaint: Masked Medical Image Inpainting with Diffusion Models for Mitigation of Spurious Correlations
por: Jin, Qixuan, et al.
Publicado: (2024)
por: Jin, Qixuan, et al.
Publicado: (2024)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
por: Pramanik, Rishav, et al.
Publicado: (2024)
por: Pramanik, Rishav, et al.
Publicado: (2024)
Learning Real-World Action-Video Dynamics with Heterogeneous Masked Autoregression
por: Wang, Lirui, et al.
Publicado: (2025)
por: Wang, Lirui, et al.
Publicado: (2025)
Mamba-3D as Masked Autoencoders for Accurate and Data-Efficient Analysis of Medical Ultrasound Videos
por: Zhou, Jiaheng, et al.
Publicado: (2025)
por: Zhou, Jiaheng, et al.
Publicado: (2025)
Self-Masking Networks for Unsupervised Adaptation
por: Warmerdam, Alfonso Taboada, et al.
Publicado: (2024)
por: Warmerdam, Alfonso Taboada, et al.
Publicado: (2024)
Keypoint Aware Masked Image Modelling
por: Krishna, Madhava, et al.
Publicado: (2024)
por: Krishna, Madhava, et al.
Publicado: (2024)
Robust Representation Learning in Masked Autoencoders
por: Shrivastava, Anika, et al.
Publicado: (2026)
por: Shrivastava, Anika, et al.
Publicado: (2026)
Masked Conditioning for Deep Generative Models
por: Mueller, Phillip, et al.
Publicado: (2025)
por: Mueller, Phillip, et al.
Publicado: (2025)
MixMask: Revisiting Masking Strategy for Siamese ConvNets
por: Vishniakov, Kirill, et al.
Publicado: (2022)
por: Vishniakov, Kirill, et al.
Publicado: (2022)
CanvasMAR: Improving Masked Autoregressive Video Prediction With Canvas
por: Li, Zian, et al.
Publicado: (2025)
por: Li, Zian, et al.
Publicado: (2025)
Task-customized Masked AutoEncoder via Mixture of Cluster-conditional Experts
por: Liu, Zhili, et al.
Publicado: (2024)
por: Liu, Zhili, et al.
Publicado: (2024)
Instruction-Guided Visual Masking
por: Zheng, Jinliang, et al.
Publicado: (2024)
por: Zheng, Jinliang, et al.
Publicado: (2024)
MCM: Multi-layer Concept Map for Efficient Concept Learning from Masked Images
por: Sun, Yuwei, et al.
Publicado: (2025)
por: Sun, Yuwei, et al.
Publicado: (2025)
Ejemplares similares
-
Block Selective Reprogramming for On-device Training of Vision Transformers
por: Sarkar, Sreetama, et al.
Publicado: (2024) -
Region Masking to Accelerate Video Processing on Neuromorphic Hardware
por: Sarkar, Sreetama, et al.
Publicado: (2025) -
Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression
por: Sarkar, Sreetama, et al.
Publicado: (2025) -
RedVTP: Training-Free Acceleration of Diffusion Vision-Language Models Inference via Masked Token-Guided Visual Token Pruning
por: Xu, Jingqi, et al.
Publicado: (2025) -
Linearizing Models for Efficient yet Robust Private Inference
por: Sarkar, Sreetama, et al.
Publicado: (2024)