Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Yogesh, Mishra, Anand |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mutually-Aware Feature Learning for Few-Shot Object Counting
by: Jeon, Yerim, et al.
Published: (2024)
by: Jeon, Yerim, et al.
Published: (2024)
CDFormer: Cross-Domain Few-Shot Object Detection Transformer Against Feature Confusion
by: Meng, Boyuan, et al.
Published: (2025)
by: Meng, Boyuan, et al.
Published: (2025)
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
by: Li, Wenxi, et al.
Published: (2025)
by: Li, Wenxi, et al.
Published: (2025)
Context-Aware Temporal Embedding of Objects in Video Data
by: Farhan, Ahnaf, et al.
Published: (2024)
by: Farhan, Ahnaf, et al.
Published: (2024)
Re-Scoring Using Image-Language Similarity for Few-Shot Object Detection
by: Jung, Min Jae, et al.
Published: (2023)
by: Jung, Min Jae, et al.
Published: (2023)
NTIRE 2025 Challenge on Cross-Domain Few-Shot Object Detection: Methods and Results
by: Fu, Yuqian, et al.
Published: (2025)
by: Fu, Yuqian, et al.
Published: (2025)
The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results
by: Qiu, Xingyu, et al.
Published: (2026)
by: Qiu, Xingyu, et al.
Published: (2026)
Eyes on Target: Gaze-Aware Object Detection in Egocentric Video
by: Lall, Vishakha, et al.
Published: (2025)
by: Lall, Vishakha, et al.
Published: (2025)
Hybrid Spiking Vision Transformer for Object Detection with Event Cameras
by: Xu, Qi, et al.
Published: (2025)
by: Xu, Qi, et al.
Published: (2025)
Analyzing the Impact of Low-Rank Adaptation for Cross-Domain Few-Shot Object Detection in Aerial Images
by: Talaoubrid, Hicham, et al.
Published: (2025)
by: Talaoubrid, Hicham, et al.
Published: (2025)
Enhance Then Search: An Augmentation-Search Strategy with Foundation Models for Cross-Domain Few-Shot Object Detection
by: Pan, Jiancheng, et al.
Published: (2025)
by: Pan, Jiancheng, et al.
Published: (2025)
GiPL: Generative augmented iterative Pseudo-Labeling for Cross-Domain Few-Shot Object Detection
by: Liu, Jiacong, et al.
Published: (2026)
by: Liu, Jiacong, et al.
Published: (2026)
Object-Shot Enhanced Grounding Network for Egocentric Video
by: Feng, Yisen, et al.
Published: (2025)
by: Feng, Yisen, et al.
Published: (2025)
Evaluating the Energy Efficiency of Few-Shot Learning for Object Detection in Industrial Settings
by: Tsoumplekas, Georgios, et al.
Published: (2024)
by: Tsoumplekas, Georgios, et al.
Published: (2024)
Few-Shot LoRA Adaptation of a Flow-Matching Foundation Model for Cross-Spectral Object Detection
by: Clouser, Maxim, et al.
Published: (2026)
by: Clouser, Maxim, et al.
Published: (2026)
TGBFormer: Transformer-GraphFormer Blender Network for Video Object Detection
by: Qi, Qiang, et al.
Published: (2025)
by: Qi, Qiang, et al.
Published: (2025)
VideoLLM Benchmarks and Evaluation: A Survey
by: Kumar, Yogesh
Published: (2025)
by: Kumar, Yogesh
Published: (2025)
Dynamic Object Queries for Transformer-based Incremental Object Detection
by: Zhang, Jichuan, et al.
Published: (2024)
by: Zhang, Jichuan, et al.
Published: (2024)
Small Object Few-shot Segmentation for Vision-based Industrial Inspection
by: Zhang, Zilong, et al.
Published: (2024)
by: Zhang, Zilong, et al.
Published: (2024)
Object-conditioned Bag of Instances for Few-Shot Personalized Instance Recognition
by: Michieli, Umberto, et al.
Published: (2024)
by: Michieli, Umberto, et al.
Published: (2024)
SAM-PM: Enhancing Video Camouflaged Object Detection using Spatio-Temporal Attention
by: Meeran, Muhammad Nawfal, et al.
Published: (2024)
by: Meeran, Muhammad Nawfal, et al.
Published: (2024)
Source-Free Object Detection with Detection Transformer
by: Yao, Huizai, et al.
Published: (2025)
by: Yao, Huizai, et al.
Published: (2025)
Knowledge Amalgamation for Object Detection with Transformers
by: Zhang, Haofei, et al.
Published: (2022)
by: Zhang, Haofei, et al.
Published: (2022)
TinyVLM: Zero-Shot Object Detection on Microcontrollers via Vision-Language Distillation with Matryoshka Embeddings
by: Wilson, Bibin
Published: (2026)
by: Wilson, Bibin
Published: (2026)
Object Aware Egocentric Online Action Detection
by: An, Joungbin, et al.
Published: (2024)
by: An, Joungbin, et al.
Published: (2024)
On Moving Object Segmentation from Monocular Video with Transformers
by: Homeyer, Christian, et al.
Published: (2024)
by: Homeyer, Christian, et al.
Published: (2024)
TinyViT-Batten: Few-Shot Vision Transformer with Explainable Attention for Early Batten-Disease Detection on Pediatric MRI
by: Uppalapati, Khartik, et al.
Published: (2025)
by: Uppalapati, Khartik, et al.
Published: (2025)
Object Detection for Vehicle Dashcams using Transformers
by: Mustafa, Osama, et al.
Published: (2024)
by: Mustafa, Osama, et al.
Published: (2024)
Topology-Aware CLIP Few-Shot Learning
by: Huang, Dazhi
Published: (2025)
by: Huang, Dazhi
Published: (2025)
Cross-domain Few-shot Object Detection with Multi-modal Textual Enrichment
by: Shangguan, Zeyu, et al.
Published: (2025)
by: Shangguan, Zeyu, et al.
Published: (2025)
EfficientFSL: Enhancing Few-Shot Classification via Query-Only Tuning in Vision Transformers
by: Liao, Wenwen, et al.
Published: (2026)
by: Liao, Wenwen, et al.
Published: (2026)
Spatial-Frequency Aware for Object Detection in RAW Image
by: Ye, Zhuohua, et al.
Published: (2025)
by: Ye, Zhuohua, et al.
Published: (2025)
Trajectory-Aware Adaptive Inference in Object Detection Models
by: Papanikolaou, Grigorios, et al.
Published: (2026)
by: Papanikolaou, Grigorios, et al.
Published: (2026)
Towards Efficient and General-Purpose Few-Shot Misclassification Detection for Vision-Language Models
by: Zeng, Fanhu, et al.
Published: (2025)
by: Zeng, Fanhu, et al.
Published: (2025)
Hands-on Evaluation of Visual Transformers for Object Recognition and Detection
by: Vlachogiannis, Dimitrios N., et al.
Published: (2025)
by: Vlachogiannis, Dimitrios N., et al.
Published: (2025)
Mask-RadarNet: Enhancing Transformer With Spatial-Temporal Semantic Context for Radar Object Detection in Autonomous Driving
by: Wu, Yuzhi, et al.
Published: (2024)
by: Wu, Yuzhi, et al.
Published: (2024)
Stability Plasticity Decoupled Fine-tuning For Few-shot end-to-end Object Detection
by: Yin, Yuantao, et al.
Published: (2024)
by: Yin, Yuantao, et al.
Published: (2024)
Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving Safety
by: Shriram, Shashank, et al.
Published: (2025)
by: Shriram, Shashank, et al.
Published: (2025)
Improving Zero-Shot Object-Level Change Detection by Incorporating Visual Correspondence
by: Nguyen, Hung Huy, et al.
Published: (2025)
by: Nguyen, Hung Huy, et al.
Published: (2025)
Beyond the Doors of Perception: Vision Transformers Represent Relations Between Objects
by: Lepori, Michael A., et al.
Published: (2024)
by: Lepori, Michael A., et al.
Published: (2024)
Similar Items
-
Mutually-Aware Feature Learning for Few-Shot Object Counting
by: Jeon, Yerim, et al.
Published: (2024) -
CDFormer: Cross-Domain Few-Shot Object Detection Transformer Against Feature Confusion
by: Meng, Boyuan, et al.
Published: (2025) -
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
by: Li, Wenxi, et al.
Published: (2025) -
Context-Aware Temporal Embedding of Objects in Video Data
by: Farhan, Ahnaf, et al.
Published: (2024) -
Re-Scoring Using Image-Language Similarity for Few-Shot Object Detection
by: Jung, Min Jae, et al.
Published: (2023)