Beyond still images: Temporal features and input variance resilience
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fadaei, Amir Hosein, Dehaqani, Mohammad-Reza A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification
von: Ho, Darryl, et al.
Veröffentlicht: (2025)
von: Ho, Darryl, et al.
Veröffentlicht: (2025)
TAG-Head: Time-Aligned Graph Head for Plug-and-Play Fine-grained Action Recognition
von: Hassan, Imtiaz Ul, et al.
Veröffentlicht: (2026)
von: Hassan, Imtiaz Ul, et al.
Veröffentlicht: (2026)
High-Frequency Semantics and Geometric Priors for End-to-End Detection Transformers in Challenging UAV Imagery
von: Peng, Hongxing, et al.
Veröffentlicht: (2025)
von: Peng, Hongxing, et al.
Veröffentlicht: (2025)
PhysicsNeRF: Physics-Guided 3D Reconstruction from Sparse Views
von: Barhdadi, Mohamed Rayan, et al.
Veröffentlicht: (2025)
von: Barhdadi, Mohamed Rayan, et al.
Veröffentlicht: (2025)
Deep Learning-based Depth Estimation Methods from Monocular Image and Videos: A Comprehensive Survey
von: Rajapaksha, Uchitha, et al.
Veröffentlicht: (2024)
von: Rajapaksha, Uchitha, et al.
Veröffentlicht: (2024)
CG-HOI: Contact-Guided 3D Human-Object Interaction Generation
von: Diller, Christian, et al.
Veröffentlicht: (2023)
von: Diller, Christian, et al.
Veröffentlicht: (2023)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
Adapting SAM with Dynamic Similarity Graphs for Few-Shot Parameter-Efficient Small Dense Object Detection: A Case Study of Chickpea Pods in Field Conditions
von: Jiang, Xintong, et al.
Veröffentlicht: (2025)
von: Jiang, Xintong, et al.
Veröffentlicht: (2025)
Sign language recognition based on deep learning and low-cost handcrafted descriptors
von: Carneiro, Alvaro Leandro Cavalcante, et al.
Veröffentlicht: (2024)
von: Carneiro, Alvaro Leandro Cavalcante, et al.
Veröffentlicht: (2024)
Capacity Constraint Analysis Using Object Detection for Smart Manufacturing
von: Ahmad, Hafiz Mughees, et al.
Veröffentlicht: (2024)
von: Ahmad, Hafiz Mughees, et al.
Veröffentlicht: (2024)
SH17: A Dataset for Human Safety and Personal Protective Equipment Detection in Manufacturing Industry
von: Ahmad, Hafiz Mughees, et al.
Veröffentlicht: (2024)
von: Ahmad, Hafiz Mughees, et al.
Veröffentlicht: (2024)
FutureHuman3D: Forecasting Complex Long-Term 3D Human Behavior from Video Observations
von: Diller, Christian, et al.
Veröffentlicht: (2022)
von: Diller, Christian, et al.
Veröffentlicht: (2022)
Decoupling Vision and Language: Codebook Anchored Visual Adaptation
von: Wu, Jason, et al.
Veröffentlicht: (2026)
von: Wu, Jason, et al.
Veröffentlicht: (2026)
FlowDet: Overcoming Perspective and Scale Challenges in Real-Time End-to-End Traffic Detection
von: Wang, Zixing, et al.
Veröffentlicht: (2025)
von: Wang, Zixing, et al.
Veröffentlicht: (2025)
UGOD: Uncertainty-Guided Differentiable Opacity and Soft Dropout for Enhanced Sparse-View 3DGS
von: Guo, Zhihao, et al.
Veröffentlicht: (2025)
von: Guo, Zhihao, et al.
Veröffentlicht: (2025)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
von: He, Yuankai, et al.
Veröffentlicht: (2025)
von: He, Yuankai, et al.
Veröffentlicht: (2025)
Detecting AI-Generated Videos with Spiking Neural Networks
von: Jang, Minsuk, et al.
Veröffentlicht: (2026)
von: Jang, Minsuk, et al.
Veröffentlicht: (2026)
Butter: Frequency Consistency and Hierarchical Fusion for Autonomous Driving Object Detection
von: Lin, Xiaojian, et al.
Veröffentlicht: (2025)
von: Lin, Xiaojian, et al.
Veröffentlicht: (2025)
Scene Detection Policies and Keyframe Extraction Strategies for Large-Scale Video Analysis
von: Korolkov, Vasilii
Veröffentlicht: (2025)
von: Korolkov, Vasilii
Veröffentlicht: (2025)
Prompt Sensitivity in Vision-Language Grounding: How Small Changes in Wording Affect Object Detection
von: Deka, Dawar Jyoti, et al.
Veröffentlicht: (2026)
von: Deka, Dawar Jyoti, et al.
Veröffentlicht: (2026)
SpectralCA: Bi-Directional Cross-Attention for Next-Generation UAV Hyperspectral Vision
von: Brovko, D. V.
Veröffentlicht: (2025)
von: Brovko, D. V.
Veröffentlicht: (2025)
Object detection in adverse weather conditions for autonomous vehicles using Instruct Pix2Pix
von: Gurbindo, Unai, et al.
Veröffentlicht: (2025)
von: Gurbindo, Unai, et al.
Veröffentlicht: (2025)
Implementing Adaptations for Vision AutoRegressive Model
von: Shaikh, Kaif, et al.
Veröffentlicht: (2025)
von: Shaikh, Kaif, et al.
Veröffentlicht: (2025)
SelvaMask: Segmenting Trees in Tropical Forests and Beyond
von: Duguay, Simon-Olivier, et al.
Veröffentlicht: (2026)
von: Duguay, Simon-Olivier, et al.
Veröffentlicht: (2026)
Domain-Adaptive Pretraining Improves Primate Behavior Recognition
von: Mueller, Felix B., et al.
Veröffentlicht: (2025)
von: Mueller, Felix B., et al.
Veröffentlicht: (2025)
IMKD: Intensity-Aware Multi-Level Knowledge Distillation for Camera-Radar Fusion
von: Mishra, Shashank, et al.
Veröffentlicht: (2025)
von: Mishra, Shashank, et al.
Veröffentlicht: (2025)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
von: Wang, Yiming, et al.
Veröffentlicht: (2026)
von: Wang, Yiming, et al.
Veröffentlicht: (2026)
Mistake Attribution: Fine-Grained Mistake Understanding in Egocentric Videos
von: Li, Yayuan, et al.
Veröffentlicht: (2025)
von: Li, Yayuan, et al.
Veröffentlicht: (2025)
SelvaBox: A high-resolution dataset for tropical tree crown detection
von: Baudchon, Hugo, et al.
Veröffentlicht: (2025)
von: Baudchon, Hugo, et al.
Veröffentlicht: (2025)
DeltaVLM: Interactive Remote Sensing Image Change Analysis via Instruction-guided Difference Perception
von: Deng, Pei, et al.
Veröffentlicht: (2025)
von: Deng, Pei, et al.
Veröffentlicht: (2025)
Image-Based Leopard Seal Recognition: Approaches and Challenges in Current Automated Systems
von: Salazar, Jorge Yero, et al.
Veröffentlicht: (2024)
von: Salazar, Jorge Yero, et al.
Veröffentlicht: (2024)
Dense Motion Captioning
von: Xu, Shiyao, et al.
Veröffentlicht: (2025)
von: Xu, Shiyao, et al.
Veröffentlicht: (2025)
TD3Net: A temporal densely connected multi-dilated convolutional network for lipreading
von: Lee, Byung Hoon, et al.
Veröffentlicht: (2025)
von: Lee, Byung Hoon, et al.
Veröffentlicht: (2025)
CoMatcher: Multi-View Collaborative Feature Matching
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
MoDE: Mixture of Diffusion Experts for Any Occluded Face Recognition
von: Fan, Qiannan, et al.
Veröffentlicht: (2025)
von: Fan, Qiannan, et al.
Veröffentlicht: (2025)
NeuroGaze-Distill: Brain-informed Distillation and Depression-Inspired Geometric Priors for Robust Facial Emotion Recognition
von: Li, Zilin, et al.
Veröffentlicht: (2025)
von: Li, Zilin, et al.
Veröffentlicht: (2025)
LiftAvatar: Kinematic-Space Completion for Expression-Controlled 3D Gaussian Avatar Animation
von: Wei, Hualiang, et al.
Veröffentlicht: (2026)
von: Wei, Hualiang, et al.
Veröffentlicht: (2026)
Beyond Few-shot Object Detection: A Detailed Survey
von: Chudasama, Vishal, et al.
Veröffentlicht: (2024)
von: Chudasama, Vishal, et al.
Veröffentlicht: (2024)
FeedbackSTS-Det: Sparse Frames-Based Spatio-Temporal Semantic Feedback Network for Moving Infrared Small Target Detection
von: Huang, Yian, et al.
Veröffentlicht: (2026)
von: Huang, Yian, et al.
Veröffentlicht: (2026)
A deep learning approach to track eye movements based on events
von: Seth, Chirag, et al.
Veröffentlicht: (2025)
von: Seth, Chirag, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification
von: Ho, Darryl, et al.
Veröffentlicht: (2025) -
TAG-Head: Time-Aligned Graph Head for Plug-and-Play Fine-grained Action Recognition
von: Hassan, Imtiaz Ul, et al.
Veröffentlicht: (2026) -
High-Frequency Semantics and Geometric Priors for End-to-End Detection Transformers in Challenging UAV Imagery
von: Peng, Hongxing, et al.
Veröffentlicht: (2025) -
PhysicsNeRF: Physics-Guided 3D Reconstruction from Sparse Views
von: Barhdadi, Mohamed Rayan, et al.
Veröffentlicht: (2025) -
Deep Learning-based Depth Estimation Methods from Monocular Image and Videos: A Comprehensive Survey
von: Rajapaksha, Uchitha, et al.
Veröffentlicht: (2024)