Learning Discriminative Spatio-temporal Representations for Semi-supervised Action Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yu, Zhou, Sanping, Xia, Kun, Wang, Le |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DSER: Spectral Epipolar Representation for Efficient Light Field Depth Estimation
von: Mohammad, Noor Islam S., et al.
Veröffentlicht: (2025)
von: Mohammad, Noor Islam S., et al.
Veröffentlicht: (2025)
DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations
von: Jin, Hang, et al.
Veröffentlicht: (2025)
von: Jin, Hang, et al.
Veröffentlicht: (2025)
Sequence Matters: Harnessing Video Models in 3D Super-Resolution
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024)
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024)
Hierarchical Spatial Algorithms for High-Resolution Image Quantization and Feature Extraction
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
Deep Learning Approaches for Human Action Recognition in Video Data
von: Xie, Yufei
Veröffentlicht: (2024)
von: Xie, Yufei
Veröffentlicht: (2024)
Real Time Human Detection by Unmanned Aerial Vehicles
von: Guettala, Walid, et al.
Veröffentlicht: (2024)
von: Guettala, Walid, et al.
Veröffentlicht: (2024)
When Less is Enough: Adaptive Token Reduction for Efficient Image Representation
von: Allakhverdov, Eduard, et al.
Veröffentlicht: (2025)
von: Allakhverdov, Eduard, et al.
Veröffentlicht: (2025)
A large-scale, physically-based synthetic dataset for satellite pose estimation
von: Velkei, Szabolcs, et al.
Veröffentlicht: (2025)
von: Velkei, Szabolcs, et al.
Veröffentlicht: (2025)
ShapBPT: Image Feature Attributions Using Data-Aware Binary Partition Trees
von: Rashid, Muhammad, et al.
Veröffentlicht: (2026)
von: Rashid, Muhammad, et al.
Veröffentlicht: (2026)
OpenFusion++: An Open-vocabulary Real-time Scene Understanding System
von: Jin, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Jin, Xiaofeng, et al.
Veröffentlicht: (2025)
Image Reconstruction as a Tool for Feature Analysis
von: Allakhverdov, Eduard, et al.
Veröffentlicht: (2025)
von: Allakhverdov, Eduard, et al.
Veröffentlicht: (2025)
Extraction Of Cumulative Blobs From Dynamic Gestures
von: Naulakha, Rishabh, et al.
Veröffentlicht: (2025)
von: Naulakha, Rishabh, et al.
Veröffentlicht: (2025)
Zero-Shot Multi-Criteria Visual Quality Inspection for Semi-Controlled Industrial Environments via Real-Time 3D Digital Twin Simulation
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2025)
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2025)
Does CLIP perceive art the same way we do?
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
UrbanAlign: Post-hoc Semantic Calibration for VLM-Human Preference Alignment
von: Zhang, Yecheng, et al.
Veröffentlicht: (2026)
von: Zhang, Yecheng, et al.
Veröffentlicht: (2026)
Sat-JEPA-Diff: Bridging Self-Supervised Learning and Generative Diffusion for Remote Sensing
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
Advancing Brain Tumor Segmentation via Attention-based 3D U-Net Architecture and Digital Image Processing
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
WildfireVLM: AI-powered Analysis for Early Wildfire Detection and Risk Assessment Using Satellite Imagery
von: Ayanzadeh, Aydin, et al.
Veröffentlicht: (2026)
von: Ayanzadeh, Aydin, et al.
Veröffentlicht: (2026)
Dual-sensing driving detection model
von: K, Leon C. C., et al.
Veröffentlicht: (2025)
von: K, Leon C. C., et al.
Veröffentlicht: (2025)
UVLM: A Universal Vision-Language Model Loader for Reproducible Multimodal Benchmarking
von: Perez, Joan, et al.
Veröffentlicht: (2026)
von: Perez, Joan, et al.
Veröffentlicht: (2026)
Splat and Distill: Augmenting Teachers with Feed-Forward 3D Reconstruction For 3D-Aware Distillation
von: Shavin, David, et al.
Veröffentlicht: (2026)
von: Shavin, David, et al.
Veröffentlicht: (2026)
Deep Learning in Automated Power Line Inspection: A Review
von: Faisal, Md. Ahasan Atick, et al.
Veröffentlicht: (2025)
von: Faisal, Md. Ahasan Atick, et al.
Veröffentlicht: (2025)
Corn Ear Detection and Orientation Estimation Using Deep Learning
von: Sprague, Nathan, et al.
Veröffentlicht: (2024)
von: Sprague, Nathan, et al.
Veröffentlicht: (2024)
Multi-scale Temporal Prediction via Incremental Generation and Multi-agent Collaboration
von: Zeng, Zhitao, et al.
Veröffentlicht: (2025)
von: Zeng, Zhitao, et al.
Veröffentlicht: (2025)
GFLAN: Generative Functional Layouts
von: Abouagour, Mohamed, et al.
Veröffentlicht: (2025)
von: Abouagour, Mohamed, et al.
Veröffentlicht: (2025)
Extrapolating and Decoupling Image-to-Video Generation Models: Motion Modeling is Easier Than You Think
von: Tian, Jie, et al.
Veröffentlicht: (2025)
von: Tian, Jie, et al.
Veröffentlicht: (2025)
Fast 3D point clouds retrieval for Large-scale 3D Place Recognition
von: Zede, Chahine-Nicolas, et al.
Veröffentlicht: (2025)
von: Zede, Chahine-Nicolas, et al.
Veröffentlicht: (2025)
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
von: Zeng, Zhitao, et al.
Veröffentlicht: (2025)
von: Zeng, Zhitao, et al.
Veröffentlicht: (2025)
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
Harmony: A Joint Self-Supervised and Weakly-Supervised Framework for Learning General Purpose Visual Representations
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2024)
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2024)
Gr-IoU: Ground-Intersection over Union for Robust Multi-Object Tracking with 3D Geometric Constraints
von: Toida, Keisuke, et al.
Veröffentlicht: (2024)
von: Toida, Keisuke, et al.
Veröffentlicht: (2024)
Consistent Zero-shot 3D Texture Synthesis Using Geometry-aware Diffusion and Temporal Video Models
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
ClustViT: Clustering-based Token Merging for Semantic Segmentation
von: Montello, Fabio, et al.
Veröffentlicht: (2025)
von: Montello, Fabio, et al.
Veröffentlicht: (2025)
LatentForensics: Towards frugal deepfake detection in the StyleGAN latent space
von: Delmas, Matthieu, et al.
Veröffentlicht: (2023)
von: Delmas, Matthieu, et al.
Veröffentlicht: (2023)
Synthetic Industrial Object Detection: GenAI vs. Feature-Based Methods
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2025)
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2025)
One-to-Normal: Anomaly Personalization for Few-shot Anomaly Detection
von: Li, Yiyue, et al.
Veröffentlicht: (2025)
von: Li, Yiyue, et al.
Veröffentlicht: (2025)
PlaneSAM: Multimodal Plane Instance Segmentation Using the Segment Anything Model
von: Deng, Zhongchen, et al.
Veröffentlicht: (2024)
von: Deng, Zhongchen, et al.
Veröffentlicht: (2024)
SynthRender and IRIS: Open-Source Framework and Dataset for Bidirectional Sim-Real Transfer in Industrial Object Perception
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2026)
von: Araya-Martinez, Jose Moises, et al.
Veröffentlicht: (2026)
DiffYOLO: Object Detection for Anti-Noise via YOLO and Diffusion Models
von: Liu, Yichen, et al.
Veröffentlicht: (2024)
von: Liu, Yichen, et al.
Veröffentlicht: (2024)
Video-CoE: Reinforcing Video Event Prediction via Chain of Events
von: Su, Qile, et al.
Veröffentlicht: (2026)
von: Su, Qile, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
DSER: Spectral Epipolar Representation for Efficient Light Field Depth Estimation
von: Mohammad, Noor Islam S., et al.
Veröffentlicht: (2025) -
DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations
von: Jin, Hang, et al.
Veröffentlicht: (2025) -
Sequence Matters: Harnessing Video Models in 3D Super-Resolution
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2024) -
Hierarchical Spatial Algorithms for High-Resolution Image Quantization and Feature Extraction
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025) -
Deep Learning Approaches for Human Action Recognition in Video Data
von: Xie, Yufei
Veröffentlicht: (2024)