Dual DETRs for Multi-Label Temporal Action Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhu, Yuhan, Zhang, Guozhen, Tan, Jing, Wu, Gangshan, Wang, Limin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Open-Vocabulary Spatio-Temporal Action Detection
di: Wu, Tao, et al.
Pubblicazione: (2024)
di: Wu, Tao, et al.
Pubblicazione: (2024)
Efficient Test-Time Prompt Tuning for Vision-Language Models
di: Zhu, Yuhan, et al.
Pubblicazione: (2024)
di: Zhu, Yuhan, et al.
Pubblicazione: (2024)
AWT: Transferring Vision-Language Models via Augmentation, Weighting, and Transportation
di: Zhu, Yuhan, et al.
Pubblicazione: (2024)
di: Zhu, Yuhan, et al.
Pubblicazione: (2024)
STMixer: A One-Stage Sparse Action Detector
di: Wu, Tao, et al.
Pubblicazione: (2024)
di: Wu, Tao, et al.
Pubblicazione: (2024)
SportsHHI: A Dataset for Human-Human Interaction Detection in Sports Videos
di: Wu, Tao, et al.
Pubblicazione: (2024)
di: Wu, Tao, et al.
Pubblicazione: (2024)
Integrating Diverse Assignment Strategies into DETRs
di: Zhang, Yiwei, et al.
Pubblicazione: (2026)
di: Zhang, Yiwei, et al.
Pubblicazione: (2026)
DETRs Beat YOLOs on Real-time Object Detection
di: Zhao, Yian, et al.
Pubblicazione: (2023)
di: Zhao, Yian, et al.
Pubblicazione: (2023)
Joint Modeling of Feature, Correspondence, and a Compressed Memory for Video Object Segmentation
di: Zhang, Jiaming, et al.
Pubblicazione: (2023)
di: Zhang, Jiaming, et al.
Pubblicazione: (2023)
MotionRAG: Motion Retrieval-Augmented Image-to-Video Generation
di: Zhu, Chenhui, et al.
Pubblicazione: (2025)
di: Zhu, Chenhui, et al.
Pubblicazione: (2025)
Motion-Aware Generative Frame Interpolation
di: Zhang, Guozhen, et al.
Pubblicazione: (2025)
di: Zhang, Guozhen, et al.
Pubblicazione: (2025)
Unbiased Regression Loss for DETRs
di: Edric, et al.
Pubblicazione: (2024)
di: Edric, et al.
Pubblicazione: (2024)
MixFormerV2: Efficient Fully Transformer Tracking
di: Cui, Yutao, et al.
Pubblicazione: (2023)
di: Cui, Yutao, et al.
Pubblicazione: (2023)
Spatial-Temporal Human-Object Interaction Detection
di: Sun, Xu, et al.
Pubblicazione: (2025)
di: Sun, Xu, et al.
Pubblicazione: (2025)
Ranking-based Adaptive Query Generation for DETRs in Crowded Pedestrian Detection
di: Gao, Feng, et al.
Pubblicazione: (2023)
di: Gao, Feng, et al.
Pubblicazione: (2023)
D-FINE: Redefine Regression Task in DETRs as Fine-grained Distribution Refinement
di: Peng, Yansong, et al.
Pubblicazione: (2024)
di: Peng, Yansong, et al.
Pubblicazione: (2024)
One-Stage Open-Vocabulary Temporal Action Detection Leveraging Temporal Multi-scale and Action Label Features
di: Nguyen, Trung Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Trung Thanh, et al.
Pubblicazione: (2024)
Sparse Global Matching for Video Frame Interpolation with Large Motion
di: Liu, Chunxu, et al.
Pubblicazione: (2024)
di: Liu, Chunxu, et al.
Pubblicazione: (2024)
ZeroI2V: Zero-Cost Adaptation of Pre-trained Transformers from Image to Video
di: Li, Xinhao, et al.
Pubblicazione: (2023)
di: Li, Xinhao, et al.
Pubblicazione: (2023)
Enhancing DETRs Variants through Improved Content Query and Similar Query Aggregation
di: Zhang, Yingying, et al.
Pubblicazione: (2024)
di: Zhang, Yingying, et al.
Pubblicazione: (2024)
GLAD: Generative Language-Assisted Visual Tracking for Low-Semantic Templates
di: Luo, Xingyu, et al.
Pubblicazione: (2026)
di: Luo, Xingyu, et al.
Pubblicazione: (2026)
Asymmetric Masked Distillation for Pre-Training Small Foundation Models
di: Zhao, Zhiyu, et al.
Pubblicazione: (2023)
di: Zhao, Zhiyu, et al.
Pubblicazione: (2023)
Probabilistic Temporal Masked Attention for Cross-view Online Action Detection
di: Xie, Liping, et al.
Pubblicazione: (2025)
di: Xie, Liping, et al.
Pubblicazione: (2025)
CLoCKDistill: Consistent Location-and-Context-aware Knowledge Distillation for DETRs
di: Lan, Qizhen, et al.
Pubblicazione: (2025)
di: Lan, Qizhen, et al.
Pubblicazione: (2025)
FDDet: Frequency-Decoupling for Boundary Refinement in Temporal Action Detection
di: Zhu, Xinnan, et al.
Pubblicazione: (2025)
di: Zhu, Xinnan, et al.
Pubblicazione: (2025)
Temporal2Seq: A Unified Framework for Temporal Video Understanding Tasks
di: Yang, Min, et al.
Pubblicazione: (2024)
di: Yang, Min, et al.
Pubblicazione: (2024)
An Effective-Efficient Approach for Dense Multi-Label Action Detection
di: Sardari, Faegheh, et al.
Pubblicazione: (2024)
di: Sardari, Faegheh, et al.
Pubblicazione: (2024)
Adapting Short-Term Transformers for Action Detection in Untrimmed Videos
di: Yang, Min, et al.
Pubblicazione: (2023)
di: Yang, Min, et al.
Pubblicazione: (2023)
Arbitrary Generative Video Interpolation
di: Zhang, Guozhen, et al.
Pubblicazione: (2025)
di: Zhang, Guozhen, et al.
Pubblicazione: (2025)
AnomalyR1: A GRPO-based End-to-end MLLM for Industrial Anomaly Detection
di: Chao, Yuhao, et al.
Pubblicazione: (2025)
di: Chao, Yuhao, et al.
Pubblicazione: (2025)
RGB-D Video Object Segmentation via Enhanced Multi-store Feature Memory
di: Xu, Boyue, et al.
Pubblicazione: (2025)
di: Xu, Boyue, et al.
Pubblicazione: (2025)
Multi-Granularity Hand Action Detection
di: Zhe, Ting, et al.
Pubblicazione: (2023)
di: Zhe, Ting, et al.
Pubblicazione: (2023)
Spatial-Temporal Perception with Causal Inference for Naturalistic Driving Action Recognition
di: Chang, Qing, et al.
Pubblicazione: (2025)
di: Chang, Qing, et al.
Pubblicazione: (2025)
Harnessing Temporal Causality for Advanced Temporal Action Detection
di: Liu, Shuming, et al.
Pubblicazione: (2024)
di: Liu, Shuming, et al.
Pubblicazione: (2024)
Test-time Ego-Exo-centric Adaptation for Action Anticipation via Multi-Label Prototype Growing and Dual-Clue Consistency
di: Shi, Zhaofeng, et al.
Pubblicazione: (2026)
di: Shi, Zhaofeng, et al.
Pubblicazione: (2026)
Denoise and Align: Diffusion-Driven Foreground Knowledge Prompting for Open-Vocabulary Temporal Action Detection
di: Zhu, Sa, et al.
Pubblicazione: (2026)
di: Zhu, Sa, et al.
Pubblicazione: (2026)
StableDrag: Stable Dragging for Point-based Image Editing
di: Cui, Yutao, et al.
Pubblicazione: (2024)
di: Cui, Yutao, et al.
Pubblicazione: (2024)
Sketch and Refine: Towards Fast and Accurate Lane Detection
di: Chen, Chao, et al.
Pubblicazione: (2024)
di: Chen, Chao, et al.
Pubblicazione: (2024)
TempR1: Improving Temporal Understanding of MLLMs via Temporal-Aware Multi-Task Reinforcement Learning
di: Wu, Tao, et al.
Pubblicazione: (2025)
di: Wu, Tao, et al.
Pubblicazione: (2025)
Density-Guided Label Smoothing for Temporal Localization of Driving Actions
di: Alkanat, Tunc, et al.
Pubblicazione: (2024)
di: Alkanat, Tunc, et al.
Pubblicazione: (2024)
Benchmarking the Robustness of Temporal Action Detection Models Against Temporal Corruptions
di: Zeng, Runhao, et al.
Pubblicazione: (2024)
di: Zeng, Runhao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Open-Vocabulary Spatio-Temporal Action Detection
di: Wu, Tao, et al.
Pubblicazione: (2024) -
Efficient Test-Time Prompt Tuning for Vision-Language Models
di: Zhu, Yuhan, et al.
Pubblicazione: (2024) -
AWT: Transferring Vision-Language Models via Augmentation, Weighting, and Transportation
di: Zhu, Yuhan, et al.
Pubblicazione: (2024) -
STMixer: A One-Stage Sparse Action Detector
di: Wu, Tao, et al.
Pubblicazione: (2024) -
SportsHHI: A Dataset for Human-Human Interaction Detection in Sports Videos
di: Wu, Tao, et al.
Pubblicazione: (2024)