Action tube generation by person query matching for spatio-temporal action detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Omi, Kazuki, Oshima, Jion, Tamaki, Toru |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Query matching for spatio-temporal action detection with query-based object detector
di: Hori, Shimon, et al.
Pubblicazione: (2024)
di: Hori, Shimon, et al.
Pubblicazione: (2024)
Shift and matching queries for video semantic segmentation
di: Mizuno, Tsubasa, et al.
Pubblicazione: (2024)
di: Mizuno, Tsubasa, et al.
Pubblicazione: (2024)
Multi-model learning by sequential reading of untrimmed videos for action recognition
di: Kamiya, Kodai, et al.
Pubblicazione: (2024)
di: Kamiya, Kodai, et al.
Pubblicazione: (2024)
MoExDA: Domain Adaptation for Edge-based Action Recognition
di: Sugimoto, Takuya, et al.
Pubblicazione: (2025)
di: Sugimoto, Takuya, et al.
Pubblicazione: (2025)
Disentangling Static and Dynamic Information for Reducing Static Bias in Action Recognition
di: Kobayashi, Masato, et al.
Pubblicazione: (2025)
di: Kobayashi, Masato, et al.
Pubblicazione: (2025)
Can masking background and object reduce static bias for zero-shot action recognition?
di: Fukuzawa, Takumi, et al.
Pubblicazione: (2025)
di: Fukuzawa, Takumi, et al.
Pubblicazione: (2025)
Reflective Dialogue between Teacher and Solver Agents for Video Question Answering
di: Murakawa, Takuya, et al.
Pubblicazione: (2026)
di: Murakawa, Takuya, et al.
Pubblicazione: (2026)
Disentangling spatio-temporal knowledge for weakly supervised object detection and segmentation in surgical video
di: Liao, Guiqiu, et al.
Pubblicazione: (2024)
di: Liao, Guiqiu, et al.
Pubblicazione: (2024)
Shot2Tactic-Caption: Multi-Scale Captioning of Badminton Videos for Tactical Understanding
di: Ding, Ning, et al.
Pubblicazione: (2025)
di: Ding, Ning, et al.
Pubblicazione: (2025)
BFMD: A Full-Match Badminton Dense Dataset for Dense Shot Captioning
di: Ding, Ning, et al.
Pubblicazione: (2026)
di: Ding, Ning, et al.
Pubblicazione: (2026)
Online pre-training with long-form videos
di: Kato, Itsuki, et al.
Pubblicazione: (2024)
di: Kato, Itsuki, et al.
Pubblicazione: (2024)
Fine-grained length controllable video captioning with ordinal embeddings
di: Nitta, Tomoya, et al.
Pubblicazione: (2024)
di: Nitta, Tomoya, et al.
Pubblicazione: (2024)
TadML: A fast temporal action detection with Mechanics-MLP
di: Deng, Bowen, et al.
Pubblicazione: (2022)
di: Deng, Bowen, et al.
Pubblicazione: (2022)
Separating Shared and Domain-Specific LoRAs for Multi-Domain Learning
di: Takama, Yusaku, et al.
Pubblicazione: (2025)
di: Takama, Yusaku, et al.
Pubblicazione: (2025)
M3DDM+: An improved video outpainting by a modified masking strategy
di: Murakawa, Takuya, et al.
Pubblicazione: (2026)
di: Murakawa, Takuya, et al.
Pubblicazione: (2026)
ST-Gait++: Leveraging spatio-temporal convolutions for gait-based emotion recognition on videos
di: Lima, Maria Luísa, et al.
Pubblicazione: (2024)
di: Lima, Maria Luísa, et al.
Pubblicazione: (2024)
Global spatio-temporal downscaling of ERA5 precipitation through generative AI
di: Glawion, Luca, et al.
Pubblicazione: (2024)
di: Glawion, Luca, et al.
Pubblicazione: (2024)
MM-SEAL: A Large-scale Video Dataset of Multi-person Multi-grained Spatio-temporally Action Localization
di: Chen, Shimin, et al.
Pubblicazione: (2022)
di: Chen, Shimin, et al.
Pubblicazione: (2022)
Detection of moving objects through turbulent media. Decomposition of Oscillatory vs Non-Oscillatory spatio-temporal vector fields
di: Gilles, Jerome, et al.
Pubblicazione: (2024)
di: Gilles, Jerome, et al.
Pubblicazione: (2024)
Covariant spatio-temporal receptive fields for spiking neural networks
di: Pedersen, Jens Egholm, et al.
Pubblicazione: (2024)
di: Pedersen, Jens Egholm, et al.
Pubblicazione: (2024)
Rethinking temporal self-similarity for repetitive action counting
di: Luo, Yanan, et al.
Pubblicazione: (2024)
di: Luo, Yanan, et al.
Pubblicazione: (2024)
$^R$FLAV: Rolling Flow matching for infinite Audio Video generation
di: Ergasti, Alex, et al.
Pubblicazione: (2025)
di: Ergasti, Alex, et al.
Pubblicazione: (2025)
Two-stream joint matching method based on contrastive learning for few-shot action recognition
di: Deng, Long, et al.
Pubblicazione: (2024)
di: Deng, Long, et al.
Pubblicazione: (2024)
Local2Global query Alignment for Video Instance Segmentation
di: Koner, Rajat, et al.
Pubblicazione: (2025)
di: Koner, Rajat, et al.
Pubblicazione: (2025)
UKDM: Underwater keypoint detection and matching using underwater image enhancement techniques
di: Diaz-Garcia, Pedro, et al.
Pubblicazione: (2025)
di: Diaz-Garcia, Pedro, et al.
Pubblicazione: (2025)
Large-scale unsupervised spatio-temporal semantic analysis of vast regions from satellite images sequences
di: Echegoyen, Carlos, et al.
Pubblicazione: (2022)
di: Echegoyen, Carlos, et al.
Pubblicazione: (2022)
Symmetrical Joint Learning Support-query Prototypes for Few-shot Segmentation
di: Li, Qun, et al.
Pubblicazione: (2024)
di: Li, Qun, et al.
Pubblicazione: (2024)
Station2Radar: query conditioned gaussian splatting for precipitation field
di: Kim, Doyi, et al.
Pubblicazione: (2026)
di: Kim, Doyi, et al.
Pubblicazione: (2026)
INFANiTE: Implicit Neural representation for high-resolution Fetal brain spatio-temporal Atlas learNing from clinical Thick-slicE MRI
di: Hu, Xiaotian, et al.
Pubblicazione: (2026)
di: Hu, Xiaotian, et al.
Pubblicazione: (2026)
Spatio-temporal Transformers for Action Unit Classification with Event Cameras
di: Cultrera, Luca, et al.
Pubblicazione: (2024)
di: Cultrera, Luca, et al.
Pubblicazione: (2024)
MSTAR: Box-free Multi-query Scene Text Retrieval with Attention Recycling
di: Yin, Liang, et al.
Pubblicazione: (2025)
di: Yin, Liang, et al.
Pubblicazione: (2025)
IDPro: Flexible Interactive Video Object Segmentation by ID-queried Concurrent Propagation
di: Li, Kexin, et al.
Pubblicazione: (2024)
di: Li, Kexin, et al.
Pubblicazione: (2024)
Interlaced dynamic XCT reconstruction with spatio-temporal implicit neural representations
di: Boulanger, Mathias, et al.
Pubblicazione: (2025)
di: Boulanger, Mathias, et al.
Pubblicazione: (2025)
FineParser: A Fine-grained Spatio-temporal Action Parser for Human-centric Action Quality Assessment
di: Xu, Jinglin, et al.
Pubblicazione: (2024)
di: Xu, Jinglin, et al.
Pubblicazione: (2024)
Spatio-temporal Decoupled Knowledge Compensator for Few-Shot Action Recognition
di: Qu, Hongyu, et al.
Pubblicazione: (2026)
di: Qu, Hongyu, et al.
Pubblicazione: (2026)
OmViD: Omni-supervised active learning for video action detection
di: Rana, Aayush, et al.
Pubblicazione: (2025)
di: Rana, Aayush, et al.
Pubblicazione: (2025)
Hybrid-Tower: Fine-grained Pseudo-query Interaction and Generation for Text-to-Video Retrieval
di: Lan, Bangxiang, et al.
Pubblicazione: (2025)
di: Lan, Bangxiang, et al.
Pubblicazione: (2025)
Extended multi-stream temporal-attention module for skeleton-based human action recognition (HAR)
di: Mehmood, Faisal, et al.
Pubblicazione: (2024)
di: Mehmood, Faisal, et al.
Pubblicazione: (2024)
Action-guided generation of 3D functionality segmentation data
di: Corsetti, Jaime, et al.
Pubblicazione: (2025)
di: Corsetti, Jaime, et al.
Pubblicazione: (2025)
Anomaly detection in satellite imagery through temporal inpainting
di: Rouet-Leduc, Bertrand, et al.
Pubblicazione: (2025)
di: Rouet-Leduc, Bertrand, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Query matching for spatio-temporal action detection with query-based object detector
di: Hori, Shimon, et al.
Pubblicazione: (2024) -
Shift and matching queries for video semantic segmentation
di: Mizuno, Tsubasa, et al.
Pubblicazione: (2024) -
Multi-model learning by sequential reading of untrimmed videos for action recognition
di: Kamiya, Kodai, et al.
Pubblicazione: (2024) -
MoExDA: Domain Adaptation for Edge-based Action Recognition
di: Sugimoto, Takuya, et al.
Pubblicazione: (2025) -
Disentangling Static and Dynamic Information for Reducing Static Bias in Action Recognition
di: Kobayashi, Masato, et al.
Pubblicazione: (2025)