Multi-model learning by sequential reading of untrimmed videos for action recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kamiya, Kodai, Tamaki, Toru |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Online pre-training with long-form videos
von: Kato, Itsuki, et al.
Veröffentlicht: (2024)
von: Kato, Itsuki, et al.
Veröffentlicht: (2024)
Can masking background and object reduce static bias for zero-shot action recognition?
von: Fukuzawa, Takumi, et al.
Veröffentlicht: (2025)
von: Fukuzawa, Takumi, et al.
Veröffentlicht: (2025)
Shift and matching queries for video semantic segmentation
von: Mizuno, Tsubasa, et al.
Veröffentlicht: (2024)
von: Mizuno, Tsubasa, et al.
Veröffentlicht: (2024)
Fine-grained length controllable video captioning with ordinal embeddings
von: Nitta, Tomoya, et al.
Veröffentlicht: (2024)
von: Nitta, Tomoya, et al.
Veröffentlicht: (2024)
Query matching for spatio-temporal action detection with query-based object detector
von: Hori, Shimon, et al.
Veröffentlicht: (2024)
von: Hori, Shimon, et al.
Veröffentlicht: (2024)
Action tube generation by person query matching for spatio-temporal action detection
von: Omi, Kazuki, et al.
Veröffentlicht: (2025)
von: Omi, Kazuki, et al.
Veröffentlicht: (2025)
View while Moving: Efficient Video Recognition in Long-untrimmed Videos
von: Tian, Ye, et al.
Veröffentlicht: (2023)
von: Tian, Ye, et al.
Veröffentlicht: (2023)
M3DDM+: An improved video outpainting by a modified masking strategy
von: Murakawa, Takuya, et al.
Veröffentlicht: (2026)
von: Murakawa, Takuya, et al.
Veröffentlicht: (2026)
Shot2Tactic-Caption: Multi-Scale Captioning of Badminton Videos for Tactical Understanding
von: Ding, Ning, et al.
Veröffentlicht: (2025)
von: Ding, Ning, et al.
Veröffentlicht: (2025)
Reflective Dialogue between Teacher and Solver Agents for Video Question Answering
von: Murakawa, Takuya, et al.
Veröffentlicht: (2026)
von: Murakawa, Takuya, et al.
Veröffentlicht: (2026)
Contrastive learning-based video quality assessment-jointed video vision transformer for video recognition
von: Sun, Jian, et al.
Veröffentlicht: (2026)
von: Sun, Jian, et al.
Veröffentlicht: (2026)
Can VLMs be used on videos for action recognition? LLMs are Visual Reasoning Coordinators
von: Lunia, Harsh
Veröffentlicht: (2024)
von: Lunia, Harsh
Veröffentlicht: (2024)
Separating Shared and Domain-Specific LoRAs for Multi-Domain Learning
von: Takama, Yusaku, et al.
Veröffentlicht: (2025)
von: Takama, Yusaku, et al.
Veröffentlicht: (2025)
Deep learning for action spotting in association football videos
von: Giancola, Silvio, et al.
Veröffentlicht: (2024)
von: Giancola, Silvio, et al.
Veröffentlicht: (2024)
MoExDA: Domain Adaptation for Edge-based Action Recognition
von: Sugimoto, Takuya, et al.
Veröffentlicht: (2025)
von: Sugimoto, Takuya, et al.
Veröffentlicht: (2025)
BFMD: A Full-Match Badminton Dense Dataset for Dense Shot Captioning
von: Ding, Ning, et al.
Veröffentlicht: (2026)
von: Ding, Ning, et al.
Veröffentlicht: (2026)
Disentangling Static and Dynamic Information for Reducing Static Bias in Action Recognition
von: Kobayashi, Masato, et al.
Veröffentlicht: (2025)
von: Kobayashi, Masato, et al.
Veröffentlicht: (2025)
Two-stream joint matching method based on contrastive learning for few-shot action recognition
von: Deng, Long, et al.
Veröffentlicht: (2024)
von: Deng, Long, et al.
Veröffentlicht: (2024)
ADHD diagnosis based on action characteristics recorded in videos using machine learning
von: Li, Yichun, et al.
Veröffentlicht: (2024)
von: Li, Yichun, et al.
Veröffentlicht: (2024)
OmViD: Omni-supervised active learning for video action detection
von: Rana, Aayush, et al.
Veröffentlicht: (2025)
von: Rana, Aayush, et al.
Veröffentlicht: (2025)
Multi-modality action recognition based on dual feature shift in vehicle cabin monitoring
von: Lin, Dan, et al.
Veröffentlicht: (2024)
von: Lin, Dan, et al.
Veröffentlicht: (2024)
Egocentric zone-aware action recognition across environments
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
von: Peirone, Simone Alberto, et al.
Veröffentlicht: (2024)
The impact of Compositionality in Zero-shot Multi-label action recognition for Object-based tasks
von: Calabrese, Carmela, et al.
Veröffentlicht: (2024)
von: Calabrese, Carmela, et al.
Veröffentlicht: (2024)
EALD-MLLM: Emotion Analysis in Long-sequential and De-identity videos with Multi-modal Large Language Model
von: Li, Deng, et al.
Veröffentlicht: (2024)
von: Li, Deng, et al.
Veröffentlicht: (2024)
Cross-modal learning for plankton recognition
von: Kareinen, Joona, et al.
Veröffentlicht: (2026)
von: Kareinen, Joona, et al.
Veröffentlicht: (2026)
CFE-PPAR: Compression-friendly encryption for privacy-preserving action recognition leveraging video transformers
von: Lin, Haiwei, et al.
Veröffentlicht: (2026)
von: Lin, Haiwei, et al.
Veröffentlicht: (2026)
Inpainting-Driven Mask Optimization for Object Removal
von: Shimosato, Kodai, et al.
Veröffentlicht: (2024)
von: Shimosato, Kodai, et al.
Veröffentlicht: (2024)
Deep kernel video approximation for unsupervised action segmentation
von: Pintea, Silvia L., et al.
Veröffentlicht: (2026)
von: Pintea, Silvia L., et al.
Veröffentlicht: (2026)
Real-time 3D human action recognition based on Hyperpoint sequence
von: Li, Xing, et al.
Veröffentlicht: (2021)
von: Li, Xing, et al.
Veröffentlicht: (2021)
Deeply Dual Supervised learning for melanoma recognition
von: Polma, Rujosh, et al.
Veröffentlicht: (2025)
von: Polma, Rujosh, et al.
Veröffentlicht: (2025)
ST-Gait++: Leveraging spatio-temporal convolutions for gait-based emotion recognition on videos
von: Lima, Maria Luísa, et al.
Veröffentlicht: (2024)
von: Lima, Maria Luísa, et al.
Veröffentlicht: (2024)
SMART: Scene-motion-aware human action recognition framework for mental disorder group
von: Lai, Zengyuan, et al.
Veröffentlicht: (2024)
von: Lai, Zengyuan, et al.
Veröffentlicht: (2024)
PO-GUISE+: Pose and object guided transformer token selection for efficient driver action recognition
von: Pizarro, Ricardo, et al.
Veröffentlicht: (2024)
von: Pizarro, Ricardo, et al.
Veröffentlicht: (2024)
Extended multi-stream temporal-attention module for skeleton-based human action recognition (HAR)
von: Mehmood, Faisal, et al.
Veröffentlicht: (2024)
von: Mehmood, Faisal, et al.
Veröffentlicht: (2024)
Using joint angles based on the international biomechanical standards for human action recognition and related tasks
von: Schlegel, Kevin, et al.
Veröffentlicht: (2024)
von: Schlegel, Kevin, et al.
Veröffentlicht: (2024)
From Macro to Micro: Boosting micro-expression recognition via pre-training on macro-expression videos
von: Li, Hanting, et al.
Veröffentlicht: (2024)
von: Li, Hanting, et al.
Veröffentlicht: (2024)
Deep video representation learning: a survey
von: Ravanbakhsh, Elham, et al.
Veröffentlicht: (2024)
von: Ravanbakhsh, Elham, et al.
Veröffentlicht: (2024)
Genetic Information Analysis of Age-Related Macular Degeneration Fellow Eye Using Multi-Modal Selective ViT
von: Furukawa, Yoichi, et al.
Veröffentlicht: (2024)
von: Furukawa, Yoichi, et al.
Veröffentlicht: (2024)
Novel Anomaly Detection Scenarios and Evaluation Metrics to Address the Ambiguity in the Definition of Normal Samples
von: Saito, Reiji, et al.
Veröffentlicht: (2026)
von: Saito, Reiji, et al.
Veröffentlicht: (2026)
Brain Hematoma Marker Recognition Using Multitask Learning: SwinTransformer and Swin-Unet
von: Hirata, Kodai, et al.
Veröffentlicht: (2025)
von: Hirata, Kodai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Online pre-training with long-form videos
von: Kato, Itsuki, et al.
Veröffentlicht: (2024) -
Can masking background and object reduce static bias for zero-shot action recognition?
von: Fukuzawa, Takumi, et al.
Veröffentlicht: (2025) -
Shift and matching queries for video semantic segmentation
von: Mizuno, Tsubasa, et al.
Veröffentlicht: (2024) -
Fine-grained length controllable video captioning with ordinal embeddings
von: Nitta, Tomoya, et al.
Veröffentlicht: (2024) -
Query matching for spatio-temporal action detection with query-based object detector
von: Hori, Shimon, et al.
Veröffentlicht: (2024)