STAT: Towards Generalizable Temporal Action Localization
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Yangcen, Liu, Ziyi, Zhai, Yuanhao, Li, Wen, Doerman, David, Yuan, Junsong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Bridge the Gap: From Weak to Full Supervision for Temporal Action Localization with PseudoFormer
por: Liu, Ziyi, et al.
Publicado: (2025)
por: Liu, Ziyi, et al.
Publicado: (2025)
Simultaneous Detection and Interaction Reasoning for Object-Centric Action Recognition
por: Li, Xunsong, et al.
Publicado: (2024)
por: Li, Xunsong, et al.
Publicado: (2024)
Lens Privacy Sealing: A New Benchmark and Method for Physical Privacy-Preserving Action Recognition
por: Liu, Mengyuan, et al.
Publicado: (2026)
por: Liu, Mengyuan, et al.
Publicado: (2026)
Structure-Aware Human Body Reshaping with Adaptive Affinity-Graph Network
por: Deng, Qiwen, et al.
Publicado: (2024)
por: Deng, Qiwen, et al.
Publicado: (2024)
Towards Completeness: A Generalizable Action Proposal Generator for Zero-Shot Temporal Action Localization
por: Du, Jia-Run, et al.
Publicado: (2024)
por: Du, Jia-Run, et al.
Publicado: (2024)
UST-SSM: Unified Spatio-Temporal State Space Models for Point Cloud Video Modeling
por: Li, Peiming, et al.
Publicado: (2025)
por: Li, Peiming, et al.
Publicado: (2025)
3D Skeleton-Based Action Recognition: A Review
por: Liu, Mengyuan, et al.
Publicado: (2025)
por: Liu, Mengyuan, et al.
Publicado: (2025)
Recognizing Actions from Robotic View for Natural Human-Robot Interaction
por: Wang, Ziyi, et al.
Publicado: (2025)
por: Wang, Ziyi, et al.
Publicado: (2025)
IDOL: Unified Dual-Modal Latent Diffusion for Human-Centric Joint Video-Depth Generation
por: Zhai, Yuanhao, et al.
Publicado: (2024)
por: Zhai, Yuanhao, et al.
Publicado: (2024)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
por: Zhai, Yuanhao, et al.
Publicado: (2024)
por: Zhai, Yuanhao, et al.
Publicado: (2024)
Exploring the Temporal Consistency for Point-Level Weakly-Supervised Temporal Action Localization
por: Ma, Yunchuan, et al.
Publicado: (2026)
por: Ma, Yunchuan, et al.
Publicado: (2026)
Temporal Action Localization with Cross Layer Task Decoupling and Refinement
por: Li, Qiang, et al.
Publicado: (2024)
por: Li, Qiang, et al.
Publicado: (2024)
RubricRL: Simple Generalizable Rewards for Text-to-Image Generation
por: Feng, Xuelu, et al.
Publicado: (2025)
por: Feng, Xuelu, et al.
Publicado: (2025)
Towards Mitigating Modality Bias in Vision-Language Models for Temporal Action Localization
por: Li, Jiaqi, et al.
Publicado: (2026)
por: Li, Jiaqi, et al.
Publicado: (2026)
PathDiff: Histopathology Image Synthesis with Unpaired Text and Mask Conditions
por: Bhosale, Mahesh, et al.
Publicado: (2025)
por: Bhosale, Mahesh, et al.
Publicado: (2025)
Boosting Point-supervised Temporal Action Localization via Text Refinement and Alignment
por: Ma, Yunchuan, et al.
Publicado: (2026)
por: Ma, Yunchuan, et al.
Publicado: (2026)
Towards Adaptive Pseudo-label Learning for Semi-Supervised Temporal Action Localization
por: Zhou, Feixiang, et al.
Publicado: (2024)
por: Zhou, Feixiang, et al.
Publicado: (2024)
End-to-End Streaming Video Temporal Action Segmentation with Reinforce Learning
por: Zhang, Jinrong, et al.
Publicado: (2023)
por: Zhang, Jinrong, et al.
Publicado: (2023)
Informative Sample Selection Model for Skeleton-based Action Recognition with Limited Training Samples
por: Tu, Zhigang, et al.
Publicado: (2025)
por: Tu, Zhigang, et al.
Publicado: (2025)
Forecasting Future Videos from Novel Views via Disentangled 3D Scene Representation
por: Yarram, Sudhir, et al.
Publicado: (2024)
por: Yarram, Sudhir, et al.
Publicado: (2024)
CountLLM: Towards Generalizable Repetitive Action Counting via Large Language Model
por: Yao, Ziyu, et al.
Publicado: (2025)
por: Yao, Ziyu, et al.
Publicado: (2025)
SADA: Semantic adversarial unsupervised domain adaptation for Temporal Action Localization
por: Pujol-Perich, David, et al.
Publicado: (2023)
por: Pujol-Perich, David, et al.
Publicado: (2023)
The Evolution of Dataset Distillation: Toward Scalable and Generalizable Solutions
por: Liu, Ping, et al.
Publicado: (2025)
por: Liu, Ping, et al.
Publicado: (2025)
EgoLoc: A Generalizable Solution for Temporal Interaction Localization in Egocentric Videos
por: Ma, Junyi, et al.
Publicado: (2025)
por: Ma, Junyi, et al.
Publicado: (2025)
DeepLocalization: Using change point detection for Temporal Action Localization
por: Rahman, Mohammed Shaiqur, et al.
Publicado: (2024)
por: Rahman, Mohammed Shaiqur, et al.
Publicado: (2024)
REArtGS++: Generalizable Articulation Reconstruction with Temporal Geometry Constraint via Planar Gaussian Splatting
por: Wu, Di, et al.
Publicado: (2025)
por: Wu, Di, et al.
Publicado: (2025)
Online Temporal Action Localization with Memory-Augmented Transformer
por: Song, Youngkil, et al.
Publicado: (2024)
por: Song, Youngkil, et al.
Publicado: (2024)
Test-Time Zero-Shot Temporal Action Localization
por: Liberatori, Benedetta, et al.
Publicado: (2024)
por: Liberatori, Benedetta, et al.
Publicado: (2024)
Technical Report for ActivityNet Challenge 2022 -- Temporal Action Localization
por: Chen, Shimin, et al.
Publicado: (2024)
por: Chen, Shimin, et al.
Publicado: (2024)
Multi-task Learning with Extended Temporal Shift Module for Temporal Action Localization
por: Duong, Anh-Kiet, et al.
Publicado: (2025)
por: Duong, Anh-Kiet, et al.
Publicado: (2025)
InstructGIE: Towards Generalizable Image Editing
por: Meng, Zichong, et al.
Publicado: (2024)
por: Meng, Zichong, et al.
Publicado: (2024)
Rethinking Pseudo-Label Guided Learning for Weakly Supervised Temporal Action Localization from the Perspective of Noise Correction
por: Zhang, Quan, et al.
Publicado: (2025)
por: Zhang, Quan, et al.
Publicado: (2025)
Harnessing Temporal Causality for Advanced Temporal Action Detection
por: Liu, Shuming, et al.
Publicado: (2024)
por: Liu, Shuming, et al.
Publicado: (2024)
MiVLA: Towards Generalizable Vision-Language-Action Model with Human-Robot Mutual Imitation Pre-training
por: Yin, Zhenhan, et al.
Publicado: (2025)
por: Yin, Zhenhan, et al.
Publicado: (2025)
CLIP-AE: CLIP-assisted Cross-view Audio-Visual Enhancement for Unsupervised Temporal Action Localization
por: Xia, Rui, et al.
Publicado: (2025)
por: Xia, Rui, et al.
Publicado: (2025)
Density-Guided Label Smoothing for Temporal Localization of Driving Actions
por: Alkanat, Tunc, et al.
Publicado: (2024)
por: Alkanat, Tunc, et al.
Publicado: (2024)
Open-Vocabulary Temporal Action Localization using Multimodal Guidance
por: Gupta, Akshita, et al.
Publicado: (2024)
por: Gupta, Akshita, et al.
Publicado: (2024)
Video Self-Stitching Graph Network for Temporal Action Localization
por: Zhao, Chen, et al.
Publicado: (2020)
por: Zhao, Chen, et al.
Publicado: (2020)
Masked Diffusion Vision-Language Models for Temporal Action Localization
por: Wang, Fengshun, et al.
Publicado: (2026)
por: Wang, Fengshun, et al.
Publicado: (2026)
TBT-Former: Learning Temporal Boundary Distributions for Action Localization
por: Rathnayaka, Thisara, et al.
Publicado: (2025)
por: Rathnayaka, Thisara, et al.
Publicado: (2025)
Ejemplares similares
-
Bridge the Gap: From Weak to Full Supervision for Temporal Action Localization with PseudoFormer
por: Liu, Ziyi, et al.
Publicado: (2025) -
Simultaneous Detection and Interaction Reasoning for Object-Centric Action Recognition
por: Li, Xunsong, et al.
Publicado: (2024) -
Lens Privacy Sealing: A New Benchmark and Method for Physical Privacy-Preserving Action Recognition
por: Liu, Mengyuan, et al.
Publicado: (2026) -
Structure-Aware Human Body Reshaping with Adaptive Affinity-Graph Network
por: Deng, Qiwen, et al.
Publicado: (2024) -
Towards Completeness: A Generalizable Action Proposal Generator for Zero-Shot Temporal Action Localization
por: Du, Jia-Run, et al.
Publicado: (2024)