Is Temporal Prompting All We Need For Limited Labeled Action Recognition?
Fuente:
arXiv
Guardado en:
| Autores principales: | Gowda, Shreyank N, Gao, Boyan, Gu, Xiao, Jin, Xiaobo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Telling Stories for Common Sense Zero-Shot Action Recognition
por: Gowda, Shreyank N, et al.
Publicado: (2023)
por: Gowda, Shreyank N, et al.
Publicado: (2023)
FE-Adapter: Adapting Image-based Emotion Classifiers to Videos
por: Gowda, Shreyank N, et al.
Publicado: (2024)
por: Gowda, Shreyank N, et al.
Publicado: (2024)
Prototype-Enhanced Confidence Modeling for Cross-Modal Medical Image-Report Retrieval
por: Gowda, Shreyank N, et al.
Publicado: (2025)
por: Gowda, Shreyank N, et al.
Publicado: (2025)
Performance is not All You Need: Sustainability Considerations for Algorithms
por: Li, Xiang, et al.
Publicado: (2025)
por: Li, Xiang, et al.
Publicado: (2025)
Continual Learning Improves Zero-Shot Action Recognition
por: Gowda, Shreyank N, et al.
Publicado: (2024)
por: Gowda, Shreyank N, et al.
Publicado: (2024)
Reimagining Reality: A Comprehensive Survey of Video Inpainting Techniques
por: Gowda, Shreyank N, et al.
Publicado: (2024)
por: Gowda, Shreyank N, et al.
Publicado: (2024)
FATE: A Prompt-Tuning-Based Semi-Supervised Learning Framework for Extremely Limited Labeled Data
por: Liu, Hezhao, et al.
Publicado: (2025)
por: Liu, Hezhao, et al.
Publicado: (2025)
Adversarial Augmentation Training Makes Action Recognition Models More Robust to Realistic Video Distribution Shifts
por: Kim, Kiyoon, et al.
Publicado: (2024)
por: Kim, Kiyoon, et al.
Publicado: (2024)
Interpretable Zero-shot Learning with Infinite Class Concepts
por: Ye, Zihan, et al.
Publicado: (2025)
por: Ye, Zihan, et al.
Publicado: (2025)
CC-SAM: SAM with Cross-feature Attention and Context for Ultrasound Image Segmentation
por: Gowda, Shreyank N, et al.
Publicado: (2024)
por: Gowda, Shreyank N, et al.
Publicado: (2024)
Masks and Manuscripts: Advancing Medical Pre-training with End-to-End Masking and Narrative Structuring
por: Gowda, Shreyank N, et al.
Publicado: (2024)
por: Gowda, Shreyank N, et al.
Publicado: (2024)
Distribution-Based Masked Medical Vision-Language Model Using Structured Reports
por: Gowda, Shreyank N, et al.
Publicado: (2025)
por: Gowda, Shreyank N, et al.
Publicado: (2025)
Watt For What: Rethinking Deep Learning's Energy-Performance Relationship
por: Gowda, Shreyank N, et al.
Publicado: (2023)
por: Gowda, Shreyank N, et al.
Publicado: (2023)
Bridging the Projection Gap: Overcoming Projection Bias Through Parameterized Distance Learning
por: Zhang, Chong, et al.
Publicado: (2023)
por: Zhang, Chong, et al.
Publicado: (2023)
Twin Trigger Generative Networks for Backdoor Attacks against Object Detection
por: Li, Zhiying, et al.
Publicado: (2024)
por: Li, Zhiying, et al.
Publicado: (2024)
Anyone Can Jailbreak: Prompt-Based Attacks on LLMs and T2Is
por: Mustafa, Ahmed B, et al.
Publicado: (2025)
por: Mustafa, Ahmed B, et al.
Publicado: (2025)
ZeroDiff: Solidified Visual-Semantic Correlation in Zero-Shot Learning
por: Ye, Zihan, et al.
Publicado: (2024)
por: Ye, Zihan, et al.
Publicado: (2024)
Adaptive Data Dropout: Towards Self-Regulated Learning in Deep Neural Networks
por: Gahir, Amar, et al.
Publicado: (2026)
por: Gahir, Amar, et al.
Publicado: (2026)
Compression as an Adversarial Amplifier Through Decision Space Reduction
por: Evans, Lewis, et al.
Publicado: (2026)
por: Evans, Lewis, et al.
Publicado: (2026)
ZeroDiff++: Substantial Unseen Visual-semantic Correlation in Zero-shot Learning
por: Ye, Zihan, et al.
Publicado: (2026)
por: Ye, Zihan, et al.
Publicado: (2026)
CAPT: Class-Aware Prompt Tuning for Federated Long-Tailed Learning with Vision-Language Model
por: Hou, Shihao, et al.
Publicado: (2025)
por: Hou, Shihao, et al.
Publicado: (2025)
SECOS: Semantic Capture for Rigorous Classification in Open-World Semi-Supervised Learning
por: Liu, Hezhao, et al.
Publicado: (2026)
por: Liu, Hezhao, et al.
Publicado: (2026)
Adversarial Robustness in Zero-Shot Learning:An Empirical Study on Class and Concept-Level Vulnerabilities
por: Peng, Zhiyuan, et al.
Publicado: (2025)
por: Peng, Zhiyuan, et al.
Publicado: (2025)
Low-Effort Jailbreak Attacks Against Text-to-Image Safety Filters
por: Mustafa, Ahmed B, et al.
Publicado: (2026)
por: Mustafa, Ahmed B, et al.
Publicado: (2026)
Positive Label Is All You Need for Multi-Label Classification
por: Yuan, Zhixiang, et al.
Publicado: (2023)
por: Yuan, Zhixiang, et al.
Publicado: (2023)
Towards Micro-Action Recognition with Limited Annotations: An Asynchronous Pseudo Labeling and Training Approach
por: Zhang, Yan, et al.
Publicado: (2025)
por: Zhang, Yan, et al.
Publicado: (2025)
EgoPrompt: Prompt Learning for Egocentric Action Recognition
por: Lyu, Huaihai, et al.
Publicado: (2025)
por: Lyu, Huaihai, et al.
Publicado: (2025)
Density-Guided Label Smoothing for Temporal Localization of Driving Actions
por: Alkanat, Tunc, et al.
Publicado: (2024)
por: Alkanat, Tunc, et al.
Publicado: (2024)
Synchronization is All You Need: Exocentric-to-Egocentric Transfer for Temporal Action Segmentation with Unlabeled Synchronized Video Pairs
por: Quattrocchi, Camillo, et al.
Publicado: (2023)
por: Quattrocchi, Camillo, et al.
Publicado: (2023)
Prompt-guided Disentangled Representation for Action Recognition
por: Wu, Tianci, et al.
Publicado: (2025)
por: Wu, Tianci, et al.
Publicado: (2025)
Dual DETRs for Multi-Label Temporal Action Detection
por: Zhu, Yuhan, et al.
Publicado: (2024)
por: Zhu, Yuhan, et al.
Publicado: (2024)
Principles of Visual Tokens for Efficient Video Understanding
por: Hao, Xinyue, et al.
Publicado: (2024)
por: Hao, Xinyue, et al.
Publicado: (2024)
Leveraging Temporal Contextualization for Video Action Recognition
por: Kim, Minji, et al.
Publicado: (2024)
por: Kim, Minji, et al.
Publicado: (2024)
Novel Semantic Prompting for Zero-Shot Action Recognition
por: Iqbal, Salman, et al.
Publicado: (2026)
por: Iqbal, Salman, et al.
Publicado: (2026)
TDS-CLIP: Temporal Difference Side Network for Efficient VideoAction Recognition
por: Wang, Bin, et al.
Publicado: (2024)
por: Wang, Bin, et al.
Publicado: (2024)
One-Stage Open-Vocabulary Temporal Action Detection Leveraging Temporal Multi-scale and Action Label Features
por: Nguyen, Trung Thanh, et al.
Publicado: (2024)
por: Nguyen, Trung Thanh, et al.
Publicado: (2024)
Beyond Label Semantics: Language-Guided Action Anatomy for Few-shot Action Recognition
por: Qian, Zefeng, et al.
Publicado: (2025)
por: Qian, Zefeng, et al.
Publicado: (2025)
TP-UNet: Temporal Prompt Guided UNet for Medical Image Segmentation
por: Wang, Ranmin, et al.
Publicado: (2024)
por: Wang, Ranmin, et al.
Publicado: (2024)
Modelling Spatio-Temporal Interactions For Compositional Action Recognition
por: Rajendiran, Ramanathan, et al.
Publicado: (2023)
por: Rajendiran, Ramanathan, et al.
Publicado: (2023)
AM Flow: Adapters for Temporal Processing in Action Recognition
por: Agrawal, Tanay, et al.
Publicado: (2024)
por: Agrawal, Tanay, et al.
Publicado: (2024)
Ejemplares similares
-
Telling Stories for Common Sense Zero-Shot Action Recognition
por: Gowda, Shreyank N, et al.
Publicado: (2023) -
FE-Adapter: Adapting Image-based Emotion Classifiers to Videos
por: Gowda, Shreyank N, et al.
Publicado: (2024) -
Prototype-Enhanced Confidence Modeling for Cross-Modal Medical Image-Report Retrieval
por: Gowda, Shreyank N, et al.
Publicado: (2025) -
Performance is not All You Need: Sustainability Considerations for Algorithms
por: Li, Xiang, et al.
Publicado: (2025) -
Continual Learning Improves Zero-Shot Action Recognition
por: Gowda, Shreyank N, et al.
Publicado: (2024)