Saved in:
| Main Authors: | Kumar, Prerana, Giese, Martin A. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.16675 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Zero-shot Compositional Action Recognition with Neural Logic Constraints
by: Ye, Gefan, et al.
Published: (2025)
by: Ye, Gefan, et al.
Published: (2025)
ActionHub: A Large-scale Action Video Description Dataset for Zero-shot Action Recognition
by: Zhou, Jiaming, et al.
Published: (2024)
by: Zhou, Jiaming, et al.
Published: (2024)
Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP
by: Yu, Yating, et al.
Published: (2024)
by: Yu, Yating, et al.
Published: (2024)
Part-aware Unified Representation of Language and Skeleton for Zero-shot Action Recognition
by: Zhu, Anqi, et al.
Published: (2024)
by: Zhu, Anqi, et al.
Published: (2024)
Enhancing Spatio-Temporal Zero-shot Action Recognition with Language-driven Description Attributes
by: Kim, Yehna, et al.
Published: (2025)
by: Kim, Yehna, et al.
Published: (2025)
Trajectory-aligned Space-time Tokens for Few-shot Action Recognition
by: Kumar, Pulkit, et al.
Published: (2024)
by: Kumar, Pulkit, et al.
Published: (2024)
Text-Enhanced Zero-Shot Action Recognition: A training-free approach
by: Bosetti, Massimo, et al.
Published: (2024)
by: Bosetti, Massimo, et al.
Published: (2024)
SCALE: Semantic- and Confidence-Aware Conditional Variational Autoencoder for Zero-shot Skeleton-based Action Recognition
by: Oraki, Soroush, et al.
Published: (2026)
by: Oraki, Soroush, et al.
Published: (2026)
World Action Models are Zero-shot Policies
by: Ye, Seonghyeon, et al.
Published: (2026)
by: Ye, Seonghyeon, et al.
Published: (2026)
Hierarchical Compositional Representations for Few-shot Action Recognition
by: Li, Changzhen, et al.
Published: (2022)
by: Li, Changzhen, et al.
Published: (2022)
A Comprehensive Review of Few-shot Action Recognition
by: Wanyan, Yuyang, et al.
Published: (2024)
by: Wanyan, Yuyang, et al.
Published: (2024)
Zero-shot Action Localization via the Confidence of Large Vision-Language Models
by: Aklilu, Josiah, et al.
Published: (2024)
by: Aklilu, Josiah, et al.
Published: (2024)
RobustGait: Robustness Analysis for Appearance Based Gait Recognition
by: Sayera, Reeshoon, et al.
Published: (2025)
by: Sayera, Reeshoon, et al.
Published: (2025)
Bridging the Skeleton-Text Modality Gap: Diffusion-Powered Modality Alignment for Zero-shot Skeleton-based Action Recognition
by: Do, Jeonghyeok, et al.
Published: (2024)
by: Do, Jeonghyeok, et al.
Published: (2024)
CLIP-guided Prototype Modulating for Few-shot Action Recognition
by: Wang, Xiang, et al.
Published: (2023)
by: Wang, Xiang, et al.
Published: (2023)
Active Generation Network of Human Skeleton for Action Recognition
by: Liu, Long, et al.
Published: (2024)
by: Liu, Long, et al.
Published: (2024)
Beyond Label Semantics: Language-Guided Action Anatomy for Few-shot Action Recognition
by: Qian, Zefeng, et al.
Published: (2025)
by: Qian, Zefeng, et al.
Published: (2025)
InterControl: Zero-shot Human Interaction Generation by Controlling Every Joint
by: Wang, Zhenzhi, et al.
Published: (2023)
by: Wang, Zhenzhi, et al.
Published: (2023)
Adaptive Prototype Model for Attribute-based Multi-label Few-shot Action Recognition
by: Xiao, Juefeng, et al.
Published: (2025)
by: Xiao, Juefeng, et al.
Published: (2025)
Task-Adapter: Task-specific Adaptation of Image Models for Few-shot Action Recognition
by: Cao, Congqi, et al.
Published: (2024)
by: Cao, Congqi, et al.
Published: (2024)
GPT-4V with Emotion: A Zero-shot Benchmark for Generalized Emotion Recognition
by: Lian, Zheng, et al.
Published: (2023)
by: Lian, Zheng, et al.
Published: (2023)
Zero-shot Prompt-based Video Encoder for Surgical Gesture Recognition
by: Rao, Mingxing, et al.
Published: (2024)
by: Rao, Mingxing, et al.
Published: (2024)
AvatarGO: Zero-shot 4D Human-Object Interaction Generation and Animation
by: Cao, Yukang, et al.
Published: (2024)
by: Cao, Yukang, et al.
Published: (2024)
ActionCOMET: A Zero-shot Approach to Learn Image-specific Commonsense Concepts about Actions
by: Sampat, Shailaja Keyur, et al.
Published: (2024)
by: Sampat, Shailaja Keyur, et al.
Published: (2024)
Hierarchical Material Recognition from Local Appearance
by: Beveridge, Matthew, et al.
Published: (2025)
by: Beveridge, Matthew, et al.
Published: (2025)
Zero-shot Compound Expression Recognition with Visual Language Model at the 6th ABAW Challenge
by: Wang, Jiahe, et al.
Published: (2024)
by: Wang, Jiahe, et al.
Published: (2024)
Active Multimodal Distillation for Few-shot Action Recognition
by: Feng, Weijia, et al.
Published: (2025)
by: Feng, Weijia, et al.
Published: (2025)
Register and [CLS] tokens yield a decoupling of local and global features in large ViTs
by: Lappe, Alexander, et al.
Published: (2025)
by: Lappe, Alexander, et al.
Published: (2025)
Zero-Shot Action Recognition in Surveillance Videos
by: Pereira, Joao, et al.
Published: (2024)
by: Pereira, Joao, et al.
Published: (2024)
Human Action Recognition without Human
by: Kataoka, Hirokatsu, et al.
Published: (2016)
by: Kataoka, Hirokatsu, et al.
Published: (2016)
Novel Semantic Prompting for Zero-Shot Action Recognition
by: Iqbal, Salman, et al.
Published: (2026)
by: Iqbal, Salman, et al.
Published: (2026)
Continual Learning Improves Zero-Shot Action Recognition
by: Gowda, Shreyank N, et al.
Published: (2024)
by: Gowda, Shreyank N, et al.
Published: (2024)
Joint Image-Instance Spatial-Temporal Attention for Few-shot Action Recognition
by: Qian, Zefeng, et al.
Published: (2025)
by: Qian, Zefeng, et al.
Published: (2025)
Manifold Induced Biases for Zero-shot and Few-shot Detection of Generated Images
by: Brokman, Jonathan, et al.
Published: (2025)
by: Brokman, Jonathan, et al.
Published: (2025)
Zero-shot HOI Detection with MLLM-based Detector-agnostic Interaction Recognition
by: Xuan, Shiyu, et al.
Published: (2026)
by: Xuan, Shiyu, et al.
Published: (2026)
Training-free Zero-shot Composed Image Retrieval with Local Concept Reranking
by: Sun, Shitong, et al.
Published: (2023)
by: Sun, Shitong, et al.
Published: (2023)
YOWOv3: An Efficient and Generalized Framework for Human Action Detection and Recognition
by: Dang, Duc Manh Nguyen, et al.
Published: (2024)
by: Dang, Duc Manh Nguyen, et al.
Published: (2024)
FLORA: Formal Language Model Enables Robust Training-free Zero-shot Object Referring Analysis
by: Chen, Zhe, et al.
Published: (2025)
by: Chen, Zhe, et al.
Published: (2025)
Telling Stories for Common Sense Zero-Shot Action Recognition
by: Gowda, Shreyank N, et al.
Published: (2023)
by: Gowda, Shreyank N, et al.
Published: (2023)
Head Similarity: Modeling Structured Whole-Head Appearance Beyond Face Recognition
by: Wang, Yingfeng, et al.
Published: (2026)
by: Wang, Yingfeng, et al.
Published: (2026)
Similar Items
-
Zero-shot Compositional Action Recognition with Neural Logic Constraints
by: Ye, Gefan, et al.
Published: (2025) -
ActionHub: A Large-scale Action Video Description Dataset for Zero-shot Action Recognition
by: Zhou, Jiaming, et al.
Published: (2024) -
Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP
by: Yu, Yating, et al.
Published: (2024) -
Part-aware Unified Representation of Language and Skeleton for Zero-shot Action Recognition
by: Zhu, Anqi, et al.
Published: (2024) -
Enhancing Spatio-Temporal Zero-shot Action Recognition with Language-driven Description Attributes
by: Kim, Yehna, et al.
Published: (2025)