PromptGAR: Flexible Promptive Group Activity Recognition
Fuente:
arXiv
Guardado en:
| Autores principales: | Jin, Zhangyu, Feng, Andrew, Chemburkar, Ankur, De Melo, Celso M. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SoGAR: Self-supervised Spatiotemporal Attention-based Social Group Activity Recognition
por: Chappa, Naga VS Raviteja, et al.
Publicado: (2023)
por: Chappa, Naga VS Raviteja, et al.
Publicado: (2023)
LiGAR: LiDAR-Guided Hierarchical Transformer for Multi-Modal Group Activity Recognition
por: Chappa, Naga Venkata Sai Raviteja, et al.
Publicado: (2024)
por: Chappa, Naga Venkata Sai Raviteja, et al.
Publicado: (2024)
SAT-SKYLINES: 3D Building Generation from Satellite Imagery and Coarse Geometric Priors
por: Jin, Zhangyu, et al.
Publicado: (2025)
por: Jin, Zhangyu, et al.
Publicado: (2025)
Zero-Shot Scene Understanding for Automatic Target Recognition Using Large Vision-Language Models
por: Ranasinghe, Yasiru, et al.
Publicado: (2025)
por: Ranasinghe, Yasiru, et al.
Publicado: (2025)
GDPO-Listener: Expressive Interactive Head Generation via Auto-Regressive Flow Matching and Group reward-Decoupled Policy Optimization
por: Jin, Zhangyu, et al.
Publicado: (2026)
por: Jin, Zhangyu, et al.
Publicado: (2026)
SHARDeg: A Benchmark for Skeletal Human Action Recognition in Degraded Scenarios
por: Malzard, Simon, et al.
Publicado: (2025)
por: Malzard, Simon, et al.
Publicado: (2025)
An Evaluation of Large Pre-Trained Models for Gesture Recognition using Synthetic Videos
por: Reddy, Arun, et al.
Publicado: (2024)
por: Reddy, Arun, et al.
Publicado: (2024)
ActivityCLIP: Enhancing Group Activity Recognition by Mining Complementary Information from Text to Supplement Image Modality
por: Xu, Guoliang, et al.
Publicado: (2024)
por: Xu, Guoliang, et al.
Publicado: (2024)
Flexible and Efficient Spatio-Temporal Transformer for Sequential Visual Place Recognition
por: Kiu, Yu, et al.
Publicado: (2025)
por: Kiu, Yu, et al.
Publicado: (2025)
Pixels or Positions? Benchmarking Modalities in Group Activity Recognition
por: Karki, Drishya, et al.
Publicado: (2025)
por: Karki, Drishya, et al.
Publicado: (2025)
STMT: A Spatial-Temporal Mesh Transformer for MoCap-Based Action Recognition
por: Zhu, Xiaoyu, et al.
Publicado: (2023)
por: Zhu, Xiaoyu, et al.
Publicado: (2023)
Discrete Facial Encoding: : A Framework for Data-driven Facial Display Discovery
por: Tran, Minh, et al.
Publicado: (2025)
por: Tran, Minh, et al.
Publicado: (2025)
Skeleton-based Group Activity Recognition via Spatial-Temporal Panoramic Graph
por: Li, Zhengcen, et al.
Publicado: (2024)
por: Li, Zhengcen, et al.
Publicado: (2024)
Decoupled Prompt-Adapter Tuning for Continual Activity Recognition
por: Fu, Di, et al.
Publicado: (2024)
por: Fu, Di, et al.
Publicado: (2024)
Deep Correlated Prompting for Visual Recognition with Missing Modalities
por: Hu, Lianyu, et al.
Publicado: (2024)
por: Hu, Lianyu, et al.
Publicado: (2024)
AHA -- Predicting What Matters Next: Online Highlight Detection Without Looking Ahead
por: Chang, Aiden, et al.
Publicado: (2025)
por: Chang, Aiden, et al.
Publicado: (2025)
Synergistic Prompting for Robust Visual Recognition with Missing Modalities
por: Zhang, Zhihui, et al.
Publicado: (2025)
por: Zhang, Zhihui, et al.
Publicado: (2025)
Is Temporal Prompting All We Need For Limited Labeled Action Recognition?
por: Gowda, Shreyank N, et al.
Publicado: (2025)
por: Gowda, Shreyank N, et al.
Publicado: (2025)
EgoPrompt: Prompt Learning for Egocentric Action Recognition
por: Lyu, Huaihai, et al.
Publicado: (2025)
por: Lyu, Huaihai, et al.
Publicado: (2025)
VicKAM: Visual Conceptual Knowledge Guided Action Map for Weakly Supervised Group Activity Recognition
por: Wang, Zhuming, et al.
Publicado: (2025)
por: Wang, Zhuming, et al.
Publicado: (2025)
Emotion Recognition from the perspective of Activity Recognition
por: Nagendra, Savinay, et al.
Publicado: (2024)
por: Nagendra, Savinay, et al.
Publicado: (2024)
F-ViTA: Foundation Model Guided Visible to Thermal Translation
por: Paranjape, Jay N., et al.
Publicado: (2025)
por: Paranjape, Jay N., et al.
Publicado: (2025)
A Mamba-based Siamese Network for Remote Sensing Change Detection
por: Paranjape, Jay N., et al.
Publicado: (2024)
por: Paranjape, Jay N., et al.
Publicado: (2024)
Multimodal Prompt Learning with Missing Modalities for Sentiment Analysis and Emotion Recognition
por: Guo, Zirun, et al.
Publicado: (2024)
por: Guo, Zirun, et al.
Publicado: (2024)
Synthetic-to-Real Domain Adaptation for Action Recognition: A Dataset and Baseline Performances
por: Reddy, Arun V., et al.
Publicado: (2023)
por: Reddy, Arun V., et al.
Publicado: (2023)
Group Activity Recognition using Unreliable Tracked Pose
por: Thilakarathne, Haritha, et al.
Publicado: (2024)
por: Thilakarathne, Haritha, et al.
Publicado: (2024)
Design and Analysis of Efficient Attention in Transformers for Social Group Activity Recognition
por: Tamura, Masato
Publicado: (2024)
por: Tamura, Masato
Publicado: (2024)
Improving Object Detection by Modifying Synthetic Data with Explainable AI
por: Mital, Nitish, et al.
Publicado: (2024)
por: Mital, Nitish, et al.
Publicado: (2024)
Research on Defect Detection Method of Motor Control Board Based on Image Processing
por: Huang, Jingde, et al.
Publicado: (2025)
por: Huang, Jingde, et al.
Publicado: (2025)
Text to Image for Multi-Label Image Recognition with Joint Prompt-Adapter Learning
por: Feng, Chun-Mei, et al.
Publicado: (2025)
por: Feng, Chun-Mei, et al.
Publicado: (2025)
Prompt-guided Disentangled Representation for Action Recognition
por: Wu, Tianci, et al.
Publicado: (2025)
por: Wu, Tianci, et al.
Publicado: (2025)
Visual Prompting in LLMs for Enhancing Emotion Recognition
por: Zhang, Qixuan, et al.
Publicado: (2024)
por: Zhang, Qixuan, et al.
Publicado: (2024)
Teaching Prompts to Coordinate: Hierarchical Layer-Grouped Prompt Tuning for Continual Learning
por: Jiang, Shengqin, et al.
Publicado: (2025)
por: Jiang, Shengqin, et al.
Publicado: (2025)
A Human-in-the-Loop Framework for Efficient Prompt Selection in Microscopy Vision-Language Models
por: Kandiyana, Abhiram, et al.
Publicado: (2026)
por: Kandiyana, Abhiram, et al.
Publicado: (2026)
Referring Change Detection in Remote Sensing Imagery
por: Korkmaz, Yilmaz, et al.
Publicado: (2025)
por: Korkmaz, Yilmaz, et al.
Publicado: (2025)
Diff-Prompt: Diffusion-Driven Prompt Generator with Mask Supervision
por: Yan, Weicai, et al.
Publicado: (2025)
por: Yan, Weicai, et al.
Publicado: (2025)
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
por: Wang, Zhifeng, et al.
Publicado: (2025)
por: Wang, Zhifeng, et al.
Publicado: (2025)
Novel Semantic Prompting for Zero-Shot Action Recognition
por: Iqbal, Salman, et al.
Publicado: (2026)
por: Iqbal, Salman, et al.
Publicado: (2026)
HyperTokens: Controlling Token Dynamics for Continual Video-Language Understanding
por: Nguyen, Toan, et al.
Publicado: (2026)
por: Nguyen, Toan, et al.
Publicado: (2026)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
por: Wang, Xiao, et al.
Publicado: (2023)
por: Wang, Xiao, et al.
Publicado: (2023)
Ejemplares similares
-
SoGAR: Self-supervised Spatiotemporal Attention-based Social Group Activity Recognition
por: Chappa, Naga VS Raviteja, et al.
Publicado: (2023) -
LiGAR: LiDAR-Guided Hierarchical Transformer for Multi-Modal Group Activity Recognition
por: Chappa, Naga Venkata Sai Raviteja, et al.
Publicado: (2024) -
SAT-SKYLINES: 3D Building Generation from Satellite Imagery and Coarse Geometric Priors
por: Jin, Zhangyu, et al.
Publicado: (2025) -
Zero-Shot Scene Understanding for Automatic Target Recognition Using Large Vision-Language Models
por: Ranasinghe, Yasiru, et al.
Publicado: (2025) -
GDPO-Listener: Expressive Interactive Head Generation via Auto-Regressive Flow Matching and Group reward-Decoupled Policy Optimization
por: Jin, Zhangyu, et al.
Publicado: (2026)