Two-Stream Spatial-Temporal Transformer Framework for Person Identification via Natural Conversational Keypoints
Fuente:
arXiv
Salvato in:
| Autori principali: | Chapariniya, Masoumeh, Ranjbar, Hossein, Vukovic, Teodora, Ebling, Sarah, Dellwo, Volker |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Beyond Appearance: Transformer-based Person Identification from Conversational Dynamics
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2025)
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2025)
Multimodal Emotion Recognition and Sentiment Analysis in Multi-Party Conversation Contexts
di: Farhadipour, Aref, et al.
Pubblicazione: (2025)
di: Farhadipour, Aref, et al.
Pubblicazione: (2025)
Foundation Model Embeddings Meet Blended Emotions: A Multimodal Fusion Approach for the BLEMORE Challenge
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2026)
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2026)
Comparative Analysis of Modality Fusion Approaches for Audio-Visual Person Identification and Verification
di: Farhadipour, Aref, et al.
Pubblicazione: (2024)
di: Farhadipour, Aref, et al.
Pubblicazione: (2024)
Micro-Expression-Aware Avatar Fingerprinting via Inter-Frame Feature Differencing
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2026)
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2026)
Investigating Identity Signals in Conversational Facial Dynamics via Disentangled Expression Features
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2025)
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2025)
Adaptive Multimodal Person Recognition: A Robust Framework for Handling Missing Modalities
di: Farhadipour, Aref, et al.
Pubblicazione: (2025)
di: Farhadipour, Aref, et al.
Pubblicazione: (2025)
CL-UZH submission to the NIST SRE 2024 Speaker Recognition Evaluation
di: Farhadipour, Aref, et al.
Pubblicazione: (2025)
di: Farhadipour, Aref, et al.
Pubblicazione: (2025)
Deep Neural Networks for Automatic Speaker Recognition Do Not Learn Supra-Segmental Temporal Features
di: Neururer, Daniel, et al.
Pubblicazione: (2023)
di: Neururer, Daniel, et al.
Pubblicazione: (2023)
Keypoint Promptable Re-Identification
di: Somers, Vladimir, et al.
Pubblicazione: (2024)
di: Somers, Vladimir, et al.
Pubblicazione: (2024)
Towards Language-Independent Face-Voice Association with Multimodal Foundation Models
di: Farhadipour, Aref, et al.
Pubblicazione: (2025)
di: Farhadipour, Aref, et al.
Pubblicazione: (2025)
Continuous Sign Language Recognition Using Intra-inter Gloss Attention
di: Ranjbar, Hossein, et al.
Pubblicazione: (2024)
di: Ranjbar, Hossein, et al.
Pubblicazione: (2024)
LatentKeypointGAN: Controlling Images via Latent Keypoints
di: He, Xingzhe, et al.
Pubblicazione: (2021)
di: He, Xingzhe, et al.
Pubblicazione: (2021)
An Open-World, Diverse, Cross-Spatial-Temporal Benchmark for Dynamic Wild Person Re-Identification
di: Zhang, Lei, et al.
Pubblicazione: (2024)
di: Zhang, Lei, et al.
Pubblicazione: (2024)
Multi-Stream Keypoint Attention Network for Sign Language Recognition and Translation
di: Guan, Mo, et al.
Pubblicazione: (2024)
di: Guan, Mo, et al.
Pubblicazione: (2024)
Expressive Keypoints for Skeleton-based Action Recognition via Skeleton Transformation
di: Yang, Yijie, et al.
Pubblicazione: (2024)
di: Yang, Yijie, et al.
Pubblicazione: (2024)
Design and Identification of Keypoint Patches in Unstructured Environments
di: Park, Taewook, et al.
Pubblicazione: (2024)
di: Park, Taewook, et al.
Pubblicazione: (2024)
Fractional Correspondence Framework in Detection Transformer
di: Zareapoor, Masoumeh, et al.
Pubblicazione: (2025)
di: Zareapoor, Masoumeh, et al.
Pubblicazione: (2025)
Skeleton-Guided Spatial-Temporal Feature Learning for Video-Based Visible-Infrared Person Re-Identification
di: Jiang, Wenjia, et al.
Pubblicazione: (2024)
di: Jiang, Wenjia, et al.
Pubblicazione: (2024)
Categorical Keypoint Positional Embedding for Robust Animal Re-Identification
di: Lin, Yuhao, et al.
Pubblicazione: (2024)
di: Lin, Yuhao, et al.
Pubblicazione: (2024)
Good Keypoints for the Two-View Geometry Estimation Problem
di: Pakulev, Konstantin, et al.
Pubblicazione: (2025)
di: Pakulev, Konstantin, et al.
Pubblicazione: (2025)
ISLR101: an Iranian Word-Level Sign Language Recognition Dataset
di: Ranjbar, Hossein, et al.
Pubblicazione: (2025)
di: Ranjbar, Hossein, et al.
Pubblicazione: (2025)
KeyRe-ID: Keypoint-Guided Person Re-Identification using Part-Aware Representation in Videos
di: Kim, Jinseong, et al.
Pubblicazione: (2025)
di: Kim, Jinseong, et al.
Pubblicazione: (2025)
StreamSTGS: Streaming Spatial and Temporal Gaussian Grids for Real-Time Free-Viewpoint Video
di: Ke, Zhihui, et al.
Pubblicazione: (2025)
di: Ke, Zhihui, et al.
Pubblicazione: (2025)
FLAMe: Federated Learning with Attention Mechanism using Spatio-Temporal Keypoint Transformers for Pedestrian Fall Detection in Smart Cities
di: Kim, Byeonghun, et al.
Pubblicazione: (2024)
di: Kim, Byeonghun, et al.
Pubblicazione: (2024)
SRPose: Two-view Relative Pose Estimation with Sparse Keypoints
di: Yin, Rui, et al.
Pubblicazione: (2024)
di: Yin, Rui, et al.
Pubblicazione: (2024)
Progressive Cross-Stream Cooperation in Spatial and Temporal Domain for Action Localization
di: Su, Rui, et al.
Pubblicazione: (2019)
di: Su, Rui, et al.
Pubblicazione: (2019)
Learning Structure-Supporting Dependencies via Keypoint Interactive Transformer for General Mammal Pose Estimation
di: Xu, Tianyang, et al.
Pubblicazione: (2025)
di: Xu, Tianyang, et al.
Pubblicazione: (2025)
AAformer: Auto-Aligned Transformer for Person Re-Identification
di: Zhu, Kuan, et al.
Pubblicazione: (2021)
di: Zhu, Kuan, et al.
Pubblicazione: (2021)
Automatic Temporal Segmentation for Post-Stroke Rehabilitation: A Keypoint Detection and Temporal Segmentation Approach for Small Datasets
di: Lee, Jisoo, et al.
Pubblicazione: (2025)
di: Lee, Jisoo, et al.
Pubblicazione: (2025)
VideoINSTA: Zero-shot Long Video Understanding via Informative Spatial-Temporal Reasoning with LLMs
di: Liao, Ruotong, et al.
Pubblicazione: (2024)
di: Liao, Ruotong, et al.
Pubblicazione: (2024)
VCBench: A Streaming Counting Benchmark for Spatial-Temporal State Maintenance in Long Videos
di: Liu, Pengyiang, et al.
Pubblicazione: (2026)
di: Liu, Pengyiang, et al.
Pubblicazione: (2026)
PersonViT: Large-scale Self-supervised Vision Transformer for Person Re-Identification
di: Hu, Bin, et al.
Pubblicazione: (2024)
di: Hu, Bin, et al.
Pubblicazione: (2024)
Efficient Multi-Person Motion Prediction by Lightweight Spatial and Temporal Interactions
di: Zheng, Yuanhong, et al.
Pubblicazione: (2025)
di: Zheng, Yuanhong, et al.
Pubblicazione: (2025)
TSDW: A Tri-Stream Dynamic Weight Network for Cloth-Changing Person Re-Identification
di: He, Ruiqi, et al.
Pubblicazione: (2025)
di: He, Ruiqi, et al.
Pubblicazione: (2025)
Test-Time Adaptation via Cache Personalization for Facial Expression Recognition in Videos
di: Sharafi, Masoumeh, et al.
Pubblicazione: (2026)
di: Sharafi, Masoumeh, et al.
Pubblicazione: (2026)
Motion Manipulation via Unsupervised Keypoint Positioning in Face Animation
di: Li, Hong, et al.
Pubblicazione: (2026)
di: Li, Hong, et al.
Pubblicazione: (2026)
Keypoint Counting Classifiers: Turning Vision Transformers into Self-Explainable Models Without Training
di: Wickstrøm, Kristoffer, et al.
Pubblicazione: (2025)
di: Wickstrøm, Kristoffer, et al.
Pubblicazione: (2025)
Exploring Stronger Transformer Representation Learning for Occluded Person Re-Identification
di: Ji, Zhangjian, et al.
Pubblicazione: (2024)
di: Ji, Zhangjian, et al.
Pubblicazione: (2024)
Dynamic Patch-aware Enrichment Transformer for Occluded Person Re-Identification
di: Zhang, Xin, et al.
Pubblicazione: (2024)
di: Zhang, Xin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Beyond Appearance: Transformer-based Person Identification from Conversational Dynamics
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2025) -
Multimodal Emotion Recognition and Sentiment Analysis in Multi-Party Conversation Contexts
di: Farhadipour, Aref, et al.
Pubblicazione: (2025) -
Foundation Model Embeddings Meet Blended Emotions: A Multimodal Fusion Approach for the BLEMORE Challenge
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2026) -
Comparative Analysis of Modality Fusion Approaches for Audio-Visual Person Identification and Verification
di: Farhadipour, Aref, et al.
Pubblicazione: (2024) -
Micro-Expression-Aware Avatar Fingerprinting via Inter-Frame Feature Differencing
di: Chapariniya, Masoumeh, et al.
Pubblicazione: (2026)