Beyond Pedestrians: Caption-Guided CLIP Framework for High-Difficulty Video-based Person Re-Identification
Fuente:
arXiv
Saved in:
| Main Authors: | Hamano, Shogo, Wakasugi, Shunya, Sato, Tatsuhito, Nakamura, Sayaka |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CLIP-SCGI: Synthesized Caption-Guided Inversion for Person Re-Identification
by: Han, Qianru, et al.
Published: (2024)
by: Han, Qianru, et al.
Published: (2024)
Person Re-Identification System at Semantic Level based on Pedestrian Attributes Ontology
by: Ly, Ngoc Q., et al.
Published: (2025)
by: Ly, Ngoc Q., et al.
Published: (2025)
CLIP-based Camera-Agnostic Feature Learning for Intra-camera Person Re-Identification
by: Tan, Xuan, et al.
Published: (2024)
by: Tan, Xuan, et al.
Published: (2024)
S3-CLIP: Video Super Resolution for Person-ReID
by: Endrei, Tamas, et al.
Published: (2026)
by: Endrei, Tamas, et al.
Published: (2026)
From Global to Local: Rethinking CLIP Feature Aggregation for Person Re-Identification
by: Zheng, Aotian, et al.
Published: (2026)
by: Zheng, Aotian, et al.
Published: (2026)
KeyRe-ID: Keypoint-Guided Person Re-Identification using Part-Aware Representation in Videos
by: Kim, Jinseong, et al.
Published: (2025)
by: Kim, Jinseong, et al.
Published: (2025)
A CLIP-based Uncertainty Modal Modeling (UMM) Framework for Pedestrian Re-Identification in Autonomous Driving
by: Li, Jialin, et al.
Published: (2025)
by: Li, Jialin, et al.
Published: (2025)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
by: Wang, Xiao, et al.
Published: (2023)
by: Wang, Xiao, et al.
Published: (2023)
Parrot Captions Teach CLIP to Spot Text
by: Lin, Yiqi, et al.
Published: (2023)
by: Lin, Yiqi, et al.
Published: (2023)
CLIP-DFGS: A Hard Sample Mining Method for CLIP in Generalizable Person Re-Identification
by: Zhao, Huazhong, et al.
Published: (2024)
by: Zhao, Huazhong, et al.
Published: (2024)
Discriminative Pedestrian Features and Gated Channel Attention for Clothes-Changing Person Re-Identification
by: Ding, Yongkang, et al.
Published: (2024)
by: Ding, Yongkang, et al.
Published: (2024)
Video CLIP Model for Multi-View Echocardiography Interpretation
by: Takizawa, Ryo, et al.
Published: (2025)
by: Takizawa, Ryo, et al.
Published: (2025)
No Captions, No Problem: Captionless 3D-CLIP Alignment with Hard Negatives via CLIP Knowledge and LLMs
by: Sbrolli, Cristian, et al.
Published: (2024)
by: Sbrolli, Cristian, et al.
Published: (2024)
VeCLIP: Improving CLIP Training via Visual-enriched Captions
by: Lai, Zhengfeng, et al.
Published: (2023)
by: Lai, Zhengfeng, et al.
Published: (2023)
Motif Guided Graph Transformer with Combinatorial Skeleton Prototype Learning for Skeleton-Based Person Re-Identification
by: Rao, Haocong, et al.
Published: (2024)
by: Rao, Haocong, et al.
Published: (2024)
Diverse Representation Embedding for Lifelong Person Re-Identification
by: Liu, Shiben, et al.
Published: (2024)
by: Liu, Shiben, et al.
Published: (2024)
CLIP-Driven Semantic Discovery Network for Visible-Infrared Person Re-Identification
by: Yu, Xiaoyan, et al.
Published: (2024)
by: Yu, Xiaoyan, et al.
Published: (2024)
Dual-Prompt CLIP with Hybrid Visual Encoders for Occluded Person Re-Identification
by: Ji, Zhangjian, et al.
Published: (2026)
by: Ji, Zhangjian, et al.
Published: (2026)
Updating CLIP to Prefer Descriptions Over Captions
by: Zur, Amir, et al.
Published: (2024)
by: Zur, Amir, et al.
Published: (2024)
CLIP-based Synergistic Knowledge Transfer for Text-based Person Retrieval
by: Liu, Yating, et al.
Published: (2023)
by: Liu, Yating, et al.
Published: (2023)
MAMS: Model-Agnostic Module Selection Framework for Video Captioning
by: Lee, Sangho, et al.
Published: (2025)
by: Lee, Sangho, et al.
Published: (2025)
Parameter Hierarchical Optimization for Visible-Infrared Person Re-Identification
by: YU, Zeng, et al.
Published: (2024)
by: YU, Zeng, et al.
Published: (2024)
Domain Consistency Representation Learning for Lifelong Person Re-Identification
by: Liu, Shiben, et al.
Published: (2024)
by: Liu, Shiben, et al.
Published: (2024)
One-Shot Knowledge Transfer for Scalable Person Re-Identification
by: Li, Longhua, et al.
Published: (2025)
by: Li, Longhua, et al.
Published: (2025)
FedKLPR: KL-Guided Pruning-Aware Federated Learning for Person Re-Identification
by: Yu, Po-Hsien, et al.
Published: (2025)
by: Yu, Po-Hsien, et al.
Published: (2025)
Deep Learning for Video-based Person Re-Identification: A Survey
by: Islam, Khawar
Published: (2023)
by: Islam, Khawar
Published: (2023)
AGIC: Attention-Guided Image Captioning to Improve Caption Relevance
by: Teja, L. D. M. S. Sai, et al.
Published: (2025)
by: Teja, L. D. M. S. Sai, et al.
Published: (2025)
CLIP-Driven Cloth-Agnostic Feature Learning for Cloth-Changing Person Re-Identification
by: Li, Shuang, et al.
Published: (2024)
by: Li, Shuang, et al.
Published: (2024)
Captured by Captions: On Memorization and its Mitigation in CLIP Models
by: Wang, Wenhao, et al.
Published: (2025)
by: Wang, Wenhao, et al.
Published: (2025)
Dense Video Captioning using Graph-based Sentence Summarization
by: Zhang, Zhiwang, et al.
Published: (2025)
by: Zhang, Zhiwang, et al.
Published: (2025)
Hierarchical Identity Learning for Unsupervised Visible-Infrared Person Re-Identification
by: Shi, Haonan, et al.
Published: (2025)
by: Shi, Haonan, et al.
Published: (2025)
Disentangled Representations for Short-Term and Long-Term Person Re-Identification
by: Eom, Chanho, et al.
Published: (2024)
by: Eom, Chanho, et al.
Published: (2024)
Hierarchical Prompt Learning for Image- and Text-Based Person Re-Identification
by: Zhou, Linhan, et al.
Published: (2025)
by: Zhou, Linhan, et al.
Published: (2025)
A Culturally-Aware Benchmark for Person Re-Identification in Modest Attire
by: Moghaddam, Alireza Sedighi, et al.
Published: (2024)
by: Moghaddam, Alireza Sedighi, et al.
Published: (2024)
CLIPS: An Enhanced CLIP Framework for Learning with Synthetic Captions
by: Liu, Yanqing, et al.
Published: (2024)
by: Liu, Yanqing, et al.
Published: (2024)
SAIL: Similarity-Aware Guidance and Inter-Caption Augmentation-based Learning for Weakly-Supervised Dense Video Captioning
by: Kim, Ye-Chan, et al.
Published: (2026)
by: Kim, Ye-Chan, et al.
Published: (2026)
AIDA-ReID: Adaptive Intermediate Domain Adaptation for Generalizable and Source-Free Person Re-Identification
by: Iqbal, Sundas, et al.
Published: (2026)
by: Iqbal, Sundas, et al.
Published: (2026)
CountCLIP -- [Re] Teaching CLIP to Count to Ten
by: Mestha, Harshvardhan, et al.
Published: (2024)
by: Mestha, Harshvardhan, et al.
Published: (2024)
Frame-Difference Guided Dynamic Region Perception for CLIP Adaptation in Text-Video Retrieval
by: Yu, Jiaao, et al.
Published: (2025)
by: Yu, Jiaao, et al.
Published: (2025)
Revisiting Image Captioning Training Paradigm via Direct CLIP-based Optimization
by: Moratelli, Nicholas, et al.
Published: (2024)
by: Moratelli, Nicholas, et al.
Published: (2024)
Similar Items
-
CLIP-SCGI: Synthesized Caption-Guided Inversion for Person Re-Identification
by: Han, Qianru, et al.
Published: (2024) -
Person Re-Identification System at Semantic Level based on Pedestrian Attributes Ontology
by: Ly, Ngoc Q., et al.
Published: (2025) -
CLIP-based Camera-Agnostic Feature Learning for Intra-camera Person Re-Identification
by: Tan, Xuan, et al.
Published: (2024) -
S3-CLIP: Video Super Resolution for Person-ReID
by: Endrei, Tamas, et al.
Published: (2026) -
From Global to Local: Rethinking CLIP Feature Aggregation for Person Re-Identification
by: Zheng, Aotian, et al.
Published: (2026)