Saved in:
| Main Authors: | Lin, Rifen, Wang, Alex Jinpeng, Mo, Jiawei, Li, Min |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.13150 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MoChat: Joints-Grouped Spatio-Temporal Grounding LLM for Multi-Turn Motion Comprehension and Description
by: Mo, Jiawei, et al.
Published: (2024)
by: Mo, Jiawei, et al.
Published: (2024)
Text Speaks Louder than Vision: ASCII Art Reveals Textual Biases in Vision-Language Models
by: Wang, Zhaochen, et al.
Published: (2025)
by: Wang, Zhaochen, et al.
Published: (2025)
Disentangled Concepts Speak Louder Than Words: Explainable Video Action Recognition
by: Lee, Jongseo, et al.
Published: (2025)
by: Lee, Jongseo, et al.
Published: (2025)
Skeleton-Guided Spatial-Temporal Feature Learning for Video-Based Visible-Infrared Person Re-Identification
by: Jiang, Wenjia, et al.
Published: (2024)
by: Jiang, Wenjia, et al.
Published: (2024)
ShapeSpeak: Body Shape-Aware Textual Alignment for Visible-Infrared Person Re-Identification
by: Yan, Shuanglin, et al.
Published: (2025)
by: Yan, Shuanglin, et al.
Published: (2025)
A Dual-stage Prompt-driven Privacy-preserving Paradigm for Person Re-Identification
by: Li, Ruolin, et al.
Published: (2025)
by: Li, Ruolin, et al.
Published: (2025)
Motif Guided Graph Transformer with Combinatorial Skeleton Prototype Learning for Skeleton-Based Person Re-Identification
by: Rao, Haocong, et al.
Published: (2024)
by: Rao, Haocong, et al.
Published: (2024)
Clothes-Changing Person Re-identification Based On Skeleton Dynamics
by: Joseph, Asaf, et al.
Published: (2025)
by: Joseph, Asaf, et al.
Published: (2025)
Color Space Learning for Cross-Color Person Re-Identification
by: Nie, Jiahao, et al.
Published: (2024)
by: Nie, Jiahao, et al.
Published: (2024)
Images Speak Louder Than Scores: Failure Mode Escape for Enhancing Generative Quality
by: Shao, Jie, et al.
Published: (2025)
by: Shao, Jie, et al.
Published: (2025)
Mix-Modality Person Re-Identification: A New and Practical Paradigm
by: Liu, Wei, et al.
Published: (2024)
by: Liu, Wei, et al.
Published: (2024)
Video-Level Language-Driven Video-Based Visible-Infrared Person Re-Identification
by: Li, Shuang, et al.
Published: (2025)
by: Li, Shuang, et al.
Published: (2025)
Hierarchical Prompt Learning for Image- and Text-Based Person Re-Identification
by: Zhou, Linhan, et al.
Published: (2025)
by: Zhou, Linhan, et al.
Published: (2025)
Controllable Complex Human Motion Video Generation via Text-to-Skeleton Cascades
by: Taghipour, Ashkan, et al.
Published: (2026)
by: Taghipour, Ashkan, et al.
Published: (2026)
Marrying Text-to-Motion Generation with Skeleton-Based Action Recognition
by: Kuang, Jidong, et al.
Published: (2026)
by: Kuang, Jidong, et al.
Published: (2026)
Aligned Divergent Pathways for Omni-Domain Generalized Person Re-Identification
by: Ang, Eugene P. W., et al.
Published: (2024)
by: Ang, Eugene P. W., et al.
Published: (2024)
TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering
by: Mao, Dongxing, et al.
Published: (2026)
by: Mao, Dongxing, et al.
Published: (2026)
Unleashing the Potential of Tracklets for Unsupervised Video Person Re-Identification
by: Meng, Nanxing, et al.
Published: (2024)
by: Meng, Nanxing, et al.
Published: (2024)
Skeleton-to-Image Encoding: Enabling Skeleton Representation Learning via Vision-Pretrained Models
by: Yang, Siyuan, et al.
Published: (2026)
by: Yang, Siyuan, et al.
Published: (2026)
MotionBooth: Motion-Aware Customized Text-to-Video Generation
by: Wu, Jianzong, et al.
Published: (2024)
by: Wu, Jianzong, et al.
Published: (2024)
Causal Bootstrapped Alignment for Unsupervised Video-Based Visible-Infrared Person Re-Identification
by: Li, Shuang, et al.
Published: (2026)
by: Li, Shuang, et al.
Published: (2026)
A Survey on 3D Skeleton Based Person Re-Identification: Taxonomy, Advances, Challenges, and Interdisciplinary Prospects
by: Rao, Haocong, et al.
Published: (2024)
by: Rao, Haocong, et al.
Published: (2024)
SA-Person: Text-Based Person Retrieval with Scene-aware Re-ranking
by: Xu, Yingjia, et al.
Published: (2025)
by: Xu, Yingjia, et al.
Published: (2025)
KeyRe-ID: Keypoint-Guided Person Re-Identification using Part-Aware Representation in Videos
by: Kim, Jinseong, et al.
Published: (2025)
by: Kim, Jinseong, et al.
Published: (2025)
When Images Speak Louder: Mitigating Language Bias-induced Hallucinations in VLMs through Cross-Modal Guidance
by: Cao, Jinjin, et al.
Published: (2025)
by: Cao, Jinjin, et al.
Published: (2025)
Leveraging Visual Tokens for Extended Text Contexts in Multi-Modal Learning
by: Wang, Alex Jinpeng, et al.
Published: (2024)
by: Wang, Alex Jinpeng, et al.
Published: (2024)
Minimizing the Pretraining Gap: Domain-aligned Text-Based Person Retrieval
by: Yang, Shuyu, et al.
Published: (2025)
by: Yang, Shuyu, et al.
Published: (2025)
Thinking Before Matching: A Reinforcement Reasoning Paradigm Towards General Person Re-Identification
by: Zhang, Quan, et al.
Published: (2026)
by: Zhang, Quan, et al.
Published: (2026)
Beyond Words: Advancing Long-Text Image Generation via Multimodal Autoregressive Models
by: Wang, Alex Jinpeng, et al.
Published: (2025)
by: Wang, Alex Jinpeng, et al.
Published: (2025)
X-ReID: Multi-granularity Information Interaction for Video-Based Visible-Infrared Person Re-Identification
by: Yu, Chenyang, et al.
Published: (2025)
by: Yu, Chenyang, et al.
Published: (2025)
Skeleton Motion Words for Unsupervised Skeleton-Based Temporal Action Segmentation
by: Gökay, Uzay, et al.
Published: (2025)
by: Gökay, Uzay, et al.
Published: (2025)
TextAtlas5M: A Large-scale Dataset for Dense Text Image Generation
by: Wang, Alex Jinpeng, et al.
Published: (2025)
by: Wang, Alex Jinpeng, et al.
Published: (2025)
MSP-ReID: Hairstyle-Robust Cloth-Changing Person Re-Identification
by: He, Xiangyang, et al.
Published: (2026)
by: He, Xiangyang, et al.
Published: (2026)
Images Speak Louder than Words: Understanding and Mitigating Bias in Vision-Language Model from a Causal Mediation Perspective
by: Weng, Zhaotian, et al.
Published: (2024)
by: Weng, Zhaotian, et al.
Published: (2024)
View-Aware Semantic Alignment for Aerial-Ground Person Re-Identification
by: Zhang, Quan, et al.
Published: (2026)
by: Zhang, Quan, et al.
Published: (2026)
Clothes-Changing Person Re-Identification with Feasibility-Aware Intermediary Matching
by: Zhao, Jiahe, et al.
Published: (2024)
by: Zhao, Jiahe, et al.
Published: (2024)
Local-Aware Global Attention Network for Person Re-Identification Based on Body and Hand Images
by: Baisa, Nathanael L.
Published: (2022)
by: Baisa, Nathanael L.
Published: (2022)
Deep Learning for Video-based Person Re-Identification: A Survey
by: Islam, Khawar
Published: (2023)
by: Islam, Khawar
Published: (2023)
VILLS -- Video-Image Learning to Learn Semantics for Person Re-Identification
by: Huang, Siyuan, et al.
Published: (2023)
by: Huang, Siyuan, et al.
Published: (2023)
Learning to Learn Transferable Generative Attack for Person Re-Identification
by: Bian, Yuan, et al.
Published: (2024)
by: Bian, Yuan, et al.
Published: (2024)
Similar Items
-
MoChat: Joints-Grouped Spatio-Temporal Grounding LLM for Multi-Turn Motion Comprehension and Description
by: Mo, Jiawei, et al.
Published: (2024) -
Text Speaks Louder than Vision: ASCII Art Reveals Textual Biases in Vision-Language Models
by: Wang, Zhaochen, et al.
Published: (2025) -
Disentangled Concepts Speak Louder Than Words: Explainable Video Action Recognition
by: Lee, Jongseo, et al.
Published: (2025) -
Skeleton-Guided Spatial-Temporal Feature Learning for Video-Based Visible-Infrared Person Re-Identification
by: Jiang, Wenjia, et al.
Published: (2024) -
ShapeSpeak: Body Shape-Aware Textual Alignment for Visible-Infrared Person Re-Identification
by: Yan, Shuanglin, et al.
Published: (2025)