Following the Clues: Experiments on Person Re-ID using Cross-Modal Intelligence
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Aufschläger, Robert, Shoeb, Youssef, Nowzad, Azarm, Heigl, Michael, Bally, Fabian, Schramm, Martin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Neural Networks for Intelligent Data-Driven Development
von: Shoeb, Youssef, et al.
Veröffentlicht: (2025)
von: Shoeb, Youssef, et al.
Veröffentlicht: (2025)
Out-of-Distribution Segmentation in Autonomous Driving: Problems and State of the Art
von: Shoeb, Youssef, et al.
Veröffentlicht: (2025)
von: Shoeb, Youssef, et al.
Veröffentlicht: (2025)
ClustEm4Ano: Clustering Text Embeddings of Nominal Textual Attributes for Microdata Anonymization
von: Aufschläger, Robert, et al.
Veröffentlicht: (2024)
von: Aufschläger, Robert, et al.
Veröffentlicht: (2024)
Segment-Level Road Obstacle Detection Using Visual Foundation Model Priors and Likelihood Ratios
von: Shoeb, Youssef, et al.
Veröffentlicht: (2024)
von: Shoeb, Youssef, et al.
Veröffentlicht: (2024)
Towards Context-Aware Image Anonymization with Multi-Agent Reasoning
von: Aufschläger, Robert, et al.
Veröffentlicht: (2026)
von: Aufschläger, Robert, et al.
Veröffentlicht: (2026)
CMCC-ReID: Cross-Modality Clothing-Change Person Re-Identification
von: Xu, Haoxuan, et al.
Veröffentlicht: (2026)
von: Xu, Haoxuan, et al.
Veröffentlicht: (2026)
Building Blocks for Robust and Effective Semi-Supervised Real-World Object Detection
von: Sbeyti, Moussa Kassem, et al.
Veröffentlicht: (2025)
von: Sbeyti, Moussa Kassem, et al.
Veröffentlicht: (2025)
MLLMReID: Multimodal Large Language Model-based Person Re-identification
von: Yang, Shan, et al.
Veröffentlicht: (2024)
von: Yang, Shan, et al.
Veröffentlicht: (2024)
PriMod4AI: Lifecycle-Aware Privacy Threat Modeling for AI Systems using LLM
von: Savaliya, Gautam, et al.
Veröffentlicht: (2026)
von: Savaliya, Gautam, et al.
Veröffentlicht: (2026)
FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-Identification
von: Sun, Zhen, et al.
Veröffentlicht: (2025)
von: Sun, Zhen, et al.
Veröffentlicht: (2025)
How Could Generative AI Support Compliance with the EU AI Act? A Review for Safe Automated Driving Perception
von: Keser, Mert, et al.
Veröffentlicht: (2024)
von: Keser, Mert, et al.
Veröffentlicht: (2024)
A Likelihood Ratio-Based Approach to Segmenting Unknown Objects
von: Nayal, Nazir, et al.
Veröffentlicht: (2024)
von: Nayal, Nazir, et al.
Veröffentlicht: (2024)
Efficient Bilateral Cross-Modality Cluster Matching for Unsupervised Visible-Infrared Person ReID
von: Cheng, De, et al.
Veröffentlicht: (2023)
von: Cheng, De, et al.
Veröffentlicht: (2023)
Identity Clue Refinement and Enhancement for Visible-Infrared Person Re-Identification
von: Zhang, Guoqing, et al.
Veröffentlicht: (2025)
von: Zhang, Guoqing, et al.
Veröffentlicht: (2025)
Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration Rate
von: Huang, Qidong, et al.
Veröffentlicht: (2024)
von: Huang, Qidong, et al.
Veröffentlicht: (2024)
FSMR: A Feature Swapping Multi-modal Reasoning Approach with Joint Textual and Visual Clues
von: Li, Shuang, et al.
Veröffentlicht: (2024)
von: Li, Shuang, et al.
Veröffentlicht: (2024)
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
von: Csizmadia, Daniel, et al.
Veröffentlicht: (2025)
von: Csizmadia, Daniel, et al.
Veröffentlicht: (2025)
Cross-Modality Perturbation Synergy Attack for Person Re-identification
von: Gong, Yunpeng, et al.
Veröffentlicht: (2024)
von: Gong, Yunpeng, et al.
Veröffentlicht: (2024)
Cross-Modal Rationale Transfer for Explainable Humanitarian Classification on Social Media
von: Nguyen, Thi Huyen, et al.
Veröffentlicht: (2026)
von: Nguyen, Thi Huyen, et al.
Veröffentlicht: (2026)
Cross-Modal Obfuscation for Jailbreak Attacks on Large Vision-Language Models
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
Cross-Modal Retrieval for Motion and Text via DropTriple Loss
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models
von: Zhu, Tinghui, et al.
Veröffentlicht: (2024)
von: Zhu, Tinghui, et al.
Veröffentlicht: (2024)
Dynamic Clue Bottlenecks: Towards Interpretable-by-Design Visual Question Answering
von: Fu, Xingyu, et al.
Veröffentlicht: (2023)
von: Fu, Xingyu, et al.
Veröffentlicht: (2023)
Exploring Typographic Visual Prompts Injection Threats in Cross-Modality Generation Models
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
CoLA: Cross-Modal Low-rank Adaptation for Multimodal Downstream Tasks
von: Suharitdamrong, Wish, et al.
Veröffentlicht: (2026)
von: Suharitdamrong, Wish, et al.
Veröffentlicht: (2026)
Cross-Modal Adapter for Vision-Language Retrieval
von: Jiang, Haojun, et al.
Veröffentlicht: (2022)
von: Jiang, Haojun, et al.
Veröffentlicht: (2022)
LLM2CLIP: Powerful Language Model Unlocks Richer Cross-Modality Representation
von: Huang, Weiquan, et al.
Veröffentlicht: (2024)
von: Huang, Weiquan, et al.
Veröffentlicht: (2024)
Enhancing Chest X-ray Classification through Knowledge Injection in Cross-Modality Learning
von: Yan, Yang, et al.
Veröffentlicht: (2025)
von: Yan, Yang, et al.
Veröffentlicht: (2025)
Leveraging Entity Information for Cross-Modality Correlation Learning: The Entity-Guided Multimodal Summarization
von: Zhang, Yanghai, et al.
Veröffentlicht: (2024)
von: Zhang, Yanghai, et al.
Veröffentlicht: (2024)
EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data
von: Lin, Dongyan, et al.
Veröffentlicht: (2026)
von: Lin, Dongyan, et al.
Veröffentlicht: (2026)
CROME: Cross-Modal Adapters for Efficient Multimodal LLM
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2024)
Do Text Edits Generalize to Visual Generation? Benchmarking Cross-Modal Knowledge Editing in UMMs
von: Gao, Xin, et al.
Veröffentlicht: (2026)
von: Gao, Xin, et al.
Veröffentlicht: (2026)
Synthetic-To-Real Video Person Re-ID
von: Zhang, Xiangqun, et al.
Veröffentlicht: (2024)
von: Zhang, Xiangqun, et al.
Veröffentlicht: (2024)
LinguDistill: Recovering Linguistic Ability in Vision-Language Models via Selective Cross-Modal Distillation
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2026)
von: Irawan, Patrick Amadeus, et al.
Veröffentlicht: (2026)
X-VILA: Cross-Modality Alignment for Large Language Model
von: Ye, Hanrong, et al.
Veröffentlicht: (2024)
von: Ye, Hanrong, et al.
Veröffentlicht: (2024)
Vision-Language Models Create Cross-Modal Task Representations
von: Luo, Grace, et al.
Veröffentlicht: (2024)
von: Luo, Grace, et al.
Veröffentlicht: (2024)
Text-to-Image Cross-Modal Generation: A Systematic Review
von: Żelaszczyk, Maciej, et al.
Veröffentlicht: (2024)
von: Żelaszczyk, Maciej, et al.
Veröffentlicht: (2024)
Investigating Cross-Modal Skill Injection: Scenarios, Methods, and Hyperparameters
von: Xu, Zhiyu, et al.
Veröffentlicht: (2026)
von: Xu, Zhiyu, et al.
Veröffentlicht: (2026)
Region-R1: Reinforcing Query-Side Region Cropping for Multi-Modal Re-Ranking
von: Hu, Chan-Wei, et al.
Veröffentlicht: (2026)
von: Hu, Chan-Wei, et al.
Veröffentlicht: (2026)
MoD-DPO: Towards Mitigating Cross-modal Hallucinations in Omni LLMs using Modality Decoupled Preference Optimization
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2026)
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Adaptive Neural Networks for Intelligent Data-Driven Development
von: Shoeb, Youssef, et al.
Veröffentlicht: (2025) -
Out-of-Distribution Segmentation in Autonomous Driving: Problems and State of the Art
von: Shoeb, Youssef, et al.
Veröffentlicht: (2025) -
ClustEm4Ano: Clustering Text Embeddings of Nominal Textual Attributes for Microdata Anonymization
von: Aufschläger, Robert, et al.
Veröffentlicht: (2024) -
Segment-Level Road Obstacle Detection Using Visual Foundation Model Priors and Likelihood Ratios
von: Shoeb, Youssef, et al.
Veröffentlicht: (2024) -
Towards Context-Aware Image Anonymization with Multi-Agent Reasoning
von: Aufschläger, Robert, et al.
Veröffentlicht: (2026)