PLOT: Text-based Person Search with Part Slot Attention for Corresponding Part Discovery
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Park, Jicheol, Kim, Dongwon, Jeong, Boseung, Kwak, Suha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Text-based Person Search via Part-level Cross-modal Correspondence
von: Park, Jicheol, et al.
Veröffentlicht: (2024)
von: Park, Jicheol, et al.
Veröffentlicht: (2024)
Learning Audio-guided Video Representation with Gated Attention for Video-Text Retrieval
von: Jeong, Boseung, et al.
Veröffentlicht: (2025)
von: Jeong, Boseung, et al.
Veröffentlicht: (2025)
Improving Sound Source Localization with Joint Slot Attention on Image and Audio
von: Kim, Inho, et al.
Veröffentlicht: (2025)
von: Kim, Inho, et al.
Veröffentlicht: (2025)
Bootstrapping Top-down Information for Self-modulating Slot Attention
von: Kim, Dongwon, et al.
Veröffentlicht: (2024)
von: Kim, Dongwon, et al.
Veröffentlicht: (2024)
Efficient and Versatile Robust Fine-Tuning of Zero-shot Models
von: Kim, Sungyeon, et al.
Veröffentlicht: (2024)
von: Kim, Sungyeon, et al.
Veröffentlicht: (2024)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
Part-Aware Bottom-Up Group Reasoning for Fine-Grained Social Interaction Detection
von: Kim, Dongkeun, et al.
Veröffentlicht: (2025)
von: Kim, Dongkeun, et al.
Veröffentlicht: (2025)
Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens
von: Kim, Dongwon, et al.
Veröffentlicht: (2025)
von: Kim, Dongwon, et al.
Veröffentlicht: (2025)
PartCo: Part-Level Correspondence Priors Enhance Category Discovery
von: Cendra, Fernando Julio, et al.
Veröffentlicht: (2025)
von: Cendra, Fernando Julio, et al.
Veröffentlicht: (2025)
Structured State-Space Regularization for Generation-Friendly Image Tokenization
von: Lee, Jinsung, et al.
Veröffentlicht: (2026)
von: Lee, Jinsung, et al.
Veröffentlicht: (2026)
Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
RePL: Pseudo-label Refinement for Semi-supervised LiDAR Semantic Segmentation
von: Kwon, Donghyeon, et al.
Veröffentlicht: (2026)
von: Kwon, Donghyeon, et al.
Veröffentlicht: (2026)
Tempered Self-Similarity Alignment for Physically Plausible Video Generation
von: Kim, Manjin, et al.
Veröffentlicht: (2026)
von: Kim, Manjin, et al.
Veröffentlicht: (2026)
Adaptive Slot Attention: Object Discovery with Dynamic Slot Number
von: Fan, Ke, et al.
Veröffentlicht: (2024)
von: Fan, Ke, et al.
Veröffentlicht: (2024)
MINDiff: Mask-Integrated Negative Attention for Controlling Overfitting in Text-to-Image Personalization
von: Jeong, Seulgi, et al.
Veröffentlicht: (2025)
von: Jeong, Seulgi, et al.
Veröffentlicht: (2025)
CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models
von: Kwon, Minkyung, et al.
Veröffentlicht: (2025)
von: Kwon, Minkyung, et al.
Veröffentlicht: (2025)
FREST: Feature RESToration for Semantic Segmentation under Multiple Adverse Conditions
von: Lee, Sohyun, et al.
Veröffentlicht: (2024)
von: Lee, Sohyun, et al.
Veröffentlicht: (2024)
Dynamic Uncertainty Learning with Noisy Correspondence for Text-Based Person Search
von: Xie, Zequn, et al.
Veröffentlicht: (2025)
von: Xie, Zequn, et al.
Veröffentlicht: (2025)
Online Temporal Action Localization with Memory-Augmented Transformer
von: Song, Youngkil, et al.
Veröffentlicht: (2024)
von: Song, Youngkil, et al.
Veröffentlicht: (2024)
Active Label Correction for Semantic Segmentation with Foundation Models
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
Towards More Practical Group Activity Detection: A New Benchmark and Model
von: Kim, Dongkeun, et al.
Veröffentlicht: (2023)
von: Kim, Dongkeun, et al.
Veröffentlicht: (2023)
InterPartAbility: Text-Guided Part Matching for Interpretable Person Re-Identification
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2026)
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2026)
Extreme Point Supervised Instance Segmentation
von: Lee, Hyeonjun, et al.
Veröffentlicht: (2024)
von: Lee, Hyeonjun, et al.
Veröffentlicht: (2024)
Classification Matters: Improving Video Action Detection with Class-Specific Attention
von: Lee, Jinsung, et al.
Veröffentlicht: (2024)
von: Lee, Jinsung, et al.
Veröffentlicht: (2024)
Fine-Grained Image-Text Correspondence with Cost Aggregation for Open-Vocabulary Part Segmentation
von: Choi, Jiho, et al.
Veröffentlicht: (2025)
von: Choi, Jiho, et al.
Veröffentlicht: (2025)
Learning Unified Distance Metric Across Diverse Data Distributions with Parameter-Efficient Transfer Learning
von: Kim, Sungyeon, et al.
Veröffentlicht: (2023)
von: Kim, Sungyeon, et al.
Veröffentlicht: (2023)
Articulate That Object Part (ATOP): 3D Part Articulation via Text and Motion Personalization
von: Vora, Aditya, et al.
Veröffentlicht: (2025)
von: Vora, Aditya, et al.
Veröffentlicht: (2025)
Knowing Where to Focus: Attention-Guided Alignment for Text-based Person Search
von: Tan, Lei, et al.
Veröffentlicht: (2024)
von: Tan, Lei, et al.
Veröffentlicht: (2024)
GroupCoOp: Group-robust Fine-tuning via Group Prompt Learning
von: Kim, Nayeong, et al.
Veröffentlicht: (2025)
von: Kim, Nayeong, et al.
Veröffentlicht: (2025)
VIRO: Robust and Efficient Neuro-Symbolic Reasoning with Verification for Referring Expression Comprehension
von: Park, Hyejin, et al.
Veröffentlicht: (2026)
von: Park, Hyejin, et al.
Veröffentlicht: (2026)
Guided Slot Attention for Unsupervised Video Object Segmentation
von: Lee, Minhyeok, et al.
Veröffentlicht: (2023)
von: Lee, Minhyeok, et al.
Veröffentlicht: (2023)
Slot-VAE: Object-Centric Scene Generation with Slot Attention
von: Wang, Yanbo, et al.
Veröffentlicht: (2023)
von: Wang, Yanbo, et al.
Veröffentlicht: (2023)
ReFlex: Text-Guided Editing of Real Images in Rectified Flow via Mid-Step Feature Extraction and Attention Adaptation
von: Kim, Jimyeong, et al.
Veröffentlicht: (2025)
von: Kim, Jimyeong, et al.
Veröffentlicht: (2025)
Part-Attention Based Model Make Occluded Person Re-Identification Stronger
von: Chen, Zhihao, et al.
Veröffentlicht: (2024)
von: Chen, Zhihao, et al.
Veröffentlicht: (2024)
GaRA-SAM: Robustifying Segment Anything Model with Gated-Rank Adaptation
von: Lee, Sohyun, et al.
Veröffentlicht: (2025)
von: Lee, Sohyun, et al.
Veröffentlicht: (2025)
ActFusion: a Unified Diffusion Model for Action Segmentation and Anticipation
von: Gong, Dayoung, et al.
Veröffentlicht: (2024)
von: Gong, Dayoung, et al.
Veröffentlicht: (2024)
Slot Attention-based Feature Filtering for Few-Shot Learning
von: Rodenas, Javier, et al.
Veröffentlicht: (2025)
von: Rodenas, Javier, et al.
Veröffentlicht: (2025)
AutoPartGen: Autogressive 3D Part Generation and Discovery
von: Chen, Minghao, et al.
Veröffentlicht: (2025)
von: Chen, Minghao, et al.
Veröffentlicht: (2025)
Semi-supervised Text-based Person Search
von: Gao, Daming, et al.
Veröffentlicht: (2024)
von: Gao, Daming, et al.
Veröffentlicht: (2024)
Predicting Video Slot Attention Queries from Random Slot-Feature Pairs
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Improving Text-based Person Search via Part-level Cross-modal Correspondence
von: Park, Jicheol, et al.
Veröffentlicht: (2024) -
Learning Audio-guided Video Representation with Gated Attention for Video-Text Retrieval
von: Jeong, Boseung, et al.
Veröffentlicht: (2025) -
Improving Sound Source Localization with Joint Slot Attention on Image and Audio
von: Kim, Inho, et al.
Veröffentlicht: (2025) -
Bootstrapping Top-down Information for Self-modulating Slot Attention
von: Kim, Dongwon, et al.
Veröffentlicht: (2024) -
Efficient and Versatile Robust Fine-Tuning of Zero-shot Models
von: Kim, Sungyeon, et al.
Veröffentlicht: (2024)