TAI++: Text as Image for Multi-Label Image Classification by Co-Learning Transferable Prompt
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Xiangyu, Jiang, Qing-Yuan, Yang, Yang, Wu, Yi-Feng, Chen, Qing-Guo, Lu, Jianfeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Label Test-Time Adaptation with Bound Entropy Minimization
by: Wu, Xiangyu, et al.
Published: (2025)
by: Wu, Xiangyu, et al.
Published: (2025)
Text as Any-Modality for Zero-Shot Classification by Consistent Prompt Tuning
by: Wu, Xiangyu, et al.
Published: (2025)
by: Wu, Xiangyu, et al.
Published: (2025)
Multimodal Classification via Total Correlation Maximization
by: Yu, Feng, et al.
Published: (2026)
by: Yu, Feng, et al.
Published: (2026)
CDUL: CLIP-Driven Unsupervised Learning for Multi-Label Image Classification
by: Abdelfattah, Rabab, et al.
Published: (2023)
by: Abdelfattah, Rabab, et al.
Published: (2023)
Text to Image for Multi-Label Image Recognition with Joint Prompt-Adapter Learning
by: Feng, Chun-Mei, et al.
Published: (2025)
by: Feng, Chun-Mei, et al.
Published: (2025)
Adaptive Debiasing Tsallis Entropy for Test-Time Adaptation
by: Wu, Xiangyu, et al.
Published: (2026)
by: Wu, Xiangyu, et al.
Published: (2026)
The Solution for the CVPR2023 NICE Image Captioning Challenge
by: Wu, Xiangyu, et al.
Published: (2023)
by: Wu, Xiangyu, et al.
Published: (2023)
Controllable Generation with Text-to-Image Diffusion Models: A Survey
by: Cao, Pu, et al.
Published: (2024)
by: Cao, Pu, et al.
Published: (2024)
CoRe: Context-Regularized Text Embedding Learning for Text-to-Image Personalization
by: Wu, Feize, et al.
Published: (2024)
by: Wu, Feize, et al.
Published: (2024)
Rethinking Multimodal Learning from the Perspective of Mitigating Classification Ability Disproportion
by: Jiang, QingYuan, et al.
Published: (2025)
by: Jiang, QingYuan, et al.
Published: (2025)
LLM-enhanced Action-aware Multi-modal Prompt Tuning for Image-Text Matching
by: Tian, Mengxiao, et al.
Published: (2025)
by: Tian, Mengxiao, et al.
Published: (2025)
Dynamic Prompt Optimizing for Text-to-Image Generation
by: Mo, Wenyi, et al.
Published: (2024)
by: Mo, Wenyi, et al.
Published: (2024)
Category-Prompt Refined Feature Learning for Long-Tailed Multi-Label Image Classification
by: Yan, Jiexuan, et al.
Published: (2024)
by: Yan, Jiexuan, et al.
Published: (2024)
Image Super-Resolution with Text Prompt Diffusion
by: Chen, Zheng, et al.
Published: (2023)
by: Chen, Zheng, et al.
Published: (2023)
Second Place Solution of WSDM2023 Toloka Visual Question Answering Challenge
by: Wu, Xiangyu, et al.
Published: (2024)
by: Wu, Xiangyu, et al.
Published: (2024)
Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation
by: Zhao, Chenxi, et al.
Published: (2026)
by: Zhao, Chenxi, et al.
Published: (2026)
Graph Attention Transformer Network for Multi-Label Image Classification
by: Yuan, Jin, et al.
Published: (2022)
by: Yuan, Jin, et al.
Published: (2022)
A Benchmark for Multi-Lingual Vision-Language Learning in Remote Sensing Image Captioning
by: Zhou, Qing, et al.
Published: (2025)
by: Zhou, Qing, et al.
Published: (2025)
SP-Det: Self-Prompted Dual-Text Fusion for Generalized Multi-Label Lesion Detection
by: Xu, Qing, et al.
Published: (2025)
by: Xu, Qing, et al.
Published: (2025)
EPIC: Efficient Prompt Interaction for Text-Image Classification
by: Yu, Xinyao, et al.
Published: (2025)
by: Yu, Xinyao, et al.
Published: (2025)
Multi-rater Prompting for Ambiguous Medical Image Segmentation
by: Wang, Jinhong, et al.
Published: (2024)
by: Wang, Jinhong, et al.
Published: (2024)
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
by: Wu, Mingrui, et al.
Published: (2025)
by: Wu, Mingrui, et al.
Published: (2025)
TIP-I2V: A Million-Scale Real Text and Image Prompt Dataset for Image-to-Video Generation
by: Wang, Wenhao, et al.
Published: (2024)
by: Wang, Wenhao, et al.
Published: (2024)
CosalPure: Learning Concept from Group Images for Robust Co-Saliency Detection
by: Zhu, Jiayi, et al.
Published: (2024)
by: Zhu, Jiayi, et al.
Published: (2024)
Contrastive Prompts Improve Disentanglement in Text-to-Image Diffusion Models
by: Wu, Chen, et al.
Published: (2024)
by: Wu, Chen, et al.
Published: (2024)
Dual-Perspective Semantic-Aware Representation Blending for Multi-Label Image Recognition with Partial Labels
by: Pu, Tao, et al.
Published: (2022)
by: Pu, Tao, et al.
Published: (2022)
Pathology-knowledge Enhanced Multi-instance Prompt Learning for Few-shot Whole Slide Image Classification
by: Qu, Linhao, et al.
Published: (2024)
by: Qu, Linhao, et al.
Published: (2024)
Verify Claimed Text-to-Image Models via Boundary-Aware Prompt Optimization
by: Zhao, Zidong, et al.
Published: (2026)
by: Zhao, Zidong, et al.
Published: (2026)
PromptCIR: Blind Compressed Image Restoration with Prompt Learning
by: Li, Bingchen, et al.
Published: (2024)
by: Li, Bingchen, et al.
Published: (2024)
Language-Image Alignment with Fixed Text Encoders
by: Yang, Jingfeng, et al.
Published: (2025)
by: Yang, Jingfeng, et al.
Published: (2025)
StyleInject: Parameter Efficient Tuning of Text-to-Image Diffusion Models
by: Zhou, Mohan, et al.
Published: (2024)
by: Zhou, Mohan, et al.
Published: (2024)
Multimodal Classification via Modal-Aware Interactive Enhancement
by: Jiang, Qing-Yuan, et al.
Published: (2024)
by: Jiang, Qing-Yuan, et al.
Published: (2024)
Multi-Branch Non-Homogeneous Image Dehazing via Concentration Partitioning and Image Fusion
by: Zhang, Yingming, et al.
Published: (2026)
by: Zhang, Yingming, et al.
Published: (2026)
ImageDoctor: Diagnosing Text-to-Image Generation via Grounded Image Reasoning
by: Guo, Yuxiang, et al.
Published: (2025)
by: Guo, Yuxiang, et al.
Published: (2025)
Learning Semantic-Aware Threshold for Multi-Label Image Recognition with Partial Labels
by: Ruan, Haoxian, et al.
Published: (2025)
by: Ruan, Haoxian, et al.
Published: (2025)
Co-Seg: Mutual Prompt-Guided Collaborative Learning for Tissue and Nuclei Segmentation
by: Xu, Qing, et al.
Published: (2025)
by: Xu, Qing, et al.
Published: (2025)
CoMM: A Coherent Interleaved Image-Text Dataset for Multimodal Understanding and Generation
by: Chen, Wei, et al.
Published: (2024)
by: Chen, Wei, et al.
Published: (2024)
Multiple Code Hashing for Efficient Image Retrieval
by: Li, Ming-Wei, et al.
Published: (2020)
by: Li, Ming-Wei, et al.
Published: (2020)
Self-Paced Learning for Images of Antinuclear Antibodies
by: Jiang, Yiyang, et al.
Published: (2025)
by: Jiang, Yiyang, et al.
Published: (2025)
Evaluating Image Caption via Cycle-consistent Text-to-Image Generation
by: Cui, Tianyu, et al.
Published: (2025)
by: Cui, Tianyu, et al.
Published: (2025)
Similar Items
-
Multi-Label Test-Time Adaptation with Bound Entropy Minimization
by: Wu, Xiangyu, et al.
Published: (2025) -
Text as Any-Modality for Zero-Shot Classification by Consistent Prompt Tuning
by: Wu, Xiangyu, et al.
Published: (2025) -
Multimodal Classification via Total Correlation Maximization
by: Yu, Feng, et al.
Published: (2026) -
CDUL: CLIP-Driven Unsupervised Learning for Multi-Label Image Classification
by: Abdelfattah, Rabab, et al.
Published: (2023) -
Text to Image for Multi-Label Image Recognition with Joint Prompt-Adapter Learning
by: Feng, Chun-Mei, et al.
Published: (2025)