Visual-Semantic Decomposition and Partial Alignment for Document-based Zero-Shot Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Qu, Xiangyan, Yu, Jing, Gai, Keke, Zhuang, Jiamin, Tang, Yuanmin, Xiong, Gang, Gou, Gaopeng, Wu, Qi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Missing Target-Relevant Information Prediction with World Model for Accurate Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2025)
by: Tang, Yuanmin, et al.
Published: (2025)
Denoise-I2W: Mapping Images to Denoising Words for Accurate Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2024)
by: Tang, Yuanmin, et al.
Published: (2024)
MADS: Multi-Attribute Document Supervision for Zero-Shot Image Classification
by: Qu, Xiangyan, et al.
Published: (2025)
by: Qu, Xiangyan, et al.
Published: (2025)
T2VParser: Adaptive Decomposition Tokens for Partial Alignment in Text to Video Retrieval
by: Li, Yili, et al.
Published: (2025)
by: Li, Yili, et al.
Published: (2025)
ProAPO: Progressively Automatic Prompt Optimization for Visual Classification
by: Qu, Xiangyan, et al.
Published: (2025)
by: Qu, Xiangyan, et al.
Published: (2025)
Reason-before-Retrieve: One-Stage Reflective Chain-of-Thoughts for Training-Free Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2024)
by: Tang, Yuanmin, et al.
Published: (2024)
IIU: Independent Inference Units for Knowledge-based Visual Question Answering
by: Li, Yili, et al.
Published: (2024)
by: Li, Yili, et al.
Published: (2024)
From Scale to Speed: Adaptive Test-Time Scaling for Image Editing
by: Qu, Xiangyan, et al.
Published: (2026)
by: Qu, Xiangyan, et al.
Published: (2026)
T2VIndexer: A Generative Video Indexer for Efficient Text-Video Retrieval
by: Li, Yili, et al.
Published: (2024)
by: Li, Yili, et al.
Published: (2024)
Watermarking Visual Concepts for Diffusion Models
by: Lei, Liangqi, et al.
Published: (2024)
by: Lei, Liangqi, et al.
Published: (2024)
ALIEN: Analytic Latent Watermarking for Controllable Generation
by: Lei, Liangqi, et al.
Published: (2026)
by: Lei, Liangqi, et al.
Published: (2026)
DecETT: Accurate App Fingerprinting Under Encrypted Tunnels via Dual Decouple-based Semantic Enhancement
by: Gu, Zheyuan, et al.
Published: (2025)
by: Gu, Zheyuan, et al.
Published: (2025)
Respond to Change with Constancy: Instruction-tuning with LLM for Non-I.I.D. Network Traffic Classification
by: Lin, Xinjie, et al.
Published: (2025)
by: Lin, Xinjie, et al.
Published: (2025)
A Vision-Language Pre-training Model-Guided Approach for Mitigating Backdoor Attacks in Federated Learning
by: Gai, Keke, et al.
Published: (2025)
by: Gai, Keke, et al.
Published: (2025)
Secure and Efficient Watermarking for Latent Diffusion Models in Model Distribution Scenarios
by: Lei, Liangqi, et al.
Published: (2025)
by: Lei, Liangqi, et al.
Published: (2025)
Adaptive Prototype Knowledge Transfer for Federated Learning with Mixed Modalities and Heterogeneous Tasks
by: Gai, Keke, et al.
Published: (2025)
by: Gai, Keke, et al.
Published: (2025)
AGATE: Stealthy Black-box Watermarking for Multimodal Model Copyright Protection
by: Gao, Jianbo, et al.
Published: (2025)
by: Gao, Jianbo, et al.
Published: (2025)
Vertical Federated Continual Learning via Evolving Prototype Knowledge
by: Wang, Shuo, et al.
Published: (2025)
by: Wang, Shuo, et al.
Published: (2025)
PCDiff: Proactive Control for Ownership Protection in Diffusion Models with Watermark Compatibility
by: Gai, Keke, et al.
Published: (2025)
by: Gai, Keke, et al.
Published: (2025)
Deep Semantic-Visual Alignment for Zero-Shot Remote Sensing Image Scene Classification
by: Xu, Wenjia, et al.
Published: (2024)
by: Xu, Wenjia, et al.
Published: (2024)
Binary Linear Tree Commitment-based Ownership Protection for Distributed Machine Learning
by: Xie, Tianxiu, et al.
Published: (2024)
by: Xie, Tianxiu, et al.
Published: (2024)
Visual and Semantic Prompt Collaboration for Generalized Zero-Shot Learning
by: Jiang, Huajie, et al.
Published: (2025)
by: Jiang, Huajie, et al.
Published: (2025)
EVA: Mixture-of-Experts Semantic Variant Alignment for Compositional Zero-Shot Learning
by: Zhang, Xiao, et al.
Published: (2025)
by: Zhang, Xiao, et al.
Published: (2025)
Language-Driven Visual Consensus for Zero-Shot Semantic Segmentation
by: Zhang, Zicheng, et al.
Published: (2024)
by: Zhang, Zicheng, et al.
Published: (2024)
A2-DIDM: Privacy-preserving Accumulator-enabled Auditing for Distributed Identity of DNN Model
by: Xie, Tianxiu, et al.
Published: (2024)
by: Xie, Tianxiu, et al.
Published: (2024)
DiffuseTrace: A Transparent and Flexible Watermarking Scheme for Latent Diffusion Model
by: Lei, Liangqi, et al.
Published: (2024)
by: Lei, Liangqi, et al.
Published: (2024)
MalRAG: A Retrieval-Augmented LLM Framework for Open-set Malicious Traffic Identification
by: Luo, Xiang, et al.
Published: (2025)
by: Luo, Xiang, et al.
Published: (2025)
SLVC-DIDA: Signature-less Verifiable Credential-based Issuer-hiding and Multi-party Authentication for Decentralized Identity
by: Xie, Tianxiu, et al.
Published: (2025)
by: Xie, Tianxiu, et al.
Published: (2025)
Semantic-Space-Intervened Diffusive Alignment for Visual Classification
by: Li, Zixuan, et al.
Published: (2025)
by: Li, Zixuan, et al.
Published: (2025)
Enhancing Zero-Shot Brain Tumor Subtype Classification via Fine-Grained Patch-Text Alignment
by: Gan, Lubin, et al.
Published: (2025)
by: Gan, Lubin, et al.
Published: (2025)
Zero-Shot Skeleton-based Action Recognition with Dual Visual-Text Alignment
by: Kuang, Jidong, et al.
Published: (2024)
by: Kuang, Jidong, et al.
Published: (2024)
Zero-Shot Hashing Based on Reconstruction With Part Alignment
by: Jiang, Yan, et al.
Published: (2025)
by: Jiang, Yan, et al.
Published: (2025)
ZeroDiff: Solidified Visual-Semantic Correlation in Zero-Shot Learning
by: Ye, Zihan, et al.
Published: (2024)
by: Ye, Zihan, et al.
Published: (2024)
Semantically Guided Dynamic Visual Prototype Refinement for Compositional Zero-Shot Learning
by: Peng, Zhong, et al.
Published: (2025)
by: Peng, Zhong, et al.
Published: (2025)
RANGER: A Monocular Zero-Shot Semantic Navigation Framework through Visual Contextual Adaptation
by: Yu, Ming-Ming, et al.
Published: (2025)
by: Yu, Ming-Ming, et al.
Published: (2025)
Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation
by: Li, Yunheng, et al.
Published: (2024)
by: Li, Yunheng, et al.
Published: (2024)
ZeroDDI: A Zero-Shot Drug-Drug Interaction Event Prediction Method with Semantic Enhanced Learning and Dual-Modal Uniform Alignment
by: Wang, Ziyan, et al.
Published: (2024)
by: Wang, Ziyan, et al.
Published: (2024)
Prompt-based Visual Alignment for Zero-shot Policy Transfer
by: Gao, Haihan, et al.
Published: (2024)
by: Gao, Haihan, et al.
Published: (2024)
SVIP: Semantically Contextualized Visual Patches for Zero-Shot Learning
by: Chen, Zhi, et al.
Published: (2025)
by: Chen, Zhi, et al.
Published: (2025)
Visual-Semantic Graph Matching Net for Zero-Shot Learning
by: Duan, Bowen, et al.
Published: (2024)
by: Duan, Bowen, et al.
Published: (2024)
Similar Items
-
Missing Target-Relevant Information Prediction with World Model for Accurate Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2025) -
Denoise-I2W: Mapping Images to Denoising Words for Accurate Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2024) -
MADS: Multi-Attribute Document Supervision for Zero-Shot Image Classification
by: Qu, Xiangyan, et al.
Published: (2025) -
T2VParser: Adaptive Decomposition Tokens for Partial Alignment in Text to Video Retrieval
by: Li, Yili, et al.
Published: (2025) -
ProAPO: Progressively Automatic Prompt Optimization for Visual Classification
by: Qu, Xiangyan, et al.
Published: (2025)