Craft: Cross-modal Aligned Features Improve Robustness of Prompt Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Jingchen, Sharma, Rohan, Lokhande, Vishnu Suresh, Chen, Changyou |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Latent Diffusion Unlearning: Protecting Against Unauthorized Personalization Through Trajectory Shifted Perturbations
by: Devulapally, Naresh Kumar, et al.
Published: (2025)
by: Devulapally, Naresh Kumar, et al.
Published: (2025)
Enhancing Diffusion Posterior Sampling for Inverse Problems by Integrating Crafted Measurements
by: Zhou, Shijie, et al.
Published: (2024)
by: Zhou, Shijie, et al.
Published: (2024)
ParallelEdits: Efficient Multi-object Image Editing
by: Huang, Mingzhen, et al.
Published: (2024)
by: Huang, Mingzhen, et al.
Published: (2024)
Pooling Image Datasets With Multiple Covariate Shift and Imbalance
by: Chytas, Sotirios Panagiotis, et al.
Published: (2024)
by: Chytas, Sotirios Panagiotis, et al.
Published: (2024)
Uncertainty-Aware Knowledge Distillation for Multimodal Large Language Models
by: Sun, Jingchen, et al.
Published: (2026)
by: Sun, Jingchen, et al.
Published: (2026)
SCING:Towards More Efficient and Robust Person Re-Identification through Selective Cross-modal Prompt Tuning
by: Xie, Yunfei, et al.
Published: (2025)
by: Xie, Yunfei, et al.
Published: (2025)
ManifoldGD: Training-Free Hierarchical Manifold Guidance for Diffusion-Based Dataset Distillation
by: Roy, Ayush, et al.
Published: (2026)
by: Roy, Ayush, et al.
Published: (2026)
Score-Control for Hallucination Reduction in Diffusion Models
by: Bhosale, Mahesh, et al.
Published: (2026)
by: Bhosale, Mahesh, et al.
Published: (2026)
Is Exchangeability better than I.I.D to handle Data Distribution Shifts while Pooling Data for Data-scarce Medical image segmentation?
by: Roy, Ayush, et al.
Published: (2025)
by: Roy, Ayush, et al.
Published: (2025)
Progressive Multi-modal Conditional Prompt Tuning
by: Qiu, Xiaoyu, et al.
Published: (2024)
by: Qiu, Xiaoyu, et al.
Published: (2024)
Forget Less by Learning Together through Concept Consolidation
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
Cross-modal Prompting for Balanced Incomplete Multi-modal Emotion Recognition
by: He, Wen-Jue, et al.
Published: (2025)
by: He, Wen-Jue, et al.
Published: (2025)
Prompt2Craft: Generating Functional Craft Assemblies with LLMs
by: Isume, Vitor Hideyo, et al.
Published: (2025)
by: Isume, Vitor Hideyo, et al.
Published: (2025)
Forget Less by Learning from Parents Through Hierarchical Relationships
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
by: Kaushik, Arjun Ramesh, et al.
Published: (2026)
Model-Agnostic Gender Bias Control for Text-to-Image Generation via Sparse Autoencoder
by: Wu, Chao, et al.
Published: (2025)
by: Wu, Chao, et al.
Published: (2025)
EasyCraft: A Robust and Efficient Framework for Automatic Avatar Crafting
by: Wang, Suzhen, et al.
Published: (2025)
by: Wang, Suzhen, et al.
Published: (2025)
IRisPath: Enhancing Costmap for Off-Road Navigation with Robust IR-RGB Fusion for Improved Day and Night Traversability
by: Sharma, Saksham, et al.
Published: (2024)
by: Sharma, Saksham, et al.
Published: (2024)
iVPT: Improving Task-relevant Information Sharing in Visual Prompt Tuning by Cross-layer Dynamic Connection
by: Zhou, Nan, et al.
Published: (2024)
by: Zhou, Nan, et al.
Published: (2024)
Improving Zero-Shot ObjectNav with Generative Communication
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
Hierarchical Cross-modal Prompt Learning for Vision-Language Models
by: Zheng, Hao, et al.
Published: (2025)
by: Zheng, Hao, et al.
Published: (2025)
Patch-Prompt Aligned Bayesian Prompt Tuning for Vision-Language Models
by: Liu, Xinyang, et al.
Published: (2023)
by: Liu, Xinyang, et al.
Published: (2023)
XoFTR: Cross-modal Feature Matching Transformer
by: Tuzcuoğlu, Önder, et al.
Published: (2024)
by: Tuzcuoğlu, Önder, et al.
Published: (2024)
A High-Quality Text-Rich Image Instruction Tuning Dataset via Hybrid Instruction Generation
by: Zhou, Shijie, et al.
Published: (2024)
by: Zhou, Shijie, et al.
Published: (2024)
Multi-modal Semantic Understanding with Contrastive Cross-modal Feature Alignment
by: Zhang, Ming, et al.
Published: (2024)
by: Zhang, Ming, et al.
Published: (2024)
CIBR: Cross-modal Information Bottleneck Regularization for Robust CLIP Generalization
by: Ji, Yingrui, et al.
Published: (2025)
by: Ji, Yingrui, et al.
Published: (2025)
Learning Robust Anymodal Segmentor with Unimodal and Cross-modal Distillation
by: Zheng, Xu, et al.
Published: (2024)
by: Zheng, Xu, et al.
Published: (2024)
Aligned Novel View Image and Geometry Synthesis via Cross-modal Attention Instillation
by: Kwak, Min-Seop, et al.
Published: (2025)
by: Kwak, Min-Seop, et al.
Published: (2025)
Your Text Encoder Can Be An Object-Level Watermarking Controller
by: Devulapally, Naresh Kumar, et al.
Published: (2025)
by: Devulapally, Naresh Kumar, et al.
Published: (2025)
Text-guided Feature Disentanglement for Cross-modal Gait Recognition
by: Lu, Zhiyang, et al.
Published: (2026)
by: Lu, Zhiyang, et al.
Published: (2026)
DMPT: Decoupled Modality-aware Prompt Tuning for Multi-modal Object Re-identification
by: Lin, Minghui, et al.
Published: (2025)
by: Lin, Minghui, et al.
Published: (2025)
LLM-enhanced Action-aware Multi-modal Prompt Tuning for Image-Text Matching
by: Tian, Mengxiao, et al.
Published: (2025)
by: Tian, Mengxiao, et al.
Published: (2025)
Computed Tomography Visual Question Answering with Cross-modal Feature Graphing
by: Tian, Yuanhe, et al.
Published: (2025)
by: Tian, Yuanhe, et al.
Published: (2025)
Align3D-AD: Cross-Modal Feature Alignment and Dual-Prompt Learning for Zero-shot 3D Anomaly Detection
by: Bai, Letian, et al.
Published: (2026)
by: Bai, Letian, et al.
Published: (2026)
AlignMamba: Enhancing Multimodal Mamba with Local and Global Cross-modal Alignment
by: Li, Yan, et al.
Published: (2024)
by: Li, Yan, et al.
Published: (2024)
Cross-modal Offset-guided Dynamic Alignment and Fusion for Weakly Aligned UAV Object Detection
by: Zongzhen, Liu, et al.
Published: (2025)
by: Zongzhen, Liu, et al.
Published: (2025)
CrossTracker: Robust Multi-modal 3D Multi-Object Tracking via Cross Correction
by: Gu, Lipeng, et al.
Published: (2024)
by: Gu, Lipeng, et al.
Published: (2024)
TAPT: Test-Time Adversarial Prompt Tuning for Robust Inference in Vision-Language Models
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
From Points to Clouds: Learning Robust Semantic Distributions for Multi-modal Prompts
by: Li, Weiran, et al.
Published: (2025)
by: Li, Weiran, et al.
Published: (2025)
Decouple before Align: Visual Disentanglement Enhances Prompt Tuning
by: Zhang, Fei, et al.
Published: (2025)
by: Zhang, Fei, et al.
Published: (2025)
Exploring Interpretability for Visual Prompt Tuning with Cross-layer Concepts
by: Wang, Yubin, et al.
Published: (2025)
by: Wang, Yubin, et al.
Published: (2025)
Similar Items
-
Latent Diffusion Unlearning: Protecting Against Unauthorized Personalization Through Trajectory Shifted Perturbations
by: Devulapally, Naresh Kumar, et al.
Published: (2025) -
Enhancing Diffusion Posterior Sampling for Inverse Problems by Integrating Crafted Measurements
by: Zhou, Shijie, et al.
Published: (2024) -
ParallelEdits: Efficient Multi-object Image Editing
by: Huang, Mingzhen, et al.
Published: (2024) -
Pooling Image Datasets With Multiple Covariate Shift and Imbalance
by: Chytas, Sotirios Panagiotis, et al.
Published: (2024) -
Uncertainty-Aware Knowledge Distillation for Multimodal Large Language Models
by: Sun, Jingchen, et al.
Published: (2026)