PromptHub: Enhancing Multi-Prompt Visual In-Context Learning with Locality-Aware Fusion, Concentration and Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Tianci, Wang, Jinpeng, Qin, Shiyu, Lian, Niu, Feng, Yan, Chen, Bin, Yuan, Chun, Xia, Shu-Tao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Love Me, Love My Label: Rethinking the Role of Labels in Prompt Retrieval for Visual In-Context Learning
by: Luo, Tianci, et al.
Published: (2026)
by: Luo, Tianci, et al.
Published: (2026)
Embracing Collaboration Over Competition: Condensing Multiple Prompts for Visual In-Context Learning
by: Wang, Jinpeng, et al.
Published: (2025)
by: Wang, Jinpeng, et al.
Published: (2025)
MambaVC: Learned Visual Compression with Selective State Spaces
by: Qin, Shiyu, et al.
Published: (2024)
by: Qin, Shiyu, et al.
Published: (2024)
Textual and Visual Prompt Fusion for Image Editing via Step-Wise Alignment
by: Feng, Zhanbo, et al.
Published: (2023)
by: Feng, Zhanbo, et al.
Published: (2023)
SOUPLE: Enhancing Audio-Visual Localization and Segmentation with Learnable Prompt Contexts
by: Nguyen, Khanh Binh, et al.
Published: (2026)
by: Nguyen, Khanh Binh, et al.
Published: (2026)
AutoSSVH: Exploring Automated Frame Sampling for Efficient Self-Supervised Video Hashing
by: Lian, Niu, et al.
Published: (2025)
by: Lian, Niu, et al.
Published: (2025)
Visual Prompting in LLMs for Enhancing Emotion Recognition
by: Zhang, Qixuan, et al.
Published: (2024)
by: Zhang, Qixuan, et al.
Published: (2024)
Imagine Before Concentration: Diffusion-Guided Registers Enhance Partially Relevant Video Retrieval
by: Li, Jun, et al.
Published: (2026)
by: Li, Jun, et al.
Published: (2026)
Memory-SAM: Human-Prompt-Free Tongue Segmentation via Retrieval-to-Prompt
by: Chae, Joongwon, et al.
Published: (2025)
by: Chae, Joongwon, et al.
Published: (2025)
VisualPrompter: Semantic-Aware Prompt Optimization with Visual Feedback for Text-to-Image Synthesis
by: Wu, Shiyu, et al.
Published: (2025)
by: Wu, Shiyu, et al.
Published: (2025)
Visually Prompted Benchmarks Are Surprisingly Fragile
by: Feng, Haiwen, et al.
Published: (2025)
by: Feng, Haiwen, et al.
Published: (2025)
Anatomy-Aware Text-Visual Fusion with Dual-Perspective Prompts for Fine-Grained Lumbar Spine Segmentation
by: Lian, Sheng, et al.
Published: (2025)
by: Lian, Sheng, et al.
Published: (2025)
Efficient Self-Supervised Video Hashing with Selective State Spaces
by: Wang, Jinpeng, et al.
Published: (2024)
by: Wang, Jinpeng, et al.
Published: (2024)
HLFormer: Enhancing Partially Relevant Video Retrieval with Hyperbolic Learning
by: Li, Jun, et al.
Published: (2025)
by: Li, Jun, et al.
Published: (2025)
PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification
by: Luo, Qiuming, et al.
Published: (2026)
by: Luo, Qiuming, et al.
Published: (2026)
BadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIP
by: Bai, Jiawang, et al.
Published: (2023)
by: Bai, Jiawang, et al.
Published: (2023)
KEPIL: Knowledge-Enhanced Prompt-Image Learning for Prompt-Robust Disease Detection
by: Luo, Haozhe, et al.
Published: (2026)
by: Luo, Haozhe, et al.
Published: (2026)
ROCKET-1: Mastering Open-World Interaction with Visual-Temporal Context Prompting
by: Cai, Shaofei, et al.
Published: (2024)
by: Cai, Shaofei, et al.
Published: (2024)
Enhancing Visual Forced Alignment with Local Context-Aware Feature Extraction and Multi-Task Learning
by: He, Yi, et al.
Published: (2025)
by: He, Yi, et al.
Published: (2025)
Towards Global Optimal Visual In-Context Learning Prompt Selection
by: Xu, Chengming, et al.
Published: (2024)
by: Xu, Chengming, et al.
Published: (2024)
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
by: Wang, Zhifeng, et al.
Published: (2025)
by: Wang, Zhifeng, et al.
Published: (2025)
IF-Bench: Benchmarking and Enhancing MLLMs for Infrared Images with Generative Visual Prompting
by: Zhang, Tao, et al.
Published: (2025)
by: Zhang, Tao, et al.
Published: (2025)
E-InMeMo: Enhanced Prompting for Visual In-Context Learning
by: Zhang, Jiahao, et al.
Published: (2025)
by: Zhang, Jiahao, et al.
Published: (2025)
From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents
by: Lian, Niu, et al.
Published: (2026)
by: Lian, Niu, et al.
Published: (2026)
3D-LMVIC: Learning-based Multi-View Image Coding with 3D Gaussian Geometric Priors
by: Huang, Yujun, et al.
Published: (2024)
by: Huang, Yujun, et al.
Published: (2024)
PromptDx: Differentiable Prompt Tuning for Multimodal In-Context Alzheimer's Diagnosis
by: Zhong, Lujia, et al.
Published: (2026)
by: Zhong, Lujia, et al.
Published: (2026)
Visual Prompt Selection for In-Context Learning Segmentation
by: Suo, Wei, et al.
Published: (2024)
by: Suo, Wei, et al.
Published: (2024)
GMMFormer v2: An Uncertainty-aware Framework for Partially Relevant Video Retrieval
by: Wang, Yuting, et al.
Published: (2024)
by: Wang, Yuting, et al.
Published: (2024)
Attention to the Burstiness in Visual Prompt Tuning!
by: Wang, Yuzhu, et al.
Published: (2025)
by: Wang, Yuzhu, et al.
Published: (2025)
Integrated Structural Prompt Learning for Vision-Language Models
by: Wang, Jiahui, et al.
Published: (2025)
by: Wang, Jiahui, et al.
Published: (2025)
Text-guided Visual Prompt DINO for Generic Segmentation
by: Guan, Yuchen, et al.
Published: (2025)
by: Guan, Yuchen, et al.
Published: (2025)
ARGUS: Defending LLM Agents Against Context-Aware Prompt Injection
by: Weng, Shihao, et al.
Published: (2026)
by: Weng, Shihao, et al.
Published: (2026)
Multimodal Prompt Alignment for Facial Expression Recognition
by: Ma, Fuyan, et al.
Published: (2025)
by: Ma, Fuyan, et al.
Published: (2025)
Kernel-Aware Graph Prompt Learning for Few-Shot Anomaly Detection
by: Tao, Fenfang, et al.
Published: (2024)
by: Tao, Fenfang, et al.
Published: (2024)
DiffusionUavLoc: Visually Prompted Diffusion for Cross-View UAV Localization
by: Liu, Tao, et al.
Published: (2025)
by: Liu, Tao, et al.
Published: (2025)
Explicit Uncertainty Modeling for Active CLIP Adaptation with Dual Prompt Tuning
by: Wang, Qian-Wei, et al.
Published: (2026)
by: Wang, Qian-Wei, et al.
Published: (2026)
PromptMAD: Cross-Modal Prompting for Multi-Class Visual Anomaly Localization
by: McCain, Duncan, et al.
Published: (2026)
by: McCain, Duncan, et al.
Published: (2026)
Tuning Vision-Language Models with Candidate Labels by Prompt Alignment
by: Zhang, Zhifang, et al.
Published: (2024)
by: Zhang, Zhifang, et al.
Published: (2024)
AutoV: Loss-Oriented Ranking for Visual Prompt Retrieval in LVLMs
by: Zhang, Yuan, et al.
Published: (2025)
by: Zhang, Yuan, et al.
Published: (2025)
Context-Aware Pragmatic Metacognitive Prompting for Sarcasm Detection
by: Iskandardinata, Michael, et al.
Published: (2025)
by: Iskandardinata, Michael, et al.
Published: (2025)
Similar Items
-
Love Me, Love My Label: Rethinking the Role of Labels in Prompt Retrieval for Visual In-Context Learning
by: Luo, Tianci, et al.
Published: (2026) -
Embracing Collaboration Over Competition: Condensing Multiple Prompts for Visual In-Context Learning
by: Wang, Jinpeng, et al.
Published: (2025) -
MambaVC: Learned Visual Compression with Selective State Spaces
by: Qin, Shiyu, et al.
Published: (2024) -
Textual and Visual Prompt Fusion for Image Editing via Step-Wise Alignment
by: Feng, Zhanbo, et al.
Published: (2023) -
SOUPLE: Enhancing Audio-Visual Localization and Segmentation with Learnable Prompt Contexts
by: Nguyen, Khanh Binh, et al.
Published: (2026)