Visual Prompt Selection for In-Context Learning Segmentation
Fuente:
arXiv
Salvato in:
| Autori principali: | Suo, Wei, Lai, Lanqing, Sun, Mengyang, Zhang, Hanwang, Wang, Peng, Zhang, Yanning |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Octopus: Alleviating Hallucination via Dynamic Contrastive Decoding
di: Suo, Wei, et al.
Pubblicazione: (2025)
di: Suo, Wei, et al.
Pubblicazione: (2025)
Visual Prompt Tuning in Null Space for Continual Learning
di: Lu, Yue, et al.
Pubblicazione: (2024)
di: Lu, Yue, et al.
Pubblicazione: (2024)
Hallucination-aware intermediate representation edit in large vision-language models
di: Suo, Wei, et al.
Pubblicazione: (2026)
di: Suo, Wei, et al.
Pubblicazione: (2026)
Exploring Task-Level Optimal Prompts for Visual In-Context Learning
di: Zhu, Yan, et al.
Pubblicazione: (2025)
di: Zhu, Yan, et al.
Pubblicazione: (2025)
One-shot In-context Part Segmentation
di: Dai, Zhenqi, et al.
Pubblicazione: (2025)
di: Dai, Zhenqi, et al.
Pubblicazione: (2025)
Weakly Supervised Video Anomaly Detection and Localization with Spatio-Temporal Prompts
di: Wu, Peng, et al.
Pubblicazione: (2024)
di: Wu, Peng, et al.
Pubblicazione: (2024)
Selective Visual Prompting in Vision Mamba
di: Yao, Yifeng, et al.
Pubblicazione: (2024)
di: Yao, Yifeng, et al.
Pubblicazione: (2024)
Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models
di: Suo, Wei, et al.
Pubblicazione: (2024)
di: Suo, Wei, et al.
Pubblicazione: (2024)
Deep Learning for Video Anomaly Detection: A Review
di: Wu, Peng, et al.
Pubblicazione: (2024)
di: Wu, Peng, et al.
Pubblicazione: (2024)
Spatio-Temporal Context Prompting for Zero-Shot Action Detection
di: Huang, Wei-Jhe, et al.
Pubblicazione: (2024)
di: Huang, Wei-Jhe, et al.
Pubblicazione: (2024)
Semi-Supervised Semantic Segmentation Based on Pseudo-Labels: A Survey
di: Ran, Lingyan, et al.
Pubblicazione: (2024)
di: Ran, Lingyan, et al.
Pubblicazione: (2024)
Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning
di: Li, Zejun, et al.
Pubblicazione: (2025)
di: Li, Zejun, et al.
Pubblicazione: (2025)
Dual-Modal Prompting for Sketch-Based Image Retrieval
di: Gao, Liying, et al.
Pubblicazione: (2024)
di: Gao, Liying, et al.
Pubblicazione: (2024)
P4Q: Learning to Prompt for Quantization in Visual-language Models
di: Sun, Huixin, et al.
Pubblicazione: (2024)
di: Sun, Huixin, et al.
Pubblicazione: (2024)
Visual Position Prompt for MLLM based Visual Grounding
di: Tang, Wei, et al.
Pubblicazione: (2025)
di: Tang, Wei, et al.
Pubblicazione: (2025)
Visual Instance-aware Prompt Tuning
di: Xiao, Xi, et al.
Pubblicazione: (2025)
di: Xiao, Xi, et al.
Pubblicazione: (2025)
You Only Need One Color Space: An Efficient Network for Low-light Image Enhancement
di: Yan, Qingsen, et al.
Pubblicazione: (2024)
di: Yan, Qingsen, et al.
Pubblicazione: (2024)
ContextGS: Compact 3D Gaussian Splatting with Anchor Level Context Model
di: Wang, Yufei, et al.
Pubblicazione: (2024)
di: Wang, Yufei, et al.
Pubblicazione: (2024)
SAMAug: Point Prompt Augmentation for Segment Anything Model
di: Dai, Haixing, et al.
Pubblicazione: (2023)
di: Dai, Haixing, et al.
Pubblicazione: (2023)
How Do Optical Flow and Textual Prompts Collaborate to Assist in Audio-Visual Semantic Segmentation?
di: Lee, Yujian, et al.
Pubblicazione: (2026)
di: Lee, Yujian, et al.
Pubblicazione: (2026)
From Trial to Triumph: Advancing Long Video Understanding via Visual Context Sample Scaling and Self-reward Alignment
di: Suo, Yucheng, et al.
Pubblicazione: (2025)
di: Suo, Yucheng, et al.
Pubblicazione: (2025)
C3L: Content Correlated Vision-Language Instruction Tuning Data Generation via Contrastive Learning
di: Ma, Ji, et al.
Pubblicazione: (2024)
di: Ma, Ji, et al.
Pubblicazione: (2024)
Merging Context Clustering with Visual State Space Models for Medical Image Segmentation
di: Zhu, Yun, et al.
Pubblicazione: (2025)
di: Zhu, Yun, et al.
Pubblicazione: (2025)
Towards Video Anomaly Retrieval from Video Anomaly Detection: New Benchmarks and Model
di: Wu, Peng, et al.
Pubblicazione: (2023)
di: Wu, Peng, et al.
Pubblicazione: (2023)
Generalizable Object Re-Identification via Visual In-Context Prompting
di: Huang, Zhizhong, et al.
Pubblicazione: (2025)
di: Huang, Zhizhong, et al.
Pubblicazione: (2025)
A Plug-and-Play Method for Rare Human-Object Interactions Detection by Bridging Domain Gap
di: Zhang, Lijun, et al.
Pubblicazione: (2024)
di: Zhang, Lijun, et al.
Pubblicazione: (2024)
ROCKET-1: Mastering Open-World Interaction with Visual-Temporal Context Prompting
di: Cai, Shaofei, et al.
Pubblicazione: (2024)
di: Cai, Shaofei, et al.
Pubblicazione: (2024)
PromptDx: Differentiable Prompt Tuning for Multimodal In-Context Alzheimer's Diagnosis
di: Zhong, Lujia, et al.
Pubblicazione: (2026)
di: Zhong, Lujia, et al.
Pubblicazione: (2026)
Ref-AVS: Refer and Segment Objects in Audio-Visual Scenes
di: Wang, Yaoting, et al.
Pubblicazione: (2024)
di: Wang, Yaoting, et al.
Pubblicazione: (2024)
Learning Local and Global Temporal Contexts for Video Semantic Segmentation
di: Sun, Guolei, et al.
Pubblicazione: (2022)
di: Sun, Guolei, et al.
Pubblicazione: (2022)
Refer to Any Segmentation Mask Group With Vision-Language Prompts
di: Cao, Shengcao, et al.
Pubblicazione: (2025)
di: Cao, Shengcao, et al.
Pubblicazione: (2025)
Training-Free Text-Guided Image Editing with Visual Autoregressive Model
di: Wang, Yufei, et al.
Pubblicazione: (2025)
di: Wang, Yufei, et al.
Pubblicazione: (2025)
Dysen-VDM: Empowering Dynamics-aware Text-to-Video Diffusion with LLMs
di: Fei, Hao, et al.
Pubblicazione: (2023)
di: Fei, Hao, et al.
Pubblicazione: (2023)
GeoDiffusion: Text-Prompted Geometric Control for Object Detection Data Generation
di: Chen, Kai, et al.
Pubblicazione: (2023)
di: Chen, Kai, et al.
Pubblicazione: (2023)
Decouple before Align: Visual Disentanglement Enhances Prompt Tuning
di: Zhang, Fei, et al.
Pubblicazione: (2025)
di: Zhang, Fei, et al.
Pubblicazione: (2025)
Curriculum Prompting Foundation Models for Medical Image Segmentation
di: Zheng, Xiuqi, et al.
Pubblicazione: (2024)
di: Zheng, Xiuqi, et al.
Pubblicazione: (2024)
Seeking Consensus: Geometric-Semantic On-the-Fly Recalibration for Open-Vocabulary Remote Sensing Semantic Segmentation
di: Wang, Guanchun, et al.
Pubblicazione: (2026)
di: Wang, Guanchun, et al.
Pubblicazione: (2026)
Language-Driven Visual Consensus for Zero-Shot Semantic Segmentation
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers
di: Ma, Ji, et al.
Pubblicazione: (2025)
di: Ma, Ji, et al.
Pubblicazione: (2025)
Understanding and Mitigating Hallucinations in Multimodal Chain-of-Thought Models
di: Ma, Ji, et al.
Pubblicazione: (2026)
di: Ma, Ji, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Octopus: Alleviating Hallucination via Dynamic Contrastive Decoding
di: Suo, Wei, et al.
Pubblicazione: (2025) -
Visual Prompt Tuning in Null Space for Continual Learning
di: Lu, Yue, et al.
Pubblicazione: (2024) -
Hallucination-aware intermediate representation edit in large vision-language models
di: Suo, Wei, et al.
Pubblicazione: (2026) -
Exploring Task-Level Optimal Prompts for Visual In-Context Learning
di: Zhu, Yan, et al.
Pubblicazione: (2025) -
One-shot In-context Part Segmentation
di: Dai, Zhenqi, et al.
Pubblicazione: (2025)