AdaViPro: Region-based Adaptive Visual Prompt for Large-Scale Models Adapting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Mengyu, Tian, Ye, Zhang, Lanshan, Liang, Xiao, Ran, Xuming, Wang, Wendong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
View while Moving: Efficient Video Recognition in Long-untrimmed Videos
von: Tian, Ye, et al.
Veröffentlicht: (2023)
von: Tian, Ye, et al.
Veröffentlicht: (2023)
ViT-AdaLA: Adapting Vision Transformers with Linear Attention
von: Li, Yifan, et al.
Veröffentlicht: (2026)
von: Li, Yifan, et al.
Veröffentlicht: (2026)
Large Vision-Language Models Get Lost in Attention
von: Xi, Gongli, et al.
Veröffentlicht: (2026)
von: Xi, Gongli, et al.
Veröffentlicht: (2026)
AdaFusion: Prompt-Guided Inference with Adaptive Fusion of Pathology Foundation Models
von: Xiao, Yuxiang, et al.
Veröffentlicht: (2025)
von: Xiao, Yuxiang, et al.
Veröffentlicht: (2025)
AdaCoder: Adaptive Prompt Compression for Programmatic Visual Question Answering
von: Ukai, Mahiro, et al.
Veröffentlicht: (2024)
von: Ukai, Mahiro, et al.
Veröffentlicht: (2024)
Causality-guided Prompt Learning for Vision-language Models via Visual Granulation
von: Gao, Mengyu, et al.
Veröffentlicht: (2025)
von: Gao, Mengyu, et al.
Veröffentlicht: (2025)
AdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detection
von: Cao, Yunkang, et al.
Veröffentlicht: (2024)
von: Cao, Yunkang, et al.
Veröffentlicht: (2024)
LaViC: Adapting Large Vision-Language Models to Visually-Aware Conversational Recommendation
von: Jeon, Hyunsik, et al.
Veröffentlicht: (2025)
von: Jeon, Hyunsik, et al.
Veröffentlicht: (2025)
U-REPA: Aligning Diffusion U-Nets to ViTs
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
AdaSCALE: Adaptive Scaling for OOD Detection
von: Regmi, Sudarshan
Veröffentlicht: (2025)
von: Regmi, Sudarshan
Veröffentlicht: (2025)
Global Position Aware Group Choreography using Large Language Model
von: Pang, Haozhou, et al.
Veröffentlicht: (2025)
von: Pang, Haozhou, et al.
Veröffentlicht: (2025)
AdaViP: Aligning Multi-modal LLMs via Adaptive Vision-enhanced Preference Optimization
von: Lu, Jinda, et al.
Veröffentlicht: (2025)
von: Lu, Jinda, et al.
Veröffentlicht: (2025)
PrAda: Few-Shot Visual Adaptation for Text-Prompted Segmentation
von: Rosi, Gabriele, et al.
Veröffentlicht: (2026)
von: Rosi, Gabriele, et al.
Veröffentlicht: (2026)
LoopViT: Scaling Visual ARC with Looped Transformers
von: Shu, Wen-Jie, et al.
Veröffentlicht: (2026)
von: Shu, Wen-Jie, et al.
Veröffentlicht: (2026)
ChAda-ViT : Channel Adaptive Attention for Joint Representation Learning of Heterogeneous Microscopy Images
von: Bourriez, Nicolas, et al.
Veröffentlicht: (2023)
von: Bourriez, Nicolas, et al.
Veröffentlicht: (2023)
AdaptVision: Efficient Vision-Language Models via Adaptive Visual Acquisition
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
von: Lin, Zichuan, et al.
Veröffentlicht: (2025)
AdaViewPlanner: Adapting Video Diffusion Models for Viewpoint Planning in 4D Scenes
von: Li, Yu, et al.
Veröffentlicht: (2025)
von: Li, Yu, et al.
Veröffentlicht: (2025)
Adapting to Distribution Shift by Visual Domain Prompt Generation
von: Chi, Zhixiang, et al.
Veröffentlicht: (2024)
von: Chi, Zhixiang, et al.
Veröffentlicht: (2024)
Janus-Pro-R1: Advancing Collaborative Visual Comprehension and Generation via Reinforcement Learning
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
AdaQual-Diff: Diffusion-Based Image Restoration via Adaptive Quality Prompting
von: Su, Xin, et al.
Veröffentlicht: (2025)
von: Su, Xin, et al.
Veröffentlicht: (2025)
GRASP: Guided Region-Aware Sparse Prompting for Adapting MLLMs to Remote Sensing
von: Sun, Qigan, et al.
Veröffentlicht: (2026)
von: Sun, Qigan, et al.
Veröffentlicht: (2026)
AdaGlimpse: Active Visual Exploration with Arbitrary Glimpse Position and Scale
von: Pardyl, Adam, et al.
Veröffentlicht: (2024)
von: Pardyl, Adam, et al.
Veröffentlicht: (2024)
ProSAM: Enhancing the Robustness of SAM-based Visual Reference Segmentation with Probabilistic Prompts
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2025)
AVM: Towards Structure-Preserving Neural Response Modeling in the Visual Cortex Across Stimuli and Individuals
von: Xu, Qi, et al.
Veröffentlicht: (2025)
von: Xu, Qi, et al.
Veröffentlicht: (2025)
AdaSVD: Adaptive Singular Value Decomposition for Large Language Models
von: Li, Zhiteng, et al.
Veröffentlicht: (2025)
von: Li, Zhiteng, et al.
Veröffentlicht: (2025)
AdaDiff: Accelerating Diffusion Models through Step-Wise Adaptive Computation
von: Tang, Shengkun, et al.
Veröffentlicht: (2023)
von: Tang, Shengkun, et al.
Veröffentlicht: (2023)
ViKey: Enhancing Temporal Understanding in Videos via Visual Prompting
von: Lee, Yeonkyung, et al.
Veröffentlicht: (2026)
von: Lee, Yeonkyung, et al.
Veröffentlicht: (2026)
AdaptViG: Adaptive Vision GNN with Exponential Decay Gating
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
von: Munir, Mustafa, et al.
Veröffentlicht: (2025)
ProAPO: Progressively Automatic Prompt Optimization for Visual Classification
von: Qu, Xiangyan, et al.
Veröffentlicht: (2025)
von: Qu, Xiangyan, et al.
Veröffentlicht: (2025)
FairSeg: A Large-Scale Medical Image Segmentation Dataset for Fairness Learning Using Segment Anything Model with Fair Error-Bound Scaling
von: Tian, Yu, et al.
Veröffentlicht: (2023)
von: Tian, Yu, et al.
Veröffentlicht: (2023)
ViPO: Visual Preference Optimization at Scale
von: Li, Ming, et al.
Veröffentlicht: (2026)
von: Li, Ming, et al.
Veröffentlicht: (2026)
ProEdit: Inversion-based Editing From Prompts Done Right
von: Ouyang, Zhi, et al.
Veröffentlicht: (2025)
von: Ouyang, Zhi, et al.
Veröffentlicht: (2025)
AdaContour: Adaptive Contour Descriptor with Hierarchical Representation
von: Ding, Tianyu, et al.
Veröffentlicht: (2024)
von: Ding, Tianyu, et al.
Veröffentlicht: (2024)
Prompt-Aware Adapter: Towards Learning Adaptive Visual Tokens for Multimodal Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
AdaViT: Adaptive Vision Transformer for Flexible Pretrain and Finetune with Variable 3D Medical Image Modalities
von: Das, Badhan Kumar, et al.
Veröffentlicht: (2025)
von: Das, Badhan Kumar, et al.
Veröffentlicht: (2025)
AdaFV: Rethinking of Visual-Language alignment for VLM acceleration
von: Han, Jiayi, et al.
Veröffentlicht: (2025)
von: Han, Jiayi, et al.
Veröffentlicht: (2025)
AdaDeDup: Adaptive Hybrid Data Pruning for Efficient Large-Scale Object Detection Training
von: Kang, Feiyang, et al.
Veröffentlicht: (2025)
von: Kang, Feiyang, et al.
Veröffentlicht: (2025)
ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts
von: Cai, Mu, et al.
Veröffentlicht: (2023)
von: Cai, Mu, et al.
Veröffentlicht: (2023)
AdaRadar: Rate Adaptive Spectral Compression for Radar-based Perception
von: Park, Jinho, et al.
Veröffentlicht: (2026)
von: Park, Jinho, et al.
Veröffentlicht: (2026)
AdaMerging: Adaptive Model Merging for Multi-Task Learning
von: Yang, Enneng, et al.
Veröffentlicht: (2023)
von: Yang, Enneng, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
View while Moving: Efficient Video Recognition in Long-untrimmed Videos
von: Tian, Ye, et al.
Veröffentlicht: (2023) -
ViT-AdaLA: Adapting Vision Transformers with Linear Attention
von: Li, Yifan, et al.
Veröffentlicht: (2026) -
Large Vision-Language Models Get Lost in Attention
von: Xi, Gongli, et al.
Veröffentlicht: (2026) -
AdaFusion: Prompt-Guided Inference with Adaptive Fusion of Pathology Foundation Models
von: Xiao, Yuxiang, et al.
Veröffentlicht: (2025) -
AdaCoder: Adaptive Prompt Compression for Programmatic Visual Question Answering
von: Ukai, Mahiro, et al.
Veröffentlicht: (2024)