Enhancing Visual Prompting through Expanded Transformation Space and Overfitting Mitigation
Fuente:
arXiv
Saved in:
| Main Author: | Enomoto, Shohei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pseudo Multi-Source Domain Generalization: Bridging the Gap Between Single and Multi-Source Domain Generalization
by: Enomoto, Shohei
Published: (2025)
by: Enomoto, Shohei
Published: (2025)
EntProp: High Entropy Propagation for Improving Accuracy and Robustness
by: Enomoto, Shohei
Published: (2024)
by: Enomoto, Shohei
Published: (2024)
MultiModal Fine-tuning with Synthetic Captions
by: Enomoto, Shohei, et al.
Published: (2026)
by: Enomoto, Shohei, et al.
Published: (2026)
Dynamic Test-Time Augmentation via Differentiable Functions
by: Enomoto, Shohei, et al.
Published: (2022)
by: Enomoto, Shohei, et al.
Published: (2022)
Test-time Similarity Modification for Person Re-identification toward Temporal Distribution Shift
by: Adachi, Kazuki, et al.
Published: (2024)
by: Adachi, Kazuki, et al.
Published: (2024)
Incoherent Deformation, Not Capacity: Diagnosing and Mitigating Overfitting in Dynamic Gaussian Splatting
by: Droby, Ahmad
Published: (2026)
by: Droby, Ahmad
Published: (2026)
Improving Image Coding for Machines through Optimizing Encoder via Auxiliary Loss
by: Iino, Kei, et al.
Published: (2024)
by: Iino, Kei, et al.
Published: (2024)
Multi-dimensional Visual Prompt Enhanced Image Restoration via Mamba-Transformer Aggregation
by: Jiang, Aiwen, et al.
Published: (2024)
by: Jiang, Aiwen, et al.
Published: (2024)
PANICL: Mitigating Over-Reliance on Single Prompt in Visual In-Context Learning
by: Zhang, Jiahao, et al.
Published: (2025)
by: Zhang, Jiahao, et al.
Published: (2025)
Learning Visual Prompts for Guiding the Attention of Vision Transformers
by: Rezaei, Razieh, et al.
Published: (2024)
by: Rezaei, Razieh, et al.
Published: (2024)
Visual Prompting in LLMs for Enhancing Emotion Recognition
by: Zhang, Qixuan, et al.
Published: (2024)
by: Zhang, Qixuan, et al.
Published: (2024)
User-in-the-Loop View Sampling with Error Peaking Visualization
by: Yasunaga, Ayaka, et al.
Published: (2025)
by: Yasunaga, Ayaka, et al.
Published: (2025)
Adversarially Diversified Rehearsal Memory (ADRM): Mitigating Memory Overfitting Challenge in Continual Learning
by: Khan, Hikmat, et al.
Published: (2024)
by: Khan, Hikmat, et al.
Published: (2024)
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
by: Wang, Zhifeng, et al.
Published: (2025)
by: Wang, Zhifeng, et al.
Published: (2025)
Prompt Generation Networks for Input-Space Adaptation of Frozen Vision Transformers
by: Loedeman, Jochem, et al.
Published: (2022)
by: Loedeman, Jochem, et al.
Published: (2022)
DA-VPT: Semantic-Guided Visual Prompt Tuning for Vision Transformers
by: Ren, Li, et al.
Published: (2025)
by: Ren, Li, et al.
Published: (2025)
Unveil Benign Overfitting for Transformer in Vision: Training Dynamics, Convergence, and Generalization
by: Jiang, Jiarui, et al.
Published: (2024)
by: Jiang, Jiarui, et al.
Published: (2024)
Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts
by: Özdemir, Övgü, et al.
Published: (2024)
by: Özdemir, Övgü, et al.
Published: (2024)
SEP: Self-Enhanced Prompt Tuning for Visual-Language Model
by: Yao, Hantao, et al.
Published: (2024)
by: Yao, Hantao, et al.
Published: (2024)
Attend and Enrich: Enhanced Visual Prompt for Zero-Shot Learning
by: Liu, Man, et al.
Published: (2024)
by: Liu, Man, et al.
Published: (2024)
Retrieval-Enhanced Visual Prompt Learning for Few-shot Classification
by: Rong, Jintao, et al.
Published: (2023)
by: Rong, Jintao, et al.
Published: (2023)
Expanding Zero-Shot Object Counting with Rich Prompts
by: Zhu, Huilin, et al.
Published: (2025)
by: Zhu, Huilin, et al.
Published: (2025)
EDIT: Enhancing Vision Transformers by Mitigating Attention Sink through an Encoder-Decoder Architecture
by: Feng, Wenfeng, et al.
Published: (2025)
by: Feng, Wenfeng, et al.
Published: (2025)
E-InMeMo: Enhanced Prompting for Visual In-Context Learning
by: Zhang, Jiahao, et al.
Published: (2025)
by: Zhang, Jiahao, et al.
Published: (2025)
IF-Bench: Benchmarking and Enhancing MLLMs for Infrared Images with Generative Visual Prompting
by: Zhang, Tao, et al.
Published: (2025)
by: Zhang, Tao, et al.
Published: (2025)
Enhancing Medical Visual Grounding via Knowledge-guided Spatial Prompts
by: Gao, Yifan, et al.
Published: (2026)
by: Gao, Yifan, et al.
Published: (2026)
ViKey: Enhancing Temporal Understanding in Videos via Visual Prompting
by: Lee, Yeonkyung, et al.
Published: (2026)
by: Lee, Yeonkyung, et al.
Published: (2026)
Hints of Prompt: Enhancing Visual Representation for Multimodal LLMs in Autonomous Driving
by: Zhou, Hao, et al.
Published: (2024)
by: Zhou, Hao, et al.
Published: (2024)
SOUPLE: Enhancing Audio-Visual Localization and Segmentation with Learnable Prompt Contexts
by: Nguyen, Khanh Binh, et al.
Published: (2026)
by: Nguyen, Khanh Binh, et al.
Published: (2026)
Exploring Multimodal Diffusion Transformers for Enhanced Prompt-based Image Editing
by: Shin, Joonghyuk, et al.
Published: (2025)
by: Shin, Joonghyuk, et al.
Published: (2025)
Visual Prompt Tuning in Null Space for Continual Learning
by: Lu, Yue, et al.
Published: (2024)
by: Lu, Yue, et al.
Published: (2024)
Enhancing Multimodal Large Language Models with Multi-instance Visual Prompt Generator for Visual Representation Enrichment
by: Zhong, Wenliang, et al.
Published: (2024)
by: Zhong, Wenliang, et al.
Published: (2024)
Beyond Overfitting: Doubly Adaptive Dropout for Generalizable AU Detection
by: Li, Yong, et al.
Published: (2025)
by: Li, Yong, et al.
Published: (2025)
AVCD: Mitigating Hallucinations in Audio-Visual Large Language Models through Contrastive Decoding
by: Jung, Chaeyoung, et al.
Published: (2025)
by: Jung, Chaeyoung, et al.
Published: (2025)
Mitigating Overfitting in Medical Imaging: Self-Supervised Pretraining vs. ImageNet Transfer Learning for Dermatological Diagnosis
by: Matas, Iván, et al.
Published: (2025)
by: Matas, Iván, et al.
Published: (2025)
Inferring Neural Signed Distance Functions by Overfitting on Single Noisy Point Clouds through Finetuning Data-Driven based Priors
by: Chen, Chao, et al.
Published: (2024)
by: Chen, Chao, et al.
Published: (2024)
Token Coordinated Prompt Attention is Needed for Visual Prompting
by: Liu, Zichen, et al.
Published: (2025)
by: Liu, Zichen, et al.
Published: (2025)
Explicit Visual Prompts for Visual Object Tracking
by: Shi, Liangtao, et al.
Published: (2024)
by: Shi, Liangtao, et al.
Published: (2024)
Discovering and Mitigating Visual Biases through Keyword Explanation
by: Kim, Younghyun, et al.
Published: (2023)
by: Kim, Younghyun, et al.
Published: (2023)
Hallucination Mitigation Prompts Long-term Video Understanding
by: Sun, Yiwei, et al.
Published: (2024)
by: Sun, Yiwei, et al.
Published: (2024)
Similar Items
-
Pseudo Multi-Source Domain Generalization: Bridging the Gap Between Single and Multi-Source Domain Generalization
by: Enomoto, Shohei
Published: (2025) -
EntProp: High Entropy Propagation for Improving Accuracy and Robustness
by: Enomoto, Shohei
Published: (2024) -
MultiModal Fine-tuning with Synthetic Captions
by: Enomoto, Shohei, et al.
Published: (2026) -
Dynamic Test-Time Augmentation via Differentiable Functions
by: Enomoto, Shohei, et al.
Published: (2022) -
Test-time Similarity Modification for Person Re-identification toward Temporal Distribution Shift
by: Adachi, Kazuki, et al.
Published: (2024)