EV-CLIP: Efficient Visual Prompt Adaptation for CLIP in Few-shot Action Recognition under Visual Challenges
Fuente:
arXiv
Saved in:
| Main Authors: | Jon, Hyo Jin, Jin, Longbin, Kim, Eun Yi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LVP-CLIP:Revisiting CLIP for Continual Learning with Label Vector Pool
by: Ma, Yue, et al.
Published: (2024)
by: Ma, Yue, et al.
Published: (2024)
One-to-Normal: Anomaly Personalization for Few-shot Anomaly Detection
by: Li, Yiyue, et al.
Published: (2025)
by: Li, Yiyue, et al.
Published: (2025)
CLIP-guided Prototype Modulating for Few-shot Action Recognition
by: Wang, Xiang, et al.
Published: (2023)
by: Wang, Xiang, et al.
Published: (2023)
Joint Neural Networks for One-shot Object Recognition and Detection
by: Vargas, Camilo J., et al.
Published: (2024)
by: Vargas, Camilo J., et al.
Published: (2024)
Motion Consistency Loss for Monocular Visual Odometry with Attention-Based Deep Learning
by: Françani, André O., et al.
Published: (2024)
by: Françani, André O., et al.
Published: (2024)
TALON: Test-time Adaptive Learning for On-the-Fly Category Discovery
by: Wu, Yanan, et al.
Published: (2026)
by: Wu, Yanan, et al.
Published: (2026)
Less Detail, Better Answers: Degradation-Driven Prompting for VQA
by: Han, Haoxuan, et al.
Published: (2026)
by: Han, Haoxuan, et al.
Published: (2026)
QCFace: Image Quality Control for boosting Face Representation & Recognition
by: Doan-Ngo, Duc-Phuong, et al.
Published: (2025)
by: Doan-Ngo, Duc-Phuong, et al.
Published: (2025)
Deep Learning Approaches for Human Action Recognition in Video Data
by: Xie, Yufei
Published: (2024)
by: Xie, Yufei
Published: (2024)
CLIP's Visual Embedding Projector is a Few-shot Cornucopia
by: Fahes, Mohammad, et al.
Published: (2024)
by: Fahes, Mohammad, et al.
Published: (2024)
Learning Discriminative Spatio-temporal Representations for Semi-supervised Action Recognition
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
MA-FSAR: Multimodal Adaptation of CLIP for Few-Shot Action Recognition
by: Xing, Jiazheng, et al.
Published: (2023)
by: Xing, Jiazheng, et al.
Published: (2023)
Multi-view Distillation based on Multi-modal Fusion for Few-shot Action Recognition(CLIP-$\mathrm{M^2}$DF)
by: Guo, Fei, et al.
Published: (2024)
by: Guo, Fei, et al.
Published: (2024)
Efficient Solution of Point-Line Absolute Pose
by: Hruby, Petr, et al.
Published: (2024)
by: Hruby, Petr, et al.
Published: (2024)
Isolated Sign Language Recognition with Segmentation and Pose Estimation
by: Perkins, Daniel, et al.
Published: (2025)
by: Perkins, Daniel, et al.
Published: (2025)
Transformer-Based Model for Monocular Visual Odometry: A Video Understanding Approach
by: Françani, André O., et al.
Published: (2023)
by: Françani, André O., et al.
Published: (2023)
Does CLIP perceive art the same way we do?
by: Asperti, Andrea, et al.
Published: (2025)
by: Asperti, Andrea, et al.
Published: (2025)
CardiacCLIP: Video-based CLIP Adaptation for LVEF Prediction in a Few-shot Manner
by: Du, Yao, et al.
Published: (2025)
by: Du, Yao, et al.
Published: (2025)
RailSafeNet: Visual Scene Understanding for Tram Safety
by: Valach, Ondřej, et al.
Published: (2025)
by: Valach, Ondřej, et al.
Published: (2025)
SPARF: Large-Scale Learning of 3D Sparse Radiance Fields from Few Input Images
by: Hamdi, Abdullah, et al.
Published: (2022)
by: Hamdi, Abdullah, et al.
Published: (2022)
MadCLIP: Few-shot Medical Anomaly Detection with CLIP
by: Shiri, Mahshid, et al.
Published: (2025)
by: Shiri, Mahshid, et al.
Published: (2025)
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions
by: Zong, Chang, et al.
Published: (2025)
by: Zong, Chang, et al.
Published: (2025)
GenCLIP: Generalizing CLIP Prompts for Zero-shot Anomaly Detection
by: Kim, Donghyeong, et al.
Published: (2025)
by: Kim, Donghyeong, et al.
Published: (2025)
FFaceNeRF: Few-shot Face Editing in Neural Radiance Fields
by: Yun, Kwan, et al.
Published: (2025)
by: Yun, Kwan, et al.
Published: (2025)
PAT-VCM: Plug-and-Play Auxiliary Tokens for Video Coding for Machines
by: Jiang, Wei, et al.
Published: (2026)
by: Jiang, Wei, et al.
Published: (2026)
Interactive Image Selection and Training for Brain Tumor Segmentation Network
by: Cerqueira, Matheus A., et al.
Published: (2024)
by: Cerqueira, Matheus A., et al.
Published: (2024)
MvBody: Multi-View-Based Hybrid Transformer Using Optical 3D Body Scan for Explainable Cesarean Section Prediction
by: Cheng, Ruting, et al.
Published: (2025)
by: Cheng, Ruting, et al.
Published: (2025)
MotionFollower: Editing Video Motion via Lightweight Score-Guided Diffusion
by: Tu, Shuyuan, et al.
Published: (2024)
by: Tu, Shuyuan, et al.
Published: (2024)
Performance Decay in Deepfake Detection: The Limitations of Training on Outdated Data
by: Richings, Jack, et al.
Published: (2025)
by: Richings, Jack, et al.
Published: (2025)
LRVS-Fashion: Extending Visual Search with Referring Instructions
by: Lepage, Simon, et al.
Published: (2023)
by: Lepage, Simon, et al.
Published: (2023)
Multimodal Integration Challenges in Emotionally Expressive Child Avatars for Training Applications
by: Salehi, Pegah, et al.
Published: (2025)
by: Salehi, Pegah, et al.
Published: (2025)
SealD-NeRF: Interactive Pixel-Level Editing for Dynamic Scenes by Neural Radiance Fields
by: Huang, Zhentao, et al.
Published: (2024)
by: Huang, Zhentao, et al.
Published: (2024)
An inpainting approach to manipulate asymmetry in pre-operative breast images
by: Montenegro, Helena, et al.
Published: (2025)
by: Montenegro, Helena, et al.
Published: (2025)
Gaussian-Constrained LeJEPA Representations for Unsupervised Scene Discovery and Pose Consistency
by: Mostafa, Mohsen
Published: (2026)
by: Mostafa, Mohsen
Published: (2026)
Real-time Object and Event Detection Service through Computer Vision and Edge Computing
by: Mendes, Marcos, et al.
Published: (2025)
by: Mendes, Marcos, et al.
Published: (2025)
Not all Views are Created Equal: Analyzing Viewpoint Instabilities in Vision Foundation Models
by: Michalkiewicz, Mateusz, et al.
Published: (2024)
by: Michalkiewicz, Mateusz, et al.
Published: (2024)
Learning through Creation: A Hash-Free Framework for On-the-Fly Category Discovery
by: Zhang, Bohan, et al.
Published: (2026)
by: Zhang, Bohan, et al.
Published: (2026)
Spatial-ViLT: Enhancing Visual Spatial Reasoning through Multi-Task Learning
by: Islam, Chashi Mahiul, et al.
Published: (2025)
by: Islam, Chashi Mahiul, et al.
Published: (2025)
Depth Priors in Removal Neural Radiance Fields
by: Guo, Zhihao, et al.
Published: (2024)
by: Guo, Zhihao, et al.
Published: (2024)
Zero-Shot Multi-Criteria Visual Quality Inspection for Semi-Controlled Industrial Environments via Real-Time 3D Digital Twin Simulation
by: Araya-Martinez, Jose Moises, et al.
Published: (2025)
by: Araya-Martinez, Jose Moises, et al.
Published: (2025)
Similar Items
-
LVP-CLIP:Revisiting CLIP for Continual Learning with Label Vector Pool
by: Ma, Yue, et al.
Published: (2024) -
One-to-Normal: Anomaly Personalization for Few-shot Anomaly Detection
by: Li, Yiyue, et al.
Published: (2025) -
CLIP-guided Prototype Modulating for Few-shot Action Recognition
by: Wang, Xiang, et al.
Published: (2023) -
Joint Neural Networks for One-shot Object Recognition and Detection
by: Vargas, Camilo J., et al.
Published: (2024) -
Motion Consistency Loss for Monocular Visual Odometry with Attention-Based Deep Learning
by: Françani, André O., et al.
Published: (2024)