Multi-Scale Visual Prompting for Lightweight Small-Image Classification
Fuente:
arXiv
Guardado en:
| Autor principal: | Khazem, Salim |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
TopoLoRA-SAM: Topology-Aware Parameter-Efficient Adaptation of Foundation Segmenters for Thin-Structure and Cross-Domain Binary Semantic Segmentation
por: Khazem, Salim
Publicado: (2026)
por: Khazem, Salim
Publicado: (2026)
SAFE-KD: Risk-Controlled Early-Exit Distillation for Vision Backbones
por: Khazem, Salim
Publicado: (2026)
por: Khazem, Salim
Publicado: (2026)
Margin and Consistency Supervision for Calibrated and Robust Vision Models
por: Khazem, Salim
Publicado: (2026)
por: Khazem, Salim
Publicado: (2026)
AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers
por: Khazem, Salim
Publicado: (2026)
por: Khazem, Salim
Publicado: (2026)
PolygoNet: Leveraging Simplified Polygonal Representation for Effective Image Classification
por: Khazem, Salim, et al.
Publicado: (2025)
por: Khazem, Salim, et al.
Publicado: (2025)
Cyclical Temporal Encoding and Hybrid Deep Ensembles for Multistep Energy Forecasting
por: Khazem, Salim, et al.
Publicado: (2025)
por: Khazem, Salim, et al.
Publicado: (2025)
BYO-Eval: Build Your Own Dataset for Fine-Grained Visual Assessment of Multimodal Language Models
por: Arnould, Ludovic, et al.
Publicado: (2025)
por: Arnould, Ludovic, et al.
Publicado: (2025)
MC-RFM: Geometry-Aware Few-Shot Adaptation via Mixed-Curvature Riemannian Flow Matching
por: Khazem, Salim, et al.
Publicado: (2026)
por: Khazem, Salim, et al.
Publicado: (2026)
PoTATO: A Dataset for Analyzing Polarimetric Traces of Afloat Trash Objects
por: Batista, Luis Felipe Wolf, et al.
Publicado: (2024)
por: Batista, Luis Felipe Wolf, et al.
Publicado: (2024)
PromptSR: Cascade Prompting for Lightweight Image Super-Resolution
por: Liu, Wenyang, et al.
Publicado: (2025)
por: Liu, Wenyang, et al.
Publicado: (2025)
Women Sport Actions Dataset for Visual Classification Using Small Scale Training Data
por: Ray, Palash, et al.
Publicado: (2025)
por: Ray, Palash, et al.
Publicado: (2025)
ISTD-YOLO: A Multi-Scale Lightweight High-Performance Infrared Small Target Detection Algorithm
por: Zhang, Shang, et al.
Publicado: (2025)
por: Zhang, Shang, et al.
Publicado: (2025)
HydraMix: Multi-Image Feature Mixing for Small Data Image Classification
por: Reinders, Christoph, et al.
Publicado: (2025)
por: Reinders, Christoph, et al.
Publicado: (2025)
Pre-training of Lightweight Vision Transformers on Small Datasets with Minimally Scaled Images
por: Tan, Jen Hong
Publicado: (2024)
por: Tan, Jen Hong
Publicado: (2024)
TAI++: Text as Image for Multi-Label Image Classification by Co-Learning Transferable Prompt
por: Wu, Xiangyu, et al.
Publicado: (2024)
por: Wu, Xiangyu, et al.
Publicado: (2024)
IMC-Net: A Lightweight Content-Conditioned Encoder with Multi-Pass Processing for Image Classification
por: Li, YiZhou
Publicado: (2025)
por: Li, YiZhou
Publicado: (2025)
Lightweight Adaptive Feature De-drifting for Compressed Image Classification
por: Peng, Long, et al.
Publicado: (2024)
por: Peng, Long, et al.
Publicado: (2024)
Category-Prompt Refined Feature Learning for Long-Tailed Multi-Label Image Classification
por: Yan, Jiexuan, et al.
Publicado: (2024)
por: Yan, Jiexuan, et al.
Publicado: (2024)
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification
por: Luo, Qiuming, et al.
Publicado: (2026)
por: Luo, Qiuming, et al.
Publicado: (2026)
ProAPO: Progressively Automatic Prompt Optimization for Visual Classification
por: Qu, Xiangyan, et al.
Publicado: (2025)
por: Qu, Xiangyan, et al.
Publicado: (2025)
Retrieval-Enhanced Visual Prompt Learning for Few-shot Classification
por: Rong, Jintao, et al.
Publicado: (2023)
por: Rong, Jintao, et al.
Publicado: (2023)
Lightweight Vision Transformer with Window and Spatial Attention for Food Image Classification
por: Gao, Xinle, et al.
Publicado: (2025)
por: Gao, Xinle, et al.
Publicado: (2025)
MixerSENet: A Lightweight Framework for Efficient Hyperspectral Image Classification
por: Alkhatib, Mohammed Q., et al.
Publicado: (2026)
por: Alkhatib, Mohammed Q., et al.
Publicado: (2026)
VeloxNet: Efficient Spatial Gating for Lightweight Embedded Image Classification
por: Ferdaus, Md Meftahul, et al.
Publicado: (2026)
por: Ferdaus, Md Meftahul, et al.
Publicado: (2026)
Reliable Deep Learning for Small-Scale Classifications: Experiments on Real-World Image Datasets from Bangladesh
por: Suny, Alfe, et al.
Publicado: (2026)
por: Suny, Alfe, et al.
Publicado: (2026)
DualCap: Enhancing Lightweight Image Captioning via Dual Retrieval with Similar Scenes Visual Prompts
por: Li, Binbin, et al.
Publicado: (2025)
por: Li, Binbin, et al.
Publicado: (2025)
MobileMamba: Lightweight Multi-Receptive Visual Mamba Network
por: He, Haoyang, et al.
Publicado: (2024)
por: He, Haoyang, et al.
Publicado: (2024)
Multi-dimensional Visual Prompt Enhanced Image Restoration via Mamba-Transformer Aggregation
por: Jiang, Aiwen, et al.
Publicado: (2024)
por: Jiang, Aiwen, et al.
Publicado: (2024)
PVLR: Prompt-driven Visual-Linguistic Representation Learning for Multi-Label Image Recognition
por: Tan, Hao, et al.
Publicado: (2024)
por: Tan, Hao, et al.
Publicado: (2024)
Visual Textualization for Image Prompted Object Detection
por: Wu, Yongjian, et al.
Publicado: (2025)
por: Wu, Yongjian, et al.
Publicado: (2025)
EPIC: Efficient Prompt Interaction for Text-Image Classification
por: Yu, Xinyao, et al.
Publicado: (2025)
por: Yu, Xinyao, et al.
Publicado: (2025)
X-Prompt: Multi-modal Visual Prompt for Video Object Segmentation
por: Guo, Pinxue, et al.
Publicado: (2024)
por: Guo, Pinxue, et al.
Publicado: (2024)
Convolutional Networks as Extremely Small Foundation Models: Visual Prompting and Theoretical Perspective
por: Wangni, Jianqiao
Publicado: (2024)
por: Wangni, Jianqiao
Publicado: (2024)
Visually Consistent Hierarchical Image Classification
por: Park, Seulki, et al.
Publicado: (2024)
por: Park, Seulki, et al.
Publicado: (2024)
TriLiteNet: Lightweight Model for Multi-Task Visual Perception
por: Che, Quang-Huy, et al.
Publicado: (2025)
por: Che, Quang-Huy, et al.
Publicado: (2025)
ViPCap: Retrieval Text-Based Visual Prompts for Lightweight Image Captioning
por: Kim, Taewhan, et al.
Publicado: (2024)
por: Kim, Taewhan, et al.
Publicado: (2024)
MAPLE: Multi-scale Attribute-enhanced Prompt Learning for Few-shot Whole Slide Image Classification
por: Zhou, Junjie, et al.
Publicado: (2025)
por: Zhou, Junjie, et al.
Publicado: (2025)
MSCPT: Few-shot Whole Slide Image Classification with Multi-scale and Context-focused Prompt Tuning
por: Han, Minghao, et al.
Publicado: (2024)
por: Han, Minghao, et al.
Publicado: (2024)
Pathology-knowledge Enhanced Multi-instance Prompt Learning for Few-shot Whole Slide Image Classification
por: Qu, Linhao, et al.
Publicado: (2024)
por: Qu, Linhao, et al.
Publicado: (2024)
Ejemplares similares
-
TopoLoRA-SAM: Topology-Aware Parameter-Efficient Adaptation of Foundation Segmenters for Thin-Structure and Cross-Domain Binary Semantic Segmentation
por: Khazem, Salim
Publicado: (2026) -
SAFE-KD: Risk-Controlled Early-Exit Distillation for Vision Backbones
por: Khazem, Salim
Publicado: (2026) -
Margin and Consistency Supervision for Calibrated and Robust Vision Models
por: Khazem, Salim
Publicado: (2026) -
AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers
por: Khazem, Salim
Publicado: (2026) -
PolygoNet: Leveraging Simplified Polygonal Representation for Effective Image Classification
por: Khazem, Salim, et al.
Publicado: (2025)