Convolutional Networks as Extremely Small Foundation Models: Visual Prompting and Theoretical Perspective
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Wangni, Jianqiao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GPT Carry-On: Training Foundation Model for Customization Could Be Simple, Scalable and Affordable
von: Wangni, Jianqiao
Veröffentlicht: (2025)
von: Wangni, Jianqiao
Veröffentlicht: (2025)
Convolutional Initialization for Data-Efficient Vision Transformers
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2024)
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2024)
Language Models and Cycle Consistency for Self-Reflective Machine Translation
von: Wangni, Jianqiao
Veröffentlicht: (2024)
von: Wangni, Jianqiao
Veröffentlicht: (2024)
Visual Prompting Upgrades Neural Network Sparsification: A Data-Model Perspective
von: Jin, Can, et al.
Veröffentlicht: (2023)
von: Jin, Can, et al.
Veröffentlicht: (2023)
Sampling Foundational Transformer: A Theoretical Perspective
von: Nguyen, Viet Anh, et al.
Veröffentlicht: (2024)
von: Nguyen, Viet Anh, et al.
Veröffentlicht: (2024)
Multi-Scale Visual Prompting for Lightweight Small-Image Classification
von: Khazem, Salim
Veröffentlicht: (2025)
von: Khazem, Salim
Veröffentlicht: (2025)
Convolutional Prompting meets Language Models for Continual Learning
von: Roy, Anurag, et al.
Veröffentlicht: (2024)
von: Roy, Anurag, et al.
Veröffentlicht: (2024)
Robust Adaptation of Foundation Models with Black-Box Visual Prompting
von: Oh, Changdae, et al.
Veröffentlicht: (2024)
von: Oh, Changdae, et al.
Veröffentlicht: (2024)
Accelerating Vision Foundation Models with Drop-in Depthwise Convolution
von: Scribano, Carmelo, et al.
Veröffentlicht: (2026)
von: Scribano, Carmelo, et al.
Veröffentlicht: (2026)
Feature Visualization in 3D Convolutional Neural Networks
von: Li, Chunpeng, et al.
Veröffentlicht: (2025)
von: Li, Chunpeng, et al.
Veröffentlicht: (2025)
Emergent Extreme-View Geometry in 3D Foundation Models
von: Zhang, Yiwen, et al.
Veröffentlicht: (2025)
von: Zhang, Yiwen, et al.
Veröffentlicht: (2025)
Trading Positional Complexity vs. Deepness in Coordinate Networks
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2022)
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2022)
RRCANet: Recurrent Reusable-Convolution Attention Network for Infrared Small Target Detection
von: Liu, Yongxian, et al.
Veröffentlicht: (2025)
von: Liu, Yongxian, et al.
Veröffentlicht: (2025)
A Taxonomy and Library for Visualizing Learned Features in Convolutional Neural Networks
von: Grün, Felix, et al.
Veröffentlicht: (2016)
von: Grün, Felix, et al.
Veröffentlicht: (2016)
Evaluating Pre-trained Convolutional Neural Networks and Foundation Models as Feature Extractors for Content-based Medical Image Retrieval
von: Mahbod, Amirreza, et al.
Veröffentlicht: (2024)
von: Mahbod, Amirreza, et al.
Veröffentlicht: (2024)
SUNY: A Visual Interpretation Framework for Convolutional Neural Networks from a Necessary and Sufficient Perspective
von: Xuan, Xiwei, et al.
Veröffentlicht: (2023)
von: Xuan, Xiwei, et al.
Veröffentlicht: (2023)
Foundation Cures Personalization: Improving Personalized Models' Prompt Consistency via Hidden Foundation Knowledge
von: Cai, Yiyang, et al.
Veröffentlicht: (2024)
von: Cai, Yiyang, et al.
Veröffentlicht: (2024)
Convolutional Prompting for Broad-Domain Retinal Vessel Segmentation
von: Wei, Qijie, et al.
Veröffentlicht: (2024)
von: Wei, Qijie, et al.
Veröffentlicht: (2024)
Asymmetric Masked Distillation for Pre-Training Small Foundation Models
von: Zhao, Zhiyu, et al.
Veröffentlicht: (2023)
von: Zhao, Zhiyu, et al.
Veröffentlicht: (2023)
Audio-Visual Intelligence in Large Foundation Models
von: Qin, You, et al.
Veröffentlicht: (2026)
von: Qin, You, et al.
Veröffentlicht: (2026)
Attention Lattice Adapter: Visual Explanation Generation for Visual Foundation Model
von: Hirano, Shinnosuke, et al.
Veröffentlicht: (2025)
von: Hirano, Shinnosuke, et al.
Veröffentlicht: (2025)
Detection of Virus and Small Cell Patches in Foci Images Using Switchable Convolution and Feature Pyramid Networks
von: Singh, Amrita, et al.
Veröffentlicht: (2026)
von: Singh, Amrita, et al.
Veröffentlicht: (2026)
InfoSyncNet: Information Synchronization Temporal Convolutional Network for Visual Speech Recognition
von: Xue, Junxiao, et al.
Veröffentlicht: (2025)
von: Xue, Junxiao, et al.
Veröffentlicht: (2025)
BAPLe: Backdoor Attacks on Medical Foundational Models using Prompt Learning
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
Visual Prompt Engineering for Vision Language Models in Radiology
von: Denner, Stefan, et al.
Veröffentlicht: (2024)
von: Denner, Stefan, et al.
Veröffentlicht: (2024)
PF3Det: A Prompted Foundation Feature Assisted Visual LiDAR 3D Detector
von: Li, Kaidong, et al.
Veröffentlicht: (2025)
von: Li, Kaidong, et al.
Veröffentlicht: (2025)
Defective Convolutional Networks
von: Luo, Tiange, et al.
Veröffentlicht: (2019)
von: Luo, Tiange, et al.
Veröffentlicht: (2019)
XS-VID: An Extremely Small Video Object Detection Dataset
von: Guo, Jiahao, et al.
Veröffentlicht: (2024)
von: Guo, Jiahao, et al.
Veröffentlicht: (2024)
Structured Initialization for Attention in Vision Transformers
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2024)
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2024)
Probing the 3D Awareness of Visual Foundation Models
von: Banani, Mohamed El, et al.
Veröffentlicht: (2024)
von: Banani, Mohamed El, et al.
Veröffentlicht: (2024)
Visual Instruction Pretraining for Domain-Specific Foundation Models
von: Li, Yuxuan, et al.
Veröffentlicht: (2025)
von: Li, Yuxuan, et al.
Veröffentlicht: (2025)
ARFC-WAHNet: Adaptive Receptive Field Convolution and Wavelet-Attentive Hierarchical Network for Infrared Small Target Detection
von: Cui, Xingye, et al.
Veröffentlicht: (2025)
von: Cui, Xingye, et al.
Veröffentlicht: (2025)
Efficient Higher-order Convolution for Small Kernels in Deep Learning
von: Wen, Zuocheng, et al.
Veröffentlicht: (2024)
von: Wen, Zuocheng, et al.
Veröffentlicht: (2024)
$ShiftwiseConv:$ Small Convolutional Kernel with Large Kernel Effect
von: Li, Dachong, et al.
Veröffentlicht: (2024)
von: Li, Dachong, et al.
Veröffentlicht: (2024)
Dynamic High-frequency Convolution for Infrared Small Target Detection
von: Li, Ruojing, et al.
Veröffentlicht: (2026)
von: Li, Ruojing, et al.
Veröffentlicht: (2026)
A Foundation Model for DAS Signal Recognition and Visual Prompt Tuning of the Pre-trained Model for Downstream Tasks
von: Gui, Kun, et al.
Veröffentlicht: (2025)
von: Gui, Kun, et al.
Veröffentlicht: (2025)
Anatomy-Aware Text-Visual Fusion with Dual-Perspective Prompts for Fine-Grained Lumbar Spine Segmentation
von: Lian, Sheng, et al.
Veröffentlicht: (2025)
von: Lian, Sheng, et al.
Veröffentlicht: (2025)
Explicit Visual Prompts for Visual Object Tracking
von: Shi, Liangtao, et al.
Veröffentlicht: (2024)
von: Shi, Liangtao, et al.
Veröffentlicht: (2024)
Token Coordinated Prompt Attention is Needed for Visual Prompting
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
AdaFusion: Prompt-Guided Inference with Adaptive Fusion of Pathology Foundation Models
von: Xiao, Yuxiang, et al.
Veröffentlicht: (2025)
von: Xiao, Yuxiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GPT Carry-On: Training Foundation Model for Customization Could Be Simple, Scalable and Affordable
von: Wangni, Jianqiao
Veröffentlicht: (2025) -
Convolutional Initialization for Data-Efficient Vision Transformers
von: Zheng, Jianqiao, et al.
Veröffentlicht: (2024) -
Language Models and Cycle Consistency for Self-Reflective Machine Translation
von: Wangni, Jianqiao
Veröffentlicht: (2024) -
Visual Prompting Upgrades Neural Network Sparsification: A Data-Model Perspective
von: Jin, Can, et al.
Veröffentlicht: (2023) -
Sampling Foundational Transformer: A Theoretical Perspective
von: Nguyen, Viet Anh, et al.
Veröffentlicht: (2024)