TiC: Exploring Vision Transformer in Convolution
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Song, Wang, Qingzhong, Bian, Jiang, Xiong, Haoyi |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ALPS: An Auto-Labeling and Pre-training Scheme for Remote Sensing Segmentation With Segment Anything Model
by: Zhang, Song, et al.
Published: (2024)
by: Zhang, Song, et al.
Published: (2024)
Learning Discriminative Features for Crowd Counting
by: Chen, Yuehai, et al.
Published: (2023)
by: Chen, Yuehai, et al.
Published: (2023)
TiC-CLIP: Continual Training of CLIP Models
by: Garg, Saurabh, et al.
Published: (2023)
by: Garg, Saurabh, et al.
Published: (2023)
P2ANet: A Dataset and Benchmark for Dense Action Detection from Table Tennis Match Broadcasting Videos
by: Bian, Jiang, et al.
Published: (2022)
by: Bian, Jiang, et al.
Published: (2022)
Enhancing Learnable Descriptive Convolutional Vision Transformer for Face Anti-Spoofing
by: Huanga, Pei-Kai, et al.
Published: (2025)
by: Huanga, Pei-Kai, et al.
Published: (2025)
On Convolutional Vision Transformers for Yield Prediction
by: Inderka, Alvin, et al.
Published: (2024)
by: Inderka, Alvin, et al.
Published: (2024)
Depth-Wise Convolutions in Vision Transformers for Efficient Training on Small Datasets
by: Zhang, Tianxiao, et al.
Published: (2024)
by: Zhang, Tianxiao, et al.
Published: (2024)
Convolutional Initialization for Data-Efficient Vision Transformers
by: Zheng, Jianqiao, et al.
Published: (2024)
by: Zheng, Jianqiao, et al.
Published: (2024)
A Benchmark for Vision-Centric HD Mapping by V2I Systems
by: Fan, Miao, et al.
Published: (2025)
by: Fan, Miao, et al.
Published: (2025)
ACC-ViT : Atrous Convolution's Comeback in Vision Transformers
by: Ibtehaz, Nabil, et al.
Published: (2024)
by: Ibtehaz, Nabil, et al.
Published: (2024)
Coordinative Learning with Ordinal and Relational Priors for Volumetric Medical Image Segmentation
by: Wang, Haoyi
Published: (2025)
by: Wang, Haoyi
Published: (2025)
LM-MCVT: A Lightweight Multi-modal Multi-view Convolutional-Vision Transformer Approach for 3D Object Recognition
by: Xiong, Songsong, et al.
Published: (2025)
by: Xiong, Songsong, et al.
Published: (2025)
A Fusion Model for Artwork Identification Based on Convolutional Neural Networks and Transformers
by: Wang, Zhenyu, et al.
Published: (2025)
by: Wang, Zhenyu, et al.
Published: (2025)
Vision Transformers and Convolutional Neural Networks for Land Use Scene Classification
by: Kulkarni, Arun D.
Published: (2026)
by: Kulkarni, Arun D.
Published: (2026)
CAS-ViT: Convolutional Additive Self-attention Vision Transformers for Efficient Mobile Applications
by: Zhang, Tianfang, et al.
Published: (2024)
by: Zhang, Tianfang, et al.
Published: (2024)
ViT-CoMer: Vision Transformer with Convolutional Multi-scale Feature Interaction for Dense Predictions
by: Xia, Chunlong, et al.
Published: (2024)
by: Xia, Chunlong, et al.
Published: (2024)
Joint Multi-scale Gated Transformer and Prior-guided Convolutional Network for Learned Image Compression
by: Chen, Zhengxin, et al.
Published: (2025)
by: Chen, Zhengxin, et al.
Published: (2025)
Diffusion Models in 3D Vision: A Survey
by: Wang, Zhen, et al.
Published: (2024)
by: Wang, Zhen, et al.
Published: (2024)
big.LITTLE Vision Transformer for Efficient Visual Recognition
by: Guo, He, et al.
Published: (2024)
by: Guo, He, et al.
Published: (2024)
GenConViT: Deepfake Video Detection Using Generative Convolutional Vision Transformer
by: Deressa, Deressa Wodajo, et al.
Published: (2023)
by: Deressa, Deressa Wodajo, et al.
Published: (2023)
PDC-ViT : Source Camera Identification using Pixel Difference Convolution and Vision Transformer
by: Elharrouss, Omar, et al.
Published: (2025)
by: Elharrouss, Omar, et al.
Published: (2025)
Exploring Compositionality in Vision Transformers using Wavelet Representations
by: Purushottamdas, Akshad Shyam, et al.
Published: (2025)
by: Purushottamdas, Akshad Shyam, et al.
Published: (2025)
Not All Noises Are Created Equally:Diffusion Noise Selection and Optimization
by: Qi, Zipeng, et al.
Published: (2024)
by: Qi, Zipeng, et al.
Published: (2024)
Group-based Distinctive Image Captioning with Memory Difference Encoding and Attention
by: Wang, Jiuniu, et al.
Published: (2025)
by: Wang, Jiuniu, et al.
Published: (2025)
An Experimental Study on Exploring Strong Lightweight Vision Transformers via Masked Image Modeling Pre-Training
by: Gao, Jin, et al.
Published: (2024)
by: Gao, Jin, et al.
Published: (2024)
AIQViT: Architecture-Informed Post-Training Quantization for Vision Transformers
by: Jiang, Runqing, et al.
Published: (2025)
by: Jiang, Runqing, et al.
Published: (2025)
Exploring Token-Level Augmentation in Vision Transformer for Semi-Supervised Semantic Segmentation
by: Zhang, Dengke, et al.
Published: (2025)
by: Zhang, Dengke, et al.
Published: (2025)
GaussTR: Foundation Model-Aligned Gaussian Transformer for Self-Supervised 3D Spatial Understanding
by: Jiang, Haoyi, et al.
Published: (2024)
by: Jiang, Haoyi, et al.
Published: (2024)
DTC: A Deformable Transposed Convolution Module for Medical Image Segmentation
by: Sun, Chengkun, et al.
Published: (2026)
by: Sun, Chengkun, et al.
Published: (2026)
Other Tokens Matter: Exploring Global and Local Features of Vision Transformers for Object Re-Identification
by: Wang, Yingquan, et al.
Published: (2024)
by: Wang, Yingquan, et al.
Published: (2024)
Beyond Grids: Exploring Elastic Input Sampling for Vision Transformers
by: Pardyl, Adam, et al.
Published: (2023)
by: Pardyl, Adam, et al.
Published: (2023)
ViTGaze: Gaze Following with Interaction Features in Vision Transformers
by: Song, Yuehao, et al.
Published: (2024)
by: Song, Yuehao, et al.
Published: (2024)
WeakTr: Exploring Plain Vision Transformer for Weakly-supervised Semantic Segmentation
by: Zhu, Lianghui, et al.
Published: (2023)
by: Zhu, Lianghui, et al.
Published: (2023)
A Timely Survey on Vision Transformer for Deepfake Detection
by: Wang, Zhikan, et al.
Published: (2024)
by: Wang, Zhikan, et al.
Published: (2024)
RFAConv: Receptive-Field Attention Convolution for Improving Convolutional Neural Networks
by: Zhang, Xin, et al.
Published: (2023)
by: Zhang, Xin, et al.
Published: (2023)
Demystify Transformers & Convolutions in Modern Image Deep Networks
by: Hu, Xiaowei, et al.
Published: (2022)
by: Hu, Xiaowei, et al.
Published: (2022)
A Lightweight Convolution and Vision Transformer integrated model with Multi-scale Self-attention Mechanism
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
LeMeViT: Efficient Vision Transformer with Learnable Meta Tokens for Remote Sensing Image Interpretation
by: Jiang, Wentao, et al.
Published: (2024)
by: Jiang, Wentao, et al.
Published: (2024)
KAConvNet: Kolmogorov-Arnold Convolutional Networks for Vision Recognition
by: Liu, Zhaoxiang, et al.
Published: (2026)
by: Liu, Zhaoxiang, et al.
Published: (2026)
Hierarchical Vision Transformer Enhanced by Graph Convolutional Network for Image Classification
by: Jiao, Haibin
Published: (2026)
by: Jiao, Haibin
Published: (2026)
Similar Items
-
ALPS: An Auto-Labeling and Pre-training Scheme for Remote Sensing Segmentation With Segment Anything Model
by: Zhang, Song, et al.
Published: (2024) -
Learning Discriminative Features for Crowd Counting
by: Chen, Yuehai, et al.
Published: (2023) -
TiC-CLIP: Continual Training of CLIP Models
by: Garg, Saurabh, et al.
Published: (2023) -
P2ANet: A Dataset and Benchmark for Dense Action Detection from Table Tennis Match Broadcasting Videos
by: Bian, Jiang, et al.
Published: (2022) -
Enhancing Learnable Descriptive Convolutional Vision Transformer for Face Anti-Spoofing
by: Huanga, Pei-Kai, et al.
Published: (2025)