Scaling Up Your Kernels: Large Kernel Design in ConvNets towards Universal Representations
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yiyuan, Ding, Xiaohan, Yue, Xiangyu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UniRepLKNet: A Universal Perception Large-Kernel ConvNet for Audio, Video, Point Cloud, Time-Series and Image Recognition
by: Ding, Xiaohan, et al.
Published: (2023)
by: Ding, Xiaohan, et al.
Published: (2023)
iFormer: Integrating ConvNet and Transformer for Mobile Application
by: Zheng, Chuanyang
Published: (2025)
by: Zheng, Chuanyang
Published: (2025)
MixMask: Revisiting Masking Strategy for Siamese ConvNets
by: Vishniakov, Kirill, et al.
Published: (2022)
by: Vishniakov, Kirill, et al.
Published: (2022)
RepVGG-GELAN: Enhanced GELAN with VGG-STYLE ConvNets for Brain Tumour Detection
by: Balakrishnan, Thennarasi, et al.
Published: (2024)
by: Balakrishnan, Thennarasi, et al.
Published: (2024)
Masking Improves Contrastive Self-Supervised Learning for ConvNets, and Saliency Tells You Where
by: Chin, Zhi-Yi, et al.
Published: (2023)
by: Chin, Zhi-Yi, et al.
Published: (2023)
Vision Search Assistant: Empower Vision-Language Models as Multimodal Search Engines
by: Zhang, Zhixin, et al.
Published: (2024)
by: Zhang, Zhixin, et al.
Published: (2024)
Multimodal Pathway: Improve Transformers with Irrelevant Data from Other Modalities
by: Zhang, Yiyuan, et al.
Published: (2024)
by: Zhang, Yiyuan, et al.
Published: (2024)
Conv-Adapter: Exploring Parameter Efficient Transfer Learning for ConvNets
by: Chen, Hao, et al.
Published: (2022)
by: Chen, Hao, et al.
Published: (2022)
InteractiveVideo: User-Centric Controllable Video Generation with Synergistic Multimodal Instructions
by: Zhang, Yiyuan, et al.
Published: (2024)
by: Zhang, Yiyuan, et al.
Published: (2024)
Revealing the Dark Secrets of Extremely Large Kernel ConvNets on Robustness
by: Chen, Honghao, et al.
Published: (2024)
by: Chen, Honghao, et al.
Published: (2024)
Explore the Limits of Omni-modal Pretraining at Scale
by: Zhang, Yiyuan, et al.
Published: (2024)
by: Zhang, Yiyuan, et al.
Published: (2024)
PeLK: Parameter-efficient Large Kernel ConvNets with Peripheral Convolution
by: Chen, Honghao, et al.
Published: (2024)
by: Chen, Honghao, et al.
Published: (2024)
Surface EMG-Based Inter-Session/Inter-Subject Gesture Recognition by Leveraging Lightweight All-ConvNet and Transfer Learning
by: Islam, Md. Rabiul, et al.
Published: (2023)
by: Islam, Md. Rabiul, et al.
Published: (2023)
Subspace Kernel Learning on Tensor Sequences
by: Wang, Lei, et al.
Published: (2026)
by: Wang, Lei, et al.
Published: (2026)
KernelDNA: Dynamic Kernel Sharing via Decoupled Naive Adapters
by: Huang, Haiduo, et al.
Published: (2025)
by: Huang, Haiduo, et al.
Published: (2025)
KernelWarehouse: Rethinking the Design of Dynamic Convolution
by: Li, Chao, et al.
Published: (2024)
by: Li, Chao, et al.
Published: (2024)
ConvNet vs Transformer, Supervised vs CLIP: Beyond ImageNet Accuracy
by: Vishniakov, Kirill, et al.
Published: (2023)
by: Vishniakov, Kirill, et al.
Published: (2023)
Ultrafast-and-Ultralight ConvNet-Based Intelligent Monitoring System for Diagnosing Early-Stage Mpox Anytime and Anywhere
by: Yue, Yubiao, et al.
Published: (2023)
by: Yue, Yubiao, et al.
Published: (2023)
BCFPL: Binary classification ConvNet based Fast Parking space recognition with Low resolution image
by: Zhang, Shuo, et al.
Published: (2024)
by: Zhang, Shuo, et al.
Published: (2024)
Human-annotated label noise and their impact on ConvNets for remote sensing image scene classification
by: Peng, Longkang, et al.
Published: (2023)
by: Peng, Longkang, et al.
Published: (2023)
Online Vectorized HD Map Construction using Geometry
by: Zhang, Zhixin, et al.
Published: (2023)
by: Zhang, Zhixin, et al.
Published: (2023)
MedNeXt: Transformer-driven Scaling of ConvNets for Medical Image Segmentation
by: Roy, Saikat, et al.
Published: (2023)
by: Roy, Saikat, et al.
Published: (2023)
Scaling Up Deep Clustering Methods Beyond ImageNet-1K
by: Adaloglou, Nikolas, et al.
Published: (2024)
by: Adaloglou, Nikolas, et al.
Published: (2024)
Multimodal Long Video Modeling Based on Temporal Dynamic Context
by: Hao, Haoran, et al.
Published: (2025)
by: Hao, Haoran, et al.
Published: (2025)
Designing Concise ConvNets with Columnar Stages
by: Kumar, Ashish, et al.
Published: (2024)
by: Kumar, Ashish, et al.
Published: (2024)
ProKeR: A Kernel Perspective on Few-Shot Adaptation of Large Vision-Language Models
by: Bendou, Yassir, et al.
Published: (2025)
by: Bendou, Yassir, et al.
Published: (2025)
Improving the Effectiveness and Efficiency of Stochastic Neighbour Embedding with Isolation Kernel
by: Zhu, Ye, et al.
Published: (2019)
by: Zhu, Ye, et al.
Published: (2019)
CKGAN: Training Generative Adversarial Networks Using Characteristic Kernel Integral Probability Metrics
by: Zhang, Kuntian, et al.
Published: (2025)
by: Zhang, Kuntian, et al.
Published: (2025)
MedNeXt-v2: Scaling 3D ConvNeXts for Large-Scale Supervised Representation Learning in Medical Image Segmentation
by: Roy, Saikat, et al.
Published: (2025)
by: Roy, Saikat, et al.
Published: (2025)
Self-Attention through Kernel-Eigen Pair Sparse Variational Gaussian Processes
by: Chen, Yingyi, et al.
Published: (2024)
by: Chen, Yingyi, et al.
Published: (2024)
OverLoCK: An Overview-first-Look-Closely-next ConvNet with Context-Mixing Dynamic Kernels
by: Lou, Meng, et al.
Published: (2025)
by: Lou, Meng, et al.
Published: (2025)
When Training-Free NAS Meets Vision Transformer: A Neural Tangent Kernel Perspective
by: Zhou, Qiqi, et al.
Published: (2024)
by: Zhou, Qiqi, et al.
Published: (2024)
Reviving ConvNeXt for Efficient Convolutional Diffusion Models
by: Kwon, Taesung, et al.
Published: (2026)
by: Kwon, Taesung, et al.
Published: (2026)
Graph Your Own Prompt
by: Ding, Xi, et al.
Published: (2025)
by: Ding, Xi, et al.
Published: (2025)
SleepNet and DreamNet: Enriching and Reconstructing Representations for Consolidated Visual Classification
by: Ni, Mingze, et al.
Published: (2024)
by: Ni, Mingze, et al.
Published: (2024)
InceptionNeXt: When Inception Meets ConvNeXt
by: Yu, Weihao, et al.
Published: (2023)
by: Yu, Weihao, et al.
Published: (2023)
ConvNets for Counting: Object Detection of Transient Phenomena in Steelpan Drums
by: Hawley, Scott H., et al.
Published: (2021)
by: Hawley, Scott H., et al.
Published: (2021)
Align Your Query: Representation Alignment for Multimodality Medical Object Detection
by: Seo, Ara, et al.
Published: (2025)
by: Seo, Ara, et al.
Published: (2025)
RPN: Reconciled Polynomial Network Towards Unifying PGMs, Kernel SVMs, MLP and KAN
by: Zhang, Jiawei
Published: (2024)
by: Zhang, Jiawei
Published: (2024)
Scaling 4D Representations
by: Carreira, João, et al.
Published: (2024)
by: Carreira, João, et al.
Published: (2024)
Similar Items
-
UniRepLKNet: A Universal Perception Large-Kernel ConvNet for Audio, Video, Point Cloud, Time-Series and Image Recognition
by: Ding, Xiaohan, et al.
Published: (2023) -
iFormer: Integrating ConvNet and Transformer for Mobile Application
by: Zheng, Chuanyang
Published: (2025) -
MixMask: Revisiting Masking Strategy for Siamese ConvNets
by: Vishniakov, Kirill, et al.
Published: (2022) -
RepVGG-GELAN: Enhanced GELAN with VGG-STYLE ConvNets for Brain Tumour Detection
by: Balakrishnan, Thennarasi, et al.
Published: (2024) -
Masking Improves Contrastive Self-Supervised Learning for ConvNets, and Saliency Tells You Where
by: Chin, Zhi-Yi, et al.
Published: (2023)