DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Ai, Yuang, Fan, Qihang, Hu, Xuefeng, Yang, Zhenheng, He, Ran, Huang, Huaibo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rectifying Magnitude Neglect in Linear Attention
by: Fan, Qihang, et al.
Published: (2025)
by: Fan, Qihang, et al.
Published: (2025)
Random Wins All: Rethinking Grouping Strategies for Vision Tokens
by: Fan, Qihang, et al.
Published: (2026)
by: Fan, Qihang, et al.
Published: (2026)
LoRA-IR: Taming Low-Rank Experts for Efficient All-in-One Image Restoration
by: Ai, Yuang, et al.
Published: (2024)
by: Ai, Yuang, et al.
Published: (2024)
Breaking Complexity Barriers: High-Resolution Image Restoration with Rank Enhanced Linear Attention
by: Ai, Yuang, et al.
Published: (2025)
by: Ai, Yuang, et al.
Published: (2025)
Conv-Adapter: Exploring Parameter Efficient Transfer Learning for ConvNets
by: Chen, Hao, et al.
Published: (2022)
by: Chen, Hao, et al.
Published: (2022)
Expand and Prune: Maximizing Trajectory Diversity for Effective GRPO in Generative Models
by: Ge, Shiran, et al.
Published: (2025)
by: Ge, Shiran, et al.
Published: (2025)
InfiMM-WebMath-40B: Advancing Multimodal Pre-Training for Enhanced Mathematical Reasoning
by: Han, Xiaotian, et al.
Published: (2024)
by: Han, Xiaotian, et al.
Published: (2024)
Breaking the Low-Rank Dilemma of Linear Attention
by: Fan, Qihang, et al.
Published: (2024)
by: Fan, Qihang, et al.
Published: (2024)
Designing Concise ConvNets with Columnar Stages
by: Kumar, Ashish, et al.
Published: (2024)
by: Kumar, Ashish, et al.
Published: (2024)
Semantic Equitable Clustering: A Simple and Effective Strategy for Clustering Vision Tokens
by: Fan, Qihang, et al.
Published: (2024)
by: Fan, Qihang, et al.
Published: (2024)
Lightweight Vision Transformer with Bidirectional Interaction
by: Fan, Qihang, et al.
Published: (2023)
by: Fan, Qihang, et al.
Published: (2023)
Pick-or-Mix: Dynamic Channel Sampling for ConvNets
by: Kumar, Ashish, et al.
Published: (2024)
by: Kumar, Ashish, et al.
Published: (2024)
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image Restoration
by: Ai, Yuang, et al.
Published: (2023)
by: Ai, Yuang, et al.
Published: (2023)
Uncertainty-Aware Source-Free Adaptive Image Super-Resolution with Wavelet Augmentation Transformer
by: Ai, Yuang, et al.
Published: (2023)
by: Ai, Yuang, et al.
Published: (2023)
Revealing the Dark Secrets of Extremely Large Kernel ConvNets on Robustness
by: Chen, Honghao, et al.
Published: (2024)
by: Chen, Honghao, et al.
Published: (2024)
Efficient Deformable ConvNets: Rethinking Dynamic and Sparse Operator for Vision Applications
by: Xiong, Yuwen, et al.
Published: (2024)
by: Xiong, Yuwen, et al.
Published: (2024)
DiCo: Disentangled Concept Representation for Text-to-image Person Re-identification
by: Kim, Giyeol, et al.
Published: (2026)
by: Kim, Giyeol, et al.
Published: (2026)
DECO: Unleashing the Potential of ConvNets for Query-based Detection and Segmentation
by: Chen, Xinghao, et al.
Published: (2023)
by: Chen, Xinghao, et al.
Published: (2023)
Learning to Generate Parameters of ConvNets for Unseen Image Data
by: Wang, Shiye, et al.
Published: (2023)
by: Wang, Shiye, et al.
Published: (2023)
VideoMAC: Video Masked Autoencoders Meet ConvNets
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
PeLK: Parameter-efficient Large Kernel ConvNets with Peripheral Convolution
by: Chen, Honghao, et al.
Published: (2024)
by: Chen, Honghao, et al.
Published: (2024)
BitDance: Scaling Autoregressive Generative Models with Binary Tokens
by: Ai, Yuang, et al.
Published: (2026)
by: Ai, Yuang, et al.
Published: (2026)
RMT: Retentive Networks Meet Vision Transformers
by: Fan, Qihang, et al.
Published: (2023)
by: Fan, Qihang, et al.
Published: (2023)
Advancing Vision Transformer with Enhanced Spatial Priors
by: Fan, Qihang, et al.
Published: (2026)
by: Fan, Qihang, et al.
Published: (2026)
UniConvNet: Expanding Effective Receptive Field while Maintaining Asymptotically Gaussian Distribution for ConvNets of Any Scale
by: Wang, Yuhao, et al.
Published: (2025)
by: Wang, Yuhao, et al.
Published: (2025)
iFormer: Integrating ConvNet and Transformer for Mobile Application
by: Zheng, Chuanyang
Published: (2025)
by: Zheng, Chuanyang
Published: (2025)
MixMask: Revisiting Masking Strategy for Siamese ConvNets
by: Vishniakov, Kirill, et al.
Published: (2022)
by: Vishniakov, Kirill, et al.
Published: (2022)
WaveDH: Wavelet Sub-bands Guided ConvNet for Efficient Image Dehazing
by: Hwang, Seongmin, et al.
Published: (2024)
by: Hwang, Seongmin, et al.
Published: (2024)
Universal Neural Architecture Space: Covering ConvNets, Transformers and Everything in Between
by: Týbl, Ondřej, et al.
Published: (2025)
by: Týbl, Ondřej, et al.
Published: (2025)
Infrared Object Detection with Ultra Small ConvNets: Is ImageNet Pretraining Still Useful?
by: Muralidharan, Srikanth, et al.
Published: (2025)
by: Muralidharan, Srikanth, et al.
Published: (2025)
Vision Transformer with Sparse Scan Prior
by: Zhang, Yuguang, et al.
Published: (2024)
by: Zhang, Yuguang, et al.
Published: (2024)
ConvNet vs Transformer, Supervised vs CLIP: Beyond ImageNet Accuracy
by: Vishniakov, Kirill, et al.
Published: (2023)
by: Vishniakov, Kirill, et al.
Published: (2023)
Embedded ConvNet Ensembles: A Lightweight Approach to Recognize Arabic Handwritten Characters
by: Khayati, Mohsine El, et al.
Published: (2026)
by: Khayati, Mohsine El, et al.
Published: (2026)
Masking Improves Contrastive Self-Supervised Learning for ConvNets, and Saliency Tells You Where
by: Chin, Zhi-Yi, et al.
Published: (2023)
by: Chin, Zhi-Yi, et al.
Published: (2023)
Threats to Arabic Handwriting Recognition: Investigating Black-Box Adversarial Attacks on embedded ConvNet models
by: Khayati, Mohsine EL, et al.
Published: (2026)
by: Khayati, Mohsine EL, et al.
Published: (2026)
MobileIE: An Extremely Lightweight and Effective ConvNet for Real-Time Image Enhancement on Mobile Devices
by: Yan, Hailong, et al.
Published: (2025)
by: Yan, Hailong, et al.
Published: (2025)
PaPr: Training-Free One-Step Patch Pruning with Lightweight ConvNets for Faster Inference
by: Mahmud, Tanvir, et al.
Published: (2024)
by: Mahmud, Tanvir, et al.
Published: (2024)
Human-annotated label noise and their impact on ConvNets for remote sensing image scene classification
by: Peng, Longkang, et al.
Published: (2023)
by: Peng, Longkang, et al.
Published: (2023)
ConvNets for Counting: Object Detection of Transient Phenomena in Steelpan Drums
by: Hawley, Scott H., et al.
Published: (2021)
by: Hawley, Scott H., et al.
Published: (2021)
OverLoCK: An Overview-first-Look-Closely-next ConvNet with Context-Mixing Dynamic Kernels
by: Lou, Meng, et al.
Published: (2025)
by: Lou, Meng, et al.
Published: (2025)
Similar Items
-
Rectifying Magnitude Neglect in Linear Attention
by: Fan, Qihang, et al.
Published: (2025) -
Random Wins All: Rethinking Grouping Strategies for Vision Tokens
by: Fan, Qihang, et al.
Published: (2026) -
LoRA-IR: Taming Low-Rank Experts for Efficient All-in-One Image Restoration
by: Ai, Yuang, et al.
Published: (2024) -
Breaking Complexity Barriers: High-Resolution Image Restoration with Rank Enhanced Linear Attention
by: Ai, Yuang, et al.
Published: (2025) -
Conv-Adapter: Exploring Parameter Efficient Transfer Learning for ConvNets
by: Chen, Hao, et al.
Published: (2022)