SCHEME: Scalable Channel Mixer for Vision Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Sridhar, Deepak, Li, Yunsheng, Vasconcelos, Nuno |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompt Sliders for Fine-Grained Control, Editing and Erasing of Concepts in Diffusion Models
by: Sridhar, Deepak, et al.
Published: (2024)
by: Sridhar, Deepak, et al.
Published: (2024)
Adapting Diffusion Models for Improved Prompt Compliance and Controllable Image Synthesis
by: Sridhar, Deepak, et al.
Published: (2024)
by: Sridhar, Deepak, et al.
Published: (2024)
ViT-DD: Multi-Task Vision Transformer for Semi-Supervised Driver Distraction Detection
by: Ma, Yunsheng, et al.
Published: (2022)
by: Ma, Yunsheng, et al.
Published: (2022)
ParFormer: A Vision Transformer with Parallel Mixer and Sparse Channel Attention Patch Embedding
by: Setyawan, Novendra, et al.
Published: (2024)
by: Setyawan, Novendra, et al.
Published: (2024)
CS-Mixer: A Cross-Scale Vision MLP Model with Spatial-Channel Mixing
by: Cui, Jonathan, et al.
Published: (2023)
by: Cui, Jonathan, et al.
Published: (2023)
ChebMixer: Efficient Graph Representation Learning with MLP Mixer
by: Kui, Xiaoyan, et al.
Published: (2024)
by: Kui, Xiaoyan, et al.
Published: (2024)
Video Reasoning without Training
by: Sridhar, Deepak, et al.
Published: (2025)
by: Sridhar, Deepak, et al.
Published: (2025)
HSViT: Horizontally Scalable Vision Transformer
by: Xu, Chenhao, et al.
Published: (2024)
by: Xu, Chenhao, et al.
Published: (2024)
Exploring Synergistic Ensemble Learning: Uniting CNNs, MLP-Mixers, and Vision Transformers to Enhance Image Classification
by: Bashar, Mk, et al.
Published: (2025)
by: Bashar, Mk, et al.
Published: (2025)
Improving image synthesis with diffusion-negative sampling
by: Desai, Alakh, et al.
Published: (2024)
by: Desai, Alakh, et al.
Published: (2024)
Diffusion Models with Adaptive Negative Sampling Without External Resources
by: Desai, Alakh, et al.
Published: (2025)
by: Desai, Alakh, et al.
Published: (2025)
MACMD: Multi-dilated Contextual Attention and Channel Mixer Decoding for Medical Image Segmentation
by: Maurya, Lalit, et al.
Published: (2025)
by: Maurya, Lalit, et al.
Published: (2025)
STS-Mixer: Spatio-Temporal-Spectral Mixer for 4D Point Cloud Video Understanding
by: Li, Wenhao, et al.
Published: (2026)
by: Li, Wenhao, et al.
Published: (2026)
GraViT: Transfer Learning with Vision Transformers and MLP-Mixer for Strong Gravitational Lens Discovery
by: Parlange, René, et al.
Published: (2025)
by: Parlange, René, et al.
Published: (2025)
Isolated Channel Vision Transformers: From Single-Channel Pretraining to Multi-Channel Finetuning
by: Lian, Wenyi, et al.
Published: (2025)
by: Lian, Wenyi, et al.
Published: (2025)
Adapting Dual-encoder Vision-language Models for Paraphrased Retrieval
by: Cheng, Jiacheng, et al.
Published: (2024)
by: Cheng, Jiacheng, et al.
Published: (2024)
Windowed-FourierMixer: Enhancing Clutter-Free Room Modeling with Fourier Transform
by: Henriques, Bruno, et al.
Published: (2024)
by: Henriques, Bruno, et al.
Published: (2024)
Fairness and Bias Mitigation in Computer Vision: A Survey
by: Dehdashtian, Sepehr, et al.
Published: (2024)
by: Dehdashtian, Sepehr, et al.
Published: (2024)
NOVO: Unlearning-Compliant Vision Transformers
by: Roy, Soumya, et al.
Published: (2025)
by: Roy, Soumya, et al.
Published: (2025)
EditAR: Unified Conditional Generation with Autoregressive Models
by: Mu, Jiteng, et al.
Published: (2025)
by: Mu, Jiteng, et al.
Published: (2025)
Diffusion-based Data Augmentation for Object Counting Problems
by: Wang, Zhen, et al.
Published: (2024)
by: Wang, Zhen, et al.
Published: (2024)
MixerFlow: MLP-Mixer meets Normalising Flows
by: English, Eshant, et al.
Published: (2023)
by: English, Eshant, et al.
Published: (2023)
The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer
by: Lei, Weixian, et al.
Published: (2025)
by: Lei, Weixian, et al.
Published: (2025)
EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality
by: Lee, Sanghyeok, et al.
Published: (2024)
by: Lee, Sanghyeok, et al.
Published: (2024)
ATOM: Attention Mixer for Efficient Dataset Distillation
by: Khaki, Samir, et al.
Published: (2024)
by: Khaki, Samir, et al.
Published: (2024)
ProTeCt: Prompt Tuning for Taxonomic Open Set Classification
by: Wu, Tz-Ying, et al.
Published: (2023)
by: Wu, Tz-Ying, et al.
Published: (2023)
Long-Tailed Anomaly Detection with Learnable Class Names
by: Ho, Chih-Hui, et al.
Published: (2024)
by: Ho, Chih-Hui, et al.
Published: (2024)
VistaFormer: Scalable Vision Transformers for Satellite Image Time Series Segmentation
by: MacDonald, Ezra, et al.
Published: (2024)
by: MacDonald, Ezra, et al.
Published: (2024)
Elastic Attention Cores for Scalable Vision Transformers
by: Song, Alan Z., et al.
Published: (2026)
by: Song, Alan Z., et al.
Published: (2026)
MixerCSeg: An Efficient Mixer Architecture for Crack Segmentation via Decoupled Mamba Attention
by: Zhao, Zilong, et al.
Published: (2026)
by: Zhao, Zilong, et al.
Published: (2026)
GroupedMixer: An Entropy Model with Group-wise Token-Mixers for Learned Image Compression
by: Li, Daxin, et al.
Published: (2024)
by: Li, Daxin, et al.
Published: (2024)
Hyperspectral Image Classification using Spectral-Spatial Mixer Network
by: Alkhatib, Mohammed Q.
Published: (2025)
by: Alkhatib, Mohammed Q.
Published: (2025)
MixerMDM: Learnable Composition of Human Motion Diffusion Models
by: Ruiz-Ponce, Pablo, et al.
Published: (2025)
by: Ruiz-Ponce, Pablo, et al.
Published: (2025)
Lightweight Transformer-Driven Segmentation of Hotspots and Snail Trails in Solar PV Thermal Imagery
by: Joshi, Deepak, et al.
Published: (2025)
by: Joshi, Deepak, et al.
Published: (2025)
MixerSENet: A Lightweight Framework for Efficient Hyperspectral Image Classification
by: Alkhatib, Mohammed Q., et al.
Published: (2026)
by: Alkhatib, Mohammed Q., et al.
Published: (2026)
FFNet: MetaMixer-based Efficient Convolutional Mixer Design
by: Yun, Seokju, et al.
Published: (2024)
by: Yun, Seokju, et al.
Published: (2024)
VARS: Vision-based Assessment of Risk in Security Systems
by: Gupta, Pranav, et al.
Published: (2024)
by: Gupta, Pranav, et al.
Published: (2024)
Representative Attention For Vision Transformers
by: Li, Yuntong, et al.
Published: (2026)
by: Li, Yuntong, et al.
Published: (2026)
OU-CoViT: Copula-Enhanced Bi-Channel Multi-Task Vision Transformers with Dual Adaptation for OU-UWF Images
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
DisentangleFormer: Spatial-Channel Decoupling for Multi-Channel Vision
by: Liao, Jiashu, et al.
Published: (2025)
by: Liao, Jiashu, et al.
Published: (2025)
Similar Items
-
Prompt Sliders for Fine-Grained Control, Editing and Erasing of Concepts in Diffusion Models
by: Sridhar, Deepak, et al.
Published: (2024) -
Adapting Diffusion Models for Improved Prompt Compliance and Controllable Image Synthesis
by: Sridhar, Deepak, et al.
Published: (2024) -
ViT-DD: Multi-Task Vision Transformer for Semi-Supervised Driver Distraction Detection
by: Ma, Yunsheng, et al.
Published: (2022) -
ParFormer: A Vision Transformer with Parallel Mixer and Sparse Channel Attention Patch Embedding
by: Setyawan, Novendra, et al.
Published: (2024) -
CS-Mixer: A Cross-Scale Vision MLP Model with Spatial-Channel Mixing
by: Cui, Jonathan, et al.
Published: (2023)