Caterpillar: A Pure-MLP Architecture with Shifted-Pillars-Concatenation
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Jin, Shi, Xiaoshuang, Wang, Zhiyuan, Xu, Kaidi, Shen, Heng Tao, Zhu, Xiaofeng |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Conformal Lesion Segmentation for 3D Medical Images
by: Tan, Binyu, et al.
Published: (2025)
by: Tan, Binyu, et al.
Published: (2025)
ACT-Diffusion: Efficient Adversarial Consistency Training for One-step Diffusion Models
by: Kong, Fei, et al.
Published: (2023)
by: Kong, Fei, et al.
Published: (2023)
LRR-Bench: Left, Right or Rotate? Vision-Language models Still Struggle With Spatial Understanding Tasks
by: Kong, Fei, et al.
Published: (2025)
by: Kong, Fei, et al.
Published: (2025)
SpiralMLP: A Lightweight Vision MLP Architecture
by: Mu, Haojie, et al.
Published: (2024)
by: Mu, Haojie, et al.
Published: (2024)
evMLP: An Efficient Event-Driven MLP Architecture for Vision
by: Zheng, Zhentan
Published: (2025)
by: Zheng, Zhentan
Published: (2025)
TruthPrInt: Mitigating Large Vision-Language Models Object Hallucination Via Latent Truthful-Guided Pre-Intervention
by: Duan, Jinhao, et al.
Published: (2025)
by: Duan, Jinhao, et al.
Published: (2025)
Iterative Filter Pruning for Concatenation-based CNN Architectures
by: Pavlitska, Svetlana, et al.
Published: (2024)
by: Pavlitska, Svetlana, et al.
Published: (2024)
On Efficient Variants of Segment Anything Model: A Survey
by: Sun, Xiaorui, et al.
Published: (2024)
by: Sun, Xiaorui, et al.
Published: (2024)
PointMT: Efficient Point Cloud Analysis with Hybrid MLP-Transformer Architecture
by: Zheng, Qiang, et al.
Published: (2024)
by: Zheng, Qiang, et al.
Published: (2024)
Adaptive MLP Pruning for Large Vision Transformers
by: Shen, Chengchao
Published: (2026)
by: Shen, Chengchao
Published: (2026)
Lateralization MLP: A Simple Brain-inspired Architecture for Diffusion
by: Hu, Zizhao, et al.
Published: (2024)
by: Hu, Zizhao, et al.
Published: (2024)
Background Prompt for Few-Shot Out-of-Distribution Detection
by: Cai, Songyue, et al.
Published: (2025)
by: Cai, Songyue, et al.
Published: (2025)
MLPHand: Real Time Multi-View 3D Hand Mesh Reconstruction via MLP Modeling
by: Yang, Jian, et al.
Published: (2024)
by: Yang, Jian, et al.
Published: (2024)
Selective Feature Connection Mechanism: Concatenating Multi-layer CNN Features with a Feature Selector
by: Du, Chen, et al.
Published: (2018)
by: Du, Chen, et al.
Published: (2018)
MLP Can Be A Good Transformer Learner
by: Lin, Sihao, et al.
Published: (2024)
by: Lin, Sihao, et al.
Published: (2024)
GraphMLP: A Graph MLP-Like Architecture for 3D Human Pose Estimation
by: Li, Wenhao, et al.
Published: (2022)
by: Li, Wenhao, et al.
Published: (2022)
Diversity-Guided MLP Reduction for Efficient Large Vision Transformers
by: Shen, Chengchao, et al.
Published: (2025)
by: Shen, Chengchao, et al.
Published: (2025)
KAN or MLP? Point Cloud Shows the Way Forward
by: Shi, Yan, et al.
Published: (2025)
by: Shi, Yan, et al.
Published: (2025)
D2-MLP: Dynamic Decomposed MLP Mixer for Medical Image Segmentation
by: Yang, Jin, et al.
Published: (2024)
by: Yang, Jin, et al.
Published: (2024)
An Efficient MLP-based Point-guided Segmentation Network for Ore Images with Ambiguous Boundary
by: Sun, Guodong, et al.
Published: (2024)
by: Sun, Guodong, et al.
Published: (2024)
PillarTrack:Boosting Pillar Representation for Transformer-based 3D Single Object Tracking on Point Clouds
by: Xu, Weisheng, et al.
Published: (2024)
by: Xu, Weisheng, et al.
Published: (2024)
TimeColor: Flexible Reference Colorization via Temporal Concatenation
by: Sadihin, Bryan Constantine, et al.
Published: (2026)
by: Sadihin, Bryan Constantine, et al.
Published: (2026)
PDSE: A Multiple Lesion Detector for CT Images using PANet and Deformable Squeeze-and-Excitation Block
by: Fan, Di, et al.
Published: (2025)
by: Fan, Di, et al.
Published: (2025)
Layer as Puzzle Pieces: Compressing Large Language Models through Layer Concatenation
by: Wang, Fei, et al.
Published: (2025)
by: Wang, Fei, et al.
Published: (2025)
SiT-MLP: A Simple MLP with Point-wise Topology Feature Learning for Skeleton-based Action Recognition
by: Zhang, Shaojie, et al.
Published: (2023)
by: Zhang, Shaojie, et al.
Published: (2023)
MM-UNet: A Mixed MLP Architecture for Improved Ophthalmic Image Segmentation
by: Xiao, Zunjie, et al.
Published: (2024)
by: Xiao, Zunjie, et al.
Published: (2024)
Window Token Concatenation for Efficient Visual Large Language Models
by: Li, Yifan, et al.
Published: (2025)
by: Li, Yifan, et al.
Published: (2025)
Is Vanilla MLP in Neural Radiance Field Enough for Few-shot View Synthesis?
by: Zhu, Hanxin, et al.
Published: (2024)
by: Zhu, Hanxin, et al.
Published: (2024)
Fast and Lightweight Novel View Synthesis with Differentiable Multiplane Image
by: Zhang, Kaidi, et al.
Published: (2026)
by: Zhang, Kaidi, et al.
Published: (2026)
Are GNNs Worth the Effort for IoT Botnet Detection? A Comparative Study of VAE-GNN vs. ViT-MLP and VAE-MLP Approaches
by: Wasswa, Hassan, et al.
Published: (2025)
by: Wasswa, Hassan, et al.
Published: (2025)
PillarGen: Enhancing Radar Point Cloud Density and Quality via Pillar-based Point Generation Network
by: Kim, Jisong, et al.
Published: (2024)
by: Kim, Jisong, et al.
Published: (2024)
Cross-DINO: Cross the Deep MLP and Transformer for Small Object Detection
by: Cao, Guiping, et al.
Published: (2025)
by: Cao, Guiping, et al.
Published: (2025)
Towards Generalized Range-View LiDAR Segmentation in Adverse Weather
by: Yang, Longyu, et al.
Published: (2025)
by: Yang, Longyu, et al.
Published: (2025)
Synthetic Thermal and RGB Videos for Automatic Pain Assessment utilizing a Vision-MLP Architecture
by: Gkikas, Stefanos, et al.
Published: (2024)
by: Gkikas, Stefanos, et al.
Published: (2024)
VQ-Flow: Taming Normalizing Flows for Multi-Class Anomaly Detection via Hierarchical Vector Quantization
by: Zhou, Yixuan, et al.
Published: (2024)
by: Zhou, Yixuan, et al.
Published: (2024)
PosMLP-Video: Spatial and Temporal Relative Position Encoding for Efficient Video Recognition
by: Hao, Yanbin, et al.
Published: (2024)
by: Hao, Yanbin, et al.
Published: (2024)
ChebMixer: Efficient Graph Representation Learning with MLP Mixer
by: Kui, Xiaoyan, et al.
Published: (2024)
by: Kui, Xiaoyan, et al.
Published: (2024)
PillarMamba: Learning Local-Global Context for Roadside Point Cloud via Hybrid State Space Model
by: Zhang, Zhang, et al.
Published: (2025)
by: Zhang, Zhang, et al.
Published: (2025)
Study of Dropout in PointPillars with 3D Object Detection
by: Sun, Xiaoxiang, et al.
Published: (2024)
by: Sun, Xiaoxiang, et al.
Published: (2024)
Prototype-based Aleatoric Uncertainty Quantification for Cross-modal Retrieval
by: Li, Hao, et al.
Published: (2023)
by: Li, Hao, et al.
Published: (2023)
Similar Items
-
Conformal Lesion Segmentation for 3D Medical Images
by: Tan, Binyu, et al.
Published: (2025) -
ACT-Diffusion: Efficient Adversarial Consistency Training for One-step Diffusion Models
by: Kong, Fei, et al.
Published: (2023) -
LRR-Bench: Left, Right or Rotate? Vision-Language models Still Struggle With Spatial Understanding Tasks
by: Kong, Fei, et al.
Published: (2025) -
SpiralMLP: A Lightweight Vision MLP Architecture
by: Mu, Haojie, et al.
Published: (2024) -
evMLP: An Efficient Event-Driven MLP Architecture for Vision
by: Zheng, Zhentan
Published: (2025)