KAN-FPN-Stem:A KAN-Enhanced Feature Pyramid Stem for Boosting ViT-based Pose Estimation
Fuente:
arXiv
Guardado en:
| Autor principal: | Tang, HaoNan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hyb-KAN ViT: Hybrid Kolmogorov-Arnold Networks Augmented Vision Transformer
por: Dey, Sainath, et al.
Publicado: (2025)
por: Dey, Sainath, et al.
Publicado: (2025)
LR-FPN: Enhancing Remote Sensing Object Detection with Location Refined Feature Pyramid Network
por: Li, Hanqian, et al.
Publicado: (2024)
por: Li, Hanqian, et al.
Publicado: (2024)
FastPose-ViT: A Vision Transformer for Real-Time Spacecraft Pose Estimation
por: Ancey, Pierre, et al.
Publicado: (2025)
por: Ancey, Pierre, et al.
Publicado: (2025)
Poseidon: A ViT-based Architecture for Multi-Frame Pose Estimation with Adaptive Frame Weighting and Multi-Scale Feature Fusion
por: Pace, Cesare Davide, et al.
Publicado: (2025)
por: Pace, Cesare Davide, et al.
Publicado: (2025)
Motion Aware ViT-based Framework for Monocular 6-DoF Spacecraft Pose Estimation
por: Sosa, Jose, et al.
Publicado: (2025)
por: Sosa, Jose, et al.
Publicado: (2025)
Demystifying KAN for Vision Tasks: The RepKAN Approach
por: Cheon, Minjong
Publicado: (2026)
por: Cheon, Minjong
Publicado: (2026)
FGAA-FPN: Foreground-Guided Angle-Aware Feature Pyramid Network for Oriented Object Detection
por: Ma, Jialin
Publicado: (2026)
por: Ma, Jialin
Publicado: (2026)
Harnessing the Computation Redundancy in ViTs to Boost Adversarial Transferability
por: Liu, Jiani, et al.
Publicado: (2025)
por: Liu, Jiani, et al.
Publicado: (2025)
DiffKAN-Inpainting: KAN-based Diffusion model for brain tumor inpainting
por: Tao, Tianli, et al.
Publicado: (2025)
por: Tao, Tianli, et al.
Publicado: (2025)
KAN See Your Face
por: Han, Dong, et al.
Publicado: (2024)
por: Han, Dong, et al.
Publicado: (2024)
Semi-KAN: KAN Provides an Effective Representation for Semi-Supervised Learning in Medical Image Segmentation
por: Ye, Zanting, et al.
Publicado: (2025)
por: Ye, Zanting, et al.
Publicado: (2025)
Light-ResKAN: A Parameter-Sharing Lightweight KAN with Gram Polynomials for Efficient SAR Image Recognition
por: Yi, Pan, et al.
Publicado: (2026)
por: Yi, Pan, et al.
Publicado: (2026)
Learning to Anchor Visual Odometry: KAN-Based Pose Regression for Planetary Landing
por: Luo, Xubo, et al.
Publicado: (2025)
por: Luo, Xubo, et al.
Publicado: (2025)
UKAN-EP: Enhancing U-KAN with Efficient Attention and Pyramid Aggregation for 3D Multi-Modal MRI Brain Tumor Segmentation
por: Chen, Yanbing, et al.
Publicado: (2024)
por: Chen, Yanbing, et al.
Publicado: (2024)
KAN See In the Dark
por: Ning, Aoxiang, et al.
Publicado: (2024)
por: Ning, Aoxiang, et al.
Publicado: (2024)
KAN We Flow? Advancing Robotic Manipulation with 3D Flow Matching via KAN & RWKV
por: Chen, Zhihao, et al.
Publicado: (2026)
por: Chen, Zhihao, et al.
Publicado: (2026)
TFS-ViT: Token-Level Feature Stylization for Domain Generalization
por: Noori, Mehrdad, et al.
Publicado: (2023)
por: Noori, Mehrdad, et al.
Publicado: (2023)
Learning KAN-based Implicit Neural Representations for Deformable Image Registration
por: Drozdov, Nikita, et al.
Publicado: (2025)
por: Drozdov, Nikita, et al.
Publicado: (2025)
S$^2$-FPN: Scale-ware Strip Attention Guided Feature Pyramid Network for Real-time Semantic Segmentation
por: Elhassan, Mohammed A. M., et al.
Publicado: (2022)
por: Elhassan, Mohammed A. M., et al.
Publicado: (2022)
KAN or MLP? Point Cloud Shows the Way Forward
por: Shi, Yan, et al.
Publicado: (2025)
por: Shi, Yan, et al.
Publicado: (2025)
ViTCAE: ViT-based Class-conditioned Autoencoder
por: Jebraeeli, Vahid, et al.
Publicado: (2025)
por: Jebraeeli, Vahid, et al.
Publicado: (2025)
Deeper Inside Deep ViT
por: Hong, Sungrae
Publicado: (2025)
por: Hong, Sungrae
Publicado: (2025)
K-U-KAN: Koopman-Enhanced U-KAN for 3D Dental Reconstruction from a Single Panoramic X-ray Radiograph
por: Parida, Bikram Keshari, et al.
Publicado: (2026)
por: Parida, Bikram Keshari, et al.
Publicado: (2026)
ViT-CoMer: Vision Transformer with Convolutional Multi-scale Feature Interaction for Dense Predictions
por: Xia, Chunlong, et al.
Publicado: (2024)
por: Xia, Chunlong, et al.
Publicado: (2024)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
por: Zhong, Yunshan, et al.
Publicado: (2023)
por: Zhong, Yunshan, et al.
Publicado: (2023)
Foggy Crowd Counting: Combining Physical Priors and KAN-Graph
por: Wang, Yuhao, et al.
Publicado: (2025)
por: Wang, Yuhao, et al.
Publicado: (2025)
ABFR-KAN: Kolmogorov-Arnold Networks for Functional Brain Analysis
por: Ward, Tyler, et al.
Publicado: (2026)
por: Ward, Tyler, et al.
Publicado: (2026)
LiFT: A Surprisingly Simple Lightweight Feature Transform for Dense ViT Descriptors
por: Suri, Saksham, et al.
Publicado: (2024)
por: Suri, Saksham, et al.
Publicado: (2024)
RepViT: Revisiting Mobile CNN From ViT Perspective
por: Wang, Ao, et al.
Publicado: (2023)
por: Wang, Ao, et al.
Publicado: (2023)
SCKAN: Structural Consensus-based KAN Prototype Learning for Semi-Supervised Pancreas Segmentation
por: Liu, Yuqi, et al.
Publicado: (2026)
por: Liu, Yuqi, et al.
Publicado: (2026)
A Hybrid Framework Bridging CNN and ViT based on Theory of Evidence for Diabetic Retinopathy Grading
por: Qiu, Junlai, et al.
Publicado: (2025)
por: Qiu, Junlai, et al.
Publicado: (2025)
HS-FPN: High Frequency and Spatial Perception FPN for Tiny Object Detection
por: Shi, Zican, et al.
Publicado: (2024)
por: Shi, Zican, et al.
Publicado: (2024)
Trustworthy Longitudinal Brain MRI Completion: A Deformation-Based Approach with KAN-Enhanced Diffusion Model
por: Tao, Tianli, et al.
Publicado: (2026)
por: Tao, Tianli, et al.
Publicado: (2026)
EA-ViT: Efficient Adaptation for Elastic Vision Transformer
por: Zhu, Chen, et al.
Publicado: (2025)
por: Zhu, Chen, et al.
Publicado: (2025)
QuadKAN: KAN-Enhanced Quadruped Motion Control via End-to-End Reinforcement Learning
por: Wang, Yinuo, et al.
Publicado: (2025)
por: Wang, Yinuo, et al.
Publicado: (2025)
SpyroPose: SE(3) Pyramids for Object Pose Distribution Estimation
por: Haugaard, Rasmus Laurvig, et al.
Publicado: (2023)
por: Haugaard, Rasmus Laurvig, et al.
Publicado: (2023)
MSCGC-KAN: Multi-scale Causal Graph Convolution and Kolmogorov-Arnold Feature Mapping for EEG Emotion Recognition
por: Gong, Haoliang, et al.
Publicado: (2026)
por: Gong, Haoliang, et al.
Publicado: (2026)
MedVKAN: Efficient Feature Extraction with Mamba and KAN for Medical Image Segmentation
por: Zhu, Hancan, et al.
Publicado: (2025)
por: Zhu, Hancan, et al.
Publicado: (2025)
Dynamic Tuning Towards Parameter and Inference Efficiency for ViT Adaptation
por: Zhao, Wangbo, et al.
Publicado: (2024)
por: Zhao, Wangbo, et al.
Publicado: (2024)
MedKAN: An Advanced Kolmogorov-Arnold Network for Medical Image Classification
por: Yang, Zhuoqin, et al.
Publicado: (2025)
por: Yang, Zhuoqin, et al.
Publicado: (2025)
Ejemplares similares
-
Hyb-KAN ViT: Hybrid Kolmogorov-Arnold Networks Augmented Vision Transformer
por: Dey, Sainath, et al.
Publicado: (2025) -
LR-FPN: Enhancing Remote Sensing Object Detection with Location Refined Feature Pyramid Network
por: Li, Hanqian, et al.
Publicado: (2024) -
FastPose-ViT: A Vision Transformer for Real-Time Spacecraft Pose Estimation
por: Ancey, Pierre, et al.
Publicado: (2025) -
Poseidon: A ViT-based Architecture for Multi-Frame Pose Estimation with Adaptive Frame Weighting and Multi-Scale Feature Fusion
por: Pace, Cesare Davide, et al.
Publicado: (2025) -
Motion Aware ViT-based Framework for Monocular 6-DoF Spacecraft Pose Estimation
por: Sosa, Jose, et al.
Publicado: (2025)