SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Wenxi, Guo, Yuchen, Zheng, Jilai, Lin, Haozhe, Ma, Chao, Fang, Lu, Yang, Xiaokang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Bootstrapping SparseFormers from Vision Foundation Models
di: Gao, Ziteng, et al.
Pubblicazione: (2023)
di: Gao, Ziteng, et al.
Pubblicazione: (2023)
SaccadeDet: A Novel Dual-Stage Architecture for Rapid and Accurate Detection in Gigapixel Images
di: Li, Wenxi, et al.
Pubblicazione: (2024)
di: Li, Wenxi, et al.
Pubblicazione: (2024)
SparseOcc: Rethinking Sparse Latent Representation for Vision-Based Semantic Occupancy Prediction
di: Tang, Pin, et al.
Pubblicazione: (2024)
di: Tang, Pin, et al.
Pubblicazione: (2024)
SparseVoxFormer: Sparse Voxel-based Transformer for Multi-modal 3D Object Detection
di: Son, Hyeongseok, et al.
Pubblicazione: (2025)
di: Son, Hyeongseok, et al.
Pubblicazione: (2025)
Few-Shot Object Detection with Sparse Context Transformers
di: Mei, Jie, et al.
Pubblicazione: (2024)
di: Mei, Jie, et al.
Pubblicazione: (2024)
LiteFusion: Taming 3D Object Detectors from Vision-Based to Multi-Modal with Minimal Adaptation
di: Ren, Xiangxuan, et al.
Pubblicazione: (2025)
di: Ren, Xiangxuan, et al.
Pubblicazione: (2025)
ParFormer: A Vision Transformer with Parallel Mixer and Sparse Channel Attention Patch Embedding
di: Setyawan, Novendra, et al.
Pubblicazione: (2024)
di: Setyawan, Novendra, et al.
Pubblicazione: (2024)
HeightFormer: Learning Height Prediction in Voxel Features for Roadside Vision Centric 3D Object Detection via Transformer
di: Zhang, Zhang, et al.
Pubblicazione: (2025)
di: Zhang, Zhang, et al.
Pubblicazione: (2025)
MedFormer: Hierarchical Medical Vision Transformer with Content-Aware Dual Sparse Selection Attention
di: Xia, Zunhui, et al.
Pubblicazione: (2025)
di: Xia, Zunhui, et al.
Pubblicazione: (2025)
AF-CLIP: Zero-Shot Anomaly Detection via Anomaly-Focused CLIP Adaptation
di: Fang, Qingqing, et al.
Pubblicazione: (2025)
di: Fang, Qingqing, et al.
Pubblicazione: (2025)
Blur-aware Spatio-temporal Sparse Transformer for Video Deblurring
di: Zhang, Huicong, et al.
Pubblicazione: (2024)
di: Zhang, Huicong, et al.
Pubblicazione: (2024)
PosFormer: Recognizing Complex Handwritten Mathematical Expression with Position Forest Transformer
di: Guan, Tongkun, et al.
Pubblicazione: (2024)
di: Guan, Tongkun, et al.
Pubblicazione: (2024)
Scene Adaptive Sparse Transformer for Event-based Object Detection
di: Peng, Yansong, et al.
Pubblicazione: (2024)
di: Peng, Yansong, et al.
Pubblicazione: (2024)
Sparse Generation: Making Pseudo Labels Sparse for Point Weakly Supervised Object Detection on Low Data Volume
di: Shang, Chuyang, et al.
Pubblicazione: (2024)
di: Shang, Chuyang, et al.
Pubblicazione: (2024)
Automotive Object Detection via Learning Sparse Events by Spiking Neurons
di: Zhang, Hu, et al.
Pubblicazione: (2023)
di: Zhang, Hu, et al.
Pubblicazione: (2023)
SPG: Sparse-Projected Guides with Sparse Autoencoders for Zero-Shot Anomaly Detection
di: Nanaumi, Tomoyasu, et al.
Pubblicazione: (2026)
di: Nanaumi, Tomoyasu, et al.
Pubblicazione: (2026)
Proto-Former: Unified Facial Landmark Detection by Prototype Transformer
di: Hu, Shengkai, et al.
Pubblicazione: (2025)
di: Hu, Shengkai, et al.
Pubblicazione: (2025)
RhythmFormer: Extracting Patterned rPPG Signals based on Periodic Sparse Attention
di: Zou, Bochao, et al.
Pubblicazione: (2024)
di: Zou, Bochao, et al.
Pubblicazione: (2024)
Vision Transformer with Sparse Scan Prior
di: Zhang, Yuguang, et al.
Pubblicazione: (2024)
di: Zhang, Yuguang, et al.
Pubblicazione: (2024)
SMamba: Sparse Mamba for Event-based Object Detection
di: Yang, Nan, et al.
Pubblicazione: (2025)
di: Yang, Nan, et al.
Pubblicazione: (2025)
SparseAlign: A Fully Sparse Framework for Cooperative Object Detection
di: Yuan, Yunshuang, et al.
Pubblicazione: (2025)
di: Yuan, Yunshuang, et al.
Pubblicazione: (2025)
Interpretable and Testable Vision Features via Sparse Autoencoders
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
di: Stevens, Samuel, et al.
Pubblicazione: (2025)
SparseDet: A Simple and Effective Framework for Fully Sparse LiDAR-based 3D Object Detection
di: Liu, Lin, et al.
Pubblicazione: (2024)
di: Liu, Lin, et al.
Pubblicazione: (2024)
Sparse Semi-DETR: Sparse Learnable Queries for Semi-Supervised Object Detection
di: Shehzadi, Tahira, et al.
Pubblicazione: (2024)
di: Shehzadi, Tahira, et al.
Pubblicazione: (2024)
Fully Sparse Fusion for 3D Object Detection
di: Li, Yingyan, et al.
Pubblicazione: (2023)
di: Li, Yingyan, et al.
Pubblicazione: (2023)
Angle of Arrival Estimation with Transformer: A Sparse and Gridless Method with Zero-Shot Capability
di: Zhu, Zhaoxuan, et al.
Pubblicazione: (2024)
di: Zhu, Zhaoxuan, et al.
Pubblicazione: (2024)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
di: Kumar, Yogesh, et al.
Pubblicazione: (2025)
di: Kumar, Yogesh, et al.
Pubblicazione: (2025)
Improving Adversarial Transferability on Vision Transformers via Forward Propagation Refinement
di: Ren, Yuchen, et al.
Pubblicazione: (2025)
di: Ren, Yuchen, et al.
Pubblicazione: (2025)
SPWOOD: Sparse Partial Weakly-Supervised Oriented Object Detection
di: Zhang, Wei, et al.
Pubblicazione: (2026)
di: Zhang, Wei, et al.
Pubblicazione: (2026)
SparseDrive: End-to-End Autonomous Driving via Sparse Scene Representation
di: Sun, Wenchao, et al.
Pubblicazione: (2024)
di: Sun, Wenchao, et al.
Pubblicazione: (2024)
SigFormer: Sparse Signal-Guided Transformer for Multi-Modal Human Action Segmentation
di: Liu, Qi, et al.
Pubblicazione: (2023)
di: Liu, Qi, et al.
Pubblicazione: (2023)
S$^2$Teacher: Step-by-step Teacher for Sparsely Annotated Oriented Object Detection
di: Lin, Yu, et al.
Pubblicazione: (2025)
di: Lin, Yu, et al.
Pubblicazione: (2025)
Active-SAOOD: Active Sparsely Annotated Oriented Object Detection in Remote Sensing Images
di: Lin, Yu, et al.
Pubblicazione: (2026)
di: Lin, Yu, et al.
Pubblicazione: (2026)
SparseDFF: Sparse-View Feature Distillation for One-Shot Dexterous Manipulation
di: Wang, Qianxu, et al.
Pubblicazione: (2023)
di: Wang, Qianxu, et al.
Pubblicazione: (2023)
Corner2Net: Detecting Objects as Cascade Corners
di: Liu, Chenglong, et al.
Pubblicazione: (2024)
di: Liu, Chenglong, et al.
Pubblicazione: (2024)
SP3D: Boosting Sparsely-Supervised 3D Object Detection via Accurate Cross-Modal Semantic Prompts
di: Zhao, Shijia, et al.
Pubblicazione: (2025)
di: Zhao, Shijia, et al.
Pubblicazione: (2025)
Point2Insert: Video Object Insertion via Sparse Point Guidance
di: Zhou, Yu, et al.
Pubblicazione: (2026)
di: Zhou, Yu, et al.
Pubblicazione: (2026)
Learning Class Prototypes for Unified Sparse Supervised 3D Object Detection
di: Zhu, Yun, et al.
Pubblicazione: (2025)
di: Zhu, Yun, et al.
Pubblicazione: (2025)
SNN-Driven Multimodal Human Action Recognition via Sparse Spatial-Temporal Data Fusion
di: Zheng, Naichuan, et al.
Pubblicazione: (2025)
di: Zheng, Naichuan, et al.
Pubblicazione: (2025)
SparseLIF: High-Performance Sparse LiDAR-Camera Fusion for 3D Object Detection
di: Zhang, Hongcheng, et al.
Pubblicazione: (2024)
di: Zhang, Hongcheng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Bootstrapping SparseFormers from Vision Foundation Models
di: Gao, Ziteng, et al.
Pubblicazione: (2023) -
SaccadeDet: A Novel Dual-Stage Architecture for Rapid and Accurate Detection in Gigapixel Images
di: Li, Wenxi, et al.
Pubblicazione: (2024) -
SparseOcc: Rethinking Sparse Latent Representation for Vision-Based Semantic Occupancy Prediction
di: Tang, Pin, et al.
Pubblicazione: (2024) -
SparseVoxFormer: Sparse Voxel-based Transformer for Multi-modal 3D Object Detection
di: Son, Hyeongseok, et al.
Pubblicazione: (2025) -
Few-Shot Object Detection with Sparse Context Transformers
di: Mei, Jie, et al.
Pubblicazione: (2024)