FASTer: Focal Token Acquiring-and-Scaling Transformer for Long-term 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Dang, Chenxu, Duan, Zaipeng, An, Pei, Zhang, Xinmin, Hu, Xuzhong, Ma, Jie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dual-Domain Homogeneous Fusion with Cross-Modal Mamba and Progressive Decoder for 3D Object Detection
by: Hu, Xuzhong, et al.
Published: (2025)
by: Hu, Xuzhong, et al.
Published: (2025)
SDGOCC: Semantic and Depth-Guided Bird's-Eye View Transformation for 3D Multimodal Occupancy Prediction
by: Duan, Zaipeng, et al.
Published: (2025)
by: Duan, Zaipeng, et al.
Published: (2025)
Multimodal Point Cloud Semantic Segmentation With Virtual Point Enhancement
by: Duan, Zaipeng, et al.
Published: (2025)
by: Duan, Zaipeng, et al.
Published: (2025)
FASTer: Toward Efficient Autoregressive Vision Language Action Modeling via Neural Action Tokenization
by: Liu, Yicheng, et al.
Published: (2025)
by: Liu, Yicheng, et al.
Published: (2025)
SFMNet: Sparse Focal Modulation for 3D Object Detection
by: Shrout, Oren, et al.
Published: (2025)
by: Shrout, Oren, et al.
Published: (2025)
PVTransformer: Point-to-Voxel Transformer for Scalable 3D Object Detection
by: Leng, Zhaoqi, et al.
Published: (2024)
by: Leng, Zhaoqi, et al.
Published: (2024)
Learning Monocular Depth from Focus with Event Focal Stack
by: Jiang, Chenxu, et al.
Published: (2024)
by: Jiang, Chenxu, et al.
Published: (2024)
Approaching Outside: Scaling Unsupervised 3D Object Detection from 2D Scene
by: Zhang, Ruiyang, et al.
Published: (2024)
by: Zhang, Ruiyang, et al.
Published: (2024)
FocalPose++: Focal Length and Object Pose Estimation via Render and Compare
by: Cífka, Martin, et al.
Published: (2023)
by: Cífka, Martin, et al.
Published: (2023)
SparseWorld: A Flexible, Adaptive, and Efficient 4D Occupancy World Model Powered by Sparse and Dynamic Queries
by: Dang, Chenxu, et al.
Published: (2025)
by: Dang, Chenxu, et al.
Published: (2025)
Scenes as Tokens: Multi-Scale Normal Distributions Transform Tokenizer for General 3D Vision-Language Understanding
by: Tang, Yutao, et al.
Published: (2025)
by: Tang, Yutao, et al.
Published: (2025)
Adaptive Dual Uncertainty Optimization: Boosting Monocular 3D Object Detection under Test-Time Shifts
by: Hu, Zixuan, et al.
Published: (2025)
by: Hu, Zixuan, et al.
Published: (2025)
GFT: Gradient Focal Transformer
by: Kriuk, Boris, et al.
Published: (2025)
by: Kriuk, Boris, et al.
Published: (2025)
Visual Prototype Conditioned Focal Region Generation for UAV-Based Object Detection
by: Li, Wenhao, et al.
Published: (2026)
by: Li, Wenhao, et al.
Published: (2026)
Long-term Pre-training for Temporal Action Detection with Transformers
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
Hierarchical Graph Interaction Transformer with Dynamic Token Clustering for Camouflaged Object Detection
by: Yao, Siyuan, et al.
Published: (2024)
by: Yao, Siyuan, et al.
Published: (2024)
Temporal Structure Matters for Efficient Test-Time Adaptation in Wearable Human Activity Recognition
by: Zhou, Zishu, et al.
Published: (2026)
by: Zhou, Zishu, et al.
Published: (2026)
Every Dataset Counts: Scaling up Monocular 3D Object Detection with Joint Datasets Training
by: Ma, Fulong, et al.
Published: (2023)
by: Ma, Fulong, et al.
Published: (2023)
Cubify Anything: Scaling Indoor 3D Object Detection
by: Lazarow, Justin, et al.
Published: (2024)
by: Lazarow, Justin, et al.
Published: (2024)
TokenMotion: Motion-Guided Vision Transformer for Video Camouflaged Object Detection Via Learnable Token Selection
by: Yu, Zifan, et al.
Published: (2023)
by: Yu, Zifan, et al.
Published: (2023)
FocalOrder: Focal Preference Optimization for Reading Order Detection
by: Liu, Fuyuan, et al.
Published: (2026)
by: Liu, Fuyuan, et al.
Published: (2026)
SpikeSMOKE: Spiking Neural Networks for Monocular 3D Object Detection with Cross-Scale Gated Coding
by: Chen, Xuemei, et al.
Published: (2025)
by: Chen, Xuemei, et al.
Published: (2025)
Gaussian-Det: Learning Closed-Surface Gaussians for 3D Object Detection
by: Yan, Hongru, et al.
Published: (2024)
by: Yan, Hongru, et al.
Published: (2024)
DAOcc: 3D Object Detection Assisted Multi-Sensor Fusion for 3D Occupancy Prediction
by: Yang, Zhen, et al.
Published: (2024)
by: Yang, Zhen, et al.
Published: (2024)
MonoVQD: Monocular 3D Object Detection with Variational Query Denoising and Self-Distillation
by: Vu, Kiet Dang, et al.
Published: (2025)
by: Vu, Kiet Dang, et al.
Published: (2025)
MonoCD: Monocular 3D Object Detection with Complementary Depths
by: Yan, Longfei, et al.
Published: (2024)
by: Yan, Longfei, et al.
Published: (2024)
Efficient Multi-View 3D Object Detection by Dynamic Token Selection and Fine-Tuning
by: Nazir, Danish, et al.
Published: (2026)
by: Nazir, Danish, et al.
Published: (2026)
Long-RVOS: A Comprehensive Benchmark for Long-term Referring Video Object Segmentation
by: Liang, Tianming, et al.
Published: (2025)
by: Liang, Tianming, et al.
Published: (2025)
ObjFiller3D: Scaling 3D Object Inpainting to Dense Multi-View Consistency
by: Feng, Haitang, et al.
Published: (2025)
by: Feng, Haitang, et al.
Published: (2025)
Hierarchical Neural Collapse Detection Transformer for Class Incremental Object Detection
by: Pham, Duc Thanh, et al.
Published: (2025)
by: Pham, Duc Thanh, et al.
Published: (2025)
StereoDETR: Stereo-based Transformer for 3D Object Detection
by: Mu, Shiyi, et al.
Published: (2025)
by: Mu, Shiyi, et al.
Published: (2025)
LLM-Assisted Semantic Guidance for Sparsely Annotated Remote Sensing Object Detection
by: Liao, Wei, et al.
Published: (2025)
by: Liao, Wei, et al.
Published: (2025)
Investigating Long-term Training for Remote Sensing Object Detection
by: Park, JongHyun, et al.
Published: (2024)
by: Park, JongHyun, et al.
Published: (2024)
Rectify the Regression Bias in Long-Tailed Object Detection
by: Zhu, Ke, et al.
Published: (2024)
by: Zhu, Ke, et al.
Published: (2024)
FQ-PETR: Fully Quantized Position Embedding Transformation for Multi-View 3D Object Detection
by: Yu, Jiangyong, et al.
Published: (2025)
by: Yu, Jiangyong, et al.
Published: (2025)
FUN: A Focal U-Net Combining Reconstruction and Object Detection for Snapshot Spectral Imaging
by: Gao, Dahua, et al.
Published: (2026)
by: Gao, Dahua, et al.
Published: (2026)
Mono3DV: Monocular 3D Object Detection with 3D-Aware Bipartite Matching and Variational Query DeNoising
by: Vu, Kiet Dang, et al.
Published: (2026)
by: Vu, Kiet Dang, et al.
Published: (2026)
Cross-Layer Feature Pyramid Transformer for Small Object Detection in Aerial Images
by: Du, Zewen, et al.
Published: (2024)
by: Du, Zewen, et al.
Published: (2024)
STENet: Superpixel Token Enhancing Network for RGB-D Salient Object Detection
by: Chen, Jianlin, et al.
Published: (2026)
by: Chen, Jianlin, et al.
Published: (2026)
Domain Generalization of 3D Object Detection by Density-Resampling
by: Li, Shuangzhi, et al.
Published: (2023)
by: Li, Shuangzhi, et al.
Published: (2023)
Similar Items
-
Dual-Domain Homogeneous Fusion with Cross-Modal Mamba and Progressive Decoder for 3D Object Detection
by: Hu, Xuzhong, et al.
Published: (2025) -
SDGOCC: Semantic and Depth-Guided Bird's-Eye View Transformation for 3D Multimodal Occupancy Prediction
by: Duan, Zaipeng, et al.
Published: (2025) -
Multimodal Point Cloud Semantic Segmentation With Virtual Point Enhancement
by: Duan, Zaipeng, et al.
Published: (2025) -
FASTer: Toward Efficient Autoregressive Vision Language Action Modeling via Neural Action Tokenization
by: Liu, Yicheng, et al.
Published: (2025) -
SFMNet: Sparse Focal Modulation for 3D Object Detection
by: Shrout, Oren, et al.
Published: (2025)