Focus Through Motion: RGB-Event Collaborative Token Sparsification for Efficient Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Nan, Wang, Yang, Liu, Zhanwen, Dai, Yuchao, Liu, Yang, Zhao, Xiangmo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PSTTS: A Plug-and-Play Token Selector for Efficient Event-based Spatio-temporal Representation Learning
by: Zhao, Xiangmo, et al.
Published: (2025)
by: Zhao, Xiangmo, et al.
Published: (2025)
Enhancing Traffic Object Detection in Variable Illumination with RGB-Event Fusion
by: Liu, Zhanwen, et al.
Published: (2023)
by: Liu, Zhanwen, et al.
Published: (2023)
SMamba: Sparse Mamba for Event-based Object Detection
by: Yang, Nan, et al.
Published: (2025)
by: Yang, Nan, et al.
Published: (2025)
Beyond conventional vision: RGB-event fusion for robust object detection in dynamic traffic scenarios
by: Liu, Zhanwen, et al.
Published: (2025)
by: Liu, Zhanwen, et al.
Published: (2025)
DPMambaIR: All-in-One Image Restoration via Degradation-Aware Prompt State Space Model
by: Liu, Zhanwen, et al.
Published: (2025)
by: Liu, Zhanwen, et al.
Published: (2025)
Boosting Visual Recognition in Real-world Degradations via Unsupervised Feature Enhancement Module with Deep Channel Prior
by: Liu, Zhanwen, et al.
Published: (2024)
by: Liu, Zhanwen, et al.
Published: (2024)
Multi-scale Temporal Fusion Transformer for Incomplete Vehicle Trajectory Prediction
by: Liu, Zhanwen, et al.
Published: (2024)
by: Liu, Zhanwen, et al.
Published: (2024)
MSTF: Multiscale Transformer for Incomplete Trajectory Prediction
by: Liu, Zhanwen, et al.
Published: (2024)
by: Liu, Zhanwen, et al.
Published: (2024)
Bidirectional Image-Event Guided Fusion Framework for Low-Light Image Enhancement
by: Liu, Zhanwen, et al.
Published: (2025)
by: Liu, Zhanwen, et al.
Published: (2025)
SparseDiT: Token Sparsification for Efficient Diffusion Transformer
by: Chang, Shuning, et al.
Published: (2024)
by: Chang, Shuning, et al.
Published: (2024)
Learnable Motion-Focused Tokenization for Effective and Efficient Video Unsupervised Domain Adaptation
by: Liu, Tzu Ling, et al.
Published: (2026)
by: Liu, Tzu Ling, et al.
Published: (2026)
Geometry-Aware 3D Salient Object Detection Network
by: Wang, Chen, et al.
Published: (2025)
by: Wang, Chen, et al.
Published: (2025)
PEOD: A Pixel-Aligned Event-RGB Benchmark for Object Detection under Challenging Conditions
by: Cui, Luoping, et al.
Published: (2025)
by: Cui, Luoping, et al.
Published: (2025)
UCDNet: Multi-UAV Collaborative 3D Object Detection Network by Reliable Feature Mapping
by: Tian, Pengju, et al.
Published: (2024)
by: Tian, Pengju, et al.
Published: (2024)
Video Token Sparsification for Efficient Multimodal LLMs in Autonomous Driving
by: Ma, Yunsheng, et al.
Published: (2024)
by: Ma, Yunsheng, et al.
Published: (2024)
Detecting Every Object from Events
by: Zhang, Haitian, et al.
Published: (2024)
by: Zhang, Haitian, et al.
Published: (2024)
Event-based Graph Representation with Spatial and Motion Vectors for Asynchronous Object Detection
by: Verma, Aayush Atul, et al.
Published: (2025)
by: Verma, Aayush Atul, et al.
Published: (2025)
Removal then Selection: A Coarse-to-Fine Fusion Perspective for RGB-Infrared Object Detection
by: Zhao, Tianyi, et al.
Published: (2024)
by: Zhao, Tianyi, et al.
Published: (2024)
RGBD Objects in the Wild: Scaling Real-World 3D Object Learning from RGB-D Videos
by: Xia, Hongchi, et al.
Published: (2024)
by: Xia, Hongchi, et al.
Published: (2024)
Modality-Specific Hierarchical Enhancement for RGB-D Camouflaged Object Detection
by: Niu, Yuzhen, et al.
Published: (2026)
by: Niu, Yuzhen, et al.
Published: (2026)
A Unified Structure for Efficient RGB and RGB-D Salient Object Detection
by: Peng, Peng, et al.
Published: (2020)
by: Peng, Peng, et al.
Published: (2020)
STENet: Superpixel Token Enhancing Network for RGB-D Salient Object Detection
by: Chen, Jianlin, et al.
Published: (2026)
by: Chen, Jianlin, et al.
Published: (2026)
A Gated Cross-domain Collaborative Network for Underwater Object Detection
by: Dai, Linhui, et al.
Published: (2023)
by: Dai, Linhui, et al.
Published: (2023)
SAMSOD: Rethinking SAM Optimization for RGB-T Salient Object Detection
by: Liu, Zhengyi, et al.
Published: (2025)
by: Liu, Zhengyi, et al.
Published: (2025)
UVCPNet: A UAV-Vehicle Collaborative Perception Network for 3D Object Detection
by: Wang, Yuchao, et al.
Published: (2024)
by: Wang, Yuchao, et al.
Published: (2024)
Instance-Level Moving Object Segmentation from a Single Image with Events
by: Wan, Zhexiong, et al.
Published: (2025)
by: Wan, Zhexiong, et al.
Published: (2025)
FocusDiffuser: Perceiving Local Disparities for Camouflaged Object Detection
by: Zhao, Jianwei, et al.
Published: (2024)
by: Zhao, Jianwei, et al.
Published: (2024)
Lightweight RGB-D Salient Object Detection from a Speed-Accuracy Tradeoff Perspective
by: Duan, Songsong, et al.
Published: (2025)
by: Duan, Songsong, et al.
Published: (2025)
RGB-T Object Detection via Group Shuffled Multi-receptive Attention and Multi-modal Supervision
by: Wang, Jinzhong, et al.
Published: (2024)
by: Wang, Jinzhong, et al.
Published: (2024)
LEOD: Label-Efficient Object Detection for Event Cameras
by: Wu, Ziyi, et al.
Published: (2023)
by: Wu, Ziyi, et al.
Published: (2023)
3D Focusing-and-Matching Network for Multi-Instance Point Cloud Registration
by: Zhang, Liyuan, et al.
Published: (2024)
by: Zhang, Liyuan, et al.
Published: (2024)
SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference
by: Zhang, Yuan, et al.
Published: (2024)
by: Zhang, Yuan, et al.
Published: (2024)
ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification
by: He, Yefei, et al.
Published: (2024)
by: He, Yefei, et al.
Published: (2024)
Active Object Detection with Knowledge Aggregation and Distillation from Large Models
by: Yang, Dejie, et al.
Published: (2024)
by: Yang, Dejie, et al.
Published: (2024)
TokenMotion: Motion-Guided Vision Transformer for Video Camouflaged Object Detection Via Learnable Token Selection
by: Yu, Zifan, et al.
Published: (2023)
by: Yu, Zifan, et al.
Published: (2023)
Unleashing the Power of Motion and Depth: A Selective Fusion Strategy for RGB-D Video Salient Object Detection
by: He, Jiahao, et al.
Published: (2025)
by: He, Jiahao, et al.
Published: (2025)
Spatial Orthogonal Refinement for Robust RGB-Event Visual Object Tracking
by: Huang, Dexing, et al.
Published: (2026)
by: Huang, Dexing, et al.
Published: (2026)
Salient Object Detection in RGB-D Videos
by: Mou, Ao, et al.
Published: (2023)
by: Mou, Ao, et al.
Published: (2023)
Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion Models
by: Li, Liulei, et al.
Published: (2024)
by: Li, Liulei, et al.
Published: (2024)
METEOR: Multi-Encoder Collaborative Token Pruning for Efficient Vision Language Models
by: Liu, Yuchen, et al.
Published: (2025)
by: Liu, Yuchen, et al.
Published: (2025)
Similar Items
-
PSTTS: A Plug-and-Play Token Selector for Efficient Event-based Spatio-temporal Representation Learning
by: Zhao, Xiangmo, et al.
Published: (2025) -
Enhancing Traffic Object Detection in Variable Illumination with RGB-Event Fusion
by: Liu, Zhanwen, et al.
Published: (2023) -
SMamba: Sparse Mamba for Event-based Object Detection
by: Yang, Nan, et al.
Published: (2025) -
Beyond conventional vision: RGB-event fusion for robust object detection in dynamic traffic scenarios
by: Liu, Zhanwen, et al.
Published: (2025) -
DPMambaIR: All-in-One Image Restoration via Degradation-Aware Prompt State Space Model
by: Liu, Zhanwen, et al.
Published: (2025)