PSTTS: A Plug-and-Play Token Selector for Efficient Event-based Spatio-temporal Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Xiangmo, Yang, Nan, Wang, Yang, Liu, Zhanwen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Focus Through Motion: RGB-Event Collaborative Token Sparsification for Efficient Object Detection
by: Yang, Nan, et al.
Published: (2025)
by: Yang, Nan, et al.
Published: (2025)
SMamba: Sparse Mamba for Event-based Object Detection
by: Yang, Nan, et al.
Published: (2025)
by: Yang, Nan, et al.
Published: (2025)
Enhancing Traffic Object Detection in Variable Illumination with RGB-Event Fusion
by: Liu, Zhanwen, et al.
Published: (2023)
by: Liu, Zhanwen, et al.
Published: (2023)
Beyond conventional vision: RGB-event fusion for robust object detection in dynamic traffic scenarios
by: Liu, Zhanwen, et al.
Published: (2025)
by: Liu, Zhanwen, et al.
Published: (2025)
Boosting Visual Recognition in Real-world Degradations via Unsupervised Feature Enhancement Module with Deep Channel Prior
by: Liu, Zhanwen, et al.
Published: (2024)
by: Liu, Zhanwen, et al.
Published: (2024)
Multi-scale Temporal Fusion Transformer for Incomplete Vehicle Trajectory Prediction
by: Liu, Zhanwen, et al.
Published: (2024)
by: Liu, Zhanwen, et al.
Published: (2024)
MSTF: Multiscale Transformer for Incomplete Trajectory Prediction
by: Liu, Zhanwen, et al.
Published: (2024)
by: Liu, Zhanwen, et al.
Published: (2024)
DPMambaIR: All-in-One Image Restoration via Degradation-Aware Prompt State Space Model
by: Liu, Zhanwen, et al.
Published: (2025)
by: Liu, Zhanwen, et al.
Published: (2025)
Bidirectional Image-Event Guided Fusion Framework for Low-Light Image Enhancement
by: Liu, Zhanwen, et al.
Published: (2025)
by: Liu, Zhanwen, et al.
Published: (2025)
Vote&Mix: Plug-and-Play Token Reduction for Efficient Vision Transformer
by: Peng, Shuai, et al.
Published: (2024)
by: Peng, Shuai, et al.
Published: (2024)
FastDriveVLA: Efficient End-to-End Driving via Plug-and-Play Reconstruction-based Token Pruning
by: Cao, Jiajun, et al.
Published: (2025)
by: Cao, Jiajun, et al.
Published: (2025)
A Plug-and-Play Learning-based IMU Bias Factor for Robust Visual-Inertial Odometry
by: Yi, Yang, et al.
Published: (2025)
by: Yi, Yang, et al.
Published: (2025)
Text-to-Image Rectified Flow as Plug-and-Play Priors
by: Yang, Xiaofeng, et al.
Published: (2024)
by: Yang, Xiaofeng, et al.
Published: (2024)
Plug-and-Play Context Feature Reuse for Efficient Masked Generation
by: Liu, Xuejie, et al.
Published: (2025)
by: Liu, Xuejie, et al.
Published: (2025)
Plug and Play Active Learning for Object Detection
by: Yang, Chenhongyi, et al.
Published: (2022)
by: Yang, Chenhongyi, et al.
Published: (2022)
VisionSelector: End-to-End Learnable Visual Token Compression for Efficient Multimodal LLMs
by: Zhu, Jiaying, et al.
Published: (2025)
by: Zhu, Jiaying, et al.
Published: (2025)
FasterVAR: Plug-and-Play Acceleration for Visual Autoregressive Models
by: Li, Senmao, et al.
Published: (2025)
by: Li, Senmao, et al.
Published: (2025)
Adaptive Part Learning for Fine-Grained Generalized Category Discovery: A Plug-and-Play Enhancement
by: Dai, Qiyuan, et al.
Published: (2025)
by: Dai, Qiyuan, et al.
Published: (2025)
Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation
by: Zhu, Lunjie, et al.
Published: (2026)
by: Zhu, Lunjie, et al.
Published: (2026)
Plug-and-Play Diffusion Distillation
by: Hsiao, Yi-Ting, et al.
Published: (2024)
by: Hsiao, Yi-Ting, et al.
Published: (2024)
CleanStyle: Plug-and-Play Style Conditioning Purification for Text-to-Image Stylization
by: Feng, Xiaoman, et al.
Published: (2026)
by: Feng, Xiaoman, et al.
Published: (2026)
Segmentation as A Plug-and-Play Capability for Frozen Multimodal LLMs
by: Liu, Jiazhen, et al.
Published: (2025)
by: Liu, Jiazhen, et al.
Published: (2025)
STAC: Plug-and-Play Spatio-Temporal Aware Cache Compression for Streaming 3D Reconstruction
by: Wang, Runze, et al.
Published: (2026)
by: Wang, Runze, et al.
Published: (2026)
PAT-VCM: Plug-and-Play Auxiliary Tokens for Video Coding for Machines
by: Jiang, Wei, et al.
Published: (2026)
by: Jiang, Wei, et al.
Published: (2026)
Diffusion Sampling Path Tells More: An Efficient Plug-and-Play Strategy for Sample Filtering
by: Wang, Sixian, et al.
Published: (2025)
by: Wang, Sixian, et al.
Published: (2025)
Plug-and-Play Logit Fusion for Heterogeneous Pathology Foundation Models
by: Huang, Gexin, et al.
Published: (2026)
by: Huang, Gexin, et al.
Published: (2026)
Spatio-temporal Transformers for Action Unit Classification with Event Cameras
by: Cultrera, Luca, et al.
Published: (2024)
by: Cultrera, Luca, et al.
Published: (2024)
ARM: A Learnable, Plug-and-Play Module for CLIP-based Open-vocabulary Semantic Segmentation
by: Liu, Ziquan, et al.
Published: (2025)
by: Liu, Ziquan, et al.
Published: (2025)
Spatio-temporal Sign Language Representation and Translation
by: Hamidullah, Yasser, et al.
Published: (2025)
by: Hamidullah, Yasser, et al.
Published: (2025)
Plug-and-Play 1.x-Bit KV Cache Quantization for Video Large Language Models
by: Tao, Keda, et al.
Published: (2025)
by: Tao, Keda, et al.
Published: (2025)
Stand-In: A Lightweight and Plug-and-Play Identity Control for Video Generation
by: Xue, Bowen, et al.
Published: (2025)
by: Xue, Bowen, et al.
Published: (2025)
SIS-Challenge: Event-based Spatio-temporal Instance Segmentation Challenge at the CVPR 2025 Event-based Vision Workshop
by: Hamann, Friedhelm, et al.
Published: (2025)
by: Hamann, Friedhelm, et al.
Published: (2025)
Plug-and-Play Grounding of Reasoning in Multimodal Large Language Models
by: Chen, Jiaxing, et al.
Published: (2024)
by: Chen, Jiaxing, et al.
Published: (2024)
Plug-and-Play Versatile Compressed Video Enhancement
by: Zeng, Huimin, et al.
Published: (2025)
by: Zeng, Huimin, et al.
Published: (2025)
PhysDepth: Plug-and-Play Physical Refinement for Monocular Depth Estimation in Challenging Environments
by: Peng, Kebin, et al.
Published: (2024)
by: Peng, Kebin, et al.
Published: (2024)
Prune2Drive: A Plug-and-Play Framework for Accelerating Vision-Language Models in Autonomous Driving
by: Xiong, Minhao, et al.
Published: (2025)
by: Xiong, Minhao, et al.
Published: (2025)
RL-Selector: Reinforcement Learning-Guided Data Selection via Redundancy Assessment
by: Yang, Suorong, et al.
Published: (2025)
by: Yang, Suorong, et al.
Published: (2025)
Poppy: Polarization-based Plug-and-Play Guidance for Enhancing Monocular Normal Estimation
by: Kim, Irene, et al.
Published: (2026)
by: Kim, Irene, et al.
Published: (2026)
NFCDS: A Plug-and-Play Noise Frequency-Controlled Diffusion Sampling Strategy for Image Restoration
by: Wang, Zhen, et al.
Published: (2026)
by: Wang, Zhen, et al.
Published: (2026)
CBNet: A Plug-and-Play Network for Segmentation-Based Scene Text Detection
by: Zhao, Xi, et al.
Published: (2022)
by: Zhao, Xi, et al.
Published: (2022)
Similar Items
-
Focus Through Motion: RGB-Event Collaborative Token Sparsification for Efficient Object Detection
by: Yang, Nan, et al.
Published: (2025) -
SMamba: Sparse Mamba for Event-based Object Detection
by: Yang, Nan, et al.
Published: (2025) -
Enhancing Traffic Object Detection in Variable Illumination with RGB-Event Fusion
by: Liu, Zhanwen, et al.
Published: (2023) -
Beyond conventional vision: RGB-event fusion for robust object detection in dynamic traffic scenarios
by: Liu, Zhanwen, et al.
Published: (2025) -
Boosting Visual Recognition in Real-world Degradations via Unsupervised Feature Enhancement Module with Deep Channel Prior
by: Liu, Zhanwen, et al.
Published: (2024)