SpikeGen: Decoupled "Rods and Cones" Visual Representation Processing with Latent Generative Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Dai, Gaole, Dong, Menghang, Zhang, Rongyu, An, Ruichuan, Zhang, Shanghang, Huang, Tiejun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpikeNVS: Enhancing Novel View Synthesis from Blurry Images via Spike Camera
by: Dai, Gaole, et al.
Published: (2024)
by: Dai, Gaole, et al.
Published: (2024)
SpikePingpong: Spike Vision-based Fast-Slow Pingpong Robot System
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
MoSA: Mixture of Sparse Adapters for Visual Efficient Tuning
by: Zhang, Qizhe, et al.
Published: (2023)
by: Zhang, Qizhe, et al.
Published: (2023)
Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain
by: Luo, Yulin, et al.
Published: (2025)
by: Luo, Yulin, et al.
Published: (2025)
Unsupervised Spike Depth Estimation via Cross-modality Cross-domain Knowledge Transfer
by: Liu, Jiaming, et al.
Published: (2022)
by: Liu, Jiaming, et al.
Published: (2022)
MoVE-KD: Knowledge Distillation for VLMs with Mixture of Visual Encoders
by: Cao, Jiajun, et al.
Published: (2025)
by: Cao, Jiajun, et al.
Published: (2025)
M$^{2}$Chat: Empowering VLM for Multimodal LLM Interleaved Text-Image Generation
by: Chi, Xiaowei, et al.
Published: (2023)
by: Chi, Xiaowei, et al.
Published: (2023)
Decomposing the Neurons: Activation Sparsity via Mixture of Experts for Continual Test Time Adaptation
by: Zhang, Rongyu, et al.
Published: (2024)
by: Zhang, Rongyu, et al.
Published: (2024)
Rethinking High-speed Image Reconstruction Framework with Spike Camera
by: Chen, Kang, et al.
Published: (2025)
by: Chen, Kang, et al.
Published: (2025)
Draw-and-Understand: Leveraging Visual Prompts to Enable MLLMs to Comprehend What You Want
by: Lin, Weifeng, et al.
Published: (2024)
by: Lin, Weifeng, et al.
Published: (2024)
SpikeStereoNet: A Brain-Inspired Framework for Stereo Depth Estimation from Spike Streams
by: Gao, Zhuoheng, et al.
Published: (2025)
by: Gao, Zhuoheng, et al.
Published: (2025)
Training-free Regional Prompting for Diffusion Transformers
by: Chen, Anthony, et al.
Published: (2024)
by: Chen, Anthony, et al.
Published: (2024)
SpikeReveal: Unlocking Temporal Sequences from Real Blurry Inputs with Spike Streams
by: Chen, Kang, et al.
Published: (2024)
by: Chen, Kang, et al.
Published: (2024)
SpikeGS: 3D Gaussian Splatting from Spike Streams with High-Speed Camera Motion
by: Zhang, Jiyuan, et al.
Published: (2024)
by: Zhang, Jiyuan, et al.
Published: (2024)
Driving in Spikes: An Entropy-Guided Object Detector for Spike Cameras
by: Liu, Ziyan, et al.
Published: (2025)
by: Liu, Ziyan, et al.
Published: (2025)
UniCTokens: Boosting Personalized Understanding and Generation via Unified Concept Tokens
by: An, Ruichuan, et al.
Published: (2025)
by: An, Ruichuan, et al.
Published: (2025)
LLM as Dataset Analyst: Subpopulation Structure Discovery with Large Language Model
by: Luo, Yulin, et al.
Published: (2024)
by: Luo, Yulin, et al.
Published: (2024)
SCSim: A Realistic Spike Cameras Simulator
by: Hu, Liwen, et al.
Published: (2024)
by: Hu, Liwen, et al.
Published: (2024)
SpikeGrasp: A Benchmark for 6-DoF Grasp Pose Detection from Stereo Spike Streams
by: Gao, Zhuoheng, et al.
Published: (2025)
by: Gao, Zhuoheng, et al.
Published: (2025)
Spike Imaging Velocimetry: Dense Motion Estimation of Fluids Using Spike Cameras
by: Zhang, Yunzhong, et al.
Published: (2025)
by: Zhang, Yunzhong, et al.
Published: (2025)
SpikeMM: Flexi-Magnification of High-Speed Micro-Motions
by: Zhang, Baoyue, et al.
Published: (2024)
by: Zhang, Baoyue, et al.
Published: (2024)
SpikeTrack: A Spike-driven Framework for Efficient Visual Tracking
by: Zhang, Qiuyang, et al.
Published: (2026)
by: Zhang, Qiuyang, et al.
Published: (2026)
Towards Low-latency Event-based Visual Recognition with Hybrid Step-wise Distillation Spiking Neural Networks
by: Zhong, Xian, et al.
Published: (2024)
by: Zhong, Xian, et al.
Published: (2024)
Proactive Gradient Conflict Mitigation in Multi-Task Learning: A Sparse Training Perspective
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
Learning from Different Samples: A Source-free Framework for Semi-supervised Domain Adaptation
by: Huang, Xinyang, et al.
Published: (2024)
by: Huang, Xinyang, et al.
Published: (2024)
SpikeDerain: Unveiling Clear Videos from Rainy Sequences Using Color Spike Streams
by: Liang, Hanwen, et al.
Published: (2025)
by: Liang, Hanwen, et al.
Published: (2025)
Spike-NeRF: Neural Radiance Field Based On Spike Camera
by: Guo, Yijia, et al.
Published: (2024)
by: Guo, Yijia, et al.
Published: (2024)
SPKLIP: Aligning Spike Video Streams with Natural Language
by: Gao, Yongchang, et al.
Published: (2025)
by: Gao, Yongchang, et al.
Published: (2025)
MoASE++: Mixture of Activation Sparsity Experts with Domain-Adaptive On-policy Distillation for Continual Test Time Adaptation
by: Zhang, Ronyu, et al.
Published: (2026)
by: Zhang, Ronyu, et al.
Published: (2026)
BEVUDA++: Geometric-aware Unsupervised Domain Adaptation for Multi-View 3D Object Detection
by: Zhang, Rongyu, et al.
Published: (2025)
by: Zhang, Rongyu, et al.
Published: (2025)
BEVUDA: Multi-geometric Space Alignments for Domain Adaptive BEV 3D Object Detection
by: Liu, Jiaming, et al.
Published: (2022)
by: Liu, Jiaming, et al.
Published: (2022)
SpikeCV: Open a Continuous Computer Vision Era
by: Zheng, Yajing, et al.
Published: (2023)
by: Zheng, Yajing, et al.
Published: (2023)
Agent Skills Should Go Beyond Text: The Case for Visual Skills
by: Xu, Binxiao, et al.
Published: (2026)
by: Xu, Binxiao, et al.
Published: (2026)
ThinkGen: Generalized Thinking for Visual Generation
by: Jiao, Siyu, et al.
Published: (2025)
by: Jiao, Siyu, et al.
Published: (2025)
Learning to Robustly Reconstruct Low-light Dynamic Scenes from Spike Streams
by: Hu, Liwen, et al.
Published: (2024)
by: Hu, Liwen, et al.
Published: (2024)
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning
by: Li, Wenrui, et al.
Published: (2025)
by: Li, Wenrui, et al.
Published: (2025)
Exploring Image Representation with Decoupled Classical Visual Descriptors
by: Qu, Chenyuan, et al.
Published: (2025)
by: Qu, Chenyuan, et al.
Published: (2025)
Towards High-performance Spiking Transformers from ANN to SNN Conversion
by: Huang, Zihan, et al.
Published: (2025)
by: Huang, Zihan, et al.
Published: (2025)
Brain-Inspired Multimodal Spiking Neural Network for Image-Text Retrieval
by: Zong, Xintao, et al.
Published: (2026)
by: Zong, Xintao, et al.
Published: (2026)
Hybrid Latent Reasoning with Decoupled Policy Optimization
by: Cheng, Tao, et al.
Published: (2026)
by: Cheng, Tao, et al.
Published: (2026)
Similar Items
-
SpikeNVS: Enhancing Novel View Synthesis from Blurry Images via Spike Camera
by: Dai, Gaole, et al.
Published: (2024) -
SpikePingpong: Spike Vision-based Fast-Slow Pingpong Robot System
by: Wang, Hao, et al.
Published: (2025) -
MoSA: Mixture of Sparse Adapters for Visual Efficient Tuning
by: Zhang, Qizhe, et al.
Published: (2023) -
Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain
by: Luo, Yulin, et al.
Published: (2025) -
Unsupervised Spike Depth Estimation via Cross-modality Cross-domain Knowledge Transfer
by: Liu, Jiaming, et al.
Published: (2022)