Spiking Variational Graph Representation Inference for Video Summarization
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Wenrui, Han, Wei, Deng, Liang-Jian, Xiong, Ruiqin, Fan, Xiaopeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpikeMba: Multi-Modal Spiking Saliency Mamba for Temporal Video Grounding
by: Li, Wenrui, et al.
Published: (2024)
by: Li, Wenrui, et al.
Published: (2024)
Language-Guided Graph Representation Learning for Video Summarization
by: Li, Wenrui, et al.
Published: (2025)
by: Li, Wenrui, et al.
Published: (2025)
Spiking Tucker Fusion Transformer for Audio-Visual Zero-Shot Learning
by: Li, Wenrui, et al.
Published: (2024)
by: Li, Wenrui, et al.
Published: (2024)
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning
by: Li, Wenrui, et al.
Published: (2025)
by: Li, Wenrui, et al.
Published: (2025)
PAGCNet: A Pose-Aware and Geometry Constrained Framework for Panoramic Depth Estimation
by: Ning, Kanglin, et al.
Published: (2025)
by: Ning, Kanglin, et al.
Published: (2025)
VideoSAGE: Video Summarization with Graph Representation Learning
by: Chaves, Jose M. Rojas, et al.
Published: (2024)
by: Chaves, Jose M. Rojas, et al.
Published: (2024)
Spatio-Temporal Distortion Aware Omnidirectional Video Super-Resolution
by: An, Hongyu, et al.
Published: (2024)
by: An, Hongyu, et al.
Published: (2024)
Hyperbolic-constraint Point Cloud Reconstruction from Single RGB-D Images
by: Li, Wenrui, et al.
Published: (2024)
by: Li, Wenrui, et al.
Published: (2024)
Spatial Degradation-Aware and Temporal Consistent Diffusion Model for Compressed Video Super-Resolution
by: An, Hongyu, et al.
Published: (2025)
by: An, Hongyu, et al.
Published: (2025)
SFOD: Spiking Fusion Object Detector
by: Fan, Yimeng, et al.
Published: (2024)
by: Fan, Yimeng, et al.
Published: (2024)
T-GVC: Trajectory-Guided Generative Video Coding at Ultra-Low Bitrates
by: Wang, Zhitao, et al.
Published: (2025)
by: Wang, Zhitao, et al.
Published: (2025)
All-in-One Video Restoration under Smoothly Evolving Unknown Weather Degradations
by: Li, Wenrui, et al.
Published: (2026)
by: Li, Wenrui, et al.
Published: (2026)
SpikeCV: Open a Continuous Computer Vision Era
by: Zheng, Yajing, et al.
Published: (2023)
by: Zheng, Yajing, et al.
Published: (2023)
Riemann-based Multi-scale Attention Reasoning Network for Text-3D Retrieval
by: Li, Wenrui, et al.
Published: (2024)
by: Li, Wenrui, et al.
Published: (2024)
Noise Conditional Variational Score Distillation
by: Peng, Xinyu, et al.
Published: (2025)
by: Peng, Xinyu, et al.
Published: (2025)
Hyperbolic Hierarchical Alignment Reasoning Network for Text-3D Retrieval
by: Li, Wenrui, et al.
Published: (2025)
by: Li, Wenrui, et al.
Published: (2025)
Towards Accurate Single Panoramic 3D Detection: A Semantic Gaussian Centric Approach
by: Ning, Kanglin, et al.
Published: (2026)
by: Ning, Kanglin, et al.
Published: (2026)
Digging into Intrinsic Contextual Information for High-fidelity 3D Point Cloud Completion
by: Chu, Jisheng, et al.
Published: (2024)
by: Chu, Jisheng, et al.
Published: (2024)
RoamScene3D: Immersive Text-to-3D Scene Generation via Adaptive Object-aware Roaming
by: Chu, Jisheng, et al.
Published: (2026)
by: Chu, Jisheng, et al.
Published: (2026)
LiftImage3D: Lifting Any Single Image to 3D Gaussians with Video Generation Priors
by: Chen, Yabo, et al.
Published: (2024)
by: Chen, Yabo, et al.
Published: (2024)
Error-Propagation-Free Learned Video Compression With Dual-Domain Progressive Temporal Alignment
by: Li, Han, et al.
Published: (2025)
by: Li, Han, et al.
Published: (2025)
MRT: Learning Compact Representations with Mixed RWKV-Transformer for Extreme Image Compression
by: Liu, Han, et al.
Published: (2025)
by: Liu, Han, et al.
Published: (2025)
METEOR: Multi-Encoder Collaborative Token Pruning for Efficient Vision Language Models
by: Liu, Yuchen, et al.
Published: (2025)
by: Liu, Yuchen, et al.
Published: (2025)
Hyperbolic Distillation: Geometry-Guided Cross-Modal Transfer for Robust 3D Object Detection
by: Ning, Kanglin, et al.
Published: (2026)
by: Ning, Kanglin, et al.
Published: (2026)
SpikeDerain: Unveiling Clear Videos from Rainy Sequences Using Color Spike Streams
by: Liang, Hanwen, et al.
Published: (2025)
by: Liang, Hanwen, et al.
Published: (2025)
Temporal-adaptive Weight Quantization for Spiking Neural Networks
by: Zhang, Han, et al.
Published: (2025)
by: Zhang, Han, et al.
Published: (2025)
Towards Holistic Modeling for Video Frame Interpolation with Auto-regressive Diffusion Transformers
by: Peng, Xinyu, et al.
Published: (2026)
by: Peng, Xinyu, et al.
Published: (2026)
SpikePoint: An Efficient Point-based Spiking Neural Network for Event Cameras Action Recognition
by: Ren, Hongwei, et al.
Published: (2023)
by: Ren, Hongwei, et al.
Published: (2023)
On Exploring PDE Modeling for Point Cloud Video Representation Learning
by: Huang, Zhuoxu, et al.
Published: (2024)
by: Huang, Zhuoxu, et al.
Published: (2024)
Point Cloud Denoising With Fine-Granularity Dynamic Graph Convolutional Networks
by: Xu, Wenqiang, et al.
Published: (2024)
by: Xu, Wenqiang, et al.
Published: (2024)
Video Summarization with Large Language Models
by: Lee, Min Jung, et al.
Published: (2025)
by: Lee, Min Jung, et al.
Published: (2025)
GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking
by: Cheng, Zixu, et al.
Published: (2026)
by: Cheng, Zixu, et al.
Published: (2026)
RouteWinFormer: A Route-Window Transformer for Middle-range Attention in Image Restoration
by: Li, Qifan, et al.
Published: (2025)
by: Li, Qifan, et al.
Published: (2025)
Video Summarization using Denoising Diffusion Probabilistic Model
by: Shang, Zirui, et al.
Published: (2024)
by: Shang, Zirui, et al.
Published: (2024)
SAEN-BGS: Energy-Efficient Spiking AutoEncoder Network for Background Subtraction
by: Zhang, Zhixuan, et al.
Published: (2025)
by: Zhang, Zhixuan, et al.
Published: (2025)
Toward Real-Time Surgical Scene Segmentation via a Spike-Driven Video Transformer with Spike-Informed Pretraining
by: Zou, Shihao, et al.
Published: (2025)
by: Zou, Shihao, et al.
Published: (2025)
SceneDreamer360: Text-Driven 3D-Consistent Scene Generation with Panoramic Gaussian Splatting
by: Li, Wenrui, et al.
Published: (2024)
by: Li, Wenrui, et al.
Published: (2024)
Dense Video Captioning using Graph-based Sentence Summarization
by: Zhang, Zhiwang, et al.
Published: (2025)
by: Zhang, Zhiwang, et al.
Published: (2025)
Language-guided Recursive Spatiotemporal Graph Modeling for Video Summarization
by: Park, Jungin, et al.
Published: (2025)
by: Park, Jungin, et al.
Published: (2025)
Diffusion-Driven Progressive Target Manipulation for Source-Free Domain Adaptation
by: Huang, Yuyang, et al.
Published: (2025)
by: Huang, Yuyang, et al.
Published: (2025)
Similar Items
-
SpikeMba: Multi-Modal Spiking Saliency Mamba for Temporal Video Grounding
by: Li, Wenrui, et al.
Published: (2024) -
Language-Guided Graph Representation Learning for Video Summarization
by: Li, Wenrui, et al.
Published: (2025) -
Spiking Tucker Fusion Transformer for Audio-Visual Zero-Shot Learning
by: Li, Wenrui, et al.
Published: (2024) -
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning
by: Li, Wenrui, et al.
Published: (2025) -
PAGCNet: A Pose-Aware and Geometry Constrained Framework for Panoramic Depth Estimation
by: Ning, Kanglin, et al.
Published: (2025)