V2V: Scaling Event-Based Vision through Efficient Video-to-Voxel Simulation
Fuente:
arXiv
Saved in:
| Main Authors: | Lou, Hanyue, Liang, Jinxiu, Teng, Minggui, Wang, Yi, Shi, Boxin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
E2VIDiff: Perceptual Events-to-Video Reconstruction using Diffusion Priors
by: Liang, Jinxiu, et al.
Published: (2024)
by: Liang, Jinxiu, et al.
Published: (2024)
EvTurb: Event Camera Guided Turbulence Removal
by: Liu, Yixing, et al.
Published: (2025)
by: Liu, Yixing, et al.
Published: (2025)
Reconstruction as a Bridge for Event-Based Visual Question Answering
by: Lou, Hanyue, et al.
Published: (2025)
by: Lou, Hanyue, et al.
Published: (2025)
Learning to Deblur Polarized Images
by: Zhou, Chu, et al.
Published: (2024)
by: Zhou, Chu, et al.
Published: (2024)
Language-guided Image Reflection Separation
by: Zhong, Haofeng, et al.
Published: (2024)
by: Zhong, Haofeng, et al.
Published: (2024)
PanoWan: Lifting Diffusion Video Generation Models to 360° with Latitude/Longitude-aware Mechanisms
by: Xia, Yifei, et al.
Published: (2025)
by: Xia, Yifei, et al.
Published: (2025)
V2CE: Video to Continuous Events Simulator
by: Zhang, Zhongyang, et al.
Published: (2023)
by: Zhang, Zhongyang, et al.
Published: (2023)
OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation
by: Liang, Sen, et al.
Published: (2025)
by: Liang, Sen, et al.
Published: (2025)
V-CAST: Video Curvature-Aware Spatio-Temporal Pruning for Efficient Video Large Language Models
by: Lin, Xinying, et al.
Published: (2026)
by: Lin, Xinying, et al.
Published: (2026)
InstanceV: Instance-Level Video Generation
by: Chen, Yuheng, et al.
Published: (2025)
by: Chen, Yuheng, et al.
Published: (2025)
TIP-I2V: A Million-Scale Real Text and Image Prompt Dataset for Image-to-Video Generation
by: Wang, Wenhao, et al.
Published: (2024)
by: Wang, Wenhao, et al.
Published: (2024)
T2VEval: Benchmark Dataset and Objective Evaluation Method for T2V-generated Videos
by: Qi, Zelu, et al.
Published: (2025)
by: Qi, Zelu, et al.
Published: (2025)
SweepEvGS: Event-Based 3D Gaussian Splatting for Macro and Micro Radiance Field Rendering from a Single Sweep
by: Wu, Jingqian, et al.
Published: (2024)
by: Wu, Jingqian, et al.
Published: (2024)
Event Voxel Set Transformer for Spatiotemporal Representation Learning on Event Streams
by: Xie, Bochen, et al.
Published: (2023)
by: Xie, Bochen, et al.
Published: (2023)
Complementing Event Streams and RGB Frames for Hand Mesh Reconstruction
by: Jiang, Jianping, et al.
Published: (2024)
by: Jiang, Jianping, et al.
Published: (2024)
InstructAV2AV: Instruction-Guided Audio-Video Joint Editing
by: Zheng, Haojie, et al.
Published: (2026)
by: Zheng, Haojie, et al.
Published: (2026)
EventFlash: Towards Efficient MLLMs for Event-Based Vision
by: Liu, Shaoyu, et al.
Published: (2026)
by: Liu, Shaoyu, et al.
Published: (2026)
LaSe-E2V: Towards Language-guided Semantic-Aware Event-to-Video Reconstruction
by: Chen, Kanghao, et al.
Published: (2024)
by: Chen, Kanghao, et al.
Published: (2024)
LiteVoxel: Low-memory Intelligent Thresholding for Efficient Voxel Rasterization
by: Lee, Jee Won, et al.
Published: (2025)
by: Lee, Jee Won, et al.
Published: (2025)
Dark-EvGS: Event Camera as an Eye for Radiance Field in the Dark
by: Wu, Jingqian, et al.
Published: (2025)
by: Wu, Jingqian, et al.
Published: (2025)
L-C4: Language-Based Video Colorization for Creative and Consistent Color
by: Chang, Zheng, et al.
Published: (2024)
by: Chang, Zheng, et al.
Published: (2024)
OpenEvents V1: Large-Scale Benchmark Dataset for Multimodal Event Grounding
by: Nguyen, Hieu, et al.
Published: (2025)
by: Nguyen, Hieu, et al.
Published: (2025)
NeuV-SLAM: Fast Neural Multiresolution Voxel Optimization for RGBD Dense SLAM
by: Guo, Wenzhi, et al.
Published: (2024)
by: Guo, Wenzhi, et al.
Published: (2024)
CamI2V: Camera-Controlled Image-to-Video Diffusion Model
by: Zheng, Guangcong, et al.
Published: (2024)
by: Zheng, Guangcong, et al.
Published: (2024)
Voxel-Aggregated Feature Synthesis: Efficient Dense Mapping for Simulated 3D Reasoning
by: Burns, Owen, et al.
Published: (2024)
by: Burns, Owen, et al.
Published: (2024)
StableV2V: Stablizing Shape Consistency in Video-to-Video Editing
by: Liu, Chang, et al.
Published: (2024)
by: Liu, Chang, et al.
Published: (2024)
SIMS-V: Simulated Instruction-Tuning for Spatial Video Understanding
by: Brown, Ellis, et al.
Published: (2025)
by: Brown, Ellis, et al.
Published: (2025)
ViLA: Efficient Video-Language Alignment for Video Question Answering
by: Wang, Xijun, et al.
Published: (2023)
by: Wang, Xijun, et al.
Published: (2023)
Arbitrary-Scale Point Cloud Upsampling by Voxel-Based Network with Latent Geometric-Consistent Learning
by: Du, Hang, et al.
Published: (2024)
by: Du, Hang, et al.
Published: (2024)
CyberV: Cybernetics for Test-time Scaling in Video Understanding
by: Meng, Jiahao, et al.
Published: (2025)
by: Meng, Jiahao, et al.
Published: (2025)
AVI-Edit: Audio-sync Video Instance Editing with Granularity-Aware Mask Refiner
by: Zheng, Haojie, et al.
Published: (2025)
by: Zheng, Haojie, et al.
Published: (2025)
PolarVSR: A Unified Framework and Benchmark for Continuous Space-Time Polarization Video Reconstruction
by: Li, Chenggong, et al.
Published: (2026)
by: Li, Chenggong, et al.
Published: (2026)
IE2Video: Adapting Pretrained Diffusion Models for Event-Based Video Reconstruction
by: Torbunov, Dmitrii, et al.
Published: (2025)
by: Torbunov, Dmitrii, et al.
Published: (2025)
EventSTU: Event-Guided Efficient Spatio-Temporal Understanding for Video Large Language Models
by: Xu, Wenhao, et al.
Published: (2025)
by: Xu, Wenhao, et al.
Published: (2025)
DynamicScaler: Seamless and Scalable Video Generation for Panoramic Scenes
by: Liu, Jinxiu, et al.
Published: (2024)
by: Liu, Jinxiu, et al.
Published: (2024)
G$^2$V$^2$former: Graph Guided Video Vision Transformer for Face Anti-Spoofing
by: Yang, Jingyi, et al.
Published: (2024)
by: Yang, Jingyi, et al.
Published: (2024)
StreamForest: Efficient Online Video Understanding with Persistent Event Memory
by: Zeng, Xiangyu, et al.
Published: (2025)
by: Zeng, Xiangyu, et al.
Published: (2025)
GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization
by: Pan, Yu, et al.
Published: (2026)
by: Pan, Yu, et al.
Published: (2026)
PolarVLM: Bridging the Semantic-Physical Gap in Vision-Language Models
by: Li, Yuliang, et al.
Published: (2026)
by: Li, Yuliang, et al.
Published: (2026)
Motion-I2V: Consistent and Controllable Image-to-Video Generation with Explicit Motion Modeling
by: Shi, Xiaoyu, et al.
Published: (2024)
by: Shi, Xiaoyu, et al.
Published: (2024)
Similar Items
-
E2VIDiff: Perceptual Events-to-Video Reconstruction using Diffusion Priors
by: Liang, Jinxiu, et al.
Published: (2024) -
EvTurb: Event Camera Guided Turbulence Removal
by: Liu, Yixing, et al.
Published: (2025) -
Reconstruction as a Bridge for Event-Based Visual Question Answering
by: Lou, Hanyue, et al.
Published: (2025) -
Learning to Deblur Polarized Images
by: Zhou, Chu, et al.
Published: (2024) -
Language-guided Image Reflection Separation
by: Zhong, Haofeng, et al.
Published: (2024)