Looking Backward: Streaming Video-to-Video Translation with Feature Banks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Feng, Kodaira, Akio, Xu, Chenfeng, Tomizuka, Masayoshi, Keutzer, Kurt, Marculescu, Diana |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Immiscible Diffusion: Accelerating Diffusion Training with Noise Assignment
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
FlowVid: Taming Imperfect Optical Flows for Consistent Video-to-Video Synthesis
von: Liang, Feng, et al.
Veröffentlicht: (2023)
von: Liang, Feng, et al.
Veröffentlicht: (2023)
Improved Immiscible Diffusion: Accelerate Diffusion Training by Reducing Its Miscibility
von: Li, Yiheng, et al.
Veröffentlicht: (2025)
von: Li, Yiheng, et al.
Veröffentlicht: (2025)
Vision-Language Models Learn Super Images for Efficient Partially Relevant Video Retrieval
von: Nishimura, Taichi, et al.
Veröffentlicht: (2023)
von: Nishimura, Taichi, et al.
Veröffentlicht: (2023)
StreamDiffusion: A Pipeline-level Solution for Real-time Interactive Generation
von: Kodaira, Akio, et al.
Veröffentlicht: (2023)
von: Kodaira, Akio, et al.
Veröffentlicht: (2023)
SSNVC: Single Stream Neural Video Compression with Implicit Temporal Information
von: Wang, Feng, et al.
Veröffentlicht: (2024)
von: Wang, Feng, et al.
Veröffentlicht: (2024)
A Lesson in Splats: Teacher-Guided Diffusion for 3D Gaussian Splats Generation with 2D Supervision
von: Peng, Chensheng, et al.
Veröffentlicht: (2024)
von: Peng, Chensheng, et al.
Veröffentlicht: (2024)
Viewport Prediction for Volumetric Video Streaming by Exploring Video Saliency and Trajectory Information
von: Li, Jie, et al.
Veröffentlicht: (2023)
von: Li, Jie, et al.
Veröffentlicht: (2023)
Adaptive 3D Gaussian Splatting Video Streaming
von: Gong, Han, et al.
Veröffentlicht: (2025)
von: Gong, Han, et al.
Veröffentlicht: (2025)
StreamingEval: A Unified Evaluation Protocol towards Realistic Streaming Video Understanding
von: Tang, Guowei, et al.
Veröffentlicht: (2026)
von: Tang, Guowei, et al.
Veröffentlicht: (2026)
StreamDiT: Real-Time Streaming Text-to-Video Generation
von: Kodaira, Akio, et al.
Veröffentlicht: (2025)
von: Kodaira, Akio, et al.
Veröffentlicht: (2025)
SwinGS: Sliding Window Gaussian Splatting for Volumetric Video Streaming with Arbitrary Length
von: Liu, Bangya, et al.
Veröffentlicht: (2024)
von: Liu, Bangya, et al.
Veröffentlicht: (2024)
High-Quality Live Video Streaming via Transcoding Time Prediction and Preset Selection
von: Shahre-Babak, Zahra Nabizadeh, et al.
Veröffentlicht: (2023)
von: Shahre-Babak, Zahra Nabizadeh, et al.
Veröffentlicht: (2023)
StreamDiffusionV2: A Streaming System for Dynamic and Interactive Video Generation
von: Feng, Tianrui, et al.
Veröffentlicht: (2025)
von: Feng, Tianrui, et al.
Veröffentlicht: (2025)
Joint Flow And Feature Refinement Using Attention For Video Restoration
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
Accelerated Event-Based Feature Detection and Compression for Surveillance Video Systems
von: Freeman, Andrew C., et al.
Veröffentlicht: (2023)
von: Freeman, Andrew C., et al.
Veröffentlicht: (2023)
GMFVAD: Using Grained Multi-modal Feature to Improve Video Anomaly Detection
von: Dai, Guangyu, et al.
Veröffentlicht: (2025)
von: Dai, Guangyu, et al.
Veröffentlicht: (2025)
Compression of 3D Gaussian Splatting with Optimized Feature Planes and Standard Video Codecs
von: Lee, Soonbin, et al.
Veröffentlicht: (2025)
von: Lee, Soonbin, et al.
Veröffentlicht: (2025)
In-Loop Filtering Using Learned Look-Up Tables for Video Coding
von: Li, Zhuoyuan, et al.
Veröffentlicht: (2025)
von: Li, Zhuoyuan, et al.
Veröffentlicht: (2025)
VideoZeroBench: Probing the Limits of Video MLLMs with Spatio-Temporal Evidence Verification
von: Meng, Jiahao, et al.
Veröffentlicht: (2026)
von: Meng, Jiahao, et al.
Veröffentlicht: (2026)
NeR-SC: Adapting Neural Video Representation to Screen Content
von: Shi, Ruohan, et al.
Veröffentlicht: (2026)
von: Shi, Ruohan, et al.
Veröffentlicht: (2026)
Consistency-aware Fake Videos Detection on Short Video Platforms
von: Wang, Junxi, et al.
Veröffentlicht: (2025)
von: Wang, Junxi, et al.
Veröffentlicht: (2025)
When Video Coding Meets Multimodal Large Language Models: A Unified Paradigm for Video Coding
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
von: Zhang, Pingping, et al.
Veröffentlicht: (2024)
Hallo-Live: Real-Time Streaming Joint Audio-Video Avatar Generation with Asynchronous Dual-Stream and Human-Centric Preference Distillation
von: Li, Chunyu, et al.
Veröffentlicht: (2026)
von: Li, Chunyu, et al.
Veröffentlicht: (2026)
Zero-shot Video Moment Retrieval via Off-the-shelf Multimodal Large Language Models
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
Video Compression Meets Video Generation: Latent Inter-Frame Pruning with Attention Recovery
von: Menn, Dennis, et al.
Veröffentlicht: (2026)
von: Menn, Dennis, et al.
Veröffentlicht: (2026)
Rate-aware Compression for NeRF-based Volumetric Video
von: Zhang, Zhiyu, et al.
Veröffentlicht: (2024)
von: Zhang, Zhiyu, et al.
Veröffentlicht: (2024)
Generative Frame Sampler for Long Video Understanding
von: Yao, Linli, et al.
Veröffentlicht: (2025)
von: Yao, Linli, et al.
Veröffentlicht: (2025)
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering
von: Meng, Yiran, et al.
Veröffentlicht: (2025)
von: Meng, Yiran, et al.
Veröffentlicht: (2025)
WVSC: Wireless Video Semantic Communication with Multi-frame Compensation
von: Xie, Bingyan, et al.
Veröffentlicht: (2025)
von: Xie, Bingyan, et al.
Veröffentlicht: (2025)
A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming
von: Zhou, Pengyuan, et al.
Veröffentlicht: (2024)
von: Zhou, Pengyuan, et al.
Veröffentlicht: (2024)
MMBench-Video: A Long-Form Multi-Shot Benchmark for Holistic Video Understanding
von: Fang, Xinyu, et al.
Veröffentlicht: (2024)
von: Fang, Xinyu, et al.
Veröffentlicht: (2024)
VideoMem: Constructing, Analyzing, Predicting Short-term and Long-term Video Memorability
von: Cohendet, Romain, et al.
Veröffentlicht: (2018)
von: Cohendet, Romain, et al.
Veröffentlicht: (2018)
VCEval: Rethinking What is a Good Educational Video and How to Automatically Evaluate It
von: Zhu, Xiaoxuan, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaoxuan, et al.
Veröffentlicht: (2024)
A Human-Annotated Video Dataset for Training and Evaluation of 360-Degree Video Summarization Methods
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2024)
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2024)
ChatVTG: Video Temporal Grounding via Chat with Video Dialogue Large Language Models
von: Qu, Mengxue, et al.
Veröffentlicht: (2024)
von: Qu, Mengxue, et al.
Veröffentlicht: (2024)
KeyVideoLLM: Towards Large-scale Video Keyframe Selection
von: Liang, Hao, et al.
Veröffentlicht: (2024)
von: Liang, Hao, et al.
Veröffentlicht: (2024)
Hierarchical Action Recognition: A Contrastive Video-Language Approach with Hierarchical Interactions
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
Unbiased Video Scene Graph Generation via Visual and Semantic Dual Debiasing
von: Li, Yanjun, et al.
Veröffentlicht: (2025)
von: Li, Yanjun, et al.
Veröffentlicht: (2025)
Generalized Video Anomaly Event Detection: Systematic Taxonomy and Comparison of Deep Models
von: Liu, Yang, et al.
Veröffentlicht: (2023)
von: Liu, Yang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Immiscible Diffusion: Accelerating Diffusion Training with Noise Assignment
von: Li, Yiheng, et al.
Veröffentlicht: (2024) -
FlowVid: Taming Imperfect Optical Flows for Consistent Video-to-Video Synthesis
von: Liang, Feng, et al.
Veröffentlicht: (2023) -
Improved Immiscible Diffusion: Accelerate Diffusion Training by Reducing Its Miscibility
von: Li, Yiheng, et al.
Veröffentlicht: (2025) -
Vision-Language Models Learn Super Images for Efficient Partially Relevant Video Retrieval
von: Nishimura, Taichi, et al.
Veröffentlicht: (2023) -
StreamDiffusion: A Pipeline-level Solution for Real-time Interactive Generation
von: Kodaira, Akio, et al.
Veröffentlicht: (2023)