RAPTOR: Real-Time High-Resolution UAV Video Prediction with Efficient Video Attention
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Zhan, Guo, Zile, Zhu, Enze, Zhang, Peirong, Liu, Xiaoxuan, Wang, Lei, Zhang, Yidan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MotionScape: A Large-Scale Real-World Highly Dynamic UAV Video Dataset for World Models
von: Guo, Zile, et al.
Veröffentlicht: (2026)
von: Guo, Zile, et al.
Veröffentlicht: (2026)
UNetMamba: An Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
von: Zhu, Enze, et al.
Veröffentlicht: (2024)
von: Zhu, Enze, et al.
Veröffentlicht: (2024)
RealViformer: Investigating Attention for Real-World Video Super-Resolution
von: Zhang, Yuehan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuehan, et al.
Veröffentlicht: (2024)
DOVE: Efficient One-Step Diffusion Model for Real-World Video Super-Resolution
von: Chen, Zheng, et al.
Veröffentlicht: (2025)
von: Chen, Zheng, et al.
Veröffentlicht: (2025)
FlashVideo: Flowing Fidelity to Detail for Efficient High-Resolution Video Generation
von: Zhang, Shilong, et al.
Veröffentlicht: (2025)
von: Zhang, Shilong, et al.
Veröffentlicht: (2025)
UltraGen: High-Resolution Video Generation with Hierarchical Attention
von: Hu, Teng, et al.
Veröffentlicht: (2025)
von: Hu, Teng, et al.
Veröffentlicht: (2025)
MonarchRT: Efficient Attention for Real-Time Video Generation
von: Agarwal, Krish, et al.
Veröffentlicht: (2026)
von: Agarwal, Krish, et al.
Veröffentlicht: (2026)
PixelWizard: Towards Efficient High-Fidelity Video Generation at Ultra-Large Spatial Resolution
von: Li, Wenxue, et al.
Veröffentlicht: (2026)
von: Li, Wenxue, et al.
Veröffentlicht: (2026)
DiffST: Spatiotemporal-Aware Diffusion for Real-World Space-Time Video Super-Resolution
von: Chen, Zheng, et al.
Veröffentlicht: (2026)
von: Chen, Zheng, et al.
Veröffentlicht: (2026)
QuantVSR: Low-Bit Post-Training Quantization for Real-World Video Super-Resolution
von: Chai, Bowen, et al.
Veröffentlicht: (2025)
von: Chai, Bowen, et al.
Veröffentlicht: (2025)
Input-Aware Sparse Attention for Real-Time Co-Speech Video Generation
von: Lu, Beijia, et al.
Veröffentlicht: (2025)
von: Lu, Beijia, et al.
Veröffentlicht: (2025)
STAR: Spatial-Temporal Augmentation with Text-to-Video Models for Real-World Video Super-Resolution
von: Xie, Rui, et al.
Veröffentlicht: (2025)
von: Xie, Rui, et al.
Veröffentlicht: (2025)
MuseTalk: Real-Time High-Fidelity Video Dubbing via Spatio-Temporal Sampling
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
von: Zhang, Yue, et al.
Veröffentlicht: (2024)
TMP: Temporal Motion Propagation for Online Video Super-Resolution
von: Zhang, Zhengqiang, et al.
Veröffentlicht: (2023)
von: Zhang, Zhengqiang, et al.
Veröffentlicht: (2023)
AlignVid: Training-Free Attention Scaling for Semantic Fidelity in Text-Guided Image-to-Video Generation
von: Liu, Yexin, et al.
Veröffentlicht: (2025)
von: Liu, Yexin, et al.
Veröffentlicht: (2025)
FlashVSR: Towards Real-Time Diffusion-Based Streaming Video Super-Resolution
von: Zhuang, Junhao, et al.
Veröffentlicht: (2025)
von: Zhuang, Junhao, et al.
Veröffentlicht: (2025)
Improved Adversarial Diffusion Compression for Real-World Video Super-Resolution
von: Chen, Bin, et al.
Veröffentlicht: (2026)
von: Chen, Bin, et al.
Veröffentlicht: (2026)
Towards Redundancy Reduction in Diffusion Models for Efficient Video Super-Resolution
von: Guo, Jinpei, et al.
Veröffentlicht: (2025)
von: Guo, Jinpei, et al.
Veröffentlicht: (2025)
Flash-VStream: Efficient Real-Time Understanding for Long Video Streams
von: Zhang, Haoji, et al.
Veröffentlicht: (2025)
von: Zhang, Haoji, et al.
Veröffentlicht: (2025)
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs
von: Xu, Sicheng, et al.
Veröffentlicht: (2026)
von: Xu, Sicheng, et al.
Veröffentlicht: (2026)
STDAN: Deformable Attention Network for Space-Time Video Super-Resolution
von: Wang, Hai, et al.
Veröffentlicht: (2022)
von: Wang, Hai, et al.
Veröffentlicht: (2022)
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers
von: Xie, Enze, et al.
Veröffentlicht: (2024)
von: Xie, Enze, et al.
Veröffentlicht: (2024)
Bidirectional Sparse Attention for Faster Video Diffusion Training
von: Zhan, Chenlu, et al.
Veröffentlicht: (2025)
von: Zhan, Chenlu, et al.
Veröffentlicht: (2025)
DC-VideoGen: Efficient Video Generation with Deep Compression Video Autoencoder
von: Chen, Junyu, et al.
Veröffentlicht: (2025)
von: Chen, Junyu, et al.
Veröffentlicht: (2025)
LSGQuant: Layer-Sensitivity Guided Quantization for One-Step Diffusion Real-World Video Super-Resolution
von: Wu, Tianxing, et al.
Veröffentlicht: (2026)
von: Wu, Tianxing, et al.
Veröffentlicht: (2026)
Real-Time Vehicle Detection and Urban Traffic Behavior Analysis Based on UAV Traffic Videos on Mobile Devices
von: Zhu, Yuan, et al.
Veröffentlicht: (2024)
von: Zhu, Yuan, et al.
Veröffentlicht: (2024)
SwiftI2V: Efficient High-Resolution Image-to-Video Generation via Conditional Segment-wise Generation
von: Liu, YaoYang, et al.
Veröffentlicht: (2026)
von: Liu, YaoYang, et al.
Veröffentlicht: (2026)
InstaVSR: Taming Diffusion for Efficient and Temporally Consistent Video Super-Resolution
von: Hu, Jintong, et al.
Veröffentlicht: (2026)
von: Hu, Jintong, et al.
Veröffentlicht: (2026)
VideoGrain: Modulating Space-Time Attention for Multi-grained Video Editing
von: Yang, Xiangpeng, et al.
Veröffentlicht: (2025)
von: Yang, Xiangpeng, et al.
Veröffentlicht: (2025)
Video Super-Resolution Transformer with Masked Inter&Intra-Frame Attention
von: Zhou, Xingyu, et al.
Veröffentlicht: (2024)
von: Zhou, Xingyu, et al.
Veröffentlicht: (2024)
Video Generation with Predictive Latents
von: Zhao, Yian, et al.
Veröffentlicht: (2026)
von: Zhao, Yian, et al.
Veröffentlicht: (2026)
HiStream: Efficient High-Resolution Video Generation via Redundancy-Eliminated Streaming
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
von: Qiu, Haonan, et al.
Veröffentlicht: (2025)
Real-Time Motion-Controllable Autoregressive Video Diffusion
von: Zhao, Kesen, et al.
Veröffentlicht: (2025)
von: Zhao, Kesen, et al.
Veröffentlicht: (2025)
Efficient-vDiT: Efficient Video Diffusion Transformers With Attention Tile
von: Ding, Hangliang, et al.
Veröffentlicht: (2025)
von: Ding, Hangliang, et al.
Veröffentlicht: (2025)
Video-T1: Test-Time Scaling for Video Generation
von: Liu, Fangfu, et al.
Veröffentlicht: (2025)
von: Liu, Fangfu, et al.
Veröffentlicht: (2025)
Turbo2K: Towards Ultra-Efficient and High-Quality 2K Video Synthesis
von: Ren, Jingjing, et al.
Veröffentlicht: (2025)
von: Ren, Jingjing, et al.
Veröffentlicht: (2025)
RealisVSR: Detail-enhanced Diffusion for Real-World 4K Video Super-Resolution
von: Zhao, Weisong, et al.
Veröffentlicht: (2025)
von: Zhao, Weisong, et al.
Veröffentlicht: (2025)
Fast Video Generation with Sliding Tile Attention
von: Zhang, Peiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Peiyuan, et al.
Veröffentlicht: (2025)
GeoViS: Geospatially Rewarded Visual Search for Remote Sensing Visual Grounding
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
von: Zhang, Peirong, et al.
Veröffentlicht: (2025)
FCVSR: A Frequency-aware Method for Compressed Video Super-Resolution
von: Zhu, Qiang, et al.
Veröffentlicht: (2025)
von: Zhu, Qiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MotionScape: A Large-Scale Real-World Highly Dynamic UAV Video Dataset for World Models
von: Guo, Zile, et al.
Veröffentlicht: (2026) -
UNetMamba: An Efficient UNet-Like Mamba for Semantic Segmentation of High-Resolution Remote Sensing Images
von: Zhu, Enze, et al.
Veröffentlicht: (2024) -
RealViformer: Investigating Attention for Real-World Video Super-Resolution
von: Zhang, Yuehan, et al.
Veröffentlicht: (2024) -
DOVE: Efficient One-Step Diffusion Model for Real-World Video Super-Resolution
von: Chen, Zheng, et al.
Veröffentlicht: (2025) -
FlashVideo: Flowing Fidelity to Detail for Efficient High-Resolution Video Generation
von: Zhang, Shilong, et al.
Veröffentlicht: (2025)