Shot-Aware Frame Sampling for Video Understanding
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhao, Mengyu, Fu, Di, Xie, Yongyu, Zhang, Jiaxing, Yuan, Zhigang, Jalali, Shirin, Cao, Yong |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Untrained Neural Nets for Snapshot Compressive Imaging: Theory and Algorithms
par: Zhao, Mengyu, et autres
Publié: (2024)
par: Zhao, Mengyu, et autres
Publié: (2024)
Motion-Aware Video Frame Interpolation
par: Han, Pengfei, et autres
Publié: (2024)
par: Han, Pengfei, et autres
Publié: (2024)
Chain-of-Frames: Advancing Video Understanding in Multimodal LLMs via Frame-Aware Reasoning
par: Ghazanfari, Sara, et autres
Publié: (2025)
par: Ghazanfari, Sara, et autres
Publié: (2025)
Monte Carlo Maximum Likelihood Reconstruction for Digital Holography with Speckle
par: Chen, Xi, et autres
Publié: (2026)
par: Chen, Xi, et autres
Publié: (2026)
Shot Segmentation Based on Von Neumann Entropy for Key Frame Extraction
par: Zhang, Xueqing, et autres
Publié: (2024)
par: Zhang, Xueqing, et autres
Publié: (2024)
RB-FT: Rationale-Bootstrapped Fine-Tuning for Video Classification
par: Xu, Meilong, et autres
Publié: (2025)
par: Xu, Meilong, et autres
Publié: (2025)
GIFT: Global Irreplaceability Frame Targeting for Efficient Video Understanding
par: Ma, Junpeng, et autres
Publié: (2026)
par: Ma, Junpeng, et autres
Publié: (2026)
Frame by Familiar Frame: Understanding Replication in Video Diffusion Models
par: Rahman, Aimon, et autres
Publié: (2024)
par: Rahman, Aimon, et autres
Publié: (2024)
Zero-Shot Long-Form Video Understanding through Screenplay
par: Wu, Yongliang, et autres
Publié: (2024)
par: Wu, Yongliang, et autres
Publié: (2024)
Event-Anchored Frame Selection for Effective Long-Video Understanding
par: Chen, Wang, et autres
Publié: (2026)
par: Chen, Wang, et autres
Publié: (2026)
PMQ-VE: Progressive Multi-Frame Quantization for Video Enhancement
par: Feng, ZhanFeng, et autres
Publié: (2025)
par: Feng, ZhanFeng, et autres
Publié: (2025)
Wavelet-based Frame Selection by Detecting Semantic Boundary for Long Video Understanding
par: Chen, Wang, et autres
Publié: (2026)
par: Chen, Wang, et autres
Publié: (2026)
Progress-Aware Video Frame Captioning
par: Xue, Zihui, et autres
Publié: (2024)
par: Xue, Zihui, et autres
Publié: (2024)
Enhancing Video Inpainting with Aligned Frame Interval Guidance
par: Xie, Ming, et autres
Publié: (2025)
par: Xie, Ming, et autres
Publié: (2025)
SRVAU-R1: Enhancing Video Anomaly Understanding via Reflection-Aware Learning
par: Zhao, Zihao, et autres
Publié: (2026)
par: Zhao, Zihao, et autres
Publié: (2026)
Object-Aware Video Matting with Cross-Frame Guidance
par: Zhang, Huayu, et autres
Publié: (2025)
par: Zhang, Huayu, et autres
Publié: (2025)
Think-Clip-Sample: Slow-Fast Frame Selection for Video Understanding
par: Tan, Wenhui, et autres
Publié: (2026)
par: Tan, Wenhui, et autres
Publié: (2026)
Generative Frame Sampler for Long Video Understanding
par: Yao, Linli, et autres
Publié: (2025)
par: Yao, Linli, et autres
Publié: (2025)
Enabling DBSCAN for Very Large-Scale High-Dimensional Spaces
par: Wang, Yongyu
Publié: (2024)
par: Wang, Yongyu
Publié: (2024)
Adversarial-Robustness-Guided Graph Pruning
par: Wang, Yongyu
Publié: (2024)
par: Wang, Yongyu
Publié: (2024)
Continual Text-to-Video Retrieval with Frame Fusion and Task-Aware Routing
par: Zhao, Zecheng, et autres
Publié: (2025)
par: Zhao, Zecheng, et autres
Publié: (2025)
Incentivizing Temporal-Awareness in Egocentric Video Understanding Models
par: Xu, Zhiyang, et autres
Publié: (2026)
par: Xu, Zhiyang, et autres
Publié: (2026)
Frames2Residual: Spatiotemporal Decoupling for Self-Supervised Video Denoising
par: Ji, Mingjie, et autres
Publié: (2026)
par: Ji, Mingjie, et autres
Publié: (2026)
Improving LLM Video Understanding with 16 Frames Per Second
par: Li, Yixuan, et autres
Publié: (2025)
par: Li, Yixuan, et autres
Publié: (2025)
Zero-Shot Video Restoration and Enhancement with Assistance of Video Diffusion Models
par: Cao, Cong, et autres
Publié: (2026)
par: Cao, Cong, et autres
Publié: (2026)
Prompt-aware of Frame Sampling for Efficient Text-Video Retrieval
par: Zhang, Deyu, et autres
Publié: (2025)
par: Zhang, Deyu, et autres
Publié: (2025)
KFS-Bench: Comprehensive Evaluation of Key Frame Sampling in Long Video Understanding
par: Li, Zongyao, et autres
Publié: (2025)
par: Li, Zongyao, et autres
Publié: (2025)
DreaMontage: Arbitrary Frame-Guided One-Shot Video Generation
par: Liu, Jiawei, et autres
Publié: (2025)
par: Liu, Jiawei, et autres
Publié: (2025)
Zero-Shot Video Translation and Editing with Frame Spatial-Temporal Correspondence
par: Yang, Shuai, et autres
Publié: (2025)
par: Yang, Shuai, et autres
Publié: (2025)
Active Sampling for Ultra-Low-Bit-Rate Video Compression via Conditional Controlled Diffusion
par: Javadi, Amirhosein, et autres
Publié: (2026)
par: Javadi, Amirhosein, et autres
Publié: (2026)
iMOVE: Instance-Motion-Aware Video Understanding
par: Li, Jiaze, et autres
Publié: (2025)
par: Li, Jiaze, et autres
Publié: (2025)
DynImg: Key Frames with Visual Prompts are Good Representation for Multi-Modal Video Understanding
par: Bao, Xiaoyi, et autres
Publié: (2025)
par: Bao, Xiaoyi, et autres
Publié: (2025)
Few-Step Diffusion Sampling Through Instance-Aware Discretizations
par: Yuan, Liangyu, et autres
Publié: (2026)
par: Yuan, Liangyu, et autres
Publié: (2026)
VGDFR: Diffusion-based Video Generation with Dynamic Latent Frame Rate
par: Yuan, Zhihang, et autres
Publié: (2025)
par: Yuan, Zhihang, et autres
Publié: (2025)
Motion-Aware Generative Frame Interpolation
par: Zhang, Guozhen, et autres
Publié: (2025)
par: Zhang, Guozhen, et autres
Publié: (2025)
EDEN: Enhanced Diffusion for High-quality Large-motion Video Frame Interpolation
par: Zhang, Zihao, et autres
Publié: (2025)
par: Zhang, Zihao, et autres
Publié: (2025)
Video-MTR: Reinforced Multi-Turn Reasoning for Long Video Understanding
par: Xie, Yuan, et autres
Publié: (2025)
par: Xie, Yuan, et autres
Publié: (2025)
Informative Sample Selection Model for Skeleton-based Action Recognition with Limited Training Samples
par: Tu, Zhigang, et autres
Publié: (2025)
par: Tu, Zhigang, et autres
Publié: (2025)
Resolving Task Objective Conflicts in Unified Model via Task-Aware Mixture-of-Experts
par: Zhang, Jiaxing, et autres
Publié: (2025)
par: Zhang, Jiaxing, et autres
Publié: (2025)
End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling
par: Liang, Jianxin, et autres
Publié: (2024)
par: Liang, Jianxin, et autres
Publié: (2024)
Documents similaires
-
Untrained Neural Nets for Snapshot Compressive Imaging: Theory and Algorithms
par: Zhao, Mengyu, et autres
Publié: (2024) -
Motion-Aware Video Frame Interpolation
par: Han, Pengfei, et autres
Publié: (2024) -
Chain-of-Frames: Advancing Video Understanding in Multimodal LLMs via Frame-Aware Reasoning
par: Ghazanfari, Sara, et autres
Publié: (2025) -
Monte Carlo Maximum Likelihood Reconstruction for Digital Holography with Speckle
par: Chen, Xi, et autres
Publié: (2026) -
Shot Segmentation Based on Von Neumann Entropy for Key Frame Extraction
par: Zhang, Xueqing, et autres
Publié: (2024)