Dynamic-Aware Video Distillation: Optimizing Temporal Resolution Based on Video Semantics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yinjie, Zhao, Heng, Wen, Bihan, Ong, Yew-Soon, Zhou, Joey Tianyi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Video Set Distillation: Information Diversification and Temporal Densification
von: Zhao, Yinjie, et al.
Veröffentlicht: (2024)
von: Zhao, Yinjie, et al.
Veröffentlicht: (2024)
Agentic Spatio-Temporal Grounding via Collaborative Reasoning
von: Zhao, Heng, et al.
Veröffentlicht: (2026)
von: Zhao, Heng, et al.
Veröffentlicht: (2026)
Cognitive Inception: Agentic Reasoning against Visual Deceptions by Injecting Skepticism
von: Zhao, Yinjie, et al.
Veröffentlicht: (2025)
von: Zhao, Yinjie, et al.
Veröffentlicht: (2025)
AEGIS: Authenticity Evaluation Benchmark for AI-Generated Video Sequences
von: Li, Jieyu, et al.
Veröffentlicht: (2025)
von: Li, Jieyu, et al.
Veröffentlicht: (2025)
SiamNAS: Siamese Surrogate Model for Dominance Relation Prediction in Multi-objective Neural Architecture Search
von: Zhou, Yuyang, et al.
Veröffentlicht: (2025)
von: Zhou, Yuyang, et al.
Veröffentlicht: (2025)
SweetTok: Semantic-Aware Spatial-Temporal Tokenizer for Compact Video Discretization
von: Tan, Zhentao, et al.
Veröffentlicht: (2024)
von: Tan, Zhentao, et al.
Veröffentlicht: (2024)
Possibilistic Predictive Uncertainty for Deep Learning
von: Ni, Yao, et al.
Veröffentlicht: (2026)
von: Ni, Yao, et al.
Veröffentlicht: (2026)
VITED: Video Temporal Evidence Distillation
von: Lu, Yujie, et al.
Veröffentlicht: (2025)
von: Lu, Yujie, et al.
Veröffentlicht: (2025)
Compressed-Domain-Aware Online Video Super-Resolution
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
Unified and Dynamic Graph for Temporal Character Grouping in Long Videos
von: Shu, Xiujun, et al.
Veröffentlicht: (2023)
von: Shu, Xiujun, et al.
Veröffentlicht: (2023)
Knowing Your Target: Target-Aware Transformer Makes Better Spatio-Temporal Video Grounding
von: Gu, Xin, et al.
Veröffentlicht: (2025)
von: Gu, Xin, et al.
Veröffentlicht: (2025)
Facial Dynamics in Video: Instruction Tuning for Improved Facial Expression Perception and Contextual Awareness
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
Context-Aware Temporal Embedding of Objects in Video Data
von: Farhan, Ahnaf, et al.
Veröffentlicht: (2024)
von: Farhan, Ahnaf, et al.
Veröffentlicht: (2024)
Cluster-based Video Summarization with Temporal Context Awareness
von: Huynh-Lam, Hai-Dang, et al.
Veröffentlicht: (2024)
von: Huynh-Lam, Hai-Dang, et al.
Veröffentlicht: (2024)
Learning Local and Global Temporal Contexts for Video Semantic Segmentation
von: Sun, Guolei, et al.
Veröffentlicht: (2022)
von: Sun, Guolei, et al.
Veröffentlicht: (2022)
Adaptive Video Distillation: Mitigating Oversaturation and Temporal Collapse in Few-Step Generation
von: You, Yuyang, et al.
Veröffentlicht: (2026)
von: You, Yuyang, et al.
Veröffentlicht: (2026)
VideoStir: Understanding Long Videos via Spatio-Temporally Structured and Intent-Aware RAG
von: Fu, Honghao, et al.
Veröffentlicht: (2026)
von: Fu, Honghao, et al.
Veröffentlicht: (2026)
Temporal Aware Pruning for Efficient Diffusion-based Video Generation
von: Li, Sheng, et al.
Veröffentlicht: (2026)
von: Li, Sheng, et al.
Veröffentlicht: (2026)
Hierarchically Robust Zero-shot Vision-language Models
von: Dong, Junhao, et al.
Veröffentlicht: (2026)
von: Dong, Junhao, et al.
Veröffentlicht: (2026)
One-Step Diffusion for Detail-Rich and Temporally Consistent Video Super-Resolution
von: Sun, Yujing, et al.
Veröffentlicht: (2025)
von: Sun, Yujing, et al.
Veröffentlicht: (2025)
FADE: Frequency-Aware Diffusion Model Factorization for Video Editing
von: Zhu, Yixuan, et al.
Veröffentlicht: (2025)
von: Zhu, Yixuan, et al.
Veröffentlicht: (2025)
Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification
von: Qin, Minghao, et al.
Veröffentlicht: (2025)
von: Qin, Minghao, et al.
Veröffentlicht: (2025)
Semantic-Aware Adaptive Visual Memory for Streaming Video Understanding
von: Wu, Hang, et al.
Veröffentlicht: (2026)
von: Wu, Hang, et al.
Veröffentlicht: (2026)
Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events
von: Liu, Xiaolin, et al.
Veröffentlicht: (2026)
von: Liu, Xiaolin, et al.
Veröffentlicht: (2026)
VideoMind: A Chain-of-LoRA Agent for Temporal-Grounded Video Reasoning
von: Liu, Ye, et al.
Veröffentlicht: (2025)
von: Liu, Ye, et al.
Veröffentlicht: (2025)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
GVD: Guiding Video Diffusion Model for Scalable Video Distillation
von: Li, Kunyang, et al.
Veröffentlicht: (2025)
von: Li, Kunyang, et al.
Veröffentlicht: (2025)
VideoTemp-o3: Harmonizing Temporal Grounding and Video Understanding in Agentic Thinking-with-Videos
von: Liu, Wenqi, et al.
Veröffentlicht: (2026)
von: Liu, Wenqi, et al.
Veröffentlicht: (2026)
DraftAttention: Fast Video Diffusion via Low-Resolution Attention Guidance
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
von: Shen, Xuan, et al.
Veröffentlicht: (2025)
LSA: Localized Semantic Alignment for Enhancing Temporal Consistency in Traffic Video Generation
von: Karimov, Mirlan, et al.
Veröffentlicht: (2026)
von: Karimov, Mirlan, et al.
Veröffentlicht: (2026)
Geometry-Aware Semantic Reasoning for Training Free Video Anomaly Detection
von: Zia, Ali, et al.
Veröffentlicht: (2026)
von: Zia, Ali, et al.
Veröffentlicht: (2026)
Personalized Video Summarization by Multimodal Video Understanding
von: Chen, Brian, et al.
Veröffentlicht: (2024)
von: Chen, Brian, et al.
Veröffentlicht: (2024)
DropletVideo: A Dataset and Approach to Explore Integral Spatio-Temporal Consistent Video Generation
von: Zhang, Runze, et al.
Veröffentlicht: (2025)
von: Zhang, Runze, et al.
Veröffentlicht: (2025)
Video-As-Prompt: Unified Semantic Control for Video Generation
von: Bian, Yuxuan, et al.
Veröffentlicht: (2025)
von: Bian, Yuxuan, et al.
Veröffentlicht: (2025)
Task-Aware KV Compression For Cost-Effective Long Video Understanding
von: Qin, Minghao, et al.
Veröffentlicht: (2025)
von: Qin, Minghao, et al.
Veröffentlicht: (2025)
ESA: Energy-Based Shot Assembly Optimization for Automatic Video Editing
von: Chen, Yaosen, et al.
Veröffentlicht: (2025)
von: Chen, Yaosen, et al.
Veröffentlicht: (2025)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
von: Kumar, Yogesh, et al.
Veröffentlicht: (2025)
von: Kumar, Yogesh, et al.
Veröffentlicht: (2025)
LEMON: How Well Do MLLMs Perform Temporal Multimodal Understanding on Instructional Videos?
von: Yu, Zhuang, et al.
Veröffentlicht: (2026)
von: Yu, Zhuang, et al.
Veröffentlicht: (2026)
A Survey on Backbones for Deep Video Action Recognition
von: Tang, Zixuan, et al.
Veröffentlicht: (2024)
von: Tang, Zixuan, et al.
Veröffentlicht: (2024)
Direct Distillation between Different Domains
von: Tang, Jialiang, et al.
Veröffentlicht: (2024)
von: Tang, Jialiang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Video Set Distillation: Information Diversification and Temporal Densification
von: Zhao, Yinjie, et al.
Veröffentlicht: (2024) -
Agentic Spatio-Temporal Grounding via Collaborative Reasoning
von: Zhao, Heng, et al.
Veröffentlicht: (2026) -
Cognitive Inception: Agentic Reasoning against Visual Deceptions by Injecting Skepticism
von: Zhao, Yinjie, et al.
Veröffentlicht: (2025) -
AEGIS: Authenticity Evaluation Benchmark for AI-Generated Video Sequences
von: Li, Jieyu, et al.
Veröffentlicht: (2025) -
SiamNAS: Siamese Surrogate Model for Dominance Relation Prediction in Multi-objective Neural Architecture Search
von: Zhou, Yuyang, et al.
Veröffentlicht: (2025)