Gespeichert in:
| Hauptverfasser: | Wu, Jingchao, Kang, Zejian, Liu, Haibo, Fei, Yuanchen, Huang, Xiangru |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2512.11321 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models
von: Zheng, Kai, et al.
Veröffentlicht: (2026)
von: Zheng, Kai, et al.
Veröffentlicht: (2026)
SemanticFace: Semantic Facial Action Estimation via Semantic Distillation in Interpretable Space
von: Kang, Zejian, et al.
Veröffentlicht: (2026)
von: Kang, Zejian, et al.
Veröffentlicht: (2026)
SuperFace: Preference-Aligned Facial Expression Estimation Beyond Pseudo Supervision
von: Kang, Zejian, et al.
Veröffentlicht: (2026)
von: Kang, Zejian, et al.
Veröffentlicht: (2026)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
von: Lyu, Tianle, et al.
Veröffentlicht: (2025)
von: Lyu, Tianle, et al.
Veröffentlicht: (2025)
KeyVID: Keyframe-Aware Video Diffusion for Audio-Synchronized Visual Animation
von: Wang, Xingrui, et al.
Veröffentlicht: (2025)
von: Wang, Xingrui, et al.
Veröffentlicht: (2025)
Placing Human Animations into 3D Scenes by Learning Interaction- and Geometry-Driven Keyframes
von: Mullen Jr, James F., et al.
Veröffentlicht: (2022)
von: Mullen Jr, James F., et al.
Veröffentlicht: (2022)
Exploring the Role of Synthetic Data Augmentation in Controllable Human-Centric Video Generation
von: Fei, Yuanchen, et al.
Veröffentlicht: (2026)
von: Fei, Yuanchen, et al.
Veröffentlicht: (2026)
Occlusion-Aware Physics-Semantic Keyframe Selection for Robust Video Editing
von: Liu, Lin, et al.
Veröffentlicht: (2026)
von: Liu, Lin, et al.
Veröffentlicht: (2026)
DanceCamAnimator: Keyframe-Based Controllable 3D Dance Camera Synthesis
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
Learning Semantic Facial Descriptors for Accurate Face Animation
von: Zhu, Lei, et al.
Veröffentlicht: (2025)
von: Zhu, Lei, et al.
Veröffentlicht: (2025)
Coarse-to-Fine 3D Keyframe Transporter
von: Zhu, Xupeng, et al.
Veröffentlicht: (2025)
von: Zhu, Xupeng, et al.
Veröffentlicht: (2025)
Controllable Human-centric Keyframe Interpolation with Generative Prior
von: Guo, Zujin, et al.
Veröffentlicht: (2025)
von: Guo, Zujin, et al.
Veröffentlicht: (2025)
The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection
von: He, Qingdong, et al.
Veröffentlicht: (2025)
von: He, Qingdong, et al.
Veröffentlicht: (2025)
CineVerse: Consistent Keyframe Synthesis for Cinematic Scene Composition
von: Phung, Quynh, et al.
Veröffentlicht: (2025)
von: Phung, Quynh, et al.
Veröffentlicht: (2025)
Threading Keyframe with Narratives: MLLMs as Strong Long Video Comprehenders
von: Fang, Bo, et al.
Veröffentlicht: (2025)
von: Fang, Bo, et al.
Veröffentlicht: (2025)
KFFocus: Highlighting Keyframes for Enhanced Video Understanding
von: Nie, Ming, et al.
Veröffentlicht: (2025)
von: Nie, Ming, et al.
Veröffentlicht: (2025)
Agentic Keyframe Search for Video Question Answering
von: Fan, Sunqi, et al.
Veröffentlicht: (2025)
von: Fan, Sunqi, et al.
Veröffentlicht: (2025)
SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models
von: He, Ziheng, et al.
Veröffentlicht: (2026)
von: He, Ziheng, et al.
Veröffentlicht: (2026)
Generative Motion Infilling From Imprecisely Timed Keyframes
von: Goel, Purvi, et al.
Veröffentlicht: (2025)
von: Goel, Purvi, et al.
Veröffentlicht: (2025)
KS-APR: Keyframe Selection for Robust Absolute Pose Regression
von: Liu, Changkun, et al.
Veröffentlicht: (2023)
von: Liu, Changkun, et al.
Veröffentlicht: (2023)
Keyframe-Based Feed-Forward Visual Odometry
von: Dai, Weichen, et al.
Veröffentlicht: (2026)
von: Dai, Weichen, et al.
Veröffentlicht: (2026)
Large Model based Sequential Keyframe Extraction for Video Summarization
von: Tan, Kailong, et al.
Veröffentlicht: (2024)
von: Tan, Kailong, et al.
Veröffentlicht: (2024)
Less is More: Improving Motion Diffusion Models with Sparse Keyframes
von: Bae, Jinseok, et al.
Veröffentlicht: (2025)
von: Bae, Jinseok, et al.
Veröffentlicht: (2025)
SignSparK: Efficient Multilingual Sign Language Production via Sparse Keyframe Learning
von: Low, Jianhe, et al.
Veröffentlicht: (2026)
von: Low, Jianhe, et al.
Veröffentlicht: (2026)
Compact Keyframe-Optimized Multi-Agent Gaussian Splatting SLAM
von: Li, Monica M. Q., et al.
Veröffentlicht: (2026)
von: Li, Monica M. Q., et al.
Veröffentlicht: (2026)
SparseOIT: Improving Order-Independent Transparency 3DGS via Active Set Method
von: Yang, Wentao, et al.
Veröffentlicht: (2026)
von: Yang, Wentao, et al.
Veröffentlicht: (2026)
VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA
von: He, Haibin, et al.
Veröffentlicht: (2026)
von: He, Haibin, et al.
Veröffentlicht: (2026)
Range-Agnostic Multi-View Depth Estimation With Keyframe Selection
von: Conti, Andrea, et al.
Veröffentlicht: (2024)
von: Conti, Andrea, et al.
Veröffentlicht: (2024)
Generative Inbetweening: Adapting Image-to-Video Models for Keyframe Interpolation
von: Wang, Xiaojuan, et al.
Veröffentlicht: (2024)
von: Wang, Xiaojuan, et al.
Veröffentlicht: (2024)
From Captions to Keyframes: KeyScore for Multimodal Frame Scoring and Video-Language Understanding
von: Lin, Shih-Yao, et al.
Veröffentlicht: (2025)
von: Lin, Shih-Yao, et al.
Veröffentlicht: (2025)
Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval
von: Shlapentokh-Rothman, Michal, et al.
Veröffentlicht: (2026)
von: Shlapentokh-Rothman, Michal, et al.
Veröffentlicht: (2026)
AdaFlow: Efficient Long Video Editing via Adaptive Attention Slimming And Keyframe Selection
von: Zhang, Shuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Shuheng, et al.
Veröffentlicht: (2025)
Adaptive Keyframe Sampling for Long Video Understanding
von: Tang, Xi, et al.
Veröffentlicht: (2025)
von: Tang, Xi, et al.
Veröffentlicht: (2025)
FOCUS: Efficient Keyframe Selection for Long Video Understanding
von: Zhu, Zirui, et al.
Veröffentlicht: (2025)
von: Zhu, Zirui, et al.
Veröffentlicht: (2025)
KeyVideoLLM: Towards Large-scale Video Keyframe Selection
von: Liang, Hao, et al.
Veröffentlicht: (2024)
von: Liang, Hao, et al.
Veröffentlicht: (2024)
Solving Short-Term Relocalization Problems In Monocular Keyframe Visual SLAM Using Spatial And Semantic Data
von: Kamal, Azmyin Md., et al.
Veröffentlicht: (2024)
von: Kamal, Azmyin Md., et al.
Veröffentlicht: (2024)
SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
von: Yu, Jiongze, et al.
Veröffentlicht: (2026)
von: Yu, Jiongze, et al.
Veröffentlicht: (2026)
Less Is More, but Where? Dynamic Token Compression via LLM-Guided Keyframe Prior
von: Li, Yulin, et al.
Veröffentlicht: (2025)
von: Li, Yulin, et al.
Veröffentlicht: (2025)
ToonComposer: Streamlining Cartoon Production with Generative Post-Keyframing
von: Li, Lingen, et al.
Veröffentlicht: (2025)
von: Li, Lingen, et al.
Veröffentlicht: (2025)
DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes
von: Song, Zhende, et al.
Veröffentlicht: (2024)
von: Song, Zhende, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models
von: Zheng, Kai, et al.
Veröffentlicht: (2026) -
SemanticFace: Semantic Facial Action Estimation via Semantic Distillation in Interpretable Space
von: Kang, Zejian, et al.
Veröffentlicht: (2026) -
SuperFace: Preference-Aligned Facial Expression Estimation Beyond Pseudo Supervision
von: Kang, Zejian, et al.
Veröffentlicht: (2026) -
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
von: Lyu, Tianle, et al.
Veröffentlicht: (2025) -
KeyVID: Keyframe-Aware Video Diffusion for Audio-Synchronized Visual Animation
von: Wang, Xingrui, et al.
Veröffentlicht: (2025)