Enregistré dans:
| Auteurs principaux: | Wu, Jingchao, Kang, Zejian, Liu, Haibo, Fei, Yuanchen, Huang, Xiangru |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2512.11321 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models
par: Zheng, Kai, et autres
Publié: (2026)
par: Zheng, Kai, et autres
Publié: (2026)
SemanticFace: Semantic Facial Action Estimation via Semantic Distillation in Interpretable Space
par: Kang, Zejian, et autres
Publié: (2026)
par: Kang, Zejian, et autres
Publié: (2026)
SuperFace: Preference-Aligned Facial Expression Estimation Beyond Pseudo Supervision
par: Kang, Zejian, et autres
Publié: (2026)
par: Kang, Zejian, et autres
Publié: (2026)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
par: Lyu, Tianle, et autres
Publié: (2025)
par: Lyu, Tianle, et autres
Publié: (2025)
KeyVID: Keyframe-Aware Video Diffusion for Audio-Synchronized Visual Animation
par: Wang, Xingrui, et autres
Publié: (2025)
par: Wang, Xingrui, et autres
Publié: (2025)
Placing Human Animations into 3D Scenes by Learning Interaction- and Geometry-Driven Keyframes
par: Mullen Jr, James F., et autres
Publié: (2022)
par: Mullen Jr, James F., et autres
Publié: (2022)
Exploring the Role of Synthetic Data Augmentation in Controllable Human-Centric Video Generation
par: Fei, Yuanchen, et autres
Publié: (2026)
par: Fei, Yuanchen, et autres
Publié: (2026)
Occlusion-Aware Physics-Semantic Keyframe Selection for Robust Video Editing
par: Liu, Lin, et autres
Publié: (2026)
par: Liu, Lin, et autres
Publié: (2026)
DanceCamAnimator: Keyframe-Based Controllable 3D Dance Camera Synthesis
par: Wang, Zixuan, et autres
Publié: (2024)
par: Wang, Zixuan, et autres
Publié: (2024)
Learning Semantic Facial Descriptors for Accurate Face Animation
par: Zhu, Lei, et autres
Publié: (2025)
par: Zhu, Lei, et autres
Publié: (2025)
Coarse-to-Fine 3D Keyframe Transporter
par: Zhu, Xupeng, et autres
Publié: (2025)
par: Zhu, Xupeng, et autres
Publié: (2025)
Controllable Human-centric Keyframe Interpolation with Generative Prior
par: Guo, Zujin, et autres
Publié: (2025)
par: Guo, Zujin, et autres
Publié: (2025)
The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection
par: He, Qingdong, et autres
Publié: (2025)
par: He, Qingdong, et autres
Publié: (2025)
CineVerse: Consistent Keyframe Synthesis for Cinematic Scene Composition
par: Phung, Quynh, et autres
Publié: (2025)
par: Phung, Quynh, et autres
Publié: (2025)
Threading Keyframe with Narratives: MLLMs as Strong Long Video Comprehenders
par: Fang, Bo, et autres
Publié: (2025)
par: Fang, Bo, et autres
Publié: (2025)
KFFocus: Highlighting Keyframes for Enhanced Video Understanding
par: Nie, Ming, et autres
Publié: (2025)
par: Nie, Ming, et autres
Publié: (2025)
Agentic Keyframe Search for Video Question Answering
par: Fan, Sunqi, et autres
Publié: (2025)
par: Fan, Sunqi, et autres
Publié: (2025)
SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models
par: He, Ziheng, et autres
Publié: (2026)
par: He, Ziheng, et autres
Publié: (2026)
Generative Motion Infilling From Imprecisely Timed Keyframes
par: Goel, Purvi, et autres
Publié: (2025)
par: Goel, Purvi, et autres
Publié: (2025)
KS-APR: Keyframe Selection for Robust Absolute Pose Regression
par: Liu, Changkun, et autres
Publié: (2023)
par: Liu, Changkun, et autres
Publié: (2023)
Keyframe-Based Feed-Forward Visual Odometry
par: Dai, Weichen, et autres
Publié: (2026)
par: Dai, Weichen, et autres
Publié: (2026)
Large Model based Sequential Keyframe Extraction for Video Summarization
par: Tan, Kailong, et autres
Publié: (2024)
par: Tan, Kailong, et autres
Publié: (2024)
Less is More: Improving Motion Diffusion Models with Sparse Keyframes
par: Bae, Jinseok, et autres
Publié: (2025)
par: Bae, Jinseok, et autres
Publié: (2025)
SignSparK: Efficient Multilingual Sign Language Production via Sparse Keyframe Learning
par: Low, Jianhe, et autres
Publié: (2026)
par: Low, Jianhe, et autres
Publié: (2026)
Compact Keyframe-Optimized Multi-Agent Gaussian Splatting SLAM
par: Li, Monica M. Q., et autres
Publié: (2026)
par: Li, Monica M. Q., et autres
Publié: (2026)
SparseOIT: Improving Order-Independent Transparency 3DGS via Active Set Method
par: Yang, Wentao, et autres
Publié: (2026)
par: Yang, Wentao, et autres
Publié: (2026)
VTAgent: Agentic Keyframe Anchoring for Evidence-Aware Video TextVQA
par: He, Haibin, et autres
Publié: (2026)
par: He, Haibin, et autres
Publié: (2026)
Range-Agnostic Multi-View Depth Estimation With Keyframe Selection
par: Conti, Andrea, et autres
Publié: (2024)
par: Conti, Andrea, et autres
Publié: (2024)
Generative Inbetweening: Adapting Image-to-Video Models for Keyframe Interpolation
par: Wang, Xiaojuan, et autres
Publié: (2024)
par: Wang, Xiaojuan, et autres
Publié: (2024)
From Captions to Keyframes: KeyScore for Multimodal Frame Scoring and Video-Language Understanding
par: Lin, Shih-Yao, et autres
Publié: (2025)
par: Lin, Shih-Yao, et autres
Publié: (2025)
Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval
par: Shlapentokh-Rothman, Michal, et autres
Publié: (2026)
par: Shlapentokh-Rothman, Michal, et autres
Publié: (2026)
AdaFlow: Efficient Long Video Editing via Adaptive Attention Slimming And Keyframe Selection
par: Zhang, Shuheng, et autres
Publié: (2025)
par: Zhang, Shuheng, et autres
Publié: (2025)
Adaptive Keyframe Sampling for Long Video Understanding
par: Tang, Xi, et autres
Publié: (2025)
par: Tang, Xi, et autres
Publié: (2025)
FOCUS: Efficient Keyframe Selection for Long Video Understanding
par: Zhu, Zirui, et autres
Publié: (2025)
par: Zhu, Zirui, et autres
Publié: (2025)
KeyVideoLLM: Towards Large-scale Video Keyframe Selection
par: Liang, Hao, et autres
Publié: (2024)
par: Liang, Hao, et autres
Publié: (2024)
Solving Short-Term Relocalization Problems In Monocular Keyframe Visual SLAM Using Spatial And Semantic Data
par: Kamal, Azmyin Md., et autres
Publié: (2024)
par: Kamal, Azmyin Md., et autres
Publié: (2024)
SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
par: Yu, Jiongze, et autres
Publié: (2026)
par: Yu, Jiongze, et autres
Publié: (2026)
Less Is More, but Where? Dynamic Token Compression via LLM-Guided Keyframe Prior
par: Li, Yulin, et autres
Publié: (2025)
par: Li, Yulin, et autres
Publié: (2025)
ToonComposer: Streamlining Cartoon Production with Generative Post-Keyframing
par: Li, Lingen, et autres
Publié: (2025)
par: Li, Lingen, et autres
Publié: (2025)
DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes
par: Song, Zhende, et autres
Publié: (2024)
par: Song, Zhende, et autres
Publié: (2024)
Documents similaires
-
AudioFace: Language-Assisted Speech-Driven Facial Animation with Multimodal Language Models
par: Zheng, Kai, et autres
Publié: (2026) -
SemanticFace: Semantic Facial Action Estimation via Semantic Distillation in Interpretable Space
par: Kang, Zejian, et autres
Publié: (2026) -
SuperFace: Preference-Aligned Facial Expression Estimation Beyond Pseudo Supervision
par: Kang, Zejian, et autres
Publié: (2026) -
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
par: Lyu, Tianle, et autres
Publié: (2025) -
KeyVID: Keyframe-Aware Video Diffusion for Audio-Synchronized Visual Animation
par: Wang, Xingrui, et autres
Publié: (2025)