SIDQL: An Efficient Keyframe Extraction and Motion Reconstruction Framework in Motion Capture
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Xuling, Zhang, Ziru, Wang, Yuyang, Lee, Lik-hang, Hui, Pan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Audio Matters Too! Enhancing Markerless Motion Capture with Audio Signals for String Performance Capture
di: Jin, Yitong, et al.
Pubblicazione: (2024)
di: Jin, Yitong, et al.
Pubblicazione: (2024)
StableMoFusion: Towards Robust and Efficient Diffusion-based Motion Generation Framework
di: Huang, Yiheng, et al.
Pubblicazione: (2024)
di: Huang, Yiheng, et al.
Pubblicazione: (2024)
MotionPro: A Precise Motion Controller for Image-to-Video Generation
di: Zhang, Zhongwei, et al.
Pubblicazione: (2025)
di: Zhang, Zhongwei, et al.
Pubblicazione: (2025)
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
di: Wang, Sen, et al.
Pubblicazione: (2024)
di: Wang, Sen, et al.
Pubblicazione: (2024)
AMD: Autoregressive Motion Diffusion
di: Han, Bo, et al.
Pubblicazione: (2023)
di: Han, Bo, et al.
Pubblicazione: (2023)
Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding
di: Wang, Shaoguang, et al.
Pubblicazione: (2026)
di: Wang, Shaoguang, et al.
Pubblicazione: (2026)
Stepwise Schema-Guided Prompting Framework with Parameter Efficient Instruction Tuning for Multimedia Event Extraction
di: Yuan, Xiang, et al.
Pubblicazione: (2025)
di: Yuan, Xiang, et al.
Pubblicazione: (2025)
Harmony-Aware Music-driven Motion Synthesis with Perceptual Constraint on UGC Datasets
di: Wu, Xinyi, et al.
Pubblicazione: (2025)
di: Wu, Xinyi, et al.
Pubblicazione: (2025)
Efficient Sub-pixel Motion Compensation in Learned Video Codecs
di: Ladune, Théo, et al.
Pubblicazione: (2025)
di: Ladune, Théo, et al.
Pubblicazione: (2025)
Text-controlled Motion Mamba: Text-Instructed Temporal Grounding of Human Motion
di: Wang, Xinghan, et al.
Pubblicazione: (2024)
di: Wang, Xinghan, et al.
Pubblicazione: (2024)
MotionBeat: Motion-Aligned Music Representation via Embodied Contrastive Learning and Bar-Equivariant Contact-Aware Encoding
di: Wang, Xuanchen, et al.
Pubblicazione: (2025)
di: Wang, Xuanchen, et al.
Pubblicazione: (2025)
PP-Motion: Physical-Perceptual Fidelity Evaluation for Human Motion Generation
di: Zhao, Sihan, et al.
Pubblicazione: (2025)
di: Zhao, Sihan, et al.
Pubblicazione: (2025)
TriPSS: A Tri-Modal Keyframe Extraction Framework Using Perceptual, Structural, and Semantic Representations
di: Cakmak, Mert Can, et al.
Pubblicazione: (2025)
di: Cakmak, Mert Can, et al.
Pubblicazione: (2025)
AV1 Motion Vector Fidelity and Application for Efficient Optical Flow
di: Zouein, Julien, et al.
Pubblicazione: (2025)
di: Zouein, Julien, et al.
Pubblicazione: (2025)
MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
di: Zhang, Yuang, et al.
Pubblicazione: (2024)
di: Zhang, Yuang, et al.
Pubblicazione: (2024)
Compression Metadata-assisted RoI Extraction and Adaptive Inference for Efficient Video Analytics
di: Wang, Chengzhi, et al.
Pubblicazione: (2025)
di: Wang, Chengzhi, et al.
Pubblicazione: (2025)
Mesquite MoCap: Democratizing Real-Time Motion Capture with Affordable, Bodyworn IoT Sensors and WebXR SLAM
di: Vanani, Poojan, et al.
Pubblicazione: (2025)
di: Vanani, Poojan, et al.
Pubblicazione: (2025)
SpeechEE: A Novel Benchmark for Speech Event Extraction
di: Wang, Bin, et al.
Pubblicazione: (2024)
di: Wang, Bin, et al.
Pubblicazione: (2024)
PlanMoGPT: Flow-Enhanced Progressive Planning for Text to Motion Synthesis
di: Jin, Chuhao, et al.
Pubblicazione: (2025)
di: Jin, Chuhao, et al.
Pubblicazione: (2025)
Human Motion Video Generation: A Survey
di: Xue, Haiwei, et al.
Pubblicazione: (2025)
di: Xue, Haiwei, et al.
Pubblicazione: (2025)
Recognizing Everything from All Modalities at Once: Grounded Multimodal Universal Information Extraction
di: Zhang, Meishan, et al.
Pubblicazione: (2024)
di: Zhang, Meishan, et al.
Pubblicazione: (2024)
KeyVideoLLM: Towards Large-scale Video Keyframe Selection
di: Liang, Hao, et al.
Pubblicazione: (2024)
di: Liang, Hao, et al.
Pubblicazione: (2024)
MeMo: Attentional Momentum for Real-time Audio-visual Speaker Extraction under Impaired Visual Conditions
di: Li, Junjie, et al.
Pubblicazione: (2025)
di: Li, Junjie, et al.
Pubblicazione: (2025)
MagicMotion: Controllable Video Generation with Dense-to-Sparse Trajectory Guidance
di: Li, Quanhao, et al.
Pubblicazione: (2025)
di: Li, Quanhao, et al.
Pubblicazione: (2025)
DanceCamAnimator: Keyframe-Based Controllable 3D Dance Camera Synthesis
di: Wang, Zixuan, et al.
Pubblicazione: (2024)
di: Wang, Zixuan, et al.
Pubblicazione: (2024)
MotionDreamer: One-to-Many Motion Synthesis with Localized Generative Masked Transformer
di: Wang, Yilin, et al.
Pubblicazione: (2025)
di: Wang, Yilin, et al.
Pubblicazione: (2025)
REArtGS: Reconstructing and Generating Articulated Objects via 3D Gaussian Splatting with Geometric and Motion Constraints
di: Wu, Di, et al.
Pubblicazione: (2025)
di: Wu, Di, et al.
Pubblicazione: (2025)
LocoMotion: Learning Motion-Focused Video-Language Representations
di: Doughty, Hazel, et al.
Pubblicazione: (2024)
di: Doughty, Hazel, et al.
Pubblicazione: (2024)
Multimodal Cyber-physical Interaction in XR: Hybrid Doctoral Thesis Defense
di: Alhilal, Ahmad, et al.
Pubblicazione: (2026)
di: Alhilal, Ahmad, et al.
Pubblicazione: (2026)
CueNet: Robust Audio-Visual Speaker Extraction through Cross-Modal Cue Mining and Interaction
di: Wang, Jiadong, et al.
Pubblicazione: (2026)
di: Wang, Jiadong, et al.
Pubblicazione: (2026)
OTCR: Optimal Transmission, Compression and Representation for Multimodal Information Extraction
di: Li, Yang, et al.
Pubblicazione: (2025)
di: Li, Yang, et al.
Pubblicazione: (2025)
Multi-modal and Metadata Capture Model for Micro Video Popularity Prediction
di: Lu, Jiacheng, et al.
Pubblicazione: (2025)
di: Lu, Jiacheng, et al.
Pubblicazione: (2025)
FAST-ME: Foundation-aware Adaptive Stopping for Motion Estimation for Efficient IoT Video Analysis
di: Panagidi, Kakia, et al.
Pubblicazione: (2026)
di: Panagidi, Kakia, et al.
Pubblicazione: (2026)
Benchmarking and Improving LVLMs on Event Extraction from Multimedia Documents
di: Xing, Fuyu, et al.
Pubblicazione: (2025)
di: Xing, Fuyu, et al.
Pubblicazione: (2025)
ViMo: Generating Motions from Casual Videos
di: Qiu, Liangdong, et al.
Pubblicazione: (2024)
di: Qiu, Liangdong, et al.
Pubblicazione: (2024)
MotionCtrl: A Unified and Flexible Motion Controller for Video Generation
di: Wang, Zhouxia, et al.
Pubblicazione: (2023)
di: Wang, Zhouxia, et al.
Pubblicazione: (2023)
Generating Attribute-Aware Human Motions from Textual Prompt
di: Wang, Xinghan, et al.
Pubblicazione: (2025)
di: Wang, Xinghan, et al.
Pubblicazione: (2025)
Multimodal Graph-Based Variational Mixture of Experts Network for Zero-Shot Multimodal Information Extraction
di: Zhou, Baohang, et al.
Pubblicazione: (2025)
di: Zhou, Baohang, et al.
Pubblicazione: (2025)
High-Fidelity 3D Gaussian Human Reconstruction via Region-Aware Initialization and Geometric Priors
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
Towards Robust and Controllable Text-to-Motion via Masked Autoregressive Diffusion
di: Zhang, Zongye, et al.
Pubblicazione: (2025)
di: Zhang, Zongye, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Audio Matters Too! Enhancing Markerless Motion Capture with Audio Signals for String Performance Capture
di: Jin, Yitong, et al.
Pubblicazione: (2024) -
StableMoFusion: Towards Robust and Efficient Diffusion-based Motion Generation Framework
di: Huang, Yiheng, et al.
Pubblicazione: (2024) -
MotionPro: A Precise Motion Controller for Image-to-Video Generation
di: Zhang, Zhongwei, et al.
Pubblicazione: (2025) -
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
di: Wang, Sen, et al.
Pubblicazione: (2024) -
AMD: Autoregressive Motion Diffusion
di: Han, Bo, et al.
Pubblicazione: (2023)