ViMo: Generating Motions from Casual Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Qiu, Liangdong, Yu, Chengxing, Li, Yanran, Wang, Zhao, Huang, Haibin, Ma, Chongyang, Zhang, Di, Wan, Pengfei, Han, Xiaoguang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Audio Generators Become Good Listeners: Generative Features for Understanding Tasks
by: Xie, Zeyu, et al.
Published: (2025)
by: Xie, Zeyu, et al.
Published: (2025)
SemanticVocoder: Bridging Audio Generation and Audio Understanding via Semantic Latents
by: Xie, Zeyu, et al.
Published: (2026)
by: Xie, Zeyu, et al.
Published: (2026)
From Benchmarks to Business Impact: Deploying IBM Generalist Agent in Enterprise Production
by: Shlomov, Segev, et al.
Published: (2025)
by: Shlomov, Segev, et al.
Published: (2025)
Any four real numbers are on all fours with analogy
by: Lepage, Yves, et al.
Published: (2024)
by: Lepage, Yves, et al.
Published: (2024)
Self-Composing Neural Operators with Depth and Accuracy Scaling via Adaptive Train-and-Unroll Approach
by: He, Juncai, et al.
Published: (2025)
by: He, Juncai, et al.
Published: (2025)
GraphCNNpred: A stock market indices prediction using a Graph based deep learning system
by: Jin, Yuhui
Published: (2024)
by: Jin, Yuhui
Published: (2024)
On the Robustness of Decision-Focused Learning
by: Farhat, Yehya
Published: (2023)
by: Farhat, Yehya
Published: (2023)
Big Data Intelligence Using Distributed Deep Neural Networks
by: Ongati, Felix, et al.
Published: (2019)
by: Ongati, Felix, et al.
Published: (2019)
Verifiable Dropout: Turning Randomness into a Verifiable Claim
by: Lee, Kichang, et al.
Published: (2025)
by: Lee, Kichang, et al.
Published: (2025)
Simple Prompt Injection Attacks Can Leak Personal Data Observed by LLM Agents During Task Execution
by: Alizadeh, Meysam, et al.
Published: (2025)
by: Alizadeh, Meysam, et al.
Published: (2025)
Detection of Chagas Disease from the ECG: The George B. Moody PhysioNet Challenge 2025
by: Reyna, Matthew A., et al.
Published: (2025)
by: Reyna, Matthew A., et al.
Published: (2025)
FakeSound: Deepfake General Audio Detection
by: Xie, Zeyu, et al.
Published: (2024)
by: Xie, Zeyu, et al.
Published: (2024)
RMFlow: Refined Mean Flow by a Noise-Injection Step for Multimodal Generation
by: Huang, Yuhao, et al.
Published: (2026)
by: Huang, Yuhao, et al.
Published: (2026)
Moving Object Proposals with Deep Learned Optical Flow for Video Object Segmentation
by: Shi, Ge, et al.
Published: (2024)
by: Shi, Ge, et al.
Published: (2024)
Adaptive Dataset Quantization: A New Direction for Dataset Pruning
by: Yu, Chenyue, et al.
Published: (2025)
by: Yu, Chenyue, et al.
Published: (2025)
STAR: Speech-to-Audio Generation via Representation Learning
by: Xie, Zeyu, et al.
Published: (2025)
by: Xie, Zeyu, et al.
Published: (2025)
Distributional Drift Adaptation with Temporal Conditional Variational Autoencoder for Multivariate Time Series Forecasting
by: He, Hui, et al.
Published: (2022)
by: He, Hui, et al.
Published: (2022)
PicoAudio2: Temporal Controllable Text-to-Audio Generation with Natural Language Description
by: Zheng, Zihao, et al.
Published: (2025)
by: Zheng, Zihao, et al.
Published: (2025)
PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation
by: Xie, Zeyu, et al.
Published: (2024)
by: Xie, Zeyu, et al.
Published: (2024)
MuChin: A Chinese Colloquial Description Benchmark for Evaluating Language Models in the Field of Music
by: Wang, Zihao, et al.
Published: (2024)
by: Wang, Zihao, et al.
Published: (2024)
Exploring ChatGPT and its Impact on Society
by: Haque, Md. Asraful, et al.
Published: (2024)
by: Haque, Md. Asraful, et al.
Published: (2024)
FakeSound2: A Benchmark for Explainable and Generalizable Deepfake Sound Detection
by: Xie, Zeyu, et al.
Published: (2025)
by: Xie, Zeyu, et al.
Published: (2025)
AudioTime: A Temporally-aligned Audio-text Benchmark Dataset
by: Xie, Zeyu, et al.
Published: (2024)
by: Xie, Zeyu, et al.
Published: (2024)
Contemporary Agent Technology: LLM-Driven Advancements vs Classic Multi-Agent Systems
by: Bădică, Costin, et al.
Published: (2025)
by: Bădică, Costin, et al.
Published: (2025)
CAST-TTS: A Simple Cross-Attention Framework for Unified Timbre Control in TTS
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
MuDiT & MuSiT: Alignment with Colloquial Expression in Description-to-Song Generation
by: Wang, Zihao, et al.
Published: (2024)
by: Wang, Zihao, et al.
Published: (2024)
SaMoye: Zero-shot Singing Voice Conversion Model Based on Feature Disentanglement and Enhancement
by: Wang, Zihao, et al.
Published: (2024)
by: Wang, Zihao, et al.
Published: (2024)
Whole-Song Hierarchical Generation of Symbolic Music Using Cascaded Diffusion Models
by: Wang, Ziyu, et al.
Published: (2024)
by: Wang, Ziyu, et al.
Published: (2024)
RoNFA: Robust Neural Field-based Approach for Few-Shot Image Classification with Noisy Labels
by: Xiang, Nan, et al.
Published: (2025)
by: Xiang, Nan, et al.
Published: (2025)
Robust Multivariate Time Series Forecasting against Intra- and Inter-Series Transitional Shift
by: He, Hui, et al.
Published: (2024)
by: He, Hui, et al.
Published: (2024)
AI-Driven Innovations in Modern Cloud Computing
by: Kumar, Animesh
Published: (2024)
by: Kumar, Animesh
Published: (2024)
Enabling Trustworthy Federated Learning in Industrial IoT: Bridging the Gap Between Interpretability and Robustness
by: Jagatheesaperumal, Senthil Kumar, et al.
Published: (2024)
by: Jagatheesaperumal, Senthil Kumar, et al.
Published: (2024)
Redefining Finance: The Influence of Artificial Intelligence (AI) and Machine Learning (ML)
by: Kumar, Animesh
Published: (2024)
by: Kumar, Animesh
Published: (2024)
PairHuman: A High-Fidelity Photographic Dataset for Customized Dual-Person Generation
by: Pan, Ting, et al.
Published: (2025)
by: Pan, Ting, et al.
Published: (2025)
Koopman operator learning using invertible neural networks
by: Meng, Yuhuang, et al.
Published: (2023)
by: Meng, Yuhuang, et al.
Published: (2023)
Comparative Analysis of Lightweight Deep Learning Models for Memory-Constrained Devices
by: Shahriar, Tasnim
Published: (2025)
by: Shahriar, Tasnim
Published: (2025)
AdamHD: Decoupled Huber Decay Regularization for Language Model Pre-Training
by: Guo, Fu-Ming, et al.
Published: (2025)
by: Guo, Fu-Ming, et al.
Published: (2025)
Reinforced Inverse Scattering
by: Jiang, Hanyang, et al.
Published: (2022)
by: Jiang, Hanyang, et al.
Published: (2022)
Internal noise in hardware deep and recurrent neural networks helps with learning
by: Kolesnikov, Ivan, et al.
Published: (2025)
by: Kolesnikov, Ivan, et al.
Published: (2025)
Conversion rate prediction in online advertising: modeling techniques, performance evaluation and future directions
by: Xue, Tao, et al.
Published: (2025)
by: Xue, Tao, et al.
Published: (2025)
Similar Items
-
When Audio Generators Become Good Listeners: Generative Features for Understanding Tasks
by: Xie, Zeyu, et al.
Published: (2025) -
SemanticVocoder: Bridging Audio Generation and Audio Understanding via Semantic Latents
by: Xie, Zeyu, et al.
Published: (2026) -
From Benchmarks to Business Impact: Deploying IBM Generalist Agent in Enterprise Production
by: Shlomov, Segev, et al.
Published: (2025) -
Any four real numbers are on all fours with analogy
by: Lepage, Yves, et al.
Published: (2024) -
Self-Composing Neural Operators with Depth and Accuracy Scaling via Adaptive Train-and-Unroll Approach
by: He, Juncai, et al.
Published: (2025)