Motion Graph Unleashed: A Novel Approach to Video Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Zhong, Yiqi, Liang, Luming, Tang, Bohan, Zharkov, Ilya, Neumann, Ulrich |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FORA: Fast-Forward Caching in Diffusion Transformer Acceleration
by: Selvaraju, Pratheba, et al.
Published: (2024)
by: Selvaraju, Pratheba, et al.
Published: (2024)
CaesarNeRF: Calibrated Semantic Representation for Few-shot Generalizable Neural Rendering
by: Zhu, Haidong, et al.
Published: (2023)
by: Zhu, Haidong, et al.
Published: (2023)
AdaContour: Adaptive Contour Descriptor with Hierarchical Representation
by: Ding, Tianyu, et al.
Published: (2024)
by: Ding, Tianyu, et al.
Published: (2024)
Meta-Optimization for Higher Model Generalizability in Single-Image Depth Prediction
by: Wu, Cho-Ying, et al.
Published: (2023)
by: Wu, Cho-Ying, et al.
Published: (2023)
ProCrop: Learning Aesthetic Image Cropping from Professional Compositions
by: Zhang, Ke, et al.
Published: (2025)
by: Zhang, Ke, et al.
Published: (2025)
DREAM: Diffusion Rectification and Estimation-Adaptive Models
by: Zhou, Jinxin, et al.
Published: (2023)
by: Zhou, Jinxin, et al.
Published: (2023)
S3Editor: A Sparse Semantic-Disentangled Self-Training Framework for Face Video Editing
by: Wang, Guangzhi, et al.
Published: (2024)
by: Wang, Guangzhi, et al.
Published: (2024)
Collaborative Uncertainty Benefits Multi-Agent Multi-Modal Trajectory Forecasting
by: Tang, Bohan, et al.
Published: (2022)
by: Tang, Bohan, et al.
Published: (2022)
Moaw: Unleashing Motion Awareness for Video Diffusion Models
by: Zhang, Tianqi, et al.
Published: (2026)
by: Zhang, Tianqi, et al.
Published: (2026)
Boosting Generalizability towards Zero-Shot Cross-Dataset Single-Image Indoor Depth by Meta-Initialization
by: Wu, Cho-Ying, et al.
Published: (2024)
by: Wu, Cho-Ying, et al.
Published: (2024)
Self-Supervised Bird's Eye View Motion Prediction with Cross-Modality Signals
by: Fang, Shaoheng, et al.
Published: (2024)
by: Fang, Shaoheng, et al.
Published: (2024)
Distinguish Any Fake Videos: Unleashing the Power of Large-scale Data and Motion Features
by: Ji, Lichuan, et al.
Published: (2024)
by: Ji, Lichuan, et al.
Published: (2024)
Video Motion Graphs
by: Liu, Haiyang, et al.
Published: (2025)
by: Liu, Haiyang, et al.
Published: (2025)
Cat-AIR: Content and Task-Aware All-in-One Image Restoration
by: Jiang, Jiachen, et al.
Published: (2025)
by: Jiang, Jiachen, et al.
Published: (2025)
InfoMotion: A Graph-Based Approach to Video Dataset Distillation for Echocardiography
by: Li, Zhe, et al.
Published: (2025)
by: Li, Zhe, et al.
Published: (2025)
Motion Mamba: Efficient and Long Sequence Motion Generation
by: Zhang, Zeyu, et al.
Published: (2024)
by: Zhang, Zeyu, et al.
Published: (2024)
Unleashing the Power of Motion and Depth: A Selective Fusion Strategy for RGB-D Video Salient Object Detection
by: He, Jiahao, et al.
Published: (2025)
by: He, Jiahao, et al.
Published: (2025)
HMPDM: A Diffusion Model for Driving Video Prediction with Historical Motion Priors
by: Li, Ke, et al.
Published: (2026)
by: Li, Ke, et al.
Published: (2026)
HESSO: Towards Automatic Efficient and User Friendly Any Neural Network Training and Pruning
by: Chen, Tianyi, et al.
Published: (2024)
by: Chen, Tianyi, et al.
Published: (2024)
OFER: Occluded Face Expression Reconstruction
by: Selvaraju, Pratheba, et al.
Published: (2024)
by: Selvaraju, Pratheba, et al.
Published: (2024)
Unveiling Context-Related Anomalies: Knowledge Graph Empowered Decoupling of Scene and Action for Human-Related Video Anomaly Detection
by: Chen, Chenglizhao, et al.
Published: (2024)
by: Chen, Chenglizhao, et al.
Published: (2024)
COCO is "ALL'' You Need for Visual Instruction Fine-tuning
by: Han, Xiaotian, et al.
Published: (2024)
by: Han, Xiaotian, et al.
Published: (2024)
OWL: A Novel Approach to Machine Perception During Motion
by: Raviv, Daniel, et al.
Published: (2026)
by: Raviv, Daniel, et al.
Published: (2026)
Unleashing Vision-Language Semantics for Deepfake Video Detection
by: Zhu, Jiawen, et al.
Published: (2026)
by: Zhu, Jiawen, et al.
Published: (2026)
Unleashing Generalization of End-to-End Autonomous Driving with Controllable Long Video Generation
by: Ma, Enhui, et al.
Published: (2024)
by: Ma, Enhui, et al.
Published: (2024)
InfiniMotion: Mamba Boosts Memory in Transformer for Arbitrary Long Motion Generation
by: Zhang, Zeyu, et al.
Published: (2024)
by: Zhang, Zeyu, et al.
Published: (2024)
MotionCharacter: Fine-Grained Motion Controllable Human Video Generation
by: Fang, Haopeng, et al.
Published: (2024)
by: Fang, Haopeng, et al.
Published: (2024)
Unleashing the Potential of SAM2 for Biomedical Images and Videos: A Survey
by: Zhang, Yichi, et al.
Published: (2024)
by: Zhang, Yichi, et al.
Published: (2024)
PAM: A Pose-Appearance-Motion Engine for Sim-to-Real HOI Video Generation
by: Gao, Mingju, et al.
Published: (2026)
by: Gao, Mingju, et al.
Published: (2026)
4DEquine: Disentangling Motion and Appearance for 4D Equine Reconstruction from Monocular Video
by: Lyu, Jin, et al.
Published: (2026)
by: Lyu, Jin, et al.
Published: (2026)
COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing
by: Wang, Jiangshan, et al.
Published: (2024)
by: Wang, Jiangshan, et al.
Published: (2024)
Unleashing the Potential of Tracklets for Unsupervised Video Person Re-Identification
by: Meng, Nanxing, et al.
Published: (2024)
by: Meng, Nanxing, et al.
Published: (2024)
Unleashing Hour-Scale Video Training for Long Video-Language Understanding
by: Lin, Jingyang, et al.
Published: (2025)
by: Lin, Jingyang, et al.
Published: (2025)
Motion Inversion for Video Customization
by: Wang, Luozhou, et al.
Published: (2024)
by: Wang, Luozhou, et al.
Published: (2024)
MVGD-Net: A Novel Motion-aware Video Glass Surface Detection Network
by: Lu, Yiwei, et al.
Published: (2026)
by: Lu, Yiwei, et al.
Published: (2026)
Unleashing Network Potentials for Semantic Scene Completion
by: Wang, Fengyun, et al.
Published: (2024)
by: Wang, Fengyun, et al.
Published: (2024)
E2VIDiff: Perceptual Events-to-Video Reconstruction using Diffusion Priors
by: Liang, Jinxiu, et al.
Published: (2024)
by: Liang, Jinxiu, et al.
Published: (2024)
FocusGraph: Graph-Structured Frame Selection for Embodied Long Video Question Answering
by: Zemskova, Tatiana, et al.
Published: (2026)
by: Zemskova, Tatiana, et al.
Published: (2026)
Unleashing Video Language Models for Fine-grained HRCT Report Generation
by: Fang, Yingying, et al.
Published: (2026)
by: Fang, Yingying, et al.
Published: (2026)
Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation
by: Huang, Shaofei, et al.
Published: (2024)
by: Huang, Shaofei, et al.
Published: (2024)
Similar Items
-
FORA: Fast-Forward Caching in Diffusion Transformer Acceleration
by: Selvaraju, Pratheba, et al.
Published: (2024) -
CaesarNeRF: Calibrated Semantic Representation for Few-shot Generalizable Neural Rendering
by: Zhu, Haidong, et al.
Published: (2023) -
AdaContour: Adaptive Contour Descriptor with Hierarchical Representation
by: Ding, Tianyu, et al.
Published: (2024) -
Meta-Optimization for Higher Model Generalizability in Single-Image Depth Prediction
by: Wu, Cho-Ying, et al.
Published: (2023) -
ProCrop: Learning Aesthetic Image Cropping from Professional Compositions
by: Zhang, Ke, et al.
Published: (2025)