TIMotion: Temporal and Interactive Framework for Efficient Human-Human Motion Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yabiao, Wang, Shuo, Zhang, Jiangning, Fan, Ke, Wu, Jiafu, Xue, Zhucun, Liu, Yong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MARRS: Masked Autoregressive Unit-based Reaction Synthesis
von: Wang, Yabiao, et al.
Veröffentlicht: (2025)
von: Wang, Yabiao, et al.
Veröffentlicht: (2025)
Advancing Narrative Long Video Generation via Training-Free Identity-Aware Memory
von: Liu, Jinzhuo, et al.
Veröffentlicht: (2026)
von: Liu, Jinzhuo, et al.
Veröffentlicht: (2026)
GPT-4V-AD: Exploring Grounding Potential of VQA-oriented GPT-4V for Zero-shot Anomaly Detection
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
Improving Autoregressive Visual Generation with Cluster-Oriented Token Prediction
von: Hu, Teng, et al.
Veröffentlicht: (2025)
von: Hu, Teng, et al.
Veröffentlicht: (2025)
EMOv2: Pushing 5M Vision Model Frontier
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
HumanVideo-MME: Benchmarking MLLMs for Human-Centric Video Understanding
von: Cai, Yuxuan, et al.
Veröffentlicht: (2025)
von: Cai, Yuxuan, et al.
Veröffentlicht: (2025)
FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis
von: Fan, Ke, et al.
Veröffentlicht: (2024)
von: Fan, Ke, et al.
Veröffentlicht: (2024)
Textual Decomposition Then Sub-motion-space Scattering for Open-Vocabulary Motion Generation
von: Fan, Ke, et al.
Veröffentlicht: (2024)
von: Fan, Ke, et al.
Veröffentlicht: (2024)
PVG: Progressive Vision Graph for Vision Recognition
von: Wu, Jiafu, et al.
Veröffentlicht: (2023)
von: Wu, Jiafu, et al.
Veröffentlicht: (2023)
A Comprehensive Library for Benchmarking Multi-class Visual Anomaly Detection
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
PiT: Progressive Diffusion Transformer
von: Wu, Jiafu, et al.
Veröffentlicht: (2025)
von: Wu, Jiafu, et al.
Veröffentlicht: (2025)
Transform Trained Transformer: Accelerating Naive 4K Video Generation Over 10$\times$
von: Zhang, Jiangning, et al.
Veröffentlicht: (2025)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2025)
MotionMaster: Training-free Camera Motion Transfer For Video Generation
von: Hu, Teng, et al.
Veröffentlicht: (2024)
von: Hu, Teng, et al.
Veröffentlicht: (2024)
SPIKE: An Adaptive Dual Controller Framework for Cost-Efficient Long-Horizon Game Agents
von: Jiang, Wencan, et al.
Veröffentlicht: (2026)
von: Jiang, Wencan, et al.
Veröffentlicht: (2026)
AdaVideoRAG: Omni-Contextual Adaptive Retrieval-Augmented Efficient Long Video Understanding
von: Xue, Zhucun, et al.
Veröffentlicht: (2025)
von: Xue, Zhucun, et al.
Veröffentlicht: (2025)
UltraVideo: High-Quality UHD Video Dataset with Comprehensive Captions
von: Xue, Zhucun, et al.
Veröffentlicht: (2025)
von: Xue, Zhucun, et al.
Veröffentlicht: (2025)
UltraLBM-UNet: Ultralight Bidirectional Mamba-based Model for Skin Lesion Segmentation
von: Fan, Linxuan, et al.
Veröffentlicht: (2025)
von: Fan, Linxuan, et al.
Veröffentlicht: (2025)
Semantic Frame Interpolation
von: Hong, Yijia, et al.
Veröffentlicht: (2025)
von: Hong, Yijia, et al.
Veröffentlicht: (2025)
Learning Feature Inversion for Multi-class Anomaly Detection under General-purpose COCO-AD Benchmark
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Mao, Xiaofeng, et al.
Veröffentlicht: (2024)
InstanceV: Instance-Level Video Generation
von: Chen, Yuheng, et al.
Veröffentlicht: (2025)
von: Chen, Yuheng, et al.
Veröffentlicht: (2025)
InterMamba: Efficient Human-Human Interaction Generation with Adaptive Spatio-Temporal Mamba
von: Wu, Zizhao, et al.
Veröffentlicht: (2025)
von: Wu, Zizhao, et al.
Veröffentlicht: (2025)
Efficient and Scalable Monocular Human-Object Interaction Motion Reconstruction
von: Wen, Boran, et al.
Veröffentlicht: (2025)
von: Wen, Boran, et al.
Veröffentlicht: (2025)
Motion Avatar: Generate Human and Animal Avatars with Arbitrary Motion
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
LLaVA-KD: A Framework of Distilling Multimodal Large Language Models
von: Cai, Yuxuan, et al.
Veröffentlicht: (2024)
von: Cai, Yuxuan, et al.
Veröffentlicht: (2024)
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
von: Wang, Qilin, et al.
Veröffentlicht: (2024)
von: Wang, Qilin, et al.
Veröffentlicht: (2024)
Motion-Agent: A Conversational Framework for Human Motion Generation with LLMs
von: Wu, Qi, et al.
Veröffentlicht: (2024)
von: Wu, Qi, et al.
Veröffentlicht: (2024)
SaRA: High-Efficient Diffusion Model Fine-tuning with Progressive Sparse Low-Rank Adaptation
von: Hu, Teng, et al.
Veröffentlicht: (2024)
von: Hu, Teng, et al.
Veröffentlicht: (2024)
Evolution of Optimization Methods: Algorithms, Scenarios, and Evaluations
von: Zhang, Tong, et al.
Veröffentlicht: (2026)
von: Zhang, Tong, et al.
Veröffentlicht: (2026)
PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud Learning
von: He, Qingdong, et al.
Veröffentlicht: (2024)
von: He, Qingdong, et al.
Veröffentlicht: (2024)
UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy
von: Xu, Yicheng, et al.
Veröffentlicht: (2026)
von: Xu, Yicheng, et al.
Veröffentlicht: (2026)
Identity-Preserving Text-to-Video Generation Guided by Simple yet Effective Spatial-Temporal Decoupled Representations
von: Wang, Yuji, et al.
Veröffentlicht: (2025)
von: Wang, Yuji, et al.
Veröffentlicht: (2025)
SwiftVideo: A Unified Framework for Few-Step Video Generation through Trajectory-Distribution Alignment
von: Sun, Yanxiao, et al.
Veröffentlicht: (2025)
von: Sun, Yanxiao, et al.
Veröffentlicht: (2025)
Language-Guided Transformer Tokenizer for Human Motion Generation
von: Yan, Sheng, et al.
Veröffentlicht: (2026)
von: Yan, Sheng, et al.
Veröffentlicht: (2026)
CLIP-AD: A Language-Guided Staged Dual-Path Model for Zero-shot Anomaly Detection
von: Chen, Xuhai, et al.
Veröffentlicht: (2023)
von: Chen, Xuhai, et al.
Veröffentlicht: (2023)
Dynamic Worlds, Dynamic Humans: Generating Virtual Human-Scene Interaction Motion in Dynamic Scenes
von: Wang, Yin, et al.
Veröffentlicht: (2026)
von: Wang, Yin, et al.
Veröffentlicht: (2026)
Generating Human Interaction Motions in Scenes with Text Control
von: Yi, Hongwei, et al.
Veröffentlicht: (2024)
von: Yi, Hongwei, et al.
Veröffentlicht: (2024)
AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
von: Hu, Teng, et al.
Veröffentlicht: (2023)
von: Hu, Teng, et al.
Veröffentlicht: (2023)
Multi-Dimensional Knowledge Profiling with Large-Scale Literature Database and Hierarchical Retrieval
von: Xue, Zhucun, et al.
Veröffentlicht: (2026)
von: Xue, Zhucun, et al.
Veröffentlicht: (2026)
HumanDreamer: Generating Controllable Human-Motion Videos via Decoupled Generation
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MARRS: Masked Autoregressive Unit-based Reaction Synthesis
von: Wang, Yabiao, et al.
Veröffentlicht: (2025) -
Advancing Narrative Long Video Generation via Training-Free Identity-Aware Memory
von: Liu, Jinzhuo, et al.
Veröffentlicht: (2026) -
GPT-4V-AD: Exploring Grounding Potential of VQA-oriented GPT-4V for Zero-shot Anomaly Detection
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023) -
Improving Autoregressive Visual Generation with Cluster-Oriented Token Prediction
von: Hu, Teng, et al.
Veröffentlicht: (2025) -
EMOv2: Pushing 5M Vision Model Frontier
von: Zhang, Jiangning, et al.
Veröffentlicht: (2024)