Learning Long-term Motion Embeddings for Efficient Kinematics Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Stracke, Nick, Bauer, Kolja, Baumann, Stefan Andreas, Bautista, Miguel Angel, Susskind, Josh, Ommer, Björn |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CTRLorALTer: Conditional LoRAdapter for Efficient 0-Shot Control & Altering of T2I Models
von: Stracke, Nick, et al.
Veröffentlicht: (2024)
von: Stracke, Nick, et al.
Veröffentlicht: (2024)
CleanDIFT: Diffusion Features without Noise
von: Stracke, Nick, et al.
Veröffentlicht: (2024)
von: Stracke, Nick, et al.
Veröffentlicht: (2024)
What If : Understanding Motion Through Sparse Interactions
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2025)
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2025)
RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video
von: Prestel, Ulrich, et al.
Veröffentlicht: (2026)
von: Prestel, Ulrich, et al.
Veröffentlicht: (2026)
Adapting Self-Supervised Representations as a Latent Space for Efficient Generation
von: Gui, Ming, et al.
Veröffentlicht: (2025)
von: Gui, Ming, et al.
Veröffentlicht: (2025)
Probabilistic Precipitation Nowcasting with Rectified Flow Transformers
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2026)
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2026)
Boosting Latent Diffusion with Flow Matching
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2023)
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2023)
Continuous, Subject-Specific Attribute Control in T2I Models by Identifying Semantic Directions
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2024)
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2024)
MaskFlow: Discrete Flows For Flexible and Efficient Long Video Generation
von: Fuest, Michael, et al.
Veröffentlicht: (2025)
von: Fuest, Michael, et al.
Veröffentlicht: (2025)
DisMo: Disentangled Motion Representations for Open-World Motion Transfer
von: Ressler-Antal, Thomas, et al.
Veröffentlicht: (2025)
von: Ressler-Antal, Thomas, et al.
Veröffentlicht: (2025)
TREAD: Token Routing for Efficient Architecture-agnostic Diffusion Training
von: Krause, Felix, et al.
Veröffentlicht: (2025)
von: Krause, Felix, et al.
Veröffentlicht: (2025)
Envisioning the Future, One Step at a Time
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2026)
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2026)
Guiding Token-Sparse Diffusion Models
von: Krause, Felix, et al.
Veröffentlicht: (2026)
von: Krause, Felix, et al.
Veröffentlicht: (2026)
CAGE: Unsupervised Visual Composition and Animation for Controllable Video Generation
von: Davtyan, Aram, et al.
Veröffentlicht: (2024)
von: Davtyan, Aram, et al.
Veröffentlicht: (2024)
Unsupervised View-Invariant Human Posture Representation
von: Sardari, Faegheh, et al.
Veröffentlicht: (2021)
von: Sardari, Faegheh, et al.
Veröffentlicht: (2021)
[MASK] is All You Need
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
Lagrangian Motion Fields for Long-term Motion Generation
von: Yang, Yifei, et al.
Veröffentlicht: (2024)
von: Yang, Yifei, et al.
Veröffentlicht: (2024)
Pseudo-Generalized Dynamic View Synthesis from a Video
von: Zhao, Xiaoming, et al.
Veröffentlicht: (2023)
von: Zhao, Xiaoming, et al.
Veröffentlicht: (2023)
DepthFM: Fast Monocular Depth Estimation with Flow Matching
von: Gui, Ming, et al.
Veröffentlicht: (2024)
von: Gui, Ming, et al.
Veröffentlicht: (2024)
STARFlow-V: End-to-End Video Generative Modeling with Normalizing Flows
von: Gu, Jiatao, et al.
Veröffentlicht: (2025)
von: Gu, Jiatao, et al.
Veröffentlicht: (2025)
Stable-Pose: Leveraging Transformers for Pose-Guided Text-to-Image Generation
von: Wang, Jiajun, et al.
Veröffentlicht: (2024)
von: Wang, Jiajun, et al.
Veröffentlicht: (2024)
Normalizing Flows are Capable Generative Models
von: Zhai, Shuangfei, et al.
Veröffentlicht: (2024)
von: Zhai, Shuangfei, et al.
Veröffentlicht: (2024)
3D Shape Tokenization via Latent Flow Matching
von: Chang, Jen-Hao Rick, et al.
Veröffentlicht: (2024)
von: Chang, Jen-Hao Rick, et al.
Veröffentlicht: (2024)
Overcoming the Pitfalls of Vision-Language Model Finetuning for OOD Generalization
von: Zang, Yuhang, et al.
Veröffentlicht: (2024)
von: Zang, Yuhang, et al.
Veröffentlicht: (2024)
Motion Mamba: Efficient and Long Sequence Motion Generation
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2024)
Denoising, Fast and Slow: Difficulty-Aware Adaptive Sampling for Image Generation
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2026)
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2026)
Learning Context-Adaptive Motion Priors for Masked Motion Diffusion Models with Efficient Kinematic Attention Aggregation
von: Jiang, Junkun, et al.
Veröffentlicht: (2026)
von: Jiang, Junkun, et al.
Veröffentlicht: (2026)
Controllable Long-term Motion Generation with Extended Joint Targets
von: Lee, Eunjong, et al.
Veröffentlicht: (2025)
von: Lee, Eunjong, et al.
Veröffentlicht: (2025)
Distillation of Diffusion Features for Semantic Correspondence
von: Fundel, Frank, et al.
Veröffentlicht: (2024)
von: Fundel, Frank, et al.
Veröffentlicht: (2024)
ZigMa: A DiT-style Zigzag Mamba Diffusion Model
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
World-consistent Video Diffusion with Explicit 3D Modeling
von: Zhang, Qihang, et al.
Veröffentlicht: (2024)
von: Zhang, Qihang, et al.
Veröffentlicht: (2024)
Diff2Flow: Training Flow Matching Models via Diffusion Model Alignment
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2025)
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2025)
Text-Conditional JEPA for Learning Semantically Rich Visual Representations
von: Huang, Chen, et al.
Veröffentlicht: (2026)
von: Huang, Chen, et al.
Veröffentlicht: (2026)
Does VLM Classification Benefit from LLM Description Semantics?
von: Ma, Pingchuan, et al.
Veröffentlicht: (2024)
von: Ma, Pingchuan, et al.
Veröffentlicht: (2024)
AdvMT: Adversarial Motion Transformer for Long-term Human Motion Prediction
von: Idrees, Sarmad, et al.
Veröffentlicht: (2024)
von: Idrees, Sarmad, et al.
Veröffentlicht: (2024)
Aggregate-and-Adapt Natural Language Prompts for Downstream Generalization of CLIP
von: Huang, Chen, et al.
Veröffentlicht: (2024)
von: Huang, Chen, et al.
Veröffentlicht: (2024)
Diffusion Models and Representation Learning: A Survey
von: Fuest, Michael, et al.
Veröffentlicht: (2024)
von: Fuest, Michael, et al.
Veröffentlicht: (2024)
Dyadic Mamba: Long-term Dyadic Human Motion Synthesis
von: Tanke, Julian, et al.
Veröffentlicht: (2025)
von: Tanke, Julian, et al.
Veröffentlicht: (2025)
Scalable Pre-training of Large Autoregressive Image Models
von: El-Nouby, Alaaeldin, et al.
Veröffentlicht: (2024)
von: El-Nouby, Alaaeldin, et al.
Veröffentlicht: (2024)
STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation
von: Shen, Ying, et al.
Veröffentlicht: (2026)
von: Shen, Ying, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CTRLorALTer: Conditional LoRAdapter for Efficient 0-Shot Control & Altering of T2I Models
von: Stracke, Nick, et al.
Veröffentlicht: (2024) -
CleanDIFT: Diffusion Features without Noise
von: Stracke, Nick, et al.
Veröffentlicht: (2024) -
What If : Understanding Motion Through Sparse Interactions
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2025) -
RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video
von: Prestel, Ulrich, et al.
Veröffentlicht: (2026) -
Adapting Self-Supervised Representations as a Latent Space for Efficient Generation
von: Gui, Ming, et al.
Veröffentlicht: (2025)