What If : Understanding Motion Through Sparse Interactions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Baumann, Stefan Andreas, Stracke, Nick, Phan, Timy, Ommer, Björn |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Probabilistic Precipitation Nowcasting with Rectified Flow Transformers
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2026)
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2026)
RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video
von: Prestel, Ulrich, et al.
Veröffentlicht: (2026)
von: Prestel, Ulrich, et al.
Veröffentlicht: (2026)
CleanDIFT: Diffusion Features without Noise
von: Stracke, Nick, et al.
Veröffentlicht: (2024)
von: Stracke, Nick, et al.
Veröffentlicht: (2024)
Learning Long-term Motion Embeddings for Efficient Kinematics Generation
von: Stracke, Nick, et al.
Veröffentlicht: (2026)
von: Stracke, Nick, et al.
Veröffentlicht: (2026)
CTRLorALTer: Conditional LoRAdapter for Efficient 0-Shot Control & Altering of T2I Models
von: Stracke, Nick, et al.
Veröffentlicht: (2024)
von: Stracke, Nick, et al.
Veröffentlicht: (2024)
TREAD: Token Routing for Efficient Architecture-agnostic Diffusion Training
von: Krause, Felix, et al.
Veröffentlicht: (2025)
von: Krause, Felix, et al.
Veröffentlicht: (2025)
Boosting Latent Diffusion with Flow Matching
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2023)
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2023)
Continuous, Subject-Specific Attribute Control in T2I Models by Identifying Semantic Directions
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2024)
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2024)
Adapting Self-Supervised Representations as a Latent Space for Efficient Generation
von: Gui, Ming, et al.
Veröffentlicht: (2025)
von: Gui, Ming, et al.
Veröffentlicht: (2025)
Guiding Token-Sparse Diffusion Models
von: Krause, Felix, et al.
Veröffentlicht: (2026)
von: Krause, Felix, et al.
Veröffentlicht: (2026)
DisMo: Disentangled Motion Representations for Open-World Motion Transfer
von: Ressler-Antal, Thomas, et al.
Veröffentlicht: (2025)
von: Ressler-Antal, Thomas, et al.
Veröffentlicht: (2025)
Envisioning the Future, One Step at a Time
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2026)
von: Baumann, Stefan Andreas, et al.
Veröffentlicht: (2026)
Unsupervised View-Invariant Human Posture Representation
von: Sardari, Faegheh, et al.
Veröffentlicht: (2021)
von: Sardari, Faegheh, et al.
Veröffentlicht: (2021)
[MASK] is All You Need
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
MaskFlow: Discrete Flows For Flexible and Efficient Long Video Generation
von: Fuest, Michael, et al.
Veröffentlicht: (2025)
von: Fuest, Michael, et al.
Veröffentlicht: (2025)
DepthFM: Fast Monocular Depth Estimation with Flow Matching
von: Gui, Ming, et al.
Veröffentlicht: (2024)
von: Gui, Ming, et al.
Veröffentlicht: (2024)
CAGE: Unsupervised Visual Composition and Animation for Controllable Video Generation
von: Davtyan, Aram, et al.
Veröffentlicht: (2024)
von: Davtyan, Aram, et al.
Veröffentlicht: (2024)
ZigMa: A DiT-style Zigzag Mamba Diffusion Model
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
von: Hu, Vincent Tao, et al.
Veröffentlicht: (2024)
Distillation of Diffusion Features for Semantic Correspondence
von: Fundel, Frank, et al.
Veröffentlicht: (2024)
von: Fundel, Frank, et al.
Veröffentlicht: (2024)
Stable-Pose: Leveraging Transformers for Pose-Guided Text-to-Image Generation
von: Wang, Jiajun, et al.
Veröffentlicht: (2024)
von: Wang, Jiajun, et al.
Veröffentlicht: (2024)
Diff2Flow: Training Flow Matching Models via Diffusion Model Alignment
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2025)
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2025)
Does VLM Classification Benefit from LLM Description Semantics?
von: Ma, Pingchuan, et al.
Veröffentlicht: (2024)
von: Ma, Pingchuan, et al.
Veröffentlicht: (2024)
Denoising, Fast and Slow: Difficulty-Aware Adaptive Sampling for Image Generation
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2026)
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2026)
Scaling Image Tokenizers with Grouped Spherical Quantization
von: Wang, Jiangtao, et al.
Veröffentlicht: (2024)
von: Wang, Jiangtao, et al.
Veröffentlicht: (2024)
MAVFusion: Efficient Infrared and Visible Video Fusion via Motion-Aware Sparse Interaction
von: Li, Xilai, et al.
Veröffentlicht: (2026)
von: Li, Xilai, et al.
Veröffentlicht: (2026)
Motion2Motion: Cross-topology Motion Transfer with Sparse Correspondence
von: Chen, Ling-Hao, et al.
Veröffentlicht: (2025)
von: Chen, Ling-Hao, et al.
Veröffentlicht: (2025)
Vision At Night: Exploring Biologically Inspired Preprocessing For Improved Robustness Via Color And Contrast Transformations
von: Stracke, Lorena, et al.
Veröffentlicht: (2025)
von: Stracke, Lorena, et al.
Veröffentlicht: (2025)
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
von: Zhang, Wanyue, et al.
Veröffentlicht: (2025)
von: Zhang, Wanyue, et al.
Veröffentlicht: (2025)
Unimotion: Unifying 3D Human Motion Synthesis and Understanding
von: Li, Chuqiao, et al.
Veröffentlicht: (2024)
von: Li, Chuqiao, et al.
Veröffentlicht: (2024)
Quantum Denoising Diffusion Models
von: Kölle, Michael, et al.
Veröffentlicht: (2024)
von: Kölle, Michael, et al.
Veröffentlicht: (2024)
Diffusion Models and Representation Learning: A Survey
von: Fuest, Michael, et al.
Veröffentlicht: (2024)
von: Fuest, Michael, et al.
Veröffentlicht: (2024)
DISPLAY: Directable Human-Object Interaction Video Generation via Sparse Motion Guidance and Multi-Task Auxiliary
von: Guan, Jiazhi, et al.
Veröffentlicht: (2026)
von: Guan, Jiazhi, et al.
Veröffentlicht: (2026)
Diffusion Bridge Networks Simulate Clinical-grade PET from MRI for Dementia Diagnostics
von: Li, Yitong, et al.
Veröffentlicht: (2025)
von: Li, Yitong, et al.
Veröffentlicht: (2025)
Fast and Robust Normal Estimation for Sparse LiDAR Scans
von: Bogoslavskyi, Igor, et al.
Veröffentlicht: (2024)
von: Bogoslavskyi, Igor, et al.
Veröffentlicht: (2024)
Disentangling Latent Embeddings with Sparse Linear Concept Subspaces (SLiCS)
von: Li, Zhi, et al.
Veröffentlicht: (2025)
von: Li, Zhi, et al.
Veröffentlicht: (2025)
Vector Field Synthesis with Sparse Streamlines Using Diffusion Model
von: Phan, Nguyen K., et al.
Veröffentlicht: (2026)
von: Phan, Nguyen K., et al.
Veröffentlicht: (2026)
Unsupervised Semantic Segmentation Facilitates Model Understanding
von: Yu, Xiaoyan, et al.
Veröffentlicht: (2026)
von: Yu, Xiaoyan, et al.
Veröffentlicht: (2026)
From Sparse Signal to Smooth Motion: Real-Time Motion Generation with Rolling Prediction Models
von: Barquero, German, et al.
Veröffentlicht: (2025)
von: Barquero, German, et al.
Veröffentlicht: (2025)
MotionSight: Boosting Fine-Grained Motion Understanding in Multimodal LLMs
von: Du, Yipeng, et al.
Veröffentlicht: (2025)
von: Du, Yipeng, et al.
Veröffentlicht: (2025)
Encoder-Free Human Motion Understanding via Structured Motion Descriptions
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
von: Zhang, Yao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Probabilistic Precipitation Nowcasting with Rectified Flow Transformers
von: Schusterbauer, Johannes, et al.
Veröffentlicht: (2026) -
RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video
von: Prestel, Ulrich, et al.
Veröffentlicht: (2026) -
CleanDIFT: Diffusion Features without Noise
von: Stracke, Nick, et al.
Veröffentlicht: (2024) -
Learning Long-term Motion Embeddings for Efficient Kinematics Generation
von: Stracke, Nick, et al.
Veröffentlicht: (2026) -
CTRLorALTer: Conditional LoRAdapter for Efficient 0-Shot Control & Altering of T2I Models
von: Stracke, Nick, et al.
Veröffentlicht: (2024)