Saved in:
| Main Authors: | Stracke, Nick, Baumann, Stefan Andreas, Susskind, Joshua M., Bautista, Miguel Angel, Ommer, Björn |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.07913 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Long-term Motion Embeddings for Efficient Kinematics Generation
by: Stracke, Nick, et al.
Published: (2026)
by: Stracke, Nick, et al.
Published: (2026)
RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video
by: Prestel, Ulrich, et al.
Published: (2026)
by: Prestel, Ulrich, et al.
Published: (2026)
What If : Understanding Motion Through Sparse Interactions
by: Baumann, Stefan Andreas, et al.
Published: (2025)
by: Baumann, Stefan Andreas, et al.
Published: (2025)
CleanDIFT: Diffusion Features without Noise
by: Stracke, Nick, et al.
Published: (2024)
by: Stracke, Nick, et al.
Published: (2024)
Continuous, Subject-Specific Attribute Control in T2I Models by Identifying Semantic Directions
by: Baumann, Stefan Andreas, et al.
Published: (2024)
by: Baumann, Stefan Andreas, et al.
Published: (2024)
Probabilistic Precipitation Nowcasting with Rectified Flow Transformers
by: Schusterbauer, Johannes, et al.
Published: (2026)
by: Schusterbauer, Johannes, et al.
Published: (2026)
Adapting Self-Supervised Representations as a Latent Space for Efficient Generation
by: Gui, Ming, et al.
Published: (2025)
by: Gui, Ming, et al.
Published: (2025)
Boosting Latent Diffusion with Flow Matching
by: Schusterbauer, Johannes, et al.
Published: (2023)
by: Schusterbauer, Johannes, et al.
Published: (2023)
TREAD: Token Routing for Efficient Architecture-agnostic Diffusion Training
by: Krause, Felix, et al.
Published: (2025)
by: Krause, Felix, et al.
Published: (2025)
Envisioning the Future, One Step at a Time
by: Baumann, Stefan Andreas, et al.
Published: (2026)
by: Baumann, Stefan Andreas, et al.
Published: (2026)
Manifold Diffusion Fields
by: Elhag, Ahmed A., et al.
Published: (2023)
by: Elhag, Ahmed A., et al.
Published: (2023)
ActAlign: Zero-Shot Fine-Grained Video Classification via Language-Guided Sequence Alignment
by: Aghdam, Amir, et al.
Published: (2025)
by: Aghdam, Amir, et al.
Published: (2025)
EDGS: Eliminating Densification for Efficient Convergence of 3DGS
by: Kotovenko, Dmytro, et al.
Published: (2025)
by: Kotovenko, Dmytro, et al.
Published: (2025)
INRFlow: Flow Matching for INRs in Ambient Space
by: Wang, Yuyang, et al.
Published: (2024)
by: Wang, Yuyang, et al.
Published: (2024)
Swallowing the Bitter Pill: Simplified Scalable Conformer Generation
by: Wang, Yuyang, et al.
Published: (2023)
by: Wang, Yuyang, et al.
Published: (2023)
ZigMa: A DiT-style Zigzag Mamba Diffusion Model
by: Hu, Vincent Tao, et al.
Published: (2024)
by: Hu, Vincent Tao, et al.
Published: (2024)
Guiding Token-Sparse Diffusion Models
by: Krause, Felix, et al.
Published: (2026)
by: Krause, Felix, et al.
Published: (2026)
MaskFlow: Discrete Flows For Flexible and Efficient Long Video Generation
by: Fuest, Michael, et al.
Published: (2025)
by: Fuest, Michael, et al.
Published: (2025)
DisMo: Disentangled Motion Representations for Open-World Motion Transfer
by: Ressler-Antal, Thomas, et al.
Published: (2025)
by: Ressler-Antal, Thomas, et al.
Published: (2025)
[MASK] is All You Need
by: Hu, Vincent Tao, et al.
Published: (2024)
by: Hu, Vincent Tao, et al.
Published: (2024)
Pseudo-Generalized Dynamic View Synthesis from a Video
by: Zhao, Xiaoming, et al.
Published: (2023)
by: Zhao, Xiaoming, et al.
Published: (2023)
World-consistent Video Diffusion with Explicit 3D Modeling
by: Zhang, Qihang, et al.
Published: (2024)
by: Zhang, Qihang, et al.
Published: (2024)
SimpleFold: Folding Proteins is Simpler than You Think
by: Wang, Yuyang, et al.
Published: (2025)
by: Wang, Yuyang, et al.
Published: (2025)
CAGE: Unsupervised Visual Composition and Animation for Controllable Video Generation
by: Davtyan, Aram, et al.
Published: (2024)
by: Davtyan, Aram, et al.
Published: (2024)
Die kalendarische Altersgrenze im Rentensystem: Willkür oder Gleichheit?
by: Stracke, Elmar
Published: (2025)
by: Stracke, Elmar
Published: (2025)
Benchmarking Deep Learning-Based Low-Dose CT Image Denoising Algorithms
by: Eulig, Elias, et al.
Published: (2024)
by: Eulig, Elias, et al.
Published: (2024)
Unsupervised View-Invariant Human Posture Representation
by: Sardari, Faegheh, et al.
Published: (2021)
by: Sardari, Faegheh, et al.
Published: (2021)
Benchmarking deep learning‐based low‐dose CT image denoising algorithms
by: Elias Eulig, et al.
Published: (2024)
by: Elias Eulig, et al.
Published: (2024)
DepthFM: Fast Monocular Depth Estimation with Flow Matching
by: Gui, Ming, et al.
Published: (2024)
by: Gui, Ming, et al.
Published: (2024)
Ethical Regulation of AI & Education (AI&ED): Needs and Benefits
by: Stracke, Christian M.
Published: (2025)
by: Stracke, Christian M.
Published: (2025)
Scalable Pre-training of Large Autoregressive Image Models
by: El-Nouby, Alaaeldin, et al.
Published: (2024)
by: El-Nouby, Alaaeldin, et al.
Published: (2024)
Diff2Flow: Training Flow Matching Models via Diffusion Model Alignment
by: Schusterbauer, Johannes, et al.
Published: (2025)
by: Schusterbauer, Johannes, et al.
Published: (2025)
STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation
by: Shen, Ying, et al.
Published: (2026)
by: Shen, Ying, et al.
Published: (2026)
Distillation of Diffusion Features for Semantic Correspondence
by: Fundel, Frank, et al.
Published: (2024)
by: Fundel, Frank, et al.
Published: (2024)
Stable-Pose: Leveraging Transformers for Pose-Guided Text-to-Image Generation
by: Wang, Jiajun, et al.
Published: (2024)
by: Wang, Jiajun, et al.
Published: (2024)
Reconstructing and analyzing the invariances of low‐dose CT image denoising networks
by: Elias Eulig, et al.
Published: (2024)
by: Elias Eulig, et al.
Published: (2024)
3D Shape Tokenization via Latent Flow Matching
by: Chang, Jen-Hao Rick, et al.
Published: (2024)
by: Chang, Jen-Hao Rick, et al.
Published: (2024)
Kaleido Diffusion: Improving Conditional Diffusion Models with Autoregressive Latent Modeling
by: Gu, Jiatao, et al.
Published: (2024)
by: Gu, Jiatao, et al.
Published: (2024)
Does VLM Classification Benefit from LLM Description Semantics?
by: Ma, Pingchuan, et al.
Published: (2024)
by: Ma, Pingchuan, et al.
Published: (2024)
Scaling Image Tokenizers with Grouped Spherical Quantization
by: Wang, Jiangtao, et al.
Published: (2024)
by: Wang, Jiangtao, et al.
Published: (2024)
Similar Items
-
Learning Long-term Motion Embeddings for Efficient Kinematics Generation
by: Stracke, Nick, et al.
Published: (2026) -
RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video
by: Prestel, Ulrich, et al.
Published: (2026) -
What If : Understanding Motion Through Sparse Interactions
by: Baumann, Stefan Andreas, et al.
Published: (2025) -
CleanDIFT: Diffusion Features without Noise
by: Stracke, Nick, et al.
Published: (2024) -
Continuous, Subject-Specific Attribute Control in T2I Models by Identifying Semantic Directions
by: Baumann, Stefan Andreas, et al.
Published: (2024)