FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing
Fuente:
arXiv
Saved in:
| Main Authors: | Cong, Yuren, Xu, Mengmeng, Simon, Christian, Chen, Shoufa, Ren, Jiawei, Xie, Yanping, Perez-Rua, Juan-Manuel, Rosenhahn, Bodo, Xiang, Tao, He, Sen |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GenTron: Diffusion Transformers for Image and Video Generation
by: Chen, Shoufa, et al.
Published: (2023)
by: Chen, Shoufa, et al.
Published: (2023)
SPAN: Learning Similarity between Scene Graphs and Images with Transformers
by: Cong, Yuren, et al.
Published: (2023)
by: Cong, Yuren, et al.
Published: (2023)
FDSG: Forecasting Dynamic Scene Graphs
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
Segment Any Object Model (SAOM): Real-to-Simulation Fine-Tuning Strategy for Multi-Class Multi-Instance Segmentation
by: Khan, Mariia, et al.
Published: (2024)
by: Khan, Mariia, et al.
Published: (2024)
HiStream: Efficient High-Resolution Video Generation via Redundancy-Eliminated Streaming
by: Qiu, Haonan, et al.
Published: (2025)
by: Qiu, Haonan, et al.
Published: (2025)
Improving 3D Foot Motion Reconstruction in Markerless Monocular Human Motion Capture
by: Wehrbein, Tom, et al.
Published: (2026)
by: Wehrbein, Tom, et al.
Published: (2026)
PARSAC: Accelerating Robust Multi-Model Fitting with Parallel Sample Consensus
by: Kluger, Florian, et al.
Published: (2024)
by: Kluger, Florian, et al.
Published: (2024)
Pruning by Block Benefit: Exploring the Properties of Vision Transformer Blocks during Domain Adaptation
by: Glandorf, Patrick, et al.
Published: (2025)
by: Glandorf, Patrick, et al.
Published: (2025)
Neural Random Forest Imitation
by: Reinders, Christoph, et al.
Published: (2019)
by: Reinders, Christoph, et al.
Published: (2019)
Multi-Flow: Multi-View-Enriched Normalizing Flows for Industrial Anomaly Detection
by: Kruse, Mathis, et al.
Published: (2025)
by: Kruse, Mathis, et al.
Published: (2025)
S4ConvD: Adaptive Scaling and Frequency Adjustment for Energy-Efficient Sensor Networks in Smart Buildings
by: Schaller, Melanie, et al.
Published: (2025)
by: Schaller, Melanie, et al.
Published: (2025)
Interpretable Decision-Making for End-to-End Autonomous Driving
by: Mirzaie, Mona, et al.
Published: (2025)
by: Mirzaie, Mona, et al.
Published: (2025)
Quantum Normalizing Flows for Anomaly Detection
by: Rosenhahn, Bodo, et al.
Published: (2024)
by: Rosenhahn, Bodo, et al.
Published: (2024)
Hyper-VolTran: Fast and Generalizable One-Shot Image to 3D Object Structure via HyperNetworks
by: Simon, Christian, et al.
Published: (2023)
by: Simon, Christian, et al.
Published: (2023)
Grouping Nodes With Known Value Differences: A Lossless UCT-based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Q-SENN: Quantized Self-Explaining Neural Networks
by: Norrenbrock, Thomas, et al.
Published: (2023)
by: Norrenbrock, Thomas, et al.
Published: (2023)
Mastering Zero-Shot Interactions in Cooperative and Competitive Simultaneous Games
by: Mahlau, Yannik, et al.
Published: (2024)
by: Mahlau, Yannik, et al.
Published: (2024)
AUPO -- Abstracted Until Proven Otherwise: A Reward Distribution Based Abstraction Algorithm
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Video Patch Pruning: Efficient Video Instance Segmentation via Early Token Reduction
by: Glandorf, Patrick, et al.
Published: (2026)
by: Glandorf, Patrick, et al.
Published: (2026)
Discovering State Equivalences in UCT Search Trees By Action Pruning
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
UncertainSAM: Fast and Efficient Uncertainty Quantification of the Segment Anything Model
by: Kaiser, Timo, et al.
Published: (2025)
by: Kaiser, Timo, et al.
Published: (2025)
A Rotation-Invariant Embedded Platform for (Neural) Cellular Automata
by: Woiwode, Dominik, et al.
Published: (2025)
by: Woiwode, Dominik, et al.
Published: (2025)
BUSSARD: Normalizing Flows for Bijective Universal Scene-Specific Anomalous Relationship Detection
by: Schween, Melissa, et al.
Published: (2026)
by: Schween, Melissa, et al.
Published: (2026)
Naga: Vedic Encoding for Deep State Space Models
by: Schaller, Melanie, et al.
Published: (2025)
by: Schaller, Melanie, et al.
Published: (2025)
Cell Tracking according to Biological Needs -- Strong Mitosis-aware Multi-Hypothesis Tracker with Aleatoric Uncertainty
by: Kaiser, Timo, et al.
Published: (2024)
by: Kaiser, Timo, et al.
Published: (2024)
Investigating Intra-Abstraction Policies For Non-exact Abstraction Algorithms
by: Schmöcker, Robin, et al.
Published: (2025)
by: Schmöcker, Robin, et al.
Published: (2025)
Numerical field optimization for enhanced efficiency in time-reversible gradient computation of open-source GPU-accelerated FDTD simulations
by: Mahlau, Yannik, et al.
Published: (2026)
by: Mahlau, Yannik, et al.
Published: (2026)
Benchmarking M-LTSF: Frequency and Noise-Based Evaluation of Multivariate Long Time Series Forecasting Models
by: Janssen, Nick, et al.
Published: (2025)
by: Janssen, Nick, et al.
Published: (2025)
DINO-QPM: Adapting Visual Foundation Models for Globally Interpretable Image Classification
by: Zimmermann, Robert, et al.
Published: (2026)
by: Zimmermann, Robert, et al.
Published: (2026)
HydraMix: Multi-Image Feature Mixing for Small Data Image Classification
by: Reinders, Christoph, et al.
Published: (2025)
by: Reinders, Christoph, et al.
Published: (2025)
CHOTA: A Higher Order Accuracy Metric for Cell Tracking
by: Kaiser, Timo, et al.
Published: (2024)
by: Kaiser, Timo, et al.
Published: (2024)
From Gameplay Traces to Game Mechanics: Causal Induction with Large Language Models
by: Jiwatode, Mohit, et al.
Published: (2026)
by: Jiwatode, Mohit, et al.
Published: (2026)
Learning Flow Fields in Attention for Controllable Person Image Generation
by: Zhou, Zijian, et al.
Published: (2024)
by: Zhou, Zijian, et al.
Published: (2024)
Optimization Driven Quantum Circuit Reduction
by: Rosenhahn, Bodo, et al.
Published: (2025)
by: Rosenhahn, Bodo, et al.
Published: (2025)
Neural Guided Sampling for Quantum Circuit Optimization
by: Rosenhahn, Bodo, et al.
Published: (2025)
by: Rosenhahn, Bodo, et al.
Published: (2025)
Stochastic Neural Networks for Quantum Devices
by: Rosenhahn, Bodo, et al.
Published: (2026)
by: Rosenhahn, Bodo, et al.
Published: (2026)
Hyper-parameter tuning for text guided image editing
by: Zhang, Shiwen
Published: (2024)
by: Zhang, Shiwen
Published: (2024)
Utilizing Uncertainty in 2D Pose Detectors for Probabilistic 3D Human Mesh Recovery
by: Wehrbein, Tom, et al.
Published: (2024)
by: Wehrbein, Tom, et al.
Published: (2024)
SplatPose & Detect: Pose-Agnostic 3D Anomaly Detection
by: Kruse, Mathis, et al.
Published: (2024)
by: Kruse, Mathis, et al.
Published: (2024)
Personalized 3D Human Pose and Shape Refinement
by: Wehrbein, Tom, et al.
Published: (2024)
by: Wehrbein, Tom, et al.
Published: (2024)
Similar Items
-
GenTron: Diffusion Transformers for Image and Video Generation
by: Chen, Shoufa, et al.
Published: (2023) -
SPAN: Learning Similarity between Scene Graphs and Images with Transformers
by: Cong, Yuren, et al.
Published: (2023) -
FDSG: Forecasting Dynamic Scene Graphs
by: Yang, Yi, et al.
Published: (2025) -
Segment Any Object Model (SAOM): Real-to-Simulation Fine-Tuning Strategy for Multi-Class Multi-Instance Segmentation
by: Khan, Mariia, et al.
Published: (2024) -
HiStream: Efficient High-Resolution Video Generation via Redundancy-Eliminated Streaming
by: Qiu, Haonan, et al.
Published: (2025)