Aligning Latent Geometry for Spherical Flow Matching in Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Meral, Tuna Han Salih, Oktay, Kaan, Yesiltepe, Hidir, Akan, Adil Kaan, Yanardag, Pinar |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Infinity-RoPE: Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout
by: Yesiltepe, Hidir, et al.
Published: (2025)
by: Yesiltepe, Hidir, et al.
Published: (2025)
VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion
by: Yesiltepe, Hidir, et al.
Published: (2026)
by: Yesiltepe, Hidir, et al.
Published: (2026)
MotionFlow: Attention-Driven Motion Transfer in Video Diffusion Models
by: Meral, Tuna Han Salih, et al.
Published: (2024)
by: Meral, Tuna Han Salih, et al.
Published: (2024)
MotionShop: Zero-Shot Motion Transfer in Video Diffusion Models with Mixture of Score Guidance
by: Yesiltepe, Hidir, et al.
Published: (2024)
by: Yesiltepe, Hidir, et al.
Published: (2024)
LoRAShop: Training-Free Multi-Concept Image Generation and Editing with Rectified Flow Transformers
by: Dalva, Yusuf, et al.
Published: (2025)
by: Dalva, Yusuf, et al.
Published: (2025)
Dynamic View Synthesis as an Inverse Problem
by: Yesiltepe, Hidir, et al.
Published: (2025)
by: Yesiltepe, Hidir, et al.
Published: (2025)
GANTASTIC: GAN-based Transfer of Interpretable Directions for Disentangled Image Editing in Text-to-Image Diffusion Models
by: Dalva, Yusuf, et al.
Published: (2024)
by: Dalva, Yusuf, et al.
Published: (2024)
MIST: Mitigating Intersectional Bias with Disentangled Cross-Attention Editing in Text-to-Image Diffusion Models
by: Yesiltepe, Hidir, et al.
Published: (2024)
by: Yesiltepe, Hidir, et al.
Published: (2024)
The Curious Case of End Token: A Zero-Shot Disentangled Image Editing using CLIP
by: Yesiltepe, Hidir, et al.
Published: (2024)
by: Yesiltepe, Hidir, et al.
Published: (2024)
Contrastive Test-Time Composition of Multiple LoRA Models for Image Generation
by: Meral, Tuna Han Salih, et al.
Published: (2024)
by: Meral, Tuna Han Salih, et al.
Published: (2024)
Learning Object-Centric Representations Based on Slots in Real World Scenarios
by: Akan, Adil Kaan
Published: (2025)
by: Akan, Adil Kaan
Published: (2025)
Stylebreeder: Exploring and Democratizing Artistic Styles through Text-to-Image Models
by: Zheng, Matthew, et al.
Published: (2024)
by: Zheng, Matthew, et al.
Published: (2024)
Compositional Video Synthesis by Temporal Object-Centric Learning
by: Akan, Adil Kaan, et al.
Published: (2025)
by: Akan, Adil Kaan, et al.
Published: (2025)
Slot-Guided Adaptation of Pre-trained Diffusion Models for Object-Centric Learning and Compositional Generation
by: Akan, Adil Kaan, et al.
Published: (2025)
by: Akan, Adil Kaan, et al.
Published: (2025)
ConceptAttention: Diffusion Transformers Learn Highly Interpretable Features
by: Helbling, Alec, et al.
Published: (2025)
by: Helbling, Alec, et al.
Published: (2025)
DTG-Restore: Training-Free Diffusion Refinement for Generative Video Super-Resolution
by: Yesiltepe, Hidir, et al.
Published: (2026)
by: Yesiltepe, Hidir, et al.
Published: (2026)
Conditional Information Gain Trellis
by: Bicici, Ufuk Can, et al.
Published: (2024)
by: Bicici, Ufuk Can, et al.
Published: (2024)
AdaState: Self-Evolving Anchors for Streaming Video Generation
by: Dalva, Yusuf, et al.
Published: (2026)
by: Dalva, Yusuf, et al.
Published: (2026)
CREA: A Collaborative Multi-Agent Framework for Creative Image Editing and Generation
by: Venkatesh, Kavana, et al.
Published: (2025)
by: Venkatesh, Kavana, et al.
Published: (2025)
FluxSpace: Disentangled Semantic Editing in Rectified Flow Transformers
by: Dalva, Yusuf, et al.
Published: (2024)
by: Dalva, Yusuf, et al.
Published: (2024)
From Zero to Hero: Training-Free Custom Concept Spawning in World Models
by: Akdemir, Kiymet, et al.
Published: (2026)
by: Akdemir, Kiymet, et al.
Published: (2026)
ORACLE: Leveraging Mutual Information for Consistent Character Generation with LoRAs in Diffusion Models
by: Akdemir, Kiymet, et al.
Published: (2024)
by: Akdemir, Kiymet, et al.
Published: (2024)
Audit & Repair: An Agentic Framework for Consistent Story Visualization in Text-to-Image Diffusion Models
by: Akdemir, Kiymet, et al.
Published: (2025)
by: Akdemir, Kiymet, et al.
Published: (2025)
Explaining in Diffusion: Explaining a Classifier Through Hierarchical Semantics with Text-to-Image Diffusion Models
by: Kazimi, Tahira, et al.
Published: (2024)
by: Kazimi, Tahira, et al.
Published: (2024)
Diverse Video Generation with Determinantal Point Process-Guided Policy Optimization
by: Kazimi, Tahira, et al.
Published: (2025)
by: Kazimi, Tahira, et al.
Published: (2025)
LoRAverse: A Submodular Framework to Retrieve Diverse Adapters for Diffusion Models
by: Sonmezer, Mert, et al.
Published: (2025)
by: Sonmezer, Mert, et al.
Published: (2025)
SAR-to-RGB Translation with Latent Diffusion for Earth Observation
by: Aydin, Kaan, et al.
Published: (2025)
by: Aydin, Kaan, et al.
Published: (2025)
Personalized Image Editing in Text-to-Image Diffusion Models via Collaborative Direct Preference Optimization
by: Dunlop, Connor, et al.
Published: (2025)
by: Dunlop, Connor, et al.
Published: (2025)
RAVEL: Rare Concept Generation and Editing via Graph-driven Relational Guidance
by: Venkatesh, Kavana, et al.
Published: (2024)
by: Venkatesh, Kavana, et al.
Published: (2024)
Edge Reliability Gap in Vision-Language Models: Quantifying Failure Modes of Compressed VLMs Under Visual Corruption
by: Erol, Mehmet Kaan
Published: (2026)
by: Erol, Mehmet Kaan
Published: (2026)
LayerFusion: Harmonized Multi-Layer Text-to-Image Generation with Generative Priors
by: Dalva, Yusuf, et al.
Published: (2024)
by: Dalva, Yusuf, et al.
Published: (2024)
Plot'n Polish: Zero-shot Story Visualization and Disentangled Editing with Text-to-Image Diffusion Models
by: Akdemir, Kiymet, et al.
Published: (2025)
by: Akdemir, Kiymet, et al.
Published: (2025)
Geometry-Aware Image Flow Matching
by: Lee, Junho, et al.
Published: (2026)
by: Lee, Junho, et al.
Published: (2026)
LoRACLR: Contrastive Adaptation for Customization of Diffusion Models
by: Simsar, Enis, et al.
Published: (2024)
by: Simsar, Enis, et al.
Published: (2024)
Editing Physiological Signals in Videos Using Latent Representations
by: Zhou, Tianwen, et al.
Published: (2025)
by: Zhou, Tianwen, et al.
Published: (2025)
LatentFM: A Latent Flow Matching Approach for Generative Medical Image Segmentation
by: Ngoc, Huynh Trinh, et al.
Published: (2025)
by: Ngoc, Huynh Trinh, et al.
Published: (2025)
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
by: Dalva, Yusuf, et al.
Published: (2025)
by: Dalva, Yusuf, et al.
Published: (2025)
Geometry Fidelity for Spherical Images
by: Christensen, Anders, et al.
Published: (2024)
by: Christensen, Anders, et al.
Published: (2024)
FlowIID: Single-Step Intrinsic Image Decomposition via Latent Flow Matching
by: Singla, Mithlesh, et al.
Published: (2026)
by: Singla, Mithlesh, et al.
Published: (2026)
RAW-Flow: Advancing RGB-to-RAW Image Reconstruction with Deterministic Latent Flow Matching
by: Liu, Zhen, et al.
Published: (2026)
by: Liu, Zhen, et al.
Published: (2026)
Similar Items
-
Infinity-RoPE: Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout
by: Yesiltepe, Hidir, et al.
Published: (2025) -
VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion
by: Yesiltepe, Hidir, et al.
Published: (2026) -
MotionFlow: Attention-Driven Motion Transfer in Video Diffusion Models
by: Meral, Tuna Han Salih, et al.
Published: (2024) -
MotionShop: Zero-Shot Motion Transfer in Video Diffusion Models with Mixture of Score Guidance
by: Yesiltepe, Hidir, et al.
Published: (2024) -
LoRAShop: Training-Free Multi-Concept Image Generation and Editing with Rectified Flow Transformers
by: Dalva, Yusuf, et al.
Published: (2025)