Cross-View Meets Diffusion: Aerial Image Synthesis with Geometry and Text Guidance
Fuente:
arXiv
Salvato in:
| Autori principali: | Arrabi, Ahmad, Zhang, Xiaohan, Sultani, Waqas, Chen, Chen, Wshah, Safwan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GeoFlow: Real-Time Fine-Grained Cross-View Geolocalization via Iterative Flow Prediction
di: Lehyeh, Ayesh Abu, et al.
Pubblicazione: (2026)
di: Lehyeh, Ayesh Abu, et al.
Pubblicazione: (2026)
Turning Generators into Retrievers: Unlocking MLLMs for Natural Language-Guided Geo-Localization
di: Chen, Yuqi, et al.
Pubblicazione: (2026)
di: Chen, Yuqi, et al.
Pubblicazione: (2026)
GeoDTR+: Toward generic cross-view geolocalization via geometric disentanglement
di: Zhang, Xiaohan, et al.
Pubblicazione: (2023)
di: Zhang, Xiaohan, et al.
Pubblicazione: (2023)
Geo$^\textbf{2}$: Geometry-Guided Cross-view Geo-Localization and Image Synthesis
di: Zhang, Yancheng, et al.
Pubblicazione: (2026)
di: Zhang, Yancheng, et al.
Pubblicazione: (2026)
Autonomous Skeletal Landmark Localization towards Agentic C-Arm Control
di: Jung, Jay, et al.
Pubblicazione: (2026)
di: Jung, Jay, et al.
Pubblicazione: (2026)
VICI: VLM-Instructed Cross-view Image-localisation
di: Zhang, Xiaohan, et al.
Pubblicazione: (2025)
di: Zhang, Xiaohan, et al.
Pubblicazione: (2025)
Automated C-Arm Positioning via Conformal Landmark Localization
di: Arrabi, Ahmad, et al.
Pubblicazione: (2025)
di: Arrabi, Ahmad, et al.
Pubblicazione: (2025)
C-arm Guidance: A Self-supervised Approach To Automated Positioning During Stroke Thrombectomy
di: Arrabi, Ahmad, et al.
Pubblicazione: (2025)
di: Arrabi, Ahmad, et al.
Pubblicazione: (2025)
Self-Cross Diffusion Guidance for Text-to-Image Synthesis of Similar Subjects
di: Qiu, Weimin, et al.
Pubblicazione: (2024)
di: Qiu, Weimin, et al.
Pubblicazione: (2024)
MIAdapt: Source-free Few-shot Domain Adaptive Object Detection for Microscopic Images
di: Dilawar, Nimra, et al.
Pubblicazione: (2025)
di: Dilawar, Nimra, et al.
Pubblicazione: (2025)
CountDiffusion: Text-to-Image Synthesis with Training-Free Counting-Guidance Diffusion
di: Li, Yanyu, et al.
Pubblicazione: (2025)
di: Li, Yanyu, et al.
Pubblicazione: (2025)
Few-Shot Domain Adaptive Object Detection for Microscopic Images
di: Inayat, Sumayya, et al.
Pubblicazione: (2024)
di: Inayat, Sumayya, et al.
Pubblicazione: (2024)
Joint Stream: Malignant Region Learning for Breast Cancer Diagnosis
di: Rehman, Abdul, et al.
Pubblicazione: (2024)
di: Rehman, Abdul, et al.
Pubblicazione: (2024)
OrthoPhys: Physically Plausible Video Generation with Orthogonal-View Geometry Guidance
di: Wang, Cong, et al.
Pubblicazione: (2026)
di: Wang, Cong, et al.
Pubblicazione: (2026)
Aligned Novel View Image and Geometry Synthesis via Cross-modal Attention Instillation
di: Kwak, Min-Seop, et al.
Pubblicazione: (2025)
di: Kwak, Min-Seop, et al.
Pubblicazione: (2025)
Geometry-guided Cross-view Diffusion for One-to-many Cross-view Image Synthesis
di: Lin, Tao Jun, et al.
Pubblicazione: (2024)
di: Lin, Tao Jun, et al.
Pubblicazione: (2024)
G-NeRF: Geometry-enhanced Novel View Synthesis from Single-View Images
di: Huang, Zixiong, et al.
Pubblicazione: (2024)
di: Huang, Zixiong, et al.
Pubblicazione: (2024)
A Dense Reward View on Aligning Text-to-Image Diffusion with Preference
di: Yang, Shentao, et al.
Pubblicazione: (2024)
di: Yang, Shentao, et al.
Pubblicazione: (2024)
Isolated Diffusion: Optimizing Multi-Concept Text-to-Image Generation Training-Freely with Isolated Diffusion Guidance
di: Zhu, Jingyuan, et al.
Pubblicazione: (2024)
di: Zhu, Jingyuan, et al.
Pubblicazione: (2024)
HawkI: Homography & Mutual Information Guidance for 3D-free Single Image to Aerial View
di: Kothandaraman, Divya, et al.
Pubblicazione: (2023)
di: Kothandaraman, Divya, et al.
Pubblicazione: (2023)
Counting Guidance for High Fidelity Text-to-Image Synthesis
di: Kang, Wonjun, et al.
Pubblicazione: (2023)
di: Kang, Wonjun, et al.
Pubblicazione: (2023)
Skyeyes: Ground Roaming using Aerial View Images
di: Gao, Zhiyuan, et al.
Pubblicazione: (2024)
di: Gao, Zhiyuan, et al.
Pubblicazione: (2024)
Segmentation-Free Guidance for Text-to-Image Diffusion Models
di: Azarian, Kambiz, et al.
Pubblicazione: (2024)
di: Azarian, Kambiz, et al.
Pubblicazione: (2024)
CrossViewDiff: A Cross-View Diffusion Model for Satellite-to-Street View Synthesis
di: Li, Weijia, et al.
Pubblicazione: (2024)
di: Li, Weijia, et al.
Pubblicazione: (2024)
AerialMegaDepth: Learning Aerial-Ground Reconstruction and View Synthesis
di: Vuong, Khiem, et al.
Pubblicazione: (2025)
di: Vuong, Khiem, et al.
Pubblicazione: (2025)
Manifold-Optimal Guidance: A Unified Riemannian Control View of Diffusion Guidance
di: Jia, Zexi, et al.
Pubblicazione: (2026)
di: Jia, Zexi, et al.
Pubblicazione: (2026)
DreamView: Injecting View-specific Text Guidance into Text-to-3D Generation
di: Yan, Junkai, et al.
Pubblicazione: (2024)
di: Yan, Junkai, et al.
Pubblicazione: (2024)
Injecting Image Guidance into Text-Conditioned Diffusion Models at Inference
di: Żywot, Agata, et al.
Pubblicazione: (2026)
di: Żywot, Agata, et al.
Pubblicazione: (2026)
Uni-Hema: Unified Model for Digital Hematopathology
di: Rehman, Abdul, et al.
Pubblicazione: (2025)
di: Rehman, Abdul, et al.
Pubblicazione: (2025)
GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis
di: Kang, Minjun, et al.
Pubblicazione: (2026)
di: Kang, Minjun, et al.
Pubblicazione: (2026)
Enhancing Novel View Synthesis via Geometry Grounded Set Diffusion
di: Zanjani, Farhad G., et al.
Pubblicazione: (2026)
di: Zanjani, Farhad G., et al.
Pubblicazione: (2026)
WAVE: Warp-Based View Guidance for Consistent Novel View Synthesis Using a Single Image
di: Park, Jiwoo, et al.
Pubblicazione: (2025)
di: Park, Jiwoo, et al.
Pubblicazione: (2025)
Structural Energy Guidance for View-Consistent Text-to-3D Generation
di: Zhang, Qing, et al.
Pubblicazione: (2026)
di: Zhang, Qing, et al.
Pubblicazione: (2026)
EarthBridge: A Solution for 4th Multi-modal Aerial View Image Challenge Translation Track
di: Chen, Zhenyuan, et al.
Pubblicazione: (2026)
di: Chen, Zhenyuan, et al.
Pubblicazione: (2026)
Tuning-Free Image Customization with Image and Text Guidance
di: Li, Pengzhi, et al.
Pubblicazione: (2024)
di: Li, Pengzhi, et al.
Pubblicazione: (2024)
Unbiased Image Synthesis via Manifold Guidance in Diffusion Models
di: Su, Xingzhe, et al.
Pubblicazione: (2023)
di: Su, Xingzhe, et al.
Pubblicazione: (2023)
Plasticine3D: 3D Non-Rigid Editing with Text Guidance by Multi-View Embedding Optimization
di: Chen, Yige, et al.
Pubblicazione: (2023)
di: Chen, Yige, et al.
Pubblicazione: (2023)
Consistent-1-to-3: Consistent Image to 3D View Synthesis via Geometry-aware Diffusion Models
di: Ye, Jianglong, et al.
Pubblicazione: (2023)
di: Ye, Jianglong, et al.
Pubblicazione: (2023)
Aerial-NeRF: Adaptive Spatial Partitioning and Sampling for Large-Scale Aerial Rendering
di: Zhang, Xiaohan, et al.
Pubblicazione: (2024)
di: Zhang, Xiaohan, et al.
Pubblicazione: (2024)
ICG-MVSNet: Learning Intra-view and Cross-view Relationships for Guidance in Multi-View Stereo
di: Hu, Yuxi, et al.
Pubblicazione: (2025)
di: Hu, Yuxi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GeoFlow: Real-Time Fine-Grained Cross-View Geolocalization via Iterative Flow Prediction
di: Lehyeh, Ayesh Abu, et al.
Pubblicazione: (2026) -
Turning Generators into Retrievers: Unlocking MLLMs for Natural Language-Guided Geo-Localization
di: Chen, Yuqi, et al.
Pubblicazione: (2026) -
GeoDTR+: Toward generic cross-view geolocalization via geometric disentanglement
di: Zhang, Xiaohan, et al.
Pubblicazione: (2023) -
Geo$^\textbf{2}$: Geometry-Guided Cross-view Geo-Localization and Image Synthesis
di: Zhang, Yancheng, et al.
Pubblicazione: (2026) -
Autonomous Skeletal Landmark Localization towards Agentic C-Arm Control
di: Jung, Jay, et al.
Pubblicazione: (2026)