Cross-View Meets Diffusion: Aerial Image Synthesis with Geometry and Text Guidance
Fuente:
arXiv
Guardado en:
| Autores principales: | Arrabi, Ahmad, Zhang, Xiaohan, Sultani, Waqas, Chen, Chen, Wshah, Safwan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GeoFlow: Real-Time Fine-Grained Cross-View Geolocalization via Iterative Flow Prediction
por: Lehyeh, Ayesh Abu, et al.
Publicado: (2026)
por: Lehyeh, Ayesh Abu, et al.
Publicado: (2026)
Turning Generators into Retrievers: Unlocking MLLMs for Natural Language-Guided Geo-Localization
por: Chen, Yuqi, et al.
Publicado: (2026)
por: Chen, Yuqi, et al.
Publicado: (2026)
GeoDTR+: Toward generic cross-view geolocalization via geometric disentanglement
por: Zhang, Xiaohan, et al.
Publicado: (2023)
por: Zhang, Xiaohan, et al.
Publicado: (2023)
Geo$^\textbf{2}$: Geometry-Guided Cross-view Geo-Localization and Image Synthesis
por: Zhang, Yancheng, et al.
Publicado: (2026)
por: Zhang, Yancheng, et al.
Publicado: (2026)
Autonomous Skeletal Landmark Localization towards Agentic C-Arm Control
por: Jung, Jay, et al.
Publicado: (2026)
por: Jung, Jay, et al.
Publicado: (2026)
VICI: VLM-Instructed Cross-view Image-localisation
por: Zhang, Xiaohan, et al.
Publicado: (2025)
por: Zhang, Xiaohan, et al.
Publicado: (2025)
Automated C-Arm Positioning via Conformal Landmark Localization
por: Arrabi, Ahmad, et al.
Publicado: (2025)
por: Arrabi, Ahmad, et al.
Publicado: (2025)
C-arm Guidance: A Self-supervised Approach To Automated Positioning During Stroke Thrombectomy
por: Arrabi, Ahmad, et al.
Publicado: (2025)
por: Arrabi, Ahmad, et al.
Publicado: (2025)
Self-Cross Diffusion Guidance for Text-to-Image Synthesis of Similar Subjects
por: Qiu, Weimin, et al.
Publicado: (2024)
por: Qiu, Weimin, et al.
Publicado: (2024)
MIAdapt: Source-free Few-shot Domain Adaptive Object Detection for Microscopic Images
por: Dilawar, Nimra, et al.
Publicado: (2025)
por: Dilawar, Nimra, et al.
Publicado: (2025)
CountDiffusion: Text-to-Image Synthesis with Training-Free Counting-Guidance Diffusion
por: Li, Yanyu, et al.
Publicado: (2025)
por: Li, Yanyu, et al.
Publicado: (2025)
Few-Shot Domain Adaptive Object Detection for Microscopic Images
por: Inayat, Sumayya, et al.
Publicado: (2024)
por: Inayat, Sumayya, et al.
Publicado: (2024)
Joint Stream: Malignant Region Learning for Breast Cancer Diagnosis
por: Rehman, Abdul, et al.
Publicado: (2024)
por: Rehman, Abdul, et al.
Publicado: (2024)
OrthoPhys: Physically Plausible Video Generation with Orthogonal-View Geometry Guidance
por: Wang, Cong, et al.
Publicado: (2026)
por: Wang, Cong, et al.
Publicado: (2026)
Aligned Novel View Image and Geometry Synthesis via Cross-modal Attention Instillation
por: Kwak, Min-Seop, et al.
Publicado: (2025)
por: Kwak, Min-Seop, et al.
Publicado: (2025)
Geometry-guided Cross-view Diffusion for One-to-many Cross-view Image Synthesis
por: Lin, Tao Jun, et al.
Publicado: (2024)
por: Lin, Tao Jun, et al.
Publicado: (2024)
G-NeRF: Geometry-enhanced Novel View Synthesis from Single-View Images
por: Huang, Zixiong, et al.
Publicado: (2024)
por: Huang, Zixiong, et al.
Publicado: (2024)
A Dense Reward View on Aligning Text-to-Image Diffusion with Preference
por: Yang, Shentao, et al.
Publicado: (2024)
por: Yang, Shentao, et al.
Publicado: (2024)
Isolated Diffusion: Optimizing Multi-Concept Text-to-Image Generation Training-Freely with Isolated Diffusion Guidance
por: Zhu, Jingyuan, et al.
Publicado: (2024)
por: Zhu, Jingyuan, et al.
Publicado: (2024)
HawkI: Homography & Mutual Information Guidance for 3D-free Single Image to Aerial View
por: Kothandaraman, Divya, et al.
Publicado: (2023)
por: Kothandaraman, Divya, et al.
Publicado: (2023)
Counting Guidance for High Fidelity Text-to-Image Synthesis
por: Kang, Wonjun, et al.
Publicado: (2023)
por: Kang, Wonjun, et al.
Publicado: (2023)
Skyeyes: Ground Roaming using Aerial View Images
por: Gao, Zhiyuan, et al.
Publicado: (2024)
por: Gao, Zhiyuan, et al.
Publicado: (2024)
Segmentation-Free Guidance for Text-to-Image Diffusion Models
por: Azarian, Kambiz, et al.
Publicado: (2024)
por: Azarian, Kambiz, et al.
Publicado: (2024)
CrossViewDiff: A Cross-View Diffusion Model for Satellite-to-Street View Synthesis
por: Li, Weijia, et al.
Publicado: (2024)
por: Li, Weijia, et al.
Publicado: (2024)
AerialMegaDepth: Learning Aerial-Ground Reconstruction and View Synthesis
por: Vuong, Khiem, et al.
Publicado: (2025)
por: Vuong, Khiem, et al.
Publicado: (2025)
Manifold-Optimal Guidance: A Unified Riemannian Control View of Diffusion Guidance
por: Jia, Zexi, et al.
Publicado: (2026)
por: Jia, Zexi, et al.
Publicado: (2026)
DreamView: Injecting View-specific Text Guidance into Text-to-3D Generation
por: Yan, Junkai, et al.
Publicado: (2024)
por: Yan, Junkai, et al.
Publicado: (2024)
Injecting Image Guidance into Text-Conditioned Diffusion Models at Inference
por: Żywot, Agata, et al.
Publicado: (2026)
por: Żywot, Agata, et al.
Publicado: (2026)
Uni-Hema: Unified Model for Digital Hematopathology
por: Rehman, Abdul, et al.
Publicado: (2025)
por: Rehman, Abdul, et al.
Publicado: (2025)
GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis
por: Kang, Minjun, et al.
Publicado: (2026)
por: Kang, Minjun, et al.
Publicado: (2026)
Enhancing Novel View Synthesis via Geometry Grounded Set Diffusion
por: Zanjani, Farhad G., et al.
Publicado: (2026)
por: Zanjani, Farhad G., et al.
Publicado: (2026)
WAVE: Warp-Based View Guidance for Consistent Novel View Synthesis Using a Single Image
por: Park, Jiwoo, et al.
Publicado: (2025)
por: Park, Jiwoo, et al.
Publicado: (2025)
Structural Energy Guidance for View-Consistent Text-to-3D Generation
por: Zhang, Qing, et al.
Publicado: (2026)
por: Zhang, Qing, et al.
Publicado: (2026)
EarthBridge: A Solution for 4th Multi-modal Aerial View Image Challenge Translation Track
por: Chen, Zhenyuan, et al.
Publicado: (2026)
por: Chen, Zhenyuan, et al.
Publicado: (2026)
Tuning-Free Image Customization with Image and Text Guidance
por: Li, Pengzhi, et al.
Publicado: (2024)
por: Li, Pengzhi, et al.
Publicado: (2024)
Unbiased Image Synthesis via Manifold Guidance in Diffusion Models
por: Su, Xingzhe, et al.
Publicado: (2023)
por: Su, Xingzhe, et al.
Publicado: (2023)
Plasticine3D: 3D Non-Rigid Editing with Text Guidance by Multi-View Embedding Optimization
por: Chen, Yige, et al.
Publicado: (2023)
por: Chen, Yige, et al.
Publicado: (2023)
Consistent-1-to-3: Consistent Image to 3D View Synthesis via Geometry-aware Diffusion Models
por: Ye, Jianglong, et al.
Publicado: (2023)
por: Ye, Jianglong, et al.
Publicado: (2023)
Aerial-NeRF: Adaptive Spatial Partitioning and Sampling for Large-Scale Aerial Rendering
por: Zhang, Xiaohan, et al.
Publicado: (2024)
por: Zhang, Xiaohan, et al.
Publicado: (2024)
ICG-MVSNet: Learning Intra-view and Cross-view Relationships for Guidance in Multi-View Stereo
por: Hu, Yuxi, et al.
Publicado: (2025)
por: Hu, Yuxi, et al.
Publicado: (2025)
Ejemplares similares
-
GeoFlow: Real-Time Fine-Grained Cross-View Geolocalization via Iterative Flow Prediction
por: Lehyeh, Ayesh Abu, et al.
Publicado: (2026) -
Turning Generators into Retrievers: Unlocking MLLMs for Natural Language-Guided Geo-Localization
por: Chen, Yuqi, et al.
Publicado: (2026) -
GeoDTR+: Toward generic cross-view geolocalization via geometric disentanglement
por: Zhang, Xiaohan, et al.
Publicado: (2023) -
Geo$^\textbf{2}$: Geometry-Guided Cross-view Geo-Localization and Image Synthesis
por: Zhang, Yancheng, et al.
Publicado: (2026) -
Autonomous Skeletal Landmark Localization towards Agentic C-Arm Control
por: Jung, Jay, et al.
Publicado: (2026)