Top2Ground: A Height-Aware Dual Conditioning Diffusion Model for Robust Aerial-to-Ground View Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Jae Joong, Benes, Bedrich |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RGB2Point: 3D Point Cloud Generation from Single RGB Images
by: Lee, Jae Joong, et al.
Published: (2024)
by: Lee, Jae Joong, et al.
Published: (2024)
Tuning-Free Amodal Segmentation via the Occlusion-Free Bias of Inpainting Models
by: Lee, Jae Joong, et al.
Published: (2025)
by: Lee, Jae Joong, et al.
Published: (2025)
WebAccessVL: Violation-Aware VLM for Web Accessibility
by: Zheng, Amber Yijia, et al.
Published: (2025)
by: Zheng, Amber Yijia, et al.
Published: (2025)
Tree-D Fusion: Simulation-Ready Tree Dataset from Single Images with Diffusion Priors
by: Lee, Jae Joong, et al.
Published: (2024)
by: Lee, Jae Joong, et al.
Published: (2024)
View-Aware Semantic Alignment for Aerial-Ground Person Re-Identification
by: Zhang, Quan, et al.
Published: (2026)
by: Zhang, Quan, et al.
Published: (2026)
AerialMegaDepth: Learning Aerial-Ground Reconstruction and View Synthesis
by: Vuong, Khiem, et al.
Published: (2025)
by: Vuong, Khiem, et al.
Published: (2025)
Skyeyes: Ground Roaming using Aerial View Images
by: Gao, Zhiyuan, et al.
Published: (2024)
by: Gao, Zhiyuan, et al.
Published: (2024)
SD-ReID: View-aware Stable Diffusion for Aerial-Ground Person Re-Identification
by: Wang, Yuhao, et al.
Published: (2025)
by: Wang, Yuhao, et al.
Published: (2025)
Language-Guided Invariance Probing of Vision-Language Models
by: Lee, Jae Joong
Published: (2025)
by: Lee, Jae Joong
Published: (2025)
AG-ReID.v2: Bridging Aerial and Ground Views for Person Re-identification
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
3DTurboQuant: Training-Free Near-Optimal Quantization for 3D Reconstruction Models
by: Lee, Jae Joong
Published: (2026)
by: Lee, Jae Joong
Published: (2026)
Satellite to GroundScape -- Large-scale Consistent Ground View Generation from Satellite Views
by: Xu, Ningli, et al.
Published: (2025)
by: Xu, Ningli, et al.
Published: (2025)
Paired-CSLiDAR: Height-Stratified Registration for Cross-Source Aerial-Ground LiDAR Pose Refinement
by: Hoover, Montana, et al.
Published: (2026)
by: Hoover, Montana, et al.
Published: (2026)
SparseGF: A Height-Aware Sparse Segmentation Framework with Context Compression for Robust Ground Filtering Across Urban to Natural Scenes
by: Qin, Nannan, et al.
Published: (2026)
by: Qin, Nannan, et al.
Published: (2026)
GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis
by: Kang, Minjun, et al.
Published: (2026)
by: Kang, Minjun, et al.
Published: (2026)
Geospecific View Generation -- Geometry-Context Aware High-resolution Ground View Inference from Satellite Views
by: Xu, Ningli, et al.
Published: (2024)
by: Xu, Ningli, et al.
Published: (2024)
Generate to Ground: Multimodal Text Conditioning Boosts Phrase Grounding in Medical Vision-Language Models
by: Nützel, Felix, et al.
Published: (2025)
by: Nützel, Felix, et al.
Published: (2025)
Vision Foundation Models for Domain Generalisable Cross-View Localisation in Planetary Ground-Aerial Robotic Teams
by: Holden, Lachlan, et al.
Published: (2026)
by: Holden, Lachlan, et al.
Published: (2026)
ProDiG: Progressive Diffusion-Guided Gaussian Splatting for Aerial to Ground Reconstruction
by: Mitra, Sirshapan, et al.
Published: (2026)
by: Mitra, Sirshapan, et al.
Published: (2026)
Open-Text Aerial Detection: A Unified Framework For Aerial Visual Grounding And Detection
by: Wei, Guoting, et al.
Published: (2026)
by: Wei, Guoting, et al.
Published: (2026)
GVDIFF: Grounded Text-to-Video Generation with Diffusion Models
by: Dou, Huanzhang, et al.
Published: (2024)
by: Dou, Huanzhang, et al.
Published: (2024)
MAG-VLAQ: Multi-modal Aerial-Ground Query Aggregation for Cross-View Place Recognition
by: Xu, Zhengyi, et al.
Published: (2026)
by: Xu, Zhengyi, et al.
Published: (2026)
Enhancing Novel View Synthesis via Geometry Grounded Set Diffusion
by: Zanjani, Farhad G., et al.
Published: (2026)
by: Zanjani, Farhad G., et al.
Published: (2026)
Leveraging BEV Paradigm for Ground-to-Aerial Image Synthesis
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
Aerial-Ground Image Feature Matching via 3D Gaussian Splatting-based Intermediate View Rendering
by: Yu, Jiangxue, et al.
Published: (2025)
by: Yu, Jiangxue, et al.
Published: (2025)
HOTFormerLoc: Hierarchical Octree Transformer for Versatile Lidar Place Recognition Across Ground and Aerial Views
by: Griffiths, Ethan, et al.
Published: (2025)
by: Griffiths, Ethan, et al.
Published: (2025)
Robust Mesh Saliency Ground Truth Acquisition in VR via View Cone Sampling and Manifold Diffusion
by: Zheng, Guoquan, et al.
Published: (2026)
by: Zheng, Guoquan, et al.
Published: (2026)
Top-Down Framework for Weakly-supervised Grounded Image Captioning
by: Cai, Chen, et al.
Published: (2023)
by: Cai, Chen, et al.
Published: (2023)
Empower Words: DualGround for Structured Phrase and Sentence-Level Temporal Grounding
by: Kang, Minseok, et al.
Published: (2025)
by: Kang, Minseok, et al.
Published: (2025)
Landscape-Awareness for Geometric View Diffusion Model
by: Chen, Yan-Ting, et al.
Published: (2026)
by: Chen, Yan-Ting, et al.
Published: (2026)
FloraForge: LLM-Assisted Procedural Generation of Editable and Analysis-Ready 3D Plant Geometric Models For Agricultural Applications
by: Hadadi, Mozhgan, et al.
Published: (2025)
by: Hadadi, Mozhgan, et al.
Published: (2025)
Dynamic Token Selection for Aerial-Ground Person Re-Identification
by: Wang, Yuhai, et al.
Published: (2024)
by: Wang, Yuhai, et al.
Published: (2024)
NuGrounding: A Multi-View 3D Visual Grounding Framework in Autonomous Driving
by: Li, Fuhao, et al.
Published: (2025)
by: Li, Fuhao, et al.
Published: (2025)
Top2Pano: Learning to Generate Indoor Panoramas from Top-Down View
by: Zhang, Zitong, et al.
Published: (2025)
by: Zhang, Zitong, et al.
Published: (2025)
RetinexDualV2: Physically-Grounded Dual Retinex for Generalized UHD Image Restoration
by: Kishawy, Mohab, et al.
Published: (2026)
by: Kishawy, Mohab, et al.
Published: (2026)
Joint Generative Modeling of Grounded Scene Graphs and Images via Diffusion Models
by: Xu, Bicheng, et al.
Published: (2024)
by: Xu, Bicheng, et al.
Published: (2024)
OracleGS: Grounding Generative Priors for Sparse-View Gaussian Splatting
by: Topaloglu, Atakan, et al.
Published: (2025)
by: Topaloglu, Atakan, et al.
Published: (2025)
ConGeo: Robust Cross-view Geo-localization across Ground View Variations
by: Mi, Li, et al.
Published: (2024)
by: Mi, Li, et al.
Published: (2024)
Text-based Aerial-Ground Person Retrieval
by: Zhou, Xinyu, et al.
Published: (2025)
by: Zhou, Xinyu, et al.
Published: (2025)
Toward Ambulatory Vision: Learning Visually-Grounded Active View Selection
by: Koo, Juil, et al.
Published: (2025)
by: Koo, Juil, et al.
Published: (2025)
Similar Items
-
RGB2Point: 3D Point Cloud Generation from Single RGB Images
by: Lee, Jae Joong, et al.
Published: (2024) -
Tuning-Free Amodal Segmentation via the Occlusion-Free Bias of Inpainting Models
by: Lee, Jae Joong, et al.
Published: (2025) -
WebAccessVL: Violation-Aware VLM for Web Accessibility
by: Zheng, Amber Yijia, et al.
Published: (2025) -
Tree-D Fusion: Simulation-Ready Tree Dataset from Single Images with Diffusion Priors
by: Lee, Jae Joong, et al.
Published: (2024) -
View-Aware Semantic Alignment for Aerial-Ground Person Re-Identification
by: Zhang, Quan, et al.
Published: (2026)