A Billion-scale Foundation Model for Remote Sensing Images
Fuente:
arXiv
Saved in:
| Main Authors: | Cha, Keumgang, Seo, Junghoon, Lee, Taekyung |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Pitfalls of $\textit{RemOve-And-Retrain}$: Data Processing Inequality Perspective
by: Song, Junhwa, et al.
Published: (2023)
by: Song, Junhwa, et al.
Published: (2023)
Pushing the Limits of Vision-Language Models in Remote Sensing without Human Annotations
by: Cha, Keumgang, et al.
Published: (2024)
by: Cha, Keumgang, et al.
Published: (2024)
SatVision-TOA: A Geospatial Foundation Model for Coarse-Resolution All-Sky Remote Sensing Imagery
by: Spradlin, Caleb S., et al.
Published: (2024)
by: Spradlin, Caleb S., et al.
Published: (2024)
StyleLipSync: Style-based Personalized Lip-sync Video Generation
by: Ki, Taekyung, et al.
Published: (2023)
by: Ki, Taekyung, et al.
Published: (2023)
Solid Waste Detection, Monitoring and Mapping in Remote Sensing Images: A Survey
by: Fraternali, Piero, et al.
Published: (2024)
by: Fraternali, Piero, et al.
Published: (2024)
Bridging Remote Sensors with Multisensor Geospatial Foundation Models
by: Han, Boran, et al.
Published: (2024)
by: Han, Boran, et al.
Published: (2024)
Aligning Text to Image in Diffusion Models is Easier Than You Think
by: Lee, Jaa-Yeon, et al.
Published: (2025)
by: Lee, Jaa-Yeon, et al.
Published: (2025)
Unsupervised Few-Shot Continual Learning for Remote Sensing Image Scene Classification
by: Ma'sum, Muhammad Anwar, et al.
Published: (2024)
by: Ma'sum, Muhammad Anwar, et al.
Published: (2024)
Seg-LSTM: Performance of xLSTM for Semantic Segmentation of Remotely Sensed Images
by: Zhu, Qinfeng, et al.
Published: (2024)
by: Zhu, Qinfeng, et al.
Published: (2024)
From Pixels to Prose: Advancing Multi-Modal Language Models for Remote Sensing
by: Sun, Xintian, et al.
Published: (2024)
by: Sun, Xintian, et al.
Published: (2024)
Mask Approximation Net: A Novel Diffusion Model Approach for Remote Sensing Change Captioning
by: Sun, Dongwei, et al.
Published: (2024)
by: Sun, Dongwei, et al.
Published: (2024)
LHRS-Bot: Empowering Remote Sensing with VGI-Enhanced Large Multimodal Language Model
by: Muhtar, Dilxat, et al.
Published: (2024)
by: Muhtar, Dilxat, et al.
Published: (2024)
Efficient Few-Shot Learning in Remote Sensing: Fusing Vision and Vision-Language Models
by: Chua, Jia Yun, et al.
Published: (2025)
by: Chua, Jia Yun, et al.
Published: (2025)
A Recipe for Improving Remote Sensing VLM Zero Shot Generalization
by: Barzilai, Aviad, et al.
Published: (2025)
by: Barzilai, Aviad, et al.
Published: (2025)
ZAYAN: Disentangled Contrastive Transformer for Tabular Remote Sensing Data
by: Habib, Al Zadid Sultan Bin, et al.
Published: (2026)
by: Habib, Al Zadid Sultan Bin, et al.
Published: (2026)
In the Search for Optimal Multi-view Learning Models for Crop Classification with Global Remote Sensing Data
by: Mena, Francisco, et al.
Published: (2024)
by: Mena, Francisco, et al.
Published: (2024)
Wavelet Latent Diffusion (Wala): Billion-Parameter 3D Generative Model with Compact Wavelet Encodings
by: Sanghi, Aditya, et al.
Published: (2024)
by: Sanghi, Aditya, et al.
Published: (2024)
ChangeAnywhere: Sample Generation for Remote Sensing Change Detection via Semantic Latent Diffusion Model
by: Tang, Kai, et al.
Published: (2024)
by: Tang, Kai, et al.
Published: (2024)
Hyperspherical Autoencoder for High-Fidelity Image Reconstruction and Generation
by: Chang, Hun, et al.
Published: (2026)
by: Chang, Hun, et al.
Published: (2026)
Rethinking Electro-Optical Vision Foundation Models for Remote Sensing Retrieval: A Controlled Comparison with Generalist VFM
by: Park, Hyobin, et al.
Published: (2026)
by: Park, Hyobin, et al.
Published: (2026)
A Foundation Model for General Moving Object Segmentation in Medical Images
by: Yan, Zhongnuo, et al.
Published: (2023)
by: Yan, Zhongnuo, et al.
Published: (2023)
Common Practices and Taxonomy in Deep Multi-view Fusion for Remote Sensing Applications
by: Mena, Francisco, et al.
Published: (2022)
by: Mena, Francisco, et al.
Published: (2022)
Probabilistic Wildfire Susceptibility from Remote Sensing Using Random Forests and SHAP
by: Cheerala, Udaya Bhasker, et al.
Published: (2025)
by: Cheerala, Udaya Bhasker, et al.
Published: (2025)
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation
by: Arkhipkin, Vladimir, et al.
Published: (2025)
by: Arkhipkin, Vladimir, et al.
Published: (2025)
Fair Foundation Models for Medical Image Analysis: Challenges and Perspectives
by: Queiroz, Dilermando, et al.
Published: (2025)
by: Queiroz, Dilermando, et al.
Published: (2025)
ViTok-v2: Scaling Native Resolution Auto-Encoders to 5 Billion Parameters
by: Hansen-Estruch, Philippe, et al.
Published: (2026)
by: Hansen-Estruch, Philippe, et al.
Published: (2026)
Beyond Visual Fidelity: Benchmarking Super-Resolution Models for Large-Scale Remote Sensing Imagery via Downstream Task Integration
by: Li, Zhili, et al.
Published: (2026)
by: Li, Zhili, et al.
Published: (2026)
PyroFocus: A Deep Learning Approach to Real-Time Wildfire Detection in Multispectral Remote Sensing Imagery
by: Moussa, Mark, et al.
Published: (2025)
by: Moussa, Mark, et al.
Published: (2025)
DarkVesselNet: Multi-Modal Remote Sensing and Trajectory Reasoning for Dark Vessel Detection
by: Sharma, Arun
Published: (2026)
by: Sharma, Arun
Published: (2026)
Evaluation and Analysis of Deep Neural Transformers and Convolutional Neural Networks on Modern Remote Sensing Datasets
by: Hurt, J. Alex, et al.
Published: (2025)
by: Hurt, J. Alex, et al.
Published: (2025)
Accessing Vision Foundation Models via ImageNet-1K
by: Zhang, Yitian, et al.
Published: (2024)
by: Zhang, Yitian, et al.
Published: (2024)
Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation
by: Seo, Hoigi, et al.
Published: (2025)
by: Seo, Hoigi, et al.
Published: (2025)
EXAONE Path 2.0: Pathology Foundation Model with End-to-End Supervision
by: Pyeon, Myeongjang, et al.
Published: (2025)
by: Pyeon, Myeongjang, et al.
Published: (2025)
Adaptive Fusion of Multi-view Remote Sensing data for Optimal Sub-field Crop Yield Prediction
by: Mena, Francisco, et al.
Published: (2024)
by: Mena, Francisco, et al.
Published: (2024)
Subimage Overlap Prediction: Task-Aligned Self-Supervised Pretraining For Semantic Segmentation In Remote Sensing Imagery
by: Sharma, Lakshay, et al.
Published: (2026)
by: Sharma, Lakshay, et al.
Published: (2026)
EXAONEPath 1.0 Patch-level Foundation Model for Pathology
by: Yun, Juseung, et al.
Published: (2024)
by: Yun, Juseung, et al.
Published: (2024)
Align Your Tangent: Training Better Consistency Models via Manifold-Aligned Tangents
by: Kim, Beomsu, et al.
Published: (2025)
by: Kim, Beomsu, et al.
Published: (2025)
Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models
by: Lai, Zhengfeng, et al.
Published: (2024)
by: Lai, Zhengfeng, et al.
Published: (2024)
MINR: Implicit Neural Representations with Masked Image Modelling
by: Lee, Sua, et al.
Published: (2025)
by: Lee, Sua, et al.
Published: (2025)
Cross-Domain Few-Shot Learning for Hyperspectral Image Classification Based on Mixup Foundation Model
by: Paeedeh, Naeem, et al.
Published: (2026)
by: Paeedeh, Naeem, et al.
Published: (2026)
Similar Items
-
On Pitfalls of $\textit{RemOve-And-Retrain}$: Data Processing Inequality Perspective
by: Song, Junhwa, et al.
Published: (2023) -
Pushing the Limits of Vision-Language Models in Remote Sensing without Human Annotations
by: Cha, Keumgang, et al.
Published: (2024) -
SatVision-TOA: A Geospatial Foundation Model for Coarse-Resolution All-Sky Remote Sensing Imagery
by: Spradlin, Caleb S., et al.
Published: (2024) -
StyleLipSync: Style-based Personalized Lip-sync Video Generation
by: Ki, Taekyung, et al.
Published: (2023) -
Solid Waste Detection, Monitoring and Mapping in Remote Sensing Images: A Survey
by: Fraternali, Piero, et al.
Published: (2024)