Atomizer: Generalizing to new modalities by breaking satellite images down to a set of scalars
Fuente:
arXiv
Saved in:
| Main Authors: | de Turckheim, Hugo Riffaud, Lobry, Sylvain, Interdonato, Roberto, Marcos, Diego |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RAMEN: Resolution-Adjustable Multimodal Encoder for Earth Observation
by: Houdré, Nicolas, et al.
Published: (2025)
by: Houdré, Nicolas, et al.
Published: (2025)
SenCLIP: Enhancing zero-shot land-use mapping for Sentinel-2 with ground-level prompting
by: Jain, Pallavi, et al.
Published: (2024)
by: Jain, Pallavi, et al.
Published: (2024)
TimeSenCLIP: A Time Series Vision-Language Model for Remote Sensing
by: Jain, Pallavi, et al.
Published: (2025)
by: Jain, Pallavi, et al.
Published: (2025)
Can SAR improve RSVQA performance?
by: Tosato, Lucrezia, et al.
Published: (2024)
by: Tosato, Lucrezia, et al.
Published: (2024)
SAR Strikes Back: A New Hope for RSVQA
by: Tosato, Lucrezia, et al.
Published: (2025)
by: Tosato, Lucrezia, et al.
Published: (2025)
Multi-modal Co-learning for Earth Observation: Enhancing single-modality models via modality collaboration
by: Mena, Francisco, et al.
Published: (2025)
by: Mena, Francisco, et al.
Published: (2025)
Visual Question Answering on Multiple Remote Sensing Image Modalities
by: Boussaid, Hichem, et al.
Published: (2025)
by: Boussaid, Hichem, et al.
Published: (2025)
Segmentation-guided Attention for Visual Question Answering from Remote Sensing Images
by: Tosato, Lucrezia, et al.
Published: (2024)
by: Tosato, Lucrezia, et al.
Published: (2024)
Checkmate: interpretable and explainable RSVQA is the endgame
by: Tosato, Lucrezia, et al.
Published: (2025)
by: Tosato, Lucrezia, et al.
Published: (2025)
AtomGS: Atomizing Gaussian Splatting for High-Fidelity Radiance Field
by: Liu, Rong, et al.
Published: (2024)
by: Liu, Rong, et al.
Published: (2024)
GroundSet: A Cadastral-Grounded Dataset for Spatial Understanding with Vector Data
by: Ferrod, Roger, et al.
Published: (2026)
by: Ferrod, Roger, et al.
Published: (2026)
IC-EO: Interpretable Code-based assistant for Earth Observation
by: Lahouel, Lamia, et al.
Published: (2026)
by: Lahouel, Lamia, et al.
Published: (2026)
Soft labelling for semantic segmentation: Bringing coherence to label down-sampling
by: Alcover-Couso, Roberto, et al.
Published: (2023)
by: Alcover-Couso, Roberto, et al.
Published: (2023)
Rethinking Cross-modal Interaction from a Top-down Perspective for Referring Video Object Segmentation
by: Liang, Chen, et al.
Published: (2021)
by: Liang, Chen, et al.
Published: (2021)
Toward task-driven satellite image super-resolution
by: Ziaja, Maciej, et al.
Published: (2025)
by: Ziaja, Maciej, et al.
Published: (2025)
Multi-modal, multi-scale representation learning for satellite imagery analysis just needs a good ALiBi
by: Kage, Patrick, et al.
Published: (2026)
by: Kage, Patrick, et al.
Published: (2026)
Image augmentation with invertible networks in interactive satellite image change detection
by: Sahbi, Hichem
Published: (2025)
by: Sahbi, Hichem
Published: (2025)
Multi-step feature fusion for natural disaster damage assessment on satellite images
by: Żarski, Mateusz, et al.
Published: (2024)
by: Żarski, Mateusz, et al.
Published: (2024)
Label-frugal satellite image change detection with generative virtual exemplar learning
by: Sahbi, Hichem
Published: (2025)
by: Sahbi, Hichem
Published: (2025)
Matching of SAR and optical images based on transformation to shared modality
by: Borisov, Alexey, et al.
Published: (2026)
by: Borisov, Alexey, et al.
Published: (2026)
Cross-modal ultra-scale learning with tri-modalities of renal biopsy images for glomerular multi-disease auxiliary diagnosis
by: Long, Kaixing, et al.
Published: (2025)
by: Long, Kaixing, et al.
Published: (2025)
Scalable neural pushbroom architectures for real-time denoising of hyperspectral images onboard satellites
by: Yi, Ziyao, et al.
Published: (2026)
by: Yi, Ziyao, et al.
Published: (2026)
FungiTastic: A multi-modal dataset and benchmark for image categorization
by: Picek, Lukas, et al.
Published: (2024)
by: Picek, Lukas, et al.
Published: (2024)
CanadaFireSat: Toward high-resolution wildfire forecasting with multiple modalities
by: Porta, Hugo, et al.
Published: (2025)
by: Porta, Hugo, et al.
Published: (2025)
UniVG: Towards UNIfied-modal Video Generation
by: Ruan, Ludan, et al.
Published: (2024)
by: Ruan, Ludan, et al.
Published: (2024)
Multi-Scale Grouped Prototypes for Interpretable Semantic Segmentation
by: Porta, Hugo, et al.
Published: (2024)
by: Porta, Hugo, et al.
Published: (2024)
Robust image segmentation model based on binary level set
by: Zhao, Wenqi
Published: (2024)
by: Zhao, Wenqi
Published: (2024)
EgoAnimate: Generating Human Animations from Egocentric top-down Views
by: Türkoglu, G. Kutay, et al.
Published: (2025)
by: Türkoglu, G. Kutay, et al.
Published: (2025)
Quantum-enhanced satellite image classification
by: Zhang, Qi, et al.
Published: (2026)
by: Zhang, Qi, et al.
Published: (2026)
A 3D generative model of pathological multi-modal MR images and segmentations
by: Fernandez, Virginia, et al.
Published: (2023)
by: Fernandez, Virginia, et al.
Published: (2023)
Comparative analysis of dual-form networks for live land monitoring using multi-modal satellite image time series
by: Dumeur, Iris, et al.
Published: (2026)
by: Dumeur, Iris, et al.
Published: (2026)
CromSS: Cross-modal pre-training with noisy labels for remote sensing image segmentation
by: Liu, Chenying, et al.
Published: (2024)
by: Liu, Chenying, et al.
Published: (2024)
Tree level change detection over Ahmedabad city using very high resolution satellite images and Deep Learning
by: Singla, Jai G, et al.
Published: (2024)
by: Singla, Jai G, et al.
Published: (2024)
Liquid: Language Models are Scalable and Unified Multi-modal Generators
by: Wu, Junfeng, et al.
Published: (2024)
by: Wu, Junfeng, et al.
Published: (2024)
Multi-modal Generation via Cross-Modal In-Context Learning
by: Kumar, Amandeep, et al.
Published: (2024)
by: Kumar, Amandeep, et al.
Published: (2024)
Enhancing Incomplete Multi-modal Brain Tumor Segmentation with Intra-modal Asymmetry and Inter-modal Dependency
by: Liu, Weide, et al.
Published: (2024)
by: Liu, Weide, et al.
Published: (2024)
PosterLLaVa: Constructing a Unified Multi-modal Layout Generator with LLM
by: Yang, Tao, et al.
Published: (2024)
by: Yang, Tao, et al.
Published: (2024)
Can Text-to-image Model Assist Multi-modal Learning for Visual Recognition with Visual Modality Missing?
by: Feng, Tiantian, et al.
Published: (2024)
by: Feng, Tiantian, et al.
Published: (2024)
MMGen: Unified Multi-modal Image Generation and Understanding in One Go
by: Wang, Jiepeng, et al.
Published: (2025)
by: Wang, Jiepeng, et al.
Published: (2025)
CoMA: Compositional Human Motion Generation with Multi-modal Agents
by: Sun, Shanlin, et al.
Published: (2024)
by: Sun, Shanlin, et al.
Published: (2024)
Similar Items
-
RAMEN: Resolution-Adjustable Multimodal Encoder for Earth Observation
by: Houdré, Nicolas, et al.
Published: (2025) -
SenCLIP: Enhancing zero-shot land-use mapping for Sentinel-2 with ground-level prompting
by: Jain, Pallavi, et al.
Published: (2024) -
TimeSenCLIP: A Time Series Vision-Language Model for Remote Sensing
by: Jain, Pallavi, et al.
Published: (2025) -
Can SAR improve RSVQA performance?
by: Tosato, Lucrezia, et al.
Published: (2024) -
SAR Strikes Back: A New Hope for RSVQA
by: Tosato, Lucrezia, et al.
Published: (2025)