Do Satellite Tasks Need Special Pretraining?
Fuente:
arXiv
Guardado en:
| Autores principales: | Vanyan, Ani, Barseghyan, Alvard, Tamazyan, Hakob, Galstyan, Tigran, Huroyan, Vahan, Hovakimyan, Naira, Khachatrian, Hrant |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Analyzing Local Representations of Self-supervised Vision Transformers
por: Vanyan, Ani, et al.
Publicado: (2023)
por: Vanyan, Ani, et al.
Publicado: (2023)
GeoCrossBench: Cross-Band Generalization for Remote Sensing
por: Tamazyan, Hakob, et al.
Publicado: (2025)
por: Tamazyan, Hakob, et al.
Publicado: (2025)
Fusion of Pervasive RF Data with Spatial Images via Vision Transformers for Enhanced Mapping in Smart Cities
por: Mkrtchyan, Rafayel, et al.
Publicado: (2025)
por: Mkrtchyan, Rafayel, et al.
Publicado: (2025)
Vision Transformers for Efficient Indoor Pathloss Radio Map Prediction
por: Mkrtchyan, Rafayel, et al.
Publicado: (2024)
por: Mkrtchyan, Rafayel, et al.
Publicado: (2024)
Residual-based Language Models are Free Boosters for Biomedical Imaging
por: Lai, Zhixin, et al.
Publicado: (2024)
por: Lai, Zhixin, et al.
Publicado: (2024)
Are Pretrained Image Matchers Good Enough for SAR-Optical Satellite Registration?
por: Corley, Isaac, et al.
Publicado: (2026)
por: Corley, Isaac, et al.
Publicado: (2026)
Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC
por: Tao, Ming, et al.
Publicado: (2024)
por: Tao, Ming, et al.
Publicado: (2024)
Effective Feature Learning for 3D Medical Registration via Domain-Specialized DINO Pretraining
por: Kats, Eytan, et al.
Publicado: (2026)
por: Kats, Eytan, et al.
Publicado: (2026)
Task Specific Pretraining with Noisy Labels for Remote Sensing Image Segmentation
por: Liu, Chenying, et al.
Publicado: (2024)
por: Liu, Chenying, et al.
Publicado: (2024)
On Good Practices for Task-Specific Distillation of Large Pretrained Visual Models
por: Marrie, Juliette, et al.
Publicado: (2024)
por: Marrie, Juliette, et al.
Publicado: (2024)
SenPa-MAE: Sensor Parameter Aware Masked Autoencoder for Multi-Satellite Self-Supervised Pretraining
por: Prexl, Jonathan, et al.
Publicado: (2024)
por: Prexl, Jonathan, et al.
Publicado: (2024)
Cross-Scale Pretraining: Enhancing Self-Supervised Learning for Low-Resolution Satellite Imagery for Semantic Segmentation
por: Waithaka, John, et al.
Publicado: (2026)
por: Waithaka, John, et al.
Publicado: (2026)
When Do We Not Need Larger Vision Models?
por: Shi, Baifeng, et al.
Publicado: (2024)
por: Shi, Baifeng, et al.
Publicado: (2024)
MTP: Advancing Remote Sensing Foundation Model via Multi-Task Pretraining
por: Wang, Di, et al.
Publicado: (2024)
por: Wang, Di, et al.
Publicado: (2024)
Do Vision Language Models Need to Process Image Tokens?
por: Ghosh, Sambit, et al.
Publicado: (2026)
por: Ghosh, Sambit, et al.
Publicado: (2026)
MAL: Cluster-Masked and Multi-Task Pretraining for Enhanced xLSTM Vision Performance
por: Huang, Wenjun, et al.
Publicado: (2024)
por: Huang, Wenjun, et al.
Publicado: (2024)
Do We Need Reformer for Vision? An Experimental Comparison with Vision Transformers
por: Bellaj, Ali El, et al.
Publicado: (2025)
por: Bellaj, Ali El, et al.
Publicado: (2025)
Making Sense Of Distributed Representations With Activation Spectroscopy
por: Reing, Kyle, et al.
Publicado: (2025)
por: Reing, Kyle, et al.
Publicado: (2025)
Seeing Beyond Redundancy: Task Complexity's Role in Vision Token Specialization in VLLMs
por: Hannan, Darryl, et al.
Publicado: (2026)
por: Hannan, Darryl, et al.
Publicado: (2026)
Do Less, Achieve More: Do We Need Every-Step Optimization for RL Fine-tuning of Diffusion Models?
por: Yan, Renye, et al.
Publicado: (2026)
por: Yan, Renye, et al.
Publicado: (2026)
Do We Need Perfect Data? Leveraging Noise for Domain Generalized Segmentation
por: Kim, Taeyeong, et al.
Publicado: (2025)
por: Kim, Taeyeong, et al.
Publicado: (2025)
Adept: Annotation-Denoising Auxiliary Tasks with Discrete Cosine Transform Map and Keypoint for Human-Centric Pretraining
por: He, Weizhen, et al.
Publicado: (2025)
por: He, Weizhen, et al.
Publicado: (2025)
CARE: Multi-Task Pretraining for Latent Continuous Action Representation in Robot Control
por: Shi, Jiaqi, et al.
Publicado: (2026)
por: Shi, Jiaqi, et al.
Publicado: (2026)
CLIP with Quality Captions: A Strong Pretraining for Vision Tasks
por: Vasu, Pavan Kumar Anasosalu, et al.
Publicado: (2024)
por: Vasu, Pavan Kumar Anasosalu, et al.
Publicado: (2024)
Self-Supervised Pretraining on Satellite Imagery: a Case Study on Label-Efficient Vehicle Detection
por: BOURCIER, Jules, et al.
Publicado: (2022)
por: BOURCIER, Jules, et al.
Publicado: (2022)
SatelliteCalculator: A Multi-Task Vision Foundation Model for Quantitative Remote Sensing Inversion
por: Yu, Zhenyu, et al.
Publicado: (2025)
por: Yu, Zhenyu, et al.
Publicado: (2025)
SLAM-AGS: Slide-Label Aware Multi-Task Pretraining Using Adaptive Gradient Surgery in Computational Cytology
por: Acerbis, Marco, et al.
Publicado: (2025)
por: Acerbis, Marco, et al.
Publicado: (2025)
Capsule Networks Do Not Need to Model Everything
por: Renzulli, Riccardo, et al.
Publicado: (2022)
por: Renzulli, Riccardo, et al.
Publicado: (2022)
EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation
por: Liu, Longfei, et al.
Publicado: (2026)
por: Liu, Longfei, et al.
Publicado: (2026)
Few-Shot Deployment of Pretrained MRI Transformers in Brain Imaging Tasks
por: Li, Mengyu, et al.
Publicado: (2025)
por: Li, Mengyu, et al.
Publicado: (2025)
Cross-Task Pretraining for Cross-Organ Cross-Scanner Adenocarcinoma Segmentation
por: Galdran, Adrian
Publicado: (2024)
por: Galdran, Adrian
Publicado: (2024)
From General to Specialized: The Need for Foundational Models in Agriculture
por: Nedungadi, Vishal, et al.
Publicado: (2025)
por: Nedungadi, Vishal, et al.
Publicado: (2025)
Leveraging Vision Language Models for Specialized Agricultural Tasks
por: Arshad, Muhammad Arbab, et al.
Publicado: (2024)
por: Arshad, Muhammad Arbab, et al.
Publicado: (2024)
Rethinking Transfer Learning for Industrial Inspection: DINOv3 vs. ImageNet Pretraining Across RGB and X-ray Tasks
por: Gharbage, Mehdi, et al.
Publicado: (2026)
por: Gharbage, Mehdi, et al.
Publicado: (2026)
How Much of a Model Do We Need? Redundancy and Slimmability in Remote Sensing Foundation Models
por: Hackel, Leonard, et al.
Publicado: (2026)
por: Hackel, Leonard, et al.
Publicado: (2026)
Parameter-Efficient Fine-Tuning of Large Pretrained Models for Instance Segmentation Tasks
por: Baker, Nermeen Abou, et al.
Publicado: (2026)
por: Baker, Nermeen Abou, et al.
Publicado: (2026)
Autoregressive Pretraining with Mamba in Vision
por: Ren, Sucheng, et al.
Publicado: (2024)
por: Ren, Sucheng, et al.
Publicado: (2024)
Emu: Generative Pretraining in Multimodality
por: Sun, Quan, et al.
Publicado: (2023)
por: Sun, Quan, et al.
Publicado: (2023)
Continuous Urban Change Detection from Satellite Image Time Series with Temporal Feature Refinement and Multi-Task Integration
por: Hafner, Sebastian, et al.
Publicado: (2024)
por: Hafner, Sebastian, et al.
Publicado: (2024)
Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting
por: Colombo, Antonio, et al.
Publicado: (2026)
por: Colombo, Antonio, et al.
Publicado: (2026)
Ejemplares similares
-
Analyzing Local Representations of Self-supervised Vision Transformers
por: Vanyan, Ani, et al.
Publicado: (2023) -
GeoCrossBench: Cross-Band Generalization for Remote Sensing
por: Tamazyan, Hakob, et al.
Publicado: (2025) -
Fusion of Pervasive RF Data with Spatial Images via Vision Transformers for Enhanced Mapping in Smart Cities
por: Mkrtchyan, Rafayel, et al.
Publicado: (2025) -
Vision Transformers for Efficient Indoor Pathloss Radio Map Prediction
por: Mkrtchyan, Rafayel, et al.
Publicado: (2024) -
Residual-based Language Models are Free Boosters for Biomedical Imaging
por: Lai, Zhixin, et al.
Publicado: (2024)