EarthVL: A Progressive Earth Vision-Language Understanding and Generation Framework
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wang, Junjue, Zhong, Yanfei, Chen, Zihang, Zheng, Zhuo, Ma, Ailong, Zhang, Liangpei |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
DisasterM3: A Remote Sensing Vision-Language Dataset for Disaster Damage Assessment and Response
par: Wang, Junjue, et autres
Publié: (2025)
par: Wang, Junjue, et autres
Publié: (2025)
Rapid Adaptation of Earth Observation Foundation Models for Segmentation
par: Selvam, Karthick Panner, et autres
Publié: (2024)
par: Selvam, Karthick Panner, et autres
Publié: (2024)
Data Augmentation in Earth Observation: A Diffusion Model Approach
par: Sousa, Tiago, et autres
Publié: (2024)
par: Sousa, Tiago, et autres
Publié: (2024)
Uncertainty and Generalizability in Foundation Models for Earth Observation
par: Ramos-Pollan, Raul, et autres
Publié: (2024)
par: Ramos-Pollan, Raul, et autres
Publié: (2024)
TimeCausality: Evaluating the Causal Ability in Time Dimension for Vision Language Models
par: Wang, Zeqing, et autres
Publié: (2025)
par: Wang, Zeqing, et autres
Publié: (2025)
Skeletonization-Based Adversarial Perturbations on Large Vision Language Model's Mathematical Text Recognition
par: Yoshida, Masatomo, et autres
Publié: (2026)
par: Yoshida, Masatomo, et autres
Publié: (2026)
DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models
par: Zhou, Yue, et autres
Publié: (2026)
par: Zhou, Yue, et autres
Publié: (2026)
DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot Learning
par: Li, Wenhao, et autres
Publié: (2026)
par: Li, Wenhao, et autres
Publié: (2026)
Evaluating the Significance of Outdoor Advertising from Driver's Perspective Using Computer Vision
par: Černeková, Zuzana, et autres
Publié: (2023)
par: Černeková, Zuzana, et autres
Publié: (2023)
Synthetic Image Detection with CLIP: Understanding and Assessing Predictive Cues
par: Willi, Marco, et autres
Publié: (2026)
par: Willi, Marco, et autres
Publié: (2026)
Research Challenges and Progress in the End-to-End V2X Cooperative Autonomous Driving Competition
par: Hao, Ruiyang, et autres
Publié: (2025)
par: Hao, Ruiyang, et autres
Publié: (2025)
Efficient Diffusion Models: A Comprehensive Survey from Principles to Practices
par: Ma, Zhiyuan, et autres
Publié: (2024)
par: Ma, Zhiyuan, et autres
Publié: (2024)
Deepfake Detection Generalization with Diffusion Noise
par: Qi, Hongyuan, et autres
Publié: (2026)
par: Qi, Hongyuan, et autres
Publié: (2026)
SETR: A Two-Stage Semantic-Enhanced Framework for Zero-Shot Composed Image Retrieval
par: Xiao, Yuqi, et autres
Publié: (2025)
par: Xiao, Yuqi, et autres
Publié: (2025)
Cost Savings from Automatic Quality Assessment of Generated Images
par: Giro-i-Nieto, Xavier, et autres
Publié: (2025)
par: Giro-i-Nieto, Xavier, et autres
Publié: (2025)
A Multi-purpose Tracking Framework for Salmon Welfare Monitoring in Challenging Environments
par: Høgstedt, Espen Uri, et autres
Publié: (2025)
par: Høgstedt, Espen Uri, et autres
Publié: (2025)
AOI-SSL: Self-Supervised Framework for Efficient Segmentation of Wire-bonded Semiconductors In Optical Inspection
par: Figueira, Joaquín, et autres
Publié: (2026)
par: Figueira, Joaquín, et autres
Publié: (2026)
Decoupled Sensitivity-Consistency Learning for Weakly Supervised Video Anomaly Detection
par: Zheng, Hantao, et autres
Publié: (2026)
par: Zheng, Hantao, et autres
Publié: (2026)
Fingerprint Membership and Identity Inference Against Generative Adversarial Networks
par: Cavasin, Saverio, et autres
Publié: (2024)
par: Cavasin, Saverio, et autres
Publié: (2024)
SPEAK: Speech-Driven Pose and Emotion-Adjustable Talking Head Generation
par: Cai, Changpeng, et autres
Publié: (2024)
par: Cai, Changpeng, et autres
Publié: (2024)
Investigation of cardinality classification for bacterial colony counting using explainable artificial intelligence
par: Zheng, Minghua, et autres
Publié: (2026)
par: Zheng, Minghua, et autres
Publié: (2026)
Learning to count small and clustered objects with application to bacterial colonies
par: Zheng, Minghua, et autres
Publié: (2026)
par: Zheng, Minghua, et autres
Publié: (2026)
MdaIF: Robust One-Stop Multi-Degradation-Aware Image Fusion with Language-Driven Semantics
par: Li, Jing, et autres
Publié: (2025)
par: Li, Jing, et autres
Publié: (2025)
Scalable and Realistic Virtual Try-on Application for Foundation Makeup with Kubelka-Munk Theory
par: Pang, Hui, et autres
Publié: (2025)
par: Pang, Hui, et autres
Publié: (2025)
Supersampling of Data from Structured-light Scanner with Deep Learning
par: Melicherčík, Martin, et autres
Publié: (2023)
par: Melicherčík, Martin, et autres
Publié: (2023)
Group Activity Recognition using Unreliable Tracked Pose
par: Thilakarathne, Haritha, et autres
Publié: (2024)
par: Thilakarathne, Haritha, et autres
Publié: (2024)
Towards Integrated Rock Support Visualisation in 3D Point Cloud of Underground Mines
par: Patra, Dibyayan, et autres
Publié: (2026)
par: Patra, Dibyayan, et autres
Publié: (2026)
Video-Based Human Pose Regression via Decoupled Space-Time Aggregation
par: He, Jijie, et autres
Publié: (2024)
par: He, Jijie, et autres
Publié: (2024)
HyperFM: An Efficient Hyperspectral Foundation Model with Spectral Grouping
par: Tushar, Zahid Hassan, et autres
Publié: (2026)
par: Tushar, Zahid Hassan, et autres
Publié: (2026)
KNN Transformer with Pyramid Prompts for Few-Shot Learning
par: Li, Wenhao, et autres
Publié: (2024)
par: Li, Wenhao, et autres
Publié: (2024)
Processing and Segmentation of Human Teeth from 2D Images using Weakly Supervised Learning
par: Kunzo, Tomáš, et autres
Publié: (2023)
par: Kunzo, Tomáš, et autres
Publié: (2023)
EUFCC-340K: A Faceted Hierarchical Dataset for Metadata Annotation in GLAM Collections
par: Net, Francesc, et autres
Publié: (2024)
par: Net, Francesc, et autres
Publié: (2024)
Cross-View-Prediction: Exploring Contrastive Feature for Hyperspectral Image Classification
par: Zhang, Anyu, et autres
Publié: (2022)
par: Zhang, Anyu, et autres
Publié: (2022)
Facial Spatiotemporal Graphs: Leveraging the 3D Facial Surface for Remote Physiological Measurement
par: Cantrill, Sam, et autres
Publié: (2026)
par: Cantrill, Sam, et autres
Publié: (2026)
Exploring Diffusion with Test-Time Training on Efficient Image Restoration
par: Lu, Rongchang, et autres
Publié: (2025)
par: Lu, Rongchang, et autres
Publié: (2025)
SAM Encoder Breach by Adversarial Simplicial Complex Triggers Downstream Model Failures
par: Qin, Yi, et autres
Publié: (2025)
par: Qin, Yi, et autres
Publié: (2025)
Tiny-YOLOSAM: Fast Hybrid Image Segmentation
par: Xu, Kenneth, et autres
Publié: (2025)
par: Xu, Kenneth, et autres
Publié: (2025)
Automated Discontinuity Set Characterisation in Enclosed Rock Face Point Clouds Using Single-Shot Filtering and Cyclic Orientation Transformation
par: Patra, Dibyayan, et autres
Publié: (2026)
par: Patra, Dibyayan, et autres
Publié: (2026)
SkeletonX: Data-Efficient Skeleton-based Action Recognition via Cross-sample Feature Aggregation
par: Zhang, Zongye, et autres
Publié: (2025)
par: Zhang, Zongye, et autres
Publié: (2025)
A Deep Learning Approach to Identify Rock Bolts in Complex 3D Point Clouds of Underground Mines Captured Using Mobile Laser Scanners
par: Patra, Dibyayan, et autres
Publié: (2025)
par: Patra, Dibyayan, et autres
Publié: (2025)
Documents similaires
-
DisasterM3: A Remote Sensing Vision-Language Dataset for Disaster Damage Assessment and Response
par: Wang, Junjue, et autres
Publié: (2025) -
Rapid Adaptation of Earth Observation Foundation Models for Segmentation
par: Selvam, Karthick Panner, et autres
Publié: (2024) -
Data Augmentation in Earth Observation: A Diffusion Model Approach
par: Sousa, Tiago, et autres
Publié: (2024) -
Uncertainty and Generalizability in Foundation Models for Earth Observation
par: Ramos-Pollan, Raul, et autres
Publié: (2024) -
TimeCausality: Evaluating the Causal Ability in Time Dimension for Vision Language Models
par: Wang, Zeqing, et autres
Publié: (2025)