MVT: Mask-Grounded Vision-Language Models for Taxonomy-Aligned Land-Cover Tagging
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Siyi, Wang, Kai, Pang, Weicong, Yang, Ruiming, Chen, Ziru, Gao, Renjun, Lau, Alexis Kai Hon, Gu, Dasa, Zhang, Chenchen, Li, Cheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LC4-DViT: Land-cover Creation for Land-cover Classification with Deformable Vision Transformer
by: Wang, Kai, et al.
Published: (2025)
by: Wang, Kai, et al.
Published: (2025)
TagAlign: Improving Vision-Language Alignment with Multi-Tag Classification
by: Liu, Qinying, et al.
Published: (2023)
by: Liu, Qinying, et al.
Published: (2023)
Myopia Control Efficacy of Grid Dimension Multiregion Spectacle Lenses: A One‐Year Randomized Double‐Masked Controlled Trial
by: Junfeng Wang, et al.
Published: (2026)
by: Junfeng Wang, et al.
Published: (2026)
PIPE: Physics-Informed Position Encoding for Alignment of Satellite Images and Time Series
by: Li, Haobo, et al.
Published: (2025)
by: Li, Haobo, et al.
Published: (2025)
FLrce: Resource-Efficient Federated Learning with Early-Stopping Strategy
by: Niu, Ziru, et al.
Published: (2023)
by: Niu, Ziru, et al.
Published: (2023)
CLLMate: A Multimodal Benchmark for Weather and Climate Events Forecasting
by: Li, Haobo, et al.
Published: (2024)
by: Li, Haobo, et al.
Published: (2024)
LLM4Tag: Automatic Tagging System for Information Retrieval via Large Language Models
by: Tang, Ruiming, et al.
Published: (2025)
by: Tang, Ruiming, et al.
Published: (2025)
Adaptive Bit Partitioning for Reconfigurable Intelligent Surface Assisted FDD Systems with Limited Feedback
by: Chen, Weicong, et al.
Published: (2020)
by: Chen, Weicong, et al.
Published: (2020)
MARS: Multi-Agent Robotic System with Multimodal Large Language Models for Assistive Intelligence
by: Gao, Renjun
Published: (2025)
by: Gao, Renjun
Published: (2025)
VERITAS: Leveraging Vision Priors and Expert Fusion to Improve Multimodal Data
by: Xu, Tingqiao, et al.
Published: (2025)
by: Xu, Tingqiao, et al.
Published: (2025)
Measuring and Aligning Abstraction in Vision-Language Models with Medical Taxonomies
by: Schaper, Ben, et al.
Published: (2026)
by: Schaper, Ben, et al.
Published: (2026)
Mask Clustering-based Annotation Engine for Large-Scale Submeter Land Cover Mapping
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs
by: Gu, Yuzhe, et al.
Published: (2025)
by: Gu, Yuzhe, et al.
Published: (2025)
Perceptual Taxonomy: Evaluating and Guiding Hierarchical Scene Reasoning in Vision-Language Models
by: Lee, Jonathan, et al.
Published: (2025)
by: Lee, Jonathan, et al.
Published: (2025)
Positive Education. The axiom of contemporary upbringing and education
by: Hanuliaková, Jana, et al.
Published: (2026)
by: Hanuliaková, Jana, et al.
Published: (2026)
Positive Education. The axiom of contemporary upbringing and education
by: Hanuliaková, Jana, et al.
Published: (2026)
by: Hanuliaková, Jana, et al.
Published: (2026)
Trees in the Eyes of Young Learners: A Study on Knowledge and Educational Methods
by: Daša Bombjaková, et al.
Published: (2025)
by: Daša Bombjaková, et al.
Published: (2025)
Records of Xinjiang ground-jay Podoces biddulphi in Taklimakan Desert, Xinjiang, China
by: Ma, Ming, et al.
Published: (2004)
by: Ma, Ming, et al.
Published: (2004)
Joint Spatial Division and Multiplexing with Customized Orthogonal Group Channels in Multi-RIS-Assisted Systems
by: Chen, Weicong, et al.
Published: (2025)
by: Chen, Weicong, et al.
Published: (2025)
Channel Customization for Low-Complexity CSI Acquisition in Multi-RIS-Assisted MIMO Systems
by: Chen, Weicong, et al.
Published: (2024)
by: Chen, Weicong, et al.
Published: (2024)
The Model Agreed, But Didn't Learn: Diagnosing Surface Compliance in Large Language Models
by: Gu, Xiaojie, et al.
Published: (2026)
by: Gu, Xiaojie, et al.
Published: (2026)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
by: Huang, Haifeng, et al.
Published: (2025)
by: Huang, Haifeng, et al.
Published: (2025)
Learning Visuomotor Policy for Multi-Robot Laser Tag Game
by: Li, Kai, et al.
Published: (2026)
by: Li, Kai, et al.
Published: (2026)
From Tags to Trees: Structuring Fine-Grained Knowledge for Controllable Data Selection in LLM Instruction Tuning
by: Niu, Zihan, et al.
Published: (2026)
by: Niu, Zihan, et al.
Published: (2026)
Save It for the "Hot" Day: An LLM-Empowered Visual Analytics System for Heat Risk Management
by: Li, Haobo, et al.
Published: (2024)
by: Li, Haobo, et al.
Published: (2024)
Boltzmann equation with mixed boundary condition
by: Chen, Hongxu, et al.
Published: (2024)
by: Chen, Hongxu, et al.
Published: (2024)
Diffusive limit of the Boltzmann equation around Rayleigh profile in the half space
by: Chen, Hongxu, et al.
Published: (2025)
by: Chen, Hongxu, et al.
Published: (2025)
On packing total coloring
by: Ferme, Jasmina, et al.
Published: (2025)
by: Ferme, Jasmina, et al.
Published: (2025)
Safer-Instruct: Aligning Language Models with Automated Preference Data
by: Shi, Taiwei, et al.
Published: (2023)
by: Shi, Taiwei, et al.
Published: (2023)
MTRec: Learning to Align with User Preferences via Mental Reward Models
by: Zhao, Mengchen, et al.
Published: (2025)
by: Zhao, Mengchen, et al.
Published: (2025)
IRSAMap:Towards Large-Scale, High-Resolution Land Cover Map Vectorization
by: Meng, Yu, et al.
Published: (2025)
by: Meng, Yu, et al.
Published: (2025)
Pseudocode-Guided Structured Reasoning for Automating Reliable Inference in Vision-Language Models
by: Ni, Weicong, et al.
Published: (2026)
by: Ni, Weicong, et al.
Published: (2026)
Taxonomy-Guided Zero-Shot Recommendations with LLMs
by: Liang, Yueqing, et al.
Published: (2024)
by: Liang, Yueqing, et al.
Published: (2024)
Entropy-aware Masking for Masked Language Modeling
by: Srinivasagan, Gokul, et al.
Published: (2026)
by: Srinivasagan, Gokul, et al.
Published: (2026)
ReMAP-DP: Reprojected Multi-view Aligned PointMaps for Diffusion Policy
by: Yang, Xinzhang, et al.
Published: (2026)
by: Yang, Xinzhang, et al.
Published: (2026)
AlignSum: Data Pyramid Hierarchical Fine-tuning for Aligning with Human Summarization Preference
by: Han, Yang, et al.
Published: (2024)
by: Han, Yang, et al.
Published: (2024)
Stretchable Electronic Facial Masks for Radiofrequency Therapy
by: Zanxin Zhou, et al.
Published: (2025)
by: Zanxin Zhou, et al.
Published: (2025)
The Incompressible Navier-Stokes-Fourier Limits from Boltzmann-Fermi-Dirac Equation for Low Regularity Data
by: Jiang, Ning, et al.
Published: (2025)
by: Jiang, Ning, et al.
Published: (2025)
MRMMIA: Membership Inference Attacks on Memory in Chat Agents
by: Chen, Kai, et al.
Published: (2026)
by: Chen, Kai, et al.
Published: (2026)
Mask Grounding for Referring Image Segmentation
by: Chng, Yong Xien, et al.
Published: (2023)
by: Chng, Yong Xien, et al.
Published: (2023)
Similar Items
-
LC4-DViT: Land-cover Creation for Land-cover Classification with Deformable Vision Transformer
by: Wang, Kai, et al.
Published: (2025) -
TagAlign: Improving Vision-Language Alignment with Multi-Tag Classification
by: Liu, Qinying, et al.
Published: (2023) -
Myopia Control Efficacy of Grid Dimension Multiregion Spectacle Lenses: A One‐Year Randomized Double‐Masked Controlled Trial
by: Junfeng Wang, et al.
Published: (2026) -
PIPE: Physics-Informed Position Encoding for Alignment of Satellite Images and Time Series
by: Li, Haobo, et al.
Published: (2025) -
FLrce: Resource-Efficient Federated Learning with Early-Stopping Strategy
by: Niu, Ziru, et al.
Published: (2023)