HCCM: Hierarchical Cross-Granularity Contrastive and Matching Learning for Natural Language-Guided Drones
Fuente:
arXiv
Saved in:
| Main Authors: | Ruan, Hao, Lin, Jinliang, Lai, Yingxin, Luo, Zhiming, Li, Shaozi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Selective Domain-Invariant Feature for Generalizable Deepfake Detection
by: Lai, Yingxin, et al.
Published: (2024)
by: Lai, Yingxin, et al.
Published: (2024)
Exploring Frequencies via Feature Mixing and Meta-Learning for Improving Adversarial Transferability
by: Weng, Juanjuan, et al.
Published: (2024)
by: Weng, Juanjuan, et al.
Published: (2024)
A brief introduction to a framework named Multilevel Guidance-Exploration Network
by: Yang, Guoqing, et al.
Published: (2023)
by: Yang, Guoqing, et al.
Published: (2023)
Improving Transferable Targeted Adversarial Attack via Normalized Logit Calibration and Truncated Feature Mixing
by: Weng, Juanjuan, et al.
Published: (2024)
by: Weng, Juanjuan, et al.
Published: (2024)
Towards Adversarial Robustness via Debiased High-Confidence Logit Alignment
by: Zhang, Kejia, et al.
Published: (2024)
by: Zhang, Kejia, et al.
Published: (2024)
Zero-Shot Co-salient Object Detection Framework
by: Xiao, Haoke, et al.
Published: (2023)
by: Xiao, Haoke, et al.
Published: (2023)
Mitigating Low-Frequency Bias: Feature Recalibration and Frequency Attention Regularization for Adversarial Robustness
by: Zhang, Kejia, et al.
Published: (2024)
by: Zhang, Kejia, et al.
Published: (2024)
Long-Tailed Out-of-Distribution Detection: Prioritizing Attention to Tail
by: He, Yina, et al.
Published: (2024)
by: He, Yina, et al.
Published: (2024)
Weakly Supervised Object Detection for Automatic Tooth-marked Tongue Recognition
by: Zhang, Yongcun, et al.
Published: (2024)
by: Zhang, Yongcun, et al.
Published: (2024)
Harmonizing Feature Maps: A Graph Convolutional Approach for Enhancing Adversarial Robustness
by: Zhang, Kejia, et al.
Published: (2024)
by: Zhang, Kejia, et al.
Published: (2024)
Towards Natural Language-Guided Drones: GeoText-1652 Benchmark with Spatial Relation Matching
by: Chu, Meng, et al.
Published: (2023)
by: Chu, Meng, et al.
Published: (2023)
SPIN: Hierarchical Segmentation with Subpart Granularity in Natural Images
by: Myers-Dean, Josh, et al.
Published: (2024)
by: Myers-Dean, Josh, et al.
Published: (2024)
Bridging Granularity Gaps: Hierarchical Semantic Learning for Cross-domain Few-shot Segmentation
by: Sun, Sujun, et al.
Published: (2025)
by: Sun, Sujun, et al.
Published: (2025)
CHASD: Language Increment-Calibrated Contrastive Decoding against Hallucination in LVLMs
by: Huang, Xiaoyi, et al.
Published: (2026)
by: Huang, Xiaoyi, et al.
Published: (2026)
Unsupervised Contrastive Learning for Efficient and Robust Spectral Shape Matching
by: Luo, Feifan, et al.
Published: (2026)
by: Luo, Feifan, et al.
Published: (2026)
$β$-CLIP: Text-Conditioned Contrastive Learning for Multi-Granular Vision-Language Alignment
by: Zohra, Fatimah, et al.
Published: (2025)
by: Zohra, Fatimah, et al.
Published: (2025)
Hierarchical Cross-modal Prompt Learning for Vision-Language Models
by: Zheng, Hao, et al.
Published: (2025)
by: Zheng, Hao, et al.
Published: (2025)
Multi-Granularity Language-Guided Training for Multi-Object Tracking
by: Li, Yuhao, et al.
Published: (2024)
by: Li, Yuhao, et al.
Published: (2024)
GuideFlow: Constraint-Guided Flow Matching for Planning in End-to-End Autonomous Driving
by: Liu, Lin, et al.
Published: (2025)
by: Liu, Lin, et al.
Published: (2025)
Hierarchical Action Recognition: A Contrastive Video-Language Approach with Hierarchical Interactions
by: Zhang, Rui, et al.
Published: (2024)
by: Zhang, Rui, et al.
Published: (2024)
Enhancing Representation in Radiography-Reports Foundation Model: A Granular Alignment Algorithm Using Masked Contrastive Learning
by: Huang, Weijian, et al.
Published: (2023)
by: Huang, Weijian, et al.
Published: (2023)
Weakly-supervised Semantic Segmentation via Dual-stream Contrastive Learning of Cross-image Contextual Information
by: Lai, Qi, et al.
Published: (2024)
by: Lai, Qi, et al.
Published: (2024)
MobileGeo: Exploring Hierarchical Knowledge Distillation for Resource-Efficient Cross-view Drone Geo-Localization
by: Sun, Jian, et al.
Published: (2025)
by: Sun, Jian, et al.
Published: (2025)
Why and How: Knowledge-Guided Learning for Cross-Spectral Image Patch Matching
by: Yu, Chuang, et al.
Published: (2024)
by: Yu, Chuang, et al.
Published: (2024)
CurveFlow: Curvature-Guided Flow Matching for Image Generation
by: Luo, Yan, et al.
Published: (2025)
by: Luo, Yan, et al.
Published: (2025)
MSP-MVS: Multi-Granularity Segmentation Prior Guided Multi-View Stereo
by: Yuan, Zhenlong, et al.
Published: (2024)
by: Yuan, Zhenlong, et al.
Published: (2024)
DDAP: Dual-Domain Anti-Personalization against Text-to-Image Diffusion Models
by: Yang, Jing, et al.
Published: (2024)
by: Yang, Jing, et al.
Published: (2024)
Contrast-X: A Multi-Modal Contrast Image Synthesis Benchmark and Universal Modality Flow Matching
by: Chen, Yifan, et al.
Published: (2026)
by: Chen, Yifan, et al.
Published: (2026)
HPT++: Hierarchically Prompting Vision-Language Models with Multi-Granularity Knowledge Generation and Improved Structure Modeling
by: Wang, Yubin, et al.
Published: (2024)
by: Wang, Yubin, et al.
Published: (2024)
Drone-assisted Road Gaussian Splatting with Cross-view Uncertainty
by: Zhang, Saining, et al.
Published: (2024)
by: Zhang, Saining, et al.
Published: (2024)
ZeroStereo: Zero-shot Stereo Matching from Single Images
by: Wang, Xianqi, et al.
Published: (2025)
by: Wang, Xianqi, et al.
Published: (2025)
MGHFT: Multi-Granularity Hierarchical Fusion Transformer for Cross-Modal Sticker Emotion Recognition
by: Chen, Jian, et al.
Published: (2025)
by: Chen, Jian, et al.
Published: (2025)
UMCL: Unimodal-generated Multimodal Contrastive Learning for Cross-compression-rate Deepfake Detection
by: Lai, Ching-Yi, et al.
Published: (2025)
by: Lai, Ching-Yi, et al.
Published: (2025)
LFTR: Learning-Free Token Reduction for Multimodal Large Language Models
by: Zhao, Zihui, et al.
Published: (2025)
by: Zhao, Zihui, et al.
Published: (2025)
Saliency-Guided Deep Learning for Bridge Defect Detection in Drone Imagery
by: Hebbache, Loucif, et al.
Published: (2025)
by: Hebbache, Loucif, et al.
Published: (2025)
VeCoR -- Velocity Contrastive Regularization for Flow Matching
by: Hong, Zong-Wei, et al.
Published: (2025)
by: Hong, Zong-Wei, et al.
Published: (2025)
Phrase Decoupling Cross-Modal Hierarchical Matching and Progressive Position Correction for Visual Grounding
by: Xie, Minghong, et al.
Published: (2024)
by: Xie, Minghong, et al.
Published: (2024)
PMCE: Probabilistic Multi-Granularity Semantics with Caption-Guided Enhancement for Few-Shot Learning
by: Wu, Jiaying, et al.
Published: (2026)
by: Wu, Jiaying, et al.
Published: (2026)
Hierarchical Banzhaf Interaction for General Video-Language Representation Learning
by: Jin, Peng, et al.
Published: (2024)
by: Jin, Peng, et al.
Published: (2024)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
Similar Items
-
Selective Domain-Invariant Feature for Generalizable Deepfake Detection
by: Lai, Yingxin, et al.
Published: (2024) -
Exploring Frequencies via Feature Mixing and Meta-Learning for Improving Adversarial Transferability
by: Weng, Juanjuan, et al.
Published: (2024) -
A brief introduction to a framework named Multilevel Guidance-Exploration Network
by: Yang, Guoqing, et al.
Published: (2023) -
Improving Transferable Targeted Adversarial Attack via Normalized Logit Calibration and Truncated Feature Mixing
by: Weng, Juanjuan, et al.
Published: (2024) -
Towards Adversarial Robustness via Debiased High-Confidence Logit Alignment
by: Zhang, Kejia, et al.
Published: (2024)