CLIP the Landscape: Automated Tagging of Crowdsourced Landscape Images
Fuente:
arXiv
Saved in:
| Main Authors: | Ilyankou, Ilya, Jongwiriyanurak, Natchapon, Cheng, Tao, Haworth, James |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multiple Object Detection and Tracking in Panoramic Videos for Cycling Safety Analysis
by: Guo, Jingwei, et al.
Published: (2024)
by: Guo, Jingwei, et al.
Published: (2024)
V-RoAst: Visual Road Assessment. Can VLM be a Road Safety Assessor Using the iRAP Standard?
by: Jongwiriyanurak, Natchapon, et al.
Published: (2024)
by: Jongwiriyanurak, Natchapon, et al.
Published: (2024)
TTD: Text-Tag Self-Distillation Enhancing Image-Text Alignment in CLIP to Alleviate Single Tag Bias
by: Jo, Sanghyun, et al.
Published: (2024)
by: Jo, Sanghyun, et al.
Published: (2024)
TagCLIP: Improving Discrimination Ability of Open-Vocabulary Semantic Segmentation
by: Li, Jingyao, et al.
Published: (2023)
by: Li, Jingyao, et al.
Published: (2023)
Subjective Portrait Region Cropping in Landscape Videos with Temporal Annotation Smoothing
by: Lee, Cheng-Han, et al.
Published: (2026)
by: Lee, Cheng-Han, et al.
Published: (2026)
Optimizing 4D Gaussians for Dynamic Scene Video from Single Landscape Images
by: Jin, In-Hwan, et al.
Published: (2025)
by: Jin, In-Hwan, et al.
Published: (2025)
Landscape-Awareness for Geometric View Diffusion Model
by: Chen, Yan-Ting, et al.
Published: (2026)
by: Chen, Yan-Ting, et al.
Published: (2026)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
by: Karimian, Banafsheh, et al.
Published: (2025)
by: Karimian, Banafsheh, et al.
Published: (2025)
A Wander Through the Multimodal Landscape: Efficient Transfer Learning via Low-rank Sequence Multimodal Adapter
by: Guo, Zirun, et al.
Published: (2024)
by: Guo, Zirun, et al.
Published: (2024)
Tag2Text: Guiding Vision-Language Model via Image Tagging
by: Huang, Xinyu, et al.
Published: (2023)
by: Huang, Xinyu, et al.
Published: (2023)
DetailCLIP: Injecting Image Details into CLIP's Feature Space
by: Zhang, Zilun, et al.
Published: (2022)
by: Zhang, Zilun, et al.
Published: (2022)
Surveying the Landscape of Image Captioning Evaluation: A Comprehensive Taxonomy, Trends and Metrics Analysis
by: Berger, Uri, et al.
Published: (2024)
by: Berger, Uri, et al.
Published: (2024)
A Contrastive Learning Framework Empowered by Attention-based Feature Adaptation for Street-View Image Classification
by: You, Qi, et al.
Published: (2026)
by: You, Qi, et al.
Published: (2026)
Determining Intent of Changes to Ascertain Fake Crowdsourced Image Services
by: Umair, Muhammad, et al.
Published: (2023)
by: Umair, Muhammad, et al.
Published: (2023)
Sketch-Guided Stylized Landscape Cinemagraph Synthesis
by: Jin, Hao, et al.
Published: (2024)
by: Jin, Hao, et al.
Published: (2024)
Earth-Agent: Unlocking the Full Landscape of Earth Observation with Agents
by: Feng, Peilin, et al.
Published: (2025)
by: Feng, Peilin, et al.
Published: (2025)
Reframing Long-Tailed Learning via Loss Landscape Geometry
by: Chen, Shenghan, et al.
Published: (2026)
by: Chen, Shenghan, et al.
Published: (2026)
Scene Depth Estimation from Traditional Oriental Landscape Paintings
by: Kang, Sungho, et al.
Published: (2024)
by: Kang, Sungho, et al.
Published: (2024)
IPAD-CLIP: Teaching CLIP to Detect Image Local Perceptual Artifacts
by: Wang, Juan, et al.
Published: (2026)
by: Wang, Juan, et al.
Published: (2026)
Instruct-CLIP: Improving Instruction-Guided Image Editing with Automated Data Refinement Using Contrastive Learning
by: Chen, Sherry X., et al.
Published: (2025)
by: Chen, Sherry X., et al.
Published: (2025)
MediCLIP: Adapting CLIP for Few-shot Medical Image Anomaly Detection
by: Zhang, Ximiao, et al.
Published: (2024)
by: Zhang, Ximiao, et al.
Published: (2024)
CLIP-AGIQA: Boosting the Performance of AI-Generated Image Quality Assessment with CLIP
by: Tang, Zhenchen, et al.
Published: (2024)
by: Tang, Zhenchen, et al.
Published: (2024)
Jumping through Local Minima: Quantization in the Loss Landscape of Vision Transformers
by: Frumkin, Natalia, et al.
Published: (2023)
by: Frumkin, Natalia, et al.
Published: (2023)
Style Ambiguity Loss Using CLIP
by: Baker, James
Published: (2024)
by: Baker, James
Published: (2024)
BadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIP
by: Bai, Jiawang, et al.
Published: (2023)
by: Bai, Jiawang, et al.
Published: (2023)
CLIP in Medical Imaging: A Survey
by: Zhao, Zihao, et al.
Published: (2023)
by: Zhao, Zihao, et al.
Published: (2023)
AeroLite: Tag-Guided Lightweight Generation of Aerial Image Captions
by: Zi, Xing, et al.
Published: (2025)
by: Zi, Xing, et al.
Published: (2025)
Language Prompt vs. Image Enhancement: Boosting Object Detection With CLIP in Hazy Environments
by: Pang, Jian, et al.
Published: (2026)
by: Pang, Jian, et al.
Published: (2026)
Landscape More Secure Than Portrait? Zooming Into the Directionality of Digital Images With Security Implications
by: Lorch, Benedikt, et al.
Published: (2024)
by: Lorch, Benedikt, et al.
Published: (2024)
Beyond Top Activations: Efficient and Reliable Crowdsourced Evaluation of Automated Interpretability
by: Oikarinen, Tuomas, et al.
Published: (2025)
by: Oikarinen, Tuomas, et al.
Published: (2025)
Mapping Farmed Landscapes from Remote Sensing
by: Conserva, Michelangelo, et al.
Published: (2025)
by: Conserva, Michelangelo, et al.
Published: (2025)
Transfer CLIP for Generalizable Image Denoising
by: Cheng, Jun, et al.
Published: (2024)
by: Cheng, Jun, et al.
Published: (2024)
Image-to-Image Translation with Diffusion Transformers and CLIP-Based Image Conditioning
by: Zhu, Qiang, et al.
Published: (2025)
by: Zhu, Qiang, et al.
Published: (2025)
EditCLIP: Representation Learning for Image Editing
by: Wang, Qian, et al.
Published: (2025)
by: Wang, Qian, et al.
Published: (2025)
Targeted Forgetting of Image Subgroups in CLIP Models
by: Zhang, Zeliang, et al.
Published: (2025)
by: Zhang, Zeliang, et al.
Published: (2025)
The CLIP Model is Secretly an Image-to-Prompt Converter
by: Ding, Yuxuan, et al.
Published: (2023)
by: Ding, Yuxuan, et al.
Published: (2023)
Benchmarking PathCLIP for Pathology Image Analysis
by: Zheng, Sunyi, et al.
Published: (2024)
by: Zheng, Sunyi, et al.
Published: (2024)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
by: Kim, Seoyeon, et al.
Published: (2023)
by: Kim, Seoyeon, et al.
Published: (2023)
AdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detection
by: Cao, Yunkang, et al.
Published: (2024)
by: Cao, Yunkang, et al.
Published: (2024)
Are Multimodal Large Language Models Good Annotators for Image Tagging?
by: Xie, Ming-Kun, et al.
Published: (2026)
by: Xie, Ming-Kun, et al.
Published: (2026)
Similar Items
-
Multiple Object Detection and Tracking in Panoramic Videos for Cycling Safety Analysis
by: Guo, Jingwei, et al.
Published: (2024) -
V-RoAst: Visual Road Assessment. Can VLM be a Road Safety Assessor Using the iRAP Standard?
by: Jongwiriyanurak, Natchapon, et al.
Published: (2024) -
TTD: Text-Tag Self-Distillation Enhancing Image-Text Alignment in CLIP to Alleviate Single Tag Bias
by: Jo, Sanghyun, et al.
Published: (2024) -
TagCLIP: Improving Discrimination Ability of Open-Vocabulary Semantic Segmentation
by: Li, Jingyao, et al.
Published: (2023) -
Subjective Portrait Region Cropping in Landscape Videos with Temporal Annotation Smoothing
by: Lee, Cheng-Han, et al.
Published: (2026)