Saved in:
| Main Authors: | Lu, Linqi, Yu, Xianshi, Reddy, Akhil Perumal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2412.08648 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SatCLIP: Global, General-Purpose Location Embeddings with Satellite Imagery
by: Klemmer, Konstantin, et al.
Published: (2023)
by: Klemmer, Konstantin, et al.
Published: (2023)
Predicting Local Climate Zones using Urban Morphometrics and Satellite Imagery
by: Majer, Hugo, et al.
Published: (2026)
by: Majer, Hugo, et al.
Published: (2026)
Generalizable Slum Detection from Satellite Imagery with Mixture-of-Experts
by: Lee, Sumin, et al.
Published: (2025)
by: Lee, Sumin, et al.
Published: (2025)
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
by: Jha, Akshita, et al.
Published: (2024)
by: Jha, Akshita, et al.
Published: (2024)
From Drone Imagery to Livability Mapping: AI-powered Environment Perception in Rural China
by: Deng, Weihuan, et al.
Published: (2025)
by: Deng, Weihuan, et al.
Published: (2025)
Beluga Whale Detection from Satellite Imagery with Point Labels
by: Zheng, Yijie, et al.
Published: (2025)
by: Zheng, Yijie, et al.
Published: (2025)
Open-Vocabulary X-ray Prohibited Item Detection via Fine-tuning CLIP
by: Lin, Shuyang, et al.
Published: (2024)
by: Lin, Shuyang, et al.
Published: (2024)
Vision-Language Models for Autonomous Driving: CLIP-Based Dynamic Scene Understanding
by: Elhenawy, Mohammed, et al.
Published: (2025)
by: Elhenawy, Mohammed, et al.
Published: (2025)
Privacy of Groups in Dense Street Imagery
by: Franchi, Matt, et al.
Published: (2025)
by: Franchi, Matt, et al.
Published: (2025)
From Global to Local: Rethinking CLIP Feature Aggregation for Person Re-Identification
by: Zheng, Aotian, et al.
Published: (2026)
by: Zheng, Aotian, et al.
Published: (2026)
CLIP-UP: CLIP-Based Unanswerable Problem Detection for Visual Question Answering
by: Vardi, Ben, et al.
Published: (2025)
by: Vardi, Ben, et al.
Published: (2025)
CAF-YOLO: A Robust Framework for Multi-Scale Lesion Detection in Biomedical Imagery
by: Chen, Zilin, et al.
Published: (2024)
by: Chen, Zilin, et al.
Published: (2024)
Language Augmentation in CLIP for Improved Anatomy Detection on Multi-modal Medical Images
by: Kakkar, Mansi, et al.
Published: (2024)
by: Kakkar, Mansi, et al.
Published: (2024)
ChaosBench: A Multi-Channel, Physics-Based Benchmark for Subseasonal-to-Seasonal Climate Prediction
by: Nathaniel, Juan, et al.
Published: (2024)
by: Nathaniel, Juan, et al.
Published: (2024)
Coverage Biases in High-Resolution Satellite Imagery
by: Musienko, Vadim, et al.
Published: (2025)
by: Musienko, Vadim, et al.
Published: (2025)
Locating Demographic Bias at the Attention-Head Level in CLIP's Vision Encoder
by: Yasser, Alaa, et al.
Published: (2026)
by: Yasser, Alaa, et al.
Published: (2026)
Deploying Rapid Damage Assessments from sUAS Imagery for Disaster Response
by: Manzini, Thomas, et al.
Published: (2025)
by: Manzini, Thomas, et al.
Published: (2025)
DeCLIP: Decoupled Prompting for CLIP-based Multi-Label Class-Incremental Learning
by: Du, Kaile, et al.
Published: (2025)
by: Du, Kaile, et al.
Published: (2025)
BadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIP
by: Bai, Jiawang, et al.
Published: (2023)
by: Bai, Jiawang, et al.
Published: (2023)
Global and Local Entailment Learning for Natural World Imagery
by: Sastry, Srikumar, et al.
Published: (2025)
by: Sastry, Srikumar, et al.
Published: (2025)
IPAD-CLIP: Teaching CLIP to Detect Image Local Perceptual Artifacts
by: Wang, Juan, et al.
Published: (2026)
by: Wang, Juan, et al.
Published: (2026)
An Effective Image Copy-Move Forgery Detection Using Entropy Information
by: Jiang, Li, et al.
Published: (2023)
by: Jiang, Li, et al.
Published: (2023)
Do Street View Imagery and Public Participation GIS align: Comparative Analysis of Urban Attractiveness
by: Malekzadeh, Milad, et al.
Published: (2025)
by: Malekzadeh, Milad, et al.
Published: (2025)
From Global to Local: Social Bias Transfer in CLIP
by: Ramos, Ryan, et al.
Published: (2025)
by: Ramos, Ryan, et al.
Published: (2025)
Diagnosing Urban Street Vitality via a Visual-Semantic and Spatiotemporal Framework for Street-Level Economics
by: Zhuo, Xinxin, et al.
Published: (2026)
by: Zhuo, Xinxin, et al.
Published: (2026)
A High Resolution Urban and Rural Settlement Map of Africa Using Deep Learning and Satellite Imagery
by: Kakooei, Mohammad, et al.
Published: (2024)
by: Kakooei, Mohammad, et al.
Published: (2024)
Granularity at Scale: Estimating Neighborhood Socioeconomic Indicators from High-Resolution Orthographic Imagery and Hybrid Learning
by: Brewer, Ethan, et al.
Published: (2023)
by: Brewer, Ethan, et al.
Published: (2023)
Breaking the Global North Stereotype: A Global South-centric Benchmark Dataset for Auditing and Mitigating Biases in Facial Recognition Systems
by: Jaiswal, Siddharth D, et al.
Published: (2024)
by: Jaiswal, Siddharth D, et al.
Published: (2024)
CLIP-driven Zero-shot Learning with Ambiguous Labels
by: Fan, Jinfu, et al.
Published: (2026)
by: Fan, Jinfu, et al.
Published: (2026)
FoCLIP: A Feature-Space Misalignment Framework for CLIP-Based Image Manipulation and Detection
by: Chen, Yulin, et al.
Published: (2025)
by: Chen, Yulin, et al.
Published: (2025)
Enhancing Geo-localization for Crowdsourced Flood Imagery via LLM-Guided Attention
by: Xu, Fengyi, et al.
Published: (2025)
by: Xu, Fengyi, et al.
Published: (2025)
BuildingView: Constructing Urban Building Exteriors Databases with Street View Imagery and Multimodal Large Language Mode
by: Li, Zongrong, et al.
Published: (2024)
by: Li, Zongrong, et al.
Published: (2024)
GlocalCLIP: Object-agnostic Global-Local Prompt Learning for Zero-shot Anomaly Detection
by: Ham, Jiyul, et al.
Published: (2024)
by: Ham, Jiyul, et al.
Published: (2024)
EmoAssist: Emotional Assistant for Visual Impairment Community
by: Qi, Xingyu, et al.
Published: (2025)
by: Qi, Xingyu, et al.
Published: (2025)
Micro-AU CLIP: Fine-Grained Contrastive Learning from Local Independence to Global Dependency for Micro-Expression Action Unit Detection
by: Wei, Jinsheng, et al.
Published: (2026)
by: Wei, Jinsheng, et al.
Published: (2026)
Who's in and who's out? A case study of multimodal CLIP-filtering in DataComp
by: Hong, Rachel, et al.
Published: (2024)
by: Hong, Rachel, et al.
Published: (2024)
Ethical Considerations for the Military Use of Artificial Intelligence in Visual Reconnaissance
by: Anneken, Mathias, et al.
Published: (2025)
by: Anneken, Mathias, et al.
Published: (2025)
PIXELMOD: Improving Soft Moderation of Visual Misleading Information on Twitter
by: Paudel, Pujan, et al.
Published: (2024)
by: Paudel, Pujan, et al.
Published: (2024)
Object Detection and Tracking
by: Pranto, Md, et al.
Published: (2025)
by: Pranto, Md, et al.
Published: (2025)
Can GPT-4 Models Detect Misleading Visualizations?
by: Alexander, Jason, et al.
Published: (2024)
by: Alexander, Jason, et al.
Published: (2024)
Similar Items
-
SatCLIP: Global, General-Purpose Location Embeddings with Satellite Imagery
by: Klemmer, Konstantin, et al.
Published: (2023) -
Predicting Local Climate Zones using Urban Morphometrics and Satellite Imagery
by: Majer, Hugo, et al.
Published: (2026) -
Generalizable Slum Detection from Satellite Imagery with Mixture-of-Experts
by: Lee, Sumin, et al.
Published: (2025) -
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
by: Jha, Akshita, et al.
Published: (2024) -
From Drone Imagery to Livability Mapping: AI-powered Environment Perception in Rural China
by: Deng, Weihuan, et al.
Published: (2025)