Incorporating Geo-Diverse Knowledge into Prompting for Increased Geographical Robustness in Object Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Buettner, Kyle, Malakouti, Sina, Li, Xiang Lorraine, Kovashka, Adriana |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Quantifying the Gaps Between Translation and Native Perception in Training for Multimodal, Multilingual Retrieval
di: Buettner, Kyle, et al.
Pubblicazione: (2024)
di: Buettner, Kyle, et al.
Pubblicazione: (2024)
A Multimodal Recaptioning Framework to Account for Perceptual Diversity Across Languages in Vision-Language Modeling
di: Buettner, Kyle, et al.
Pubblicazione: (2025)
di: Buettner, Kyle, et al.
Pubblicazione: (2025)
Role Bias in Diffusion Models: Diagnosing and Mitigating through Intermediate Decomposition
di: Malakouti, Sina, et al.
Pubblicazione: (2025)
di: Malakouti, Sina, et al.
Pubblicazione: (2025)
Culture in Action: Evaluating Text-to-Image Models through Social Activities
di: Malakouti, Sina, et al.
Pubblicazione: (2025)
di: Malakouti, Sina, et al.
Pubblicazione: (2025)
Integrating Audio Narrations to Strengthen Domain Generalization in Multimodal First-Person Action Recognition
di: Gungor, Cagri, et al.
Pubblicazione: (2024)
di: Gungor, Cagri, et al.
Pubblicazione: (2024)
Leveraging Large Models to Evaluate Novel Content: A Case Study on Advertisement Creativity
di: Hou, Zhaoyi Joey, et al.
Pubblicazione: (2025)
di: Hou, Zhaoyi Joey, et al.
Pubblicazione: (2025)
Benchmarking VLMs' Reasoning About Persuasive Atypical Images
di: Malakouti, Sina, et al.
Pubblicazione: (2024)
di: Malakouti, Sina, et al.
Pubblicazione: (2024)
Are Deep Learning Models Robust to Partial Object Occlusion in Visual Recognition Tasks?
di: Kassaw, Kaleb, et al.
Pubblicazione: (2024)
di: Kassaw, Kaleb, et al.
Pubblicazione: (2024)
GeoChain: Multimodal Chain-of-Thought for Geographic Reasoning
di: Yerramilli, Sahiti, et al.
Pubblicazione: (2025)
di: Yerramilli, Sahiti, et al.
Pubblicazione: (2025)
RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual Learning
di: Hong, Kiseong, et al.
Pubblicazione: (2025)
di: Hong, Kiseong, et al.
Pubblicazione: (2025)
ECOR: Explainable CLIP for Object Recognition
di: Rasekh, Ali, et al.
Pubblicazione: (2024)
di: Rasekh, Ali, et al.
Pubblicazione: (2024)
Increasing the Robustness of Model Predictions to Missing Sensors in Earth Observation
di: Mena, Francisco, et al.
Pubblicazione: (2024)
di: Mena, Francisco, et al.
Pubblicazione: (2024)
Meta-Prompting for Automating Zero-shot Visual Recognition with LLMs
di: Mirza, M. Jehanzeb, et al.
Pubblicazione: (2024)
di: Mirza, M. Jehanzeb, et al.
Pubblicazione: (2024)
Provably Better Explanations with Optimized Aggregation of Feature Attributions
di: Decker, Thomas, et al.
Pubblicazione: (2024)
di: Decker, Thomas, et al.
Pubblicazione: (2024)
Visual Modality Prompt for Adapting Vision-Language Object Detectors
di: Medeiros, Heitor R., et al.
Pubblicazione: (2024)
di: Medeiros, Heitor R., et al.
Pubblicazione: (2024)
Learning Generalizable Prompt for CLIP with Class Similarity Knowledge
di: Jung, Sehun, et al.
Pubblicazione: (2025)
di: Jung, Sehun, et al.
Pubblicazione: (2025)
RIDE: Retinex-Informed Decoupling for Exposing Concealed Objects
di: He, Chunming, et al.
Pubblicazione: (2026)
di: He, Chunming, et al.
Pubblicazione: (2026)
Diverse Prototypical Ensembles Improve Robustness to Subpopulation Shift
di: To, Minh Nguyen Nhat, et al.
Pubblicazione: (2025)
di: To, Minh Nguyen Nhat, et al.
Pubblicazione: (2025)
Segment Concealed Objects with Incomplete Supervision
di: He, Chunming, et al.
Pubblicazione: (2025)
di: He, Chunming, et al.
Pubblicazione: (2025)
pMoE: Prompting Diverse Experts Together Wins More in Visual Adaptation
di: Mo, Shentong, et al.
Pubblicazione: (2026)
di: Mo, Shentong, et al.
Pubblicazione: (2026)
Robust Box Prompt based SAM for Medical Image Segmentation
di: Huang, Yuhao, et al.
Pubblicazione: (2024)
di: Huang, Yuhao, et al.
Pubblicazione: (2024)
Towards Evaluating Robustness of Prompt Adherence in Text to Image Models
di: Vemishetty, Sujith, et al.
Pubblicazione: (2025)
di: Vemishetty, Sujith, et al.
Pubblicazione: (2025)
OCT-SelfNet: A Self-Supervised Framework with Multi-Modal Datasets for Generalized and Robust Retinal Disease Detection
di: Jannat, Fatema-E, et al.
Pubblicazione: (2024)
di: Jannat, Fatema-E, et al.
Pubblicazione: (2024)
OBSER: Object-Based Sub-Environment Recognition for Zero-Shot Environmental Inference
di: Choi, Won-Seok, et al.
Pubblicazione: (2025)
di: Choi, Won-Seok, et al.
Pubblicazione: (2025)
Deep Models for Multi-View 3D Object Recognition: A Review
di: Alzahrani, Mona, et al.
Pubblicazione: (2024)
di: Alzahrani, Mona, et al.
Pubblicazione: (2024)
SPARKE: Scalable Prompt-Aware Diversity and Novelty Guidance in Diffusion Models via RKE Score
di: Jalali, Mohammad, et al.
Pubblicazione: (2025)
di: Jalali, Mohammad, et al.
Pubblicazione: (2025)
Improving Zero-shot Generalization of Learned Prompts via Unsupervised Knowledge Distillation
di: Mistretta, Marco, et al.
Pubblicazione: (2024)
di: Mistretta, Marco, et al.
Pubblicazione: (2024)
PFM-VEPAR: Prompting Foundation Models for RGB-Event Camera based Pedestrian Attribute Recognition
di: Xu, Minghe, et al.
Pubblicazione: (2026)
di: Xu, Minghe, et al.
Pubblicazione: (2026)
Language-Driven Active Learning for Diverse Open-Set 3D Object Detection
di: Greer, Ross, et al.
Pubblicazione: (2024)
di: Greer, Ross, et al.
Pubblicazione: (2024)
CPA-Enhancer: Chain-of-Thought Prompted Adaptive Enhancer for Object Detection under Unknown Degradations
di: Zhang, Yuwei, et al.
Pubblicazione: (2024)
di: Zhang, Yuwei, et al.
Pubblicazione: (2024)
Aligning Logits Generatively for Principled Black-Box Knowledge Distillation
di: Ma, Jing, et al.
Pubblicazione: (2022)
di: Ma, Jing, et al.
Pubblicazione: (2022)
Federated Learning with Heterogeneous Data Handling for Robust Vehicular Object Detection
di: Khalil, Ahmad, et al.
Pubblicazione: (2024)
di: Khalil, Ahmad, et al.
Pubblicazione: (2024)
Large-image Object Detection for Fine-grained Recognition of Punches Patterns in Medieval Panel Painting
di: Bruegger, Josh, et al.
Pubblicazione: (2025)
di: Bruegger, Josh, et al.
Pubblicazione: (2025)
A Robust Incomplete Multimodal Low-Rank Adaptation Approach for Emotion Recognition
di: Zhao, Xinkui, et al.
Pubblicazione: (2025)
di: Zhao, Xinkui, et al.
Pubblicazione: (2025)
ImageNet-D: Benchmarking Neural Network Robustness on Diffusion Synthetic Object
di: Zhang, Chenshuang, et al.
Pubblicazione: (2024)
di: Zhang, Chenshuang, et al.
Pubblicazione: (2024)
Known Meets Unknown: Mitigating Overconfidence in Open Set Recognition
di: Zhao, Dongdong, et al.
Pubblicazione: (2025)
di: Zhao, Dongdong, et al.
Pubblicazione: (2025)
Residual SODAP: Residual Self-Organizing Domain-Adaptive Prompting with Structural Knowledge Preservation for Continual Learning
di: Oh, Gyutae, et al.
Pubblicazione: (2026)
di: Oh, Gyutae, et al.
Pubblicazione: (2026)
LeafLife: An Explainable Deep Learning Framework with Robustness for Grape Leaf Disease Recognition
di: Alam, B. M. Shahria, et al.
Pubblicazione: (2026)
di: Alam, B. M. Shahria, et al.
Pubblicazione: (2026)
RoHyDR: Robust Hybrid Diffusion Recovery for Incomplete Multimodal Emotion Recognition
di: Jin, Yuehan, et al.
Pubblicazione: (2025)
di: Jin, Yuehan, et al.
Pubblicazione: (2025)
One Prompt Word is Enough to Boost Adversarial Robustness for Pre-trained Vision-Language Models
di: Li, Lin, et al.
Pubblicazione: (2024)
di: Li, Lin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Quantifying the Gaps Between Translation and Native Perception in Training for Multimodal, Multilingual Retrieval
di: Buettner, Kyle, et al.
Pubblicazione: (2024) -
A Multimodal Recaptioning Framework to Account for Perceptual Diversity Across Languages in Vision-Language Modeling
di: Buettner, Kyle, et al.
Pubblicazione: (2025) -
Role Bias in Diffusion Models: Diagnosing and Mitigating through Intermediate Decomposition
di: Malakouti, Sina, et al.
Pubblicazione: (2025) -
Culture in Action: Evaluating Text-to-Image Models through Social Activities
di: Malakouti, Sina, et al.
Pubblicazione: (2025) -
Integrating Audio Narrations to Strengthen Domain Generalization in Multimodal First-Person Action Recognition
di: Gungor, Cagri, et al.
Pubblicazione: (2024)