Salvato in:
| Autore principale: | Joshi, Harsh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2412.18635 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Modular, On-Site Solutions with Lightweight Anomaly Detection for Sustainable Nutrient Management in Agriculture
di: Cohen, Abigail R., et al.
Pubblicazione: (2025)
di: Cohen, Abigail R., et al.
Pubblicazione: (2025)
Intelligent Systems in Neuroimaging: Pioneering AI Techniques for Brain Tumor Detection
di: Islam, Md. Mohaiminul, et al.
Pubblicazione: (2025)
di: Islam, Md. Mohaiminul, et al.
Pubblicazione: (2025)
Advancing Cucumber Disease Detection in Agriculture through Machine Vision and Drone Technology
di: Rahman, Syada Tasfia, et al.
Pubblicazione: (2024)
di: Rahman, Syada Tasfia, et al.
Pubblicazione: (2024)
Vision-Language Models for Autonomous Driving: CLIP-Based Dynamic Scene Understanding
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2025)
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2025)
Surgeons Are Indian Males and Speech Therapists Are White Females: Auditing Biases in Vision-Language Models for Healthcare Professionals
di: Siddiqui, Zohaib Hasan, et al.
Pubblicazione: (2025)
di: Siddiqui, Zohaib Hasan, et al.
Pubblicazione: (2025)
Identifying Implicit Social Biases in Vision-Language Models
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
di: Hamidieh, Kimia, et al.
Pubblicazione: (2024)
Vision-Language Models under Cultural and Inclusive Considerations
di: Karamolegkou, Antonia, et al.
Pubblicazione: (2024)
di: Karamolegkou, Antonia, et al.
Pubblicazione: (2024)
Urban Socio-Semantic Segmentation with Vision-Language Reasoning
di: Wang, Yu, et al.
Pubblicazione: (2026)
di: Wang, Yu, et al.
Pubblicazione: (2026)
Climatic & Anthropogenic Hazards to the Nasca World Heritage: Application of Remote Sensing, AI, and Flood Modelling
di: Sakai, Masato, et al.
Pubblicazione: (2024)
di: Sakai, Masato, et al.
Pubblicazione: (2024)
Assessing Greenspace Attractiveness with ChatGPT, Claude, and Gemini: Do AI Models Reflect Human Perceptions?
di: Malekzadeh, Milad, et al.
Pubblicazione: (2025)
di: Malekzadeh, Milad, et al.
Pubblicazione: (2025)
Locating Demographic Bias at the Attention-Head Level in CLIP's Vision Encoder
di: Yasser, Alaa, et al.
Pubblicazione: (2026)
di: Yasser, Alaa, et al.
Pubblicazione: (2026)
Beyond Translation: Cross-Cultural Meme Transcreation with Vision-Language Models
di: Zhao, Yuming, et al.
Pubblicazione: (2026)
di: Zhao, Yuming, et al.
Pubblicazione: (2026)
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
di: Madasu, Avinash, et al.
Pubblicazione: (2025)
di: Madasu, Avinash, et al.
Pubblicazione: (2025)
NuWa: Deriving Lightweight Task-Specific Vision Transformers for Edge Devices
di: Wei, Ziteng, et al.
Pubblicazione: (2025)
di: Wei, Ziteng, et al.
Pubblicazione: (2025)
The Weaponization of Computer Vision: Tracing Military-Surveillance Ties through Conference Sponsorship
di: Garcia, Noa, et al.
Pubblicazione: (2026)
di: Garcia, Noa, et al.
Pubblicazione: (2026)
Computer Vision for Multimedia Geolocation in Human Trafficking Investigation: A Systematic Literature Review
di: Bamigbade, Opeyemi, et al.
Pubblicazione: (2024)
di: Bamigbade, Opeyemi, et al.
Pubblicazione: (2024)
Edge-Enhanced Vision Transformer Framework for Accurate AI-Generated Image Detection
di: Das, Dabbrata, et al.
Pubblicazione: (2025)
di: Das, Dabbrata, et al.
Pubblicazione: (2025)
Closing the Gap: Data-Centric Fine-Tuning of Vision Language Models for the Standardized Exam Questions
di: Sert, Egemen, et al.
Pubblicazione: (2025)
di: Sert, Egemen, et al.
Pubblicazione: (2025)
Social Perception of Faces in a Vision-Language Model
di: Hausladen, Carina I., et al.
Pubblicazione: (2024)
di: Hausladen, Carina I., et al.
Pubblicazione: (2024)
An AI-Enabled Framework Within Reach for Enhancing Healthcare Sustainability and Fairness
di: Huang, Bin, et al.
Pubblicazione: (2024)
di: Huang, Bin, et al.
Pubblicazione: (2024)
Safer Prompts: Reducing Risks from Memorization in Visual Generative AI
di: Reissinger, Lena, et al.
Pubblicazione: (2025)
di: Reissinger, Lena, et al.
Pubblicazione: (2025)
AI's Blind Spots: Geographic Knowledge and Diversity Deficit in Generated Urban Scenario
di: Beneduce, Ciro, et al.
Pubblicazione: (2025)
di: Beneduce, Ciro, et al.
Pubblicazione: (2025)
Learning from Limited and Imperfect Data
di: Rangwani, Harsh
Pubblicazione: (2025)
di: Rangwani, Harsh
Pubblicazione: (2025)
Examining Monitoring System: Detecting Abnormal Behavior In Online Examinations
di: Ngo, Dinh An, et al.
Pubblicazione: (2024)
di: Ngo, Dinh An, et al.
Pubblicazione: (2024)
On Occlusions in Video Action Detection: Benchmark Datasets And Training Recipes
di: Modi, Rajat, et al.
Pubblicazione: (2024)
di: Modi, Rajat, et al.
Pubblicazione: (2024)
Networking Systems for Video Anomaly Detection: A Tutorial and Survey
di: Liu, Jing, et al.
Pubblicazione: (2024)
di: Liu, Jing, et al.
Pubblicazione: (2024)
Refusal as Silence: Gendered Disparities in Vision-Language Model Responses
di: Luo, Sha, et al.
Pubblicazione: (2024)
di: Luo, Sha, et al.
Pubblicazione: (2024)
Silicon Minds versus Human Hearts: The Wisdom of Crowds Beats the Wisdom of AI in Emotion Recognition
di: Akben, Mustafa, et al.
Pubblicazione: (2025)
di: Akben, Mustafa, et al.
Pubblicazione: (2025)
Smiling Women Pitching Down: Auditing Representational and Presentational Gender Biases in Image Generative AI
di: Sun, Luhang, et al.
Pubblicazione: (2023)
di: Sun, Luhang, et al.
Pubblicazione: (2023)
AI-based Multimodal Biometrics for Detecting Smartphone Distractions: Application to Online Learning
di: Becerra, Alvaro, et al.
Pubblicazione: (2025)
di: Becerra, Alvaro, et al.
Pubblicazione: (2025)
Adapting an Artificial Intelligence Sexually Transmitted Diseases Symptom Checker Tool for Mpox Detection: The HeHealth Experience
di: Tan, Rayner Kay Jin, et al.
Pubblicazione: (2024)
di: Tan, Rayner Kay Jin, et al.
Pubblicazione: (2024)
Aligning AI with Public Values: Deliberation and Decision-Making for Governing Multimodal LLMs in Political Video Analysis
di: Sharma, Tanusree, et al.
Pubblicazione: (2024)
di: Sharma, Tanusree, et al.
Pubblicazione: (2024)
VLLFL: A Vision-Language Model Based Lightweight Federated Learning Framework for Smart Agriculture
di: Li, Long, et al.
Pubblicazione: (2025)
di: Li, Long, et al.
Pubblicazione: (2025)
Open-Vocabulary X-ray Prohibited Item Detection via Fine-tuning CLIP
di: Lin, Shuyang, et al.
Pubblicazione: (2024)
di: Lin, Shuyang, et al.
Pubblicazione: (2024)
Demographic Bias of Expert-Level Vision-Language Foundation Models in Medical Imaging
di: Yang, Yuzhe, et al.
Pubblicazione: (2024)
di: Yang, Yuzhe, et al.
Pubblicazione: (2024)
Decoding Tourist Perception in Historic Urban Quarters with Multimodal Social Media Data: An AI-Based Framework and Evidence from Shanghai
di: Tan, Kaizhen, et al.
Pubblicazione: (2025)
di: Tan, Kaizhen, et al.
Pubblicazione: (2025)
Dataset Scale and Societal Consistency Mediate Facial Impression Bias in Vision-Language AI
di: Wolfe, Robert, et al.
Pubblicazione: (2024)
di: Wolfe, Robert, et al.
Pubblicazione: (2024)
LWMSCNN-SE: A Lightweight Multi-Scale Network for Efficient Maize Disease Classification on Edge Devices
di: Weloday, Fikadu, et al.
Pubblicazione: (2026)
di: Weloday, Fikadu, et al.
Pubblicazione: (2026)
TinyEcoWeedNet: Edge Efficient Real-Time Aerial Agricultural Weed Detection
di: Khater, Omar H., et al.
Pubblicazione: (2025)
di: Khater, Omar H., et al.
Pubblicazione: (2025)
AI-based System for Transforming text and sound to Educational Videos
di: ElAlami, M. E., et al.
Pubblicazione: (2026)
di: ElAlami, M. E., et al.
Pubblicazione: (2026)
Documenti analoghi
-
Modular, On-Site Solutions with Lightweight Anomaly Detection for Sustainable Nutrient Management in Agriculture
di: Cohen, Abigail R., et al.
Pubblicazione: (2025) -
Intelligent Systems in Neuroimaging: Pioneering AI Techniques for Brain Tumor Detection
di: Islam, Md. Mohaiminul, et al.
Pubblicazione: (2025) -
Advancing Cucumber Disease Detection in Agriculture through Machine Vision and Drone Technology
di: Rahman, Syada Tasfia, et al.
Pubblicazione: (2024) -
Vision-Language Models for Autonomous Driving: CLIP-Based Dynamic Scene Understanding
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2025) -
Surgeons Are Indian Males and Speech Therapists Are White Females: Auditing Biases in Vision-Language Models for Healthcare Professionals
di: Siddiqui, Zohaib Hasan, et al.
Pubblicazione: (2025)