Uncovering Cultural Representation Disparities in Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kadiyala, Ram Mohan Rao, Gupta, Siddhant, Purbey, Jebish, Yadav, Srishti, Debnath, Suman, Salamanca, Alejandro, Elliott, Desmond |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DSBC : Data Science task Benchmarking with Context engineering
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
Improving Multilingual Capabilities with Cultural and Local Knowledge in Large Language Models While Enhancing Native Performance
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
1-800-SHARED-TASKS at RegNLP: Lexical Reranking of Semantic Retrieval (LeSeR) for Regulatory Question Answering
by: Purbey, Jebish, et al.
Published: (2024)
by: Purbey, Jebish, et al.
Published: (2024)
SeQwen at the Financial Misinformation Detection Challenge Task: Sequential Learning for Claim Verification and Explanation Generation in Financial Domains
by: Purbey, Jebish, et al.
Published: (2024)
by: Purbey, Jebish, et al.
Published: (2024)
1-800-SHARED-TASKS @ NLU of Devanagari Script Languages: Detection of Language, Hate Speech, and Targets using LLMs
by: Purbey, Jebish, et al.
Published: (2024)
by: Purbey, Jebish, et al.
Published: (2024)
Large Language Models for Cross-lingual Emotion Detection
by: Kadiyala, Ram Mohan Rao
Published: (2024)
by: Kadiyala, Ram Mohan Rao
Published: (2024)
Efficient Test-Time Scaling for Small Vision-Language Models
by: Kaya, Mehmet Onurcan, et al.
Published: (2025)
by: Kaya, Mehmet Onurcan, et al.
Published: (2025)
Evaluation of Cultural Competence of Vision-Language Models
by: Yadav, Srishti, et al.
Published: (2025)
by: Yadav, Srishti, et al.
Published: (2025)
RKadiyala at SemEval-2024 Task 8: Black-Box Word-Level Text Boundary Detection in Partially Machine Generated Texts
by: Kadiyala, Ram Mohan Rao
Published: (2024)
by: Kadiyala, Ram Mohan Rao
Published: (2024)
LLMind: Bio-inspired Training-free Adaptive Visual Representations for Vision-Language Models
by: Debnath, Soumyaratna, et al.
Published: (2026)
by: Debnath, Soumyaratna, et al.
Published: (2026)
Robust and Fine-Grained Detection of AI Generated Texts
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models
by: Wang, Haoming, et al.
Published: (2026)
by: Wang, Haoming, et al.
Published: (2026)
Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models
by: Villegas, Danae Sánchez, et al.
Published: (2026)
by: Villegas, Danae Sánchez, et al.
Published: (2026)
Cross-modal Information Flow in Multimodal Large Language Models
by: Zhang, Zhi, et al.
Published: (2024)
by: Zhang, Zhi, et al.
Published: (2024)
Survey of Cultural Awareness in Language Models: Text and Beyond
by: Pawar, Siddhesh, et al.
Published: (2024)
by: Pawar, Siddhesh, et al.
Published: (2024)
ImageChain: Advancing Sequential Image-to-Text Reasoning in Multimodal Large Language Models
by: Villegas, Danae Sánchez, et al.
Published: (2025)
by: Villegas, Danae Sánchez, et al.
Published: (2025)
Long Story Short: Disentangling Compositionality and Long-Caption Understanding in Contrastive VLMs
by: Salazar, Israfel, et al.
Published: (2025)
by: Salazar, Israfel, et al.
Published: (2025)
Uncovering Bias in Large Vision-Language Models at Scale with Counterfactuals
by: Howard, Phillip, et al.
Published: (2024)
by: Howard, Phillip, et al.
Published: (2024)
Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization
by: Bharadwaj, Siddhant, et al.
Published: (2026)
by: Bharadwaj, Siddhant, et al.
Published: (2026)
Uncovering Bias in Large Vision-Language Models with Counterfactuals
by: Howard, Phillip, et al.
Published: (2024)
by: Howard, Phillip, et al.
Published: (2024)
Seeing What Tastes Good: Revisiting Multimodal Distributional Semantics in the Billion Parameter Era
by: Oneata, Dan, et al.
Published: (2025)
by: Oneata, Dan, et al.
Published: (2025)
Kaleidoscope: In-language Exams for Massively Multilingual Vision Evaluation
by: Salazar, Israfel, et al.
Published: (2025)
by: Salazar, Israfel, et al.
Published: (2025)
Driver Activity Classification Using Generalizable Representations from Vision-Language Models
by: Greer, Ross, et al.
Published: (2024)
by: Greer, Ross, et al.
Published: (2024)
SAMwave: Wavelet-Driven Feature Enrichment for Effective Adaptation of Segment Anything Model
by: Yadav, Saurabh, et al.
Published: (2025)
by: Yadav, Saurabh, et al.
Published: (2025)
The Spatial Blindspot of Vision-Language Models
by: Alam, Nahid, et al.
Published: (2026)
by: Alam, Nahid, et al.
Published: (2026)
Uncovering Intrinsic Capabilities: A Paradigm for Data Curation in Vision-Language Models
by: Li, Junjie, et al.
Published: (2025)
by: Li, Junjie, et al.
Published: (2025)
Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2026)
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2026)
Visual Alignment of Medical Vision-Language Models for Grounded Radiology Report Generation
by: Bose, Sarosij, et al.
Published: (2025)
by: Bose, Sarosij, et al.
Published: (2025)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
by: Bansal, Siddhant, et al.
Published: (2024)
by: Bansal, Siddhant, et al.
Published: (2024)
Adaptive Multi Scale Document Binarisation Using Vision Mamba
by: Azfar, Mohd., et al.
Published: (2024)
by: Azfar, Mohd., et al.
Published: (2024)
Cultural Counterfactuals: Evaluating Cultural Biases in Large Vision-Language Models with Counterfactual Examples
by: Howard, Phillip, et al.
Published: (2026)
by: Howard, Phillip, et al.
Published: (2026)
Joint Vision-Language Social Bias Removal for CLIP
by: Zhang, Haoyu, et al.
Published: (2024)
by: Zhang, Haoyu, et al.
Published: (2024)
1024m at SMM4H 2024: Tasks 3, 5 & 6 -- Ensembles of Transformers and Large Language Models for Medical Text Classification
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2024)
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2024)
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models
by: Yadav, Ankit, et al.
Published: (2025)
by: Yadav, Ankit, et al.
Published: (2025)
Uncovering Linguistic Fragility in Vision-Language-Action Models via Diversity-Aware Red Teaming
by: Tong, Baoshun, et al.
Published: (2026)
by: Tong, Baoshun, et al.
Published: (2026)
Seeing Syntax: Uncovering Syntactic Learning Limitations in Vision-Language Models
by: Dumpala, Sri Harsha, et al.
Published: (2024)
by: Dumpala, Sri Harsha, et al.
Published: (2024)
Refusal as Silence: Gendered Disparities in Vision-Language Model Responses
by: Luo, Sha, et al.
Published: (2024)
by: Luo, Sha, et al.
Published: (2024)
Uncovering Brain-Like Hierarchical Patterns in Vision-Language Models through fMRI-Based Neural Encoding
by: Ren, Yudan, et al.
Published: (2025)
by: Ren, Yudan, et al.
Published: (2025)
The Eye of Sherlock Holmes: Uncovering User Private Attribute Profiling via Vision-Language Model Agentic Framework
by: Liu, Feiran, et al.
Published: (2025)
by: Liu, Feiran, et al.
Published: (2025)
Foveated Reasoning: Stateful, Action-based Visual Focusing for Vision-Language Models
by: Min, Juhong, et al.
Published: (2026)
by: Min, Juhong, et al.
Published: (2026)
Similar Items
-
DSBC : Data Science task Benchmarking with Context engineering
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025) -
Improving Multilingual Capabilities with Cultural and Local Knowledge in Large Language Models While Enhancing Native Performance
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025) -
1-800-SHARED-TASKS at RegNLP: Lexical Reranking of Semantic Retrieval (LeSeR) for Regulatory Question Answering
by: Purbey, Jebish, et al.
Published: (2024) -
SeQwen at the Financial Misinformation Detection Challenge Task: Sequential Learning for Claim Verification and Explanation Generation in Financial Domains
by: Purbey, Jebish, et al.
Published: (2024) -
1-800-SHARED-TASKS @ NLU of Devanagari Script Languages: Detection of Language, Hate Speech, and Targets using LLMs
by: Purbey, Jebish, et al.
Published: (2024)