See It from My Perspective: How Language Affects Cultural Bias in Image Understanding
Fuente:
arXiv
Salvato in:
| Autori principali: | Ananthram, Amith, Stengel-Eskin, Elias, Bansal, Mohit, McKeown, Kathleen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Mining Contextualized Visual Associations from Images for Creativity Understanding
di: Sahu, Ananya, et al.
Pubblicazione: (2025)
di: Sahu, Ananya, et al.
Pubblicazione: (2025)
PoSh: Using Scene Graphs To Guide LLMs-as-a-Judge For Detailed Image Descriptions
di: Ananthram, Amith, et al.
Pubblicazione: (2025)
di: Ananthram, Amith, et al.
Pubblicazione: (2025)
Does AI See like Art Historians? Interpreting How Vision Language Models Recognize Artistic Style
di: Limpijankit, Marvin, et al.
Pubblicazione: (2026)
di: Limpijankit, Marvin, et al.
Pubblicazione: (2026)
RotBench: Evaluating Multimodal Large Language Models on Identifying Image Rotation
di: Niu, Tianyi, et al.
Pubblicazione: (2025)
di: Niu, Tianyi, et al.
Pubblicazione: (2025)
Rephrase, Augment, Reason: Visual Grounding of Questions for Vision-Language Models
di: Prasad, Archiki, et al.
Pubblicazione: (2023)
di: Prasad, Archiki, et al.
Pubblicazione: (2023)
CAPTURe: Evaluating Spatial Reasoning in Vision Language Models via Occluded Object Counting
di: Pothiraj, Atin, et al.
Pubblicazione: (2025)
di: Pothiraj, Atin, et al.
Pubblicazione: (2025)
Enhancing Multimodal Affective Analysis with Learned Live Comment Features
di: Deng, Zhaoyuan, et al.
Pubblicazione: (2024)
di: Deng, Zhaoyuan, et al.
Pubblicazione: (2024)
Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training
di: Wan, David, et al.
Pubblicazione: (2024)
di: Wan, David, et al.
Pubblicazione: (2024)
GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
CLaMR: Contextualized Late-Interaction for Multimodal Content Retrieval
di: Wan, David, et al.
Pubblicazione: (2025)
di: Wan, David, et al.
Pubblicazione: (2025)
Multimodal Fact-Level Attribution for Verifiable Reasoning
di: Wan, David, et al.
Pubblicazione: (2026)
di: Wan, David, et al.
Pubblicazione: (2026)
VideoTree: Adaptive Tree-based Video Representation for LLM Reasoning on Long Videos
di: Wang, Ziyang, et al.
Pubblicazione: (2024)
di: Wang, Ziyang, et al.
Pubblicazione: (2024)
Language Models Identify Ambiguities and Exploit Loopholes
di: Choi, Jio, et al.
Pubblicazione: (2025)
di: Choi, Jio, et al.
Pubblicazione: (2025)
LACIE: Listener-Aware Finetuning for Confidence Calibration in Large Language Models
di: Stengel-Eskin, Elias, et al.
Pubblicazione: (2024)
di: Stengel-Eskin, Elias, et al.
Pubblicazione: (2024)
DART: Leveraging Multi-Agent Disagreement for Tool Recruitment in Multimodal Reasoning
di: Sivakumaran, Nithin, et al.
Pubblicazione: (2025)
di: Sivakumaran, Nithin, et al.
Pubblicazione: (2025)
Teaching Models to Balance Resisting and Accepting Persuasion
di: Stengel-Eskin, Elias, et al.
Pubblicazione: (2024)
di: Stengel-Eskin, Elias, et al.
Pubblicazione: (2024)
MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments
di: Wang, Han, et al.
Pubblicazione: (2026)
di: Wang, Han, et al.
Pubblicazione: (2026)
ReGAL: Refactoring Programs to Discover Generalizable Abstractions
di: Stengel-Eskin, Elias, et al.
Pubblicazione: (2024)
di: Stengel-Eskin, Elias, et al.
Pubblicazione: (2024)
UPCORE: Utility-Preserving Coreset Selection for Balanced Unlearning
di: Patil, Vaidehi, et al.
Pubblicazione: (2025)
di: Patil, Vaidehi, et al.
Pubblicazione: (2025)
The Sum Leaks More Than Its Parts: Compositional Privacy Risks and Mitigations in Multi-Agent Collaboration
di: Patil, Vaidehi, et al.
Pubblicazione: (2025)
di: Patil, Vaidehi, et al.
Pubblicazione: (2025)
Summarization of Opinionated Political Documents with Varied Perspectives
di: Deas, Nicholas, et al.
Pubblicazione: (2024)
di: Deas, Nicholas, et al.
Pubblicazione: (2024)
Social Orientation: A New Feature for Dialogue Analysis
di: Morrill, Todd, et al.
Pubblicazione: (2024)
di: Morrill, Todd, et al.
Pubblicazione: (2024)
AdaCAD: Adaptively Decoding to Balance Conflicts between Contextual and Parametric Knowledge
di: Wang, Han, et al.
Pubblicazione: (2024)
di: Wang, Han, et al.
Pubblicazione: (2024)
Soft Self-Consistency Improves Language Model Agents
di: Wang, Han, et al.
Pubblicazione: (2024)
di: Wang, Han, et al.
Pubblicazione: (2024)
Multi-Attribute Steering of Language Models via Targeted Intervention
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
Data Caricatures: On the Representation of African American Language in Pretraining Corpora
di: Deas, Nicholas, et al.
Pubblicazione: (2025)
di: Deas, Nicholas, et al.
Pubblicazione: (2025)
MAGDi: Structured Distillation of Multi-Agent Interaction Graphs Improves Reasoning in Smaller Language Models
di: Chen, Justin Chih-Yao, et al.
Pubblicazione: (2024)
di: Chen, Justin Chih-Yao, et al.
Pubblicazione: (2024)
Artificial Impressions: Evaluating Large Language Model Behavior Through the Lens of Trait Impressions
di: Deas, Nicholas, et al.
Pubblicazione: (2025)
di: Deas, Nicholas, et al.
Pubblicazione: (2025)
Fundamental Problems With Model Editing: How Should Rational Belief Revision Work in LLMs?
di: Hase, Peter, et al.
Pubblicazione: (2024)
di: Hase, Peter, et al.
Pubblicazione: (2024)
LASeR: Learning to Adaptively Select Reward Models with Multi-Armed Bandits
di: Nguyen, Duy, et al.
Pubblicazione: (2024)
di: Nguyen, Duy, et al.
Pubblicazione: (2024)
Are language models rational? The case of coherence norms and belief revision
di: Hofweber, Thomas, et al.
Pubblicazione: (2024)
di: Hofweber, Thomas, et al.
Pubblicazione: (2024)
Retrieval-Augmented Generation with Conflicting Evidence
di: Wang, Han, et al.
Pubblicazione: (2025)
di: Wang, Han, et al.
Pubblicazione: (2025)
A General Framework for Inference-time Scaling and Steering of Diffusion Models
di: Singhal, Raghav, et al.
Pubblicazione: (2025)
di: Singhal, Raghav, et al.
Pubblicazione: (2025)
Reranking-based Generation for Unbiased Perspective Summarization
di: Ri, Narutatsu, et al.
Pubblicazione: (2025)
di: Ri, Narutatsu, et al.
Pubblicazione: (2025)
DataEnvGym: Data Generation Agents in Teacher Environments with Student Feedback
di: Khan, Zaid, et al.
Pubblicazione: (2024)
di: Khan, Zaid, et al.
Pubblicazione: (2024)
Task-Circuit Quantization: Leveraging Knowledge Localization and Interpretability for Compression
di: Xiao, Hanqi, et al.
Pubblicazione: (2025)
di: Xiao, Hanqi, et al.
Pubblicazione: (2025)
Seeing Isn't Knowing: Do VLMs Know When Not to Answer Spatial Questions (and Why)?
di: Zhang, Yue, et al.
Pubblicazione: (2026)
di: Zhang, Yue, et al.
Pubblicazione: (2026)
GenerationPrograms: Fine-grained Attribution with Executable Programs
di: Wan, David, et al.
Pubblicazione: (2025)
di: Wan, David, et al.
Pubblicazione: (2025)
Generalized Correctness Models: Learning Calibrated and Model-Agnostic Correctness Predictors from Historical Patterns
di: Xiao, Hanqi, et al.
Pubblicazione: (2025)
di: Xiao, Hanqi, et al.
Pubblicazione: (2025)
MAMM-Refine: A Recipe for Improving Faithfulness in Generation with Multi-Agent Collaboration
di: Wan, David, et al.
Pubblicazione: (2025)
di: Wan, David, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Mining Contextualized Visual Associations from Images for Creativity Understanding
di: Sahu, Ananya, et al.
Pubblicazione: (2025) -
PoSh: Using Scene Graphs To Guide LLMs-as-a-Judge For Detailed Image Descriptions
di: Ananthram, Amith, et al.
Pubblicazione: (2025) -
Does AI See like Art Historians? Interpreting How Vision Language Models Recognize Artistic Style
di: Limpijankit, Marvin, et al.
Pubblicazione: (2026) -
RotBench: Evaluating Multimodal Large Language Models on Identifying Image Rotation
di: Niu, Tianyi, et al.
Pubblicazione: (2025) -
Rephrase, Augment, Reason: Visual Grounding of Questions for Vision-Language Models
di: Prasad, Archiki, et al.
Pubblicazione: (2023)