Contrasting Cognitive Styles in Vision-Language Models: Holistic Attention in Japanese Versus Analytical Focus in English
Fuente:
arXiv
Guardado en:
| Autores principales: | Sabir, Ahmed, Gasper, Azinovič, Loem, Mengsay, Sharma, Rajesh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Joint Extraction Matters: Prompt-Based Visual Question Answering for Multi-Field Document Information Extraction
por: Loem, Mengsay, et al.
Publicado: (2025)
por: Loem, Mengsay, et al.
Publicado: (2025)
SAIE Framework: Support Alone Isn't Enough -- Advancing LLM Training with Adversarial Remarks
por: Loem, Mengsay, et al.
Publicado: (2023)
por: Loem, Mengsay, et al.
Publicado: (2023)
Likelihood-based Mitigation of Evaluation Bias in Large Language Models
por: Oi, Masanari, et al.
Publicado: (2024)
por: Oi, Masanari, et al.
Publicado: (2024)
Building a Large Japanese Web Corpus for Large Language Models
por: Okazaki, Naoaki, et al.
Publicado: (2024)
por: Okazaki, Naoaki, et al.
Publicado: (2024)
Exploring Gender Bias Beyond Occupational Titles
por: Sabir, Ahmed, et al.
Publicado: (2025)
por: Sabir, Ahmed, et al.
Publicado: (2025)
Continual Pre-Training for Cross-Lingual LLM Adaptation: Enhancing Japanese Language Capabilities
por: Fujii, Kazuki, et al.
Publicado: (2024)
por: Fujii, Kazuki, et al.
Publicado: (2024)
The Confidence Trap: Gender Bias and Predictive Certainty in LLMs
por: Sabir, Ahmed, et al.
Publicado: (2026)
por: Sabir, Ahmed, et al.
Publicado: (2026)
How Effectively Do LLMs Extract Feature-Sentiment Pairs from App Reviews?
por: Shah, Faiz Ali, et al.
Publicado: (2024)
por: Shah, Faiz Ali, et al.
Publicado: (2024)
Reproducibility Study of Large Language Model Bayesian Optimization
por: Rychert, Adam, et al.
Publicado: (2025)
por: Rychert, Adam, et al.
Publicado: (2025)
Attention, Please! PixelSHAP Reveals What Vision-Language Models Actually Focus On
por: Goldshmidt, Roni
Publicado: (2025)
por: Goldshmidt, Roni
Publicado: (2025)
WAON: A Large-Scale Japanese Image-Text Dataset for Cultural Adaptation in Contrastive Vision-Language Models
por: Sugiura, Issa, et al.
Publicado: (2025)
por: Sugiura, Issa, et al.
Publicado: (2025)
MuDAF: Long-Context Multi-Document Attention Focusing through Contrastive Learning on Attention Heads
por: Liu, Weihao, et al.
Publicado: (2025)
por: Liu, Weihao, et al.
Publicado: (2025)
VALOR-EVAL: Holistic Coverage and Faithfulness Evaluation of Large Vision-Language Models
por: Qiu, Haoyi, et al.
Publicado: (2024)
por: Qiu, Haoyi, et al.
Publicado: (2024)
Focus on the Core: Empowering Diffusion Large Language Models by Self-Contrast
por: Feng, Jinyuan, et al.
Publicado: (2026)
por: Feng, Jinyuan, et al.
Publicado: (2026)
Focusing on Language: Revealing and Exploiting Language Attention Heads in Multilingual Large Language Models
por: Liu, Xin, et al.
Publicado: (2025)
por: Liu, Xin, et al.
Publicado: (2025)
Capturing Human Cognitive Styles with Language: Towards an Experimental Evaluation Paradigm
por: Varadarajan, Vasudha, et al.
Publicado: (2025)
por: Varadarajan, Vasudha, et al.
Publicado: (2025)
Epistemology of Language Models: Do Language Models Have Holistic Knowledge?
por: Kim, Minsu, et al.
Publicado: (2024)
por: Kim, Minsu, et al.
Publicado: (2024)
Focus Directions Make Your Language Models Pay More Attention to Relevant Contexts
por: Zhu, Youxiang, et al.
Publicado: (2025)
por: Zhu, Youxiang, et al.
Publicado: (2025)
ConlangCrafter: Constructing Languages with a Multi-Hop LLM Pipeline
por: Alper, Morris, et al.
Publicado: (2025)
por: Alper, Morris, et al.
Publicado: (2025)
Heron-Bench: A Benchmark for Evaluating Vision Language Models in Japanese
por: Inoue, Yuichi, et al.
Publicado: (2024)
por: Inoue, Yuichi, et al.
Publicado: (2024)
Text Detoxification as Style Transfer in English and Hindi
por: Mukherjee, Sourabrata, et al.
Publicado: (2024)
por: Mukherjee, Sourabrata, et al.
Publicado: (2024)
The Logovista English-Japanese Machine Translation System
por: Wright, Barton D.
Publicado: (2026)
por: Wright, Barton D.
Publicado: (2026)
G-Loss: Graph-Guided Fine-Tuning of Language Models
por: Sharma, Aditya, et al.
Publicado: (2026)
por: Sharma, Aditya, et al.
Publicado: (2026)
Exploring the encoding of linguistic representations in the Fully-Connected Layer of generative CNNs for Speech
por: Šegedin, Bruno Ferenc, et al.
Publicado: (2025)
por: Šegedin, Bruno Ferenc, et al.
Publicado: (2025)
How Well Do LLMs Imitate Human Writing Style?
por: Jemama, Rebira, et al.
Publicado: (2025)
por: Jemama, Rebira, et al.
Publicado: (2025)
Large Linguistic Models: Investigating LLMs' metalinguistic abilities
por: Beguš, Gašper, et al.
Publicado: (2023)
por: Beguš, Gašper, et al.
Publicado: (2023)
Towards Large Language Model driven Reference-less Translation Evaluation for English and Indian Languages
por: Mujadia, Vandan, et al.
Publicado: (2024)
por: Mujadia, Vandan, et al.
Publicado: (2024)
CLoVe: Encoding Compositional Language in Contrastive Vision-Language Models
por: Castro, Santiago, et al.
Publicado: (2024)
por: Castro, Santiago, et al.
Publicado: (2024)
Unsupervised Learning and Representation of Mandarin Tonal Categories by a Generative CNN
por: Schenck, Kai, et al.
Publicado: (2025)
por: Schenck, Kai, et al.
Publicado: (2025)
AhaKV: Adaptive Holistic Attention-Driven KV Cache Eviction for Efficient Inference of Large Language Models
por: Gu, Yifeng, et al.
Publicado: (2025)
por: Gu, Yifeng, et al.
Publicado: (2025)
Investigating Spatial Attention Bias in Vision-Language Models
por: Chaudhary, Aryan, et al.
Publicado: (2025)
por: Chaudhary, Aryan, et al.
Publicado: (2025)
StyleBench: Evaluating Speech Language Models on Conversational Speaking Style Control
por: Zhao, Haishu, et al.
Publicado: (2026)
por: Zhao, Haishu, et al.
Publicado: (2026)
Contrastive Learning of English Language and Crystal Graphs for Multimodal Representation of Materials Knowledge
por: Park, Yang Jeong, et al.
Publicado: (2025)
por: Park, Yang Jeong, et al.
Publicado: (2025)
PHYBench: Holistic Evaluation of Physical Perception and Reasoning in Large Language Models
por: Qiu, Shi, et al.
Publicado: (2025)
por: Qiu, Shi, et al.
Publicado: (2025)
Focus Matters: Phase-Aware Suppression for Hallucination in Vision-Language Models
por: Kim, Sohyeon, et al.
Publicado: (2026)
por: Kim, Sohyeon, et al.
Publicado: (2026)
Japanese-English Sentence Translation Exercises Dataset for Automatic Grading
por: Miura, Naoki, et al.
Publicado: (2024)
por: Miura, Naoki, et al.
Publicado: (2024)
AHELM: A Holistic Evaluation of Audio-Language Models
por: Lee, Tony, et al.
Publicado: (2025)
por: Lee, Tony, et al.
Publicado: (2025)
Large Language Model Safety: A Holistic Survey
por: Shi, Dan, et al.
Publicado: (2024)
por: Shi, Dan, et al.
Publicado: (2024)
SmartTrim: Adaptive Tokens and Attention Pruning for Efficient Vision-Language Models
por: Wang, Zekun, et al.
Publicado: (2023)
por: Wang, Zekun, et al.
Publicado: (2023)
MedHELM: Holistic Evaluation of Large Language Models for Medical Tasks
por: Bedi, Suhana, et al.
Publicado: (2025)
por: Bedi, Suhana, et al.
Publicado: (2025)
Ejemplares similares
-
Joint Extraction Matters: Prompt-Based Visual Question Answering for Multi-Field Document Information Extraction
por: Loem, Mengsay, et al.
Publicado: (2025) -
SAIE Framework: Support Alone Isn't Enough -- Advancing LLM Training with Adversarial Remarks
por: Loem, Mengsay, et al.
Publicado: (2023) -
Likelihood-based Mitigation of Evaluation Bias in Large Language Models
por: Oi, Masanari, et al.
Publicado: (2024) -
Building a Large Japanese Web Corpus for Large Language Models
por: Okazaki, Naoaki, et al.
Publicado: (2024) -
Exploring Gender Bias Beyond Occupational Titles
por: Sabir, Ahmed, et al.
Publicado: (2025)