ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Zhipin, Leiter, Christoph, Frey, Christian, Abdalla, Mohamed Hesham Ibrahim, Grabocka, Josif, Eger, Steffen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Zhyper: Factorized Hypernetworks for Conditioned LLM Fine-Tuning
di: Abdalla, M. H. I., et al.
Pubblicazione: (2025)
di: Abdalla, M. H. I., et al.
Pubblicazione: (2025)
PrExMe! Large Scale Prompt Exploration of Open Source LLMs for Machine Translation and Summarization Evaluation
di: Leiter, Christoph, et al.
Pubblicazione: (2024)
di: Leiter, Christoph, et al.
Pubblicazione: (2024)
DeepSeek-R1 vs. o3-mini: How Well can Reasoning LLMs Evaluate MT and Summarization?
di: Larionov, Daniil, et al.
Pubblicazione: (2025)
di: Larionov, Daniil, et al.
Pubblicazione: (2025)
BMX: Boosting Natural Language Generation Metrics with Explainability
di: Leiter, Christoph, et al.
Pubblicazione: (2022)
di: Leiter, Christoph, et al.
Pubblicazione: (2022)
CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks
di: Leiter, Christoph, et al.
Pubblicazione: (2025)
di: Leiter, Christoph, et al.
Pubblicazione: (2025)
GerAV: Towards New Heights in German Authorship Verification using Fine-Tuned LLMs on a New Benchmark
di: Kiefer, Lotta, et al.
Pubblicazione: (2026)
di: Kiefer, Lotta, et al.
Pubblicazione: (2026)
Towards Explainable Evaluation Metrics for Machine Translation
di: Leiter, Christoph, et al.
Pubblicazione: (2023)
di: Leiter, Christoph, et al.
Pubblicazione: (2023)
LLLMs: A Data-Driven Survey of Evolving Research on Limitations of Large Language Models
di: Kostikova, Aida, et al.
Pubblicazione: (2025)
di: Kostikova, Aida, et al.
Pubblicazione: (2025)
USCORE: An Effective Approach to Fully Unsupervised Evaluation Metrics for Machine Translation
di: Belouadi, Jonas, et al.
Pubblicazione: (2022)
di: Belouadi, Jonas, et al.
Pubblicazione: (2022)
BatchGEMBA: Token-Efficient Machine Translation Evaluation with Batched Prompting and Prompt Compression
di: Larionov, Daniil, et al.
Pubblicazione: (2025)
di: Larionov, Daniil, et al.
Pubblicazione: (2025)
PromptOptMe: Error-Aware Prompt Compression for LLM-based MT Evaluation Metrics
di: Larionov, Daniil, et al.
Pubblicazione: (2024)
di: Larionov, Daniil, et al.
Pubblicazione: (2024)
TikZilla: Scaling Text-to-TikZ with High-Quality Data and Reinforcement Learning
di: Greisinger, Christian, et al.
Pubblicazione: (2026)
di: Greisinger, Christian, et al.
Pubblicazione: (2026)
Cross-lingual Cross-temporal Summarization: Dataset, Models, Evaluation
di: Zhang, Ran, et al.
Pubblicazione: (2023)
di: Zhang, Ran, et al.
Pubblicazione: (2023)
PlaM: Training-Free Plateau-Guided Model Merging for Better Visual Grounding in MLLMs
di: Wang, Zijing, et al.
Pubblicazione: (2026)
di: Wang, Zijing, et al.
Pubblicazione: (2026)
Do Emotions Really Affect Argument Convincingness? A Dynamic Approach with LLM-based Manipulation Checks
di: Chen, Yanran, et al.
Pubblicazione: (2025)
di: Chen, Yanran, et al.
Pubblicazione: (2025)
ByGPT5: End-to-End Style-conditioned Poetry Generation with Token-free Language Models
di: Belouadi, Jonas, et al.
Pubblicazione: (2022)
di: Belouadi, Jonas, et al.
Pubblicazione: (2022)
LLM-based multi-agent poetry generation in non-cooperative environments
di: Zhang, Ran, et al.
Pubblicazione: (2024)
di: Zhang, Ran, et al.
Pubblicazione: (2024)
Ensembling Finetuned Language Models for Text Classification
di: Arango, Sebastian Pineda, et al.
Pubblicazione: (2024)
di: Arango, Sebastian Pineda, et al.
Pubblicazione: (2024)
NLLG Quarterly arXiv Report 09/24: What are the most influential current AI Papers?
di: Leiter, Christoph, et al.
Pubblicazione: (2024)
di: Leiter, Christoph, et al.
Pubblicazione: (2024)
Language-Conditioned Visual Grounding with CLIP Multilingual
di: de Curtò, J., et al.
Pubblicazione: (2026)
di: de Curtò, J., et al.
Pubblicazione: (2026)
SpatialViz-Bench: A Cognitively-Grounded Benchmark for Diagnosing Spatial Visualization in MLLMs
di: Wang, Siting, et al.
Pubblicazione: (2025)
di: Wang, Siting, et al.
Pubblicazione: (2025)
How Good Are LLMs for Literary Translation, Really? Literary Translation Evaluation with Humans and LLMs
di: Zhang, Ran, et al.
Pubblicazione: (2024)
di: Zhang, Ran, et al.
Pubblicazione: (2024)
Evaluating Diversity in Automatic Poetry Generation
di: Chen, Yanran, et al.
Pubblicazione: (2024)
di: Chen, Yanran, et al.
Pubblicazione: (2024)
Seeing is Believing: Rich-Context Hallucination Detection for MLLMs via Backward Visual Grounding
di: Guo, Pinxue, et al.
Pubblicazione: (2025)
di: Guo, Pinxue, et al.
Pubblicazione: (2025)
Agentic Jackal: Live Execution and Semantic Value Grounding for Text-to-JQL
di: Murali, Vishnu, et al.
Pubblicazione: (2026)
di: Murali, Vishnu, et al.
Pubblicazione: (2026)
Linguistically Grounded Analysis of Language Models using Shapley Head Values
di: Fekete, Marcell, et al.
Pubblicazione: (2024)
di: Fekete, Marcell, et al.
Pubblicazione: (2024)
Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks
di: Greco, Candida M., et al.
Pubblicazione: (2026)
di: Greco, Candida M., et al.
Pubblicazione: (2026)
CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs
di: Ye, Yangfan, et al.
Pubblicazione: (2026)
di: Ye, Yangfan, et al.
Pubblicazione: (2026)
Is there really a Citation Age Bias in NLP?
di: Nguyen, Hoa, et al.
Pubblicazione: (2024)
di: Nguyen, Hoa, et al.
Pubblicazione: (2024)
VIVA: A Benchmark for Vision-Grounded Decision-Making with Human Values
di: Hu, Zhe, et al.
Pubblicazione: (2024)
di: Hu, Zhe, et al.
Pubblicazione: (2024)
Beyond Preferences: Learning Alignment Principles Grounded in Human Reasons and Values
di: Bell, Henry, et al.
Pubblicazione: (2026)
di: Bell, Henry, et al.
Pubblicazione: (2026)
ScImage: How Good Are Multimodal Large Language Models at Scientific Text-to-Image Generation?
di: Zhang, Leixin, et al.
Pubblicazione: (2024)
di: Zhang, Leixin, et al.
Pubblicazione: (2024)
Beyond Unimodal Shortcuts: MLLMs as Cross-Modal Reasoners for Grounded Named Entity Recognition
di: Ma, Jinlong, et al.
Pubblicazione: (2026)
di: Ma, Jinlong, et al.
Pubblicazione: (2026)
Evaluating Large Language Models for Structured Science Summarization in the Open Research Knowledge Graph
di: Nechakhin, Vladyslav, et al.
Pubblicazione: (2024)
di: Nechakhin, Vladyslav, et al.
Pubblicazione: (2024)
Seeing Culture: A Benchmark for Visual Reasoning and Grounding
di: Satar, Burak, et al.
Pubblicazione: (2025)
di: Satar, Burak, et al.
Pubblicazione: (2025)
xCOMET-lite: Bridging the Gap Between Efficiency and Quality in Learned MT Evaluation Metrics
di: Larionov, Daniil, et al.
Pubblicazione: (2024)
di: Larionov, Daniil, et al.
Pubblicazione: (2024)
AutomaTikZ: Text-Guided Synthesis of Scientific Vector Graphics with TikZ
di: Belouadi, Jonas, et al.
Pubblicazione: (2023)
di: Belouadi, Jonas, et al.
Pubblicazione: (2023)
Framing Matters: Addressing Framing Sensitivity in Decision-Making through Behaviorally-Grounded Value Alignment
di: Hwang, Seojin, et al.
Pubblicazione: (2026)
di: Hwang, Seojin, et al.
Pubblicazione: (2026)
DermoGPT: Open Weights and Open Data for Morphology-Grounded Dermatological Reasoning MLLMs
di: Ru, Jinghan, et al.
Pubblicazione: (2026)
di: Ru, Jinghan, et al.
Pubblicazione: (2026)
Retrieve, Then Classify: Corpus-Grounded Automation of Clinical Value Set Authoring
di: Mukherjee, Sumit, et al.
Pubblicazione: (2026)
di: Mukherjee, Sumit, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Zhyper: Factorized Hypernetworks for Conditioned LLM Fine-Tuning
di: Abdalla, M. H. I., et al.
Pubblicazione: (2025) -
PrExMe! Large Scale Prompt Exploration of Open Source LLMs for Machine Translation and Summarization Evaluation
di: Leiter, Christoph, et al.
Pubblicazione: (2024) -
DeepSeek-R1 vs. o3-mini: How Well can Reasoning LLMs Evaluate MT and Summarization?
di: Larionov, Daniil, et al.
Pubblicazione: (2025) -
BMX: Boosting Natural Language Generation Metrics with Explainability
di: Leiter, Christoph, et al.
Pubblicazione: (2022) -
CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks
di: Leiter, Christoph, et al.
Pubblicazione: (2025)