MM-Soc: Benchmarking Multimodal Large Language Models in Social Media Platforms
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Jin, Yiqiao, Choi, Minje, Verma, Gaurav, Wang, Jindong, Kumar, Srijan |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space
par: Verma, Gaurav, et autres
Publié: (2024)
par: Verma, Gaurav, et autres
Publié: (2024)
MM-SAP: A Comprehensive Benchmark for Assessing Self-Awareness of Multimodal Large Language Models in Perception
par: Wang, Yuhao, et autres
Publié: (2024)
par: Wang, Yuhao, et autres
Publié: (2024)
The Use of Multimodal Large Language Models to Detect Objects from Thermal Images: Transportation Applications
par: Ashqar, Huthaifa I., et autres
Publié: (2024)
par: Ashqar, Huthaifa I., et autres
Publié: (2024)
MM-SafetyBench: A Benchmark for Safety Evaluation of Multimodal Large Language Models
par: Liu, Xin, et autres
Publié: (2023)
par: Liu, Xin, et autres
Publié: (2023)
A Survey on Responsible Generative AI: What to Generate and What Not
par: Gu, Jindong
Publié: (2024)
par: Gu, Jindong
Publié: (2024)
Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering
par: Si, Chenglei, et autres
Publié: (2024)
par: Si, Chenglei, et autres
Publié: (2024)
Large Language Models and Provenance Metadata for Determining the Relevance of Images and Videos in News Stories
par: Peterka, Tomas, et autres
Publié: (2025)
par: Peterka, Tomas, et autres
Publié: (2025)
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
par: Qi, Peng, et autres
Publié: (2024)
par: Qi, Peng, et autres
Publié: (2024)
MM-Vet v2: A Challenging Benchmark to Evaluate Large Multimodal Models for Integrated Capabilities
par: Yu, Weihao, et autres
Publié: (2024)
par: Yu, Weihao, et autres
Publié: (2024)
Examining Gender and Racial Bias in Large Vision-Language Models Using a Novel Dataset of Parallel Images
par: Fraser, Kathleen C., et autres
Publié: (2024)
par: Fraser, Kathleen C., et autres
Publié: (2024)
GAOKAO-MM: A Chinese Human-Level Benchmark for Multimodal Models Evaluation
par: Zong, Yi, et autres
Publié: (2024)
par: Zong, Yi, et autres
Publié: (2024)
VideoNorms: Benchmarking Cultural Awareness of Video Language Models
par: Varimalla, Nikhil Reddy, et autres
Publié: (2025)
par: Varimalla, Nikhil Reddy, et autres
Publié: (2025)
ERIT Lightweight Multimodal Dataset for Elderly Emotion Recognition and Multimodal Fusion Evaluation
par: Frieske, Rita, et autres
Publié: (2024)
par: Frieske, Rita, et autres
Publié: (2024)
A Survey on Benchmarks of Multimodal Large Language Models
par: Li, Jian, et autres
Publié: (2024)
par: Li, Jian, et autres
Publié: (2024)
CARES: A Comprehensive Benchmark of Trustworthiness in Medical Vision Language Models
par: Xia, Peng, et autres
Publié: (2024)
par: Xia, Peng, et autres
Publié: (2024)
HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models
par: Kang, Zhaolu, et autres
Publié: (2025)
par: Kang, Zhaolu, et autres
Publié: (2025)
A Thousand Words or An Image: Studying the Influence of Persona Modality in Multimodal LLMs
par: Broomfield, Julius, et autres
Publié: (2025)
par: Broomfield, Julius, et autres
Publié: (2025)
Two Stage Context Learning with Large Language Models for Multimodal Stance Detection on Climate Change
par: Pangtey, Lata, et autres
Publié: (2025)
par: Pangtey, Lata, et autres
Publié: (2025)
Identifying Implicit Social Biases in Vision-Language Models
par: Hamidieh, Kimia, et autres
Publié: (2024)
par: Hamidieh, Kimia, et autres
Publié: (2024)
SAP-Bench: Benchmarking Multimodal Large Language Models in Surgical Action Planning
par: Xu, Mengya, et autres
Publié: (2025)
par: Xu, Mengya, et autres
Publié: (2025)
SPORTU: A Comprehensive Sports Understanding Benchmark for Multimodal Large Language Models
par: Xia, Haotian, et autres
Publié: (2024)
par: Xia, Haotian, et autres
Publié: (2024)
MFC-Bench: Benchmarking Multimodal Fact-Checking with Large Vision-Language Models
par: Wang, Shengkang, et autres
Publié: (2024)
par: Wang, Shengkang, et autres
Publié: (2024)
CODIS: Benchmarking Context-Dependent Visual Comprehension for Multimodal Large Language Models
par: Luo, Fuwen, et autres
Publié: (2024)
par: Luo, Fuwen, et autres
Publié: (2024)
MM-RLHF: The Next Step Forward in Multimodal LLM Alignment
par: Zhang, Yi-Fan, et autres
Publié: (2025)
par: Zhang, Yi-Fan, et autres
Publié: (2025)
xGen-MM (BLIP-3): A Family of Open Large Multimodal Models
par: Xue, Le, et autres
Publié: (2024)
par: Xue, Le, et autres
Publié: (2024)
Stable Signer: Hierarchical Sign Language Generative Model
par: Fang, Sen, et autres
Publié: (2025)
par: Fang, Sen, et autres
Publié: (2025)
MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
par: Li, Shilong, et autres
Publié: (2025)
par: Li, Shilong, et autres
Publié: (2025)
BLEnD-Vis: Benchmarking Multimodal Cultural Understanding in Vision Language Models
par: Tan, Bryan Chen Zhengyu, et autres
Publié: (2025)
par: Tan, Bryan Chen Zhengyu, et autres
Publié: (2025)
Are Vision Language Models Cross-Cultural Theory of Mind Reasoners?
par: Nazi, Zabir Al, et autres
Publié: (2025)
par: Nazi, Zabir Al, et autres
Publié: (2025)
MME-SCI: A Comprehensive and Challenging Science Benchmark for Multimodal Large Language Models
par: Ruan, Jiacheng, et autres
Publié: (2025)
par: Ruan, Jiacheng, et autres
Publié: (2025)
Res-Bench: Benchmarking the Robustness of Multimodal Large Language Models to Dynamic Resolution Input
par: Li, Chenxu, et autres
Publié: (2025)
par: Li, Chenxu, et autres
Publié: (2025)
MathScape: Benchmarking Multimodal Large Language Models in Real-World Mathematical Contexts
par: Liang, Hao, et autres
Publié: (2024)
par: Liang, Hao, et autres
Publié: (2024)
AesBench: An Expert Benchmark for Multimodal Large Language Models on Image Aesthetics Perception
par: Huang, Yipo, et autres
Publié: (2024)
par: Huang, Yipo, et autres
Publié: (2024)
Do LLMs Understand Social Knowledge? Evaluating the Sociability of Large Language Models with SocKET Benchmark
par: Choi, Minje, et autres
Publié: (2023)
par: Choi, Minje, et autres
Publié: (2023)
CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries
par: Liu, Shudong, et autres
Publié: (2025)
par: Liu, Shudong, et autres
Publié: (2025)
Benchmarking the Thinking Mode of Multimodal Large Language Models in Clinical Tasks
par: Hong, Jindong, et autres
Publié: (2025)
par: Hong, Jindong, et autres
Publié: (2025)
A Unified Framework and Dataset for Assessing Societal Bias in Vision-Language Models
par: Sathe, Ashutosh, et autres
Publié: (2024)
par: Sathe, Ashutosh, et autres
Publié: (2024)
Restoring Ancient Ideograph: A Multimodal Multitask Neural Network Approach
par: Duan, Siyu, et autres
Publié: (2024)
par: Duan, Siyu, et autres
Publié: (2024)
A Chinese Multi-label Affective Computing Dataset Based on Social Media Network Users
par: Zhou, Jingyi, et autres
Publié: (2024)
par: Zhou, Jingyi, et autres
Publié: (2024)
Learning Multimodal Cues of Children's Uncertainty
par: Cheng, Qi, et autres
Publié: (2024)
par: Cheng, Qi, et autres
Publié: (2024)
Documents similaires
-
Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space
par: Verma, Gaurav, et autres
Publié: (2024) -
MM-SAP: A Comprehensive Benchmark for Assessing Self-Awareness of Multimodal Large Language Models in Perception
par: Wang, Yuhao, et autres
Publié: (2024) -
The Use of Multimodal Large Language Models to Detect Objects from Thermal Images: Transportation Applications
par: Ashqar, Huthaifa I., et autres
Publié: (2024) -
MM-SafetyBench: A Benchmark for Safety Evaluation of Multimodal Large Language Models
par: Liu, Xin, et autres
Publié: (2023) -
A Survey on Responsible Generative AI: What to Generate and What Not
par: Gu, Jindong
Publié: (2024)