II-Bench: An Image Implication Understanding Benchmark for Multimodal Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Ziqiang, Fang, Feiteng, Feng, Xi, Du, Xinrun, Zhang, Chenhao, Wang, Zekun, Bai, Yuelin, Zhao, Qixuan, Fan, Liyang, Gan, Chengguang, Lin, Hongquan, Li, Jiaming, Ni, Yuansheng, Wu, Haihong, Narsupalli, Yaswanth, Zheng, Zhigang, Li, Chengming, Hu, Xiping, Xu, Ruifeng, Chen, Xiaojun, Yang, Min, Liu, Jiaheng, Liu, Ruibo, Huang, Wenhao, Zhang, Ge, Ni, Shiwen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enhancing Noise Robustness of Retrieval-Augmented Language Models with Adaptive Adversarial Training
di: Fang, Feiteng, et al.
Pubblicazione: (2024)
di: Fang, Feiteng, et al.
Pubblicazione: (2024)
Layer-wise Regularized Dropout for Neural Language Models
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
Forgetting before Learning: Utilizing Parametric Arithmetic for Knowledge Updating in Large Language Models
di: Ni, Shiwen, et al.
Pubblicazione: (2023)
di: Ni, Shiwen, et al.
Pubblicazione: (2023)
DeliLaw: A Chinese Legal Counselling System Based on a Large Language Model
di: Xie, Nan, et al.
Pubblicazione: (2024)
di: Xie, Nan, et al.
Pubblicazione: (2024)
E-EVAL: A Comprehensive Chinese K-12 Education Evaluation Benchmark for Large Language Models
di: Hou, Jinchang, et al.
Pubblicazione: (2024)
di: Hou, Jinchang, et al.
Pubblicazione: (2024)
Training on the Benchmark Is Not All You Need
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
MoZIP: A Multilingual Benchmark to Evaluate Large Language Models in Intellectual Property
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
CLHA: A Simple yet Effective Contrastive Learning Framework for Human Alignment
di: Fang, Feiteng, et al.
Pubblicazione: (2024)
di: Fang, Feiteng, et al.
Pubblicazione: (2024)
Can MLLMs Understand the Deep Implication Behind Chinese Images?
di: Zhang, Chenhao, et al.
Pubblicazione: (2024)
di: Zhang, Chenhao, et al.
Pubblicazione: (2024)
Expanding before Inferring: Enhancing Factuality in Large Language Models through Premature Layers Interpolation
di: Chen, Dingwei, et al.
Pubblicazione: (2025)
di: Chen, Dingwei, et al.
Pubblicazione: (2025)
COIG-CQIA: Quality is All You Need for Chinese Instruction Fine-tuning
di: Bai, Yuelin, et al.
Pubblicazione: (2024)
di: Bai, Yuelin, et al.
Pubblicazione: (2024)
Lower Layers Matter: Alleviating Hallucination via Multi-Layer Fusion Contrastive Decoding with Truthfulness Refocused
di: Chen, Dingwei, et al.
Pubblicazione: (2024)
di: Chen, Dingwei, et al.
Pubblicazione: (2024)
AgentCourt: Simulating Court with Adversarial Evolvable Lawyer Agents
di: Chen, Guhong, et al.
Pubblicazione: (2024)
di: Chen, Guhong, et al.
Pubblicazione: (2024)
ReFeR: Improving Evaluation and Reasoning through Hierarchy of Models
di: Narsupalli, Yaswanth, et al.
Pubblicazione: (2024)
di: Narsupalli, Yaswanth, et al.
Pubblicazione: (2024)
M-MRE: Extending the Mutual Reinforcement Effect to Multimodal Information Extraction
di: Gan, Chengguang, et al.
Pubblicazione: (2025)
di: Gan, Chengguang, et al.
Pubblicazione: (2025)
GuideWeb: A Benchmark for Automatic In-App Guide Generation on Real-World Web UIs
di: Gan, Chengguang, et al.
Pubblicazione: (2026)
di: Gan, Chengguang, et al.
Pubblicazione: (2026)
Quantification of Large Language Model Distillation
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
Educational-Psychological Dialogue Robot Based on Multi-Agent Collaboration
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
Small Language Model as Data Prospector for Large Language Model
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
A Survey on Large Language Model Benchmarks
di: Ni, Shiwen, et al.
Pubblicazione: (2025)
di: Ni, Shiwen, et al.
Pubblicazione: (2025)
Multi-Modality Collaborative Learning for Sentiment Analysis
di: Wang, Shanmin, et al.
Pubblicazione: (2025)
di: Wang, Shanmin, et al.
Pubblicazione: (2025)
xJailbreak: Representation Space Guided Reinforcement Learning for Interpretable LLM Jailbreaking
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
YINYANG-ALIGN: Benchmarking Contradictory Objectives and Proposing Multi-Objective Optimization based DPO for Text-to-Image Alignment
di: Das, Amitava, et al.
Pubblicazione: (2025)
di: Das, Amitava, et al.
Pubblicazione: (2025)
Pre-training, Fine-tuning and Re-ranking: A Three-Stage Framework for Legal Question Answering
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
di: Ni, Shiwen, et al.
Pubblicazione: (2024)
The Fourier Cosine Method for Discrete Probability Distributions
di: Shen, Xiaoyu, et al.
Pubblicazione: (2024)
di: Shen, Xiaoyu, et al.
Pubblicazione: (2024)
A set of nearly good real numbers to specify the ground states associated with a Hamiltonian containing non-commutable terms and the effect of the odd-channel of a pair of different bosons emerging in multi-species systems
di: He, Yanzhang, et al.
Pubblicazione: (2025)
di: He, Yanzhang, et al.
Pubblicazione: (2025)
Assessing CO2 Fluxes for European Peatlands in ORCHIDEE-PEAT With Multiple Plant Functional Types
di: Liu, Liyang
Pubblicazione: (2025)
di: Liu, Liyang
Pubblicazione: (2025)
MVAN: Multi-View Attention Networks for Fake News Detection on Social Media
di: Ni, Shiwen, et al.
Pubblicazione: (2025)
di: Ni, Shiwen, et al.
Pubblicazione: (2025)
IPBench: Benchmarking the Knowledge of Large Language Models in Intellectual Property
di: Wang, Qiyao, et al.
Pubblicazione: (2025)
di: Wang, Qiyao, et al.
Pubblicazione: (2025)
AdversariaL attacK sAfety aLIgnment(ALKALI): Safeguarding LLMs through GRACE: Geometric Representation-Aware Contrastive Enhancement- Introducing Adversarial Vulnerability Quality Index (AVQI)
di: Khanna, Danush, et al.
Pubblicazione: (2025)
di: Khanna, Danush, et al.
Pubblicazione: (2025)
Quantum Gravitational Theory Through Compression (QGTC)
di: Atmakuri, Yaswanth
Pubblicazione: (2025)
di: Atmakuri, Yaswanth
Pubblicazione: (2025)
Prompt Chaining or Stepwise Prompt? Refinement in Text Summarization
di: Sun, Shichao, et al.
Pubblicazione: (2024)
di: Sun, Shichao, et al.
Pubblicazione: (2024)
VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation
di: He, Xuan, et al.
Pubblicazione: (2024)
di: He, Xuan, et al.
Pubblicazione: (2024)
Joint Transmit Waveform and Receive Filter Design for ISAC System with Jamming
di: Shu, Yuan, et al.
Pubblicazione: (2025)
di: Shu, Yuan, et al.
Pubblicazione: (2025)
Like a moth to a flame: Do stock market bubbles exacerbate credit risks of peer‐to‐peer lending?
di: Xin Liu, et al.
Pubblicazione: (2024)
di: Xin Liu, et al.
Pubblicazione: (2024)
Automatic Paper Reviewing with Heterogeneous Graph Reasoning over LLM-Simulated Reviewer-Author Debates
di: Li, Shuaimin, et al.
Pubblicazione: (2025)
di: Li, Shuaimin, et al.
Pubblicazione: (2025)
Earley-Driven Dynamic Pruning for Efficient Structured Decoding
di: Sun, Xintong, et al.
Pubblicazione: (2025)
di: Sun, Xintong, et al.
Pubblicazione: (2025)
Scaling Test-Driven Code Generation from Functions to Classes: An Empirical Study
di: Liang, Yunhao, et al.
Pubblicazione: (2026)
di: Liang, Yunhao, et al.
Pubblicazione: (2026)
KeyInst: Keyword Instruction for Improving SQL Formulation in Text-to-SQL
di: Liu, Xiping, et al.
Pubblicazione: (2024)
di: Liu, Xiping, et al.
Pubblicazione: (2024)
EPI-SQL: Enhancing Text-to-SQL Translation with Error-Prevention Instructions
di: Liu, Xiping, et al.
Pubblicazione: (2024)
di: Liu, Xiping, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Enhancing Noise Robustness of Retrieval-Augmented Language Models with Adaptive Adversarial Training
di: Fang, Feiteng, et al.
Pubblicazione: (2024) -
Layer-wise Regularized Dropout for Neural Language Models
di: Ni, Shiwen, et al.
Pubblicazione: (2024) -
Forgetting before Learning: Utilizing Parametric Arithmetic for Knowledge Updating in Large Language Models
di: Ni, Shiwen, et al.
Pubblicazione: (2023) -
DeliLaw: A Chinese Legal Counselling System Based on a Large Language Model
di: Xie, Nan, et al.
Pubblicazione: (2024) -
E-EVAL: A Comprehensive Chinese K-12 Education Evaluation Benchmark for Large Language Models
di: Hou, Jinchang, et al.
Pubblicazione: (2024)