What is the Visual Cognition Gap between Humans and Multimodal LLMs?
Fuente:
arXiv
Salvato in:
| Autori principali: | Cao, Xu, Shen, Yifan, Lai, Bolin, Ye, Wenqian, Ma, Yunsheng, Heintz, Joerg, Chen, Jintai, Huang, Meihuan, Cao, Jianguo, Zhang, Aidong, Rehg, James M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MM-SpuBench: Towards Better Understanding of Spurious Biases in Multimodal LLMs
di: Ye, Wenqian, et al.
Pubblicazione: (2024)
di: Ye, Wenqian, et al.
Pubblicazione: (2024)
The Selective G-Bispectrum and its Inversion: Applications to G-Invariant Networks
di: Mataigne, Simon, et al.
Pubblicazione: (2024)
di: Mataigne, Simon, et al.
Pubblicazione: (2024)
RHealthTwin: Towards Responsible and Multimodal Digital Twins for Personalized Well-being
di: Ferdousi, Rahatara, et al.
Pubblicazione: (2025)
di: Ferdousi, Rahatara, et al.
Pubblicazione: (2025)
Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading
di: Sadjoli, Nicholas, et al.
Pubblicazione: (2026)
di: Sadjoli, Nicholas, et al.
Pubblicazione: (2026)
Towards Human Cognition Level-based Experiment Design for Counterfactual Explanations (XAI)
di: Suffian, Muhammad, et al.
Pubblicazione: (2022)
di: Suffian, Muhammad, et al.
Pubblicazione: (2022)
Shortest Paths in a Weighted Simplicial Complex
di: Chakraborty, Sukrit, et al.
Pubblicazione: (2025)
di: Chakraborty, Sukrit, et al.
Pubblicazione: (2025)
Vibe-Creation: The Epistemology of Human-AI Emergent Cognition
di: Levin, Ilya
Pubblicazione: (2026)
di: Levin, Ilya
Pubblicazione: (2026)
What is the $\textit{intrinsic}$ dimension of your binary data? -- and how to compute it quickly
di: Hanika, Tom, et al.
Pubblicazione: (2024)
di: Hanika, Tom, et al.
Pubblicazione: (2024)
A Human-In-The-Loop Approach for Improving Fairness in Predictive Business Process Monitoring
di: Käppel, Martin, et al.
Pubblicazione: (2025)
di: Käppel, Martin, et al.
Pubblicazione: (2025)
Modeling Membrane Degradation in PEM Electrolyzers with Physics-Informed Neural Networks
di: Polo-Molina, Alejandro, et al.
Pubblicazione: (2025)
di: Polo-Molina, Alejandro, et al.
Pubblicazione: (2025)
ConSensus: Multi-Agent Collaboration for Multimodal Sensing
di: Yoon, Hyungjun, et al.
Pubblicazione: (2026)
di: Yoon, Hyungjun, et al.
Pubblicazione: (2026)
Animation Needs Attention: A Holistic Approach to Slides Animation Comprehension with Visual-Language Models
di: Jiang, Yifan, et al.
Pubblicazione: (2025)
di: Jiang, Yifan, et al.
Pubblicazione: (2025)
IDA: Breaking Barriers in No-code UI Automation Through Large Language Models and Human-Centric Design
di: Shlomov, Segev, et al.
Pubblicazione: (2024)
di: Shlomov, Segev, et al.
Pubblicazione: (2024)
Structured Extraction of Vulnerabilities in OpenVAS and Tenable WAS Reports Using LLMs
di: Machado, Beatriz, et al.
Pubblicazione: (2025)
di: Machado, Beatriz, et al.
Pubblicazione: (2025)
AI-based modular warning machine for risk identification in proximity healthcare
di: Razzetta, Chiara, et al.
Pubblicazione: (2025)
di: Razzetta, Chiara, et al.
Pubblicazione: (2025)
Strategic inputs: feature selection from game-theoretic perspective
di: Zhao, Chi, et al.
Pubblicazione: (2025)
di: Zhao, Chi, et al.
Pubblicazione: (2025)
A Space-Efficient Algorithm for Longest Common Almost Increasing Subsequence of Two Sequences
di: Rahat, Md Tanzeem, et al.
Pubblicazione: (2025)
di: Rahat, Md Tanzeem, et al.
Pubblicazione: (2025)
The Trap of Presumed Equivalence: Artificial General Intelligence Should Not Be Assessed on the Scale of Human Intelligence
di: Dolgikh, Serge
Pubblicazione: (2024)
di: Dolgikh, Serge
Pubblicazione: (2024)
AI Governance InternationaL Evaluation Index (AGILE Index) 2025
di: Zeng, Yi, et al.
Pubblicazione: (2025)
di: Zeng, Yi, et al.
Pubblicazione: (2025)
Evaluating AI Grading on Real-World Handwritten College Mathematics: A Large-Scale Study Toward a Benchmark
di: Yu, Zhiqi, et al.
Pubblicazione: (2026)
di: Yu, Zhiqi, et al.
Pubblicazione: (2026)
A Landmark-Aware Visual Navigation Dataset
di: Johnson, Faith, et al.
Pubblicazione: (2024)
di: Johnson, Faith, et al.
Pubblicazione: (2024)
Transit Functions and Clustering Systems
di: Changat, Manoj, et al.
Pubblicazione: (2024)
di: Changat, Manoj, et al.
Pubblicazione: (2024)
Differential Parity: Relative Fairness Between Two Sets of Decisions
di: Yu, Zhe, et al.
Pubblicazione: (2021)
di: Yu, Zhe, et al.
Pubblicazione: (2021)
Freeze, Diffuse, Decode: Geometry-Aware Adaptation of Pretrained Transformer Embeddings for Antimicrobial Peptide Design
di: Gawade, Pankhil, et al.
Pubblicazione: (2025)
di: Gawade, Pankhil, et al.
Pubblicazione: (2025)
Creativity in the Age of AI: Rethinking the Role of Intentional Agency
di: Pearson, James S., et al.
Pubblicazione: (2026)
di: Pearson, James S., et al.
Pubblicazione: (2026)
Mathematical reasoning and the computer
di: Buzzard, Kevin
Pubblicazione: (2025)
di: Buzzard, Kevin
Pubblicazione: (2025)
ReMIA: a Powerful and Efficient Alternative to Membership Inference Attacks against Synthetic Data Generators
di: Scassola, Davide, et al.
Pubblicazione: (2026)
di: Scassola, Davide, et al.
Pubblicazione: (2026)
Intrinsic Rewards for Exploration without Harm from Observational Noise: A Simulation Study Based on the Free Energy Principle
di: Tinker, Theodore Jerome, et al.
Pubblicazione: (2024)
di: Tinker, Theodore Jerome, et al.
Pubblicazione: (2024)
Modeling Clinical Concern Trajectories in Language Model Agents
di: Subaharan, Sukesh, et al.
Pubblicazione: (2026)
di: Subaharan, Sukesh, et al.
Pubblicazione: (2026)
How well can a large language model explain business processes as perceived by users?
di: Fahland, Dirk, et al.
Pubblicazione: (2024)
di: Fahland, Dirk, et al.
Pubblicazione: (2024)
humancompatible.detect: a Python Toolkit for Detecting Bias in AI Models
di: Matilla, German M., et al.
Pubblicazione: (2025)
di: Matilla, German M., et al.
Pubblicazione: (2025)
Physics-Informed Neural Networks and Neural Operators for Parametric PDEs
di: Zhang, Zhuo, et al.
Pubblicazione: (2025)
di: Zhang, Zhuo, et al.
Pubblicazione: (2025)
Toward a Dynamic Stackelberg Game-Theoretic Framework for Agentic AI Defense Against LLM Jailbreaking
di: Han, Zhengye, et al.
Pubblicazione: (2025)
di: Han, Zhengye, et al.
Pubblicazione: (2025)
Application of machine learning for infrastructure reconstruction programs management
di: Khudiakov, Illia, et al.
Pubblicazione: (2025)
di: Khudiakov, Illia, et al.
Pubblicazione: (2025)
On the Optimal Memorization Capacity of Transformers
di: Kajitsuka, Tokio, et al.
Pubblicazione: (2024)
di: Kajitsuka, Tokio, et al.
Pubblicazione: (2024)
BernGraph: Probabilistic Graph Neural Networks for EHR-based Medication Recommendations
di: Piao, Xihao, et al.
Pubblicazione: (2024)
di: Piao, Xihao, et al.
Pubblicazione: (2024)
Semantic Mobile Base Station Placement
di: Soman, Kritik, et al.
Pubblicazione: (2021)
di: Soman, Kritik, et al.
Pubblicazione: (2021)
Representation Integrity in Temporal Graph Learning Methods
di: Kooshafar, Elahe
Pubblicazione: (2025)
di: Kooshafar, Elahe
Pubblicazione: (2025)
A Data-Driven Measure of Relative Uncertainty for Misclassification Detection
di: Dadalto, Eduardo, et al.
Pubblicazione: (2023)
di: Dadalto, Eduardo, et al.
Pubblicazione: (2023)
Benchmarking PNW Model for MedMNIST to 100% Accuracy
di: Deng, Bo
Pubblicazione: (2026)
di: Deng, Bo
Pubblicazione: (2026)
Documenti analoghi
-
MM-SpuBench: Towards Better Understanding of Spurious Biases in Multimodal LLMs
di: Ye, Wenqian, et al.
Pubblicazione: (2024) -
The Selective G-Bispectrum and its Inversion: Applications to G-Invariant Networks
di: Mataigne, Simon, et al.
Pubblicazione: (2024) -
RHealthTwin: Towards Responsible and Multimodal Digital Twins for Personalized Well-being
di: Ferdousi, Rahatara, et al.
Pubblicazione: (2025) -
Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading
di: Sadjoli, Nicholas, et al.
Pubblicazione: (2026) -
Towards Human Cognition Level-based Experiment Design for Counterfactual Explanations (XAI)
di: Suffian, Muhammad, et al.
Pubblicazione: (2022)