Guardado en:
| Autores principales: | Triebel, Maximilian, Menner, Marco, Helfenstein, Dominik |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2605.11223 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Online library learning in human visual puzzle solving
por: Zhao, Pinzhe, et al.
Publicado: (2026)
por: Zhao, Pinzhe, et al.
Publicado: (2026)
Language models show human-like content effects on reasoning tasks
por: Dasgupta, Ishita, et al.
Publicado: (2022)
por: Dasgupta, Ishita, et al.
Publicado: (2022)
RECALL: Rehearsal-free Continual Learning for Object Classification
por: Knauer, Markus, et al.
Publicado: (2022)
por: Knauer, Markus, et al.
Publicado: (2022)
Learning Expressive Priors for Generalization and Uncertainty Estimation in Neural Networks
por: Schnaus, Dominik, et al.
Publicado: (2023)
por: Schnaus, Dominik, et al.
Publicado: (2023)
Large Language Models show both individual and collective creativity comparable to humans
por: Sun, Luning, et al.
Publicado: (2024)
por: Sun, Luning, et al.
Publicado: (2024)
Procedurally generating rules to adapt difficulty for narrative puzzle games
por: Volden, Thomas, et al.
Publicado: (2023)
por: Volden, Thomas, et al.
Publicado: (2023)
Towards Explaining Uncertainty Estimates in Point Cloud Registration
por: Qin, Ziyuan, et al.
Publicado: (2024)
por: Qin, Ziyuan, et al.
Publicado: (2024)
Evaluation of LLMs for mathematical problem solving
por: Wang, Ruonan, et al.
Publicado: (2025)
por: Wang, Ruonan, et al.
Publicado: (2025)
Can Large Language Models generalize analogy solving like children can?
por: Stevenson, Claire E., et al.
Publicado: (2024)
por: Stevenson, Claire E., et al.
Publicado: (2024)
LLMs model how humans induce logically structured rules
por: Loo, Alyssa, et al.
Publicado: (2025)
por: Loo, Alyssa, et al.
Publicado: (2025)
Can generative AI and ChatGPT outperform humans on cognitive-demanding problem-solving tasks in science?
por: Zhai, Xiaoming, et al.
Publicado: (2024)
por: Zhai, Xiaoming, et al.
Publicado: (2024)
The promise and limits of LLMs in constructing proofs and hints for logic problems in intelligent tutoring systems
por: Tithi, Sutapa Dey, et al.
Publicado: (2025)
por: Tithi, Sutapa Dey, et al.
Publicado: (2025)
The receptron is a nonlinear threshold logic gate with intrinsic multi-dimensional selective capabilities for analog inputs
por: Paroli, B., et al.
Publicado: (2025)
por: Paroli, B., et al.
Publicado: (2025)
Tracing the ongoing emergence of human-like reasoning in Large Language Models
por: Morosi, Paolo, et al.
Publicado: (2026)
por: Morosi, Paolo, et al.
Publicado: (2026)
AlphaBeta is not as good as you think: a simple class of synthetic games for a better analysis of deterministic game-solving algorithms
por: Boige, Raphaël, et al.
Publicado: (2025)
por: Boige, Raphaël, et al.
Publicado: (2025)
A method for quantifying the generalization capabilities of generative models for solving Ising models
por: Ma, Qunlong, et al.
Publicado: (2024)
por: Ma, Qunlong, et al.
Publicado: (2024)
Among Them: A game-based framework for assessing persuasion capabilities of LLMs
por: Idziejczak, Mateusz, et al.
Publicado: (2025)
por: Idziejczak, Mateusz, et al.
Publicado: (2025)
Assessing SPARQL capabilities of Large Language Models
por: Meyer, Lars-Peter, et al.
Publicado: (2024)
por: Meyer, Lars-Peter, et al.
Publicado: (2024)
Do What? Teaching Vision-Language-Action Models to Reject the Impossible
por: Hsieh, Wen-Han, et al.
Publicado: (2025)
por: Hsieh, Wen-Han, et al.
Publicado: (2025)
Performance Review on LLM for solving leetcode problems
por: Wang, Lun, et al.
Publicado: (2025)
por: Wang, Lun, et al.
Publicado: (2025)
Logical recognition method for solving the problem of identification in the Internet of Things
por: Saymanov, Islambek
Publicado: (2024)
por: Saymanov, Islambek
Publicado: (2024)
Do Theory of Mind Benchmarks Need Explicit Human-like Reasoning in Language Models?
por: Lu, Yi-Long, et al.
Publicado: (2025)
por: Lu, Yi-Long, et al.
Publicado: (2025)
Large language models show fragile cognitive reasoning about human emotions
por: Bhattacharyya, Sree, et al.
Publicado: (2025)
por: Bhattacharyya, Sree, et al.
Publicado: (2025)
Do Vision-Language Models Respect Contextual Integrity in Location Disclosure?
por: Yang, Ruixin, et al.
Publicado: (2026)
por: Yang, Ruixin, et al.
Publicado: (2026)
Evaluating List Construction and Temporal Understanding capabilities of Large Language Models
por: Dumitru, Alexandru, et al.
Publicado: (2025)
por: Dumitru, Alexandru, et al.
Publicado: (2025)
Is AI currently capable of identifying wild oysters? A comparison of human annotators against the AI model, ODYSSEE
por: Campbell, Brendan, et al.
Publicado: (2025)
por: Campbell, Brendan, et al.
Publicado: (2025)
CrochetBench: Can Vision-Language Models Move from Describing to Doing in Crochet Domain?
por: Li, Peiyu, et al.
Publicado: (2025)
por: Li, Peiyu, et al.
Publicado: (2025)
YesBut: A High-Quality Annotated Multimodal Dataset for evaluating Satire Comprehension capability of Vision-Language Models
por: Nandy, Abhilash, et al.
Publicado: (2024)
por: Nandy, Abhilash, et al.
Publicado: (2024)
FFHFlow: Diverse and Uncertainty-Aware Dexterous Grasp Generation via Flow Variational Inference
por: Feng, Qian, et al.
Publicado: (2024)
por: Feng, Qian, et al.
Publicado: (2024)
Leveraging Vision-Language Models for Visual Grounding and Analysis of Automotive UI
por: Ernhofer, Benjamin Raphael, et al.
Publicado: (2025)
por: Ernhofer, Benjamin Raphael, et al.
Publicado: (2025)
Surgeons vs. Computer Vision: A comparative analysis on surgical phase recognition capabilities
por: Mezzina, Marco, et al.
Publicado: (2025)
por: Mezzina, Marco, et al.
Publicado: (2025)
Trinity: Unifying Class-Agnostic Terrain and Semantic Segmentation for Unstructured Outdoor Environments by Leveraging Synthetic Data
por: Müller, Marcus G, et al.
Publicado: (2026)
por: Müller, Marcus G, et al.
Publicado: (2026)
Do Vision-Language Models See Urban Scenes as People Do? An Urban Perception Benchmark
por: Mushkani, Rashid
Publicado: (2025)
por: Mushkani, Rashid
Publicado: (2025)
VideoGameBench: Can Vision-Language Models complete popular video games?
por: Zhang, Alex L., et al.
Publicado: (2025)
por: Zhang, Alex L., et al.
Publicado: (2025)
Which symbol grounding problem should we try to solve?
por: Müller, Vincent C.
Publicado: (2025)
por: Müller, Vincent C.
Publicado: (2025)
Do All Individual Layers Help? An Empirical Study of Task-Interfering Layers in Vision-Language Models
por: Liu, Zhiming, et al.
Publicado: (2026)
por: Liu, Zhiming, et al.
Publicado: (2026)
Why Do Vision Language Models Struggle To Recognize Human Emotions?
por: Agarwal, Madhav, et al.
Publicado: (2026)
por: Agarwal, Madhav, et al.
Publicado: (2026)
Do Pre-trained Vision-Language Models Encode Object States?
por: Newman, Kaleb, et al.
Publicado: (2024)
por: Newman, Kaleb, et al.
Publicado: (2024)
Evidence of interrelated cognitive-like capabilities in large language models: Indications of artificial general intelligence or achievement?
por: Ilić, David, et al.
Publicado: (2023)
por: Ilić, David, et al.
Publicado: (2023)
RomanSetu: Efficiently unlocking multilingual capabilities of Large Language Models via Romanization
por: Husain, Jaavid Aktar, et al.
Publicado: (2024)
por: Husain, Jaavid Aktar, et al.
Publicado: (2024)
Ejemplares similares
-
Online library learning in human visual puzzle solving
por: Zhao, Pinzhe, et al.
Publicado: (2026) -
Language models show human-like content effects on reasoning tasks
por: Dasgupta, Ishita, et al.
Publicado: (2022) -
RECALL: Rehearsal-free Continual Learning for Object Classification
por: Knauer, Markus, et al.
Publicado: (2022) -
Learning Expressive Priors for Generalization and Uncertainty Estimation in Neural Networks
por: Schnaus, Dominik, et al.
Publicado: (2023) -
Large Language Models show both individual and collective creativity comparable to humans
por: Sun, Luning, et al.
Publicado: (2024)