Exploring Spatial Schema Intuitions in Large Language and Vision Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wicke, Philipp, Wachowiak, Lennart |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Probing Language Models' Gesture Understanding for Enhanced Human-AI Interaction
par: Wicke, Philipp
Publié: (2024)
par: Wicke, Philipp
Publié: (2024)
Are Large Language Models Aligned with People's Social Intuitions for Human-Robot Interactions?
par: Wachowiak, Lennart, et autres
Publié: (2024)
par: Wachowiak, Lennart, et autres
Publié: (2024)
Time Course MechInterp: Analyzing the Evolution of Components and Knowledge in Large Language Models
par: Hakimi, Ahmad Dawar, et autres
Publié: (2025)
par: Hakimi, Ahmad Dawar, et autres
Publié: (2025)
Red and blue language: Word choices in the Trump & Harris 2024 presidential debate
par: Wicke, Philipp, et autres
Publié: (2024)
par: Wicke, Philipp, et autres
Publié: (2024)
SLAyiNG: Towards Queer Language Processing
par: Veloso, Leonor, et autres
Publié: (2025)
par: Veloso, Leonor, et autres
Publié: (2025)
Data Metabolism: An Efficient Data Design Schema For Vision Language Model
par: Zhang, Jingyuan, et autres
Publié: (2025)
par: Zhang, Jingyuan, et autres
Publié: (2025)
What Questions Should Robots Be Able to Answer? A Dataset of User Questions for Explainable Robotics
par: Wachowiak, Lennart, et autres
Publié: (2025)
par: Wachowiak, Lennart, et autres
Publié: (2025)
LLMs4SchemaDiscovery: A Human-in-the-Loop Workflow for Scientific Schema Mining with Large Language Models
par: Sadruddin, Sameer, et autres
Publié: (2025)
par: Sadruddin, Sameer, et autres
Publié: (2025)
An Expert Schema for Evaluating Large Language Model Errors in Scholarly Question-Answering Systems
par: Martin-Boyle, Anna, et autres
Publié: (2026)
par: Martin-Boyle, Anna, et autres
Publié: (2026)
Retrieval-Augmented Large Language Models for Schema-Constrained Clinical Information Extraction
par: Karim, A H M Rezaul, et autres
Publié: (2026)
par: Karim, A H M Rezaul, et autres
Publié: (2026)
Beyond the Vision Encoder: Identifying and Mitigating Spatial Bias in Large Vision-Language Models
par: Zhu, Yingjie, et autres
Publié: (2025)
par: Zhu, Yingjie, et autres
Publié: (2025)
Learning a Structural Causal Model for Intuition Reasoning in Conversation
par: Chen, Hang, et autres
Publié: (2023)
par: Chen, Hang, et autres
Publié: (2023)
Don't Adapt Small Language Models for Tools; Adapt Tool Schemas to the Models
par: Lee, Jonggeun, et autres
Publié: (2025)
par: Lee, Jonggeun, et autres
Publié: (2025)
Effective and Efficient Schema-aware Information Extraction Using On-Device Large Language Models
par: Wen, Zhihao, et autres
Publié: (2025)
par: Wen, Zhihao, et autres
Publié: (2025)
Broadening Access to Transportation Safety Data with Generative AI: A Schema-Grounded Framework for Spatial Natural Language Queries
par: Azhdari, Mahdi, et autres
Publié: (2026)
par: Azhdari, Mahdi, et autres
Publié: (2026)
The Death of Schema Linking? Text-to-SQL in the Age of Well-Reasoned Language Models
par: Maamari, Karime, et autres
Publié: (2024)
par: Maamari, Karime, et autres
Publié: (2024)
EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models
par: Du, Mengfei, et autres
Publié: (2024)
par: Du, Mengfei, et autres
Publié: (2024)
Reasoning Paths with Reference Objects Elicit Quantitative Spatial Reasoning in Large Vision-Language Models
par: Liao, Yuan-Hong, et autres
Publié: (2024)
par: Liao, Yuan-Hong, et autres
Publié: (2024)
Concept-Reversed Winograd Schema Challenge: Evaluating and Improving Robust Reasoning in Large Language Models via Abstraction
par: Han, Kaiqiao, et autres
Publié: (2024)
par: Han, Kaiqiao, et autres
Publié: (2024)
Beyond Shallow Heuristics: Leveraging Human Intuition for Curriculum Learning
par: Toborek, Vanessa, et autres
Publié: (2025)
par: Toborek, Vanessa, et autres
Publié: (2025)
A Layered Intuition -- Method Model with Scope Extension for LLM Reasoning
par: Su, Hong
Publié: (2025)
par: Su, Hong
Publié: (2025)
ChatSchema: A pipeline of extracting structured information with Large Multimodal Models based on schema
par: Wang, Fei, et autres
Publié: (2024)
par: Wang, Fei, et autres
Publié: (2024)
Investigating Spatial Attention Bias in Vision-Language Models
par: Chaudhary, Aryan, et autres
Publié: (2025)
par: Chaudhary, Aryan, et autres
Publié: (2025)
Exploring the Translation Mechanism of Large Language Models
par: Zhang, Hongbin, et autres
Publié: (2025)
par: Zhang, Hongbin, et autres
Publié: (2025)
Safe + Safe = Unsafe? Exploring How Safe Images Can Be Exploited to Jailbreak Large Vision-Language Models
par: Cui, Chenhang, et autres
Publié: (2024)
par: Cui, Chenhang, et autres
Publié: (2024)
Distortions in Judged Spatial Relations in Large Language Models
par: Fulman, Nir, et autres
Publié: (2024)
par: Fulman, Nir, et autres
Publié: (2024)
Misalignment of Semantic Relation Knowledge between WordNet and Human Intuition
par: Cao, Zhihan, et autres
Publié: (2024)
par: Cao, Zhihan, et autres
Publié: (2024)
SpatialLadder: Progressive Training for Spatial Reasoning in Vision-Language Models
par: Li, Hongxing, et autres
Publié: (2025)
par: Li, Hongxing, et autres
Publié: (2025)
BiasDora: Exploring Hidden Biased Associations in Vision-Language Models
par: Raj, Chahat, et autres
Publié: (2024)
par: Raj, Chahat, et autres
Publié: (2024)
Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models
par: Stogiannidis, Ilias, et autres
Publié: (2025)
par: Stogiannidis, Ilias, et autres
Publié: (2025)
Prompt4Vis: Prompting Large Language Models with Example Mining and Schema Filtering for Tabular Data Visualization
par: Li, Shuaimin, et autres
Publié: (2024)
par: Li, Shuaimin, et autres
Publié: (2024)
Exploring Forgetting in Large Language Model Pre-Training
par: Liao, Chonghua, et autres
Publié: (2024)
par: Liao, Chonghua, et autres
Publié: (2024)
Exploring Group and Symmetry Principles in Large Language Models
par: Imani, Shima, et autres
Publié: (2024)
par: Imani, Shima, et autres
Publié: (2024)
Exploring and Mitigating Fawning Hallucinations in Large Language Models
par: Shangguan, Zixuan, et autres
Publié: (2025)
par: Shangguan, Zixuan, et autres
Publié: (2025)
Exploring the Potential of Large Language Models in Computational Argumentation
par: Chen, Guizhen, et autres
Publié: (2023)
par: Chen, Guizhen, et autres
Publié: (2023)
Transfer Learning for Finetuning Large Language Models
par: Strangmann, Tobias, et autres
Publié: (2024)
par: Strangmann, Tobias, et autres
Publié: (2024)
SchemaGraphSQL: Efficient Schema Linking with Pathfinding Graph Algorithms for Text-to-SQL on Large-Scale Databases
par: Safdarian, AmirHossein, et autres
Publié: (2025)
par: Safdarian, AmirHossein, et autres
Publié: (2025)
Can Vision Replace Text in Working Memory? Evidence from Spatial n-Back in Vision-Language Models
par: Liang, Sichu, et autres
Publié: (2026)
par: Liang, Sichu, et autres
Publié: (2026)
One Script Instead of Hundreds? On Pretraining Romanized Encoder Language Models
par: Ebing, Benedikt, et autres
Publié: (2026)
par: Ebing, Benedikt, et autres
Publié: (2026)
Evaluating Spatial Understanding of Large Language Models
par: Yamada, Yutaro, et autres
Publié: (2023)
par: Yamada, Yutaro, et autres
Publié: (2023)
Documents similaires
-
Probing Language Models' Gesture Understanding for Enhanced Human-AI Interaction
par: Wicke, Philipp
Publié: (2024) -
Are Large Language Models Aligned with People's Social Intuitions for Human-Robot Interactions?
par: Wachowiak, Lennart, et autres
Publié: (2024) -
Time Course MechInterp: Analyzing the Evolution of Components and Knowledge in Large Language Models
par: Hakimi, Ahmad Dawar, et autres
Publié: (2025) -
Red and blue language: Word choices in the Trump & Harris 2024 presidential debate
par: Wicke, Philipp, et autres
Publié: (2024) -
SLAyiNG: Towards Queer Language Processing
par: Veloso, Leonor, et autres
Publié: (2025)