Smart Vision-Language Reasoners
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Roberts, Denisa, Roberts, Lucas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Which Programming Language and Model Work Best With LLM-as-a-Judge For Code Retrieval?
von: Roberts, Lucas, et al.
Veröffentlicht: (2025)
von: Roberts, Lucas, et al.
Veröffentlicht: (2025)
Gender-Neutral Large Language Models for Medical Applications: Reducing Bias in PubMed Abstracts
von: Schaefer, Elizabeth, et al.
Veröffentlicht: (2025)
von: Schaefer, Elizabeth, et al.
Veröffentlicht: (2025)
Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models
von: Moore, Kyle, et al.
Veröffentlicht: (2025)
von: Moore, Kyle, et al.
Veröffentlicht: (2025)
Image2Struct: Benchmarking Structure Extraction for Vision-Language Models
von: Roberts, Josselin Somerville, et al.
Veröffentlicht: (2024)
von: Roberts, Josselin Somerville, et al.
Veröffentlicht: (2024)
Exploitation Is All You Need... for Exploration
von: Rentschler, Micah, et al.
Veröffentlicht: (2025)
von: Rentschler, Micah, et al.
Veröffentlicht: (2025)
RL + Transformer = A General-Purpose Problem Solver
von: Rentschler, Micah, et al.
Veröffentlicht: (2025)
von: Rentschler, Micah, et al.
Veröffentlicht: (2025)
From Values to Frameworks: A Qualitative Study of Ethical Reasoning in Agentic AI Practitioners
von: Roberts, Theodore, et al.
Veröffentlicht: (2025)
von: Roberts, Theodore, et al.
Veröffentlicht: (2025)
Integrating Natural Language Processing Techniques of Text Mining Into Financial System: Applications and Limitations
von: Millo, Denisa, et al.
Veröffentlicht: (2024)
von: Millo, Denisa, et al.
Veröffentlicht: (2024)
Do Large Language Models Learn Human-Like Strategic Preferences?
von: Roberts, Jesse, et al.
Veröffentlicht: (2024)
von: Roberts, Jesse, et al.
Veröffentlicht: (2024)
Helios: A Foundational Language Model for Smart Energy Knowledge Reasoning and Application
von: Jiang, Haoyu, et al.
Veröffentlicht: (2025)
von: Jiang, Haoyu, et al.
Veröffentlicht: (2025)
Procedural Knowledge Improves Agentic LLM Workflows
von: Hsiao, Vincent, et al.
Veröffentlicht: (2025)
von: Hsiao, Vincent, et al.
Veröffentlicht: (2025)
Graph-to-Vision: Multi-graph Understanding and Reasoning using Vision-Language Models
von: Ai, Qihang, et al.
Veröffentlicht: (2025)
von: Ai, Qihang, et al.
Veröffentlicht: (2025)
Human-Centric Goal Reasoning with Ripple-Down Rules
von: Brameld, Kenji, et al.
Veröffentlicht: (2024)
von: Brameld, Kenji, et al.
Veröffentlicht: (2024)
Planning with Reasoning using Vision Language World Model
von: Chen, Delong, et al.
Veröffentlicht: (2025)
von: Chen, Delong, et al.
Veröffentlicht: (2025)
Compute Optimal Scaling of Skills: Knowledge vs Reasoning
von: Roberts, Nicholas, et al.
Veröffentlicht: (2025)
von: Roberts, Nicholas, et al.
Veröffentlicht: (2025)
LLMs as Agentic Cooperative Players in Multiplayer UNO
von: Matinez, Yago Romano, et al.
Veröffentlicht: (2025)
von: Matinez, Yago Romano, et al.
Veröffentlicht: (2025)
Can Large Language Models Make the Grade? An Empirical Study Evaluating LLMs Ability to Mark Short Answer Questions in K-12 Education
von: Henkel, Owen, et al.
Veröffentlicht: (2024)
von: Henkel, Owen, et al.
Veröffentlicht: (2024)
Landmark-Assisted Monte Carlo Planning
von: Chan, David H., et al.
Veröffentlicht: (2025)
von: Chan, David H., et al.
Veröffentlicht: (2025)
Vision Language Models Cannot Reason About Physical Transformation
von: Luo, Dezhi, et al.
Veröffentlicht: (2026)
von: Luo, Dezhi, et al.
Veröffentlicht: (2026)
Unveiling the Compositional Ability Gap in Vision-Language Reasoning Model
von: Li, Tianle, et al.
Veröffentlicht: (2025)
von: Li, Tianle, et al.
Veröffentlicht: (2025)
Visual Distraction Undermines Moral Reasoning in Vision-Language Models
von: Yang, Xinyi, et al.
Veröffentlicht: (2026)
von: Yang, Xinyi, et al.
Veröffentlicht: (2026)
Large Language Model Recall Uncertainty is Modulated by the Fan Effect
von: Roberts, Jesse, et al.
Veröffentlicht: (2024)
von: Roberts, Jesse, et al.
Veröffentlicht: (2024)
Continuous Reasoning for Vision-Language-Action
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
von: Wu, Yueh-Hua, et al.
Veröffentlicht: (2026)
Projected Task-Specific Layers for Multi-Task Reinforcement Learning
von: Roberts, Josselin Somerville, et al.
Veröffentlicht: (2023)
von: Roberts, Josselin Somerville, et al.
Veröffentlicht: (2023)
Generalized invariants meet constitutive neural networks: A novel framework for hyperelastic materials
von: Martonová, Denisa, et al.
Veröffentlicht: (2025)
von: Martonová, Denisa, et al.
Veröffentlicht: (2025)
MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO
von: Xiao, Yicheng, et al.
Veröffentlicht: (2025)
von: Xiao, Yicheng, et al.
Veröffentlicht: (2025)
Scaling, Benchmarking, and Reasoning of Vision-Language Agents for Mobile GUI Navigation
von: Qu, Heng, et al.
Veröffentlicht: (2026)
von: Qu, Heng, et al.
Veröffentlicht: (2026)
Jigsaw-Puzzles: From Seeing to Understanding to Reasoning in Vision-Language Models
von: Lyu, Zesen, et al.
Veröffentlicht: (2025)
von: Lyu, Zesen, et al.
Veröffentlicht: (2025)
VHELM: A Holistic Evaluation of Vision Language Models
von: Lee, Tony, et al.
Veröffentlicht: (2024)
von: Lee, Tony, et al.
Veröffentlicht: (2024)
Zero-shot Object Navigation with Vision-Language Models Reasoning
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
Commonsense Reasoning for Legged Robot Adaptation with Vision-Language Models
von: Chen, Annie S., et al.
Veröffentlicht: (2024)
von: Chen, Annie S., et al.
Veröffentlicht: (2024)
Probing Mechanical Reasoning in Large Vision Language Models
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
Diagnosing Causal Reasoning in Vision-Language Models via Structured Relevance Graphs
von: Pratama, Dhita Putri, et al.
Veröffentlicht: (2026)
von: Pratama, Dhita Putri, et al.
Veröffentlicht: (2026)
Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection
von: Zhou, Xiaona, et al.
Veröffentlicht: (2026)
von: Zhou, Xiaona, et al.
Veröffentlicht: (2026)
TangramSR: Can Vision-Language Models Reason in Continuous Geometric Space?
von: Zong, Yikun, et al.
Veröffentlicht: (2026)
von: Zong, Yikun, et al.
Veröffentlicht: (2026)
Pseudocode-Guided Structured Reasoning for Automating Reliable Inference in Vision-Language Models
von: Ni, Weicong, et al.
Veröffentlicht: (2026)
von: Ni, Weicong, et al.
Veröffentlicht: (2026)
TRACE: A Framework for Analyzing and Enhancing Stepwise Reasoning in Vision-Language Models
von: Imani, Shima, et al.
Veröffentlicht: (2025)
von: Imani, Shima, et al.
Veröffentlicht: (2025)
Machine Learning Techniques with Fairness for Prediction of Completion of Drug and Alcohol Rehabilitation
von: Roberts-Licklider, Karen, et al.
Veröffentlicht: (2024)
von: Roberts-Licklider, Karen, et al.
Veröffentlicht: (2024)
A Smart Multimodal Healthcare Copilot with Powerful LLM Reasoning
von: Zhao, Xuejiao, et al.
Veröffentlicht: (2025)
von: Zhao, Xuejiao, et al.
Veröffentlicht: (2025)
From Perception to Cognition: A Survey of Vision-Language Interactive Reasoning in Multimodal Large Language Models
von: Zhou, Chenyue, et al.
Veröffentlicht: (2025)
von: Zhou, Chenyue, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Which Programming Language and Model Work Best With LLM-as-a-Judge For Code Retrieval?
von: Roberts, Lucas, et al.
Veröffentlicht: (2025) -
Gender-Neutral Large Language Models for Medical Applications: Reducing Bias in PubMed Abstracts
von: Schaefer, Elizabeth, et al.
Veröffentlicht: (2025) -
Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models
von: Moore, Kyle, et al.
Veröffentlicht: (2025) -
Image2Struct: Benchmarking Structure Extraction for Vision-Language Models
von: Roberts, Josselin Somerville, et al.
Veröffentlicht: (2024) -
Exploitation Is All You Need... for Exploration
von: Rentschler, Micah, et al.
Veröffentlicht: (2025)