Bongards at the Boundary of Perception and Reasoning: Programs or Language?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Langenfeld, Cassidy, Beger, Claas, Geng, Gloria, Piriyakulkij, Wasu Top, Hu, Keya, Pu, Yewen, Ellis, Kevin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Doing Experiments and Revising Rules with Natural Language and Probabilistic Reasoning
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2024)
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2024)
Active Preference Inference using Language Models and Probabilistic Reasoning
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2023)
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2023)
Joint Learning of Hierarchical Neural Options and Abstract World Model
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2026)
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2026)
Bongard-RWR+: Real-World Representations of Fine-Grained Concepts in Bongard Problems
von: Pawlonka, Szymon, et al.
Veröffentlicht: (2025)
von: Pawlonka, Szymon, et al.
Veröffentlicht: (2025)
Reasoning Limitations of Multimodal Large Language Models. A Case Study of Bongard Problems
von: Małkiński, Mikołaj, et al.
Veröffentlicht: (2024)
von: Małkiński, Mikołaj, et al.
Veröffentlicht: (2024)
Support-Set Context Matters for Bongard Problems
von: Raghuraman, Nikhil, et al.
Veröffentlicht: (2023)
von: Raghuraman, Nikhil, et al.
Veröffentlicht: (2023)
PoE-World: Compositional World Modeling with Products of Programmatic Experts
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2025)
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2025)
Hypothesis Generation and Inductive Inference in Children and Language Models
von: Qin, Jeffrey, et al.
Veröffentlicht: (2026)
von: Qin, Jeffrey, et al.
Veröffentlicht: (2026)
CadVLM: Bridging Language and Vision in the Generation of Parametric CAD Sketches
von: Wu, Sifan, et al.
Veröffentlicht: (2024)
von: Wu, Sifan, et al.
Veröffentlicht: (2024)
Predicting Performance of Symbolic and Prompt Programs with Examples
von: Zheng, Chengqi, et al.
Veröffentlicht: (2026)
von: Zheng, Chengqi, et al.
Veröffentlicht: (2026)
Video-based Vehicle Surveillance in the Wild: License Plate, Make, and Model Recognition with Self Reflective Vision-Language Models
von: Parsa, Pouya, et al.
Veröffentlicht: (2025)
von: Parsa, Pouya, et al.
Veröffentlicht: (2025)
MIRAGE: A Multi-modal Benchmark for Spatial Perception, Reasoning, and Intelligence
von: Liu, Chonghan, et al.
Veröffentlicht: (2025)
von: Liu, Chonghan, et al.
Veröffentlicht: (2025)
AirCopBench: A Benchmark for Multi-drone Collaborative Embodied Perception and Reasoning
von: Zha, Jirong, et al.
Veröffentlicht: (2025)
von: Zha, Jirong, et al.
Veröffentlicht: (2025)
AgroNVILA: Perception-Reasoning Decoupling for Multi-view Agricultural Multimodal Large Language Models
von: Zhang, Jiarui, et al.
Veröffentlicht: (2026)
von: Zhang, Jiarui, et al.
Veröffentlicht: (2026)
Perception Before Reasoning: Two-Stage Reinforcement Learning for Visual Reasoning in Vision-Language Models
von: Chen, Yan, et al.
Veröffentlicht: (2025)
von: Chen, Yan, et al.
Veröffentlicht: (2025)
Bad Seeing or Bad Thinking? Rewarding Perception for Vision-Language Reasoning
von: Wang, Haozhe, et al.
Veröffentlicht: (2026)
von: Wang, Haozhe, et al.
Veröffentlicht: (2026)
Denoising Diffusion Variational Inference: Diffusion Models as Expressive Variational Posteriors
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2024)
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2024)
Grid-LOGAT: Grid Based Local and Global Area Transcription for Video Question Answering
von: Chowdhury, Md Intisar, et al.
Veröffentlicht: (2025)
von: Chowdhury, Md Intisar, et al.
Veröffentlicht: (2025)
When Robots Should Say "I Don't Know": Benchmarking Abstention in Embodied Question Answering
von: Wu, Tao, et al.
Veröffentlicht: (2025)
von: Wu, Tao, et al.
Veröffentlicht: (2025)
CaughtCheating: Is Your MLLM a Good Cheating Detective? Exploring the Boundary of Visual Perception and Reasoning
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
Delineating Knowledge Boundaries for Honest Large Vision-Language Models
von: Song, Junru, et al.
Veröffentlicht: (2026)
von: Song, Junru, et al.
Veröffentlicht: (2026)
The Reasoning Boundary Paradox: How Reinforcement Learning Constrains Language Models
von: Nguyen, Phuc Minh, et al.
Veröffentlicht: (2025)
von: Nguyen, Phuc Minh, et al.
Veröffentlicht: (2025)
Decoupling the Image Perception and Multimodal Reasoning for Reasoning Segmentation with Digital Twin Representations
von: Li, Yizhen, et al.
Veröffentlicht: (2025)
von: Li, Yizhen, et al.
Veröffentlicht: (2025)
Perception Tokens Enhance Visual Reasoning in Multimodal Language Models
von: Bigverdi, Mahtab, et al.
Veröffentlicht: (2024)
von: Bigverdi, Mahtab, et al.
Veröffentlicht: (2024)
Prompting Large Vision-Language Models for Compositional Reasoning
von: Ossowski, Timothy, et al.
Veröffentlicht: (2024)
von: Ossowski, Timothy, et al.
Veröffentlicht: (2024)
Think or Not? Selective Reasoning via Reinforcement Learning for Vision-Language Models
von: Wang, Jiaqi, et al.
Veröffentlicht: (2025)
von: Wang, Jiaqi, et al.
Veröffentlicht: (2025)
Discriminative Perception via Anchored Description for Reasoning Segmentation
von: Yang, Tao, et al.
Veröffentlicht: (2026)
von: Yang, Tao, et al.
Veröffentlicht: (2026)
Programmatic Video Prediction Using Large Language Models
von: Tang, Hao, et al.
Veröffentlicht: (2025)
von: Tang, Hao, et al.
Veröffentlicht: (2025)
Beyond Seeing: Evaluating Multimodal LLMs on Tool-Enabled Image Perception, Transformation, and Reasoning
von: Guo, Xingang, et al.
Veröffentlicht: (2025)
von: Guo, Xingang, et al.
Veröffentlicht: (2025)
VisuRiddles: Fine-grained Perception is a Primary Bottleneck for Multimodal Large Language Models in Abstract Visual Reasoning
von: Yan, Hao, et al.
Veröffentlicht: (2025)
von: Yan, Hao, et al.
Veröffentlicht: (2025)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
von: Xu, Tianshuo, et al.
Veröffentlicht: (2024)
Benchmarking Multi-Image Understanding in Vision and Language Models: Perception, Knowledge, Reasoning, and Multi-Hop Reasoning
von: Zhao, Bingchen, et al.
Veröffentlicht: (2024)
von: Zhao, Bingchen, et al.
Veröffentlicht: (2024)
A Cognitive Paradigm Approach to Probe the Perception-Reasoning Interface in VLMs
von: Vaishnav, Mohit, et al.
Veröffentlicht: (2025)
von: Vaishnav, Mohit, et al.
Veröffentlicht: (2025)
Guiding Perception-Reasoning Closer to Human in Blind Image Quality Assessment
von: Li, Yuan, et al.
Veröffentlicht: (2025)
von: Li, Yuan, et al.
Veröffentlicht: (2025)
Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
von: Fei, Hao, et al.
Veröffentlicht: (2024)
von: Fei, Hao, et al.
Veröffentlicht: (2024)
MediSee: Reasoning-based Pixel-level Perception in Medical Images
von: Tong, Qinyue, et al.
Veröffentlicht: (2025)
von: Tong, Qinyue, et al.
Veröffentlicht: (2025)
Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning
von: He, Hulingxiao, et al.
Veröffentlicht: (2026)
von: He, Hulingxiao, et al.
Veröffentlicht: (2026)
A Deep Learning Approach for Augmenting Perceptional Understanding of Histopathology Images
von: Hu, Xiaoqian
Veröffentlicht: (2025)
von: Hu, Xiaoqian
Veröffentlicht: (2025)
Rapid Motor Adaptation for Robotic Manipulator Arms
von: Liang, Yichao, et al.
Veröffentlicht: (2023)
von: Liang, Yichao, et al.
Veröffentlicht: (2023)
ARIADNE: A Perception-Reasoning Synergy Framework for Trustworthy Coronary Angiography Analysis
von: Jin, Zhan, et al.
Veröffentlicht: (2026)
von: Jin, Zhan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Doing Experiments and Revising Rules with Natural Language and Probabilistic Reasoning
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2024) -
Active Preference Inference using Language Models and Probabilistic Reasoning
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2023) -
Joint Learning of Hierarchical Neural Options and Abstract World Model
von: Piriyakulkij, Wasu Top, et al.
Veröffentlicht: (2026) -
Bongard-RWR+: Real-World Representations of Fine-Grained Concepts in Bongard Problems
von: Pawlonka, Szymon, et al.
Veröffentlicht: (2025) -
Reasoning Limitations of Multimodal Large Language Models. A Case Study of Bongard Problems
von: Małkiński, Mikołaj, et al.
Veröffentlicht: (2024)