Pseudocode-Guided Structured Reasoning for Automating Reliable Inference in Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ni, Weicong, Jiang, Tianbao, Wang, Linlin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization
von: Cao, Zouying, et al.
Veröffentlicht: (2025)
von: Cao, Zouying, et al.
Veröffentlicht: (2025)
Scaling Automatic Extraction of Pseudocode
von: Toksoz, Levent, et al.
Veröffentlicht: (2024)
von: Toksoz, Levent, et al.
Veröffentlicht: (2024)
ATA: Bridging Implicit Reasoning with Attention-Guided and Action-Guided Inference for Vision-Language Action Models
von: Yang, Cheng, et al.
Veröffentlicht: (2026)
von: Yang, Cheng, et al.
Veröffentlicht: (2026)
PseudoAct: Leveraging Pseudocode Synthesis for Flexible Planning and Action Control in Large Language Model Agents
von: Yihan, et al.
Veröffentlicht: (2026)
von: Yihan, et al.
Veröffentlicht: (2026)
Graph-to-Vision: Multi-graph Understanding and Reasoning using Vision-Language Models
von: Ai, Qihang, et al.
Veröffentlicht: (2025)
von: Ai, Qihang, et al.
Veröffentlicht: (2025)
PCodeTrans: Translate Decompiled Pseudocode to Compilable and Executable Equivalent
von: Cui, Yuxin, et al.
Veröffentlicht: (2026)
von: Cui, Yuxin, et al.
Veröffentlicht: (2026)
ShorterBetter: Guiding Reasoning Models to Find Optimal Inference Length for Efficient Reasoning
von: Yi, Jingyang, et al.
Veröffentlicht: (2025)
von: Yi, Jingyang, et al.
Veröffentlicht: (2025)
Pseudocode-Injection Magic: Enabling LLMs to Tackle Graph Computational Tasks
von: Gong, Chang, et al.
Veröffentlicht: (2025)
von: Gong, Chang, et al.
Veröffentlicht: (2025)
3DGSNav: Enhancing Vision-Language Model Reasoning for Object Navigation via Active 3D Gaussian Splatting
von: Zheng, Wancai, et al.
Veröffentlicht: (2026)
von: Zheng, Wancai, et al.
Veröffentlicht: (2026)
AGMark: Attention-Guided Dynamic Watermarking for Large Vision-Language Models
von: Li, Yue, et al.
Veröffentlicht: (2026)
von: Li, Yue, et al.
Veröffentlicht: (2026)
Boosting Inference with Guided Reasoning: Stochastic Exploration for Recursive Models
von: Corbett, Andrew, et al.
Veröffentlicht: (2026)
von: Corbett, Andrew, et al.
Veröffentlicht: (2026)
Diagnosing Causal Reasoning in Vision-Language Models via Structured Relevance Graphs
von: Pratama, Dhita Putri, et al.
Veröffentlicht: (2026)
von: Pratama, Dhita Putri, et al.
Veröffentlicht: (2026)
SPREG: Structured Plan Repair with Entropy-Guided Test-Time Intervention for Large Language Model Reasoning
von: Wang, Xuan, et al.
Veröffentlicht: (2026)
von: Wang, Xuan, et al.
Veröffentlicht: (2026)
Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding
von: Fang, Yixiong, et al.
Veröffentlicht: (2024)
von: Fang, Yixiong, et al.
Veröffentlicht: (2024)
DIRCR: Dual-Inference Rule-Contrastive Reasoning for Solving RAVENs
von: Zhang, Jiachen, et al.
Veröffentlicht: (2026)
von: Zhang, Jiachen, et al.
Veröffentlicht: (2026)
MALSIGHT: Exploring Malicious Source Code and Benign Pseudocode for Iterative Binary Malware Summarization
von: Lu, Haolang, et al.
Veröffentlicht: (2024)
von: Lu, Haolang, et al.
Veröffentlicht: (2024)
Jigsaw-Puzzles: From Seeing to Understanding to Reasoning in Vision-Language Models
von: Lyu, Zesen, et al.
Veröffentlicht: (2025)
von: Lyu, Zesen, et al.
Veröffentlicht: (2025)
Vision Language Models Cannot Reason About Physical Transformation
von: Luo, Dezhi, et al.
Veröffentlicht: (2026)
von: Luo, Dezhi, et al.
Veröffentlicht: (2026)
Locatability-Guided Adaptive Reasoning for Image Geo-Localization with Vision-Language Models
von: Yu, Bo, et al.
Veröffentlicht: (2026)
von: Yu, Bo, et al.
Veröffentlicht: (2026)
Improving Large Language Models Function Calling and Interpretability via Guided-Structured Templates
von: Dang, Hy, et al.
Veröffentlicht: (2025)
von: Dang, Hy, et al.
Veröffentlicht: (2025)
ACPO: Adaptive Curriculum Policy Optimization for Aligning Vision-Language Models in Complex Reasoning
von: Wang, Yunhao, et al.
Veröffentlicht: (2025)
von: Wang, Yunhao, et al.
Veröffentlicht: (2025)
Prompting Large Vision-Language Models for Compositional Reasoning
von: Ossowski, Timothy, et al.
Veröffentlicht: (2024)
von: Ossowski, Timothy, et al.
Veröffentlicht: (2024)
Visual Distraction Undermines Moral Reasoning in Vision-Language Models
von: Yang, Xinyi, et al.
Veröffentlicht: (2026)
von: Yang, Xinyi, et al.
Veröffentlicht: (2026)
Draft-and-Prune: Improving the Reliability of Auto-formalization for Logical Reasoning
von: Ni, Zhiyu, et al.
Veröffentlicht: (2026)
von: Ni, Zhiyu, et al.
Veröffentlicht: (2026)
Harness as an Asset: Enforcing Determinism via the Convergent AI Agent Framework (CAAF)
von: Zhang, Tianbao
Veröffentlicht: (2026)
von: Zhang, Tianbao
Veröffentlicht: (2026)
DisCO: Reinforcing Large Reasoning Models with Discriminative Constrained Optimization
von: Li, Gang, et al.
Veröffentlicht: (2025)
von: Li, Gang, et al.
Veröffentlicht: (2025)
Planning with Reasoning using Vision Language World Model
von: Chen, Delong, et al.
Veröffentlicht: (2025)
von: Chen, Delong, et al.
Veröffentlicht: (2025)
How Open Must Language Models be to Enable Reliable Scientific Inference?
von: Michaelov, James A., et al.
Veröffentlicht: (2026)
von: Michaelov, James A., et al.
Veröffentlicht: (2026)
Safe Semantics, Unsafe Interpretations: Tackling Implicit Reasoning Safety in Large Vision-Language Models
von: Cai, Wei, et al.
Veröffentlicht: (2025)
von: Cai, Wei, et al.
Veröffentlicht: (2025)
Logic Error Localization in Student Programming Assignments Using Pseudocode and Graph Neural Networks
von: Xu, Zhenyu, et al.
Veröffentlicht: (2024)
von: Xu, Zhenyu, et al.
Veröffentlicht: (2024)
Large Language Model for Verilog Generation with Code-Structure-Guided Reinforcement Learning
von: Wang, Ning, et al.
Veröffentlicht: (2024)
von: Wang, Ning, et al.
Veröffentlicht: (2024)
Reliable Reasoning Beyond Natural Language
von: Borazjanizadeh, Nasim, et al.
Veröffentlicht: (2024)
von: Borazjanizadeh, Nasim, et al.
Veröffentlicht: (2024)
Smart Vision-Language Reasoners
von: Roberts, Denisa, et al.
Veröffentlicht: (2024)
von: Roberts, Denisa, et al.
Veröffentlicht: (2024)
GRE Suite: Geo-localization Inference via Fine-Tuned Vision-Language Models and Enhanced Reasoning Chains
von: Wang, Chun, et al.
Veröffentlicht: (2025)
von: Wang, Chun, et al.
Veröffentlicht: (2025)
A Comprehensive Survey and Guide to Multimodal Large Language Models in Vision-Language Tasks
von: Liang, Chia Xin, et al.
Veröffentlicht: (2024)
von: Liang, Chia Xin, et al.
Veröffentlicht: (2024)
LLM-BI: Towards Fully Automated Bayesian Inference with Large Language Models
von: Huang, Yongchao
Veröffentlicht: (2025)
von: Huang, Yongchao
Veröffentlicht: (2025)
Guiding Language Model Reasoning with Planning Tokens
von: Wang, Xinyi, et al.
Veröffentlicht: (2023)
von: Wang, Xinyi, et al.
Veröffentlicht: (2023)
Large Language Models as an Indirect Reasoner: Contrapositive and Contradiction for Automated Reasoning
von: Zhang, Yanfang, et al.
Veröffentlicht: (2024)
von: Zhang, Yanfang, et al.
Veröffentlicht: (2024)
Investigating The Functional Roles of Attention Heads in Vision Language Models: Evidence for Reasoning Modules
von: Jiang, Yanbei, et al.
Veröffentlicht: (2025)
von: Jiang, Yanbei, et al.
Veröffentlicht: (2025)
Automated Hazard Detection in Construction Sites Using Large Language and Vision-Language Models
von: Sahraoui, Islem
Veröffentlicht: (2025)
von: Sahraoui, Islem
Veröffentlicht: (2025)
Ähnliche Einträge
-
PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization
von: Cao, Zouying, et al.
Veröffentlicht: (2025) -
Scaling Automatic Extraction of Pseudocode
von: Toksoz, Levent, et al.
Veröffentlicht: (2024) -
ATA: Bridging Implicit Reasoning with Attention-Guided and Action-Guided Inference for Vision-Language Action Models
von: Yang, Cheng, et al.
Veröffentlicht: (2026) -
PseudoAct: Leveraging Pseudocode Synthesis for Flexible Planning and Action Control in Large Language Model Agents
von: Yihan, et al.
Veröffentlicht: (2026) -
Graph-to-Vision: Multi-graph Understanding and Reasoning using Vision-Language Models
von: Ai, Qihang, et al.
Veröffentlicht: (2025)