Chain-of-Procedure: Hierarchical Visual-Language Reasoning for Procedural QA
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Guanhua, Yao, Yutong, Sun, Shenghe, Gao, Ci-Jun, Liu, Shudong, Chao, Lidia S., Wan, Feng, Wong, Derek F. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Not All LoRA Parameters Are Essential: Insights on Inference Necessity
von: Chen, Guanhua, et al.
Veröffentlicht: (2025)
von: Chen, Guanhua, et al.
Veröffentlicht: (2025)
A Two-Stage Prediction-Aware Contrastive Learning Framework for Multi-Intent NLU
von: Chen, Guanhua, et al.
Veröffentlicht: (2024)
von: Chen, Guanhua, et al.
Veröffentlicht: (2024)
From Scenes to Elements: Multi-Granularity Evidence Retrieval for Verifiable Multimodal RAG
von: Chen, Guanhua, et al.
Veröffentlicht: (2026)
von: Chen, Guanhua, et al.
Veröffentlicht: (2026)
SGIC: A Self-Guided Iterative Calibration Framework for RAG
von: Chen, Guanhua, et al.
Veröffentlicht: (2025)
von: Chen, Guanhua, et al.
Veröffentlicht: (2025)
Risks and NLP Design: A Case Study on Procedural Document QA
von: Haduong, Nikita, et al.
Veröffentlicht: (2024)
von: Haduong, Nikita, et al.
Veröffentlicht: (2024)
Procedural Knowledge at Scale Improves Reasoning
von: Wu, Di, et al.
Veröffentlicht: (2026)
von: Wu, Di, et al.
Veröffentlicht: (2026)
Let's Focus on Neuron: Neuron-Level Supervised Fine-tuning for Large Language Model
von: Xu, Haoyun, et al.
Veröffentlicht: (2024)
von: Xu, Haoyun, et al.
Veröffentlicht: (2024)
Can ChatGPT Really Understand Modern Chinese Poetry?
von: Wang, Shanshan, et al.
Veröffentlicht: (2026)
von: Wang, Shanshan, et al.
Veröffentlicht: (2026)
What is the Best Way for ChatGPT to Translate Poetry?
von: Wang, Shanshan, et al.
Veröffentlicht: (2024)
von: Wang, Shanshan, et al.
Veröffentlicht: (2024)
ISO-Bench: Benchmarking Multimodal Causal Reasoning in Visual-Language Models through Procedural Plans
von: Sadana, Ananya, et al.
Veröffentlicht: (2025)
von: Sadana, Ananya, et al.
Veröffentlicht: (2025)
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
von: Ma, Jingkun, et al.
Veröffentlicht: (2024)
Path Drift in Large Reasoning Models:How First-Person Commitments Override Safety
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
ProMQA-Assembly: Multimodal Procedural QA Dataset on Assembly
von: Hasegawa, Kimihiro, et al.
Veröffentlicht: (2025)
von: Hasegawa, Kimihiro, et al.
Veröffentlicht: (2025)
L0-Reasoning Bench: Evaluating Procedural Correctness in Language Models via Simple Program Execution
von: Sun, Simeng, et al.
Veröffentlicht: (2025)
von: Sun, Simeng, et al.
Veröffentlicht: (2025)
Harnessing the Reasoning Economy: A Survey of Efficient Reasoning for Large Language Models
von: Wang, Rui, et al.
Veröffentlicht: (2025)
von: Wang, Rui, et al.
Veröffentlicht: (2025)
Procedural Knowledge in Pretraining Drives Reasoning in Large Language Models
von: Ruis, Laura, et al.
Veröffentlicht: (2024)
von: Ruis, Laura, et al.
Veröffentlicht: (2024)
Procedural Dilemma Generation for Evaluating Moral Reasoning in Humans and Language Models
von: Fränken, Jan-Philipp, et al.
Veröffentlicht: (2024)
von: Fränken, Jan-Philipp, et al.
Veröffentlicht: (2024)
Intrinsic Model Weaknesses: How Priming Attacks Unveil Vulnerabilities in Large Language Models
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
Multilingual Reasoning Gym: Multilingual Scaling of Procedural Reasoning Environments
von: Dobler, Konstantin, et al.
Veröffentlicht: (2026)
von: Dobler, Konstantin, et al.
Veröffentlicht: (2026)
Are Large Reasoning Models Good Translation Evaluators? Analysis and Performance Boost
von: Zhan, Runzhe, et al.
Veröffentlicht: (2025)
von: Zhan, Runzhe, et al.
Veröffentlicht: (2025)
Prefix Text as a Yarn: Eliciting Non-English Alignment in Foundation Language Model
von: Zhan, Runzhe, et al.
Veröffentlicht: (2024)
von: Zhan, Runzhe, et al.
Veröffentlicht: (2024)
AutoPRM: Automating Procedural Supervision for Multi-Step Reasoning via Controllable Question Decomposition
von: Chen, Zhaorun, et al.
Veröffentlicht: (2024)
von: Chen, Zhaorun, et al.
Veröffentlicht: (2024)
Rethinking Prompt-based Debiasing in Large Language Models
von: Yang, Xinyi, et al.
Veröffentlicht: (2025)
von: Yang, Xinyi, et al.
Veröffentlicht: (2025)
Can Large Language Models Generalize Procedures Across Representations?
von: Lin, Fangru, et al.
Veröffentlicht: (2026)
von: Lin, Fangru, et al.
Veröffentlicht: (2026)
Hierarchical Chain-of-Thought Prompting: Enhancing LLM Reasoning Performance and Efficiency
von: Huang, Xingshuai, et al.
Veröffentlicht: (2026)
von: Huang, Xingshuai, et al.
Veröffentlicht: (2026)
Unveiling LLMs' Metaphorical Understanding: Exploring Conceptual Irrelevance, Context Leveraging and Syntactic Influence
von: Ye, Fengying, et al.
Veröffentlicht: (2025)
von: Ye, Fengying, et al.
Veröffentlicht: (2025)
Is Your Model Really A Good Math Reasoner? Evaluating Mathematical Reasoning with Checklist
von: Zhou, Zihao, et al.
Veröffentlicht: (2024)
von: Zhou, Zihao, et al.
Veröffentlicht: (2024)
LongProc: Benchmarking Long-Context Language Models on Long Procedural Generation
von: Ye, Xi, et al.
Veröffentlicht: (2025)
von: Ye, Xi, et al.
Veröffentlicht: (2025)
Legal Mathematical Reasoning with LLMs: Procedural Alignment through Two-Stage Reinforcement Learning
von: Zhang, Kepu, et al.
Veröffentlicht: (2025)
von: Zhang, Kepu, et al.
Veröffentlicht: (2025)
Evaluating LLMs' Reasoning Over Ordered Procedural Steps
von: Anika, Adrita, et al.
Veröffentlicht: (2025)
von: Anika, Adrita, et al.
Veröffentlicht: (2025)
Learning from "Silly" Questions Improves Large Language Models, But Only Slightly
von: Zhu, Tingyuan, et al.
Veröffentlicht: (2024)
von: Zhu, Tingyuan, et al.
Veröffentlicht: (2024)
HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA
von: Chen, Xinyue, et al.
Veröffentlicht: (2024)
von: Chen, Xinyue, et al.
Veröffentlicht: (2024)
Benchmarking the Detection of LLMs-Generated Modern Chinese Poetry
von: Wang, Shanshan, et al.
Veröffentlicht: (2025)
von: Wang, Shanshan, et al.
Veröffentlicht: (2025)
Decision Procedure for A Theory of String Sequences
von: Hu, Denghang, et al.
Veröffentlicht: (2025)
von: Hu, Denghang, et al.
Veröffentlicht: (2025)
Pairing Analogy-Augmented Generation with Procedural Memory for Procedural Q&A
von: Roth, K, et al.
Veröffentlicht: (2024)
von: Roth, K, et al.
Veröffentlicht: (2024)
ActPlan-1K: Benchmarking the Procedural Planning Ability of Visual Language Models in Household Activities
von: Su, Ying, et al.
Veröffentlicht: (2024)
von: Su, Ying, et al.
Veröffentlicht: (2024)
Boosting Language Models Reasoning with Chain-of-Knowledge Prompting
von: Wang, Jianing, et al.
Veröffentlicht: (2023)
von: Wang, Jianing, et al.
Veröffentlicht: (2023)
FOCUS: Forging Originality through Contrastive Use in Self-Plagiarism for Language Models
von: Lan, Kaixin, et al.
Veröffentlicht: (2024)
von: Lan, Kaixin, et al.
Veröffentlicht: (2024)
Phraselette: A Poet's Procedural Palette
von: Calderwood, Alex, et al.
Veröffentlicht: (2025)
von: Calderwood, Alex, et al.
Veröffentlicht: (2025)
ProcBench: Benchmark for Multi-Step Reasoning and Following Procedure
von: Fujisawa, Ippei, et al.
Veröffentlicht: (2024)
von: Fujisawa, Ippei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Not All LoRA Parameters Are Essential: Insights on Inference Necessity
von: Chen, Guanhua, et al.
Veröffentlicht: (2025) -
A Two-Stage Prediction-Aware Contrastive Learning Framework for Multi-Intent NLU
von: Chen, Guanhua, et al.
Veröffentlicht: (2024) -
From Scenes to Elements: Multi-Granularity Evidence Retrieval for Verifiable Multimodal RAG
von: Chen, Guanhua, et al.
Veröffentlicht: (2026) -
SGIC: A Self-Guided Iterative Calibration Framework for RAG
von: Chen, Guanhua, et al.
Veröffentlicht: (2025) -
Risks and NLP Design: A Case Study on Procedural Document QA
von: Haduong, Nikita, et al.
Veröffentlicht: (2024)