Faithful-First Reasoning, Planning, and Acting for Multimodal LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Junxian, Xu, Xinyue, Ma, Sai, Zhang, Di, Li, Sichao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EO-Gym: A Multimodal, Interactive Environment for Earth Observation Agents
von: Ma, Sai, et al.
Veröffentlicht: (2026)
von: Ma, Sai, et al.
Veröffentlicht: (2026)
Evaluating LLM Understanding via Structured Tabular Decision Simulations
von: Li, Sichao, et al.
Veröffentlicht: (2025)
von: Li, Sichao, et al.
Veröffentlicht: (2025)
Dissociation of Faithful and Unfaithful Reasoning in LLMs
von: Yee, Evelyn, et al.
Veröffentlicht: (2024)
von: Yee, Evelyn, et al.
Veröffentlicht: (2024)
Pre-Act: Multi-Step Planning and Reasoning Improves Acting in LLM Agents
von: Rawat, Mrinal, et al.
Veröffentlicht: (2025)
von: Rawat, Mrinal, et al.
Veröffentlicht: (2025)
Non-myopic Generation of Language Models for Reasoning and Planning
von: Ma, Chang, et al.
Veröffentlicht: (2024)
von: Ma, Chang, et al.
Veröffentlicht: (2024)
Acting Flatterers via LLMs Sycophancy: Combating Clickbait with LLMs Opposing-Stance Reasoning
von: Zhang, Chaowei, et al.
Veröffentlicht: (2026)
von: Zhang, Chaowei, et al.
Veröffentlicht: (2026)
Faithful or Just Plausible? Evaluating the Faithfulness of Closed-Source LLMs in Medical Reasoning
von: Afolabi, Halimat, et al.
Veröffentlicht: (2026)
von: Afolabi, Halimat, et al.
Veröffentlicht: (2026)
FloCA: Towards Faithful and Logically Consistent Flowchart Reasoning
von: Zou, Jinzi, et al.
Veröffentlicht: (2026)
von: Zou, Jinzi, et al.
Veröffentlicht: (2026)
Planning and Acting While the Clock Ticks
von: Coles, Andrew, et al.
Veröffentlicht: (2024)
von: Coles, Andrew, et al.
Veröffentlicht: (2024)
Parallelized Planning-Acting for Efficient LLM-based Multi-Agent Systems in Minecraft
von: Li, Yaoru, et al.
Veröffentlicht: (2025)
von: Li, Yaoru, et al.
Veröffentlicht: (2025)
RPTS: Tree-Structured Reasoning Process Scoring for Faithful Multimodal Evaluation
von: Wang, Haofeng, et al.
Veröffentlicht: (2025)
von: Wang, Haofeng, et al.
Veröffentlicht: (2025)
FaithCoT-Bench: Benchmarking Instance-Level Faithfulness of Chain-of-Thought Reasoning
von: Shen, Xu, et al.
Veröffentlicht: (2025)
von: Shen, Xu, et al.
Veröffentlicht: (2025)
Faithful GRPO: Improving Visual Spatial Reasoning in Multimodal Language Models via Constrained Policy Optimization
von: Kancheti, Sai Srinivas, et al.
Veröffentlicht: (2026)
von: Kancheti, Sai Srinivas, et al.
Veröffentlicht: (2026)
FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning
von: Wang, Zeyu, et al.
Veröffentlicht: (2026)
von: Wang, Zeyu, et al.
Veröffentlicht: (2026)
Adaptive Reasoning and Acting in Medical Language Agents
von: Dutta, Abhishek, et al.
Veröffentlicht: (2024)
von: Dutta, Abhishek, et al.
Veröffentlicht: (2024)
Call Me When Necessary: LLMs can Efficiently and Faithfully Reason over Structured Environments
von: Cheng, Sitao, et al.
Veröffentlicht: (2024)
von: Cheng, Sitao, et al.
Veröffentlicht: (2024)
PRACT: Optimizing Principled Reasoning and Acting of LLM Agent
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
Cogito, Ergo Ludo: An Agent that Learns to Play by Reasoning and Planning
von: Wang, Sai, et al.
Veröffentlicht: (2025)
von: Wang, Sai, et al.
Veröffentlicht: (2025)
PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment
von: Li, Yantao, et al.
Veröffentlicht: (2026)
von: Li, Yantao, et al.
Veröffentlicht: (2026)
DeFacto: Counterfactual Thinking with Images for Enforcing Evidence-Grounded and Faithful Reasoning
von: Xu, Tianrun, et al.
Veröffentlicht: (2025)
von: Xu, Tianrun, et al.
Veröffentlicht: (2025)
Diving into Self-Evolving Training for Multimodal Reasoning
von: Liu, Wei, et al.
Veröffentlicht: (2024)
von: Liu, Wei, et al.
Veröffentlicht: (2024)
LABSHIELD: A Multimodal Benchmark for Safety-Critical Reasoning and Planning in Scientific Laboratories
von: Sun, Qianpu, et al.
Veröffentlicht: (2026)
von: Sun, Qianpu, et al.
Veröffentlicht: (2026)
GraphReAct: Reasoning and Acting for Multi-step Graph Inference
von: Yu, Xingtong, et al.
Veröffentlicht: (2026)
von: Yu, Xingtong, et al.
Veröffentlicht: (2026)
Reasoning on Graphs: Faithful and Interpretable Large Language Model Reasoning
von: Luo, Linhao, et al.
Veröffentlicht: (2023)
von: Luo, Linhao, et al.
Veröffentlicht: (2023)
Plan before Solving: Problem-Aware Strategy Routing for Mathematical Reasoning with LLMs
von: Qi, Shihao, et al.
Veröffentlicht: (2025)
von: Qi, Shihao, et al.
Veröffentlicht: (2025)
SpeakRL: Synergizing Reasoning, Speaking, and Acting in Language Models with Reinforcement Learning
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
von: Acikgoz, Emre Can, et al.
Veröffentlicht: (2025)
LBM: Hierarchical Large Auto-Bidding Model via Reasoning and Acting
von: Li, Yewen, et al.
Veröffentlicht: (2026)
von: Li, Yewen, et al.
Veröffentlicht: (2026)
E3-TIR: Enhanced Experience Exploitation for Tool-Integrated Reasoning
von: Guo, Weiyang, et al.
Veröffentlicht: (2026)
von: Guo, Weiyang, et al.
Veröffentlicht: (2026)
Chain-of-Thought Degrades Visual Spatial Reasoning Capabilities of Multimodal LLMs
von: Kancheti, Sai Srinivas, et al.
Veröffentlicht: (2026)
von: Kancheti, Sai Srinivas, et al.
Veröffentlicht: (2026)
Thinking, Faithful and Stable: Mitigating Hallucinations in LLMs
von: Zou, Chelsea, et al.
Veröffentlicht: (2025)
von: Zou, Chelsea, et al.
Veröffentlicht: (2025)
rSIM: Incentivizing Reasoning Capabilities of LLMs via Reinforced Strategy Injection
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
von: Chen, Sijia, et al.
Veröffentlicht: (2025)
SFR-RAG: Towards Contextually Faithful LLMs
von: Nguyen, Xuan-Phi, et al.
Veröffentlicht: (2024)
von: Nguyen, Xuan-Phi, et al.
Veröffentlicht: (2024)
To See or To Read: User Behavior Reasoning in Multimodal LLMs
von: Dong, Tianning, et al.
Veröffentlicht: (2025)
von: Dong, Tianning, et al.
Veröffentlicht: (2025)
Hidden in Plain Sight: Reasoning in Underspecified and Misspecified Scenarios for Multimodal LLMs
von: Yan, Qianqi, et al.
Veröffentlicht: (2025)
von: Yan, Qianqi, et al.
Veröffentlicht: (2025)
How Numerical Precision Affects Arithmetical Reasoning Capabilities of LLMs
von: Feng, Guhao, et al.
Veröffentlicht: (2024)
von: Feng, Guhao, et al.
Veröffentlicht: (2024)
Evaluating Readability and Faithfulness of Concept-based Explanations
von: Li, Meng, et al.
Veröffentlicht: (2024)
von: Li, Meng, et al.
Veröffentlicht: (2024)
ORIGAMISPACE: Benchmarking Multimodal LLMs in Multi-Step Spatial Reasoning with Mathematical Constraints
von: Xu, Rui, et al.
Veröffentlicht: (2025)
von: Xu, Rui, et al.
Veröffentlicht: (2025)
Proving Olympiad Inequalities by Synergizing LLMs and Symbolic Reasoning
von: Li, Zenan, et al.
Veröffentlicht: (2025)
von: Li, Zenan, et al.
Veröffentlicht: (2025)
An Optimization Algorithm for Multimodal Data Alignment
von: Zhang, Wei, et al.
Veröffentlicht: (2025)
von: Zhang, Wei, et al.
Veröffentlicht: (2025)
AnomSeer: Reinforcing Multimodal LLMs to Reason for Time-Series Anomaly Detection
von: Zhang, Junru, et al.
Veröffentlicht: (2026)
von: Zhang, Junru, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
EO-Gym: A Multimodal, Interactive Environment for Earth Observation Agents
von: Ma, Sai, et al.
Veröffentlicht: (2026) -
Evaluating LLM Understanding via Structured Tabular Decision Simulations
von: Li, Sichao, et al.
Veröffentlicht: (2025) -
Dissociation of Faithful and Unfaithful Reasoning in LLMs
von: Yee, Evelyn, et al.
Veröffentlicht: (2024) -
Pre-Act: Multi-Step Planning and Reasoning Improves Acting in LLM Agents
von: Rawat, Mrinal, et al.
Veröffentlicht: (2025) -
Non-myopic Generation of Language Models for Reasoning and Planning
von: Ma, Chang, et al.
Veröffentlicht: (2024)