Wrong-of-Thought: An Integrated Reasoning Framework with Multi-Perspective Verification and Wrong Information
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yongheng, Chen, Qiguang, Zhou, Jingxuan, Wang, Peng, Si, Jiasheng, Wang, Jin, Lu, Wenpeng, Qin, Libo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CroPrompt: Cross-task Interactive Prompting for Zero-shot Spoken Language Understanding
von: Qin, Libo, et al.
Veröffentlicht: (2024)
von: Qin, Libo, et al.
Veröffentlicht: (2024)
Unlocking the Capabilities of Thought: A Reasoning Boundary Framework to Quantify and Optimize Chain-of-Thought
von: Chen, Qiguang, et al.
Veröffentlicht: (2024)
von: Chen, Qiguang, et al.
Veröffentlicht: (2024)
AutoCAP: Towards Automatic Cross-lingual Alignment Planning for Zero-shot Chain-of-Thought
von: Zhang, Yongheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yongheng, et al.
Veröffentlicht: (2024)
RBF++: Quantifying and Optimizing Reasoning Boundaries across Measurable and Unmeasurable Capabilities for Chain-of-Thought Reasoning
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
CCHall: A Novel Benchmark for Joint Cross-Lingual and Cross-Modal Hallucinations Detection in Large Language Models
von: Zhang, Yongheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yongheng, et al.
Veröffentlicht: (2025)
Let's Think with Images Efficiently! An Interleaved-Modal Chain-of-Thought Reasoning Framework with Dynamic and Precise Visual Thoughts
von: Liu, Xu, et al.
Veröffentlicht: (2026)
von: Liu, Xu, et al.
Veröffentlicht: (2026)
Varying Shades of Wrong: Aligning LLMs with Wrong Answers Only
von: Yao, Jihan, et al.
Veröffentlicht: (2024)
von: Yao, Jihan, et al.
Veröffentlicht: (2024)
M$^3$CoT: A Novel Benchmark for Multi-Domain Multi-step Multi-modal Chain-of-Thought
von: Chen, Qiguang, et al.
Veröffentlicht: (2024)
von: Chen, Qiguang, et al.
Veröffentlicht: (2024)
ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models
von: Zhang, Yongheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yongheng, et al.
Veröffentlicht: (2025)
CHECKWHY: Causal Fact Verification via Argument Structure
von: Si, Jiasheng, et al.
Veröffentlicht: (2024)
von: Si, Jiasheng, et al.
Veröffentlicht: (2024)
Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
Less Languages, Less Tokens: An Efficient Unified Logic Cross-lingual Chain-of-Thought Reasoning Framework
von: Zhang, Chenyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Chenyuan, et al.
Veröffentlicht: (2026)
DLPO: Towards a Robust, Efficient, and Generalizable Prompt Optimization Framework from a Deep-Learning Perspective
von: Peng, Dengyun, et al.
Veröffentlicht: (2025)
von: Peng, Dengyun, et al.
Veröffentlicht: (2025)
Visual Thoughts: A Unified Perspective of Understanding Multimodal Chain-of-Thought
von: Cheng, Zihui, et al.
Veröffentlicht: (2025)
von: Cheng, Zihui, et al.
Veröffentlicht: (2025)
Correct Prediction, Wrong Steps? Consensus Reasoning Knowledge Graph for Robust Chain-of-Thought Synthesis
von: Ling, Zipeng, et al.
Veröffentlicht: (2026)
von: Ling, Zipeng, et al.
Veröffentlicht: (2026)
Beyond Surface Reasoning: Unveiling the True Long Chain-of-Thought Capacity of Diffusion Large Language Models
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
Beware of Reasoning Overconfidence: Pitfalls in the Reasoning Process for Multi-solution Tasks
von: Guan, Jiannan, et al.
Veröffentlicht: (2025)
von: Guan, Jiannan, et al.
Veröffentlicht: (2025)
X-WebAgentBench: A Multilingual Interactive Web Benchmark for Evaluating Global Agentic System
von: Wang, Peng, et al.
Veröffentlicht: (2025)
von: Wang, Peng, et al.
Veröffentlicht: (2025)
The Molecular Structure of Thought: Mapping the Topology of Long Chain-of-Thought Reasoning
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
von: Chen, Qiguang, et al.
Veröffentlicht: (2026)
Distinguishing Right from Wrong in Debates: Attribution Analysis of Chinese Harmful Memes
von: Wang, Weiming, et al.
Veröffentlicht: (2026)
von: Wang, Weiming, et al.
Veröffentlicht: (2026)
What is Wrong with Perplexity for Long-context Language Modeling?
von: Fang, Lizhe, et al.
Veröffentlicht: (2024)
von: Fang, Lizhe, et al.
Veröffentlicht: (2024)
Wrong as Sequence Violation: The Structural Definition of Wrong as Misordering Across Reasoning, Physics, Computation, Cognition, and Ethics
von: Stewart, Arthur
Veröffentlicht: (2026)
von: Stewart, Arthur
Veröffentlicht: (2026)
SRLCG: Self-Rectified Large-Scale Code Generation with Multidimensional Chain-of-Thought and Dynamic Backtracking
von: Ma, Hongru, et al.
Veröffentlicht: (2025)
von: Ma, Hongru, et al.
Veröffentlicht: (2025)
Visual Grounding Methods for VQA are Working for the Wrong Reasons!
von: Shrestha, Robik, et al.
Veröffentlicht: (2020)
von: Shrestha, Robik, et al.
Veröffentlicht: (2020)
MM-Verify: Enhancing Multimodal Reasoning with Chain-of-Thought Verification
von: Sun, Linzhuang, et al.
Veröffentlicht: (2025)
von: Sun, Linzhuang, et al.
Veröffentlicht: (2025)
Humans Perceive Wrong Narratives from AI Reasoning Texts
von: Levy, Mosh, et al.
Veröffentlicht: (2025)
von: Levy, Mosh, et al.
Veröffentlicht: (2025)
Constructions Are So Difficult That Even Large Language Models Get Them Right for the Wrong Reasons
von: Zhou, Shijia, et al.
Veröffentlicht: (2024)
von: Zhou, Shijia, et al.
Veröffentlicht: (2024)
It's Not Easy Being Wrong: Large Language Models Struggle with Process of Elimination Reasoning
von: Balepur, Nishant, et al.
Veröffentlicht: (2023)
von: Balepur, Nishant, et al.
Veröffentlicht: (2023)
More or Less Wrong: A Benchmark for Directional Bias in LLM Comparative Reasoning
von: Shafiei, Mohammadamin, et al.
Veröffentlicht: (2025)
von: Shafiei, Mohammadamin, et al.
Veröffentlicht: (2025)
Easy Problems That LLMs Get Wrong
von: Williams, Sean, et al.
Veröffentlicht: (2024)
von: Williams, Sean, et al.
Veröffentlicht: (2024)
The Realignment Problem: When Right becomes Wrong in LLMs
von: Sharma, Aakash Sen, et al.
Veröffentlicht: (2025)
von: Sharma, Aakash Sen, et al.
Veröffentlicht: (2025)
Large Language Models Meet NLP: A Survey
von: Qin, Libo, et al.
Veröffentlicht: (2024)
von: Qin, Libo, et al.
Veröffentlicht: (2024)
Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?
von: Wen, Xueru, et al.
Veröffentlicht: (2024)
von: Wen, Xueru, et al.
Veröffentlicht: (2024)
Large Language Models Help Humans Verify Truthfulness -- Except When They Are Convincingly Wrong
von: Si, Chenglei, et al.
Veröffentlicht: (2023)
von: Si, Chenglei, et al.
Veröffentlicht: (2023)
Aware First, Think Less: Dynamic Boundary Self-Awareness Drives Extreme Reasoning Efficiency in Large Language Models
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
von: Chen, Qiguang, et al.
Veröffentlicht: (2025)
ClimateViz: A Benchmark for Statistical Reasoning and Fact Verification on Scientific Charts
von: Su, Ruiran, et al.
Veröffentlicht: (2025)
von: Su, Ruiran, et al.
Veröffentlicht: (2025)
Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Policy Evaluation
von: He, Yanjie
Veröffentlicht: (2026)
von: He, Yanjie
Veröffentlicht: (2026)
Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict
von: Chen, Yihang, et al.
Veröffentlicht: (2026)
von: Chen, Yihang, et al.
Veröffentlicht: (2026)
Not Wrong, But Untrue: LLM Overconfidence in Document-Based Queries
von: Hagar, Nick, et al.
Veröffentlicht: (2025)
von: Hagar, Nick, et al.
Veröffentlicht: (2025)
What's Wrong? Refining Meeting Summaries with LLM Feedback
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
von: Kirstein, Frederic, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CroPrompt: Cross-task Interactive Prompting for Zero-shot Spoken Language Understanding
von: Qin, Libo, et al.
Veröffentlicht: (2024) -
Unlocking the Capabilities of Thought: A Reasoning Boundary Framework to Quantify and Optimize Chain-of-Thought
von: Chen, Qiguang, et al.
Veröffentlicht: (2024) -
AutoCAP: Towards Automatic Cross-lingual Alignment Planning for Zero-shot Chain-of-Thought
von: Zhang, Yongheng, et al.
Veröffentlicht: (2024) -
RBF++: Quantifying and Optimizing Reasoning Boundaries across Measurable and Unmeasurable Capabilities for Chain-of-Thought Reasoning
von: Chen, Qiguang, et al.
Veröffentlicht: (2025) -
CCHall: A Novel Benchmark for Joint Cross-Lingual and Cross-Modal Hallucinations Detection in Large Language Models
von: Zhang, Yongheng, et al.
Veröffentlicht: (2025)