Blind Spot Navigation in Large Language Model Reasoning with Thought Space Explorer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Jinghan, Mo, Fengran, Weerasooriya, Tharindu Cyril, Ye, Xinyue, Wang, Dongjie, Fu, Yanjie, Liu, Kunpeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
StaRPO: Stability-Augmented Reinforcement Policy Optimization
von: Zhang, Jinghan, et al.
Veröffentlicht: (2026)
von: Zhang, Jinghan, et al.
Veröffentlicht: (2026)
Mitigating Shortcut Reasoning in Language Models: A Gradient-Aware Training Approach
von: Cao, Hongyu, et al.
Veröffentlicht: (2026)
von: Cao, Hongyu, et al.
Veröffentlicht: (2026)
LEKA:LLM-Enhanced Knowledge Augmentation
von: Zhang, Xinhao, et al.
Veröffentlicht: (2025)
von: Zhang, Xinhao, et al.
Veröffentlicht: (2025)
RATT: A Thought Structure for Coherent and Correct LLM Reasoning
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
Entropy-based Exploration Conduction for Multi-step Reasoning
von: Zhang, Jinghan, et al.
Veröffentlicht: (2025)
von: Zhang, Jinghan, et al.
Veröffentlicht: (2025)
Efficient Post-Training Refinement of Latent Reasoning in Large Language Models
von: Wang, Xinyuan, et al.
Veröffentlicht: (2025)
von: Wang, Xinyuan, et al.
Veröffentlicht: (2025)
SoT: Structured-of-Thought Prompting Guides Multilingual Reasoning in Large Language Models
von: Qi, Rui, et al.
Veröffentlicht: (2025)
von: Qi, Rui, et al.
Veröffentlicht: (2025)
Subasa - Adapting Language Models for Low-resourced Offensive Language Detection in Sinhala
von: Haturusinghe, Shanilka, et al.
Veröffentlicht: (2025)
von: Haturusinghe, Shanilka, et al.
Veröffentlicht: (2025)
Distilling Empathy from Large Language Models
von: Xie, Henry J., et al.
Veröffentlicht: (2025)
von: Xie, Henry J., et al.
Veröffentlicht: (2025)
Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM
von: Shetty, Samay U., et al.
Veröffentlicht: (2026)
von: Shetty, Samay U., et al.
Veröffentlicht: (2026)
Scoring with Large Language Models: A Study on Measuring Empathy of Responses in Dialogues
von: Xie, Henry J., et al.
Veröffentlicht: (2024)
von: Xie, Henry J., et al.
Veröffentlicht: (2024)
FedCoT: Communication-Efficient Federated Reasoning Enhancement for Large Language Models
von: Li, Chuan, et al.
Veröffentlicht: (2025)
von: Li, Chuan, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Feature Generation for Domain-Specific Classification
von: Zhang, Xinhao, et al.
Veröffentlicht: (2024)
von: Zhang, Xinhao, et al.
Veröffentlicht: (2024)
Evolving Demonstration Optimization for Chain-of-Thought Feature Transformation
von: Wang, Xinyuan, et al.
Veröffentlicht: (2026)
von: Wang, Xinyuan, et al.
Veröffentlicht: (2026)
Search-on-Graph: Iterative Informed Navigation for Large Language Model Reasoning on Knowledge Graphs
von: Sun, Jia Ao, et al.
Veröffentlicht: (2025)
von: Sun, Jia Ao, et al.
Veröffentlicht: (2025)
Temporal Blind Spots in Large Language Models
von: Wallat, Jonas, et al.
Veröffentlicht: (2024)
von: Wallat, Jonas, et al.
Veröffentlicht: (2024)
Vicarious Offense and Noise Audit of Offensive Speech Classifiers: Unifying Human and Machine Disagreement on What is Offensive
von: Weerasooriya, Tharindu Cyril, et al.
Veröffentlicht: (2023)
von: Weerasooriya, Tharindu Cyril, et al.
Veröffentlicht: (2023)
Linguistic Blind Spots of Large Language Models
von: Cheng, Jiali, et al.
Veröffentlicht: (2025)
von: Cheng, Jiali, et al.
Veröffentlicht: (2025)
Diversity-oriented Data Augmentation with Large Language Models
von: Wang, Zaitian, et al.
Veröffentlicht: (2025)
von: Wang, Zaitian, et al.
Veröffentlicht: (2025)
Simple Linguistic Inferences of Large Language Models (LLMs): Blind Spots and Blinds
von: Basmov, Victoria, et al.
Veröffentlicht: (2023)
von: Basmov, Victoria, et al.
Veröffentlicht: (2023)
LPI-RIT at LeWiDi-2025: Improving Distributional Predictions via Metadata and Loss Reweighting with DisCo
von: Sawkar, Mandira, et al.
Veröffentlicht: (2025)
von: Sawkar, Mandira, et al.
Veröffentlicht: (2025)
Measuring the Impact of Lexical Training Data Coverage on Hallucination Detection in Large Language Models
von: Zhang, Shuo, et al.
Veröffentlicht: (2025)
von: Zhang, Shuo, et al.
Veröffentlicht: (2025)
OpenDecoder: Open Large Language Model Decoding to Incorporate Document Quality in RAG
von: Mo, Fengran, et al.
Veröffentlicht: (2026)
von: Mo, Fengran, et al.
Veröffentlicht: (2026)
Buffer of Thoughts: Thought-Augmented Reasoning with Large Language Models
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
Rater Cohesion and Quality from a Vicarious Perspective
von: Pandita, Deepak, et al.
Veröffentlicht: (2024)
von: Pandita, Deepak, et al.
Veröffentlicht: (2024)
Can Large Language Models Understand Preferences in Personalized Recommendation?
von: Tan, Zhaoxuan, et al.
Veröffentlicht: (2025)
von: Tan, Zhaoxuan, et al.
Veröffentlicht: (2025)
Multilingual Collaborative Defense for Large Language Models
von: Li, Hongliang, et al.
Veröffentlicht: (2025)
von: Li, Hongliang, et al.
Veröffentlicht: (2025)
A User-Centric Multi-Intent Benchmark for Evaluating Large Language Models
von: Wang, Jiayin, et al.
Veröffentlicht: (2024)
von: Wang, Jiayin, et al.
Veröffentlicht: (2024)
Beyond SFT: Reinforcement Learning for Safer Large Reasoning Models with Better Reasoning Ability
von: Jia, Jinghan, et al.
Veröffentlicht: (2025)
von: Jia, Jinghan, et al.
Veröffentlicht: (2025)
ARTICLE: Annotator Reliability Through In-Context Learning
von: Dutta, Sujan, et al.
Veröffentlicht: (2024)
von: Dutta, Sujan, et al.
Veröffentlicht: (2024)
Unchecked and Overlooked: Addressing the Checkbox Blind Spot in Large Language Models with CheckboxQA
von: Turski, Michał, et al.
Veröffentlicht: (2025)
von: Turski, Michał, et al.
Veröffentlicht: (2025)
Learning to Route Queries to Heads for Attention-based Re-ranking with Large Language Models
von: Tian, Yuxing, et al.
Veröffentlicht: (2026)
von: Tian, Yuxing, et al.
Veröffentlicht: (2026)
Beyond Chains of Thought: Benchmarking Latent-Space Reasoning Abilities in Large Language Models
von: Hagendorff, Thilo, et al.
Veröffentlicht: (2025)
von: Hagendorff, Thilo, et al.
Veröffentlicht: (2025)
Fluent but Unfeeling: The Emotional Blind Spots of Language Models
von: Shu, Bangzhao, et al.
Veröffentlicht: (2025)
von: Shu, Bangzhao, et al.
Veröffentlicht: (2025)
Language Modeling Using Tensor Trains
von: Su, Zhan, et al.
Veröffentlicht: (2024)
von: Su, Zhan, et al.
Veröffentlicht: (2024)
ToW: Thoughts of Words Improve Reasoning in Large Language Models
von: Xu, Zhikun, et al.
Veröffentlicht: (2024)
von: Xu, Zhikun, et al.
Veröffentlicht: (2024)
Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning
von: Jia, Jinghan, et al.
Veröffentlicht: (2026)
von: Jia, Jinghan, et al.
Veröffentlicht: (2026)
ChatRetriever: Adapting Large Language Models for Generalized and Robust Conversational Dense Retrieval
von: Mao, Kelong, et al.
Veröffentlicht: (2024)
von: Mao, Kelong, et al.
Veröffentlicht: (2024)
WXImpactBench: A Disruptive Weather Impact Understanding Benchmark for Evaluating Large Language Models
von: Yu, Yongan, et al.
Veröffentlicht: (2025)
von: Yu, Yongan, et al.
Veröffentlicht: (2025)
MTMCS-Bench: Evaluating Contextual Safety of Multimodal Large Language Models in Multi-Turn Dialogues
von: Liu, Zheyuan, et al.
Veröffentlicht: (2026)
von: Liu, Zheyuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
StaRPO: Stability-Augmented Reinforcement Policy Optimization
von: Zhang, Jinghan, et al.
Veröffentlicht: (2026) -
Mitigating Shortcut Reasoning in Language Models: A Gradient-Aware Training Approach
von: Cao, Hongyu, et al.
Veröffentlicht: (2026) -
LEKA:LLM-Enhanced Knowledge Augmentation
von: Zhang, Xinhao, et al.
Veröffentlicht: (2025) -
RATT: A Thought Structure for Coherent and Correct LLM Reasoning
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024) -
Entropy-based Exploration Conduction for Multi-step Reasoning
von: Zhang, Jinghan, et al.
Veröffentlicht: (2025)