Reasoning or a Semblance of it? A Diagnostic Study of Transitive Reasoning in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mehrafarin, Houman, Eshghi, Arash, Konstas, Ioannis |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
When Chain-of-Thought Fails, the Solution Hides in the Hidden States
von: Mehrafarin, Houman, et al.
Veröffentlicht: (2026)
von: Mehrafarin, Houman, et al.
Veröffentlicht: (2026)
AlanaVLM: A Multimodal Embodied AI Foundation Model for Egocentric Video Understanding
von: Suglia, Alessandro, et al.
Veröffentlicht: (2024)
von: Suglia, Alessandro, et al.
Veröffentlicht: (2024)
Voices in a Crowd: Searching for Clusters of Unique Perspectives
von: Vitsakis, Nikolas, et al.
Veröffentlicht: (2024)
von: Vitsakis, Nikolas, et al.
Veröffentlicht: (2024)
Repairs in a Block World: A New Benchmark for Handling User Corrections with Multi-Modal Language Models
von: Chiyah-Garcia, Javier, et al.
Veröffentlicht: (2024)
von: Chiyah-Garcia, Javier, et al.
Veröffentlicht: (2024)
MoRFI: Monotonic Sparse Autoencoder Feature Identification
von: Dimakopoulos, Dimitris, et al.
Veröffentlicht: (2026)
von: Dimakopoulos, Dimitris, et al.
Veröffentlicht: (2026)
Erasing 'Ugly' from the Internet: Propagation of the Beauty Myth in Text-Image Models
von: Dinkar, Tanvi, et al.
Veröffentlicht: (2025)
von: Dinkar, Tanvi, et al.
Veröffentlicht: (2025)
Investigating the Role of Instruction Variety and Task Difficulty in Robotic Manipulation Tasks
von: Parekh, Amit, et al.
Veröffentlicht: (2024)
von: Parekh, Amit, et al.
Veröffentlicht: (2024)
Re-examining Sexism and Misogyny Classification with Annotator Attitudes
von: Jiang, Aiqi, et al.
Veröffentlicht: (2024)
von: Jiang, Aiqi, et al.
Veröffentlicht: (2024)
CROPE: Evaluating In-Context Adaptation of Vision and Language Models to Culture-Specific Concepts
von: Nikandrou, Malvina, et al.
Veröffentlicht: (2024)
von: Nikandrou, Malvina, et al.
Veröffentlicht: (2024)
Kwai-STaR: Transform LLMs into State-Transition Reasoners
von: Lu, Xingyu, et al.
Veröffentlicht: (2024)
von: Lu, Xingyu, et al.
Veröffentlicht: (2024)
On Code-Induced Reasoning in LLMs
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
von: Waheed, Abdul, et al.
Veröffentlicht: (2025)
ProtoReasoning: Prototypes as the Foundation for Generalizable Reasoning in LLMs
von: He, Feng, et al.
Veröffentlicht: (2025)
von: He, Feng, et al.
Veröffentlicht: (2025)
Learning Diagnostic Reasoning for Decision Support in Toxicology
von: Oberländer, Nico, et al.
Veröffentlicht: (2026)
von: Oberländer, Nico, et al.
Veröffentlicht: (2026)
Presupposition and Reasoning in Conditionals: A Theory-Based Study of Humans and LLMs
von: Azin, Tara, et al.
Veröffentlicht: (2026)
von: Azin, Tara, et al.
Veröffentlicht: (2026)
Exploring the Potential of Offline RL for Reasoning in LLMs: A Preliminary Study
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Tian, Xiaoyu, et al.
Veröffentlicht: (2025)
Reasoning Under Uncertainty: Exploring Probabilistic Reasoning Capabilities of LLMs
von: Pournemat, Mobina, et al.
Veröffentlicht: (2025)
von: Pournemat, Mobina, et al.
Veröffentlicht: (2025)
Reasoning with Graphs: Structuring Implicit Knowledge to Enhance LLMs Reasoning
von: Han, Haoyu, et al.
Veröffentlicht: (2025)
von: Han, Haoyu, et al.
Veröffentlicht: (2025)
AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP
von: Hasanaath, Ahmed, et al.
Veröffentlicht: (2025)
von: Hasanaath, Ahmed, et al.
Veröffentlicht: (2025)
Reasoning as State Transition: A Representational Analysis of Reasoning Evolution in Large Language Models
von: Zhang, Siyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Siyuan, et al.
Veröffentlicht: (2026)
LLMs as Models for Analogical Reasoning
von: Musker, Sam, et al.
Veröffentlicht: (2024)
von: Musker, Sam, et al.
Veröffentlicht: (2024)
Can NLP Tackle Hate Speech in the Real World? Stakeholder-Informed Feedback and Survey on Counterspeech
von: Dinkar, Tanvi, et al.
Veröffentlicht: (2025)
von: Dinkar, Tanvi, et al.
Veröffentlicht: (2025)
Reasoning or Not? A Comprehensive Evaluation of Reasoning LLMs for Dialogue Summarization
von: Jin, Keyan, et al.
Veröffentlicht: (2025)
von: Jin, Keyan, et al.
Veröffentlicht: (2025)
Enhanced LLM Reasoning by Optimizing Reward Functions with Search-Driven Reinforcement Learning
von: Ahmadi, Arash, et al.
Veröffentlicht: (2026)
von: Ahmadi, Arash, et al.
Veröffentlicht: (2026)
CSCBench: A PVC Diagnostic Benchmark for Commodity Supply Chain Reasoning
von: Cui, Yaxin, et al.
Veröffentlicht: (2026)
von: Cui, Yaxin, et al.
Veröffentlicht: (2026)
Reasoning about Affordances: Causal and Compositional Reasoning in LLMs
von: Gjerde, Magnus F., et al.
Veröffentlicht: (2025)
von: Gjerde, Magnus F., et al.
Veröffentlicht: (2025)
Untangling Input Language from Reasoning Language: A Diagnostic Framework for Cross-Lingual Moral Alignment in LLMs
von: Li, Nan, et al.
Veröffentlicht: (2026)
von: Li, Nan, et al.
Veröffentlicht: (2026)
Understanding and Patching Compositional Reasoning in LLMs
von: Li, Zhaoyi, et al.
Veröffentlicht: (2024)
von: Li, Zhaoyi, et al.
Veröffentlicht: (2024)
Can LLMs Reason in the Wild with Programs?
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
Speech LLMs are Contextual Reasoning Transcribers
von: Deng, Keqi, et al.
Veröffentlicht: (2026)
von: Deng, Keqi, et al.
Veröffentlicht: (2026)
Can LLMs Reason About Trust?: A Pilot Study
von: Debnath, Anushka, et al.
Veröffentlicht: (2025)
von: Debnath, Anushka, et al.
Veröffentlicht: (2025)
The Multi-Round Diagnostic RAG Framework for Emulating Clinical Reasoning
von: Sun, Penglei, et al.
Veröffentlicht: (2025)
von: Sun, Penglei, et al.
Veröffentlicht: (2025)
Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning
von: Wang, Zezhong, et al.
Veröffentlicht: (2025)
von: Wang, Zezhong, et al.
Veröffentlicht: (2025)
ALFA: Aligning LLMs to Ask Good Questions A Case Study in Clinical Reasoning
von: Li, Shuyue Stella, et al.
Veröffentlicht: (2025)
von: Li, Shuyue Stella, et al.
Veröffentlicht: (2025)
GDS Agent for Graph Algorithmic Reasoning
von: Shi, Borun, et al.
Veröffentlicht: (2025)
von: Shi, Borun, et al.
Veröffentlicht: (2025)
Can LLMs Learn by Teaching for Better Reasoning? A Preliminary Study
von: Ning, Xuefei, et al.
Veröffentlicht: (2024)
von: Ning, Xuefei, et al.
Veröffentlicht: (2024)
Tokenization Constraints in LLMs: A Study of Symbolic and Arithmetic Reasoning Limits
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
Reasoning Like a Doctor: Improving Medical Dialogue Systems via Diagnostic Reasoning Process Alignment
von: Xu, Kaishuai, et al.
Veröffentlicht: (2024)
von: Xu, Kaishuai, et al.
Veröffentlicht: (2024)
Slot Filling as a Reasoning Task for SpeechLLMs
von: Hacioglu, Kadri, et al.
Veröffentlicht: (2025)
von: Hacioglu, Kadri, et al.
Veröffentlicht: (2025)
Reasoning Gets Harder for LLMs Inside A Dialogue
von: Kartáč, Ivan, et al.
Veröffentlicht: (2026)
von: Kartáč, Ivan, et al.
Veröffentlicht: (2026)
Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples
von: Yu, Fangxu, et al.
Veröffentlicht: (2024)
von: Yu, Fangxu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
When Chain-of-Thought Fails, the Solution Hides in the Hidden States
von: Mehrafarin, Houman, et al.
Veröffentlicht: (2026) -
AlanaVLM: A Multimodal Embodied AI Foundation Model for Egocentric Video Understanding
von: Suglia, Alessandro, et al.
Veröffentlicht: (2024) -
Voices in a Crowd: Searching for Clusters of Unique Perspectives
von: Vitsakis, Nikolas, et al.
Veröffentlicht: (2024) -
Repairs in a Block World: A New Benchmark for Handling User Corrections with Multi-Modal Language Models
von: Chiyah-Garcia, Javier, et al.
Veröffentlicht: (2024) -
MoRFI: Monotonic Sparse Autoencoder Feature Identification
von: Dimakopoulos, Dimitris, et al.
Veröffentlicht: (2026)