Reasoning Primitives in Hybrid and Non-Hybrid LLMs: Do Architectural Differences Yield Advantages in State-Tracking and Recall?
Fuente:
arXiv
Saved in:
| Main Authors: | Rawat, Shivam, Flek, Lucie, Mai, Florian, Corrêa, Nicholas Kluge |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Raising Bars, Not Parameters: LilMoo Compact Language Model for Hindi
by: Fatimah, Shiza, et al.
Published: (2026)
by: Fatimah, Shiza, et al.
Published: (2026)
Superalignment with Dynamic Human Values
by: Mai, Florian, et al.
Published: (2025)
by: Mai, Florian, et al.
Published: (2025)
Plausible but Wrong: A case study on Agentic Failures in Astrophysical Workflows
by: Rawat, Shivam, et al.
Published: (2026)
by: Rawat, Shivam, et al.
Published: (2026)
Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models
by: Nickel, Christian, et al.
Published: (2026)
by: Nickel, Christian, et al.
Published: (2026)
IKnow: Instruction-Knowledge-Aware Continual Pretraining for Effective Domain Adaptation
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
Tucano 2 Cool: Better Open Source LLMs for Portuguese
by: Corrêa, Nicholas Kluge, et al.
Published: (2026)
by: Corrêa, Nicholas Kluge, et al.
Published: (2026)
Multi-Hop Reasoning for Question Answering with Hyperbolic Representations
by: Welz, Simon, et al.
Published: (2025)
by: Welz, Simon, et al.
Published: (2025)
Encoder Fine-tuning with Stochastic Sampling Outperforms Open-weight GPT in Astronomy Knowledge Extraction
by: Rawat, Shivam, et al.
Published: (2025)
by: Rawat, Shivam, et al.
Published: (2025)
Pitfalls of Conversational LLMs on News Debiasing
by: Schlicht, Ipek Baris, et al.
Published: (2024)
by: Schlicht, Ipek Baris, et al.
Published: (2024)
Can Stories Help LLMs Reason? Curating Information Space Through Narrative
by: Javadi, Vahid Sadiri, et al.
Published: (2024)
by: Javadi, Vahid Sadiri, et al.
Published: (2024)
Probing the Robustness of Theory of Mind in Large Language Models
by: Nickel, Christian, et al.
Published: (2024)
by: Nickel, Christian, et al.
Published: (2024)
Do LLMs Provide Consistent Answers to Health-Related Questions across Languages?
by: Schlicht, Ipek Baris, et al.
Published: (2025)
by: Schlicht, Ipek Baris, et al.
Published: (2025)
More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persists
by: Alavi, Khashayar, et al.
Published: (2025)
by: Alavi, Khashayar, et al.
Published: (2025)
Tailored Conversations beyond LLMs: A RL-Based Dialogue Manager
by: Galland, Lucie, et al.
Published: (2025)
by: Galland, Lucie, et al.
Published: (2025)
On the Limitations of Language Targeted Pruning: Investigating the Calibration Language Impact in Multilingual LLM Pruning
by: Kurz, Simon, et al.
Published: (2024)
by: Kurz, Simon, et al.
Published: (2024)
From Human Cognition to Neural Activations: Probing the Computational Primitives of Spatial Reasoning in LLMs
by: An, Jiyuan, et al.
Published: (2026)
by: An, Jiyuan, et al.
Published: (2026)
Understanding In-Context Learning Beyond Transformers: An Investigation of State Space and Hybrid Architectures
by: Wang, Shenran, et al.
Published: (2025)
by: Wang, Shenran, et al.
Published: (2025)
Hybrid Policy Distillation for LLMs
by: Zhu, Wenhong, et al.
Published: (2026)
by: Zhu, Wenhong, et al.
Published: (2026)
Hybrid Dialogue State Tracking for Persian Chatbots: A Language Model-Based Approach
by: Aghabagher, Samin Mahdipour, et al.
Published: (2025)
by: Aghabagher, Samin Mahdipour, et al.
Published: (2025)
Every Attention Matters: An Efficient Hybrid Architecture for Long-Context Reasoning
by: Ling Team, et al.
Published: (2025)
by: Ling Team, et al.
Published: (2025)
Tucano: Advancing Neural Text Generation for Portuguese
by: Corrêa, Nicholas Kluge, et al.
Published: (2024)
by: Corrêa, Nicholas Kluge, et al.
Published: (2024)
Do LLMs Really Adapt to Domains? An Ontology Learning Perspective
by: Mai, Huu Tan, et al.
Published: (2024)
by: Mai, Huu Tan, et al.
Published: (2024)
Recall, Retrieve and Reason: Towards Better In-Context Relation Extraction
by: Li, Guozheng, et al.
Published: (2024)
by: Li, Guozheng, et al.
Published: (2024)
USDC: A Dataset of $\underline{U}$ser $\underline{S}$tance and $\underline{D}$ogmatism in Long $\underline{C}$onversations
by: Marreddy, Mounika, et al.
Published: (2024)
by: Marreddy, Mounika, et al.
Published: (2024)
Improving Low-Resource Dialect Classification Using Retrieval-based Voice Conversion
by: Fischbach, Lea, et al.
Published: (2025)
by: Fischbach, Lea, et al.
Published: (2025)
Layerwise Recall and the Geometry of Interwoven Knowledge in LLMs
by: Lei, Ge, et al.
Published: (2025)
by: Lei, Ge, et al.
Published: (2025)
StructRAG: Boosting Knowledge Intensive Reasoning of LLMs via Inference-time Hybrid Information Structurization
by: Li, Zhuoqun, et al.
Published: (2024)
by: Li, Zhuoqun, et al.
Published: (2024)
Dynamic Normativity: Necessary and Sufficient Conditions for Value Alignment
by: Corrêa, Nicholas Kluge
Published: (2024)
by: Corrêa, Nicholas Kluge
Published: (2024)
TiEBe: Tracking Language Model Recall of Notable Worldwide Events Through Time
by: Almeida, Thales Sales, et al.
Published: (2025)
by: Almeida, Thales Sales, et al.
Published: (2025)
Do Language Models Track Entities Across State Changes?
by: Tang, Zilu, et al.
Published: (2026)
by: Tang, Zilu, et al.
Published: (2026)
Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect
by: Presa, Joao Paulo Cavalcante, et al.
Published: (2026)
by: Presa, Joao Paulo Cavalcante, et al.
Published: (2026)
How Do LLMs Perform Two-Hop Reasoning in Context?
by: Guo, Tianyu, et al.
Published: (2025)
by: Guo, Tianyu, et al.
Published: (2025)
Do LLMs Really Think Step-by-step In Implicit Reasoning?
by: Yu, Yijiong
Published: (2024)
by: Yu, Yijiong
Published: (2024)
The Chameleon Nature of LLMs: Quantifying Multi-Turn Stance Instability in Search-Enabled Language Models
by: Ratnakar, Shivam, et al.
Published: (2025)
by: Ratnakar, Shivam, et al.
Published: (2025)
Arbitrage: Efficient Reasoning via Advantage-Aware Speculation
by: Maheswaran, Monishwaran, et al.
Published: (2025)
by: Maheswaran, Monishwaran, et al.
Published: (2025)
Tiny Recursive Reasoning with Mamba-2 Attention Hybrid
by: Wang, Wenlong, et al.
Published: (2026)
by: Wang, Wenlong, et al.
Published: (2026)
KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?
by: Saha, Soumadeep, et al.
Published: (2025)
by: Saha, Soumadeep, et al.
Published: (2025)
Blended RAG: Improving RAG (Retriever-Augmented Generation) Accuracy with Semantic Search and Hybrid Query-Based Retrievers
by: Sawarkar, Kunal, et al.
Published: (2024)
by: Sawarkar, Kunal, et al.
Published: (2024)
(How) Do Language Models Track State?
by: Li, Belinda Z., et al.
Published: (2025)
by: Li, Belinda Z., et al.
Published: (2025)
Slimming Down LLMs Without Losing Their Minds
by: Qingda, et al.
Published: (2025)
by: Qingda, et al.
Published: (2025)
Similar Items
-
Raising Bars, Not Parameters: LilMoo Compact Language Model for Hindi
by: Fatimah, Shiza, et al.
Published: (2026) -
Superalignment with Dynamic Human Values
by: Mai, Florian, et al.
Published: (2025) -
Plausible but Wrong: A case study on Agentic Failures in Astrophysical Workflows
by: Rawat, Shivam, et al.
Published: (2026) -
Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models
by: Nickel, Christian, et al.
Published: (2026) -
IKnow: Instruction-Knowledge-Aware Continual Pretraining for Effective Domain Adaptation
by: Zhang, Tianyi, et al.
Published: (2025)