D-COT: Disciplined Chain-of-Thought Learning for Efficient Reasoning in Small Language Models
Fuente:
arXiv
Salvato in:
| Autore principale: | Ubukata, Shunsuke |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Why Models Know But Don't Say: Chain-of-Thought Faithfulness Divergence Between Thinking Tokens and Answers in Open-Weight Reasoning Models
di: Young, Richard J.
Pubblicazione: (2026)
di: Young, Richard J.
Pubblicazione: (2026)
TRiMS: Real-Time Tracking of Minimal Sufficient Length for Efficient Reasoning via RL
di: Bian, Tingcheng, et al.
Pubblicazione: (2026)
di: Bian, Tingcheng, et al.
Pubblicazione: (2026)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025)
di: Fadli, Samih
Pubblicazione: (2025)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
di: Han, Xudong, et al.
Pubblicazione: (2025)
di: Han, Xudong, et al.
Pubblicazione: (2025)
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
di: Kugler, Kai
Pubblicazione: (2025)
di: Kugler, Kai
Pubblicazione: (2025)
MCP: A Control-Theoretic Orchestration Framework for Synergistic Efficiency and Interpretability in Multimodal Large Language Models
di: Zhang, Luyan
Pubblicazione: (2025)
di: Zhang, Luyan
Pubblicazione: (2025)
Induce, Align, Predict: Zero-Shot Stance Detection via Cognitive Inductive Reasoning
di: Zhang, Bowen, et al.
Pubblicazione: (2025)
di: Zhang, Bowen, et al.
Pubblicazione: (2025)
Mixup Model Merge: Enhancing Model Merging Performance through Randomized Linear Interpolation
di: Zhou, Yue, et al.
Pubblicazione: (2025)
di: Zhou, Yue, et al.
Pubblicazione: (2025)
Bridging the Gap: An Intermediate Language for Enhanced and Cost-Effective Grapheme-to-Phoneme Conversion with Homographs with Multiple Pronunciations Disambiguation
di: Bertina, Abbas, et al.
Pubblicazione: (2025)
di: Bertina, Abbas, et al.
Pubblicazione: (2025)
Intention Collapse: Intention-Level Metrics for Reasoning in Language Models
di: Vera, Patricio
Pubblicazione: (2026)
di: Vera, Patricio
Pubblicazione: (2026)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
di: Liu, Zhongxin, et al.
Pubblicazione: (2025)
di: Liu, Zhongxin, et al.
Pubblicazione: (2025)
EmoLoom-2B: Fast Base-Model Screening for Emotion Classification and VAD with Lexicon-Weak Supervision and KV-Off Evaluation
di: Li, Zilin, et al.
Pubblicazione: (2026)
di: Li, Zilin, et al.
Pubblicazione: (2026)
Assessing Large Language Models on Islamic Legal Reasoning: Evidence from Inheritance Law Evaluation
di: Bouchekif, Abdessalam, et al.
Pubblicazione: (2025)
di: Bouchekif, Abdessalam, et al.
Pubblicazione: (2025)
Sample-Efficient Language Model for Hinglish Conversational AI
di: Singh, Sakshi, et al.
Pubblicazione: (2025)
di: Singh, Sakshi, et al.
Pubblicazione: (2025)
Kronecker Embeddings: Byte-Level Structured Token Representations for Parameter-Efficient Language Models
di: Shravan, Rohan
Pubblicazione: (2026)
di: Shravan, Rohan
Pubblicazione: (2026)
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
di: Yang, Shan
Pubblicazione: (2026)
di: Yang, Shan
Pubblicazione: (2026)
Calibrated Confidence Estimation for Tabular Question Answering
di: Voss, Lukas
Pubblicazione: (2026)
di: Voss, Lukas
Pubblicazione: (2026)
Character-Level Transformer for Tajik-Persian Transliteration with a Parallel Lexical Corpus
di: Arabov, Mullosharaf K.
Pubblicazione: (2026)
di: Arabov, Mullosharaf K.
Pubblicazione: (2026)
Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation
di: Resck, Lucas, et al.
Pubblicazione: (2026)
di: Resck, Lucas, et al.
Pubblicazione: (2026)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
di: Gwak, Jiho, et al.
Pubblicazione: (2025)
di: Gwak, Jiho, et al.
Pubblicazione: (2025)
On the Influence of Discourse Relations in Persuasive Texts
di: Turk, Nawar, et al.
Pubblicazione: (2025)
di: Turk, Nawar, et al.
Pubblicazione: (2025)
ImmigrationQA: A Source-Grounded Dataset and Small-Model Adaptation for U.S. Immigration Law
di: Shportun, Nazarii
Pubblicazione: (2026)
di: Shportun, Nazarii
Pubblicazione: (2026)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
di: Reddy, Sandeep, et al.
Pubblicazione: (2025)
di: Reddy, Sandeep, et al.
Pubblicazione: (2025)
Language Models Are Implicitly Continuous
di: Marro, Samuele, et al.
Pubblicazione: (2025)
di: Marro, Samuele, et al.
Pubblicazione: (2025)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
di: Hashemi, Helia, et al.
Pubblicazione: (2024)
di: Hashemi, Helia, et al.
Pubblicazione: (2024)
Your Pretrained Model Tells the Difficulty Itself: A Self-Adaptive Curriculum Learning Paradigm for Natural Language Understanding
di: Feng, Qi, et al.
Pubblicazione: (2025)
di: Feng, Qi, et al.
Pubblicazione: (2025)
The Last Word Often Wins: A Format Confound in Chain-of-Thought Corruption Studies
di: Garcia, Gabriel
Pubblicazione: (2026)
di: Garcia, Gabriel
Pubblicazione: (2026)
Machine Unlearning for Masked Diffusion Language Models
di: Lee, Georu, et al.
Pubblicazione: (2026)
di: Lee, Georu, et al.
Pubblicazione: (2026)
Truth as a Compression Artifact in Language Model Training
di: Krestnikov, Konstantin
Pubblicazione: (2026)
di: Krestnikov, Konstantin
Pubblicazione: (2026)
Combining Language and Topic Models for Hierarchical Text Classification
di: Toit, Jaco du, et al.
Pubblicazione: (2025)
di: Toit, Jaco du, et al.
Pubblicazione: (2025)
Learning the meanings of function words from grounded language using a visual question answering model
di: Portelance, Eva, et al.
Pubblicazione: (2023)
di: Portelance, Eva, et al.
Pubblicazione: (2023)
Memory Bank Compression for Continual Adaptation of Large Language Models
di: Katraouras, Thomas, et al.
Pubblicazione: (2026)
di: Katraouras, Thomas, et al.
Pubblicazione: (2026)
Statistical Scouting Finds Debate-Safe but Not Debate-Useful Cases: A Matched-Ceiling Study of Open-Weight LLM Reasoning Protocols
di: Hu, Julia, et al.
Pubblicazione: (2026)
di: Hu, Julia, et al.
Pubblicazione: (2026)
Rethinking Addressing in Language Models via Contexualized Equivariant Positional Encoding
di: Zhu, Jiajun, et al.
Pubblicazione: (2025)
di: Zhu, Jiajun, et al.
Pubblicazione: (2025)
Towards Resource-Efficient Multimodal Intelligence: Learned Routing among Specialized Expert Models
di: Saini, Mayank, et al.
Pubblicazione: (2025)
di: Saini, Mayank, et al.
Pubblicazione: (2025)
How Language Models Process Out-of-Distribution Inputs: A Two-Pathway Framework
di: Saghir, Hamidreza
Pubblicazione: (2026)
di: Saghir, Hamidreza
Pubblicazione: (2026)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
di: Nieth, Björn, et al.
Pubblicazione: (2026)
di: Nieth, Björn, et al.
Pubblicazione: (2026)
Fine-tuning of Large Language Models for Constituency Parsing Using a Sequence to Sequence Approach
di: Delgado, Francisco Jose Cortes, et al.
Pubblicazione: (2025)
di: Delgado, Francisco Jose Cortes, et al.
Pubblicazione: (2025)
TIME: Temporally Intelligent Meta-reasoning Engine for Context-Triggered Explicit Reasoning
di: Das, Susmit
Pubblicazione: (2026)
di: Das, Susmit
Pubblicazione: (2026)
Neural Activation Patterns Across Language Model Architectures: A Comprehensive Analysis of Cognitive Task Performance
di: Naser-Moghadasi, Mahdi, et al.
Pubblicazione: (2026)
di: Naser-Moghadasi, Mahdi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Why Models Know But Don't Say: Chain-of-Thought Faithfulness Divergence Between Thinking Tokens and Answers in Open-Weight Reasoning Models
di: Young, Richard J.
Pubblicazione: (2026) -
TRiMS: Real-Time Tracking of Minimal Sufficient Length for Efficient Reasoning via RL
di: Bian, Tingcheng, et al.
Pubblicazione: (2026) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025) -
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
di: Han, Xudong, et al.
Pubblicazione: (2025) -
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
di: Kugler, Kai
Pubblicazione: (2025)