Warm Up Before You Train: Unlocking General Reasoning in Resource-Constrained Settings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shrestha, Safal, Kim, Minwu, Nepal, Aadim, Shrestha, Anubhav, Ross, Keith |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reinforcement Learning vs. Distillation: Understanding Accuracy and Capability in LLM Reasoning
von: Kim, Minwu, et al.
Veröffentlicht: (2025)
von: Kim, Minwu, et al.
Veröffentlicht: (2025)
On the Limits of Layer Pruning for Generative Reasoning in Large Language Models
von: Shrestha, Safal, et al.
Veröffentlicht: (2026)
von: Shrestha, Safal, et al.
Veröffentlicht: (2026)
Training Reasoning Models on Saturated Problems via Failure-Prefix Conditioning
von: Kim, Minwu, et al.
Veröffentlicht: (2026)
von: Kim, Minwu, et al.
Veröffentlicht: (2026)
Layer Importance for Mathematical Reasoning is Forged in Pre-Training and Invariant after Post-Training
von: Nepal, Aadim, et al.
Veröffentlicht: (2025)
von: Nepal, Aadim, et al.
Veröffentlicht: (2025)
Mathematical Reasoning in Large Language Models: Assessing Logical and Arithmetic Errors across Wide Numerical Ranges
von: Shrestha, Safal, et al.
Veröffentlicht: (2025)
von: Shrestha, Safal, et al.
Veröffentlicht: (2025)
Nwāchā Munā: A Devanagari Speech Corpus and Proximal Transfer Benchmark for Nepal Bhasha ASR
von: Sharma, Rishikesh Kumar, et al.
Veröffentlicht: (2026)
von: Sharma, Rishikesh Kumar, et al.
Veröffentlicht: (2026)
Efficient Multi-Hop Question Answering over Knowledge Graphs via LLM Planning and Embedding-Guided Search
von: Shrestha, Manil, et al.
Veröffentlicht: (2025)
von: Shrestha, Manil, et al.
Veröffentlicht: (2025)
Conformal Prediction for Risk-Controlled Medical Entity Extraction Across Clinical Domains
von: Shrestha, Manil, et al.
Veröffentlicht: (2026)
von: Shrestha, Manil, et al.
Veröffentlicht: (2026)
Towards Nepali-language LLMs: Efficient GPT training with a Nepali BPE tokenizer
von: Shrestha, Adarsha, et al.
Veröffentlicht: (2025)
von: Shrestha, Adarsha, et al.
Veröffentlicht: (2025)
Scaling Up RL: Unlocking Diverse Reasoning in LLMs via Prolonged Training
von: Liu, Mingjie, et al.
Veröffentlicht: (2025)
von: Liu, Mingjie, et al.
Veröffentlicht: (2025)
Visual Grounding Methods for VQA are Working for the Wrong Reasons!
von: Shrestha, Robik, et al.
Veröffentlicht: (2020)
von: Shrestha, Robik, et al.
Veröffentlicht: (2020)
Stop Before You Fail: Operational Capability Boundaries for Mitigating Unproductive Reasoning in Large Reasoning Models
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
von: Zhang, Qingjie, et al.
Veröffentlicht: (2025)
Think Before You Lie: How Reasoning Leads to Honesty
von: Yuan, Ann, et al.
Veröffentlicht: (2026)
von: Yuan, Ann, et al.
Veröffentlicht: (2026)
SnapKV: LLM Knows What You are Looking for Before Generation
von: Li, Yuhong, et al.
Veröffentlicht: (2024)
von: Li, Yuhong, et al.
Veröffentlicht: (2024)
Think Twice Before You Write -- an Entropy-based Decoding Strategy to Enhance LLM Reasoning
von: He, Jiashu, et al.
Veröffentlicht: (2026)
von: He, Jiashu, et al.
Veröffentlicht: (2026)
Verify Before You Commit: Towards Faithful Reasoning in LLM Agents via Self-Auditing
von: Yuan, Wenhao, et al.
Veröffentlicht: (2026)
von: Yuan, Wenhao, et al.
Veröffentlicht: (2026)
Look Before You Leap: Problem Elaboration Prompting Improves Mathematical Reasoning in Large Language Models
von: Liao, Haoran, et al.
Veröffentlicht: (2024)
von: Liao, Haoran, et al.
Veröffentlicht: (2024)
ALIGN: Word Association Learning for Cultural Alignment in Large Language Models
von: Liu, Chunhua, et al.
Veröffentlicht: (2025)
von: Liu, Chunhua, et al.
Veröffentlicht: (2025)
Difficulty Estimation and Simplification of French Text Using LLMs
von: Jamet, Henri, et al.
Veröffentlicht: (2024)
von: Jamet, Henri, et al.
Veröffentlicht: (2024)
CALM Before the STORM: Unlocking Native Reasoning for Optimization Modeling
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
von: Tang, Zhengyang, et al.
Veröffentlicht: (2025)
Evaluation of Language Models in the Medical Context Under Resource-Constrained Settings
von: Posada, Andrea, et al.
Veröffentlicht: (2024)
von: Posada, Andrea, et al.
Veröffentlicht: (2024)
Look Before You Leap: Autonomous Exploration for LLM Agents
von: Ye, Ziang, et al.
Veröffentlicht: (2026)
von: Ye, Ziang, et al.
Veröffentlicht: (2026)
A Survey on LLM-Assisted Clinical Trial Recruitment
von: Ghosh, Shrestha, et al.
Veröffentlicht: (2025)
von: Ghosh, Shrestha, et al.
Veröffentlicht: (2025)
Think Before You Prune: Self-Reflective Structured Pruning for Reasoning Language Models
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
Watch Before You Answer: Learning from Visually Grounded Post-Training
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2026)
von: Zhang, Yuxuan, et al.
Veröffentlicht: (2026)
NepaliGPT: A Generative Language Model for the Nepali Language
von: Pudasaini, Shushanta, et al.
Veröffentlicht: (2025)
von: Pudasaini, Shushanta, et al.
Veröffentlicht: (2025)
Token-Driven GammaTune: Adaptive Calibration for Enhanced Speculative Decoding
von: Gautam, Aayush, et al.
Veröffentlicht: (2025)
von: Gautam, Aayush, et al.
Veröffentlicht: (2025)
Summarize Before You Speak with ARACH: A Training-Free Inference-Time Plug-In for Enhancing LLMs via Global Attention Reallocation
von: Wang, Jingtao, et al.
Veröffentlicht: (2026)
von: Wang, Jingtao, et al.
Veröffentlicht: (2026)
Unlocking Anticipatory Text Generation: A Constrained Approach for Large Language Models Decoding
von: Tu, Lifu, et al.
Veröffentlicht: (2023)
von: Tu, Lifu, et al.
Veröffentlicht: (2023)
Before You Interpret the Profile: Validity Scaling for LLM Metacognitive Self-Report
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
von: Cacioli, Jon-Paul
Veröffentlicht: (2026)
Thinking Before Constraining: A Unified Decoding Framework for Large Language Models
von: Nguyen, Ngoc Trinh Hung, et al.
Veröffentlicht: (2026)
von: Nguyen, Ngoc Trinh Hung, et al.
Veröffentlicht: (2026)
Read Before You Think: Mitigating LLM Comprehension Failures with Step-by-Step Reading
von: Han, Feijiang, et al.
Veröffentlicht: (2025)
von: Han, Feijiang, et al.
Veröffentlicht: (2025)
Do Before You Judge: Self-Reference as a Pathway to Better LLM Evaluation
von: Lin, Wei-Hsiang, et al.
Veröffentlicht: (2025)
von: Lin, Wei-Hsiang, et al.
Veröffentlicht: (2025)
Think Thrice Before You Act: Progressive Thought Refinement in Large Language Models
von: Du, Chengyu, et al.
Veröffentlicht: (2024)
von: Du, Chengyu, et al.
Veröffentlicht: (2024)
Enabling LLM Knowledge Analysis via Extensive Materialization
von: Hu, Yujia, et al.
Veröffentlicht: (2024)
von: Hu, Yujia, et al.
Veröffentlicht: (2024)
Mining the Mind: What 100M Beliefs Reveal About Frontier LLM Knowledge
von: Ghosh, Shrestha, et al.
Veröffentlicht: (2025)
von: Ghosh, Shrestha, et al.
Veröffentlicht: (2025)
Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM
von: Zhang, Shaoqing, et al.
Veröffentlicht: (2024)
von: Zhang, Shaoqing, et al.
Veröffentlicht: (2024)
Do LLMs Need Inherent Reasoning Before Reinforcement Learning? A Study in Korean Self-Correction
von: Kim, Hongjin, et al.
Veröffentlicht: (2026)
von: Kim, Hongjin, et al.
Veröffentlicht: (2026)
Think Before You Act: Decision Transformers with Working Memory
von: Kang, Jikun, et al.
Veröffentlicht: (2023)
von: Kang, Jikun, et al.
Veröffentlicht: (2023)
Unlocking the Potential of Model Merging for Low-Resource Languages
von: Tao, Mingxu, et al.
Veröffentlicht: (2024)
von: Tao, Mingxu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Reinforcement Learning vs. Distillation: Understanding Accuracy and Capability in LLM Reasoning
von: Kim, Minwu, et al.
Veröffentlicht: (2025) -
On the Limits of Layer Pruning for Generative Reasoning in Large Language Models
von: Shrestha, Safal, et al.
Veröffentlicht: (2026) -
Training Reasoning Models on Saturated Problems via Failure-Prefix Conditioning
von: Kim, Minwu, et al.
Veröffentlicht: (2026) -
Layer Importance for Mathematical Reasoning is Forged in Pre-Training and Invariant after Post-Training
von: Nepal, Aadim, et al.
Veröffentlicht: (2025) -
Mathematical Reasoning in Large Language Models: Assessing Logical and Arithmetic Errors across Wide Numerical Ranges
von: Shrestha, Safal, et al.
Veröffentlicht: (2025)