Learning From Correctness Without Prompting Makes LLM Efficient Reasoner
Fuente:
arXiv
Saved in:
| Main Authors: | Yao, Yuxuan, Wu, Han, Guo, Zhijiang, Zhou, Biyan, Gao, Jiahui, Luo, Sichun, Hou, Hanxu, Fu, Xiaojin, Song, Linqi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reasoning Meets Personalization: Unleashing the Potential of Large Reasoning Model for Personalized Generation
by: Luo, Sichun, et al.
Published: (2025)
by: Luo, Sichun, et al.
Published: (2025)
Determine-Then-Ensemble: Necessity of Top-k Union for Large Language Model Ensembling
by: Yao, Yuxuan, et al.
Published: (2024)
by: Yao, Yuxuan, et al.
Published: (2024)
RALLRec+: Retrieval Augmented Large Language Model Recommendation with Reasoning
by: Luo, Sichun, et al.
Published: (2025)
by: Luo, Sichun, et al.
Published: (2025)
RALLRec: Improving Retrieval Augmented Large Language Model Recommendation with Representation Learning
by: Xu, Jian, et al.
Published: (2025)
by: Xu, Jian, et al.
Published: (2025)
Can LLM Substitute Human Labeling? A Case Study of Fine-grained Chinese Address Entity Recognition Dataset for UAV Delivery
by: Yao, Yuxuan, et al.
Published: (2024)
by: Yao, Yuxuan, et al.
Published: (2024)
Privacy in LLM-based Recommendation: Recent Advances and Future Directions
by: Luo, Sichun, et al.
Published: (2024)
by: Luo, Sichun, et al.
Published: (2024)
Merging Beyond: Streaming LLM Updates via Activation-Guided Rotations
by: Yao, Yuxuan, et al.
Published: (2026)
by: Yao, Yuxuan, et al.
Published: (2026)
Integrating Large Language Models into Recommendation via Mutual Augmentation and Adaptive Aggregation
by: Luo, Sichun, et al.
Published: (2024)
by: Luo, Sichun, et al.
Published: (2024)
DeReason: A Difficulty-Aware Curriculum Improves Decoupled SFT-then-RL Training for General Reasoning
by: Hu, Hanxu, et al.
Published: (2026)
by: Hu, Hanxu, et al.
Published: (2026)
Unlocking Efficient Long-to-Short LLM Reasoning with Model Merging
by: Wu, Han, et al.
Published: (2025)
by: Wu, Han, et al.
Published: (2025)
Towards Quantum-Safe Federated Learning via Homomorphic Encryption: Learning with Gradients
by: Yan, Guangfeng, et al.
Published: (2024)
by: Yan, Guangfeng, et al.
Published: (2024)
Improved Quantization Strategies for Managing Heavy-tailed Gradients in Distributed Learning
by: Yan, Guangfeng, et al.
Published: (2024)
by: Yan, Guangfeng, et al.
Published: (2024)
Accelerating Data Access for Single Node in Distributed Storage Systems via MDS Codes
by: Shi, Hao, et al.
Published: (2025)
by: Shi, Hao, et al.
Published: (2025)
Reed-Solomon Codes over Cyclic Polynomial Ring with Lower Encoding/Decoding Complexity
by: Liu, Wenhao, et al.
Published: (2024)
by: Liu, Wenhao, et al.
Published: (2024)
Error Correction Decoding Algorithms of RS Codes Based on An Earlier Termination Algorithm to Find The Error Locator Polynomial
by: Jiang, Zhengyi, et al.
Published: (2024)
by: Jiang, Zhengyi, et al.
Published: (2024)
Activation-Guided Consensus Merging for Large Language Models
by: Yao, Yuxuan, et al.
Published: (2025)
by: Yao, Yuxuan, et al.
Published: (2025)
Fine-grained Conversational Decoding via Isotropic and Proximal Search
by: Yao, Yuxuan, et al.
Published: (2023)
by: Yao, Yuxuan, et al.
Published: (2023)
CodeScaler: Scaling Code LLM Training and Test-Time Inference via Reward Models
by: Zhu, Xiao, et al.
Published: (2026)
by: Zhu, Xiao, et al.
Published: (2026)
Learning to Reason with Insight for Informal Theorem Proving
by: Li, Yunhe, et al.
Published: (2026)
by: Li, Yunhe, et al.
Published: (2026)
Redundancy-Optimal Constructions of $(1,1)$-Criss-Cross Deletion Correcting Codes with Efficient Encoding/Decoding Algorithms
by: Liu, Wenhao, et al.
Published: (2026)
by: Liu, Wenhao, et al.
Published: (2026)
Routing-Aligned Fine-Tuning for Multilingual Downstream Tasks in Mixture-of-Experts Models
by: Deng, Guanzhi, et al.
Published: (2026)
by: Deng, Guanzhi, et al.
Published: (2026)
Dezentrale autonome Organisationen (DAOs) und Gesellschaftsrecht
by: Mienert, Biyan
Published: (2024)
by: Mienert, Biyan
Published: (2024)
Chain-of-Thought Reasoning Without Prompting
by: Wang, Xuezhi, et al.
Published: (2024)
by: Wang, Xuezhi, et al.
Published: (2024)
An energy-efficient spiking neural network with continuous learning for self-adaptive brain-machine interface
by: Biyan, Zhou, et al.
Published: (2025)
by: Biyan, Zhou, et al.
Published: (2025)
RecRanker: Instruction Tuning Large Language Model as Ranker for Top-k Recommendation
by: Luo, Sichun, et al.
Published: (2023)
by: Luo, Sichun, et al.
Published: (2023)
Combining SNNs with Filtering for Efficient Neural Decoding in Implantable Brain-Machine Interfaces
by: Zhou, Biyan, et al.
Published: (2023)
by: Zhou, Biyan, et al.
Published: (2023)
Accordion-Thinking: Self-Regulated Step Summaries for Efficient and Readable LLM Reasoning
by: Yang, Zhicheng, et al.
Published: (2026)
by: Yang, Zhicheng, et al.
Published: (2026)
When Corrective Hints Hurt: Prompt Design in Reasoner-Guided Repair of LLM Overcaution on Entailed Negations under OWL~2~DL
by: Qi, Yijiashun, et al.
Published: (2026)
by: Qi, Yijiashun, et al.
Published: (2026)
Grid-like Error-Correcting Codes for Matrix Multiplication with Better Correcting Capability
by: Shi, Hao, et al.
Published: (2025)
by: Shi, Hao, et al.
Published: (2025)
Quillen-Suslin Theorem for connected cochain DG algebras
by: Mao, Xuefeng, et al.
Published: (2025)
by: Mao, Xuefeng, et al.
Published: (2025)
Integrative Analyses Reveal Sigmohymena gen. nov. and the Ninth Pattern of Undulating Membranes in Hypotrichia (Protista, Ciliophora): Insights From Polygenic Analysis on Undulating Membranes Evolution
by: Qi Gao, et al.
Published: (2025)
by: Qi Gao, et al.
Published: (2025)
Efficient enzymatic synthesis of theaflavin and its production mechanism
by: Jinjin Jian, et al.
Published: (2024)
by: Jinjin Jian, et al.
Published: (2024)
Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning
by: Yang, Zhicheng, et al.
Published: (2026)
by: Yang, Zhicheng, et al.
Published: (2026)
Are LLMs Rigorous Logical Reasoners? Empowering Natural Language Proof Generation by Stepwise Decoding with Contrastive Learning
by: Su, Ying, et al.
Published: (2023)
by: Su, Ying, et al.
Published: (2023)
Bi-Chainer: Automated Large Language Models Reasoning with Bidirectional Chaining
by: Liu, Shuqi, et al.
Published: (2024)
by: Liu, Shuqi, et al.
Published: (2024)
What Makes a Good Reasoning Chain? Uncovering Structural Patterns in Long Chain-of-Thought Reasoning
by: Jiang, Gangwei, et al.
Published: (2025)
by: Jiang, Gangwei, et al.
Published: (2025)
Large Language Models are Good Multi-lingual Learners : When LLMs Meet Cross-lingual Prompts
by: Wang, Teng, et al.
Published: (2024)
by: Wang, Teng, et al.
Published: (2024)
Silent Egress: When Implicit Prompt Injection Makes LLM Agents Leak Without a Trace
by: Lan, Qianlong, et al.
Published: (2026)
by: Lan, Qianlong, et al.
Published: (2026)
Meta-Prompt Optimization for LLM-Based Sequential Decision Making
by: Kong, Mingze, et al.
Published: (2025)
by: Kong, Mingze, et al.
Published: (2025)
1bit-Merging: Dynamic Quantized Merging for Large Language Models
by: Liu, Shuqi, et al.
Published: (2025)
by: Liu, Shuqi, et al.
Published: (2025)
Similar Items
-
Reasoning Meets Personalization: Unleashing the Potential of Large Reasoning Model for Personalized Generation
by: Luo, Sichun, et al.
Published: (2025) -
Determine-Then-Ensemble: Necessity of Top-k Union for Large Language Model Ensembling
by: Yao, Yuxuan, et al.
Published: (2024) -
RALLRec+: Retrieval Augmented Large Language Model Recommendation with Reasoning
by: Luo, Sichun, et al.
Published: (2025) -
RALLRec: Improving Retrieval Augmented Large Language Model Recommendation with Representation Learning
by: Xu, Jian, et al.
Published: (2025) -
Can LLM Substitute Human Labeling? A Case Study of Fine-grained Chinese Address Entity Recognition Dataset for UAV Delivery
by: Yao, Yuxuan, et al.
Published: (2024)