K2-Think: A Parameter-Efficient Reasoning System
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cheng, Zhoujun, Fan, Richard, Hao, Shibo, Killian, Taylor W., Li, Haonan, Sun, Suqi, Ren, Hector, Moreno, Alexander, Zhang, Daqian, Zhong, Tianjun, Xiong, Yuxin, Hu, Yuanzhe, Xie, Yutao, Han, Xudong, Wang, Yuqi, Pimpalkhute, Varad, Zhuang, Yonghao, Singh, Aaryamonvikram, Liang, Xuezhi, Xie, Anze, She, Jianshu, Fan, Desai, Gao, Chengqian, Ma, Liqun, Yurochkin, Mikhail, Maggs, John, Ma, Xuezhe, He, Guowei, Hu, Zhiting, Liu, Zhengzhong, Xing, Eric P. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
IsoCompute Playbook: Optimally Scaling Sampling Compute for LLM RL
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2026)
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2026)
Concise Reasoning in the Lens of Lagrangian Optimization
von: Gao, Chengqian, et al.
Veröffentlicht: (2025)
von: Gao, Chengqian, et al.
Veröffentlicht: (2025)
K2-V2: A 360-Open, Reasoning-Enhanced LLM
von: K2 Team, et al.
Veröffentlicht: (2025)
von: K2 Team, et al.
Veröffentlicht: (2025)
Efficient Agentic Reasoning Through Self-Regulated Simulative Planning
von: Deng, Mingkai, et al.
Veröffentlicht: (2026)
von: Deng, Mingkai, et al.
Veröffentlicht: (2026)
Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2025)
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2025)
SoftQE: Learned Representations of Queries Expanded by LLMs
von: Pimpalkhute, Varad, et al.
Veröffentlicht: (2024)
von: Pimpalkhute, Varad, et al.
Veröffentlicht: (2024)
MegaMath: Pushing the Limits of Open Math Corpora
von: Zhou, Fan, et al.
Veröffentlicht: (2025)
von: Zhou, Fan, et al.
Veröffentlicht: (2025)
DISTFLASHATTN: Distributed Memory-efficient Attention for Long-context LLMs Training
von: Li, Dacheng, et al.
Veröffentlicht: (2023)
von: Li, Dacheng, et al.
Veröffentlicht: (2023)
Modeling Community Attitude through Reaction Tone: A Human-AI Collaborative Framework for Evaluating LLM Alignment with Linguistic Behaviors in Online Communities
von: Wen, Nuan, et al.
Veröffentlicht: (2026)
von: Wen, Nuan, et al.
Veröffentlicht: (2026)
LLM The Genius Paradox: A Linguistic and Math Expert's Struggle with Simple Word-based Counting Problems
von: Xu, Nan, et al.
Veröffentlicht: (2024)
von: Xu, Nan, et al.
Veröffentlicht: (2024)
DecoPrompt : Decoding Prompts Reduces Hallucinations when Large Language Models Meet False Premises
von: Xu, Nan, et al.
Veröffentlicht: (2024)
von: Xu, Nan, et al.
Veröffentlicht: (2024)
Bifurcation in a G0 Model of Hematological Stem Cells With Delay
von: Ma Suqi, et al.
Veröffentlicht: (2024)
von: Ma Suqi, et al.
Veröffentlicht: (2024)
Towards Chapter-to-Chapter Context-Aware Literary Translation via Large Language Models
von: Jin, Linghao, et al.
Veröffentlicht: (2024)
von: Jin, Linghao, et al.
Veröffentlicht: (2024)
FairMT-Bench: Benchmarking Fairness for Multi-turn Dialogue in Conversational LLMs
von: Fan, Zhiting, et al.
Veröffentlicht: (2024)
von: Fan, Zhiting, et al.
Veröffentlicht: (2024)
EMO: Frustratingly Easy Progressive Training of Extendable MoE
von: Jin, Linghao, et al.
Veröffentlicht: (2026)
von: Jin, Linghao, et al.
Veröffentlicht: (2026)
Prompt Recursive Search: A Living Framework with Adaptive Growth in LLM Auto-Prompting
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2024)
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2024)
PatentEdits: Framing Patent Novelty as Textual Entailment
von: Lee, Ryan, et al.
Veröffentlicht: (2024)
von: Lee, Ryan, et al.
Veröffentlicht: (2024)
Uniform accuracy of implicit-explicit Runge-Kutta methods for linear hyperbolic relaxation systems
von: Ma, Zhiting, et al.
Veröffentlicht: (2024)
von: Ma, Zhiting, et al.
Veröffentlicht: (2024)
Validity of relaxation models arising from numerical schemes for hyperbolic-parabolic systems
von: Ma, Zhiting, et al.
Veröffentlicht: (2025)
von: Ma, Zhiting, et al.
Veröffentlicht: (2025)
A synthetic review: natural history of amniote reproductive modes in light of comparative evolutionary genomics
von: Maggs X
Veröffentlicht: (2024)
von: Maggs X
Veröffentlicht: (2024)
On the track of unknown algae
von: Christine Maggs
Veröffentlicht: (2024)
von: Christine Maggs
Veröffentlicht: (2024)
On the Suitability of Reinforcement Fine-Tuning to Visual Tasks
von: Chen, Xiaxu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaxu, et al.
Veröffentlicht: (2025)
UP-CrackNet: Unsupervised Pixel-Wise Road Crack Detection via Adversarial Image Restoration
von: Ma, Nachuan, et al.
Veröffentlicht: (2024)
von: Ma, Nachuan, et al.
Veröffentlicht: (2024)
Dual-Phase Playtime-guided Recommendation: Interest Intensity Exploration and Multimodal Random Walks
von: Zhang, Jingmao, et al.
Veröffentlicht: (2025)
von: Zhang, Jingmao, et al.
Veröffentlicht: (2025)
Multi-Objective Recommendation in the Era of Generative AI: A Survey of Recent Progress and Future Prospects
von: Hong, Zihan, et al.
Veröffentlicht: (2025)
von: Hong, Zihan, et al.
Veröffentlicht: (2025)
Diversity Recommendation via Causal Deconfounding of Co-purchase Relations and Counterfactual Exposure
von: Zhang, Jingmao, et al.
Veröffentlicht: (2025)
von: Zhang, Jingmao, et al.
Veröffentlicht: (2025)
Harnessing Pyridinic N Vacancy Defect in Microporous Structures to Induce the Pre‐Adsorption of Oxygen and Boost Oxygen Reduction Reaction Kinetics
von: Binbin Jia, et al.
Veröffentlicht: (2025)
von: Binbin Jia, et al.
Veröffentlicht: (2025)
A Survey on Image Quality Assessment: Insights, Analysis, and Future Outlook
von: Ma, Chengqian, et al.
Veröffentlicht: (2025)
von: Ma, Chengqian, et al.
Veröffentlicht: (2025)
Spin‐Orbital Ordering Effects of Light Emission in Organic–Inorganic Hybrid Metal Halide Perovskites
von: Liqun Liu, et al.
Veröffentlicht: (2024)
von: Liqun Liu, et al.
Veröffentlicht: (2024)
LLM360 K2: Building a 65B 360-Open-Source Large Language Model from Scratch
von: Liu, Zhengzhong, et al.
Veröffentlicht: (2025)
von: Liu, Zhengzhong, et al.
Veröffentlicht: (2025)
C3: A Bilingual Benchmark for Spoken Dialogue Models Exploring Challenges in Complex Conversations
von: Ma, Chengqian, et al.
Veröffentlicht: (2025)
von: Ma, Chengqian, et al.
Veröffentlicht: (2025)
BiasGuard: A Reasoning-enhanced Bias Detection Tool For Large Language Models
von: Fan, Zhiting, et al.
Veröffentlicht: (2025)
von: Fan, Zhiting, et al.
Veröffentlicht: (2025)
Light-weight Fine-tuning Method for Defending Adversarial Noise in Pre-trained Medical Vision-Language Models
von: Han, Xu, et al.
Veröffentlicht: (2024)
von: Han, Xu, et al.
Veröffentlicht: (2024)
Learning Optimal and Sample-Efficient Decision Policies with Guarantees
von: Shao, Daqian
Veröffentlicht: (2026)
von: Shao, Daqian
Veröffentlicht: (2026)
Knowledge Graph Extension by Entity Type Recognition
von: Shi, Daqian
Veröffentlicht: (2024)
von: Shi, Daqian
Veröffentlicht: (2024)
Non-reversible Monte Carlo: an example of 'true' self-repelling motion
von: Maggs, A. C.
Veröffentlicht: (2023)
von: Maggs, A. C.
Veröffentlicht: (2023)
Event-chain Monte Carlo and the true self-avoiding walk
von: Maggs, A. C.
Veröffentlicht: (2024)
von: Maggs, A. C.
Veröffentlicht: (2024)
John Schulenberg as a developmental scholar and mentor: Personal reflections
von: Jennifer L. Maggs
Veröffentlicht: (2024)
von: Jennifer L. Maggs
Veröffentlicht: (2024)
Dynamics of a bricklayer model: multi-walker realizations of true self-avoiding motion
von: Maggs, A. C.
Veröffentlicht: (2025)
von: Maggs, A. C.
Veröffentlicht: (2025)
How Does Controllability Emerge In Language Models During Pretraining?
von: She, Jianshu, et al.
Veröffentlicht: (2025)
von: She, Jianshu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
IsoCompute Playbook: Optimally Scaling Sampling Compute for LLM RL
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2026) -
Concise Reasoning in the Lens of Lagrangian Optimization
von: Gao, Chengqian, et al.
Veröffentlicht: (2025) -
K2-V2: A 360-Open, Reasoning-Enhanced LLM
von: K2 Team, et al.
Veröffentlicht: (2025) -
Efficient Agentic Reasoning Through Self-Regulated Simulative Planning
von: Deng, Mingkai, et al.
Veröffentlicht: (2026) -
Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2025)