LLM360 K2: Building a 65B 360-Open-Source Large Language Model from Scratch
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Zhengzhong, Tan, Bowen, Wang, Hongyi, Neiswanger, Willie, Tao, Tianhua, Li, Haonan, Koto, Fajri, Wang, Yuqi, Sun, Suqi, Pangarkar, Omkar, Fan, Richard, Gu, Yi, Miller, Victor, Ma, Liqun, Tang, Liping, Ranjan, Nikhil, Zhuang, Yonghao, He, Guowei, Wang, Renxi, Deng, Mingkai, Algayres, Robin, Li, Yuanzhi, Shen, Zhiqiang, Nakov, Preslav, Xing, Eric |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
K2-V2: A 360-Open, Reasoning-Enhanced LLM
by: K2 Team, et al.
Published: (2025)
by: K2 Team, et al.
Published: (2025)
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents
by: Wang, Renxi, et al.
Published: (2025)
by: Wang, Renxi, et al.
Published: (2025)
SlimPajama-DC: Understanding Data Combinations for LLM Training
by: Shen, Zhiqiang, et al.
Published: (2023)
by: Shen, Zhiqiang, et al.
Published: (2023)
Cross-Cultural Transfer of Commonsense Reasoning in LLMs: Evidence from the Arab World
by: Almheiri, Saeed, et al.
Published: (2025)
by: Almheiri, Saeed, et al.
Published: (2025)
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
by: Laiyk, Nurkhan, et al.
Published: (2025)
by: Laiyk, Nurkhan, et al.
Published: (2025)
SimuScene: Training and Benchmarking Code Generation to Simulate Physical Scenarios
by: Wang, Yanan, et al.
Published: (2026)
by: Wang, Yanan, et al.
Published: (2026)
Cracking the Code: Multi-domain LLM Evaluation on Real-World Professional Exams in Indonesia
by: Koto, Fajri
Published: (2024)
by: Koto, Fajri
Published: (2024)
Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search
by: Zhang, Jingdong, et al.
Published: (2026)
by: Zhang, Jingdong, et al.
Published: (2026)
Instruction-Guided Poetry Generation in Arabic and Its Dialects
by: Sadallah, Abdelrahman, et al.
Published: (2026)
by: Sadallah, Abdelrahman, et al.
Published: (2026)
DiffPortrait360: Consistent Portrait Diffusion for 360 View Synthesis
by: Gu, Yuming, et al.
Published: (2025)
by: Gu, Yuming, et al.
Published: (2025)
Rethinking STS and NLI in Large Language Models
by: Wang, Yuxia, et al.
Published: (2023)
by: Wang, Yuxia, et al.
Published: (2023)
AMIR-GRPO: Inducing Implicit Preference Signals into GRPO
by: Yari, Amir Hossein, et al.
Published: (2026)
by: Yari, Amir Hossein, et al.
Published: (2026)
Unveiling Cultural Blind Spots: Analyzing the Limitations of mLLMs in Procedural Text Comprehension
by: Yari, Amir Hossein, et al.
Published: (2025)
by: Yari, Amir Hossein, et al.
Published: (2025)
Dream360: Diverse and Immersive Outdoor Virtual Scene Creation via Transformer-Based 360 Image Outpainting
by: Ai, Hao, et al.
Published: (2024)
by: Ai, Hao, et al.
Published: (2024)
360PanT: Training-Free Text-Driven 360-Degree Panorama-to-Panorama Translation
by: Wang, Hai, et al.
Published: (2024)
by: Wang, Hai, et al.
Published: (2024)
DenoiseRank: Learning to Rank by Diffusion Models
by: Wang, Ying, et al.
Published: (2026)
by: Wang, Ying, et al.
Published: (2026)
CRF360D: Monocular 360 Depth Estimation via Spherical Fully-Connected CRFs
by: Cao, Zidong, et al.
Published: (2024)
by: Cao, Zidong, et al.
Published: (2024)
Parallel Tokenizers: Rethinking Vocabulary Design for Cross-Lingual Transfer
by: Kautsar, Muhammad Dehan Al, et al.
Published: (2025)
by: Kautsar, Muhammad Dehan Al, et al.
Published: (2025)
360
Published: (2019)
Published: (2019)
360DVD: Controllable Panorama Video Generation with 360-Degree Video Diffusion Model
by: Wang, Qian, et al.
Published: (2024)
by: Wang, Qian, et al.
Published: (2024)
pteacher/360_aiu: AIU 360 app
by: Ruslan Isaev
Published: (2025)
by: Ruslan Isaev
Published: (2025)
CUBE360: Learning Cubic Field Representation for Monocular 360 Depth Estimation for Virtual Reality
by: Chang, Wenjie, et al.
Published: (2024)
by: Chang, Wenjie, et al.
Published: (2024)
Web2Code: A Large-scale Webpage-to-Code Dataset and Evaluation Framework for Multimodal LLMs
by: Yun, Sukmin, et al.
Published: (2024)
by: Yun, Sukmin, et al.
Published: (2024)
UnsafeChain: Enhancing Reasoning Model Safety via Hard Cases
by: Tomar, Raj Vardhan, et al.
Published: (2025)
by: Tomar, Raj Vardhan, et al.
Published: (2025)
How Does Prefix Matter in Reasoning Model Tuning?
by: Tomar, Raj Vardhan, et al.
Published: (2026)
by: Tomar, Raj Vardhan, et al.
Published: (2026)
Elite360M: Efficient 360 Multi-task Learning via Bi-projection Fusion and Cross-task Collaboration
by: Ai, Hao, et al.
Published: (2024)
by: Ai, Hao, et al.
Published: (2024)
Elite360D: Towards Efficient 360 Depth Estimation via Semantic- and Distance-Aware Bi-Projection Fusion
by: Ai, Hao, et al.
Published: (2024)
by: Ai, Hao, et al.
Published: (2024)
ArabicMMLU: Assessing Massive Multitask Language Understanding in Arabic
by: Koto, Fajri, et al.
Published: (2024)
by: Koto, Fajri, et al.
Published: (2024)
ICX360: In-Context eXplainability 360 Toolkit
by: Wei, Dennis, et al.
Published: (2025)
by: Wei, Dennis, et al.
Published: (2025)
Low-Resource Safety Failures Are Action Failures, Not Representation Failures
by: Aziz, Rashad, et al.
Published: (2026)
by: Aziz, Rashad, et al.
Published: (2026)
RePer-360: Releasing Perspective Priors for 360$^\circ$ Depth Estimation via Self-Modulation
by: Guan, Cheng, et al.
Published: (2026)
by: Guan, Cheng, et al.
Published: (2026)
Kidney360
Published: (2024)
Published: (2024)
IM360: Large-scale Indoor Mapping with 360 Cameras
by: Jung, Dongki, et al.
Published: (2025)
by: Jung, Dongki, et al.
Published: (2025)
360Anything: Geometry-Free Lifting of Images and Videos to 360°
by: Wu, Ziyi, et al.
Published: (2026)
by: Wu, Ziyi, et al.
Published: (2026)
Can Machines Resonate with Humans? Evaluating the Emotional and Empathic Comprehension of LMs
by: Manzoor, Muhammad Arslan, et al.
Published: (2024)
by: Manzoor, Muhammad Arslan, et al.
Published: (2024)
Inpaint360GS: Efficient Object-Aware 3D Inpainting via Gaussian Splatting for 360° Scenes
by: Wang, Shaoxiang, et al.
Published: (2025)
by: Wang, Shaoxiang, et al.
Published: (2025)
MuDRiC: Multi-Dialect Reasoning for Arabic Commonsense Validation
by: Elozeiri, Kareem, et al.
Published: (2025)
by: Elozeiri, Kareem, et al.
Published: (2025)
Head360: Learning a Parametric 3D Full-Head for Free-View Synthesis in 360°
by: He, Yuxiao, et al.
Published: (2024)
by: He, Yuxiao, et al.
Published: (2024)
360DVO: Deep Visual Odometry for Monocular 360-Degree Camera
by: Guo, Xiaopeng, et al.
Published: (2026)
by: Guo, Xiaopeng, et al.
Published: (2026)
Imagine360: Immersive 360 Video Generation from Perspective Anchor
by: Tan, Jing, et al.
Published: (2024)
by: Tan, Jing, et al.
Published: (2024)
Similar Items
-
K2-V2: A 360-Open, Reasoning-Enhanced LLM
by: K2 Team, et al.
Published: (2025) -
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents
by: Wang, Renxi, et al.
Published: (2025) -
SlimPajama-DC: Understanding Data Combinations for LLM Training
by: Shen, Zhiqiang, et al.
Published: (2023) -
Cross-Cultural Transfer of Commonsense Reasoning in LLMs: Evidence from the Arab World
by: Almheiri, Saeed, et al.
Published: (2025) -
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
by: Laiyk, Nurkhan, et al.
Published: (2025)