LLM360 K2: Building a 65B 360-Open-Source Large Language Model from Scratch
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Zhengzhong, Tan, Bowen, Wang, Hongyi, Neiswanger, Willie, Tao, Tianhua, Li, Haonan, Koto, Fajri, Wang, Yuqi, Sun, Suqi, Pangarkar, Omkar, Fan, Richard, Gu, Yi, Miller, Victor, Ma, Liqun, Tang, Liping, Ranjan, Nikhil, Zhuang, Yonghao, He, Guowei, Wang, Renxi, Deng, Mingkai, Algayres, Robin, Li, Yuanzhi, Shen, Zhiqiang, Nakov, Preslav, Xing, Eric |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
K2-V2: A 360-Open, Reasoning-Enhanced LLM
von: K2 Team, et al.
Veröffentlicht: (2025)
von: K2 Team, et al.
Veröffentlicht: (2025)
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents
von: Wang, Renxi, et al.
Veröffentlicht: (2025)
von: Wang, Renxi, et al.
Veröffentlicht: (2025)
SlimPajama-DC: Understanding Data Combinations for LLM Training
von: Shen, Zhiqiang, et al.
Veröffentlicht: (2023)
von: Shen, Zhiqiang, et al.
Veröffentlicht: (2023)
Cross-Cultural Transfer of Commonsense Reasoning in LLMs: Evidence from the Arab World
von: Almheiri, Saeed, et al.
Veröffentlicht: (2025)
von: Almheiri, Saeed, et al.
Veröffentlicht: (2025)
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2025)
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2025)
SimuScene: Training and Benchmarking Code Generation to Simulate Physical Scenarios
von: Wang, Yanan, et al.
Veröffentlicht: (2026)
von: Wang, Yanan, et al.
Veröffentlicht: (2026)
Cracking the Code: Multi-domain LLM Evaluation on Real-World Professional Exams in Indonesia
von: Koto, Fajri
Veröffentlicht: (2024)
von: Koto, Fajri
Veröffentlicht: (2024)
Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search
von: Zhang, Jingdong, et al.
Veröffentlicht: (2026)
von: Zhang, Jingdong, et al.
Veröffentlicht: (2026)
Instruction-Guided Poetry Generation in Arabic and Its Dialects
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Sadallah, Abdelrahman, et al.
Veröffentlicht: (2026)
DiffPortrait360: Consistent Portrait Diffusion for 360 View Synthesis
von: Gu, Yuming, et al.
Veröffentlicht: (2025)
von: Gu, Yuming, et al.
Veröffentlicht: (2025)
Rethinking STS and NLI in Large Language Models
von: Wang, Yuxia, et al.
Veröffentlicht: (2023)
von: Wang, Yuxia, et al.
Veröffentlicht: (2023)
AMIR-GRPO: Inducing Implicit Preference Signals into GRPO
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2026)
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2026)
Unveiling Cultural Blind Spots: Analyzing the Limitations of mLLMs in Procedural Text Comprehension
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
Dream360: Diverse and Immersive Outdoor Virtual Scene Creation via Transformer-Based 360 Image Outpainting
von: Ai, Hao, et al.
Veröffentlicht: (2024)
von: Ai, Hao, et al.
Veröffentlicht: (2024)
360PanT: Training-Free Text-Driven 360-Degree Panorama-to-Panorama Translation
von: Wang, Hai, et al.
Veröffentlicht: (2024)
von: Wang, Hai, et al.
Veröffentlicht: (2024)
DenoiseRank: Learning to Rank by Diffusion Models
von: Wang, Ying, et al.
Veröffentlicht: (2026)
von: Wang, Ying, et al.
Veröffentlicht: (2026)
CRF360D: Monocular 360 Depth Estimation via Spherical Fully-Connected CRFs
von: Cao, Zidong, et al.
Veröffentlicht: (2024)
von: Cao, Zidong, et al.
Veröffentlicht: (2024)
Parallel Tokenizers: Rethinking Vocabulary Design for Cross-Lingual Transfer
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
von: Kautsar, Muhammad Dehan Al, et al.
Veröffentlicht: (2025)
360
Veröffentlicht: (2019)
Veröffentlicht: (2019)
360DVD: Controllable Panorama Video Generation with 360-Degree Video Diffusion Model
von: Wang, Qian, et al.
Veröffentlicht: (2024)
von: Wang, Qian, et al.
Veröffentlicht: (2024)
pteacher/360_aiu: AIU 360 app
von: Ruslan Isaev
Veröffentlicht: (2025)
von: Ruslan Isaev
Veröffentlicht: (2025)
CUBE360: Learning Cubic Field Representation for Monocular 360 Depth Estimation for Virtual Reality
von: Chang, Wenjie, et al.
Veröffentlicht: (2024)
von: Chang, Wenjie, et al.
Veröffentlicht: (2024)
Web2Code: A Large-scale Webpage-to-Code Dataset and Evaluation Framework for Multimodal LLMs
von: Yun, Sukmin, et al.
Veröffentlicht: (2024)
von: Yun, Sukmin, et al.
Veröffentlicht: (2024)
UnsafeChain: Enhancing Reasoning Model Safety via Hard Cases
von: Tomar, Raj Vardhan, et al.
Veröffentlicht: (2025)
von: Tomar, Raj Vardhan, et al.
Veröffentlicht: (2025)
How Does Prefix Matter in Reasoning Model Tuning?
von: Tomar, Raj Vardhan, et al.
Veröffentlicht: (2026)
von: Tomar, Raj Vardhan, et al.
Veröffentlicht: (2026)
Elite360M: Efficient 360 Multi-task Learning via Bi-projection Fusion and Cross-task Collaboration
von: Ai, Hao, et al.
Veröffentlicht: (2024)
von: Ai, Hao, et al.
Veröffentlicht: (2024)
Elite360D: Towards Efficient 360 Depth Estimation via Semantic- and Distance-Aware Bi-Projection Fusion
von: Ai, Hao, et al.
Veröffentlicht: (2024)
von: Ai, Hao, et al.
Veröffentlicht: (2024)
ArabicMMLU: Assessing Massive Multitask Language Understanding in Arabic
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
von: Koto, Fajri, et al.
Veröffentlicht: (2024)
ICX360: In-Context eXplainability 360 Toolkit
von: Wei, Dennis, et al.
Veröffentlicht: (2025)
von: Wei, Dennis, et al.
Veröffentlicht: (2025)
Low-Resource Safety Failures Are Action Failures, Not Representation Failures
von: Aziz, Rashad, et al.
Veröffentlicht: (2026)
von: Aziz, Rashad, et al.
Veröffentlicht: (2026)
RePer-360: Releasing Perspective Priors for 360$^\circ$ Depth Estimation via Self-Modulation
von: Guan, Cheng, et al.
Veröffentlicht: (2026)
von: Guan, Cheng, et al.
Veröffentlicht: (2026)
Kidney360
Veröffentlicht: (2024)
Veröffentlicht: (2024)
IM360: Large-scale Indoor Mapping with 360 Cameras
von: Jung, Dongki, et al.
Veröffentlicht: (2025)
von: Jung, Dongki, et al.
Veröffentlicht: (2025)
360Anything: Geometry-Free Lifting of Images and Videos to 360°
von: Wu, Ziyi, et al.
Veröffentlicht: (2026)
von: Wu, Ziyi, et al.
Veröffentlicht: (2026)
Can Machines Resonate with Humans? Evaluating the Emotional and Empathic Comprehension of LMs
von: Manzoor, Muhammad Arslan, et al.
Veröffentlicht: (2024)
von: Manzoor, Muhammad Arslan, et al.
Veröffentlicht: (2024)
MuDRiC: Multi-Dialect Reasoning for Arabic Commonsense Validation
von: Elozeiri, Kareem, et al.
Veröffentlicht: (2025)
von: Elozeiri, Kareem, et al.
Veröffentlicht: (2025)
Inpaint360GS: Efficient Object-Aware 3D Inpainting via Gaussian Splatting for 360° Scenes
von: Wang, Shaoxiang, et al.
Veröffentlicht: (2025)
von: Wang, Shaoxiang, et al.
Veröffentlicht: (2025)
Head360: Learning a Parametric 3D Full-Head for Free-View Synthesis in 360°
von: He, Yuxiao, et al.
Veröffentlicht: (2024)
von: He, Yuxiao, et al.
Veröffentlicht: (2024)
360DVO: Deep Visual Odometry for Monocular 360-Degree Camera
von: Guo, Xiaopeng, et al.
Veröffentlicht: (2026)
von: Guo, Xiaopeng, et al.
Veröffentlicht: (2026)
Imagine360: Immersive 360 Video Generation from Perspective Anchor
von: Tan, Jing, et al.
Veröffentlicht: (2024)
von: Tan, Jing, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
K2-V2: A 360-Open, Reasoning-Enhanced LLM
von: K2 Team, et al.
Veröffentlicht: (2025) -
AgentFly: Extensible and Scalable Reinforcement Learning for LM Agents
von: Wang, Renxi, et al.
Veröffentlicht: (2025) -
SlimPajama-DC: Understanding Data Combinations for LLM Training
von: Shen, Zhiqiang, et al.
Veröffentlicht: (2023) -
Cross-Cultural Transfer of Commonsense Reasoning in LLMs: Evidence from the Arab World
von: Almheiri, Saeed, et al.
Veröffentlicht: (2025) -
Instruction Tuning on Public Government and Cultural Data for Low-Resource Language: a Case Study in Kazakh
von: Laiyk, Nurkhan, et al.
Veröffentlicht: (2025)