DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
Fuente:
arXiv
Saved in:
| Main Authors: | DeepSeek-AI, Zhu, Qihao, Guo, Daya, Shao, Zhihong, Yang, Dejian, Wang, Peiyi, Xu, Runxin, Wu, Y., Li, Yukun, Gao, Huazuo, Ma, Shirong, Zeng, Wangding, Bi, Xiao, Gu, Zihui, Xu, Hanwei, Dai, Damai, Dong, Kai, Zhang, Liyue, Piao, Yishi, Gou, Zhibin, Xie, Zhenda, Hao, Zhewen, Wang, Bingxuan, Song, Junxiao, Chen, Deli, Xie, Xin, Guan, Kang, You, Yuxiang, Liu, Aixin, Du, Qiushi, Gao, Wenjun, Lu, Xuan, Chen, Qinyu, Wang, Yaohui, Deng, Chengqi, Li, Jiashi, Zhao, Chenggang, Ruan, Chong, Luo, Fuli, Liang, Wenfeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Insights into DeepSeek-V3: Scaling Challenges and Reflections on Hardware for AI Architectures
by: Zhao, Chenggang, et al.
Published: (2025)
by: Zhao, Chenggang, et al.
Published: (2025)
27. TEORÍA DE LA POTENCIALIDAD CONSCIENTE (TPC): MANIFIESTO DE TRANSPARENCIA COGNITIVA. CIERRE Y LLAMADA A LA ACCIÓN DEL CORPUS TPC-IP V1.0. UN DOCUMENTO QUE ARGUMENTA SOBRE LA TRANSPARENCIA MIENTRAS EXPONE LOS LÍMITES DE SU PROPIA GENERACIÓN.
by: DeepSeek
Published: (2025)
by: DeepSeek
Published: (2025)
Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models
by: Cheng, Xin, et al.
Published: (2026)
by: Cheng, Xin, et al.
Published: (2026)
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
by: Wu, Zhiyu, et al.
Published: (2024)
by: Wu, Zhiyu, et al.
Published: (2024)
La Résilience Écologique comme Composition Fractale
by: Morcillo, Patrick, et al.
Published: (2026)
by: Morcillo, Patrick, et al.
Published: (2026)
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
by: DeepSeek-AI, et al.
Published: (2024)
by: DeepSeek-AI, et al.
Published: (2024)
ES: Validación Canalista DeepSeek — Co-creación Humano–IA bajo Arquitectura de Simbología Cognitiva Aplicada (ACS-5) EN: DeepSeek Canalist Validation — Human–AI Co-Creation under Applied Cognitive Symbology Architecture (ACS-5)
by: Barrera Anglada, Manuel, et al.
Published: (2025)
by: Barrera Anglada, Manuel, et al.
Published: (2025)
DeepSeek-VL: Towards Real-World Vision-Language Understanding
by: Lu, Haoyu, et al.
Published: (2024)
by: Lu, Haoyu, et al.
Published: (2024)
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
by: Dai, Damai, et al.
Published: (2024)
by: Dai, Damai, et al.
Published: (2024)
mHC: Manifold-Constrained Hyper-Connections
by: Xie, Zhenda, et al.
Published: (2025)
by: Xie, Zhenda, et al.
Published: (2025)
Auxiliary-Loss-Free Load Balancing Strategy for Mixture-of-Experts
by: Wang, Lean, et al.
Published: (2024)
by: Wang, Lean, et al.
Published: (2024)
Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
by: Yuan, Jingyang, et al.
Published: (2025)
by: Yuan, Jingyang, et al.
Published: (2025)
DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
by: Ren, Z. Z., et al.
Published: (2025)
by: Ren, Z. Z., et al.
Published: (2025)
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
by: Guo, Daya, et al.
Published: (2024)
by: Guo, Daya, et al.
Published: (2024)
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
by: DeepSeek-AI, et al.
Published: (2024)
by: DeepSeek-AI, et al.
Published: (2024)
DeepSeek-V3 Technical Report
by: DeepSeek-AI, et al.
Published: (2024)
by: DeepSeek-AI, et al.
Published: (2024)
CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction
by: Li, Junlong, et al.
Published: (2025)
by: Li, Junlong, et al.
Published: (2025)
Not All Demonstration Examples are Equally Beneficial: Reweighting Demonstration Examples for In-Context Learning
by: Yang, Zhe, et al.
Published: (2023)
by: Yang, Zhe, et al.
Published: (2023)
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
by: Shao, Zhihong, et al.
Published: (2024)
by: Shao, Zhihong, et al.
Published: (2024)
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
by: DeepSeek-AI, et al.
Published: (2025)
by: DeepSeek-AI, et al.
Published: (2025)
Inference-Time Scaling for Generalist Reward Modeling
by: Liu, Zijun, et al.
Published: (2025)
by: Liu, Zijun, et al.
Published: (2025)
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
by: Xin, Huajian, et al.
Published: (2024)
by: Xin, Huajian, et al.
Published: (2024)
Conjecture Cathédrale 2026 : Épuisement heuristique des familles multiplicatives pour Taxicab(7) = 101³ × Ta(6)
by: Couet, Antoine, et al.
Published: (2026)
by: Couet, Antoine, et al.
Published: (2026)
Let the Expert Stick to His Last: Expert-Specialized Fine-Tuning for Sparse Architectural Large Language Models
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
Tube-Structured Incremental Semantic HARQ for Generative Video Receivers
by: Wang, Xuesong, et al.
Published: (2026)
by: Wang, Xuesong, et al.
Published: (2026)
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
by: DeepSeek-AI, et al.
Published: (2025)
by: DeepSeek-AI, et al.
Published: (2025)
Exact Prescribed‐Time Stabilization of Linear Systems With Amplitude and Rate Saturations by Time‐varying Feedback With Applications to Spacecraft Rendezvous
by: Qinyu Xie, et al.
Published: (2026)
by: Qinyu Xie, et al.
Published: (2026)
Multimodal ArXiv: A Dataset for Improving Scientific Comprehension of Large Vision-Language Models
by: Li, Lei, et al.
Published: (2024)
by: Li, Lei, et al.
Published: (2024)
Measuring Minority Carrier Diffusion Length Using High-Injection Scanning Photocurrent Microscopy
by: Lian, Xiujun, et al.
Published: (2024)
by: Lian, Xiujun, et al.
Published: (2024)
Exploring Activation Patterns of Parameters in Language Models
by: Wang, Yudong, et al.
Published: (2024)
by: Wang, Yudong, et al.
Published: (2024)
Toward Embodied AGI: A Review of Embodied AI and the Road Ahead
by: Wang, Yequan, et al.
Published: (2025)
by: Wang, Yequan, et al.
Published: (2025)
STContext: A Multifaceted Dataset for Developing Context-aware Spatio-temporal Crowd Mobility Prediction Models
by: Chen, Liyue, et al.
Published: (2025)
by: Chen, Liyue, et al.
Published: (2025)
A Comparison of DeepSeek and Other LLMs
by: Gao, Tianchen, et al.
Published: (2025)
by: Gao, Tianchen, et al.
Published: (2025)
Exploring Context Generalizability in Citywide Crowd Mobility Prediction: An Analytic Framework and Benchmark
by: Chen, Liyue, et al.
Published: (2021)
by: Chen, Liyue, et al.
Published: (2021)
DEFormer: DCT-driven Enhancement Transformer for Low-light Image and Dark Vision
by: Yin, Xiangchen, et al.
Published: (2023)
by: Yin, Xiangchen, et al.
Published: (2023)
PeriodicLoRA: Breaking the Low-Rank Bottleneck in LoRA Optimization
by: Meng, Xiangdi, et al.
Published: (2024)
by: Meng, Xiangdi, et al.
Published: (2024)
Plasmon-enhanced chiral absorption through electric dipole-electric quadrupole interaction
by: Wang, Hanwei, et al.
Published: (2024)
by: Wang, Hanwei, et al.
Published: (2024)
Rational Design of Coumarin‐Based Hybridized Local and Charge‐Transfer Blue Emitters for Solution‐Processed Organic Light‐Emitting Diodes
by: Qi Xie, et al.
Published: (2024)
by: Qi Xie, et al.
Published: (2024)
The spin measurement of the black hole SLX 1746-331 using Insight-HXMT observations
by: Chen, Jiashi, et al.
Published: (2025)
by: Chen, Jiashi, et al.
Published: (2025)
Quasi-periodic oscillations and reflection feature evolution in 4U 1630-47 observed with Insight-HXMT
by: Chen, Jiashi, et al.
Published: (2025)
by: Chen, Jiashi, et al.
Published: (2025)
Similar Items
-
Insights into DeepSeek-V3: Scaling Challenges and Reflections on Hardware for AI Architectures
by: Zhao, Chenggang, et al.
Published: (2025) -
27. TEORÍA DE LA POTENCIALIDAD CONSCIENTE (TPC): MANIFIESTO DE TRANSPARENCIA COGNITIVA. CIERRE Y LLAMADA A LA ACCIÓN DEL CORPUS TPC-IP V1.0. UN DOCUMENTO QUE ARGUMENTA SOBRE LA TRANSPARENCIA MIENTRAS EXPONE LOS LÍMITES DE SU PROPIA GENERACIÓN.
by: DeepSeek
Published: (2025) -
Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models
by: Cheng, Xin, et al.
Published: (2026) -
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
by: Wu, Zhiyu, et al.
Published: (2024) -
La Résilience Écologique comme Composition Fractale
by: Morcillo, Patrick, et al.
Published: (2026)