A Review of DeepSeek Models' Key Innovative Techniques
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Chengen, Kantarcioglu, Murat |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Graph Generative Models Evaluation with Masked Autoencoder
by: Wang, Chengen, et al.
Published: (2025)
by: Wang, Chengen, et al.
Published: (2025)
How to Backdoor Consistency Models?
by: Wang, Chengen, et al.
Published: (2024)
by: Wang, Chengen, et al.
Published: (2024)
A Systematic Evaluation of Generative Models on Tabular Transportation Data
by: Wang, Chengen, et al.
Published: (2025)
by: Wang, Chengen, et al.
Published: (2025)
Ask ChatGPT: Caveats and Mitigations for Individual Users of AI Chatbots
by: Wang, Chengen, et al.
Published: (2025)
by: Wang, Chengen, et al.
Published: (2025)
Memory Analysis on the Training Course of DeepSeek Models
by: Zhang, Ping, et al.
Published: (2025)
by: Zhang, Ping, et al.
Published: (2025)
Quantitative Analysis of Performance Drop in DeepSeek Model Quantization
by: Zhao, Enbo, et al.
Published: (2025)
by: Zhao, Enbo, et al.
Published: (2025)
Are DeepSeek R1 And Other Reasoning Models More Faithful?
by: Chua, James, et al.
Published: (2025)
by: Chua, James, et al.
Published: (2025)
Quantifying the Capability Boundary of DeepSeek Models: An Application-Driven Performance Analysis
by: Zhao, Kaikai, et al.
Published: (2025)
by: Zhao, Kaikai, et al.
Published: (2025)
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
by: DeepSeek-AI, et al.
Published: (2024)
by: DeepSeek-AI, et al.
Published: (2024)
Brief analysis of DeepSeek R1 and its implications for Generative AI
by: Mercer, Sarah, et al.
Published: (2025)
by: Mercer, Sarah, et al.
Published: (2025)
Benchmark-Driven Selection of AI: Evidence from DeepSeek-R1
by: Spelda, Petr, et al.
Published: (2025)
by: Spelda, Petr, et al.
Published: (2025)
SMOTE-DP: Improving Privacy-Utility Tradeoff with Synthetic Data
by: Zhou, Yan, et al.
Published: (2025)
by: Zhou, Yan, et al.
Published: (2025)
Learning Joint Embeddings of Function and Process Call Graphs for Malware Detection
by: Aneja, Kartikeya, et al.
Published: (2025)
by: Aneja, Kartikeya, et al.
Published: (2025)
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
by: DeepSeek-AI, et al.
Published: (2024)
by: DeepSeek-AI, et al.
Published: (2024)
R1dacted: Investigating Local Censorship in DeepSeek's R1 Language Model
by: Naseh, Ali, et al.
Published: (2025)
by: Naseh, Ali, et al.
Published: (2025)
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
by: Guo, Daya, et al.
Published: (2024)
by: Guo, Daya, et al.
Published: (2024)
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
by: DeepSeek-AI, et al.
Published: (2025)
by: DeepSeek-AI, et al.
Published: (2025)
DeepSeek on a Trip: Inducing Targeted Visual Hallucinations via Representation Vulnerabilities
by: Islam, Chashi Mahiul, et al.
Published: (2025)
by: Islam, Chashi Mahiul, et al.
Published: (2025)
Challenges in Ensuring AI Safety in DeepSeek-R1 Models: The Shortcomings of Reinforcement Learning Strategies
by: Parmar, Manojkumar, et al.
Published: (2025)
by: Parmar, Manojkumar, et al.
Published: (2025)
Token-Hungry, Yet Precise: DeepSeek R1 Highlights the Need for Multi-Step Reasoning Over Speed in MATH
by: Evstafev, Evgenii
Published: (2025)
by: Evstafev, Evgenii
Published: (2025)
How Effective Is Constitutional AI in Small LLMs? A Study on DeepSeek-R1 and Its Peers
by: Menke, Antonio-Gabriel Chacón, et al.
Published: (2025)
by: Menke, Antonio-Gabriel Chacón, et al.
Published: (2025)
DeepSeek vs. ChatGPT vs. Claude: A Comparative Study for Scientific Computing and Scientific Machine Learning Tasks
by: Jiang, Qile, et al.
Published: (2025)
by: Jiang, Qile, et al.
Published: (2025)
DeepSeek-Inspired Exploration of RL-based LLMs and Synergy with Wireless Networks: A Survey
by: Qiao, Yu, et al.
Published: (2025)
by: Qiao, Yu, et al.
Published: (2025)
FedDAG: Clustered Federated Learning via Global Data and Gradient Integration for Heterogeneous Environments
by: Pramanik, Anik, et al.
Published: (2026)
by: Pramanik, Anik, et al.
Published: (2026)
Challenges and Applications of Large Language Models: A Comparison of GPT and DeepSeek family of models
by: Sharma, Shubham, et al.
Published: (2025)
by: Sharma, Shubham, et al.
Published: (2025)
CourtGuard: A Model-Agnostic Framework for Zero-Shot Policy Adaptation in LLM Safety
by: Suleymanov, Umid, et al.
Published: (2026)
by: Suleymanov, Umid, et al.
Published: (2026)
PROVCREATOR: Synthesizing Complex Heterogenous Graphs with Node and Edge Attributes
by: Wang, Tianhao, et al.
Published: (2025)
by: Wang, Tianhao, et al.
Published: (2025)
Adversarial Graph Neural Network Benchmarks: Towards Practical and Fair Evaluation
by: Ngo, Tran Gia Bao, et al.
Published: (2026)
by: Ngo, Tran Gia Bao, et al.
Published: (2026)
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
by: Sadik, Ahmed R., et al.
Published: (2025)
by: Sadik, Ahmed R., et al.
Published: (2025)
MHA2MLA-VLM: Enabling DeepSeek's Economical Multi-Head Latent Attention across Vision-Language Models
by: Fan, Xiaoran, et al.
Published: (2026)
by: Fan, Xiaoran, et al.
Published: (2026)
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
by: Xin, Huajian, et al.
Published: (2024)
by: Xin, Huajian, et al.
Published: (2024)
Interpreting GNN-based IDS Detections Using Provenance Graph Structural Features
by: Mukherjee, Kunal, et al.
Published: (2023)
by: Mukherjee, Kunal, et al.
Published: (2023)
Robust and Explainable Divide-and-Conquer Learning for Intrusion Detection
by: Zhou, Yan, et al.
Published: (2026)
by: Zhou, Yan, et al.
Published: (2026)
Optimal Transport-Guided Adversarial Attacks on Graph Neural Network-Based Bot Detection
by: Mukherjee, Kunal, et al.
Published: (2026)
by: Mukherjee, Kunal, et al.
Published: (2026)
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation
by: Li, Zhuohang, et al.
Published: (2024)
by: Li, Zhuohang, et al.
Published: (2024)
Matched Topological Subspace Detector
by: Liu, Chengen, et al.
Published: (2025)
by: Liu, Chengen, et al.
Published: (2025)
VERI-DPO: Evidence-Aware Alignment for Clinical Summarization via Claim Verification and Direct Preference Optimization
by: Liu, Weixin, et al.
Published: (2026)
by: Liu, Weixin, et al.
Published: (2026)
Safety Evaluation of DeepSeek Models in Chinese Contexts
by: Zhang, Wenjing, et al.
Published: (2025)
by: Zhang, Wenjing, et al.
Published: (2025)
A Brief Review for Compression and Transfer Learning Techniques in DeepFake Detection
by: Karathanasis, Andreas, et al.
Published: (2025)
by: Karathanasis, Andreas, et al.
Published: (2025)
Large Language Model (LLM) for Telecommunications: A Comprehensive Survey on Principles, Key Techniques, and Opportunities
by: Zhou, Hao, et al.
Published: (2024)
by: Zhou, Hao, et al.
Published: (2024)
Similar Items
-
Graph Generative Models Evaluation with Masked Autoencoder
by: Wang, Chengen, et al.
Published: (2025) -
How to Backdoor Consistency Models?
by: Wang, Chengen, et al.
Published: (2024) -
A Systematic Evaluation of Generative Models on Tabular Transportation Data
by: Wang, Chengen, et al.
Published: (2025) -
Ask ChatGPT: Caveats and Mitigations for Individual Users of AI Chatbots
by: Wang, Chengen, et al.
Published: (2025) -
Memory Analysis on the Training Course of DeepSeek Models
by: Zhang, Ping, et al.
Published: (2025)