DeepSeek-VL: Towards Real-World Vision-Language Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Haoyu, Liu, Wen, Zhang, Bo, Wang, Bingxuan, Dong, Kai, Liu, Bo, Sun, Jingxiang, Ren, Tongzheng, Li, Zhuoshu, Yang, Hao, Sun, Yaofeng, Deng, Chengqi, Xu, Hanwei, Xie, Zhenda, Ruan, Chong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
by: Wu, Zhiyu, et al.
Published: (2024)
by: Wu, Zhiyu, et al.
Published: (2024)
DeepSeek-OCR: Contexts Optical Compression
by: Wei, Haoran, et al.
Published: (2025)
by: Wei, Haoran, et al.
Published: (2025)
DeepSeek-OCR 2: Visual Causal Flow
by: Wei, Haoran, et al.
Published: (2026)
by: Wei, Haoran, et al.
Published: (2026)
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
by: DeepSeek-AI, et al.
Published: (2024)
by: DeepSeek-AI, et al.
Published: (2024)
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
by: Xin, Huajian, et al.
Published: (2024)
by: Xin, Huajian, et al.
Published: (2024)
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
by: DeepSeek-AI, et al.
Published: (2024)
by: DeepSeek-AI, et al.
Published: (2024)
Insights into DeepSeek-V3: Scaling Challenges and Reflections on Hardware for AI Architectures
by: Zhao, Chenggang, et al.
Published: (2025)
by: Zhao, Chenggang, et al.
Published: (2025)
DeepSeek-V3 Technical Report
by: DeepSeek-AI, et al.
Published: (2024)
by: DeepSeek-AI, et al.
Published: (2024)
Towards Understanding the Safety Boundaries of DeepSeek Models: Evaluation and Findings
by: Ying, Zonghao, et al.
Published: (2025)
by: Ying, Zonghao, et al.
Published: (2025)
DeepSeek in Education: Exploring the Transformative Potential of AI‐Driven Educational Intelligence
by: Jian Liao, et al.
Published: (2025)
by: Jian Liao, et al.
Published: (2025)
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
by: Guo, Daya, et al.
Published: (2024)
by: Guo, Daya, et al.
Published: (2024)
Safety Evaluation of DeepSeek Models in Chinese Contexts
by: Zhang, Wenjing, et al.
Published: (2025)
by: Zhang, Wenjing, et al.
Published: (2025)
Visual Merit or Linguistic Crutch? A Close Look at DeepSeek-OCR
by: Liang, Yunhao, et al.
Published: (2026)
by: Liang, Yunhao, et al.
Published: (2026)
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
by: DeepSeek-AI, et al.
Published: (2024)
by: DeepSeek-AI, et al.
Published: (2024)
Can DeepSeek Reason Like a Surgeon? An Empirical Evaluation for Vision-Language Understanding in Robotic-Assisted Surgery
by: Ma, Boyi, et al.
Published: (2025)
by: Ma, Boyi, et al.
Published: (2025)
DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
by: Xin, Huajian, et al.
Published: (2024)
by: Xin, Huajian, et al.
Published: (2024)
Research on Collaborative Governance of AIGC Applications in the DeepSeek Era
by: Shengli Deng, et al.
Published: (2025)
by: Shengli Deng, et al.
Published: (2025)
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
by: DeepSeek-AI, et al.
Published: (2025)
by: DeepSeek-AI, et al.
Published: (2025)
Safety Evaluation and Enhancement of DeepSeek Models in Chinese Contexts
by: Zhang, Wenjing, et al.
Published: (2025)
by: Zhang, Wenjing, et al.
Published: (2025)
Spectral Representation for Causal Estimation with Hidden Confounders
by: Sun, Haotian, et al.
Published: (2024)
by: Sun, Haotian, et al.
Published: (2024)
Synergizing DeepSeek's artificial intelligence innovations with brain–computer interfaces
by: Canbiao Wu, et al.
Published: (2025)
by: Canbiao Wu, et al.
Published: (2025)
An evaluation of DeepSeek Models in Biomedical Natural Language Processing
by: Zhan, Zaifu, et al.
Published: (2025)
by: Zhan, Zaifu, et al.
Published: (2025)
A Comparison of DeepSeek and Other LLMs
by: Gao, Tianchen, et al.
Published: (2025)
by: Gao, Tianchen, et al.
Published: (2025)
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
by: DeepSeek-AI, et al.
Published: (2025)
by: DeepSeek-AI, et al.
Published: (2025)
Quantitative Analysis of Performance Drop in DeepSeek Model Quantization
by: Zhao, Enbo, et al.
Published: (2025)
by: Zhao, Enbo, et al.
Published: (2025)
Quantifying the Capability Boundary of DeepSeek Models: An Application-Driven Performance Analysis
by: Zhao, Kaikai, et al.
Published: (2025)
by: Zhao, Kaikai, et al.
Published: (2025)
DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
by: Ren, Z. Z., et al.
Published: (2025)
by: Ren, Z. Z., et al.
Published: (2025)
Firebolt-VL: Efficient Vision-Language Understanding with Cross-Modality Modulation
by: Trinh, Quoc-Huy, et al.
Published: (2026)
by: Trinh, Quoc-Huy, et al.
Published: (2026)
RealSafe-R1: Safety-Aligned DeepSeek-R1 without Compromising Reasoning Capability
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
MHA2MLA-VLM: Enabling DeepSeek's Economical Multi-Head Latent Attention across Vision-Language Models
by: Fan, Xiaoran, et al.
Published: (2026)
by: Fan, Xiaoran, et al.
Published: (2026)
DeepSeek reshaping healthcare in China's tertiary hospitals
by: Chen, Jishizhan, et al.
Published: (2025)
by: Chen, Jishizhan, et al.
Published: (2025)
Memory Analysis on the Training Course of DeepSeek Models
by: Zhang, Ping, et al.
Published: (2025)
by: Zhang, Ping, et al.
Published: (2025)
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
by: Dai, Damai, et al.
Published: (2024)
by: Dai, Damai, et al.
Published: (2024)
DeepSeek Robustness Against Semantic-Character Dual-Space Mutated Prompt Injection
by: Ren, Junyu, et al.
Published: (2026)
by: Ren, Junyu, et al.
Published: (2026)
Can LLMs Assist Computer Education? an Empirical Case Study of DeepSeek
by: Xiao, Dongfu, et al.
Published: (2025)
by: Xiao, Dongfu, et al.
Published: (2025)
DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models
by: Xiong, Luolin, et al.
Published: (2025)
by: Xiong, Luolin, et al.
Published: (2025)
A DeepSeek-Powered AI System for Automated Chest Radiograph Interpretation in Clinical Practice
by: Bai, Yaowei, et al.
Published: (2025)
by: Bai, Yaowei, et al.
Published: (2025)
Stochastic Nonlinear Control via Finite-dimensional Spectral Dynamic Embedding
by: Ren, Zhaolin, et al.
Published: (2023)
by: Ren, Zhaolin, et al.
Published: (2023)
AgriGPT-VL: Agricultural Vision-Language Understanding Suite
by: Yang, Bo, et al.
Published: (2025)
by: Yang, Bo, et al.
Published: (2025)
A Review of DeepSeek Models' Key Innovative Techniques
by: Wang, Chengen, et al.
Published: (2025)
by: Wang, Chengen, et al.
Published: (2025)
Similar Items
-
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
by: Wu, Zhiyu, et al.
Published: (2024) -
DeepSeek-OCR: Contexts Optical Compression
by: Wei, Haoran, et al.
Published: (2025) -
DeepSeek-OCR 2: Visual Causal Flow
by: Wei, Haoran, et al.
Published: (2026) -
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
by: DeepSeek-AI, et al.
Published: (2024) -
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
by: Xin, Huajian, et al.
Published: (2024)