Evaluating the Performance of the DeepSeek Model in Confidential Computing Environment
Fuente:
arXiv
Salvato in:
| Autori principali: | Dong, Ben, Wang, Qian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Memory Analysis on the Training Course of DeepSeek Models
di: Zhang, Ping, et al.
Pubblicazione: (2025)
di: Zhang, Ping, et al.
Pubblicazione: (2025)
Performance of Confidential Computing GPUs
di: Ibarra, Antonio Martínez, et al.
Pubblicazione: (2025)
di: Ibarra, Antonio Martínez, et al.
Pubblicazione: (2025)
Confidential Computing on NVIDIA Hopper GPUs: A Performance Benchmark Study
di: Zhu, Jianwei, et al.
Pubblicazione: (2024)
di: Zhu, Jianwei, et al.
Pubblicazione: (2024)
DeepSeek-R1 Outperforms Gemini 2.0 Pro, OpenAI o1, and o3-mini in Bilingual Complex Ophthalmology Reasoning
di: Xu, Pusheng, et al.
Pubblicazione: (2025)
di: Xu, Pusheng, et al.
Pubblicazione: (2025)
A performance analysis of VM-based Trusted Execution Environments for Confidential Federated Learning
di: Casella, Bruno
Pubblicazione: (2025)
di: Casella, Bruno
Pubblicazione: (2025)
The Multiserver-Job Stochastic Recurrence Equation for Cloud Computing Performance Evaluation
di: Baccelli, Francois, et al.
Pubblicazione: (2026)
di: Baccelli, Francois, et al.
Pubblicazione: (2026)
A Microbenchmark Framework for Performance Evaluation of OpenMP Target Offloading
di: Atif, Mohammad, et al.
Pubblicazione: (2025)
di: Atif, Mohammad, et al.
Pubblicazione: (2025)
Confidential LLM Inference: Performance and Cost Across CPU and GPU TEEs
di: Chrapek, Marcin, et al.
Pubblicazione: (2025)
di: Chrapek, Marcin, et al.
Pubblicazione: (2025)
gpu tracker: Python Package for Tracking and Profiling GPU and Other Hardware Utilization in Both Desktop and High-Performance Computing Environments
di: Huckvale, Erik D., et al.
Pubblicazione: (2024)
di: Huckvale, Erik D., et al.
Pubblicazione: (2024)
PerfSeer: An Efficient and Accurate Deep Learning Models Performance Predictor
di: Zhao, Xinlong, et al.
Pubblicazione: (2025)
di: Zhao, Xinlong, et al.
Pubblicazione: (2025)
Performance Characterization of Containers in Edge Computing
di: Gupta, Ragini, et al.
Pubblicazione: (2025)
di: Gupta, Ragini, et al.
Pubblicazione: (2025)
Systematic Performance Evaluation Framework for LEO Mega-Constellation Satellite Networks
di: Wang, Yu, et al.
Pubblicazione: (2024)
di: Wang, Yu, et al.
Pubblicazione: (2024)
Performance Evaluation of Subroutines Call in PHP
di: Kalmukov, Yordan
Pubblicazione: (2026)
di: Kalmukov, Yordan
Pubblicazione: (2026)
A Continuous Benchmarking Infrastructure for High-Performance Computing Applications
di: Alt, Christoph, et al.
Pubblicazione: (2024)
di: Alt, Christoph, et al.
Pubblicazione: (2024)
Performance Evaluation of CMOS Annealing with Support Vector Machine
di: Fukuhara, Ryoga, et al.
Pubblicazione: (2024)
di: Fukuhara, Ryoga, et al.
Pubblicazione: (2024)
Performance Optimization of 3D Stencil Computation on ARM Scalable Vector Extension
di: Chen, Hongguang
Pubblicazione: (2025)
di: Chen, Hongguang
Pubblicazione: (2025)
Hierarchical Analyses Applied to Computer System Performance: Review and Call for Further Studies
di: Thomasian, Alexander
Pubblicazione: (2024)
di: Thomasian, Alexander
Pubblicazione: (2024)
OSCAR-P and aMLLibrary: Profiling and Predicting the Performance of FaaS-based Applications in Computing Continua
di: Sala, Roberto, et al.
Pubblicazione: (2024)
di: Sala, Roberto, et al.
Pubblicazione: (2024)
Distilled Large Language Model in Confidential Computing Environment for System-on-Chip Design
di: Ben, Dong, et al.
Pubblicazione: (2025)
di: Ben, Dong, et al.
Pubblicazione: (2025)
An Analytical Cost Model for Fast Evaluation of Multiple Compute-Engine CNN Accelerators
di: Qararyah, Fareed, et al.
Pubblicazione: (2025)
di: Qararyah, Fareed, et al.
Pubblicazione: (2025)
A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations
di: Flavin, Timothy, et al.
Pubblicazione: (2026)
di: Flavin, Timothy, et al.
Pubblicazione: (2026)
Cost-Performance Evaluation of General Compute Instances: AWS, Azure, GCP, and OCI
di: Tharwani, Jay, et al.
Pubblicazione: (2024)
di: Tharwani, Jay, et al.
Pubblicazione: (2024)
Modeling Tradeoffs between mobility, cost, and performance in Edge Computing
di: Waseem, Muhammad Danish, et al.
Pubblicazione: (2026)
di: Waseem, Muhammad Danish, et al.
Pubblicazione: (2026)
Pinching-Antenna Systems For Indoor Immersive Communications: A 3D-Modeling Based Performance Analysis
di: Wang, Yulei, et al.
Pubblicazione: (2025)
di: Wang, Yulei, et al.
Pubblicazione: (2025)
Redundant Array Computation Elimination
di: Wang, Zixuan, et al.
Pubblicazione: (2025)
di: Wang, Zixuan, et al.
Pubblicazione: (2025)
PipeWeave: Synergizing Analytical and Learning Models for Unified GPU Performance Prediction
di: Zhang, Kaixuan, et al.
Pubblicazione: (2026)
di: Zhang, Kaixuan, et al.
Pubblicazione: (2026)
Forecasting GPU Performance for Deep Learning Training and Inference
di: Lee, Seonho, et al.
Pubblicazione: (2024)
di: Lee, Seonho, et al.
Pubblicazione: (2024)
Towards Computational Performance Engineering for Unsupervised Concept Drift Detection -- Complexities, Benchmarking, Performance Analysis
di: Werner, Elias, et al.
Pubblicazione: (2023)
di: Werner, Elias, et al.
Pubblicazione: (2023)
Modeling Utilization to Identify Shared-Memory Atomic Bottlenecks
di: Dong, Rongcui, et al.
Pubblicazione: (2025)
di: Dong, Rongcui, et al.
Pubblicazione: (2025)
LCS.jl: A High-Performance, Multi-Platform Computational Model in Julia for Turbulent Particle-Laden Flows
di: Tominaga, Taketo, et al.
Pubblicazione: (2026)
di: Tominaga, Taketo, et al.
Pubblicazione: (2026)
MoE-Inference-Bench: Performance Evaluation of Mixture of Expert Large Language and Vision Models
di: Chitty-Venkata, Krishna Teja, et al.
Pubblicazione: (2025)
di: Chitty-Venkata, Krishna Teja, et al.
Pubblicazione: (2025)
Accurate Performance Modeling And Uncertainty Analysis of Lossy Compression in Scientific Applications
di: Liu, Youyuan, et al.
Pubblicazione: (2024)
di: Liu, Youyuan, et al.
Pubblicazione: (2024)
denet, A lightweight command-line tool for process monitoring in benchmarking and beyond
di: Carrillo, Ben, et al.
Pubblicazione: (2025)
di: Carrillo, Ben, et al.
Pubblicazione: (2025)
Profiling Apple Silicon Performance for ML Training
di: Feng, Dahua, et al.
Pubblicazione: (2025)
di: Feng, Dahua, et al.
Pubblicazione: (2025)
AI Application Benchmarking: Power-Aware Performance Analysis for Vision and Language Models
di: Mayr, Martin, et al.
Pubblicazione: (2026)
di: Mayr, Martin, et al.
Pubblicazione: (2026)
Evaluating Compiler Optimization Impacts on zkVM Performance
di: Gassmann, Thomas, et al.
Pubblicazione: (2025)
di: Gassmann, Thomas, et al.
Pubblicazione: (2025)
Achieving Consistent and Comparable CPU Evaluation
di: Wang, Chenxi, et al.
Pubblicazione: (2024)
di: Wang, Chenxi, et al.
Pubblicazione: (2024)
V-Seek: Accelerating LLM Reasoning on Open-hardware Server-class RISC-V Platforms
di: Rodrigo, Javier J. Poveda, et al.
Pubblicazione: (2025)
di: Rodrigo, Javier J. Poveda, et al.
Pubblicazione: (2025)
MambaCPU: Enhanced Correlation Mining with State Space Models for CPU Performance Prediction
di: Liu, Xiaoman
Pubblicazione: (2024)
di: Liu, Xiaoman
Pubblicazione: (2024)
Wasure: A Modular Toolkit for Comprehensive WebAssembly Benchmarking
di: Carissimi, Riccardo, et al.
Pubblicazione: (2026)
di: Carissimi, Riccardo, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Memory Analysis on the Training Course of DeepSeek Models
di: Zhang, Ping, et al.
Pubblicazione: (2025) -
Performance of Confidential Computing GPUs
di: Ibarra, Antonio Martínez, et al.
Pubblicazione: (2025) -
Confidential Computing on NVIDIA Hopper GPUs: A Performance Benchmark Study
di: Zhu, Jianwei, et al.
Pubblicazione: (2024) -
DeepSeek-R1 Outperforms Gemini 2.0 Pro, OpenAI o1, and o3-mini in Bilingual Complex Ophthalmology Reasoning
di: Xu, Pusheng, et al.
Pubblicazione: (2025) -
A performance analysis of VM-based Trusted Execution Environments for Confidential Federated Learning
di: Casella, Bruno
Pubblicazione: (2025)