Open-Source AI Model

DeepSeek V3

Developed by DeepSeek

Local AI Deployment Experts 24+ Years IT Infrastructure GPU Hardware In Stock

Key Capabilities

  • GPT-4 class performance at open-source level
  • Highly efficient MoE - only 37B parameters active per token
  • 128K context window
  • Strong math, coding, and Chinese language capabilities
  • Multi-head Latent Attention reduces KV cache by 93%

VRAM Requirements by Quantization

Choose the right GPU based on your performance and quality needs.

Model / QuantizationVRAM Required
full FP161.3TB+
Q4350GB
Q2170GB

Use Cases

DeepSeek V3 (671B total (37B active via MoE)) can be deployed for enterprise AI applications including document processing, code generation, data analysis, and conversational AI. License: MIT License.

Defense contractors: Section 1532 of the FY2026 NDAA (Public Law 119-60) bars contractors from using AI developed by DeepSeek or High Flyer in the performance of a Department of Defense contract, effective January 17, 2026, and the enacted text has no exemption for models run on your own hardware. Read DeepSeek and the FY2026 NDAA: what DoD contractors need to know, and see the American open-weight alternatives we recommend: Gemma 4 and GPT-OSS 20B and 120B.

Run DeepSeek V3 with Petronella

Petronella builds DeepSeek V3 inference clusters that deliver GPT-4 class performance with zero API costs. The MIT license and MoE efficiency make it the most cost-effective frontier model to self-host.

Recommended Hardware

Model SizeRecommended GPU
full FP16DGX B300 (2.3TB HBM3e) or multi-node cluster
Q4 quantizedDGX Station GB300 (384GB) or 4x RTX PRO 6000 (384GB)
Q2 quantizedDGX Spark (128GB) or 2x RTX PRO 6000 (192GB)

Deploy DeepSeek V3 On-Premises

Our team builds GPU-accelerated systems configured and optimized for DeepSeek V3. Private, secure, and fully under your control.