โ€” LLM (Open Source)

DeepSeek-V3

Last updated June 23, 2026 ยทReviewed by ToolForge Editorial

GPT-4o performance at ~5% of the cost. The most disruptive open-source LLM of 2026.

โ˜… 4.7/5ยทProduction-readyยทReleased Dec 2024 (V3.2)ยทAPI: $0.14/M input
API: $0.14/M tokens
Read full review See verdict

The open-source LLM that flipped the script.

DeepSeek-V3 is a 671B-parameter mixture-of-experts model from Chinese AI lab DeepSeek. Despite having 37B active parameters per token, it matches GPT-4o and Claude 3.5 Sonnet on most reasoning, coding, and math benchmarks โ€” and costs ~95% less via API. The model is fully open-source under MIT license, meaning you can self-host, fine-tune, or distill it for commercial use without restriction.

Who it's for: Developers shipping LLM features, startups watching their OpenAI bill, enterprises needing on-prem deployment, researchers benchmarking open models.

Key features

671B MoEMixture of experts

671B total parameters but only 37B active per forward pass โ€” gives GPT-4-class quality at Llama-3-70B inference cost. Trained on 14.8T tokens.

128K ctxLong context window

128K token context window via YaRN scaling. Handles full codebases, long legal documents, and entire books in one prompt.

MIT licenseTruly open

Weights, code, and training data documentation all released under MIT license. Self-host, fine-tune commercially, distill โ€” no restrictions.

$0.14/MCheapest frontier API

API pricing: $0.14/M input, $0.28/M output. Compare: GPT-4o ($2.50/$10), Claude 3.5 Sonnet ($3/$15). 95%+ savings at scale.

DistilledLlama & Qwen variants

DeepSeek-R1-Distill versions fine-tuned into Llama 3.1 8B/70B and Qwen 2.5 1.5B/7B/14B/32B โ€” runs on consumer GPUs.

The honest take

โœ“ What works

  • Benchmark parity with GPT-4o and Claude 3.5 Sonnet on coding (HumanEval), math (MATH), and Chinese/English reasoning (MMLU, C-Eval)
  • API is 95%+ cheaper than OpenAI/Anthropic โ€” disrupts the entire economics of LLM apps
  • MIT license: fine-tune, self-host, distill, redistribute โ€” no OpenAI-style restrictions
  • Distilled variants (1.5B-70B) let you run frontier-quality on a MacBook
  • Fast inference: 60+ tokens/sec on H100, 18 tokens/sec on M2 Ultra with MLX
  • OpenAI-compatible API โ€” drop-in replacement for most apps

โœ— What doesn't

  • Chinese government origin raises data privacy and IP concerns for some enterprises
  • Self-hosting 671B requires 8x H100s minimum โ€” not for hobbyists without hardware
  • Safety tuning differs from Western models โ€” will discuss topics GPT-4 refuses
  • Censorship on China-specific political topics (Tiananmen, Xi Jinping, Taiwan)
  • English prose still slightly weaker than Claude โ€” wins on code, math, and reasoning but loses on creative writing

Verdict

DeepSeek-V3 is the most important open-source LLM release since Llama 2. If you're building an LLM-powered product and you're not using it yet, you're likely overpaying 10-20x for OpenAI. The API is a no-brainer for cost-sensitive startups. Self-hosting requires serious hardware but gives you total control. Just be aware of the geopolitical reality and test for your use case.

๐Ÿ’ก Transparency: This review is editorially independent. We never accept payment for positive coverage. Full disclosure.

Related Tools

Try DeepSeek-V3 today

API: $0.14/M input ยท Free chat at chat.deepseek.com

Browse all tools โ†’