DeepSeek V3

Dec '24

Creator

DeepSeek

Released

2024-12-26

Intelligence

14.2

Artificial Analysis Index

Coding

23.0

Artificial Analysis Index

In $/1M

$0.36

input tokens

Out $/1M

$0.89

output tokens

Blended $/1M

$0.49

3:1 blended

Speed

0

tokens / sec

Profile

License, openness, and modality — the governance layer.

License

MIT + Model License (Commercial use allowed)

non-commercial

Weights

Open

downloadable

Modalities

Text

Parameters

671B

total

A powerful Mixture-of-Experts (MoE) language model with 671B total parameters (37B activated per token). Features Multi-head Latent Attention (MLA), auxiliary-loss-free load balancing, and multi-token prediction training. Pre-trained on 14.8T tokens with strong performance in reasoning, math, and code tasks.

Capability profile

Category strength across 20 domains, via LLM Stats.

instruction following
90legal
80finance
80language
80healthcare
80math
70general
70reasoning
70
structured output
70biology
60physics
60chemistry
60code
50long context
50frontend development
40factuality
20
Via LLM Stats

Benchmark breakdown

Independent evaluation scores, normalized to 0–100.

GPQA Diamond
56
Humanity’s Last Exam
4
MMLU-Pro
75
SciCode
35
LiveCodeBench
36
MATH-500
89
AIME 2025
26
τ²-Bench (agentic)
23
Terminal-Bench Hard
7
IFBench
35
Via Artificial Analysis
© 2026 NYSGPT2525 LLC