GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →

DeepSeek cloud models

34 models tracked via Artificial Analysis. Compare cloud performance, then find local GGUF versions in the GraySoft model catalog.

ModelIntelligenceSpeed (tok/s)
DeepSeek V4 Pro 0813 (Reasoning, Max Effort)42.159.459
DeepSeek V4 Flash Vision (Reasoning, Max Effort)41.5121.308
DeepSeek V4 Flash 0731 (Reasoning, Max Effort)40.8118.418
DeepSeek V4 Pro (Reasoning, Max Effort)35.861.264
DeepSeek V4 Pro (Reasoning, High Effort)34.662.907
DeepSeek V4 Flash (Reasoning, Max Effort)33.10
DeepSeek V4 Flash (Reasoning, High Effort)30.40
DeepSeek V3.2 (Reasoning)25.30
DeepSeek V4 Pro (Non-reasoning)24.559.136
DeepSeek V3.1 Terminus (Reasoning)23.50
DeepSeek V4 Flash (Non-reasoning)22.10
DeepSeek V3.2 Exp (Reasoning)190
DeepSeek V3.2 (Non-reasoning)18.30
DeepSeek V3.2 Speciale160
DeepSeek V3.1 Terminus (Non-reasoning)15.20
DeepSeek V3.2 Exp (Non-reasoning)15.10
DeepSeek V3.1 (Non-reasoning)14.80
DeepSeek V3.1 (Reasoning)14.50
DeepSeek R1 0528 (May '25)13.90
DeepSeek R1 (Jan '25)120
DeepSeek V3 03249.20
DeepSeek V3 (Dec '24)8.30
DeepSeek R1 Distill Qwen 32B5.30
DeepSeek R1 0528 Qwen3 8B4.70
DeepSeek R1 Distill Llama 70B4.20
DeepSeek R1 Distill Qwen 14B4.10
DeepSeek-V2.5 (Dec '24)1.20
DeepSeek-V2.51.10
DeepSeek-Coder-V210
DeepSeek R1 Distill Llama 8B10
DeepSeek LLM 67B Chat (V1)10
DeepSeek R1 Distill Qwen 1.5B10
DeepSeek Coder V2 Lite Instruct10
DeepSeek-V2-Chat10

Run models locally with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models