DeepSeek: DeepSeek R1 0528 Qwen3 8B

DeepSeek deepseek-r1-0528-qwen3-8b

Model Information
Slug deepseek-r1-0528-qwen3-8b
Aliases deepseek-r1-0528-qwen3-8b
Organization
Name DeepSeek
Description

DeepSeek-R1-0528 is a lightly upgraded release of DeepSeek R1 that taps more compute and smarter post-training tricks, pushing its reasoning and inference to the brink of flagship models like O3 and Gemini 2.5 Pro. It now tops math, programming, and logic leaderboards, showcasing a step-change in depth-of-thought. The distilled variant, DeepSeek-R1-0528-Qwen3-8B, transfers this chain-of-thought into an 8 B-parameter form, beating standard Qwen3 8B by +10 pp and tying the 235 B “thinking” giant on AIME 2024. Context: 128000

Available at 1 Provider
Provider Input Price ($/1M) Output Price ($/1M) Free
OpenRouter $0.06 $0.09