Skip to content
r/LocalLLaMA

DeepSeek releases V4: 1.6T Pro, 284B Flash, MIT license, 1M context

DeepSeek V4 just dropped, 1.6T Pro and 284B Flash, MIT license, 1M context. This is huge.

DeepSeek released two open-weight V4 models: Pro at 1.6T total with 49B active, and Flash at 284B total with 13B active; both use an MIT license and support 1M context. The RSS snippet points to a Hugging Face collection and a tech report, but the post does not disclose benchmark scores, pricing, training data size, or real inference throughput. The key thing to watch is the 1M context plus low active-parameter ratio; if evals hold, self-hosted long-context and routing economics change materially.

Why it matters: HKR-H/K/R all pass: this is a flagship DeepSeek open release with two huge MIT-licensed weights and 1M context, strong enough for same-day coverage. The score stops at 86 because the provided text does not disclose benchmarks, throughput, training data, or pricing.

Read the original ↗Export Markdown