Introducing Mellum2: JetBrains' 12B Mixture-of-Experts Model
JetBrains published a Hugging Face blog post introducing Mellum2, confirming a mixture-of-experts architecture and a 12B parameter scale; the snippet does not disclose training data, license, benchmarks, or deployment conditions.
Why it matters: HKR-H/K/R all pass, but the body only confirms 12B and MoE, with no benchmarks, license, context window, or IDE integration terms. Treat as a mid-weight model release at the lower featured band.