Skip to content
StableTechnologyReported 2026-06-01 23:45

JetBrains Introduces Mellum2: A 12B Parameter Mixture-of-Experts Model

JetBrains has announced Mellum2, a 12B-parameter Mixture-of-Experts model designed for efficient high-throughput, low-latency inference in natural language and code processing.

01

Evidence

  • HHugging Face BlogCompany2026-06-01 23:45
    Mellum2 is a 12B-parameter Mixture-of-Experts model trained from scratch on natural language and code.
    View source
  • HHugging Face BlogCompany2026-06-01 23:45
    The model activates only 2.5B parameters per token, making it efficient for high-throughput, low-latency inference.
    View source
  • HHugging Face BlogCompany2026-06-01 23:45
    Mellum2 is a 12B-parameter Mixture-of-Experts model trained from scratch on natural language and code.
    View source
  • HHugging Face BlogCompany2026-06-01 23:45
    Today we’re releasing Mellum2, an open Mixture-of-Experts model optimized for low-latency text-and-code workloads. Mellum originally started as a code completion model. With Mellum2, we extend that foundation to a broader set of natural language and software engineering tasks wh…
    View source