JetBrains Introduces Mellum2: A 12B Parameter Mixture-of-Experts Model
JetBrains has announced Mellum2, a 12B-parameter Mixture-of-Experts model designed for efficient high-throughput, low-latency inference in natural language and code processing.
JetBrains has announced Mellum2, a 12B-parameter Mixture-of-Experts model designed for efficient high-throughput, low-latency inference in natural language and code processing.
Mellum2 is a 12B-parameter Mixture-of-Experts model trained from scratch on natural language and code.
The model activates only 2.5B parameters per token, making it efficient for high-throughput, low-latency inference.
Mellum2 is a 12B-parameter Mixture-of-Experts model trained from scratch on natural language and code.
Today we’re releasing Mellum2, an open Mixture-of-Experts model optimized for low-latency text-and-code workloads. Mellum originally started as a code completion model. With Mellum2, we extend that foundation to a broader set of natural language and software engineering tasks wh…