Published event
ArtificialIntelligence Research 1 source(s)

Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains

Updated September 26, 2026 · 2:44 PM · source date June 1, 2026

Summary

Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains Team Article Published June 1, 2026 Upvote 40 Nikita Pavlichenko pavlichenko JetBrains Mellum2 is a 12B-parameter Mixture-of-Experts model trained from scratch on natural language and code. The model activates only 2.5B parameters per token, making it efficient for high-throughput, low-latency inference.

Why it matters

This Research is relevant to the technology intelligence record because it involves Hugging Face. The source article should remain the factual reference for follow-up coverage.

Key facts
  • Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains Team Article Published June 1, 2026 Upvote 40 Nikita Pavlichenko pavlichenko JetBrains Mellum2 is a 12B-parameter Mixture-of-Experts model trained from scratch on natural language and code.
  • The model activates only 2.5B parameters per token, making it efficient for high-throughput, low-latency inference.
  • Mellum2 is can be used for routing, RAG, summarization, sub-agents, high-throughput coding features, and private deployments.
  • It is released under the Apache 2.0 license.
  • Compared with similar-sized models, Mellum2 delivers competitive benchmark performance while achieving more than 2x faster inference.
  • Download the model on Hugging Face: https://huggingface.co/collections/JetBrains/mellum-2 For architecture details, training setup, benchmarks, and evaluation methodology, read the full technical report: https://arxiv.org/pdf/2605.31268 Today we’re releasing Mellum2, an open Mixture-of-Experts model optimized for low-latency text-and-code workloads.
Entities in this story
Related events