JetBrains / Mellum2-12B-A2.5B-Thinking-SFT

huggingface.co
Total runs: 365
24-hour runs: 0
7-day runs: 44
30-day runs: -8
Model's Last Updated: August 13 2026
text-generation

Introduction of Mellum2-12B-A2.5B-Thinking-SFT

Model Details of Mellum2-12B-A2.5B-Thinking-SFT

Mellum

Mellum2 Thinking (SFT)

Research artifact: the SFT-only intermediate checkpoint of the Thinking training pipeline. Use it to study or build on the reasoning SFT stage in isolation, run post-training experiments (preference tuning, RLVR, etc.), or compare against the final RL-tuned variant. For production-quality reasoning use Thinking instead.

Mellum2 Thinking (SFT) Highlights

Mellum2 Thinking (SFT) is the supervised-fine-tuning-only intermediate checkpoint of the Mellum2 Thinking training pipeline, released by JetBrains as a research artifact.

The model uses a Mixture-of-Experts architecture with 64 experts and activates 8 experts per token. It uses a combination of sliding-window and full attention layers, with a context length of 131,072 tokens.

This checkpoint was produced from Mellum2-12B-A2.5B-Base with three epochs of SFT on a corpus of reasoning-trace-bearing data. The model emits its reasoning inside <think>...</think> blocks. It is the starting point of the RLVR stage that produces the final Mellum2-12B-A2.5B-Thinking .

Mellum2 Model Family

This repository contains one checkpoint from the Mellum2 family.

Checkpoint Description
Base Pretrain Base checkpoint before long-context extension
Base Final base model
Instruct SFT Supervised instruction-tuned checkpoint
Thinking SFT Supervised thinking checkpoint
Instruct RL-tuned instruction model
Thinking RL-tuned thinking model
Model Overview

Mellum2 Thinking (SFT) has the following features:

  • Number of Layers: 28
  • Hidden Size: 2304
  • Intermediate Size: 7168
  • MoE Intermediate Size: 896
  • Number of Experts: 64
  • Number of Activated Experts: 8
  • Number of Attention Heads (GQA): 32 for Q and 4 for KV
  • Context Length: 131,072
  • Sliding Window: 1,024
  • Vocabulary Size: 98,304
  • Precision: bfloat16
Serving with vLLM
# Without tool calling
vllm serve JetBrains/Mellum2-12B-A2.5B-Thinking-SFT \
  --max-model-len 131072 \
  --reasoning-parser qwen3

# With tool calling
vllm serve JetBrains/Mellum2-12B-A2.5B-Thinking-SFT \
  --max-model-len 131072 \
  --reasoning-parser qwen3 \
  --enable-auto-tool-choice \
  --tool-call-parser hermes
Quickstart

Text-Only Input

from openai import OpenAI
# Configured by environment variables
client = OpenAI()

messages = [
    {"role": "user", "content": "Is 1024 a power of 2? Explain your reasoning."},
]

chat_response = client.chat.completions.create(
    model="JetBrains/Mellum2-12B-A2.5B-Thinking-SFT",
    messages=messages,
    max_tokens=81920,
    temperature=0.6,
    top_p=0.95,
    extra_body={
        "top_k": 20,
    },
)
print("Chat response:", chat_response)
Evaluation

Evaluation results are available in the model card. All values are self-reported by JetBrains.

For more details, see the Mellum2 Technical Report .

License

Released under the Apache 2.0 license.

Runs of JetBrains Mellum2-12B-A2.5B-Thinking-SFT on huggingface.co

365
Total runs
0
24-hour runs
5
3-day runs
44
7-day runs
-8
30-day runs

More Information About Mellum2-12B-A2.5B-Thinking-SFT huggingface.co Model

More Mellum2-12B-A2.5B-Thinking-SFT license Visit here:

https://choosealicense.com/licenses/apache-2.0

Mellum2-12B-A2.5B-Thinking-SFT huggingface.co

Mellum2-12B-A2.5B-Thinking-SFT huggingface.co is an AI model on huggingface.co that provides Mellum2-12B-A2.5B-Thinking-SFT's model effect (), which can be used instantly with this JetBrains Mellum2-12B-A2.5B-Thinking-SFT model. huggingface.co supports a free trial of the Mellum2-12B-A2.5B-Thinking-SFT model, and also provides paid use of the Mellum2-12B-A2.5B-Thinking-SFT. Support call Mellum2-12B-A2.5B-Thinking-SFT model through api, including Node.js, Python, http.

Mellum2-12B-A2.5B-Thinking-SFT huggingface.co Url

https://huggingface.co/JetBrains/Mellum2-12B-A2.5B-Thinking-SFT

JetBrains Mellum2-12B-A2.5B-Thinking-SFT online free

Mellum2-12B-A2.5B-Thinking-SFT huggingface.co is an online trial and call api platform, which integrates Mellum2-12B-A2.5B-Thinking-SFT's modeling effects, including api services, and provides a free online trial of Mellum2-12B-A2.5B-Thinking-SFT, you can try Mellum2-12B-A2.5B-Thinking-SFT online for free by clicking the link below.

JetBrains Mellum2-12B-A2.5B-Thinking-SFT online free url in huggingface.co:

https://huggingface.co/JetBrains/Mellum2-12B-A2.5B-Thinking-SFT

Mellum2-12B-A2.5B-Thinking-SFT install

Mellum2-12B-A2.5B-Thinking-SFT is an open source model from GitHub that offers a free installation service, and any user can find Mellum2-12B-A2.5B-Thinking-SFT on GitHub to install. At the same time, huggingface.co provides the effect of Mellum2-12B-A2.5B-Thinking-SFT install, users can directly use Mellum2-12B-A2.5B-Thinking-SFT installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Mellum2-12B-A2.5B-Thinking-SFT install url in huggingface.co:

https://huggingface.co/JetBrains/Mellum2-12B-A2.5B-Thinking-SFT

Url of Mellum2-12B-A2.5B-Thinking-SFT

Mellum2-12B-A2.5B-Thinking-SFT huggingface.co Url

Provider of Mellum2-12B-A2.5B-Thinking-SFT huggingface.co

JetBrains
ORGANIZATIONS

Other API from JetBrains