Rigel is a 2.3B-total, 360M-active-parameter hybrid Mamba-2 + attention mixture-of-experts base model, trained with
LM Engine
. It used about 117× fewer pretraining FLOPs than Llama-3.2-1B/3B and 53× fewer than Granite-4.2-3B, while landing within a few points of Granite-4.2-3B and Llama-3.2-3B on zero-shot averages and above Llama-3.2-1B.
This is a base model: it has no chat template and is not instruction-tuned.
Usage with LM Engine
Rigel's architecture lives in LM Engine. Installing it and importing
lm_engine.training
registers the architecture with
transformers
'
Auto
classes, after which the usual Hugging Face API works.
git clone https://github.com/open-lm-engine/lm-engine.git && cd lm-engine && uv sync --extra cuda
import torch
import lm_engine.training # registers Rigel's architecture with transformers' Auto classesfrom lm_engine.training.kernels import Kernel, enable_kernels
from transformers import AutoModelForCausalLM, AutoTokenizer
model_id = "open-lm-engine/rigel-base"
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(model_id, dtype=torch.bfloat16, device_map={"": "cuda"}).eval()
inputs = tokenizer("Rigel is a blue supergiant star in the constellation", return_tensors="pt").to("cuda")
with torch.no_grad(), enable_kernels([Kernel.mamba2_ssm, Kernel.causal_conv1d, Kernel.sonicmoe, Kernel.flash_attention_3]):
output_ids = model.generate(**inputs, max_new_tokens=64, do_sample=False)
print(tokenizer.decode(output_ids[0], skip_special_tokens=True))
Any kernel left out of the list falls back to a pure-PyTorch implementation, so the snippet also runs with an empty list, just slower.
To finetune or continue pretraining, point an LM Engine training config at this checkpoint; see the
LM Engine README
.
Citation
@misc{mishra2026rigel,
title = {Rigel Base: Reaching Llama-3.2 Quality with <1% of its Compute},
author = {Mishra, Mayank and Runwal, Bharat and Stoica, Ion and Dao, Tri and Gonzalez, Joseph E.},
year = {2026},
url = {https://open-lm-engine.github.io/blog/rigel/}
}
Runs of open-lm-engine rigel-base on huggingface.co
529
Total runs
24
24-hour runs
78
3-day runs
109
7-day runs
306
30-day runs
More Information About rigel-base huggingface.co Model
rigel-base huggingface.co
rigel-base huggingface.co is an AI model on huggingface.co that provides rigel-base's model effect (), which can be used instantly with this open-lm-engine rigel-base model. huggingface.co supports a free trial of the rigel-base model, and also provides paid use of the rigel-base. Support call rigel-base model through api, including Node.js, Python, http.
rigel-base huggingface.co is an online trial and call api platform, which integrates rigel-base's modeling effects, including api services, and provides a free online trial of rigel-base, you can try rigel-base online for free by clicking the link below.
open-lm-engine rigel-base online free url in huggingface.co:
rigel-base is an open source model from GitHub that offers a free installation service, and any user can find rigel-base on GitHub to install. At the same time, huggingface.co provides the effect of rigel-base install, users can directly use rigel-base installed effect in huggingface.co for debugging and trial. It also supports api for free installation.