We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.
Olmo is a series of
O
pen
l
anguage
mo
dels designed to enable the science of language models.
These models are pre-trained on the Dolma 3 dataset and post-trained on the Dolci datasets. We are releasing all code, checkpoints, logs (coming soon), and associated training details.
The RL-Zero family of models is an experimental set of model for the scientific exploration of RLVR training.
The quantized model is more sensitive to data types and CUDA operations. To avoid potential issues, it's recommended to pass the inputs directly to CUDA using:
inputs.input_ids.to('cuda')
Fine-tuning
Model fine-tuning can be done from the final checkpoint (the
main
revision of this model) or the base model.
We recommend fine-tuning with the open-instruct repository:
bash ./scripts/train/olmo3/rlvr_script.sh
You can override most configuration options from the command-line. For example, to override the learning rate you could launch the script like this:
Model type:
a Transformer style autoregressive language model.
Language(s) (NLP):
English
License:
This model is licensed under Apache 2.0. It is intended for research and educational use in accordance with Ai2's
Responsible Use Guidelines
.
Like any base language model or fine-tuned model without safety filtering, these models can easily be prompted by users to generate harmful and sensitive content. Such content may also be produced unintentionally, especially in cases involving bias, so we recommend that users consider the risks when applying this technology. Additionally, many statements from OLMo or any LLM are often inaccurate, so facts should be verified.
Olmo-3.1-7B-RL-Zero-Code huggingface.co is an AI model on huggingface.co that provides Olmo-3.1-7B-RL-Zero-Code's model effect (), which can be used instantly with this allenai Olmo-3.1-7B-RL-Zero-Code model. huggingface.co supports a free trial of the Olmo-3.1-7B-RL-Zero-Code model, and also provides paid use of the Olmo-3.1-7B-RL-Zero-Code. Support call Olmo-3.1-7B-RL-Zero-Code model through api, including Node.js, Python, http.
Olmo-3.1-7B-RL-Zero-Code huggingface.co is an online trial and call api platform, which integrates Olmo-3.1-7B-RL-Zero-Code's modeling effects, including api services, and provides a free online trial of Olmo-3.1-7B-RL-Zero-Code, you can try Olmo-3.1-7B-RL-Zero-Code online for free by clicking the link below.
allenai Olmo-3.1-7B-RL-Zero-Code online free url in huggingface.co:
Olmo-3.1-7B-RL-Zero-Code is an open source model from GitHub that offers a free installation service, and any user can find Olmo-3.1-7B-RL-Zero-Code on GitHub to install. At the same time, huggingface.co provides the effect of Olmo-3.1-7B-RL-Zero-Code install, users can directly use Olmo-3.1-7B-RL-Zero-Code installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
Olmo-3.1-7B-RL-Zero-Code install url in huggingface.co: