zjunlp / KnowRL-Skywork-OR1-7B-Preview

huggingface.co
Total runs: 9
24-hour runs: 0
7-day runs: 0
30-day runs: 7
Model's Last Updated: November 29 2025

Introduction of KnowRL-Skywork-OR1-7B-Preview

Model Details of KnowRL-Skywork-OR1-7B-Preview

KnowRL

Exploring Knowledgeable Reinforcement Learning for Factuality

📄arXiv • 💻GitHub Repo • 📖Dataset


Model Description

KnowRL-Skywork-OR1-7B-Preview is a slow-thinking language model that results from applying our KnowRL framework to the base model Skywork-OR1-7B-Preview .

The KnowRL (Knowledgeable Reinforcement Learning) framework is designed to mitigate hallucinations in Large Language Models (LLMs) by integrating external knowledge directly into the training process. This model undergoes a two-stage training process:

  1. Cold-Start Supervised Fine-Tuning (SFT) : The model first aligns with factual thinking patterns on a high-quality dataset.
  2. Knowledgeable Reinforcement Learning (RL) : The model is then further trained using a reward signal that explicitly encourages factual accuracy in its reasoning process, helping it learn its own knowledge boundaries.

As a result, this model demonstrates a significant reduction in hallucinations on factual benchmarks while preserving or even enhancing the strong reasoning capabilities inherited from its base model.

How to Use
Using the transformers Library

You can use this model with the transformers library for text generation tasks. It is important to follow the specific prompt format, which includes <think> and <answer> tags, to get the best results.

import torch
from transformers import AutoModelForCausalLM, AutoTokenizer

# Set the device
device = "cuda" if torch.cuda.is_available() else "cpu"

# Load the model and tokenizer
model_name = "zjunlp/KnowRL-Skywork-OR1-7B-Preview"
tokenizer = AutoTokenizer.from_pretrained(model_name)
model = AutoModelForCausalLM.from_pretrained(model_name, torch_dtype=torch.bfloat16).to(device)

# Define the prompt using the model's template
prompt = "What is the main function of the mitochondria?"
messages = [
    {"role": "user", "content": prompt}
]
text = tokenizer.apply_chat_template(messages, tokenize=False, add_generation_prompt=True)

# Generate a response
inputs = tokenizer(text, return_tensors="pt").to(device)
outputs = model.generate(**inputs, max_new_tokens=512)

# Decode and print the output
response = tokenizer.decode(outputs[0], skip_special_tokens=True)
print(response)
Using huggingface-cli

You can also download the model from the command line using huggingface-cli .

huggingface-cli download zjunlp/KnowRL-Skywork-OR1-7B-Preview --local-dir KnowRL-Skywork-OR1-7B-Preview
Training Details

The model's training process involves two distinct stages, using the data from the zjunlp/KnowRL-Train-Data dataset.

  • Stage 1: Cold-Start SFT : The base model undergoes supervised fine-tuning on the knowrl_coldstart.json dataset. This stage helps the model adopt a fact-based, slow-thinking response structure.
  • Stage 2: Knowledgeable RL : The SFT-tuned model is further trained using reinforcement learning (GRPO). The reward function combines a correctness reward with a factuality reward, which is calculated by verifying the model's thinking process against an external knowledge base. This stage uses the knowrl_RLdata.json and KnowRL_RLtrain_data_withknowledge.json files.

For complete details on the training configuration and hyperparameters, please refer to our GitHub repository .


Citation

If you find this model useful in your research, please consider citing our paper:

@article{ren2025knowrl,
  title={{KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality}}, 
  author={Ren, Baochang and Qiao, Shuofei and Yu, Wenhao and Chen, Huajun and Zhang, Ningyu},
  journal={arXiv preprint arXiv:2506.19807},
  year={2025}
}

Runs of zjunlp KnowRL-Skywork-OR1-7B-Preview on huggingface.co

9
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
7
30-day runs

More Information About KnowRL-Skywork-OR1-7B-Preview huggingface.co Model

More KnowRL-Skywork-OR1-7B-Preview license Visit here:

https://choosealicense.com/licenses/mit

KnowRL-Skywork-OR1-7B-Preview huggingface.co

KnowRL-Skywork-OR1-7B-Preview huggingface.co is an AI model on huggingface.co that provides KnowRL-Skywork-OR1-7B-Preview's model effect (), which can be used instantly with this zjunlp KnowRL-Skywork-OR1-7B-Preview model. huggingface.co supports a free trial of the KnowRL-Skywork-OR1-7B-Preview model, and also provides paid use of the KnowRL-Skywork-OR1-7B-Preview. Support call KnowRL-Skywork-OR1-7B-Preview model through api, including Node.js, Python, http.

KnowRL-Skywork-OR1-7B-Preview huggingface.co Url

https://huggingface.co/zjunlp/KnowRL-Skywork-OR1-7B-Preview

zjunlp KnowRL-Skywork-OR1-7B-Preview online free

KnowRL-Skywork-OR1-7B-Preview huggingface.co is an online trial and call api platform, which integrates KnowRL-Skywork-OR1-7B-Preview's modeling effects, including api services, and provides a free online trial of KnowRL-Skywork-OR1-7B-Preview, you can try KnowRL-Skywork-OR1-7B-Preview online for free by clicking the link below.

zjunlp KnowRL-Skywork-OR1-7B-Preview online free url in huggingface.co:

https://huggingface.co/zjunlp/KnowRL-Skywork-OR1-7B-Preview

KnowRL-Skywork-OR1-7B-Preview install

KnowRL-Skywork-OR1-7B-Preview is an open source model from GitHub that offers a free installation service, and any user can find KnowRL-Skywork-OR1-7B-Preview on GitHub to install. At the same time, huggingface.co provides the effect of KnowRL-Skywork-OR1-7B-Preview install, users can directly use KnowRL-Skywork-OR1-7B-Preview installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

KnowRL-Skywork-OR1-7B-Preview install url in huggingface.co:

https://huggingface.co/zjunlp/KnowRL-Skywork-OR1-7B-Preview

Url of KnowRL-Skywork-OR1-7B-Preview

KnowRL-Skywork-OR1-7B-Preview huggingface.co Url

Provider of KnowRL-Skywork-OR1-7B-Preview huggingface.co

zjunlp
ORGANIZATIONS

Other API from zjunlp

huggingface.co

Total runs: 1.1K
Run Growth: 363
Growth Rate: 33.21%
Updated:March 04 2024
huggingface.co

Total runs: 589
Run Growth: -601
Growth Rate: -102.04%
Updated:March 04 2024
huggingface.co

Total runs: 36
Run Growth: -50
Growth Rate: -138.89%
Updated:May 06 2024
huggingface.co

Total runs: 16
Run Growth: -14
Growth Rate: -87.50%
Updated:March 21 2023
huggingface.co

Total runs: 11
Run Growth: 0
Growth Rate: 0.00%
Updated:October 01 2025
huggingface.co

Total runs: 10
Run Growth: 3
Growth Rate: 30.00%
Updated:October 01 2025
huggingface.co

Total runs: 3
Run Growth: 0
Growth Rate: 0.00%
Updated:July 28 2023
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:December 22 2022
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:December 07 2024
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:June 09 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:May 10 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:February 17 2023