Figure: Effectiveness of
General-Reasoner
trained with diverse verifiable reasoning questions using model-based verifier compared to baseline methods on various reasoning tasks.
General-Reasoner
is a training paradigm for large language models (LLMs), designed to robustly enhance reasoning abilities across diverse domains—not just mathematics and coding, but also physics, chemistry, finance, humanities, and more.
Key features:
Zero RL Training:
Direct reinforcement learning from base LLMs, bypassing intermediate supervised stages.
Diverse Reasoning Data:
230K+ high-quality, verifiable questions sourced from the web and filtered for answer verifiability across disciplines.
Model-Based Verifier:
Compact 1.5B generative verifier model for context-aware, chain-of-thought answer validation, outperforming traditional rule-based methods.
This specific model is the General-Reasoner variant trained based on
Qwen2.5-7B-Base
.
Main Results
General-Reasoner outperforms base and supervised models on a variety of reasoning benchmarks, demonstrating robust generalization across domains:
Citation
If you feel our work is helpful, please cite:
@article{general-reasoner,
title={{G}eneral-{R}easoner: Advancing LLM Reasoning Across All Domains},
author={Xueguang Ma and Qian Liu and Dongfu Jiang and Ge Zhang and Zejun Ma and Wenhu Chen},
year={2025},
journal={arXiv:2505.14652},
url={https://arxiv.org/abs/2505.14652}
}
Runs of TIGER-Lab General-Reasoner-Qwen2.5-7B on huggingface.co
46
Total runs
0
24-hour runs
-1
3-day runs
-2
7-day runs
-327
30-day runs
More Information About General-Reasoner-Qwen2.5-7B huggingface.co Model
More General-Reasoner-Qwen2.5-7B license Visit here:
General-Reasoner-Qwen2.5-7B huggingface.co is an AI model on huggingface.co that provides General-Reasoner-Qwen2.5-7B's model effect (), which can be used instantly with this TIGER-Lab General-Reasoner-Qwen2.5-7B model. huggingface.co supports a free trial of the General-Reasoner-Qwen2.5-7B model, and also provides paid use of the General-Reasoner-Qwen2.5-7B. Support call General-Reasoner-Qwen2.5-7B model through api, including Node.js, Python, http.
General-Reasoner-Qwen2.5-7B huggingface.co is an online trial and call api platform, which integrates General-Reasoner-Qwen2.5-7B's modeling effects, including api services, and provides a free online trial of General-Reasoner-Qwen2.5-7B, you can try General-Reasoner-Qwen2.5-7B online for free by clicking the link below.
TIGER-Lab General-Reasoner-Qwen2.5-7B online free url in huggingface.co:
General-Reasoner-Qwen2.5-7B is an open source model from GitHub that offers a free installation service, and any user can find General-Reasoner-Qwen2.5-7B on GitHub to install. At the same time, huggingface.co provides the effect of General-Reasoner-Qwen2.5-7B install, users can directly use General-Reasoner-Qwen2.5-7B installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
General-Reasoner-Qwen2.5-7B install url in huggingface.co: