This model can be deployed efficiently using the
vLLM
backend.
Evaluation
The model was evaluated on GSM8K benchmarks.
Accuracy
Benchmark
GLM-5.1
GLM-5.1-MXFP4(this model)
Recovery
GSM8K (flexible-extract)
0.9522
0.9454
99.3%
Reproduction
The GSM8K results were obtained using the
lm-evaluation-harness
framework, based on the Docker image
rocm/pytorch-private:vllm_glm5_0225
, with vLLM, lm-eval compiled and installed from source inside the image.
The Docker image contains the necessary vLLM code modifications to support this model.
GLM-5.1-MXFP4 huggingface.co is an AI model on huggingface.co that provides GLM-5.1-MXFP4's model effect (), which can be used instantly with this amd GLM-5.1-MXFP4 model. huggingface.co supports a free trial of the GLM-5.1-MXFP4 model, and also provides paid use of the GLM-5.1-MXFP4. Support call GLM-5.1-MXFP4 model through api, including Node.js, Python, http.
GLM-5.1-MXFP4 huggingface.co is an online trial and call api platform, which integrates GLM-5.1-MXFP4's modeling effects, including api services, and provides a free online trial of GLM-5.1-MXFP4, you can try GLM-5.1-MXFP4 online for free by clicking the link below.
amd GLM-5.1-MXFP4 online free url in huggingface.co:
GLM-5.1-MXFP4 is an open source model from GitHub that offers a free installation service, and any user can find GLM-5.1-MXFP4 on GitHub to install. At the same time, huggingface.co provides the effect of GLM-5.1-MXFP4 install, users can directly use GLM-5.1-MXFP4 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.