AesSedai / GLM-5.3-Flash-GGUF

huggingface.co
Total runs: 1.9K
24-hour runs: -166
7-day runs: -876
30-day runs: 1.0K
Model's Last Updated: August 30 2026

Introduction of GLM-5.3-Flash-GGUF

Model Details of GLM-5.3-Flash-GGUF

Notes

This repo contains specialized MoE-quants for zai-org/GLM-5.3-Flash-BF16. The idea being that given the huge size of the FFN tensors compared to the rest of the tensors in the model, it should be possible to achieve a better quality while keeping the overall size of the entire model smaller compared to a similar naive quantization. To that end, the quantization type default is kept in high quality and the FFN UP + FFN GATE tensors are quanted down along with the FFN DOWN tensors.

Quant Size Mixture PPL 1-(Mean PPL(Q)/PPL(base)) KLD
Q5_K_M 224.28 GiB (6.01 BPW) Q8_0 / Q5_K / Q5_K / Q6_K 3.589877 ± 0.019865 +0.5529% 0.027859 ± 0.000207
Q4_K_M 188.10 GiB (5.04 BPW) Q8_0 / Q4_K / Q4_K / Q5_K 3.635356 ± 0.020204 +1.8267% 0.050181 ± 0.000333
IQ4_XS 148.24 GiB (3.97 BPW) Q8_0 / IQ3_S / IQ3_S / IQ4_XS 3.819227 ± 0.021423 +6.9770% 0.117358 ± 0.000727
IQ3_S 116.14 GiB (3.11 BPW) Q6_K / IQ2_S / IQ2_S / IQ3_S 4.387061 ± 0.025595 +22.8821% 0.283438 ± 0.001596
IQ2_S 105.81 GiB (2.83 BPW) Q6_K / IQ2_XS / IQ2_XS / IQ3_XXS 4.761384 ± 0.028305 +33.3669% 0.375406 ± 0.001984

kld_graph ppl_graph

Runs of AesSedai GLM-5.3-Flash-GGUF on huggingface.co

1.9K
Total runs
-166
24-hour runs
-1.0K
3-day runs
-876
7-day runs
1.0K
30-day runs

More Information About GLM-5.3-Flash-GGUF huggingface.co Model

GLM-5.3-Flash-GGUF huggingface.co

GLM-5.3-Flash-GGUF huggingface.co is an AI model on huggingface.co that provides GLM-5.3-Flash-GGUF's model effect (), which can be used instantly with this AesSedai GLM-5.3-Flash-GGUF model. huggingface.co supports a free trial of the GLM-5.3-Flash-GGUF model, and also provides paid use of the GLM-5.3-Flash-GGUF. Support call GLM-5.3-Flash-GGUF model through api, including Node.js, Python, http.

GLM-5.3-Flash-GGUF huggingface.co Url

https://huggingface.co/AesSedai/GLM-5.3-Flash-GGUF

AesSedai GLM-5.3-Flash-GGUF online free

GLM-5.3-Flash-GGUF huggingface.co is an online trial and call api platform, which integrates GLM-5.3-Flash-GGUF's modeling effects, including api services, and provides a free online trial of GLM-5.3-Flash-GGUF, you can try GLM-5.3-Flash-GGUF online for free by clicking the link below.

AesSedai GLM-5.3-Flash-GGUF online free url in huggingface.co:

https://huggingface.co/AesSedai/GLM-5.3-Flash-GGUF

GLM-5.3-Flash-GGUF install

GLM-5.3-Flash-GGUF is an open source model from GitHub that offers a free installation service, and any user can find GLM-5.3-Flash-GGUF on GitHub to install. At the same time, huggingface.co provides the effect of GLM-5.3-Flash-GGUF install, users can directly use GLM-5.3-Flash-GGUF installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

GLM-5.3-Flash-GGUF install url in huggingface.co:

https://huggingface.co/AesSedai/GLM-5.3-Flash-GGUF

Url of GLM-5.3-Flash-GGUF

GLM-5.3-Flash-GGUF huggingface.co Url

Provider of GLM-5.3-Flash-GGUF huggingface.co

AesSedai
ORGANIZATIONS

Other API from AesSedai

huggingface.co

Total runs: 1.3K
Run Growth: 314
Growth Rate: 23.88%
Updated:September 04 2026
huggingface.co

Total runs: 568
Run Growth: -1.1K
Growth Rate: -208.50%
Updated:August 13 2026
huggingface.co

Total runs: 214
Run Growth: 192
Growth Rate: 89.72%
Updated:February 01 2026
huggingface.co

Total runs: 108
Run Growth: -114
Growth Rate: -105.56%
Updated:February 13 2026