MLX 8-bit conversion of
Nex-N2.5-mini
, a sparse MoE language model for local inference, coding, reasoning, and long-context work. The source checkpoint supports a native context length of
262,144 tokens (256K)
.
Benchmarks
Benchmark results reported by Nex AI for the original Nex-N2.5 checkpoint and its upstream evaluation setup.
Release
Format
Quantization
Size
MLX safetensors
Affine 8-bit, group size 64
36.85 GB
This release contains the text-generation weights and tokenizer. It does not include MTP weights or a vision projector.
Nex-N2.5-mini-MLX-8bit huggingface.co is an AI model on huggingface.co that provides Nex-N2.5-mini-MLX-8bit's model effect (), which can be used instantly with this abenzerps Nex-N2.5-mini-MLX-8bit model. huggingface.co supports a free trial of the Nex-N2.5-mini-MLX-8bit model, and also provides paid use of the Nex-N2.5-mini-MLX-8bit. Support call Nex-N2.5-mini-MLX-8bit model through api, including Node.js, Python, http.
Nex-N2.5-mini-MLX-8bit huggingface.co is an online trial and call api platform, which integrates Nex-N2.5-mini-MLX-8bit's modeling effects, including api services, and provides a free online trial of Nex-N2.5-mini-MLX-8bit, you can try Nex-N2.5-mini-MLX-8bit online for free by clicking the link below.
abenzerps Nex-N2.5-mini-MLX-8bit online free url in huggingface.co:
Nex-N2.5-mini-MLX-8bit is an open source model from GitHub that offers a free installation service, and any user can find Nex-N2.5-mini-MLX-8bit on GitHub to install. At the same time, huggingface.co provides the effect of Nex-N2.5-mini-MLX-8bit install, users can directly use Nex-N2.5-mini-MLX-8bit installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
Nex-N2.5-mini-MLX-8bit install url in huggingface.co: