4-bit (group size 64,
4.649 bits/weight
) MLX quantization of
deepreinforce-ai/Ornith-1.0-35B
,
produced with
mlx-vlm
0.6.3. Full multimodal: the vision encoder is preserved and
quantized alongside the language model. For Apple Silicon. Runs in
mlx-vlm
or any MLX app.
Conversion note (MoE expert fusion)
Ornith stores its 256 MoE experts unfused (per-expert), but mlx-vlm's
qwen3_5_moe
loader expects
them fused/batched. A
sanitize
monkeypatch was required to stack the experts before conversion; without it the
conversion failed. This is a standard mlx-vlm 4-bit quant.
Ornith-1.0-35B-4bit huggingface.co is an AI model on huggingface.co that provides Ornith-1.0-35B-4bit's model effect (), which can be used instantly with this mlx-community Ornith-1.0-35B-4bit model. huggingface.co supports a free trial of the Ornith-1.0-35B-4bit model, and also provides paid use of the Ornith-1.0-35B-4bit. Support call Ornith-1.0-35B-4bit model through api, including Node.js, Python, http.
Ornith-1.0-35B-4bit huggingface.co is an online trial and call api platform, which integrates Ornith-1.0-35B-4bit's modeling effects, including api services, and provides a free online trial of Ornith-1.0-35B-4bit, you can try Ornith-1.0-35B-4bit online for free by clicking the link below.
mlx-community Ornith-1.0-35B-4bit online free url in huggingface.co:
Ornith-1.0-35B-4bit is an open source model from GitHub that offers a free installation service, and any user can find Ornith-1.0-35B-4bit on GitHub to install. At the same time, huggingface.co provides the effect of Ornith-1.0-35B-4bit install, users can directly use Ornith-1.0-35B-4bit installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
Ornith-1.0-35B-4bit install url in huggingface.co: