mlx-community / Qwopus3.8-27B-Flash-V2-4bit

huggingface.co
Total runs: 169
24-hour runs: 23
7-day runs: 40
30-day runs: 40
Model's Last Updated: September 22 2026
image-text-to-text

Introduction of Qwopus3.8-27B-Flash-V2-4bit

Model Details of Qwopus3.8-27B-Flash-V2-4bit

mlx-community/Qwopus3.8-27B-Flash-V2-4bit

This model mlx-community/Qwopus3.8-27B-Flash-V2-4bit was converted to MLX format from Jackrong/Qwopus3.8-27B-Flash-V2 using mlx-vlm version 0.4.4 .

This is a 4bit MLX quantized conversion. It keeps the source model's chat template and multimodal processor configuration for text/coding, image, and video-style inputs. The language model weights were quantized with MLX 4-bit affine quantization; the multimodal vision components are preserved for image/video inputs.

Refer to the original model card for model details, license, and intended use.

Use with mlx
pip install -U mlx-vlm
Image input
python -m mlx_vlm.generate \
  --model mlx-community/Qwopus3.8-27B-Flash-V2-4bit \
  --max-tokens 512 \
  --temperature 0.0 \
  --prompt "Describe this image." \
  --image <path_to_image>
Text / coding input
python -m mlx_vlm.generate \
  --model mlx-community/Qwopus3.8-27B-Flash-V2-4bit \
  --max-tokens 512 \
  --temperature 0.2 \
  --prompt "Write a Python function that parses a JSONL file and counts records by label."
Notes
  • This is a 4bit MLX quantized version of Jackrong/Qwopus3.8-27B-Flash-V2 .
  • The model is intended for Apple Silicon inference with MLX.
  • For multimodal usage, prefer mlx-vlm rather than plain mlx-lm .
  • License: Apache 2.0, inherited from the source model metadata.
Conversion
mlx_vlm.convert \
  --hf-path Jackrong/Qwopus3.8-27B-Flash-V2 \
  --mlx-path Qwopus3.8-27B-Flash-V2-4bit \
  --quantize \
  --q-bits 4 \
  --q-group-size 64 \
  --q-mode affine

Runs of mlx-community Qwopus3.8-27B-Flash-V2-4bit on huggingface.co

169
Total runs
23
24-hour runs
40
3-day runs
40
7-day runs
40
30-day runs

More Information About Qwopus3.8-27B-Flash-V2-4bit huggingface.co Model

More Qwopus3.8-27B-Flash-V2-4bit license Visit here:

https://choosealicense.com/licenses/apache-2.0

Qwopus3.8-27B-Flash-V2-4bit huggingface.co

Qwopus3.8-27B-Flash-V2-4bit huggingface.co is an AI model on huggingface.co that provides Qwopus3.8-27B-Flash-V2-4bit's model effect (), which can be used instantly with this mlx-community Qwopus3.8-27B-Flash-V2-4bit model. huggingface.co supports a free trial of the Qwopus3.8-27B-Flash-V2-4bit model, and also provides paid use of the Qwopus3.8-27B-Flash-V2-4bit. Support call Qwopus3.8-27B-Flash-V2-4bit model through api, including Node.js, Python, http.

Qwopus3.8-27B-Flash-V2-4bit huggingface.co Url

https://huggingface.co/mlx-community/Qwopus3.8-27B-Flash-V2-4bit

mlx-community Qwopus3.8-27B-Flash-V2-4bit online free

Qwopus3.8-27B-Flash-V2-4bit huggingface.co is an online trial and call api platform, which integrates Qwopus3.8-27B-Flash-V2-4bit's modeling effects, including api services, and provides a free online trial of Qwopus3.8-27B-Flash-V2-4bit, you can try Qwopus3.8-27B-Flash-V2-4bit online for free by clicking the link below.

mlx-community Qwopus3.8-27B-Flash-V2-4bit online free url in huggingface.co:

https://huggingface.co/mlx-community/Qwopus3.8-27B-Flash-V2-4bit

Qwopus3.8-27B-Flash-V2-4bit install

Qwopus3.8-27B-Flash-V2-4bit is an open source model from GitHub that offers a free installation service, and any user can find Qwopus3.8-27B-Flash-V2-4bit on GitHub to install. At the same time, huggingface.co provides the effect of Qwopus3.8-27B-Flash-V2-4bit install, users can directly use Qwopus3.8-27B-Flash-V2-4bit installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Qwopus3.8-27B-Flash-V2-4bit install url in huggingface.co:

https://huggingface.co/mlx-community/Qwopus3.8-27B-Flash-V2-4bit

Url of Qwopus3.8-27B-Flash-V2-4bit

Qwopus3.8-27B-Flash-V2-4bit huggingface.co Url

Provider of Qwopus3.8-27B-Flash-V2-4bit huggingface.co

mlx-community
ORGANIZATIONS

Other API from mlx-community