DJLougen / Qwable-5-27B-Coder-NVFP4

huggingface.co
Total runs: 63
24-hour runs: -1
7-day runs: -28
30-day runs: -169
Model's Last Updated: June 23 2026
text-generation

Introduction of Qwable-5-27B-Coder-NVFP4

Model Details of Qwable-5-27B-Coder-NVFP4

Qwable-5-27B-Coder banner

Qwable NVFP4 serving lane

Qwable-5-27B-Coder-NVFP4

Qwable-5-27B-Coder-NVFP4 is the compact NVIDIA ModelOpt release of Qwable: a Qwen3.6-based coder-agent tune trained first on Claude Fable 5 traces , then continued on Kimi 2.7 Coder traces .

This is the serving-oriented checkpoint: NVFP4 safetensors for runtimes that understand ModelOpt quantization. Use the GGUF repo for llama.cpp and Ollama.

Support on Ko-fi

Serving shape
Item Value
Format ModelOpt NVFP4 safetensors
Producer nvidia-modelopt 0.44.0
Quant algorithm NVFP4
Group size 16
Weight files 2 safetensors shards
Approx. size 19.7 GB total
Runtime target vLLM / TensorRT-LLM / NVIDIA ModelOpt-compatible serving
Source checkpoint DJLougen/Qwable-5-27B-Coder
BF16 source checkpoint
  -> NVIDIA ModelOpt export
      -> NVFP4 Linear targets, group size 16
      -> lm_head / SSM conv1d / visual modules excluded
      -> compact safetensors serving repo
What carries over from Qwable

The NVFP4 checkpoint is a quantized release of the same coder-agent model. The target workload remains repository navigation, patch planning, terminal-output analysis, verifier recovery, tool-shaped answers, and long-context coding prompts.

Early maintainer runs show the source Qwable checkpoint outperforming the base model on a private coder benchmark. Public benchmark details are pending, so treat that as early maintainer signal rather than a reproducible leaderboard claim.

Quantization details

From hf_quant_config.json :

Field Value
Producer modelopt
Producer version 0.44.0
Quant algorithm NVFP4
Group size 16
KV cache quantization Not set in this checkpoint
Higher-precision exclusions lm_head , SSM conv1d modules, and model.visual*

The config records 4-bit floating-point weights and input activations for Linear targets. The excluded modules are intentional: they preserve sensitive output, SSM convolution, and vision-tower components outside the main NVFP4 target set.

Quickstart

Install a runtime with support for ModelOpt NVFP4 checkpoints. Exact package versions change quickly; prefer the current vLLM or TensorRT-LLM documentation for your CUDA stack.

Download locally:

hf download DJLougen/Qwable-5-27B-Coder-NVFP4 --local-dir Qwable-5-27B-Coder-NVFP4

Example vLLM-style serving command, adjusted for your installed version and GPU topology:

vllm serve DJLougen/Qwable-5-27B-Coder-NVFP4 \
  --tensor-parallel-size 1 \
  --max-model-len 32768

If your runtime does not recognize the ModelOpt quantization config, use the BF16 source checkpoint or the GGUF release instead.

Recommended use
  • Use this repo for NVIDIA-serving experiments where NVFP4 is supported directly.
  • Use the BF16 repo for conversion, further training, or quality-ceiling evaluation.
  • Use the GGUF repo for llama.cpp and Ollama workflows.
  • Keep benchmark comparisons identical across model variants: same prompts, context, sampling, max tokens, and tool schema exposure.
Related releases
Limitations
  • Public benchmark tables are pending.
  • NVFP4 runtime compatibility depends on serving stack, CUDA version, GPU architecture, and package versions.
  • This is not a GGUF repo; llama.cpp users should use the GGUF release.
  • Quantization can change instruction following, code precision, and tool-call reliability. Validate on your own tasks.
  • Vision components are excluded from the main NVFP4 target set, but this release is marketed for coding behavior, not vision improvement.
  • Safety behavior is inherited from the base model and fine-tuning data; no separate safety alignment claim is made here.
License

Released under Apache-2.0, following the upstream base model license metadata.

Runs of DJLougen Qwable-5-27B-Coder-NVFP4 on huggingface.co

63
Total runs
-1
24-hour runs
-12
3-day runs
-28
7-day runs
-169
30-day runs

More Information About Qwable-5-27B-Coder-NVFP4 huggingface.co Model

More Qwable-5-27B-Coder-NVFP4 license Visit here:

https://choosealicense.com/licenses/apache-2.0

Qwable-5-27B-Coder-NVFP4 huggingface.co

Qwable-5-27B-Coder-NVFP4 huggingface.co is an AI model on huggingface.co that provides Qwable-5-27B-Coder-NVFP4's model effect (), which can be used instantly with this DJLougen Qwable-5-27B-Coder-NVFP4 model. huggingface.co supports a free trial of the Qwable-5-27B-Coder-NVFP4 model, and also provides paid use of the Qwable-5-27B-Coder-NVFP4. Support call Qwable-5-27B-Coder-NVFP4 model through api, including Node.js, Python, http.

Qwable-5-27B-Coder-NVFP4 huggingface.co Url

https://huggingface.co/DJLougen/Qwable-5-27B-Coder-NVFP4

DJLougen Qwable-5-27B-Coder-NVFP4 online free

Qwable-5-27B-Coder-NVFP4 huggingface.co is an online trial and call api platform, which integrates Qwable-5-27B-Coder-NVFP4's modeling effects, including api services, and provides a free online trial of Qwable-5-27B-Coder-NVFP4, you can try Qwable-5-27B-Coder-NVFP4 online for free by clicking the link below.

DJLougen Qwable-5-27B-Coder-NVFP4 online free url in huggingface.co:

https://huggingface.co/DJLougen/Qwable-5-27B-Coder-NVFP4

Qwable-5-27B-Coder-NVFP4 install

Qwable-5-27B-Coder-NVFP4 is an open source model from GitHub that offers a free installation service, and any user can find Qwable-5-27B-Coder-NVFP4 on GitHub to install. At the same time, huggingface.co provides the effect of Qwable-5-27B-Coder-NVFP4 install, users can directly use Qwable-5-27B-Coder-NVFP4 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Qwable-5-27B-Coder-NVFP4 install url in huggingface.co:

https://huggingface.co/DJLougen/Qwable-5-27B-Coder-NVFP4

Url of Qwable-5-27B-Coder-NVFP4

Qwable-5-27B-Coder-NVFP4 huggingface.co Url

Provider of Qwable-5-27B-Coder-NVFP4 huggingface.co

DJLougen
ORGANIZATIONS

Other API from DJLougen

huggingface.co

Total runs: 76
Run Growth: -874
Growth Rate: -1150.00%
Updated:April 10 2026
huggingface.co

Total runs: 26
Run Growth: -8
Growth Rate: -30.77%
Updated:April 10 2026
huggingface.co

Total runs: 11
Run Growth: -612
Growth Rate: -5563.64%
Updated:April 10 2026
huggingface.co

Total runs: 9
Run Growth: -634
Growth Rate: -7044.44%
Updated:April 10 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:January 13 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:January 13 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:May 05 2026