neuralmagic / Mistral-7B-Instruct-v0.3-GPTQ-4bit

huggingface.co
Total runs: 2.7K
24-hour runs: 0
7-day runs: 0
30-day runs: 0
Model's Last Updated: June 11 2024
text-generation

Introduction of Mistral-7B-Instruct-v0.3-GPTQ-4bit

Model Details of Mistral-7B-Instruct-v0.3-GPTQ-4bit

Model Card for Mistral-7B-Instruct-v0.3 quantized to 4bit weights

  • Weight-only quantization of Mistral-7B-Instruct-v0.3 via GPTQ to 4bits with group_size=128
  • GPTQ optimized for 99.75% accuracy recovery relative to the unquantized model

Open LLM Leaderboard evaluation scores

Mistral-7B-Instruct-v0.3 Mistral-7B-Instruct-v0.3-GPTQ-4bit
(this model)
arc-c
25-shot
63.48 63.40
mmlu
5-shot
61.13 60.89
hellaswag
10-shot
84.49 84.04
winogrande
5-shot
79.16 79.08
gsm8k
5-shot
43.37 45.41
truthfulqa
0-shot
59.65 57.48
Average
Accuracy
65.21 65.05
Recovery 100% 99.75%

vLLM Inference Performance

This model is ready for optimized inference using the Marlin mixed-precision kernels in vLLM: https://github.com/vllm-project/vllm

Simply start this model as an inference server with:

python -m vllm.entrypoints.openai.api_server --model neuralmagic/Mistral-7B-Instruct-v0.3-GPTQ-4bit

image/png

Runs of neuralmagic Mistral-7B-Instruct-v0.3-GPTQ-4bit on huggingface.co

2.7K
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
0
30-day runs

More Information About Mistral-7B-Instruct-v0.3-GPTQ-4bit huggingface.co Model

More Mistral-7B-Instruct-v0.3-GPTQ-4bit license Visit here:

https://choosealicense.com/licenses/apache-2.0

Mistral-7B-Instruct-v0.3-GPTQ-4bit huggingface.co

Mistral-7B-Instruct-v0.3-GPTQ-4bit huggingface.co is an AI model on huggingface.co that provides Mistral-7B-Instruct-v0.3-GPTQ-4bit's model effect (), which can be used instantly with this neuralmagic Mistral-7B-Instruct-v0.3-GPTQ-4bit model. huggingface.co supports a free trial of the Mistral-7B-Instruct-v0.3-GPTQ-4bit model, and also provides paid use of the Mistral-7B-Instruct-v0.3-GPTQ-4bit. Support call Mistral-7B-Instruct-v0.3-GPTQ-4bit model through api, including Node.js, Python, http.

Mistral-7B-Instruct-v0.3-GPTQ-4bit huggingface.co Url

https://huggingface.co/neuralmagic/Mistral-7B-Instruct-v0.3-GPTQ-4bit

neuralmagic Mistral-7B-Instruct-v0.3-GPTQ-4bit online free

Mistral-7B-Instruct-v0.3-GPTQ-4bit huggingface.co is an online trial and call api platform, which integrates Mistral-7B-Instruct-v0.3-GPTQ-4bit's modeling effects, including api services, and provides a free online trial of Mistral-7B-Instruct-v0.3-GPTQ-4bit, you can try Mistral-7B-Instruct-v0.3-GPTQ-4bit online for free by clicking the link below.

neuralmagic Mistral-7B-Instruct-v0.3-GPTQ-4bit online free url in huggingface.co:

https://huggingface.co/neuralmagic/Mistral-7B-Instruct-v0.3-GPTQ-4bit

Mistral-7B-Instruct-v0.3-GPTQ-4bit install

Mistral-7B-Instruct-v0.3-GPTQ-4bit is an open source model from GitHub that offers a free installation service, and any user can find Mistral-7B-Instruct-v0.3-GPTQ-4bit on GitHub to install. At the same time, huggingface.co provides the effect of Mistral-7B-Instruct-v0.3-GPTQ-4bit install, users can directly use Mistral-7B-Instruct-v0.3-GPTQ-4bit installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Mistral-7B-Instruct-v0.3-GPTQ-4bit install url in huggingface.co:

https://huggingface.co/neuralmagic/Mistral-7B-Instruct-v0.3-GPTQ-4bit

Url of Mistral-7B-Instruct-v0.3-GPTQ-4bit

Mistral-7B-Instruct-v0.3-GPTQ-4bit huggingface.co Url

Provider of Mistral-7B-Instruct-v0.3-GPTQ-4bit huggingface.co

neuralmagic
ORGANIZATIONS

Other API from neuralmagic