legraphista / granite-20b-code-instruct-IMat-GGUF

huggingface.co
Total runs: 1.8K
24-hour runs: 0
7-day runs: 1
30-day runs: 1.3K
Model's Last Updated: September 02 2024
text-generation

Introduction of granite-20b-code-instruct-IMat-GGUF

Model Details of granite-20b-code-instruct-IMat-GGUF

granite-20b-code-instruct-IMat-GGUF

Llama.cpp imatrix quantization of ibm-granite/granite-20b-code-instruct

Original Model: ibm-granite/granite-20b-code-instruct
Original dtype: BF16 ( bfloat16 )
Quantized by: llama.cpp b3649
IMatrix dataset: here


Files
IMatrix

Status: โœ… Available
Link: here

Common Quants
Filename Quant type File Size Status Uses IMatrix Is Split
granite-20b-code-instruct.Q8_0.gguf Q8_0 21.48GB โœ… Available โšช Static ๐Ÿ“ฆ No
granite-20b-code-instruct.Q6_K.gguf Q6_K 16.63GB โœ… Available โšช Static ๐Ÿ“ฆ No
granite-20b-code-instruct.Q4_K.gguf Q4_K 12.82GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.Q3_K.gguf Q3_K 10.57GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.Q2_K.gguf Q2_K 7.93GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
All Quants
Filename Quant type File Size Status Uses IMatrix Is Split
granite-20b-code-instruct.BF16.gguf BF16 40.24GB โœ… Available โšช Static ๐Ÿ“ฆ No
granite-20b-code-instruct.FP16.gguf F16 40.24GB โœ… Available โšช Static ๐Ÿ“ฆ No
granite-20b-code-instruct.Q8_0.gguf Q8_0 21.48GB โœ… Available โšช Static ๐Ÿ“ฆ No
granite-20b-code-instruct.Q6_K.gguf Q6_K 16.63GB โœ… Available โšช Static ๐Ÿ“ฆ No
granite-20b-code-instruct.Q5_K.gguf Q5_K 14.81GB โœ… Available โšช Static ๐Ÿ“ฆ No
granite-20b-code-instruct.Q5_K_S.gguf Q5_K_S 14.02GB โœ… Available โšช Static ๐Ÿ“ฆ No
granite-20b-code-instruct.Q4_K.gguf Q4_K 12.82GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.Q4_K_S.gguf Q4_K_S 11.67GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ4_NL.gguf IQ4_NL 11.55GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ4_XS.gguf IQ4_XS 10.94GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.Q3_K.gguf Q3_K 10.57GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.Q3_K_L.gguf Q3_K_L 11.74GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.Q3_K_S.gguf Q3_K_S 8.93GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ3_M.gguf IQ3_M 9.59GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ3_S.gguf IQ3_S 8.93GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ3_XS.gguf IQ3_XS 8.66GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ3_XXS.gguf IQ3_XXS 8.06GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.Q2_K.gguf Q2_K 7.93GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.Q2_K_S.gguf Q2_K_S 7.15GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ2_M.gguf IQ2_M 7.05GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ2_S.gguf IQ2_S 6.53GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ2_XS.gguf IQ2_XS 6.16GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ2_XXS.gguf IQ2_XXS 5.57GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ1_M.gguf IQ1_M 4.91GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
granite-20b-code-instruct.IQ1_S.gguf IQ1_S 4.52GB โœ… Available ๐ŸŸข IMatrix ๐Ÿ“ฆ No
Downloading using huggingface-cli

If you do not have hugginface-cli installed:

pip install -U "huggingface_hub[cli]"

Download the specific file you want:

huggingface-cli download legraphista/granite-20b-code-instruct-IMat-GGUF --include "granite-20b-code-instruct.Q8_0.gguf" --local-dir ./

If the model file is big, it has been split into multiple files. In order to download them all to a local folder, run:

huggingface-cli download legraphista/granite-20b-code-instruct-IMat-GGUF --include "granite-20b-code-instruct.Q8_0/*" --local-dir ./
# see FAQ for merging GGUF's

Inference
Simple chat template
Question:
{user_prompt}

Answer:
{assistant_response}

Question:
{next_user_prompt}

Chat template with system prompt
System:
{system_prompt}

Question:
{user_prompt}

Answer:
{assistant_response}

Question:
{next_user_prompt}

Llama.cpp
llama.cpp/main -m granite-20b-code-instruct.Q8_0.gguf --color -i -p "prompt here (according to the chat template)"

FAQ
Why is the IMatrix not applied everywhere?

According to this investigation , it appears that lower quantizations are the only ones that benefit from the imatrix input (as per hellaswag results).

How do I merge a split GGUF?
  1. Make sure you have gguf-split available
  2. Locate your GGUF chunks folder (ex: granite-20b-code-instruct.Q8_0 )
  3. Run gguf-split --merge granite-20b-code-instruct.Q8_0/granite-20b-code-instruct.Q8_0-00001-of-XXXXX.gguf granite-20b-code-instruct.Q8_0.gguf
    • Make sure to point gguf-split to the first chunk of the split.

Got a suggestion? Ping me @legraphista !

Runs of legraphista granite-20b-code-instruct-IMat-GGUF on huggingface.co

1.8K
Total runs
0
24-hour runs
0
3-day runs
1
7-day runs
1.3K
30-day runs

More Information About granite-20b-code-instruct-IMat-GGUF huggingface.co Model

More granite-20b-code-instruct-IMat-GGUF license Visit here:

https://choosealicense.com/licenses/apache-2.0

granite-20b-code-instruct-IMat-GGUF huggingface.co

granite-20b-code-instruct-IMat-GGUF huggingface.co is an AI model on huggingface.co that provides granite-20b-code-instruct-IMat-GGUF's model effect (), which can be used instantly with this legraphista granite-20b-code-instruct-IMat-GGUF model. huggingface.co supports a free trial of the granite-20b-code-instruct-IMat-GGUF model, and also provides paid use of the granite-20b-code-instruct-IMat-GGUF. Support call granite-20b-code-instruct-IMat-GGUF model through api, including Node.js, Python, http.

granite-20b-code-instruct-IMat-GGUF huggingface.co Url

https://huggingface.co/legraphista/granite-20b-code-instruct-IMat-GGUF

legraphista granite-20b-code-instruct-IMat-GGUF online free

granite-20b-code-instruct-IMat-GGUF huggingface.co is an online trial and call api platform, which integrates granite-20b-code-instruct-IMat-GGUF's modeling effects, including api services, and provides a free online trial of granite-20b-code-instruct-IMat-GGUF, you can try granite-20b-code-instruct-IMat-GGUF online for free by clicking the link below.

legraphista granite-20b-code-instruct-IMat-GGUF online free url in huggingface.co:

https://huggingface.co/legraphista/granite-20b-code-instruct-IMat-GGUF

granite-20b-code-instruct-IMat-GGUF install

granite-20b-code-instruct-IMat-GGUF is an open source model from GitHub that offers a free installation service, and any user can find granite-20b-code-instruct-IMat-GGUF on GitHub to install. At the same time, huggingface.co provides the effect of granite-20b-code-instruct-IMat-GGUF install, users can directly use granite-20b-code-instruct-IMat-GGUF installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

granite-20b-code-instruct-IMat-GGUF install url in huggingface.co:

https://huggingface.co/legraphista/granite-20b-code-instruct-IMat-GGUF

Url of granite-20b-code-instruct-IMat-GGUF

granite-20b-code-instruct-IMat-GGUF huggingface.co Url

Provider of granite-20b-code-instruct-IMat-GGUF huggingface.co

legraphista
ORGANIZATIONS

Other API from legraphista