Kexer models are a collection of open-source generative text models fine-tuned on the
Kotlin Exercices
dataset.
This is a repository for the fine-tuned
Deepseek-coder-6.7b
model in the
Hugging Face Transformers
format.
How to use
As with the base model, we can use FIM. To do this, the following format must be used:
The model was trained on one A100 GPU with following hyperparameters:
Hyperparameter
Value
warmup
10%
max_lr
1e-4
scheduler
linear
total_batch_size
256 (~130K tokens per step)
num_epochs
4
More details about fine-tuning can be found in the technical report (coming soon!).
Fine-tuning data
For tuning this model, we used 15K exmaples from the synthetically generated
Kotlin Exercices
dataset. Every example follows the HumanEval format. In total, the dataset contains about 3.5M tokens.
Evaluation
For evaluation, we used the
Kotlin HumanEval
dataset, which contains all 161 tasks from HumanEval translated into Kotlin by human experts. You can find more details about the pre-processing necessary to obtain our results, including the code for running, on the
datasets's page
.
Here are the results of our evaluation:
Model name
Kotlin HumanEval Pass Rate
Deepseek-coder-6.7B
40.99
Deepseek-coder-6.7B-kexer
55.28
Ethical considerations and limitations
Deepseek-coder-6.7B-kexer is a new technology that carries risks with use. The testing conducted to date has not covered, nor could it cover all scenarios. For these reasons, as with all LLMs, Deepseek-coder-6.7B-kexer's potential outputs cannot be predicted in advance, and the model may in some instances produce inaccurate or objectionable responses to user prompts. The model was fine-tuned on a specific data format (Kotlin tasks), and deviation from this format can also lead to inaccurate or undesirable responses to user queries. Therefore, before deploying any applications of Deepseek-coder-6.7B-kexer, developers should perform safety testing and tuning tailored to their specific applications of the model.
Runs of QuantFactory deepseek-coder-6.7B-kexer-GGUF on huggingface.co
1.0K
Total runs
18
24-hour runs
164
3-day runs
181
7-day runs
323
30-day runs
More Information About deepseek-coder-6.7B-kexer-GGUF huggingface.co Model
More deepseek-coder-6.7B-kexer-GGUF license Visit here:
deepseek-coder-6.7B-kexer-GGUF huggingface.co is an AI model on huggingface.co that provides deepseek-coder-6.7B-kexer-GGUF's model effect (), which can be used instantly with this QuantFactory deepseek-coder-6.7B-kexer-GGUF model. huggingface.co supports a free trial of the deepseek-coder-6.7B-kexer-GGUF model, and also provides paid use of the deepseek-coder-6.7B-kexer-GGUF. Support call deepseek-coder-6.7B-kexer-GGUF model through api, including Node.js, Python, http.
deepseek-coder-6.7B-kexer-GGUF huggingface.co is an online trial and call api platform, which integrates deepseek-coder-6.7B-kexer-GGUF's modeling effects, including api services, and provides a free online trial of deepseek-coder-6.7B-kexer-GGUF, you can try deepseek-coder-6.7B-kexer-GGUF online for free by clicking the link below.
QuantFactory deepseek-coder-6.7B-kexer-GGUF online free url in huggingface.co:
deepseek-coder-6.7B-kexer-GGUF is an open source model from GitHub that offers a free installation service, and any user can find deepseek-coder-6.7B-kexer-GGUF on GitHub to install. At the same time, huggingface.co provides the effect of deepseek-coder-6.7B-kexer-GGUF install, users can directly use deepseek-coder-6.7B-kexer-GGUF installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
deepseek-coder-6.7B-kexer-GGUF install url in huggingface.co: