Generating high-quality code using natural language is a high-frequency demand in the deployment of large models. Today, the IDEA Research Institute's Fengshenbang team officially open-sourced the latest code model, Ziya-Coding-34B-v1.0. We achieved a good score of 75.5 on the HumanEval Pass@1 evaluation, surpassing the score of GPT-4 (67.0) and setting a new high for known open-source models. The Fengshenbang team is providing the community with advanced large model technology and experience, helping to produce and customize more excellent vertical models, and promoting the development of the large model ecosystem.
In early September, we open-sourced the code model Ziya-Coding-15B-v1 based on StarCoder-15B. The training experience accumulated in training Ziya-Coding-15B-v1 was transferred to the training of the new version.
We collected and constructed about 450,000 instruction data covering almost all code-related tasks for the first stage of fine-tuning. This includes about 100,000 Chinese instructions and 350,000 English instructions, ensuring data diversity. When constructing the data, we made full use of high-quality non-instructional code data, used LLM to generate corresponding instructions, and expanded to obtain more high-quality code instruction data.
During the experiment, we noticed that the difficulty and correctness of code instructions are key to the successful training of code models. Therefore, we introduced a second stage of fine-tuning. We used the evol-instruct method to generate a large amount of high-difficulty, multi-requirement code instruction data, and used a code compiler as feedback to filter out code that could pass compilation. Finally, we used LLM to generate unit tests to further verify the correctness of the code. We ultimately filtered out 46k data, and on the basis of the first-stage model, we fine-tuned it with a lower learning rate to finally obtain our Ziya-coding-34B-v1.0.
"<human>: \nPlease Complete the given function below according to the docstring: \n{prompt}\n<bot>: \n"
In this process, we performed a decontamination process on the fine-tuning dataset to avoid data leakage. The pass@1 metric for HumanEval is based on the results of greedy generation.
If you are using the resource for your work, please cite the our
paper
:
@article{fengshenbang,
author = {Jiaxing Zhang and Ruyi Gan and Junjie Wang and Yuxiang Zhang and Lin Zhang and Ping Yang and Xinyu Gao and Ziwei Wu and Xiaoqun Dong and Junqing He and Jianheng Zhuo and Qi Yang and Yongfeng Huang and Xiayu Li and Yanghan Wu and Junyu Lu and Xinyu Zhu and Weifeng Chen and Ting Han and Kunhao Pan and Rui Wang and Hao Wang and Xiaojun Wu and Zhongshen Zeng and Chongpei Chen},
title = {Fengshenbang 1.0: Being the Foundation of Chinese Cognitive Intelligence},
journal = {CoRR},
volume = {abs/2209.02970},
year = {2022}
}
Ziya-Coding-34B-v1.0 huggingface.co is an AI model on huggingface.co that provides Ziya-Coding-34B-v1.0's model effect (), which can be used instantly with this IDEA-CCNL Ziya-Coding-34B-v1.0 model. huggingface.co supports a free trial of the Ziya-Coding-34B-v1.0 model, and also provides paid use of the Ziya-Coding-34B-v1.0. Support call Ziya-Coding-34B-v1.0 model through api, including Node.js, Python, http.
Ziya-Coding-34B-v1.0 huggingface.co is an online trial and call api platform, which integrates Ziya-Coding-34B-v1.0's modeling effects, including api services, and provides a free online trial of Ziya-Coding-34B-v1.0, you can try Ziya-Coding-34B-v1.0 online for free by clicking the link below.
IDEA-CCNL Ziya-Coding-34B-v1.0 online free url in huggingface.co:
Ziya-Coding-34B-v1.0 is an open source model from GitHub that offers a free installation service, and any user can find Ziya-Coding-34B-v1.0 on GitHub to install. At the same time, huggingface.co provides the effect of Ziya-Coding-34B-v1.0 install, users can directly use Ziya-Coding-34B-v1.0 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
Ziya-Coding-34B-v1.0 install url in huggingface.co: