CodeFuse-QWen-14B is a 14B Code-LLM finetuned by QLoRA of multiple code tasks on the base model StarCoder.
News and Updates
🔥🔥 2023-10-16 CodeFuse-QWen-14B has been released, achieving a pass@1 (greedy decoding) score of 48.78% on HumanEval, which is a 16% increase compared to Qwen-14b's 32.3%.
🔥🔥 2023-09-27 CodeFuse-StarCoder-15B has been released, achieving a pass@1 (greedy decoding) score of 54.9% on HumanEval, which is a 21% increase compared to StarCoder's 33.6%.
🔥🔥🔥 2023-09-26 We are pleased to announce the release of the
4-bit quantized version
of
CodeFuse-CodeLlama-34B
. Despite the quantization process, the model still achieves a remarkable 73.8% accuracy (greedy decoding) on the HumanEval pass@1 metric.
🔥🔥🔥 2023-09-11
CodeFuse-CodeLlama34B
has achived 74.4% of pass@1 (greedy decoding) on HumanEval, which is SOTA results for openspurced LLMs at present.
If you wish to see a demo of the model, you can visit ✨
CodeFuse Demo
✨✨
Performance
Code
Model
HumanEval(pass@1)
Date
CodeFuse-CodeLlama-34B
74.4%
2023.9
CodeFuse-CodeLlama-34B-4bits
73.8%
2023.9
WizardCoder-Python-34B-V1.0
73.2%
2023.8
GPT-4(zero-shot)
67.0%
2023.3
PanGu-Coder2 15B
61.6%
2023.8
CodeLlama-34b-Python
53.7%
2023.8
CodeLlama-34b
48.8%
2023.8
GPT-3.5(zero-shot)
48.1%
2022.11
OctoCoder
46.2%
2023.8
StarCoder-15B
33.6%
2023.5
Qwen-14b
32.3%
2023.10
CodeFuse-StarCoder-15B
54.9%
2023.9
CodeFuse-QWen-14B
48.78%
2023.10
NLP
Requirements
python>=3.8
pytorch>=2.0.0
transformers==4.32.0
Sentencepiece
CUDA 11.4
Inference String Format
The inference string is a concatenated string formed by combining conversation data(system, human and bot contents) in the training data format. It is used as input during the inference process.
Here is an example format of the concatenated string:
When applying inference, you always make your input string end with "<s>bot" to ask the model to generate answers.
Quickstart
pip install -r requirements.txt
import torch
from transformers import (
AutoTokenizer,
AutoModelForCausalLM
)
tokenizer = AutoTokenizer.from_pretrained('codefuse-ai/CodeFuse-QWen-14B', trust_remote_code=True)
tokenizer.padding_side = "left"
tokenizer.pad_token_id = tokenizer.convert_tokens_to_ids("<|endoftext|>")
tokenizer.eos_token_id = tokenizer.convert_tokens_to_ids("<|endoftext|>")
tokenizer.pad_token = "<|endoftext|>"
tokenizer.eos_token = "<|endoftext|>"# try 4bit loading if cuda memory not enough
model = AutoModelForCausalLM.from_pretrained(model_dir,
trust_remote_code=True,
load_in_4bit=False,
device_map="auto",
torch_dtype=torch.bfloat16)
model.eval()
HUMAN_ROLE_START_TAG = "<s>human\n"
BOT_ROLE_START_TAG = "<s>bot\n"
text = f"{HUMAN_ROLE_START_TAG}write a python function of quick sort.\n{BOT_ROLE_START_TAG}"
inputs = tokenizer(text, return_tensors='pt', padding=True, add_special_tokens=False).to("cuda")
outputs = model.generate(
inputs=inputs["input_ids"],
attention_mask=inputs["attention_mask"],
max_new_tokens=512,
top_p=0.95,
temperature=0.1,
do_sample=True,
eos_token_id=tokenizer.eos_token_id,
pad_token_id=tokenizer.pad_token_id
)
gen_text = tokenizer.batch_decode(outputs[:, inputs["input_ids"].shape[1]:], skip_special_tokens=True)
print(gen_text)
Citation
If you find our work useful or helpful for your R&D works, please feel free to cite our paper as below.
@article{mftcoder2023,
title={MFTCoder: Boosting Code LLMs with Multitask Fine-Tuning},
author={Bingchang Liu and Chaoyu Chen and Cong Liao and Zi Gong and Huan Wang and Zhichao Lei and Ming Liang and Dajun Chen and Min Shen and Hailian Zhou and Hang Yu and Jianguo Li},
year={2023},
journal={arXiv preprint arXiv},
archivePrefix={arXiv},
eprint={2311.02303}
}
CodeFuse-QWen-14B huggingface.co is an AI model on huggingface.co that provides CodeFuse-QWen-14B's model effect (), which can be used instantly with this codefuse-ai CodeFuse-QWen-14B model. huggingface.co supports a free trial of the CodeFuse-QWen-14B model, and also provides paid use of the CodeFuse-QWen-14B. Support call CodeFuse-QWen-14B model through api, including Node.js, Python, http.
CodeFuse-QWen-14B huggingface.co is an online trial and call api platform, which integrates CodeFuse-QWen-14B's modeling effects, including api services, and provides a free online trial of CodeFuse-QWen-14B, you can try CodeFuse-QWen-14B online for free by clicking the link below.
codefuse-ai CodeFuse-QWen-14B online free url in huggingface.co:
CodeFuse-QWen-14B is an open source model from GitHub that offers a free installation service, and any user can find CodeFuse-QWen-14B on GitHub to install. At the same time, huggingface.co provides the effect of CodeFuse-QWen-14B install, users can directly use CodeFuse-QWen-14B installed effect in huggingface.co for debugging and trial. It also supports api for free installation.