Daredevil-8B is a mega-merge designed to maximize MMLU. On 27 May 24, it is the Llama 3 8B model with the
highest MMLU score
.
From my experience, a high MMLU score is all you need with Llama 3 models.
It is a merge of the following models using
LazyMergekit
:
Thanks to nbeerbower, Hastagaras, openchat, Kukedlc, cstr, flammenai, and KingNish for their merges. Special thanks to Charles Goddard and Arcee.ai for MergeKit.
Daredevil-8B is the best-performing 8B model on the Open LLM Leaderboard in terms of MMLU score (27 May 24).
Nous
Daredevil-8B is the best-performing 8B model on Nous' benchmark suite (evaluation performed using
LLM AutoEval
, 27 May 24). See the entire leaderboard
here
.
models:-model:NousResearch/Meta-Llama-3-8B# No parameters necessary for base model-model:nbeerbower/llama-3-stella-8Bparameters:density:0.6weight:0.16-model:Hastagaras/llama-3-8b-okayparameters:density:0.56weight:0.1-model:nbeerbower/llama-3-gutenberg-8Bparameters:density:0.6weight:0.18-model:openchat/openchat-3.6-8b-20240522parameters:density:0.56weight:0.12-model:Kukedlc/NeuralLLaMa-3-8b-DT-v0.1parameters:density:0.58weight:0.18-model:cstr/llama3-8b-spaetzle-v20parameters:density:0.56weight:0.08-model:mlabonne/ChimeraLlama-3-8B-v3parameters:density:0.56weight:0.08-model:flammenai/Mahou-1.1-llama3-8Bparameters:density:0.55weight:0.05-model:KingNish/KingNish-Llama3-8bparameters:density:0.55weight:0.05merge_method:dare_tiesbase_model:NousResearch/Meta-Llama-3-8Bdtype:bfloat16
💻 Usage
!pip install -qU transformers accelerate
from transformers import AutoTokenizer
import transformers
import torch
model = "mlabonne/Daredevil-8B"
messages = [{"role": "user", "content": "What is a large language model?"}]
tokenizer = AutoTokenizer.from_pretrained(model)
prompt = tokenizer.apply_chat_template(messages, tokenize=False, add_generation_prompt=True)
pipeline = transformers.pipeline(
"text-generation",
model=model,
torch_dtype=torch.bfloat16,
device_map="auto",
)
outputs = pipeline(prompt, max_new_tokens=256, do_sample=True, temperature=0.7, top_k=50, top_p=0.95)
print(outputs[0]["generated_text"])
Runs of mlabonne Daredevil-8B on huggingface.co
196
Total runs
-2
24-hour runs
40
3-day runs
66
7-day runs
121
30-day runs
More Information About Daredevil-8B huggingface.co Model
Daredevil-8B huggingface.co is an AI model on huggingface.co that provides Daredevil-8B's model effect (), which can be used instantly with this mlabonne Daredevil-8B model. huggingface.co supports a free trial of the Daredevil-8B model, and also provides paid use of the Daredevil-8B. Support call Daredevil-8B model through api, including Node.js, Python, http.
Daredevil-8B huggingface.co is an online trial and call api platform, which integrates Daredevil-8B's modeling effects, including api services, and provides a free online trial of Daredevil-8B, you can try Daredevil-8B online for free by clicking the link below.
mlabonne Daredevil-8B online free url in huggingface.co:
Daredevil-8B is an open source model from GitHub that offers a free installation service, and any user can find Daredevil-8B on GitHub to install. At the same time, huggingface.co provides the effect of Daredevil-8B install, users can directly use Daredevil-8B installed effect in huggingface.co for debugging and trial. It also supports api for free installation.