mlabonne / AlphaMonarch-7B

huggingface.co
Total runs: 13.5K
24-hour runs: 0
7-day runs: 79
30-day runs: 817
Model's Last Updated: March 29 2024
text-generation

Introduction of AlphaMonarch-7B

Model Details of AlphaMonarch-7B

image/jpeg

👑 AlphaMonarch-7B

tl;dr: AlphaMonarch-7B is a new DPO merge that retains all the reasoning abilities of the very best merges and significantly improves its conversational abilities. Kind of the best of both worlds in a 7B model. 🎉

AlphaMonarch-7B is a DPO fine-tuned of mlabonne/NeuralMonarch-7B using the argilla/OpenHermes2.5-dpo-binarized-alpha preference dataset.

It is based on a merge of the following models using LazyMergekit :

Special thanks to Jon Durbin , Intel , Argilla , and Teknium for the preference datasets.

Try the demo : https://huggingface.co/spaces/mlabonne/AlphaMonarch-7B

🔍 Applications

This model uses a context window of 8k. I recommend using it with the Mistral Instruct chat template (works perfectly with LM Studio).

If you use SillyTavern, you might want to tweak the inference parameters. Here's what LM Studio uses as a reference: temp 0.8, top_k 40, top_p 0.95, min_p 0.05, repeat_penalty 1.1.

It is one of the very best 7B models in terms of instructing following and reasoning abilities and can be used for conversations, RP, and storytelling. Note that it tends to have a quite formal and sophisticated style, but it can be changed by modifying the prompt.

⚡ Quantized models

Thanks to LoneStriker for the GPTQ, AWQ, and EXL2 quants.

🏆 Evaluation
Nous

AlphaMonarch-7B is the best-performing 7B model on Nous' benchmark suite (evaluation performed using LLM AutoEval ). See the entire leaderboard here .

Model Average AGIEval GPT4All TruthfulQA Bigbench
AlphaMonarch-7B 📄 62.74 45.37 77.01 78.39 50.2
NeuralMonarch-7B 📄 62.73 45.31 76.99 78.35 50.28
Monarch-7B 📄 62.68 45.48 77.07 78.04 50.14
teknium/OpenHermes-2.5-Mistral-7B 📄 52.42 42.75 72.99 52.99 40.94
mlabonne/NeuralHermes-2.5-Mistral-7B 📄 53.51 43.67 73.24 55.37 41.76
mlabonne/NeuralBeagle14-7B 📄 60.25 46.06 76.77 70.32 47.86
mlabonne/NeuralOmniBeagle-7B 📄 62.3 45.85 77.26 76.06 50.03
eren23/dpo-binarized-NeuralTrix-7B 📄 62.5 44.57 76.34 79.81 49.27
CultriX/NeuralTrix-7B-dpo 📄 62.5 44.61 76.33 79.8 49.24
EQ-bench

AlphaMonarch-7B is also outperforming 70B and 120B parameter models on EQ-bench by Samuel J. Paech , who kindly ran the evaluations.

image/png

MT-Bench
########## First turn ##########
                                    score
model                       turn         
gpt-4                       1     8.95625
OmniBeagle-7B               1     8.31250
AlphaMonarch-7B             1     8.23750
claude-v1                   1     8.15000
NeuralMonarch-7B            1     8.09375
gpt-3.5-turbo               1     8.07500
claude-instant-v1           1     7.80000

########## Second turn ##########
                                     score
model                       turn          
gpt-4                       2     9.025000
claude-instant-v1           2     8.012658
OmniBeagle-7B               2     7.837500
gpt-3.5-turbo               2     7.812500
claude-v1                   2     7.650000
AlphaMonarch-7B             2     7.618750
NeuralMonarch-7B            2     7.375000

########## Average ##########
                                score
model                                
gpt-4                        8.990625
OmniBeagle-7B                8.075000
gpt-3.5-turbo                7.943750
AlphaMonarch-7B              7.928125
claude-instant-v1            7.905660
claude-v1                    7.900000
NeuralMonarch-7B             7.734375
NeuralBeagle14-7B            7.628125
Open LLM Leaderboard

AlphaMonarch-7B is one of the best-performing non-merge 7B models on the Open LLM Leaderboard:

image/png

🌳 Model Family Tree

image/png

💻 Usage
!pip install -qU transformers accelerate

from transformers import AutoTokenizer
import transformers
import torch

model = "mlabonne/AlphaMonarch-7B"
messages = [{"role": "user", "content": "What is a large language model?"}]

tokenizer = AutoTokenizer.from_pretrained(model)
prompt = tokenizer.apply_chat_template(messages, tokenize=False, add_generation_prompt=True)
pipeline = transformers.pipeline(
    "text-generation",
    model=model,
    torch_dtype=torch.float16,
    device_map="auto",
)

outputs = pipeline(prompt, max_new_tokens=256, do_sample=True, temperature=0.7, top_k=50, top_p=0.95)
print(outputs[0]["generated_text"])

Runs of mlabonne AlphaMonarch-7B on huggingface.co

13.5K
Total runs
0
24-hour runs
19
3-day runs
79
7-day runs
817
30-day runs

More Information About AlphaMonarch-7B huggingface.co Model

More AlphaMonarch-7B license Visit here:

https://choosealicense.com/licenses/cc-by-nc-4.0

AlphaMonarch-7B huggingface.co

AlphaMonarch-7B huggingface.co is an AI model on huggingface.co that provides AlphaMonarch-7B's model effect (), which can be used instantly with this mlabonne AlphaMonarch-7B model. huggingface.co supports a free trial of the AlphaMonarch-7B model, and also provides paid use of the AlphaMonarch-7B. Support call AlphaMonarch-7B model through api, including Node.js, Python, http.

AlphaMonarch-7B huggingface.co Url

https://huggingface.co/mlabonne/AlphaMonarch-7B

mlabonne AlphaMonarch-7B online free

AlphaMonarch-7B huggingface.co is an online trial and call api platform, which integrates AlphaMonarch-7B's modeling effects, including api services, and provides a free online trial of AlphaMonarch-7B, you can try AlphaMonarch-7B online for free by clicking the link below.

mlabonne AlphaMonarch-7B online free url in huggingface.co:

https://huggingface.co/mlabonne/AlphaMonarch-7B

AlphaMonarch-7B install

AlphaMonarch-7B is an open source model from GitHub that offers a free installation service, and any user can find AlphaMonarch-7B on GitHub to install. At the same time, huggingface.co provides the effect of AlphaMonarch-7B install, users can directly use AlphaMonarch-7B installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

AlphaMonarch-7B install url in huggingface.co:

https://huggingface.co/mlabonne/AlphaMonarch-7B

Url of AlphaMonarch-7B

AlphaMonarch-7B huggingface.co Url

Provider of AlphaMonarch-7B huggingface.co

mlabonne
ORGANIZATIONS

Other API from mlabonne

huggingface.co

Total runs: 156
Run Growth: -18
Growth Rate: -11.54%
Updated:March 04 2024
huggingface.co

Total runs: 76
Run Growth: 25
Growth Rate: 32.89%
Updated:March 04 2024