mlabonne / phixtral-4x2_8

huggingface.co
Total runs: 149
24-hour runs: 0
7-day runs: -3
30-day runs: 120
Model's Last Updated: January 16 2024
text-generation

Introduction of phixtral-4x2_8

Model Details of phixtral-4x2_8

phixtral-4x2_8

phixtral-4x2_8 is the first Mixure of Experts (MoE) made with four microsoft/phi-2 models, inspired by the mistralai/Mixtral-8x7B-v0.1 architecture. It performs better than each individual expert.

⚡ Quantized models
🏆 Evaluation

The evaluation was performed using LLM AutoEval on Nous suite.

Model AGIEval GPT4All TruthfulQA Bigbench Average
phixtral-4x2_8 33.91 70.44 48.78 37.68 47.7
dolphin-2_6-phi-2 33.12 69.85 47.39 37.2 46.89
phi-2-dpo 30.39 71.68 50.75 34.9 46.93
phi-2-sft-dpo-gpt4_en-ep1 30.61 71.13 48.74 35.23 46.43
phi-2-coder * 29.30 71.03 45.13 35.54 45.25
phi-2 27.98 70.8 44.43 35.21 44.61
* results reported by @vince62s here .

Check YALL - Yet Another LLM Leaderboard to compare it with other models.

🧩 Configuration

The model has been made with a custom version of the mergekit library (mixtral branch) and the following configuration:

base_model: cognitivecomputations/dolphin-2_6-phi-2
gate_mode: cheap_embed
experts:
  - source_model: cognitivecomputations/dolphin-2_6-phi-2
    positive_prompts: [""]
  - source_model: lxuechen/phi-2-dpo
    positive_prompts: [""]
  - source_model: Yhyu13/phi-2-sft-dpo-gpt4_en-ep1
    positive_prompts: [""]
  - source_model: mrm8488/phi-2-coder
    positive_prompts: [""]
💻 Usage

Here's a Colab notebook to run Phixtral in 4-bit precision on a free T4 GPU.

!pip install -q --upgrade transformers einops accelerate bitsandbytes

import torch
from transformers import AutoModelForCausalLM, AutoTokenizer

model_name = "phixtral-4x2_8"
instruction = '''
    def print_prime(n):
        """
        Print all primes between 1 and n
        """
'''

torch.set_default_device("cuda")

# Load the model and tokenizer
model = AutoModelForCausalLM.from_pretrained(
    f"mlabonne/{model_name}", 
    torch_dtype="auto", 
    load_in_4bit=True, 
    trust_remote_code=True
)
tokenizer = AutoTokenizer.from_pretrained(
    f"mlabonne/{model_name}", 
    trust_remote_code=True
)

# Tokenize the input string
inputs = tokenizer(
    instruction, 
    return_tensors="pt", 
    return_attention_mask=False
)

# Generate text using the model
outputs = model.generate(**inputs, max_length=200)

# Decode and print the output
text = tokenizer.batch_decode(outputs)[0]
print(text)

Inspired by mistralai/Mixtral-8x7B-v0.1 , you can specify the num_experts_per_tok and num_local_experts in the config.json file (2 and 4 by default). This configuration is automatically loaded in configuration.py .

vince62s implemented the MoE inference code in the modeling_phi.py file. In particular, see the MoE class .

🤝 Acknowledgments

A special thanks to vince62s for the inference code and the dynamic configuration of the number of experts. He was very patient and helped me to debug everything.

Thanks to Charles Goddard for the mergekit library and the implementation of the MoE for clowns .

Thanks to ehartford , lxuechen , Yhyu13 , and mrm8488 for their fine-tuned phi-2 models.

Runs of mlabonne phixtral-4x2_8 on huggingface.co

149
Total runs
0
24-hour runs
2
3-day runs
-3
7-day runs
120
30-day runs

More Information About phixtral-4x2_8 huggingface.co Model

More phixtral-4x2_8 license Visit here:

https://choosealicense.com/licenses/mit

phixtral-4x2_8 huggingface.co

phixtral-4x2_8 huggingface.co is an AI model on huggingface.co that provides phixtral-4x2_8's model effect (), which can be used instantly with this mlabonne phixtral-4x2_8 model. huggingface.co supports a free trial of the phixtral-4x2_8 model, and also provides paid use of the phixtral-4x2_8. Support call phixtral-4x2_8 model through api, including Node.js, Python, http.

phixtral-4x2_8 huggingface.co Url

https://huggingface.co/mlabonne/phixtral-4x2_8

mlabonne phixtral-4x2_8 online free

phixtral-4x2_8 huggingface.co is an online trial and call api platform, which integrates phixtral-4x2_8's modeling effects, including api services, and provides a free online trial of phixtral-4x2_8, you can try phixtral-4x2_8 online for free by clicking the link below.

mlabonne phixtral-4x2_8 online free url in huggingface.co:

https://huggingface.co/mlabonne/phixtral-4x2_8

phixtral-4x2_8 install

phixtral-4x2_8 is an open source model from GitHub that offers a free installation service, and any user can find phixtral-4x2_8 on GitHub to install. At the same time, huggingface.co provides the effect of phixtral-4x2_8 install, users can directly use phixtral-4x2_8 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

phixtral-4x2_8 install url in huggingface.co:

https://huggingface.co/mlabonne/phixtral-4x2_8

Url of phixtral-4x2_8

phixtral-4x2_8 huggingface.co Url

Provider of phixtral-4x2_8 huggingface.co

mlabonne
ORGANIZATIONS

Other API from mlabonne

huggingface.co

Total runs: 156
Run Growth: -18
Growth Rate: -11.54%
Updated:March 04 2024
huggingface.co

Total runs: 75
Run Growth: 24
Growth Rate: 32.00%
Updated:March 04 2024