Corianas / Neural-Mistral-7B

huggingface.co
Total runs: 136
24-hour runs: 0
7-day runs: -2
30-day runs: -2
Model's Last Updated: March 06 2024
text-generation

Introduction of Neural-Mistral-7B

Model Details of Neural-Mistral-7B

Model Card for Model ID

This is a DPO finetune of Mistral 7b-instruct0.2 following the article: https://towardsdatascience.com/fine-tune-a-mistral-7b-model-with-direct-preference-optimization-708042745aac

Model Details
Model Description

This is the model card of a 🤗 transformers model that has been pushed on the Hub. This model card has been automatically generated.

  • Developed by: Corianas
  • Model type: [More Information Needed]
  • License: Apache 2.0
  • **Finetuned from model: mistralai/Mistral-7B-Instruct-v0.2
Instruction format

In order to leverage instruction fine-tuning, your prompt should be surrounded by [INST] and [/INST] tokens. The very first instruction should begin with a begin of sentence id. The next instructions should not. The assistant generation will be ended by the end-of-sentence token id.

E.g.

text = "<s>[INST] What is your favourite condiment? [/INST]"
"Well, I'm quite partial to a good squeeze of fresh lemon juice. It adds just the right amount of zesty flavour to whatever I'm cooking up in the kitchen!</s> "
"[INST] Do you have mayonnaise recipes? [/INST]"

This format is available as a chat template via the apply_chat_template() method:

from transformers import AutoModelForCausalLM, AutoTokenizer

device = "cuda" # the device to load the model onto

model = AutoModelForCausalLM.from_pretrained("mistralai/Mistral-7B-Instruct-v0.2")
tokenizer = AutoTokenizer.from_pretrained("mistralai/Mistral-7B-Instruct-v0.2")

messages = [
    {"role": "user", "content": "What is your favourite condiment?"},
    {"role": "assistant", "content": "Well, I'm quite partial to a good squeeze of fresh lemon juice. It adds just the right amount of zesty flavour to whatever I'm cooking up in the kitchen!"},
    {"role": "user", "content": "Do you have mayonnaise recipes?"}
]

encodeds = tokenizer.apply_chat_template(messages, return_tensors="pt")

model_inputs = encodeds.to(device)
model.to(device)

generated_ids = model.generate(model_inputs, max_new_tokens=1000, do_sample=True)
decoded = tokenizer.batch_decode(generated_ids)
print(decoded[0])
Model Architecture

This instruction model is based on Mistral-7B-v0.1, a transformer model with the following architecture choices:

  • Grouped-Query Attention
  • Sliding-Window Attention
  • Byte-fallback BPE tokenizer
How to Get Started with the Model

Use the code below to get started with the model.

[More Information Needed]

Training Details
Training Data

Intel/orca_dpo_pairs

Training Procedure

https://medium.com/towards-data-science/fine-tune-a-mistral-7b-model-with-direct-preference-optimization-708042745aac

Preprocessing [optional]

def chatml_format(example): # Format system if len(example['system']) > 0: message = {"role": "user", "content": f"{example['system']}\n{example['question']}"} prompt = tokenizer.apply_chat_template([message], tokenize=False) else: # Format instruction message = {"role": "user", "content": example['question']} prompt = tokenizer.apply_chat_template([message], tokenize=False, add_generation_prompt=True)

# Format chosen answer
chosen = example['chosen'] + tokenizer.eos_token

# Format rejected answer
rejected = example['rejected'] + tokenizer.eos_token

return {
    "prompt": prompt,
    "chosen": chosen,
    "rejected": rejected,
}
Training Hyperparameters

training_args = TrainingArguments( per_device_train_batch_size=4, gradient_accumulation_steps=4, gradient_checkpointing=True, learning_rate=5e-5, lr_scheduler_type="cosine", max_steps=200, save_strategy="no", logging_steps=1, output_dir=new_model, optim="paged_adamw_32bit", warmup_steps=100, bf16=True, report_to="wandb", )

Evaluation
Testing Data, Factors & Metrics
Testing Data

[More Information Needed]

Factors

[More Information Needed]

Metrics

[More Information Needed]

Results

[More Information Needed]

Summary
Model Examination [optional]

[More Information Needed]

Environmental Impact

Carbon emissions can be estimated using the Machine Learning Impact calculator presented in Lacoste et al. (2019) .

  • Hardware Type: [More Information Needed]
  • Hours used: [More Information Needed]
  • Cloud Provider: [More Information Needed]
  • Compute Region: [More Information Needed]
  • Carbon Emitted: [More Information Needed]
Technical Specifications [optional]
Model Architecture and Objective

[More Information Needed]

Compute Infrastructure

[More Information Needed]

Hardware

[More Information Needed]

Software

[More Information Needed]

Citation [optional]

BibTeX:

[More Information Needed]

APA:

[More Information Needed]

Glossary [optional]

[More Information Needed]

More Information [optional]

[More Information Needed]

Model Card Authors [optional]

[More Information Needed]

Model Card Contact

[More Information Needed]

Runs of Corianas Neural-Mistral-7B on huggingface.co

136
Total runs
0
24-hour runs
-6
3-day runs
-2
7-day runs
-2
30-day runs

More Information About Neural-Mistral-7B huggingface.co Model

More Neural-Mistral-7B license Visit here:

https://choosealicense.com/licenses/apache-2.0

Neural-Mistral-7B huggingface.co

Neural-Mistral-7B huggingface.co is an AI model on huggingface.co that provides Neural-Mistral-7B's model effect (), which can be used instantly with this Corianas Neural-Mistral-7B model. huggingface.co supports a free trial of the Neural-Mistral-7B model, and also provides paid use of the Neural-Mistral-7B. Support call Neural-Mistral-7B model through api, including Node.js, Python, http.

Neural-Mistral-7B huggingface.co Url

https://huggingface.co/Corianas/Neural-Mistral-7B

Corianas Neural-Mistral-7B online free

Neural-Mistral-7B huggingface.co is an online trial and call api platform, which integrates Neural-Mistral-7B's modeling effects, including api services, and provides a free online trial of Neural-Mistral-7B, you can try Neural-Mistral-7B online for free by clicking the link below.

Corianas Neural-Mistral-7B online free url in huggingface.co:

https://huggingface.co/Corianas/Neural-Mistral-7B

Neural-Mistral-7B install

Neural-Mistral-7B is an open source model from GitHub that offers a free installation service, and any user can find Neural-Mistral-7B on GitHub to install. At the same time, huggingface.co provides the effect of Neural-Mistral-7B install, users can directly use Neural-Mistral-7B installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Neural-Mistral-7B install url in huggingface.co:

https://huggingface.co/Corianas/Neural-Mistral-7B

Url of Neural-Mistral-7B

Neural-Mistral-7B huggingface.co Url

Provider of Neural-Mistral-7B huggingface.co

Corianas
ORGANIZATIONS

Other API from Corianas

huggingface.co

Total runs: 552
Run Growth: 386
Growth Rate: 69.93%
Updated:November 18 2023
huggingface.co

Total runs: 542
Run Growth: 378
Growth Rate: 69.74%
Updated:November 18 2023
huggingface.co

Total runs: 534
Run Growth: 222
Growth Rate: 41.57%
Updated:November 18 2023
huggingface.co

Total runs: 250
Run Growth: 35
Growth Rate: 14.00%
Updated:March 26 2024
huggingface.co

Total runs: 223
Run Growth: 147
Growth Rate: 65.92%
Updated:December 06 2024
huggingface.co

Total runs: 111
Run Growth: 101
Growth Rate: 90.99%
Updated:March 30 2023
huggingface.co

Total runs: 95
Run Growth: 26
Growth Rate: 27.37%
Updated:November 18 2023
huggingface.co

Total runs: 89
Run Growth: 18
Growth Rate: 20.22%
Updated:November 18 2023
huggingface.co

Total runs: 80
Run Growth: 10
Growth Rate: 12.50%
Updated:November 18 2023
huggingface.co

Total runs: 80
Run Growth: 14
Growth Rate: 17.50%
Updated:November 18 2023
huggingface.co

Total runs: 44
Run Growth: -29
Growth Rate: -65.91%
Updated:March 20 2023
huggingface.co

Total runs: 37
Run Growth: 26
Growth Rate: 70.27%
Updated:February 06 2025
huggingface.co

Total runs: 9
Run Growth: 0
Growth Rate: 0.00%
Updated:March 22 2024
huggingface.co

Total runs: 1
Run Growth: -2
Growth Rate: -200.00%
Updated:March 26 2024
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:September 10 2023
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:June 21 2022