radlab / pLLama3-8B-creator

huggingface.co
Total runs: 14
24-hour runs: -1
7-day runs: 1
30-day runs: 5
Model's Last Updated: Tháng tám 31 2024
text-generation

Introduction of pLLama3-8B-creator

Model Details of pLLama3-8B-creator

image/png

Intro

We have released a collection of radlab/pLLama3 models, which we have trained into Polish. The trained version is able to communicate more precisely with the user than the base version of meta-llama/Meta-Llama-3 models. As part of the collection, we provide models in 8B and 70B architecture. We make models in the 8B architecture available in two configurations:

  • radlab/pLLama3-8B-creator, a model that gives fairly short, specific answers to user queries;
  • radlab/pLLama3-8B-chat, a model that is a chatty version that reflects the behavior of the original meta-llama/Meta-Llama-3-8B-Instruct model.
Dataset

In addition to the instruction datasets publicly available for Polish, we developed our own dataset, which contains about 650,000 instructions. This data was semi-automatically generated using other publicly available datasets. In addition, we developed a learning dataset for the DPO process, which contained 100k examples in which we taught the model to select correctly written versions of texts from those with language errors.

Learning

The learning process was divided into two stages:

  • Post-training on a set of 650k instructions in Polish, the fine-tuning time was set to 5 epochs.
  • After the FT stage, we retrained the model using DPO on 100k instructions of correct writing in Polish, in this case we set the learning time to 15k steps.

The models we released are the ones after FT and the DPO process.

Post-FT learning metrics:

  • eval/loss : 0.8690009713172913
  • eval/runtime : 464.5158
  • eval/samples_per_second : 8.611
  • eval/steps_per_second : 8.611

Post-DPO learning metrics:

  • eval/logits/chosen : 0.1370937079191208
  • eval/logits/rejected : 0.07430506497621536
  • eval/logps/chosen : -454.11962890625
  • eval/logps/rejected : -764.1261596679688
  • eval/loss : 0.05717926099896431
  • eval/rewards/accuracies : 0.9372459053993224
  • eval/rewards/chosen : -26.75682830810547
  • eval/rewards/margins : 32.37759780883789
  • eval/rewards/rejected : -59.134429931640625
  • eval/runtime : 1,386.3177
  • eval/samples_per_second : 2.838
  • eval/steps_per_second : 1.42
Outro

Enjoy!

Runs of radlab pLLama3-8B-creator on huggingface.co

14
Total runs
-1
24-hour runs
-1
3-day runs
1
7-day runs
5
30-day runs

More Information About pLLama3-8B-creator huggingface.co Model

More pLLama3-8B-creator license Visit here:

https://choosealicense.com/licenses/llama3

pLLama3-8B-creator huggingface.co

pLLama3-8B-creator huggingface.co is an AI model on huggingface.co that provides pLLama3-8B-creator's model effect (), which can be used instantly with this radlab pLLama3-8B-creator model. huggingface.co supports a free trial of the pLLama3-8B-creator model, and also provides paid use of the pLLama3-8B-creator. Support call pLLama3-8B-creator model through api, including Node.js, Python, http.

pLLama3-8B-creator huggingface.co Url

https://huggingface.co/radlab/pLLama3-8B-creator

radlab pLLama3-8B-creator online free

pLLama3-8B-creator huggingface.co is an online trial and call api platform, which integrates pLLama3-8B-creator's modeling effects, including api services, and provides a free online trial of pLLama3-8B-creator, you can try pLLama3-8B-creator online for free by clicking the link below.

radlab pLLama3-8B-creator online free url in huggingface.co:

https://huggingface.co/radlab/pLLama3-8B-creator

pLLama3-8B-creator install

pLLama3-8B-creator is an open source model from GitHub that offers a free installation service, and any user can find pLLama3-8B-creator on GitHub to install. At the same time, huggingface.co provides the effect of pLLama3-8B-creator install, users can directly use pLLama3-8B-creator installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

pLLama3-8B-creator install url in huggingface.co:

https://huggingface.co/radlab/pLLama3-8B-creator

Url of pLLama3-8B-creator

pLLama3-8B-creator huggingface.co Url

Provider of pLLama3-8B-creator huggingface.co

radlab
ORGANIZATIONS

Other API from radlab

huggingface.co

Total runs: 10
Run Growth: 7
Growth Rate: 70.00%
Updated:Tháng Mười 20 2024
huggingface.co

Total runs: 6
Run Growth: 1
Growth Rate: 16.67%
Updated:Tháng bảy 15 2024
huggingface.co

Total runs: 5
Run Growth: 2
Growth Rate: 40.00%
Updated:Tháng Mười 20 2024
huggingface.co

Total runs: 3
Run Growth: 1
Growth Rate: 33.33%
Updated:Tháng Mười 04 2025
huggingface.co

Total runs: 2
Run Growth: -1
Growth Rate: -50.00%
Updated:Tháng tám 31 2024
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:Tháng sáu 01 2025