beomi / kollama-33b

huggingface.co
Total runs: 6
24-hour runs: 0
7-day runs: 0
30-day runs: 0
Model's Last Updated: June 29 2023
text-generation

Introduction of kollama-33b

Model Details of kollama-33b

🚧 Note: this repo is under construction 🚧

Todo

✅ - finish

⏳ - currently working on it

  • ✅ Train new BBPE Tokenizer
  • ✅ Test train code on TPUv4 Pods (with model parallel)
  • ✅ Converting test (jax to PyTorch)
  • ✅ LM train validation on minimal dataset (1 sentence 1000 step)
  • ⏳ Build Data Shuffler (curriculum learning)
  • ⏳ Train 7B Model
  • ⏳ Train 13B Model
  • ⏳ Train 33B Model
  • Train 65B Model

KoLLaMA Model Card

KoLLaMA (33B) trained on Korean/English/Code dataset with LLaMA Architecture via JAX, with the warm support from Google TPU Research Cloud program for providing part of the computation resources.

Model details

Researcher developing the model

Junbum Lee (aka Beomi)

Model date

KoLLaMA was trained between 2023.04~

  • 33B model was trained on 2023.07~

Model version

This is alpha version of the model.

Model type

LLaMA is an auto-regressive language model, based on the transformer architecture. The model comes in different sizes: 7B, 13B, 33B and 65B parameters.

(This repo contains 33B model!)

Paper or resources for more information

More information can be found in the paper “LLaMA, Open and Efficient Foundation Language Models”, available at https://research.facebook.com/publications/llama-open-and-efficient-foundation-language-models/ .

Citations details

KoLLAMA: [TBD] LLAMA: https://research.facebook.com/publications/llama-open-and-efficient-foundation-language-models/

License

MIT

Where to send questions or comments about the model

Questions and comments about KoLLaMA can be sent via the GitHub repository of the project , by opening an issue.

Intended use

Primary intended uses

The primary use of KoLLaMA is research on Korean Opensource large language models

Primary intended users

The primary intended users of the model are researchers in natural language processing, machine learning and artificial intelligence.

Out-of-scope use cases

LLaMA is a base, or foundational, model. As such, it should not be used on downstream applications without further risk evaluation and mitigation. In particular, our model has not been trained with human feedback, and can thus generate toxic or offensive content, incorrect information or generally unhelpful answers.

Factors

Relevant factors

One of the most relevant factors for which model performance may vary is which language is used. Although we included 20 languages in the training data, most of our dataset is made of English text, and we thus expect the model to perform better for English than other languages. Relatedly, it has been shown in previous studies that performance might vary for different dialects, and we expect that it will be the case for our model.

Evaluation datasets

[TBD]

Training dataset

[TBD]

Ethical considerations

Data

The data used to train the model is collected from various sources, mostly from the Web. As such, it contains offensive, harmful and biased content. We thus expect the model to exhibit such biases from the training data.

Human life

The model is not intended to inform decisions about matters central to human life, and should not be used in such a way.

Risks and harms

Risks and harms of large language models include the generation of harmful, offensive or biased content. These models are often prone to generating incorrect information, sometimes referred to as hallucinations. We do not expect our model to be an exception in this regard.

Use cases

LLaMA is a foundational model, and as such, it should not be used for downstream applications without further investigation and mitigations of risks. These risks and potential fraught use cases include, but are not limited to: generation of misinformation and generation of harmful, biased or offensive content.

Runs of beomi kollama-33b on huggingface.co

6
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
0
30-day runs

More Information About kollama-33b huggingface.co Model

More kollama-33b license Visit here:

https://choosealicense.com/licenses/mit

kollama-33b huggingface.co

kollama-33b huggingface.co is an AI model on huggingface.co that provides kollama-33b's model effect (), which can be used instantly with this beomi kollama-33b model. huggingface.co supports a free trial of the kollama-33b model, and also provides paid use of the kollama-33b. Support call kollama-33b model through api, including Node.js, Python, http.

kollama-33b huggingface.co Url

https://huggingface.co/beomi/kollama-33b

beomi kollama-33b online free

kollama-33b huggingface.co is an online trial and call api platform, which integrates kollama-33b's modeling effects, including api services, and provides a free online trial of kollama-33b, you can try kollama-33b online for free by clicking the link below.

beomi kollama-33b online free url in huggingface.co:

https://huggingface.co/beomi/kollama-33b

kollama-33b install

kollama-33b is an open source model from GitHub that offers a free installation service, and any user can find kollama-33b on GitHub to install. At the same time, huggingface.co provides the effect of kollama-33b install, users can directly use kollama-33b installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

kollama-33b install url in huggingface.co:

https://huggingface.co/beomi/kollama-33b

Url of kollama-33b

kollama-33b huggingface.co Url

Provider of kollama-33b huggingface.co

beomi
ORGANIZATIONS

Other API from beomi

huggingface.co

Total runs: 5.8K
Run Growth: 1.1K
Growth Rate: 19.69%
Updated:May 23 2024
huggingface.co

Total runs: 3.5K
Run Growth: -3.6K
Growth Rate: -104.40%
Updated:March 30 2023
huggingface.co

Total runs: 664
Run Growth: 347
Growth Rate: 52.26%
Updated:December 27 2023
huggingface.co

Total runs: 541
Run Growth: 365
Growth Rate: 67.47%
Updated:July 08 2024
huggingface.co

Total runs: 358
Run Growth: 151
Growth Rate: 42.18%
Updated:March 30 2023
huggingface.co

Total runs: 244
Run Growth: 235
Growth Rate: 96.31%
Updated:July 08 2024
huggingface.co

Total runs: 192
Run Growth: 91
Growth Rate: 47.15%
Updated:March 26 2024
huggingface.co

Total runs: 175
Run Growth: 90
Growth Rate: 51.43%
Updated:July 20 2023
huggingface.co

Total runs: 129
Run Growth: 68
Growth Rate: 52.71%
Updated:June 28 2023
huggingface.co

Total runs: 128
Run Growth: 17
Growth Rate: 13.60%
Updated:March 26 2024
huggingface.co

Total runs: 111
Run Growth: 30
Growth Rate: 27.03%
Updated:July 11 2023
huggingface.co

Total runs: 53
Run Growth: 22
Growth Rate: 41.51%
Updated:August 21 2024
huggingface.co

Total runs: 46
Run Growth: 32
Growth Rate: 69.57%
Updated:November 13 2023
huggingface.co

Total runs: 34
Run Growth: 19
Growth Rate: 55.88%
Updated:July 08 2024
huggingface.co

Total runs: 33
Run Growth: 21
Growth Rate: 63.64%
Updated:November 23 2021
huggingface.co

Total runs: 32
Run Growth: 26
Growth Rate: 81.25%
Updated:March 01 2022
huggingface.co

Total runs: 25
Run Growth: 13
Growth Rate: 52.00%
Updated:September 15 2023
huggingface.co

Total runs: 10
Run Growth: 4
Growth Rate: 50.00%
Updated:June 11 2021
huggingface.co

Total runs: 10
Run Growth: -8
Growth Rate: -80.00%
Updated:May 21 2021
huggingface.co

Total runs: 8
Run Growth: 2
Growth Rate: 25.00%
Updated:May 07 2023
huggingface.co

Total runs: 6
Run Growth: 1
Growth Rate: 16.67%
Updated:March 08 2023
huggingface.co

Total runs: 5
Run Growth: -5
Growth Rate: -100.00%
Updated:May 04 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:February 10 2022
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:June 10 2021