CohereLabs / aya-23-8B

huggingface.co
Total runs: 8.7K
24-hour runs: 122
7-day runs: -599
30-day runs: -1.3K
Model's Last Updated: September 11 2025
text-generation

Introduction of aya-23-8B

Model Details of aya-23-8B

Model Card for Aya-23-8B

Note: This is an older version of Aya. The latest version is Aya Expanse 8B which is available here . We also have multimodal variant, Aya Vision 8B which is available here .

Try Aya Expanse and Aya Vision:

You can try out latest Aya models before downloading the weights in our hosted Hugging Face Space here .

Model Summary

Aya 23 is an open weights research release of an instruction fine-tuned model with highly advanced multilingual capabilities. Aya 23 focuses on pairing a highly performant pre-trained Command family of models with the recently released Aya Collection . The result is a powerful multilingual large language model serving 23 languages.

This model card corresponds to the 8-billion version of the Aya 23 model. We also released a 35-billion version which you can find here .

We cover 23 languages: Arabic, Chinese (simplified & traditional), Czech, Dutch, English, French, German, Greek, Hebrew, Hindi, Indonesian, Italian, Japanese, Korean, Persian, Polish, Portuguese, Romanian, Russian, Spanish, Turkish, Ukrainian, and Vietnamese

Developed by: Cohere For AI and Cohere

Usage

Please install transformers from the source repository that includes the necessary changes for this model

# pip install transformers==4.41.1
from transformers import AutoTokenizer, AutoModelForCausalLM

model_id = "CohereForAI/aya-23-8B"
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(model_id)

# Format message with the command-r-plus chat template
messages = [{"role": "user", "content": "Anneme onu ne kadar sevdiğimi anlatan bir mektup yaz"}]
input_ids = tokenizer.apply_chat_template(messages, tokenize=True, add_generation_prompt=True, return_tensors="pt")
## <BOS_TOKEN><|START_OF_TURN_TOKEN|><|USER_TOKEN|>Anneme onu ne kadar sevdiğimi anlatan bir mektup yaz<|END_OF_TURN_TOKEN|><|START_OF_TURN_TOKEN|><|CHATBOT_TOKEN|>

gen_tokens = model.generate(
    input_ids, 
    max_new_tokens=100, 
    do_sample=True, 
    temperature=0.3,
    )

gen_text = tokenizer.decode(gen_tokens[0])
print(gen_text)
Example Notebook

This notebook showcases a detailed use of Aya 23 (8B) including inference and fine-tuning with QLoRA .

Model Details

Input : Models input text only.

Output : Models generate text only.

Model Architecture : Aya-23-8B is an auto-regressive language model that uses an optimized transformer architecture. After pretraining, this model is fine-tuned (IFT) to follow human instructions.

Languages covered : The model is particularly optimized for multilinguality and supports the following languages: Arabic, Chinese (simplified & traditional), Czech, Dutch, English, French, German, Greek, Hebrew, Hindi, Indonesian, Italian, Japanese, Korean, Persian, Polish, Portuguese, Romanian, Russian, Spanish, Turkish, Ukrainian, and Vietnamese

Context length : 8192

Evaluation
multilingual benchmarks average win rates

Please refer to the Aya 23 technical report for further details about the base model, data, instruction tuning, and evaluation.

Model Card Contact

For errors or additional questions about details in this model card, contact [email protected] .

Terms of Use

We hope that the release of this model will make community-based research efforts more accessible, by releasing the weights of a highly performant multilingual model to researchers all over the world. This model is governed by a CC-BY-NC License with an acceptable use addendum, and also requires adhering to C4AI's Acceptable Use Policy .

Try the model today

You can try Aya 23 in the Cohere playground here. You can also use it in our dedicated Hugging Face Space here .

Citation info
@misc{aryabumi2024aya,
      title={Aya 23: Open Weight Releases to Further Multilingual Progress}, 
      author={Viraat Aryabumi and John Dang and Dwarak Talupuru and Saurabh Dash and David Cairuz and Hangyu Lin and Bharat Venkitesh and Madeline Smith and Kelly Marchisio and Sebastian Ruder and Acyr Locatelli and Julia Kreutzer and Nick Frosst and Phil Blunsom and Marzieh Fadaee and Ahmet Üstün and Sara Hooker},
      year={2024},
      eprint={2405.15032},
      archivePrefix={arXiv},
      primaryClass={cs.CL}
}

Runs of CohereLabs aya-23-8B on huggingface.co

8.7K
Total runs
122
24-hour runs
-549
3-day runs
-599
7-day runs
-1.3K
30-day runs

More Information About aya-23-8B huggingface.co Model

More aya-23-8B license Visit here:

https://choosealicense.com/licenses/cc-by-nc-4.0

aya-23-8B huggingface.co

aya-23-8B huggingface.co is an AI model on huggingface.co that provides aya-23-8B's model effect (), which can be used instantly with this CohereLabs aya-23-8B model. huggingface.co supports a free trial of the aya-23-8B model, and also provides paid use of the aya-23-8B. Support call aya-23-8B model through api, including Node.js, Python, http.

CohereLabs aya-23-8B online free

aya-23-8B huggingface.co is an online trial and call api platform, which integrates aya-23-8B's modeling effects, including api services, and provides a free online trial of aya-23-8B, you can try aya-23-8B online for free by clicking the link below.

CohereLabs aya-23-8B online free url in huggingface.co:

https://huggingface.co/CohereLabs/aya-23-8B

aya-23-8B install

aya-23-8B is an open source model from GitHub that offers a free installation service, and any user can find aya-23-8B on GitHub to install. At the same time, huggingface.co provides the effect of aya-23-8B install, users can directly use aya-23-8B installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

aya-23-8B install url in huggingface.co:

https://huggingface.co/CohereLabs/aya-23-8B

Url of aya-23-8B

Provider of aya-23-8B huggingface.co

CohereLabs
ORGANIZATIONS

Other API from CohereLabs

huggingface.co

Total runs: 2.8K
Run Growth: -1.3K
Growth Rate: -46.85%
Updated:September 11 2025
huggingface.co

Total runs: 267
Run Growth: -44
Growth Rate: -16.79%
Updated:September 11 2025