aws-neuron / Llama-2-7b-chat-hf-seqlen-2048-bs-2

huggingface.co
Total runs: 17
24-hour runs: -1
7-day runs: 7
30-day runs: 8
Model's Last Updated: November 10 2023
text-generation

Introduction of Llama-2-7b-chat-hf-seqlen-2048-bs-2

Model Details of Llama-2-7b-chat-hf-seqlen-2048-bs-2

Neuronx model for meta-llama/Llama-2-7b-chat-hf

This repository contains are AWS Inferentia2 and neuronx compatible checkpoint for meta-llama/Llama-2-7b-chat-hf . You can find detailed information about the base model on its Model Card .

Usage on Amazon SageMaker

coming soon

Usage with optimum-neuron

from optimum.neuron import pipeline

# Load pipeline from Hugging Face repository
pipe = pipeline("text-generation", "aws-neuron/Llama-2-7b-chat-hf-seqlen-2048-bs-2")

# We use the tokenizer's chat template to format each message - see https://huggingface.co/docs/transformers/main/en/chat_templating
messages = [
    {"role": "user", "content": "What is 2+2?"},
]
prompt = pipe.tokenizer.apply_chat_template(messages, tokenize=False, add_generation_prompt=True)
# Run generation
outputs = pipe(prompt, max_new_tokens=256, do_sample=True, temperature=0.7, top_k=50, top_p=0.95)
print(outputs[0]["generated_text"])
Compilation Arguments

compilation arguments

{
  "num_cores": 2,
  "auto_cast_type": "fp16"
}

input_shapes

{
  "sequence_length": 2048,
  "batch_size": 2
}

Runs of aws-neuron Llama-2-7b-chat-hf-seqlen-2048-bs-2 on huggingface.co

17
Total runs
-1
24-hour runs
5
3-day runs
7
7-day runs
8
30-day runs

More Information About Llama-2-7b-chat-hf-seqlen-2048-bs-2 huggingface.co Model

Llama-2-7b-chat-hf-seqlen-2048-bs-2 huggingface.co

Llama-2-7b-chat-hf-seqlen-2048-bs-2 huggingface.co is an AI model on huggingface.co that provides Llama-2-7b-chat-hf-seqlen-2048-bs-2's model effect (), which can be used instantly with this aws-neuron Llama-2-7b-chat-hf-seqlen-2048-bs-2 model. huggingface.co supports a free trial of the Llama-2-7b-chat-hf-seqlen-2048-bs-2 model, and also provides paid use of the Llama-2-7b-chat-hf-seqlen-2048-bs-2. Support call Llama-2-7b-chat-hf-seqlen-2048-bs-2 model through api, including Node.js, Python, http.

Llama-2-7b-chat-hf-seqlen-2048-bs-2 huggingface.co Url

https://huggingface.co/aws-neuron/Llama-2-7b-chat-hf-seqlen-2048-bs-2

aws-neuron Llama-2-7b-chat-hf-seqlen-2048-bs-2 online free

Llama-2-7b-chat-hf-seqlen-2048-bs-2 huggingface.co is an online trial and call api platform, which integrates Llama-2-7b-chat-hf-seqlen-2048-bs-2's modeling effects, including api services, and provides a free online trial of Llama-2-7b-chat-hf-seqlen-2048-bs-2, you can try Llama-2-7b-chat-hf-seqlen-2048-bs-2 online for free by clicking the link below.

aws-neuron Llama-2-7b-chat-hf-seqlen-2048-bs-2 online free url in huggingface.co:

https://huggingface.co/aws-neuron/Llama-2-7b-chat-hf-seqlen-2048-bs-2

Llama-2-7b-chat-hf-seqlen-2048-bs-2 install

Llama-2-7b-chat-hf-seqlen-2048-bs-2 is an open source model from GitHub that offers a free installation service, and any user can find Llama-2-7b-chat-hf-seqlen-2048-bs-2 on GitHub to install. At the same time, huggingface.co provides the effect of Llama-2-7b-chat-hf-seqlen-2048-bs-2 install, users can directly use Llama-2-7b-chat-hf-seqlen-2048-bs-2 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Llama-2-7b-chat-hf-seqlen-2048-bs-2 install url in huggingface.co:

https://huggingface.co/aws-neuron/Llama-2-7b-chat-hf-seqlen-2048-bs-2

Url of Llama-2-7b-chat-hf-seqlen-2048-bs-2

Llama-2-7b-chat-hf-seqlen-2048-bs-2 huggingface.co Url

Provider of Llama-2-7b-chat-hf-seqlen-2048-bs-2 huggingface.co

aws-neuron
ORGANIZATIONS

Other API from aws-neuron