Mxode / NanoLM-70M-Instruct-v1

huggingface.co
Total runs: 11
24-hour runs: 0
7-day runs: 3
30-day runs: -14
Model's Last Updated: September 09 2024
text-generation

Introduction of NanoLM-70M-Instruct-v1

Model Details of NanoLM-70M-Instruct-v1

NanoLM-70M-Instruct-v1

English | 简体中文

Introduction

In order to explore the potential of small models, I have attempted to build a series of them, which are available in the NanoLM Collections .

This is NanoLM-70M-Instruct-v1. The model currently supports English only .

Model Details

The tokenizer and model architecture of NanoLM-70M-Instruct-v1 are the same as SmolLM-135M , but the number of layers has been reduced from 30 to 12.

Essentially, it is a pure LLaMA architecture, specifically LlamaForCausalLM.

As a result, NanoLM-70M-Instruct-v1 has only 70 million parameters.

Despite this, NanoLM-70M-Instruct-v1 still demonstrates instruction-following capabilities.

How to use
import torch
from transformers import AutoModelForCausalLM, AutoTokenizer

model_path = 'Mxode/NanoLM-70M-Instruct-v1'

model = AutoModelForCausalLM.from_pretrained(model_path).to('cuda:0', torch.bfloat16)
tokenizer = AutoTokenizer.from_pretrained(model_path)


text = "Why is it important for entrepreneurs to prioritize financial management?"
prompt = tokenizer.apply_chat_template(
    [
        {'role': 'system', 'content': 'You are a helpful assistant.'},
        {'role': 'user', 'content': text}
    ],
    add_generation_prompt=True,
    tokenize=True,
    return_tensors='pt'
).to('cuda:0')


outputs = model.generate(
    prompt,
    max_new_tokens=1024,
    do_sample=True,
    temperature=0.7,
    repetition_penalty=1.1,
    eos_token_id=tokenizer.eos_token_id,
)
response = tokenizer.decode(outputs[0])
print(response)

Runs of Mxode NanoLM-70M-Instruct-v1 on huggingface.co

11
Total runs
0
24-hour runs
2
3-day runs
3
7-day runs
-14
30-day runs

More Information About NanoLM-70M-Instruct-v1 huggingface.co Model

More NanoLM-70M-Instruct-v1 license Visit here:

https://choosealicense.com/licenses/gpl-3.0

NanoLM-70M-Instruct-v1 huggingface.co

NanoLM-70M-Instruct-v1 huggingface.co is an AI model on huggingface.co that provides NanoLM-70M-Instruct-v1's model effect (), which can be used instantly with this Mxode NanoLM-70M-Instruct-v1 model. huggingface.co supports a free trial of the NanoLM-70M-Instruct-v1 model, and also provides paid use of the NanoLM-70M-Instruct-v1. Support call NanoLM-70M-Instruct-v1 model through api, including Node.js, Python, http.

NanoLM-70M-Instruct-v1 huggingface.co Url

https://huggingface.co/Mxode/NanoLM-70M-Instruct-v1

Mxode NanoLM-70M-Instruct-v1 online free

NanoLM-70M-Instruct-v1 huggingface.co is an online trial and call api platform, which integrates NanoLM-70M-Instruct-v1's modeling effects, including api services, and provides a free online trial of NanoLM-70M-Instruct-v1, you can try NanoLM-70M-Instruct-v1 online for free by clicking the link below.

Mxode NanoLM-70M-Instruct-v1 online free url in huggingface.co:

https://huggingface.co/Mxode/NanoLM-70M-Instruct-v1

NanoLM-70M-Instruct-v1 install

NanoLM-70M-Instruct-v1 is an open source model from GitHub that offers a free installation service, and any user can find NanoLM-70M-Instruct-v1 on GitHub to install. At the same time, huggingface.co provides the effect of NanoLM-70M-Instruct-v1 install, users can directly use NanoLM-70M-Instruct-v1 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

NanoLM-70M-Instruct-v1 install url in huggingface.co:

https://huggingface.co/Mxode/NanoLM-70M-Instruct-v1

Url of NanoLM-70M-Instruct-v1

NanoLM-70M-Instruct-v1 huggingface.co Url

Provider of NanoLM-70M-Instruct-v1 huggingface.co

Mxode
ORGANIZATIONS

Other API from Mxode