QuantFactory / Reasoning-Llama-1b-v0.1-GGUF

huggingface.co
Total runs: 926
24-hour runs: 0
7-day runs: 0
30-day runs: 0
Model's Last Updated: October 07 2024

Introduction of Reasoning-Llama-1b-v0.1-GGUF

Model Details of Reasoning-Llama-1b-v0.1-GGUF

QuantFactory Banner

QuantFactory/Reasoning-Llama-1b-v0.1-GGUF

This is quantized version of KingNish/Reasoning-Llama-1b-v0.1 created using llama.cpp

Original Model Card

Model Dexcription

It's First iteration of this model. For testing purpose its just trained on 10k rows. It performed very well than expected. It do first reasoning and than generate response on based on it but it do like o1. It do reasoning separately (Just like o1), no tags (like reflection). Below is inference code.

from transformers import AutoModelForCausalLM, AutoTokenizer

MAX_REASONING_TOKENS = 1024
MAX_RESPONSE_TOKENS = 512

model_name = "KingNish/Reasoning-Llama-1b-v0.1"

model = AutoModelForCausalLM.from_pretrained(model_name, torch_dtype="auto", device_map="auto")
tokenizer = AutoTokenizer.from_pretrained(model_name)

prompt = "Which is greater 9.9 or 9.11 ??"
messages = [
    {"role": "user", "content": prompt}
]

# Generate reasoning
reasoning_template = tokenizer.apply_chat_template(messages, tokenize=False, add_reasoning_prompt=True)
reasoning_inputs = tokenizer(reasoning_template, return_tensors="pt").to(model.device)
reasoning_ids = model.generate(**reasoning_inputs, max_new_tokens=MAX_REASONING_TOKENS)
reasoning_output = tokenizer.decode(reasoning_ids[0, reasoning_inputs.input_ids.shape[1]:], skip_special_tokens=True)

# print("REASONING: " + reasoning_output)

# Generate answer
messages.append({"role": "reasoning", "content": reasoning_output})
response_template = tokenizer.apply_chat_template(messages, tokenize=False, add_generation_prompt=True)
response_inputs = tokenizer(response_template, return_tensors="pt").to(model.device)
response_ids = model.generate(**response_inputs, max_new_tokens=MAX_RESPONSE_TOKENS)
response_output = tokenizer.decode(response_ids[0, response_inputs.input_ids.shape[1]:], skip_special_tokens=True)

print("ANSWER: " + response_output)

This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.

Runs of QuantFactory Reasoning-Llama-1b-v0.1-GGUF on huggingface.co

926
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
0
30-day runs

More Information About Reasoning-Llama-1b-v0.1-GGUF huggingface.co Model

More Reasoning-Llama-1b-v0.1-GGUF license Visit here:

https://choosealicense.com/licenses/llama3.2

Reasoning-Llama-1b-v0.1-GGUF huggingface.co

Reasoning-Llama-1b-v0.1-GGUF huggingface.co is an AI model on huggingface.co that provides Reasoning-Llama-1b-v0.1-GGUF's model effect (), which can be used instantly with this QuantFactory Reasoning-Llama-1b-v0.1-GGUF model. huggingface.co supports a free trial of the Reasoning-Llama-1b-v0.1-GGUF model, and also provides paid use of the Reasoning-Llama-1b-v0.1-GGUF. Support call Reasoning-Llama-1b-v0.1-GGUF model through api, including Node.js, Python, http.

Reasoning-Llama-1b-v0.1-GGUF huggingface.co Url

https://huggingface.co/QuantFactory/Reasoning-Llama-1b-v0.1-GGUF

QuantFactory Reasoning-Llama-1b-v0.1-GGUF online free

Reasoning-Llama-1b-v0.1-GGUF huggingface.co is an online trial and call api platform, which integrates Reasoning-Llama-1b-v0.1-GGUF's modeling effects, including api services, and provides a free online trial of Reasoning-Llama-1b-v0.1-GGUF, you can try Reasoning-Llama-1b-v0.1-GGUF online for free by clicking the link below.

QuantFactory Reasoning-Llama-1b-v0.1-GGUF online free url in huggingface.co:

https://huggingface.co/QuantFactory/Reasoning-Llama-1b-v0.1-GGUF

Reasoning-Llama-1b-v0.1-GGUF install

Reasoning-Llama-1b-v0.1-GGUF is an open source model from GitHub that offers a free installation service, and any user can find Reasoning-Llama-1b-v0.1-GGUF on GitHub to install. At the same time, huggingface.co provides the effect of Reasoning-Llama-1b-v0.1-GGUF install, users can directly use Reasoning-Llama-1b-v0.1-GGUF installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Reasoning-Llama-1b-v0.1-GGUF install url in huggingface.co:

https://huggingface.co/QuantFactory/Reasoning-Llama-1b-v0.1-GGUF

Url of Reasoning-Llama-1b-v0.1-GGUF

Reasoning-Llama-1b-v0.1-GGUF huggingface.co Url

Provider of Reasoning-Llama-1b-v0.1-GGUF huggingface.co

QuantFactory
ORGANIZATIONS

Other API from QuantFactory