deepcogito / cogito-v1-preview-llama-3B

huggingface.co
Total runs: 782
24-hour runs: -4
7-day runs: -219
30-day runs: -170
Model's Last Updated: April 08 2025
text-generation

Introduction of cogito-v1-preview-llama-3B

Model Details of cogito-v1-preview-llama-3B

Cogito v1 preview - 3B

NOTE
  • The model weights may be updated by Sunday, April 7th. However, these weights will just be a later checkpoint of the model currently being trained.
  • The base model (and therefore the model architecture) will remain the same. Similarly, the tokenizer will remain unchanged, as well as how to enable reasoning.
  • The complete description will be uploaded along with the evals results in the next few days.
Introduction

The Cogito LLMs are instruction tuned generative models (text in/text out). All models are released under an open license for commercial use.

  • Cogito models are hybrid reasoning models. You can pick when you want the model to answer normally and when you want it to think longer before answering.
  • They have significantly higher multilingual, coding and tool calling capabilities than their counterparts, and have been optimized for coding, STEM, instruction following, and general helpfulness.
  • Early testing demonstrates that Cogito v1-preview models significantly outperform their size equivalent counterparts on common industry benchmarks in the standard mode.
  • Similarly, in the reasoning mode, Cogito v1-preview models outperform their size equivalent reasoning model counterparts on common industry benchmarks.
Implementing extended thinking

This section will walk through how to use Cogito models to enable extended thinking (i.e., reasoning mode).

  • By default, the model will answer in the standard mode.
  • To enable thinking, you can do any one of the two methods:
    • Add a specific system prompt, or
    • Set enable_thinking=True during tokenization.
Method 1 - Add a specific system prompt.

To enable thinking, simply use this in the system prompt system_instruction = 'Enable deep thinking subroutine.'

If you already have a system_instruction, then use system_instruction = 'Enable deep thinking subroutine.' + '\n\n' + system_instruction .

Here is an example -

from transformers import AutoModelForCausalLM, AutoTokenizer

model_name = "deepcogito/cogito-v1-preview-llama-3B"

model = AutoModelForCausalLM.from_pretrained(
    model_name,
    torch_dtype="auto",
    device_map="auto"
)
tokenizer = AutoTokenizer.from_pretrained(model_name)

DEEP_THINKING_INSTRUCTION = "Enable deep thinking subroutine."
prompt = "Write a bash script that takes a matrix represented as a string with format '[1,2],[3,4],[5,6]' and prints the transpose in the same format."

messages = [
    {"role": "system", "content": DEEP_THINKING_INSTRUCTION},
    {"role": "user", "content": prompt}
]

text = tokenizer.apply_chat_template(
    messages,
    tokenize=False,
    add_generation_prompt=True
)
model_inputs = tokenizer([text], return_tensors="pt").to(model.device)

generated_ids = model.generate(
    **model_inputs,
    max_new_tokens=512
)
generated_ids = [
    output_ids[len(input_ids):] for input_ids, output_ids in zip(model_inputs.input_ids, generated_ids)
]

response = tokenizer.batch_decode(generated_ids, skip_special_tokens=True)[0]

Similarly, if you have a system prompt, you can append the DEEP_THINKING_INSTRUCTION to the beginning in this way -

DEEP_THINKING_INSTRUCTION = "Enable deep thinking subroutine."

system_prompt = "Reply to each prompt with only code answers - no explanations."
prompt = "Write a bash script that takes a matrix represented as a string with format '[1,2],[3,4],[5,6]' and prints the transpose in the same format."


messages = [
    {"role": "system", "content": DEEP_THINKING_INSTRUCTION + '\n\n' + system_prompt},
    {"role": "user", "content": prompt}
]

text = tokenizer.apply_chat_template(
    messages,
    tokenize=False,
    add_generation_prompt=True
)
Method 2 - Set enable_thinking=True in the tokenizer

If you are using Huggingface tokenizers, then you can simply use add the argument enable_thinking=True to the tokenization (this option is added to the chat template.) Here is an example -

from transformers import AutoModelForCausalLM, AutoTokenizer

model_name = "deepcogito/cogito-v1-preview-llama-3B"

model = AutoModelForCausalLM.from_pretrained(
    model_name,
    torch_dtype="auto",
    device_map="auto"
)
tokenizer = AutoTokenizer.from_pretrained(model_name)

prompt = "Write a bash script that takes a matrix represented as a string with format '[1,2],[3,4],[5,6]' and prints the transpose in the same format."

messages = [{"role": "user", "content": prompt}]

# Add enable_thinking=True for thinking mode.
text = tokenizer.apply_chat_template(
    messages,
    tokenize=False,
    add_generation_prompt=True,
    enable_thinking=True
)

model_inputs = tokenizer([text], return_tensors="pt").to(model.device)

generated_ids = model.generate(
    **model_inputs,
    max_new_tokens=512
)
generated_ids = [
    output_ids[len(input_ids):] for input_ids, output_ids in zip(model_inputs.input_ids, generated_ids)
]

response = tokenizer.batch_decode(generated_ids, skip_special_tokens=True)[0]

Runs of deepcogito cogito-v1-preview-llama-3B on huggingface.co

782
Total runs
-4
24-hour runs
-24
3-day runs
-219
7-day runs
-170
30-day runs

More Information About cogito-v1-preview-llama-3B huggingface.co Model

More cogito-v1-preview-llama-3B license Visit here:

https://choosealicense.com/licenses/llama3.2

cogito-v1-preview-llama-3B huggingface.co

cogito-v1-preview-llama-3B huggingface.co is an AI model on huggingface.co that provides cogito-v1-preview-llama-3B's model effect (), which can be used instantly with this deepcogito cogito-v1-preview-llama-3B model. huggingface.co supports a free trial of the cogito-v1-preview-llama-3B model, and also provides paid use of the cogito-v1-preview-llama-3B. Support call cogito-v1-preview-llama-3B model through api, including Node.js, Python, http.

cogito-v1-preview-llama-3B huggingface.co Url

https://huggingface.co/deepcogito/cogito-v1-preview-llama-3B

deepcogito cogito-v1-preview-llama-3B online free

cogito-v1-preview-llama-3B huggingface.co is an online trial and call api platform, which integrates cogito-v1-preview-llama-3B's modeling effects, including api services, and provides a free online trial of cogito-v1-preview-llama-3B, you can try cogito-v1-preview-llama-3B online for free by clicking the link below.

deepcogito cogito-v1-preview-llama-3B online free url in huggingface.co:

https://huggingface.co/deepcogito/cogito-v1-preview-llama-3B

cogito-v1-preview-llama-3B install

cogito-v1-preview-llama-3B is an open source model from GitHub that offers a free installation service, and any user can find cogito-v1-preview-llama-3B on GitHub to install. At the same time, huggingface.co provides the effect of cogito-v1-preview-llama-3B install, users can directly use cogito-v1-preview-llama-3B installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

cogito-v1-preview-llama-3B install url in huggingface.co:

https://huggingface.co/deepcogito/cogito-v1-preview-llama-3B

Url of cogito-v1-preview-llama-3B

cogito-v1-preview-llama-3B huggingface.co Url

Provider of cogito-v1-preview-llama-3B huggingface.co

deepcogito
ORGANIZATIONS

Other API from deepcogito