tiiuae / Falcon3-1B-Instruct-GPTQ-Int4

huggingface.co
Total runs: 59
24-hour runs: 0
7-day runs: 1
30-day runs: -149
Model's Last Updated: January 13 2025

Introduction of Falcon3-1B-Instruct-GPTQ-Int4

Model Details of Falcon3-1B-Instruct-GPTQ-Int4

drawing

Falcon3-1B-Instruct-GPTQ-Int4

Falcon3 family of Open Foundation Models is a set of pretrained and instruct LLMs ranging from 1B to 10B parameters.

Falcon3-1B-Instruct achieves strong results on reasoning, language understanding, instruction following, code and mathematics tasks. Falcon3-1B-Instruct supports 4 languages (English, French, Spanish, Portuguese) and a context length of up to 8K.

Model Details
  • Architecture
    • Transformer-based causal decoder-only architecture
    • 18 decoder blocks
    • Grouped Query Attention (GQA) for faster inference: 8 query heads and 4 key-value heads
    • Wider head dimension: 256
    • High RoPE value to support long context understanding: 1000042
    • Uses SwiGLU and RMSNorm
    • 8K context length
    • 131K vocab size
  • Pruned and healed using larger Falcon models (3B and 7B respectively) on only 80 Gigatokens of datasets comprising of web, code, STEM, high quality and multilingual data using 256 H100 GPU chips
  • Posttrained on 1.2 million samples of STEM, conversational, code, safety and function call data
  • Supports EN, FR, ES, PT
  • Developed by Technology Innovation Institute
  • License: TII Falcon-LLM License 2.0
  • Model Release Date: December 2024
  • Quantization: GPTQ 4-bit
Getting started
Click to expand
from transformers import AutoTokenizer, AutoModelForCausalLM

model_name = "tiiuae/Falcon3-1B-Instruct-GPTQ-Int4"

model = AutoModelForCausalLM.from_pretrained(
    model_name,
    torch_dtype="auto",
    device_map="auto"
)
tokenizer = AutoTokenizer.from_pretrained(model_name)

prompt = "How many hours in one day?"

messages = [
    {"role": "system", "content": "You are a helpful friendly assistant Falcon3 from TII, try to follow instructions as much as possible."},
    {"role": "user", "content": prompt}
]
text = tokenizer.apply_chat_template(
    messages,
    tokenize=False,
    add_generation_prompt=True
)
model_inputs = tokenizer([text], return_tensors="pt").to(model.device)

generated_ids = model.generate(
    **model_inputs,
    max_new_tokens=1024
)
generated_ids = [
    output_ids[len(input_ids):] for input_ids, output_ids in zip(model_inputs.input_ids, generated_ids)
]

response = tokenizer.batch_decode(generated_ids, skip_special_tokens=True)[0]
print(response)

Benchmarks

We report in the following table our internal pipeline benchmarks:

Benchmark Falcon3-1B-Instruct Falcon3-1B-Instruct-GPTQ-Int8 Falcon3-1B-Instruct-AWQ Falcon3-1B-Instruct-GPTQ-Int4
MMLU 43.6 43.5 43.0 42.6
MMLU-PRO 18.5 18.5 17.3 17.7
IFEval 54.9 56.1 51.2 51.4
Useful links
Technical Report

Coming soon....

Citation

If the Falcon3 family of models were helpful to your work, feel free to give us a cite.

@misc{Falcon3,
    title = {The Falcon 3 Family of Open Models},
    url = {https://huggingface.co/blog/falcon3},
    author = {Falcon-LLM Team},
    month = {December},
    year = {2024}
}

Runs of tiiuae Falcon3-1B-Instruct-GPTQ-Int4 on huggingface.co

59
Total runs
0
24-hour runs
0
3-day runs
1
7-day runs
-149
30-day runs

More Information About Falcon3-1B-Instruct-GPTQ-Int4 huggingface.co Model

More Falcon3-1B-Instruct-GPTQ-Int4 license Visit here:

https://choosealicense.com/licenses/falcon-llm-license

Falcon3-1B-Instruct-GPTQ-Int4 huggingface.co

Falcon3-1B-Instruct-GPTQ-Int4 huggingface.co is an AI model on huggingface.co that provides Falcon3-1B-Instruct-GPTQ-Int4's model effect (), which can be used instantly with this tiiuae Falcon3-1B-Instruct-GPTQ-Int4 model. huggingface.co supports a free trial of the Falcon3-1B-Instruct-GPTQ-Int4 model, and also provides paid use of the Falcon3-1B-Instruct-GPTQ-Int4. Support call Falcon3-1B-Instruct-GPTQ-Int4 model through api, including Node.js, Python, http.

Falcon3-1B-Instruct-GPTQ-Int4 huggingface.co Url

https://huggingface.co/tiiuae/Falcon3-1B-Instruct-GPTQ-Int4

tiiuae Falcon3-1B-Instruct-GPTQ-Int4 online free

Falcon3-1B-Instruct-GPTQ-Int4 huggingface.co is an online trial and call api platform, which integrates Falcon3-1B-Instruct-GPTQ-Int4's modeling effects, including api services, and provides a free online trial of Falcon3-1B-Instruct-GPTQ-Int4, you can try Falcon3-1B-Instruct-GPTQ-Int4 online for free by clicking the link below.

tiiuae Falcon3-1B-Instruct-GPTQ-Int4 online free url in huggingface.co:

https://huggingface.co/tiiuae/Falcon3-1B-Instruct-GPTQ-Int4

Falcon3-1B-Instruct-GPTQ-Int4 install

Falcon3-1B-Instruct-GPTQ-Int4 is an open source model from GitHub that offers a free installation service, and any user can find Falcon3-1B-Instruct-GPTQ-Int4 on GitHub to install. At the same time, huggingface.co provides the effect of Falcon3-1B-Instruct-GPTQ-Int4 install, users can directly use Falcon3-1B-Instruct-GPTQ-Int4 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Falcon3-1B-Instruct-GPTQ-Int4 install url in huggingface.co:

https://huggingface.co/tiiuae/Falcon3-1B-Instruct-GPTQ-Int4

Url of Falcon3-1B-Instruct-GPTQ-Int4

Falcon3-1B-Instruct-GPTQ-Int4 huggingface.co Url

Provider of Falcon3-1B-Instruct-GPTQ-Int4 huggingface.co

tiiuae
ORGANIZATIONS

Other API from tiiuae

huggingface.co

Total runs: 347.0K
Run Growth: -749.9K
Growth Rate: -216.33%
Updated:October 12 2024
huggingface.co

Total runs: 44.6K
Run Growth: -125.2K
Growth Rate: -280.36%
Updated:December 17 2024
huggingface.co

Total runs: 7.8K
Run Growth: -13.9K
Growth Rate: -172.36%
Updated:August 09 2024
huggingface.co

Total runs: 7.7K
Run Growth: 3.3K
Growth Rate: 44.06%
Updated:July 13 2023
huggingface.co

Total runs: 4.6K
Run Growth: 1.6K
Growth Rate: 34.36%
Updated:September 12 2026
huggingface.co

Total runs: 4.5K
Run Growth: -222
Growth Rate: -4.88%
Updated:December 17 2024
huggingface.co

Total runs: 2.5K
Run Growth: 1.2K
Growth Rate: 49.15%
Updated:January 21 2026
huggingface.co

Total runs: 2.4K
Run Growth: 2.4K
Growth Rate: 100.00%
Updated:December 24 2025
huggingface.co

Total runs: 384
Run Growth: -271
Growth Rate: -71.50%
Updated:November 07 2024
huggingface.co

Total runs: 72
Run Growth: 4
Growth Rate: 5.41%
Updated:September 06 2023
huggingface.co

Total runs: 44
Run Growth: 44
Growth Rate: 100.00%
Updated:March 12 2026