Run them directly with
nexa-sdk
installed
In nexa-sdk CLI:
NexaAI/gemma-3n-E2B-it-4bit-MLX
Overview
Gemma is a family of lightweight, state-of-the-art open models from Google,
built from the same research and technology used to create the Gemini models.
Gemma 3n models are designed for efficient execution on low-resource devices.
They are capable of multimodal input, handling text, image, video, and audio
input, and generating text outputs, with open weights for pre-trained and
instruction-tuned variants. These models were trained with data in over 140
spoken languages.
Gemma 3n models use selective parameter activation technology to reduce resource
requirements. This technique allows the models to operate at an effective size
of 2B and 4B parameters, which is lower than the total number of parameters they
contain. For more information on Gemma 3n's efficient parameter management
technology, see the
Gemma 3n
page.
Inputs and outputs
Input:
Text string, such as a question, a prompt, or a document to be
summarized
Images, normalized to 256x256, 512x512, or 768x768 resolution
and encoded to 256 tokens each
Audio data encoded to 6.25 tokens per second from a single channel
Total input context of 32K tokens
Output:
Generated text in response to the input, such as an answer to a
question, analysis of image content, or a summary of a document
Total output length up to 32K tokens, subtracting the request
input tokens
Benchmark Results
These models were evaluated at full precision (float32) against a large
collection of different datasets and metrics to cover different aspects of
content generation. Evaluation results marked with
IT
are for
instruction-tuned models. Evaluation results marked with
PT
are for
pre-trained models.
gemma-3n-E2B-it-4bit-MLX huggingface.co is an AI model on huggingface.co that provides gemma-3n-E2B-it-4bit-MLX's model effect (), which can be used instantly with this NexaAI gemma-3n-E2B-it-4bit-MLX model. huggingface.co supports a free trial of the gemma-3n-E2B-it-4bit-MLX model, and also provides paid use of the gemma-3n-E2B-it-4bit-MLX. Support call gemma-3n-E2B-it-4bit-MLX model through api, including Node.js, Python, http.
gemma-3n-E2B-it-4bit-MLX huggingface.co is an online trial and call api platform, which integrates gemma-3n-E2B-it-4bit-MLX's modeling effects, including api services, and provides a free online trial of gemma-3n-E2B-it-4bit-MLX, you can try gemma-3n-E2B-it-4bit-MLX online for free by clicking the link below.
NexaAI gemma-3n-E2B-it-4bit-MLX online free url in huggingface.co:
gemma-3n-E2B-it-4bit-MLX is an open source model from GitHub that offers a free installation service, and any user can find gemma-3n-E2B-it-4bit-MLX on GitHub to install. At the same time, huggingface.co provides the effect of gemma-3n-E2B-it-4bit-MLX install, users can directly use gemma-3n-E2B-it-4bit-MLX installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
gemma-3n-E2B-it-4bit-MLX install url in huggingface.co: