MLX export of
LFM2.5-350M
for Apple Silicon inference.
LFM2.5-350M is a compact multilingual base model built on LiquidAI's hybrid architecture, combining convolutional and attention layers for efficient long-context processing.
Model Details
Property
Value
Parameters
350M
Precision
8-bit
Group Size
64
Size
381 MB
Context Length
128K
Use with mlx
pip install mlx-lm
from mlx_lm import load, generate
from mlx_lm.sample_utils import make_sampler
model, tokenizer = load("LiquidAI/LFM2.5-350M-MLX-8bit")
response = generate(
model,
tokenizer,
prompt="The capital of France is",
max_tokens=100,
sampler=make_sampler(temp=0.7),
verbose=True,
)
LFM2.5-350M-MLX-8bit huggingface.co is an AI model on huggingface.co that provides LFM2.5-350M-MLX-8bit's model effect (), which can be used instantly with this LiquidAI LFM2.5-350M-MLX-8bit model. huggingface.co supports a free trial of the LFM2.5-350M-MLX-8bit model, and also provides paid use of the LFM2.5-350M-MLX-8bit. Support call LFM2.5-350M-MLX-8bit model through api, including Node.js, Python, http.
LFM2.5-350M-MLX-8bit huggingface.co is an online trial and call api platform, which integrates LFM2.5-350M-MLX-8bit's modeling effects, including api services, and provides a free online trial of LFM2.5-350M-MLX-8bit, you can try LFM2.5-350M-MLX-8bit online for free by clicking the link below.
LiquidAI LFM2.5-350M-MLX-8bit online free url in huggingface.co:
LFM2.5-350M-MLX-8bit is an open source model from GitHub that offers a free installation service, and any user can find LFM2.5-350M-MLX-8bit on GitHub to install. At the same time, huggingface.co provides the effect of LFM2.5-350M-MLX-8bit install, users can directly use LFM2.5-350M-MLX-8bit installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
LFM2.5-350M-MLX-8bit install url in huggingface.co: