MLX export of
LFM2.5-350M
for Apple Silicon inference.
LFM2.5-350M is a compact multilingual base model built on LiquidAI's hybrid architecture, combining convolutional and attention layers for efficient long-context processing.
Model Details
Property
Value
Parameters
350M
Precision
4-bit
Group Size
64
Size
212 MB
Context Length
128K
Use with mlx
pip install mlx-lm
from mlx_lm import load, generate
from mlx_lm.sample_utils import make_sampler
model, tokenizer = load("LiquidAI/LFM2.5-350M-MLX-4bit")
response = generate(
model,
tokenizer,
prompt="The capital of France is",
max_tokens=100,
sampler=make_sampler(temp=0.7),
verbose=True,
)
LFM2.5-350M-MLX-4bit huggingface.co is an AI model on huggingface.co that provides LFM2.5-350M-MLX-4bit's model effect (), which can be used instantly with this LiquidAI LFM2.5-350M-MLX-4bit model. huggingface.co supports a free trial of the LFM2.5-350M-MLX-4bit model, and also provides paid use of the LFM2.5-350M-MLX-4bit. Support call LFM2.5-350M-MLX-4bit model through api, including Node.js, Python, http.
LFM2.5-350M-MLX-4bit huggingface.co is an online trial and call api platform, which integrates LFM2.5-350M-MLX-4bit's modeling effects, including api services, and provides a free online trial of LFM2.5-350M-MLX-4bit, you can try LFM2.5-350M-MLX-4bit online for free by clicking the link below.
LiquidAI LFM2.5-350M-MLX-4bit online free url in huggingface.co:
LFM2.5-350M-MLX-4bit is an open source model from GitHub that offers a free installation service, and any user can find LFM2.5-350M-MLX-4bit on GitHub to install. At the same time, huggingface.co provides the effect of LFM2.5-350M-MLX-4bit install, users can directly use LFM2.5-350M-MLX-4bit installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
LFM2.5-350M-MLX-4bit install url in huggingface.co: