Mozilla / SFR-Embedding-Mistral-llamafile

huggingface.co
Total runs: 16
24-hour runs: 0
7-day runs: -5
30-day runs: 6
Model's Last Updated: July 03 2024
feature-extraction

Introduction of SFR-Embedding-Mistral-llamafile

Model Details of SFR-Embedding-Mistral-llamafile

SFR-Embedding-Mistral - llamafile

This repository contains executable weights (which we call llamafiles ) that run on Linux, MacOS, Windows, FreeBSD, OpenBSD, and NetBSD for AMD64 and ARM64.

Quickstart

Running the following on a desktop OS will launch a server on http://localhost:8080 to which you can send HTTP requests to in order to get embeddings:

chmod +x ggml-sfr-embedding-mistral-f16.llamafile
./ggml-sfr-embedding-mistral-f16.llamafile --server --nobrowser --embedding

Then, you can use your favorite HTTP client to call the server's /embedding endpoint:

curl \
-X POST \
-H "Content-Type: application/json" \
-d '{"content": "Hello, world!"}' \
http://localhost:8080/embedding

For further information, please see the llamafile README and the llamafile server docs .

Having trouble? See the "Gotchas" section of the README or contact us on Discord .

About llamafile

llamafile is a new format introduced by Mozilla Ocho on Nov 20th 2023. It uses Cosmopolitan Libc to turn LLM weights into runnable llama.cpp binaries that run on the stock installs of six OSes for both ARM64 and AMD64.

About Quantization Formats

Your choice of quantization format depends on three things:

  1. Will it fit in RAM or VRAM?
  2. Is your use case reading (e.g. summarization) or writing (e.g. chatbot)?
  3. llamafiles bigger than 4.30 GB are hard to run on Windows (see gotchas )

Good quants for writing (eval speed) are Q5_K_M, and Q4_0. Text generation is bounded by memory speed, so smaller quants help, but they also cause the LLM to hallucinate more.

Good quants for reading (prompt eval speed) are BF16, F16, Q4_0, and Q8_0 (ordered from fastest to slowest). Prompt evaluation is bounded by computation speed (flops) so simpler quants help.

Note: BF16 is currently only supported on CPU.

See also: https://huggingface.co/docs/hub/en/gguf#quantization-types


Model Card

See Salesforce/SFR-Embedding-Mistral

Runs of Mozilla SFR-Embedding-Mistral-llamafile on huggingface.co

16
Total runs
0
24-hour runs
0
3-day runs
-5
7-day runs
6
30-day runs

More Information About SFR-Embedding-Mistral-llamafile huggingface.co Model

More SFR-Embedding-Mistral-llamafile license Visit here:

https://choosealicense.com/licenses/cc-by-nc-4.0

SFR-Embedding-Mistral-llamafile huggingface.co

SFR-Embedding-Mistral-llamafile huggingface.co is an AI model on huggingface.co that provides SFR-Embedding-Mistral-llamafile's model effect (), which can be used instantly with this Mozilla SFR-Embedding-Mistral-llamafile model. huggingface.co supports a free trial of the SFR-Embedding-Mistral-llamafile model, and also provides paid use of the SFR-Embedding-Mistral-llamafile. Support call SFR-Embedding-Mistral-llamafile model through api, including Node.js, Python, http.

SFR-Embedding-Mistral-llamafile huggingface.co Url

https://huggingface.co/Mozilla/SFR-Embedding-Mistral-llamafile

Mozilla SFR-Embedding-Mistral-llamafile online free

SFR-Embedding-Mistral-llamafile huggingface.co is an online trial and call api platform, which integrates SFR-Embedding-Mistral-llamafile's modeling effects, including api services, and provides a free online trial of SFR-Embedding-Mistral-llamafile, you can try SFR-Embedding-Mistral-llamafile online for free by clicking the link below.

Mozilla SFR-Embedding-Mistral-llamafile online free url in huggingface.co:

https://huggingface.co/Mozilla/SFR-Embedding-Mistral-llamafile

SFR-Embedding-Mistral-llamafile install

SFR-Embedding-Mistral-llamafile is an open source model from GitHub that offers a free installation service, and any user can find SFR-Embedding-Mistral-llamafile on GitHub to install. At the same time, huggingface.co provides the effect of SFR-Embedding-Mistral-llamafile install, users can directly use SFR-Embedding-Mistral-llamafile installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

SFR-Embedding-Mistral-llamafile install url in huggingface.co:

https://huggingface.co/Mozilla/SFR-Embedding-Mistral-llamafile

Url of SFR-Embedding-Mistral-llamafile

SFR-Embedding-Mistral-llamafile huggingface.co Url

Provider of SFR-Embedding-Mistral-llamafile huggingface.co

Mozilla
ORGANIZATIONS

Other API from Mozilla

huggingface.co

Total runs: 2.8K
Run Growth: 905
Growth Rate: 32.30%
Updated:October 02 2024
huggingface.co

Total runs: 374
Run Growth: 144
Growth Rate: 38.50%
Updated:November 25 2024
huggingface.co

Total runs: 41
Run Growth: 7
Growth Rate: 17.07%
Updated:August 20 2025
huggingface.co

Total runs: 12
Run Growth: -14
Growth Rate: -116.67%
Updated:May 05 2024
huggingface.co

Total runs: 4
Run Growth: 4
Growth Rate: 100.00%
Updated:October 16 2024
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:March 04 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:April 04 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:September 15 2025