By default, an 8-bit quantized version of the model is used, but you can choose to use the full-precision (fp32) version by specifying
{ quantized: false }
in the
pipeline
function:
Note: Having a separate repo for ONNX weights is intended to be a temporary solution until WebML gains more traction. If you would like to make your models web-ready, we recommend converting to ONNX using
🤗 Optimum
and structuring your repo like this one (with ONNX weights located in a subfolder named
onnx
).
Runs of Xenova gte-small on huggingface.co
56.1K
Total runs
0
24-hour runs
958
3-day runs
3.1K
7-day runs
549
30-day runs
More Information About gte-small huggingface.co Model
gte-small huggingface.co
gte-small huggingface.co is an AI model on huggingface.co that provides gte-small's model effect (), which can be used instantly with this Xenova gte-small model. huggingface.co supports a free trial of the gte-small model, and also provides paid use of the gte-small. Support call gte-small model through api, including Node.js, Python, http.
gte-small huggingface.co is an online trial and call api platform, which integrates gte-small's modeling effects, including api services, and provides a free online trial of gte-small, you can try gte-small online for free by clicking the link below.
Xenova gte-small online free url in huggingface.co:
gte-small is an open source model from GitHub that offers a free installation service, and any user can find gte-small on GitHub to install. At the same time, huggingface.co provides the effect of gte-small install, users can directly use gte-small installed effect in huggingface.co for debugging and trial. It also supports api for free installation.