This model is still under active optimization and may exhibit the following issues:
Unexpected silence during generation.
Suboptimal speech quality or speech performance in certain cases.
Hardware Recommendation
For low-latency deployment, we recommend using at least
8 NVIDIA H20 GPUs
. This recommendation reflects the current optimization level and may change as inference performance improves.
Warm-up Requirement
The model should be warmed up after deployment. Without warm-up, the first response packet can have significantly higher latency.
DuplexOmni huggingface.co is an AI model on huggingface.co that provides DuplexOmni's model effect (), which can be used instantly with this MuyeHuang DuplexOmni model. huggingface.co supports a free trial of the DuplexOmni model, and also provides paid use of the DuplexOmni. Support call DuplexOmni model through api, including Node.js, Python, http.
DuplexOmni huggingface.co is an online trial and call api platform, which integrates DuplexOmni's modeling effects, including api services, and provides a free online trial of DuplexOmni, you can try DuplexOmni online for free by clicking the link below.
MuyeHuang DuplexOmni online free url in huggingface.co:
DuplexOmni is an open source model from GitHub that offers a free installation service, and any user can find DuplexOmni on GitHub to install. At the same time, huggingface.co provides the effect of DuplexOmni install, users can directly use DuplexOmni installed effect in huggingface.co for debugging and trial. It also supports api for free installation.