HuggingFace-WavLM-Base-Plus: Optimized for Mobile Deployment
Real-time Speech processing
HuggingFaceWavLMBasePlus is a real time speech processing backbone based on Microsoft's WavLM model.
This model is an implementation of HuggingFace-WavLM-Base-Plus found
here
.
This repository provides scripts to run HuggingFace-WavLM-Base-Plus on Qualcomm® devices.
More details on model performance across various devices, can be found
here
.
Model Details
Model Type:
Speech recognition
Model Stats:
Model checkpoint: wavlm-libri-clean-100h-base-plus
Profile Job summary of HuggingFace-WavLM-Base-Plus
--------------------------------------------------
Device: QCS8550 (Proxy) (12)
Estimated Inference Time: 848.86 ms
Estimated Peak Memory Range: 126.36-128.51 MB
Compute Units: CPU (811) | Total (811)
How does this work?
This
export script
leverages
Qualcomm® AI Hub
to optimize, validate, and deploy this model
on-device. Lets go through each step below in detail:
Step 1:
Compile model for on-device deployment
To compile a PyTorch model for on-device deployment, we first trace the model
in memory using the
jit.trace
and then call the
submit_compile_job
API.
import torch
import qai_hub as hub
from qai_hub_models.models.huggingface_wavlm_base_plus import Model
# Load the model
torch_model = Model.from_pretrained()
# Device
device = hub.Device("Samsung Galaxy S23")
# Trace model
input_shape = torch_model.get_input_spec()
sample_inputs = torch_model.sample_inputs()
pt_model = torch.jit.trace(torch_model, [torch.tensor(data[0]) for _, data in sample_inputs.items()])
# Compile model on a specific device
compile_job = hub.submit_compile_job(
model=pt_model,
device=device,
input_specs=torch_model.get_input_spec(),
)
# Get target model to run on-device
target_model = compile_job.get_target_model()
Step 2:
Performance profiling on cloud-hosted device
After compiling models from step 1. Models can be profiled model on-device using the
target_model
. Note that this scripts runs the model on a device automatically
provisioned in the cloud. Once the job is submitted, you can navigate to a
provided job URL to view a variety of on-device performance metrics.
HuggingFace-WavLM-Base-Plus huggingface.co is an AI model on huggingface.co that provides HuggingFace-WavLM-Base-Plus's model effect (), which can be used instantly with this qualcomm HuggingFace-WavLM-Base-Plus model. huggingface.co supports a free trial of the HuggingFace-WavLM-Base-Plus model, and also provides paid use of the HuggingFace-WavLM-Base-Plus. Support call HuggingFace-WavLM-Base-Plus model through api, including Node.js, Python, http.
HuggingFace-WavLM-Base-Plus huggingface.co is an online trial and call api platform, which integrates HuggingFace-WavLM-Base-Plus's modeling effects, including api services, and provides a free online trial of HuggingFace-WavLM-Base-Plus, you can try HuggingFace-WavLM-Base-Plus online for free by clicking the link below.
qualcomm HuggingFace-WavLM-Base-Plus online free url in huggingface.co:
HuggingFace-WavLM-Base-Plus is an open source model from GitHub that offers a free installation service, and any user can find HuggingFace-WavLM-Base-Plus on GitHub to install. At the same time, huggingface.co provides the effect of HuggingFace-WavLM-Base-Plus install, users can directly use HuggingFace-WavLM-Base-Plus installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
HuggingFace-WavLM-Base-Plus install url in huggingface.co: