This model helps to classify speakers from the frequency domain representation of speech recordings, obtained via Fast Fourier Transform (FFT).
The model is created by a 1D convolutional network with residual connections for audio classification.
This should be run with
TensorFlow 2.3
or higher, or
tf-nightly
.
Also, The noise samples in the dataset need to be resampled to a sampling rate of 16000 Hz before using for this model so, In order to do this, you will need to have installed
ffmpg
.
Training and evaluation data
During dataset preparation, the speech samples & background noise samples were sorted and categorized into 2 folders - audio & noise, and then noise samples were resampled to 16000Hz & then the background noise was added to the speech samples to augment the data. After that, the FFT of these samples was given to the model for the training & evaluation part.
Training procedure
Training hyperparameters
The following hyperparameters were used during training:
Runs of keras-io speaker-recognition on huggingface.co
7
Total runs
0
24-hour runs
0
3-day runs
2
7-day runs
1
30-day runs
More Information About speaker-recognition huggingface.co Model
speaker-recognition huggingface.co
speaker-recognition huggingface.co is an AI model on huggingface.co that provides speaker-recognition's model effect (), which can be used instantly with this keras-io speaker-recognition model. huggingface.co supports a free trial of the speaker-recognition model, and also provides paid use of the speaker-recognition. Support call speaker-recognition model through api, including Node.js, Python, http.
speaker-recognition huggingface.co is an online trial and call api platform, which integrates speaker-recognition's modeling effects, including api services, and provides a free online trial of speaker-recognition, you can try speaker-recognition online for free by clicking the link below.
keras-io speaker-recognition online free url in huggingface.co:
speaker-recognition is an open source model from GitHub that offers a free installation service, and any user can find speaker-recognition on GitHub to install. At the same time, huggingface.co provides the effect of speaker-recognition install, users can directly use speaker-recognition installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
speaker-recognition install url in huggingface.co: