aufklarer / CamPlusPlus-Speaker-CoreML

huggingface.co
Total runs: 474
24-hour runs: -19
7-day runs: -54
30-day runs: -203
Model's Last Updated: May 28 2026
audio-classification

Introduction of CamPlusPlus-Speaker-CoreML

Model Details of CamPlusPlus-Speaker-CoreML

CAM++ Speaker Embedding (CoreML)

CoreML-converted CAM++ (Context-Aware Masking++) speaker embedding model for Apple Silicon.

Produces 192-dimensional speaker embeddings compatible with CosyVoice3 voice cloning.

Model Details
  • Architecture : D-TDNN (Densely-connected Time Delay Neural Network) with context-aware masking and multi-granularity pooling
  • Parameters : 6.9M
  • Input : 80-dim log-mel features, variable length
  • Output : 192-dim speaker embedding
  • Format : CoreML .mlmodelc (compiled, FP16)
  • Size : ~14 MB
Input/Output
Tensor Shape Description
mel_features [1, T, 80] 80-dim log-mel spectrogram (T = 10-3000 frames)
embedding [1, 192] L2-normalizable speaker embedding
Conversion

Converted from the official campplus.onnx shipped with Fun-CosyVoice3-0.5B-2512 :

ONNX → onnx2torch (PyTorch) → torch.jit.trace → coremltools → CoreML FP16

One ONNX op patched: ReduceProd ReduceSum in stats pooling (single-element tensor, mathematically equivalent).

Verified: CoreML vs ONNX max diff = 0.015 (FP16 precision).

Usage

Used by speech-swift for CosyVoice3 voice cloning:

// Extract 192-dim speaker embedding for CosyVoice3 voice cloning
let embedding = try camPlusPlus.embed(audio: samples, sampleRate: 16000)
let audio = model.synthesize(text: "Hello", speakerEmbedding: embedding)
Original Model
License

Apache-2.0 (same as original 3D-Speaker)

Runs of aufklarer CamPlusPlus-Speaker-CoreML on huggingface.co

474
Total runs
-19
24-hour runs
-34
3-day runs
-54
7-day runs
-203
30-day runs

More Information About CamPlusPlus-Speaker-CoreML huggingface.co Model

More CamPlusPlus-Speaker-CoreML license Visit here:

https://choosealicense.com/licenses/apache-2.0

CamPlusPlus-Speaker-CoreML huggingface.co

CamPlusPlus-Speaker-CoreML huggingface.co is an AI model on huggingface.co that provides CamPlusPlus-Speaker-CoreML's model effect (), which can be used instantly with this aufklarer CamPlusPlus-Speaker-CoreML model. huggingface.co supports a free trial of the CamPlusPlus-Speaker-CoreML model, and also provides paid use of the CamPlusPlus-Speaker-CoreML. Support call CamPlusPlus-Speaker-CoreML model through api, including Node.js, Python, http.

CamPlusPlus-Speaker-CoreML huggingface.co Url

https://huggingface.co/aufklarer/CamPlusPlus-Speaker-CoreML

aufklarer CamPlusPlus-Speaker-CoreML online free

CamPlusPlus-Speaker-CoreML huggingface.co is an online trial and call api platform, which integrates CamPlusPlus-Speaker-CoreML's modeling effects, including api services, and provides a free online trial of CamPlusPlus-Speaker-CoreML, you can try CamPlusPlus-Speaker-CoreML online for free by clicking the link below.

aufklarer CamPlusPlus-Speaker-CoreML online free url in huggingface.co:

https://huggingface.co/aufklarer/CamPlusPlus-Speaker-CoreML

CamPlusPlus-Speaker-CoreML install

CamPlusPlus-Speaker-CoreML is an open source model from GitHub that offers a free installation service, and any user can find CamPlusPlus-Speaker-CoreML on GitHub to install. At the same time, huggingface.co provides the effect of CamPlusPlus-Speaker-CoreML install, users can directly use CamPlusPlus-Speaker-CoreML installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

CamPlusPlus-Speaker-CoreML install url in huggingface.co:

https://huggingface.co/aufklarer/CamPlusPlus-Speaker-CoreML

Url of CamPlusPlus-Speaker-CoreML

CamPlusPlus-Speaker-CoreML huggingface.co Url

Provider of CamPlusPlus-Speaker-CoreML huggingface.co

aufklarer
ORGANIZATIONS

Other API from aufklarer

huggingface.co

Total runs: 2.5K
Run Growth: 2.3K
Growth Rate: 94.66%
Updated:September 16 2025
huggingface.co

Total runs: 151
Run Growth: 93
Growth Rate: 61.59%
Updated:October 15 2025