acul3 / chatterbox-executorch

huggingface.co
Total runs: 291
24-hour runs: 4
7-day runs: 30
30-day runs: 93
Model's Last Updated: March 22 2026
text-to-speech

Introduction of chatterbox-executorch

Model Details of chatterbox-executorch

Chatterbox Multilingual TTS — ExecuTorch Models

Pre-exported .pte model files for running Resemble AI's Chatterbox Multilingual TTS fully on-device using ExecuTorch .

📦 Code & export scripts: acul3/chatterbox-executorch on GitHub


What's Here

9 ExecuTorch .pte files covering the complete TTS pipeline — from text input to 24kHz waveform — with zero PyTorch runtime required:

File Size Backend Precision Stage
voice_encoder.pte 7 MB portable FP32 Speaker embedding
xvector_encoder.pte 27 MB portable FP32 X-vector conditioning
t3_cond_speech_emb.pte 49 MB portable FP32 Speech token embedding
t3_cond_enc.pte 18 MB portable FP32 Text/conditioning encoder
t3_prefill.pte 1010 MB XNNPACK FP16 T3 Transformer prefill
t3_decode.pte 1002 MB XNNPACK FP16 T3 Transformer decode
s3gen_encoder.pte 178 MB portable FP32 S3Gen Conformer encoder
cfm_step.pte 274 MB XNNPACK FP32 CFM flow matching step
hifigan.pte 84 MB XNNPACK FP32 HiFiGAN vocoder
Total ~2.6 GB

Quick Download
from huggingface_hub import snapshot_download

snapshot_download(
    "acul3/chatterbox-executorch",
    local_dir="et_models",
    repo_type="model"
)

Pipeline Overview
Text → MTLTokenizer → text tokens
Reference Audio → VoiceEncoder + CAMPPlus → speaker conditioning
                          ↓
              T3 Prefill (LlamaModel, conditioned)
                          ↓
              T3 Decode (autoregressive, ~100 tokens)
                          ↓
              S3Gen Encoder (Conformer)
                          ↓
              CFM Step × 2 (flow matching)
                          ↓
              HiFiGAN (vocoder, chunked)
                          ↓
              24kHz PCM waveform 🎵

Key Technical Notes
  • T3 Decode uses a manually unrolled 30-layer Llama forward pass with static KV cache ( torch.where writes) — bypasses HF DynamicCache for torch.export compatibility
  • HiFiGAN uses a manual real-valued DFT (cosine/sine matrix multiply) — replaces torch.stft / torch.istft which XNNPACK doesn't support
  • T3 models are FP16 (XNNPACK half-precision kernels) — ~half the size of FP32 with near-identical quality
  • Fixed shapes: CFM expects T_MEL=2200 , HiFiGAN expects T_MEL=300 (use chunked processing for longer audio)

Usage

See the GitHub repo for full inference code: acul3/chatterbox-executorch

# Clone code
git clone https://github.com/acul3/chatterbox-executorch.git
cd chatterbox-executorch

# Download models (this repo)
python -c "
from huggingface_hub import snapshot_download
snapshot_download('acul3/chatterbox-executorch', local_dir='et_models', repo_type='model')
"

# Run full PTE inference
python test_true_full_pte.py

Android Integration

These models are designed for Android deployment via the ExecuTorch Android SDK . Load with:

val module = Module.load(context.filesDir.path + "/t3_prefill.pte")

With QNN/NPU delegation on a Snapdragon device, expect 10–50× speedup over the CPU timings below.

Performance (Jetson AGX Orin, CPU only)
Stage Time
Voice encoding ~1s
T3 prefill ~22s
T3 decode (~100 tokens) 800s total ( 8s/token)
S3Gen encoder ~2s
CFM (2 steps) ~40s
HiFiGAN ~10s/chunk

License

Model weights are derived from Resemble AI's Chatterbox . The export pipeline code is MIT licensed. Please refer to the original Chatterbox license for model weights usage terms.

Runs of acul3 chatterbox-executorch on huggingface.co

291
Total runs
4
24-hour runs
8
3-day runs
30
7-day runs
93
30-day runs

More Information About chatterbox-executorch huggingface.co Model

More chatterbox-executorch license Visit here:

https://choosealicense.com/licenses/apache-2.0

chatterbox-executorch huggingface.co

chatterbox-executorch huggingface.co is an AI model on huggingface.co that provides chatterbox-executorch's model effect (), which can be used instantly with this acul3 chatterbox-executorch model. huggingface.co supports a free trial of the chatterbox-executorch model, and also provides paid use of the chatterbox-executorch. Support call chatterbox-executorch model through api, including Node.js, Python, http.

chatterbox-executorch huggingface.co Url

https://huggingface.co/acul3/chatterbox-executorch

acul3 chatterbox-executorch online free

chatterbox-executorch huggingface.co is an online trial and call api platform, which integrates chatterbox-executorch's modeling effects, including api services, and provides a free online trial of chatterbox-executorch, you can try chatterbox-executorch online for free by clicking the link below.

acul3 chatterbox-executorch online free url in huggingface.co:

https://huggingface.co/acul3/chatterbox-executorch

chatterbox-executorch install

chatterbox-executorch is an open source model from GitHub that offers a free installation service, and any user can find chatterbox-executorch on GitHub to install. At the same time, huggingface.co provides the effect of chatterbox-executorch install, users can directly use chatterbox-executorch installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

chatterbox-executorch install url in huggingface.co:

https://huggingface.co/acul3/chatterbox-executorch

Url of chatterbox-executorch

chatterbox-executorch huggingface.co Url

Provider of chatterbox-executorch huggingface.co

acul3
ORGANIZATIONS

Other API from acul3

huggingface.co

Total runs: 13
Run Growth: 6
Growth Rate: 46.15%
Updated:May 27 2025
huggingface.co

Total runs: 12
Run Growth: 6
Growth Rate: 50.00%
Updated:May 27 2025
huggingface.co

Total runs: 7
Run Growth: -2
Growth Rate: -28.57%
Updated:September 02 2021
huggingface.co

Total runs: 5
Run Growth: -10
Growth Rate: -200.00%
Updated:December 05 2022
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:May 27 2025