drbaph / WavTTS

huggingface.co
Total runs: 0
24-hour runs: 0
7-day runs: 0
30-day runs: 0
Model's Last Updated: June 04 2026
text-to-speech

Introduction of WavTTS

Model Details of WavTTS

WavTTS 16k Safetensors for ComfyUI

Safetensors conversion of the official WavTTS model for use with WavTTS-ComfyUI .

Screenshot 2026-06-04 015737

This release converts the original PyTorch checkpoint into the Safetensors format for safer loading, improved compatibility with ComfyUI workflows, and reduced storage overhead.

Model Introduction

WavTTS is a zero-shot text-to-speech model that directly generates raw audio waveforms from text using reference-audio prompting.

Unlike token-based TTS systems that generate intermediate acoustic representations, WavTTS models speech directly in the waveform domain, enabling highly natural speech synthesis while preserving speaker characteristics from short reference samples.

Links
Usage

This checkpoint is intended for use with:

https://github.com/Saganaki22/WavTTS-ComfyUI

Place the model inside:

ComfyUI/models/wavtts/

Example:

ComfyUI/models/wavtts/
├── wavtts-fp32.safetensors
└── wavtts-mixed-bf16.safetensors

Restart ComfyUI and load the checkpoint using the WavTTS Load Model node.

Model Summary
Item Value
Model WavTTS
Format Safetensors
Task Zero-Shot Text-to-Speech
Sample Rate 16 kHz
Architecture Direct Waveform Generation
Conditioning Reference Audio + Reference Transcript
Intended Platform ComfyUI
Languages English, Chinese
Training Dataset Emilia Dataset
License CC BY-NC 4.0
Included Variants
File Description
wavtts-fp32.safetensors Clean FP32 inference checkpoint
wavtts-mixed-bf16.safetensors Mixed BF16 checkpoint optimized for lower VRAM usage

The original .pt checkpoint contains training-related state that is unnecessary for inference.

Safetensors releases store inference weights only, resulting in smaller file sizes and safer loading behavior.

Intended Use

This model is intended for:

  • Zero-shot text-to-speech
  • Voice continuation
  • Reference-audio conditioned generation
  • Voice adaptation from short prompts
  • ComfyUI speech workflows
  • Local offline TTS generation

A transcript of the reference audio is required by WavTTS.

Precision Notes
FP32

Recommended for maximum stability.

Mixed BF16

Recommended for reduced VRAM usage while preserving numerically sensitive tensors in FP32.

No model weights have been retrained or modified beyond precision conversion.

Architecture

WavTTS performs direct waveform modeling rather than generating intermediate acoustic tokens.

Inputs
  • Reference audio
  • Reference transcript
  • Target text
Output
  • 16 kHz synthesized waveform

The model learns speaker characteristics directly from reference audio and generates speech matching the target text in the reference voice style.

Limitations
  • Requires a transcript for the reference audio.
  • Voice similarity is not guaranteed.
  • Long generations may require chunking.
  • Audio quality depends heavily on reference quality.
  • Commercial usage may be restricted by the upstream license.
  • Numerical differences may occur between original and converted checkpoints.
Attribution
Original Authors

WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling

Official repository:

https://github.com/cwx-worst-one/WavTTS

ComfyUI Integration

https://github.com/Saganaki22/WavTTS-ComfyUI

Dataset

https://huggingface.co/datasets/amphion/Emilia-Dataset

License

This Safetensors conversion inherits the licensing terms of the original WavTTS release.

The original model weights are licensed under CC BY-NC 4.0 .

Please review the upstream model card and license terms before redistribution or commercial use.

Citation
@article{chen2026wavtts,
  title={WavTTS: Towards High-Quality Zero-Shot TTS via Direct Raw Waveform Modeling},
  author={TODO},
  journal={TODO},
  year={2026}
}

Runs of drbaph WavTTS on huggingface.co

0
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
0
30-day runs

More Information About WavTTS huggingface.co Model

WavTTS huggingface.co

WavTTS huggingface.co is an AI model on huggingface.co that provides WavTTS's model effect (), which can be used instantly with this drbaph WavTTS model. huggingface.co supports a free trial of the WavTTS model, and also provides paid use of the WavTTS. Support call WavTTS model through api, including Node.js, Python, http.

drbaph WavTTS online free

WavTTS huggingface.co is an online trial and call api platform, which integrates WavTTS's modeling effects, including api services, and provides a free online trial of WavTTS, you can try WavTTS online for free by clicking the link below.

drbaph WavTTS online free url in huggingface.co:

https://huggingface.co/drbaph/WavTTS

WavTTS install

WavTTS is an open source model from GitHub that offers a free installation service, and any user can find WavTTS on GitHub to install. At the same time, huggingface.co provides the effect of WavTTS install, users can directly use WavTTS installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

WavTTS install url in huggingface.co:

https://huggingface.co/drbaph/WavTTS

Url of WavTTS

WavTTS huggingface.co Url

Provider of WavTTS huggingface.co

drbaph
ORGANIZATIONS

Other API from drbaph

huggingface.co

Total runs: 7.3K
Run Growth: 1.8K
Growth Rate: 24.88%
Updated:January 28 2026
huggingface.co

Total runs: 832
Run Growth: -460
Growth Rate: -55.29%
Updated:March 13 2026
huggingface.co

Total runs: 509
Run Growth: -765
Growth Rate: -150.29%
Updated:March 07 2026
huggingface.co

Total runs: 133
Run Growth: 44
Growth Rate: 33.08%
Updated:February 12 2026
huggingface.co

Total runs: 89
Run Growth: 55
Growth Rate: 61.80%
Updated:March 17 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:March 17 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:June 13 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:January 28 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:April 13 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:January 28 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:March 17 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:September 10 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:June 14 2026