drbaph / FireRedTTS3-int8

huggingface.co
Total runs: 0
24-hour runs: 0
7-day runs: 0
30-day runs: 0
Model's Last Updated: August 15 2026
text-to-speech

Introduction of FireRedTTS3-int8

Model Details of FireRedTTS3-int8

FireRedTTS3-int8 (community INT8 ConvRot mirror)

INT8 ConvRot conversion of FireRedTeam/FireRedTTS3 for FireRedTTS3-ComfyUI , produced with the official comfy-kitchen quantizer ( TensorWiseINT8Layout.quantize , registry quantize_int8_convrot_weight ).

Format per quantized Linear (current ComfyUI representation):

  • weight torch.int8 , original [out, in] shape, contains the offline Hadamard-rotated weight ( W @ H^T per 256-column group)
  • weight_scale torch.float32 , [out, 1] per-output-row scale
  • bias — original float bias
  • comfy_quant — uint8 JSON: {"format": "int8_tensorwise", "convrot": true, "convrot_groupsize": 256}

At inference the companion custom node rotates activations online via comfy_kitchen.int8_linear(..., convrot=True, convrot_groupsize=256) — dynamic per-row INT8 activation quantization + INT8 GEMM, rescaled by scale_x * scale_w . No whole-weight dequantization on the hot path.

What is quantized (safe profile, group size 256)
Component Quantized Kept float
fireredtts3_base 321/332 Linears (1.73B params, 81.5% of core): all backbone_llm.layers.* , patch_encoder.blocks.* , dit.blocks.* embeddings, norms, spk_proj_* , patch_encoder.in_proj/out_proj , dit_head , dit.in_proj (1600 % 256 != 0), dit.t_embedder , dit.final_layer , stop_head , Conv1d
fireredtts3_instruct 321/331 Linears (1.73B params, 71.2% of core): same block families ( backbone_llm.model.layers.* ) same exclusions
redae nothing everything
campp nothing everything
Sizes
Core Official fp32 This repo
fireredtts3_base 8.48 GB 3.30 GB
fireredtts3_instruct 8.48 GB 3.30 GB
redae / campp / tokenizer copied through unchanged
Validation (base variant, full suite; instruct smoke-tested)
  • Per-layer weight roundtrip (official quantize -> official dequantize): worst rel-L2 0.00967 , worst cosine 0.999953 over 321 layers
  • Real-activation comparison vs fp32 through the same comfy_kitchen.int8_linear runtime: worst rel-L2 0.01162 , worst cosine 0.999932
  • On-disk structure: all 677 original keys preserved, scales fp32 [N,1] and positive, Conv1d/RedAE/CAM++ untouched
  • Runtime proof: 321 ConvRotInt8Linear modules, >42k counted INT8 ConvRot kernel calls during generation, weights stay int8 across unload/reload
  • BF16 vs INT8 generation (same seed/settings): identical patch counts (200/200), finite latents, EN/ZH ASR-verified, speaker-similarity parity (0.9007 vs 0.8991)
  • Peak VRAM 13.1 -> 8.3 GiB; generation ~1.3x slower (memory optimization, honestly reported)
Usage Disclaimer
  • The project incorporates zero-shot voice cloning functionality; Please note that this capability is intended solely for academic research purposes .
  • DO NOT use this model for ANY illegal activities ❗️❗️
  • The developers assume no liability for any misuse of this model.
  • If you identify any instances of abuse , misuse , or fraudulent activities related to this project, please report them to our team immediately.
Citation
@article{fireredtts3,
  title   = {FireRedTTS3: Unified Speech Generation and Editing with Semantically Enriched Speech Representations},
  author  = {FireRed Team},
  journal = {arXiv preprint},
  year    = {2026},
}
Acknowledgements
  • Qwen3 and Qwen2-Audio for the language model and audio understanding foundations
  • DiTAR for the patch-level diffusion autoregressive formulation
  • X-Codec for the discriminator design used in RedAE training
  • CAM++ for speaker embedding extraction
  • fastText for automatic language identification
  • WeTextProcessing (wetext) for the Chinese / English text normalization front-end

All credit to the FireRed Team — see the upstream repo and model card. Apache-2.0.

Runs of drbaph FireRedTTS3-int8 on huggingface.co

0
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
0
30-day runs

More Information About FireRedTTS3-int8 huggingface.co Model

More FireRedTTS3-int8 license Visit here:

https://choosealicense.com/licenses/apache-2.0

FireRedTTS3-int8 huggingface.co

FireRedTTS3-int8 huggingface.co is an AI model on huggingface.co that provides FireRedTTS3-int8's model effect (), which can be used instantly with this drbaph FireRedTTS3-int8 model. huggingface.co supports a free trial of the FireRedTTS3-int8 model, and also provides paid use of the FireRedTTS3-int8. Support call FireRedTTS3-int8 model through api, including Node.js, Python, http.

FireRedTTS3-int8 huggingface.co Url

https://huggingface.co/drbaph/FireRedTTS3-int8

drbaph FireRedTTS3-int8 online free

FireRedTTS3-int8 huggingface.co is an online trial and call api platform, which integrates FireRedTTS3-int8's modeling effects, including api services, and provides a free online trial of FireRedTTS3-int8, you can try FireRedTTS3-int8 online for free by clicking the link below.

drbaph FireRedTTS3-int8 online free url in huggingface.co:

https://huggingface.co/drbaph/FireRedTTS3-int8

FireRedTTS3-int8 install

FireRedTTS3-int8 is an open source model from GitHub that offers a free installation service, and any user can find FireRedTTS3-int8 on GitHub to install. At the same time, huggingface.co provides the effect of FireRedTTS3-int8 install, users can directly use FireRedTTS3-int8 installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

FireRedTTS3-int8 install url in huggingface.co:

https://huggingface.co/drbaph/FireRedTTS3-int8

Url of FireRedTTS3-int8

FireRedTTS3-int8 huggingface.co Url

Provider of FireRedTTS3-int8 huggingface.co

drbaph
ORGANIZATIONS

Other API from drbaph

huggingface.co

Total runs: 7.0K
Run Growth: 786
Growth Rate: 11.30%
Updated:January 28 2026
huggingface.co

Total runs: 1.1K
Run Growth: -146
Growth Rate: -13.46%
Updated:March 07 2026
huggingface.co

Total runs: 813
Run Growth: -493
Growth Rate: -60.64%
Updated:March 13 2026
huggingface.co

Total runs: 128
Run Growth: 29
Growth Rate: 22.66%
Updated:February 12 2026
huggingface.co

Total runs: 92
Run Growth: 47
Growth Rate: 51.09%
Updated:March 17 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:March 17 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:January 28 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:June 13 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:January 28 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:April 13 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:June 04 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:March 17 2025
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:June 14 2026
huggingface.co

Total runs: 0
Run Growth: 0
Growth Rate: 0.00%
Updated:September 10 2026