budecosystem / Tansen

huggingface.co
Total runs: 0
24-hour runs: 0
7-day runs: 0
30-day runs: 0
Model's Last Updated: September 07 2023

Introduction of Tansen

Model Details of Tansen

Tensen Logo


Democratizing access to LLMs, Multi-Modal Gen AI models for the open-source community.
Let's advance AI, together.


Tansen is a text-to-speech program built with the following priorities:

  1. Strong multi-voice capabilities.
  2. Highly realistic prosody and intonation.
  3. Speaking rate control

🎧 Demos
Demos

random_0_0.webm

random_0_1.webm

random_0_2.webm

💻 Getting Started on GitHub

Ready to dive in? Here's how you can get started with our repo on GitHub.

1️⃣ : Clone our GitHub repository

First things first, you'll need to clone our repository. Open up your terminal, navigate to the directory where you want the repository to be cloned, and run the following command:

conda create --name Tansen python=3.9 numba inflect
conda activate Tansen
conda install pytorch torchvision torchaudio pytorch-cuda=11.7 -c pytorch -c nvidia
conda install transformers=4.29.2
git clone https://github.com/BudEcosystem/Tansen.git
cd Tansen
2️⃣ : Install dependencies
python setup.py install
3️⃣ : Generate Audio
do_tts.py

This script allows you to speak a single phrase with one or more voices.

python do_tts.py --text "I'm going to speak this" --voice random --preset fast
read.py

This script provides tools for reading large amounts of text.

python Tansen/read.py --textfile <your text to be read> --voice random

This will break up the textfile into sentences, and then convert them to speech one at a time. It will output a series of spoken clips as they are generated. Once all the clips are generated, it will combine them into a single file and output that as well.

Sometimes Tansen screws up an output. You can re-generate any bad clips by re-running read.py with the --regenerate argument.

Intrested in running as as API ?

🐍 Usage in Python

Tansen can be used programmatically :

reference_clips = [utils.audio.load_audio(p, 22050) for p in clips_paths]
tts = api.TextToSpeech(use_deepspeed=True, kv_cache=True, half=True)
pcm_audio = tts.tts_with_preset("your text here", voice_samples=reference_clips, preset='fast')
Loss Curves

loss_mel_ce

loss_text_ce

Training Information

Device : A Single A100

Dataset : 876 hours

Runs of budecosystem Tansen on huggingface.co

0
Total runs
0
24-hour runs
0
3-day runs
0
7-day runs
0
30-day runs

More Information About Tansen huggingface.co Model

Tansen huggingface.co

Tansen huggingface.co is an AI model on huggingface.co that provides Tansen's model effect (), which can be used instantly with this budecosystem Tansen model. huggingface.co supports a free trial of the Tansen model, and also provides paid use of the Tansen. Support call Tansen model through api, including Node.js, Python, http.

budecosystem Tansen online free

Tansen huggingface.co is an online trial and call api platform, which integrates Tansen's modeling effects, including api services, and provides a free online trial of Tansen, you can try Tansen online for free by clicking the link below.

budecosystem Tansen online free url in huggingface.co:

https://huggingface.co/budecosystem/Tansen

Tansen install

Tansen is an open source model from GitHub that offers a free installation service, and any user can find Tansen on GitHub to install. At the same time, huggingface.co provides the effect of Tansen install, users can directly use Tansen installed effect in huggingface.co for debugging and trial. It also supports api for free installation.

Tansen install url in huggingface.co:

https://huggingface.co/budecosystem/Tansen

Url of Tansen

Provider of Tansen huggingface.co

budecosystem
ORGANIZATIONS

Other API from budecosystem

huggingface.co

Total runs: 127
Run Growth: 36
Growth Rate: 28.57%
Updated:September 02 2023
huggingface.co

Total runs: 31
Run Growth: -6
Growth Rate: -19.35%
Updated:May 07 2025
huggingface.co

Total runs: 7
Run Growth: -2
Growth Rate: -28.57%
Updated:September 05 2023