Youtube Autonomous Audio Narration Coqui Voice Module
The Audio narration Coqui voice module.
Please, check the 'pyproject.toml' file to see the dependencies.
General information
This project is based on coqui-tts (needs a license for commercial use), which is working with different models that are handled in a different way. Some models are older and others are more recent. There is, apparently, a sweet spot of dependencies versions to make all of them work together. I leave it here, but maybe we should prioritize the most modern models.
This project is, by now, using the xtts_v2 model only.
Stack estable universal
python==3.11.0.final.0torch==2.5.1torchaudio==2.5.1TTS==0.25.xtransformers==4.38.2
Which is allowing VITS, XTTS v2, CSS10, Tacotron2 and other multilingual models.
About the models
The models are downloaded into the cache if needed. In windows, this means here (C:/Users/dania/AppData/Local/tts).
You must set the TTS_HOME environment variable to choose where your models are (mine is D:/coqui-tts). If you don't set it, the fallback will use the local data folder, but I made it to raise an Exception if this happens to be able to control it.
Downloading the model needs to day yes to an input to accept the non-commercial agreement.
Instructions
To make this coqui-tts work we will need the extra = ['languages'] and the espeak-ng installed to be able to narrate.
Espeak-NG
- Go to the releases page and download the latest version (or one that fits your needs), the
.msiif Windows:
- Register the folder in which it has been installed in your system
PATH.
You will probably need this also:
Metadata
Release files for yta-audio-narration-coqui 0.2.4
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| yta_audio_narration_coqui-0.2.4.tar.gz | 9.4 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| yta_audio_narration_coqui-0.2.4-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 20.1 kB
Release files / yta_audio_narration_coqui-0.2.4.tar.gz
| Download URL | yta_audio_narration_coqui-0.2.4.tar.gz |
|---|---|
| Size | 9.4 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
ee59bc2999ba3822dd7c34cef18f716390ffa1a643a438e8506714118b816c2f
|
|
BLAKE2b-256 checksum How to use checksums |
3aba74a1707afb4551e42c2200912dd17fe426a68b557249467d4b5b8b58caec
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
poetry/2.2.0 CPython/3.9.0 Windows/10
|
Release files / yta_audio_narration_coqui-0.2.4-py3-none-any.whl
| Download URL | yta_audio_narration_coqui-0.2.4-py3-none-any.whl |
|---|---|
| Size | 10.7 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
4c6d5627315dda558bdbaaf84470d523f9501b215d31d83d3bfa1d3238962443
|
|
BLAKE2b-256 checksum How to use checksums |
0c3d11a905b2f5c9991e0d307330acf0703fb5d95161838c48ab2be9423ef1c0
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
poetry/2.2.0 CPython/3.9.0 Windows/10
|