This is a lightweight, low dependency and fast text-to-speech (TTS) implementation in Python. The models are from ESPnet and exported to ONNX. Text is tokenized using ttstokenizer (https://github.com/neuml/ttstokenizer).
The goal is to use existing high quality TTS models without a heavy install footprint.