A developer trained a tiny 20M parameter open-source text-to-speech model from scratch overnight using a single RTX 3090 GPU. The proof-of-concept model, while still exhibiting glitchy and robotic output, demonstrates rapid training feasibility for lightweight TTS systems. Training code is available on GitHub and model weights on Hugging Face.

Read original