A text-to-speech inference engine with support for voice cloning.
Speech & Audio
Transcription, speech synthesis, voice interfaces, music tools, and interface audio.
16 saved linksSearch in this topic →
Saved Links in Speech & Audio
A deep learning toolkit for text-to-speech research and applications.
An inference and training library for text-to-speech models.
An implementation of Kyutai's Moshi speech dialogue model for the MLX framework.