All tools
aiquick/text-to-speech
text-to-audioSpeech synthesis on the audio inference worker. Pick one of the built-in speakers or let the model choose, and get back a WAV clip you can play or download. Synthesis is slow — expect around a minute or two for a short passage — because each request generates the waveform from scratch.
Inference/audio/tts
Input
Result
IdleFill in the input and run the tool to see results here.