Fish Audio TTS
Natural speech synthesis with Fish Audio S2.1-Pro
Turn scripts into spoken audio using a Fish Audio voice model ID from Voice Design or voice cloning.
Model Overview
Fish Audio S2.1-Pro
Fish Audio TTS synthesizes speech with the production S2.1-Pro model. Pass a voice model ID from Fish Audio Voice Design or voice cloning for consistent timbre.
S2.1-Pro production TTS
Voice Design / clone model ID support
Prosody speed control
83-language coverage
Why use it on Moosky AI?
Model-specific controls with no subscription, clean outputs, and a workflow built for fast creative iteration.
No subscription required
Buy credits when you need them and pay only for what you generate.
Natural speech
Generate clear narration, dialogue, and voiceover audio from text.
Voice control
Pick built-in voices or reuse a custom voice design ID.
Commercial workflows
Use results in professional projects subject to platform and provider terms.
Example outputs
Featured public generations from the Moosky community using this composer.
How it works
Three focused steps from setup to finished output.
Provide a voice ID
Paste a Fish Audio voice model ID from Voice Design or cloning.
Enter your script
Type the narration or dialogue to synthesize.
Download audio
Use the MP3 in video, prototypes, or voiceover beds.
Pricing
Credit-based generation
Uses Moosky credits with cost shown before you generate. No subscription required.
Generate speechFrequently Asked Questions
How do I use a custom voice?
Create a voice with Fish Audio Voice Design or clone from a sample, then pass the returned model ID into Fish Audio TTS.
What languages are supported?
S2.1-Pro supports 83 languages with automatic language detection. Match your script language for best results.
How is Fish Audio TTS billed?
TTS is billed per UTF-8 bytes of input text for the S2.1-Pro model.