Fish Audio Voice Clone
High-fidelity voice cloning with Fish Audio S2.1-Pro
Upload a short reference sample and synthesize any script with Fish Audio voice models for natural, production-quality results.
Model Overview
Fish Audio Voice Clone
Fish Audio clones timbre from a short audio sample into a reusable voice model, then synthesizes your script with S2.1-Pro. Voice enrollment is free; synthesis is billed per UTF-8 bytes.
Fast voice model enrollment from samples
Fish Audio S2.1-Pro synthesis
Prosody speed control
Stateful profile reuse for narration
API-only, no GPU queue
Why use it on Moosky AI?
Model-specific controls with no subscription, clean outputs, and a workflow built for fast creative iteration.
No subscription required
Buy credits when you need them and pay only for what you generate.
Natural speech
Generate clear narration, dialogue, and voiceover audio from text.
Voice control
Pick built-in voices or reuse a custom voice design ID.
Commercial workflows
Use results in professional projects subject to platform and provider terms.
Example outputs
Featured public generations from the Moosky community using this composer.
How it works
Three focused steps from setup to finished output.
Upload a voice sample
Provide a clear reference recording (audio or video).
Enter your script
Type the text you want the cloned voice to speak.
Download your audio
Receive synthesized speech in the cloned timbre.
Pricing
Credit-based generation
Uses Moosky credits with cost shown before you generate. No subscription required.
Clone a voiceFrequently Asked Questions
What audio sample do I need?
A clip with at least a few seconds of clear speech. Longer uploads are trimmed automatically before enrollment.
Is voice enrollment billed?
Creating the cloned voice model is free. You are billed for speech synthesis based on UTF-8 bytes of your script.
When should I use Fish Audio voice cloning?
Use Fish Audio when you want production S2.1-Pro quality with a reusable voice model for narration and character dialogue.