Qwen TTS
Natural speech synthesis with Qwen 3 TTS
Turn scripts into spoken audio with built-in system voices or a custom voice ID from a prior Qwen Voice Design run.
Model Overview
Qwen TTS
Qwen TTS synthesizes speech from text using Alibaba Qwen 3 TTS Flash. Pick a system voice, optionally pass delivery instructions, and reuse custom voice IDs created with Qwen Voice Design.
Why use it on Moosky AI?
Model-specific controls with no subscription, clean outputs, and a workflow built for fast creative iteration.
No subscription required
Buy credits when you need them and pay only for what you generate.
Natural speech
Generate clear narration, dialogue, and voiceover audio from text.
Voice control
Pick built-in voices or reuse a custom voice design ID.
Commercial workflows
Use results in professional projects subject to platform and provider terms.
Example outputs
Featured public generations from the Moosky community using this composer.
How it works
Three focused steps from setup to finished output.
Pick a voice
Choose a built-in voice or paste a Voice Design ID.
Enter your script
Type the narration or dialogue to synthesize.
Download audio
Use the WAV in video, prototypes, or voiceover beds.
Pricing
Credit-based generation
Uses Moosky credits with cost shown before you generate. No subscription required.
Generate speechFrequently Asked Questions
How do I use a custom voice?
Create a voice with Qwen Voice Design first, then pass the returned voice ID into Qwen TTS for later scripts.
What languages are supported?
Qwen TTS supports multiple languages through the model configuration. Pick the language that matches your script before generating.
How is Qwen TTS different from Moosky Voice Clone?
Moosky Voice Clone clones timbre from an audio sample. Qwen TTS uses Alibaba system voices or designed voice profiles instead of direct sample cloning.