Your voice, cloned and reusable.
MiniMax Voice Clone builds a reusable copy of a voice from a 10-second-plus MP3, WAV or M4A sample, and generates a preview line so you can hear it immediately. Cloned voices are saved to your Zorq account and work in text-to-speech, so your whole voiceover pipeline lives in one subscription.
Sign in with Google or Apple · No credit card to start · 25+ AI models
Why creators pick MiniMax Voice Clone
A 10-second sample is enough
Upload at least 10 seconds of clear speech as MP3, WAV or M4A and the model builds the voice.
Saved to your account
Cloned voices are stored and selectable in Zorq’s text-to-speech, so any script can be read in your voice.
Pick the clone engine
Choose between 2.6 HD (best quality), 2.6 Turbo, 02 HD and 02 Turbo depending on your speed and quality needs.
Hear it immediately
Every clone generates a preview line so you can judge the likeness before using it in a project.
How to use MiniMax Voice Clone on Zorq
No install, no waitlist — it runs in the browser.
Sign in free
Sign in with Google or Apple at zorqai.com — free starter credits, no credit card needed to start.
Open voice cloning
Go to /audio and open the voice cloning tab, then pick an engine — 2.6 HD is the highest quality.
Upload a sample
Upload at least 10 seconds of clear speech in MP3, WAV or M4A format.
Clone and reuse
Run the clone, listen to the preview line, then select your saved voice in text-to-speech to read any script.
Simple pricing that covers it all
Credit-based plans for image, video, voice and more — not just one media type.
Prices shown billed annually · save up to 67% vs monthly · compare all plans
Frequently asked questions
How do I clone a voice online with Zorq?
Sign in to Zorq AI with Google or Apple, go to /audio, open the voice cloning tab and upload a 10-second-plus MP3, WAV or M4A sample of the voice. MiniMax Voice Clone builds the voice, plays a preview line, and saves it to your account for use in text-to-speech.
How many credits does voice cloning cost?
Cloning a voice costs a flat credit fee. Speaking with the cloned voice afterwards is billed at normal text-to-speech rates — MiniMax Speech 2.6 HD, which supports cloned voices, costs 2 credits per generation.
Voice Clone or Qwen3 Voice Design — which do I need?
MiniMax Voice Clone reproduces a real voice from an audio sample, which is what you want for your own voice or a brand voice you have rights to. Qwen3 Voice Design invents a new voice from a plain-language description with no sample needed — better for fictional characters and narrators.
Whose voice am I allowed to clone?
Clone your own voice, or a voice you have clear permission to use. Do not clone someone’s voice without their consent.
What sample gives the best clone?
At least 10 seconds of clean, clear speech with minimal background noise, in MP3, WAV or M4A. Natural talking works better than shouting or whispering, and the 2.6 HD engine gives the best likeness.
Try MiniMax Voice Clone now
One subscription covers this model plus 25+ more for image, video, voice and lipsync.