Voice Cloning
MiniMax Voice Cloning accepts a reference audio file and a name, then returns a voice for compatible MiniMax speech generation. Accuracy can be adjusted, and source cleanup can optionally include noise reduction and volume normalization. The best evaluation comes from testing the returned voice on a short script that differs from the reference recording.
Properties
- Requires a name and an audio reference file.
- Provides adjustable accuracy from 0 to 1.
- Offers optional noise reduction and volume normalization for the source.
- Returns a voice ID for compatible MiniMax speech generation.
Best for
- Personal voiceovers
- Consistent character voices
Tips
- Use a clean recording with one speaker and minimal background noise.
How to use this model
- 1
Confirm authorization to use the speaker's recording and synthesized voice.
- 2
Upload a clean, single-speaker reference file and assign a recognizable name.
- 3
Set accuracy, then enable noise reduction or volume normalization only when the source requires cleanup.
- 4
Create the clone and note the returned voice ID.
- 5
Test the voice ID in MiniMax Speech with a short representative script before longer use.
Parameters
Voice Name
nameA private label for this voice in your library. It is not sent to MiniMax.
Voice Recording
voice_fileUpload audio lasting 10 seconds to 5 minutes and under 20 MB. Unsupported audio formats are converted to compatible WAV.
Recording Acceptance
accuracyHow strictly MiniMax validates the recording before cloning. Keep the recommended 0.7; lower it only if a clear recording is rejected.
Noise Reduction
need_noise_reductionTurn this on when the recording contains background noise. Leave it off for a clean recording to avoid altering the voice.
Volume Normalization
need_volume_normalizationTurn this on when the recording is too quiet or its loudness varies noticeably. Leave it off when the volume is already even.