← Back to all generators

minimax/voice-cloning

Clone voices to use with Minimax's speech-02-hd and speech-02-turbo

Capabilities

No capability data available

Cost

Community model (estimated from hardware time)

Input Parameters

voice_filerequiredstring

Voice file to clone. Must be MP3, M4A, or WAV format, 10s to 5min duration, and less than 20MB.

accuracynumber

Text validation accuracy threshold (0-1)

Default: 0.7min: 0, max: 1
modelstring

The text-to-speech model to train

Default: "speech-02-turbo"
speech-2.6-turbospeech-2.6-hdspeech-02-turbospeech-02-hd
need_noise_reductionboolean

Enable noise reduction. Use this if the voice file has background noise.

Default: false
need_volume_normalizationboolean

Enable volume normalization

Default: false
Version: fff8a670880fUpdated: 7/25/202644.7K runs