← Back to all generators

google/gemini-3.1-flash-tts

Google's fast, expressive text-to-speech model with 30 voices and 70+ language support

Capabilities

No capability data available

Cost

Community model (estimated from hardware time)

Input Parameters

textrequiredstring

The text to convert to speech. Supports markup tags like " "[sigh], [laughing], [whispering], [shouting], [extremely fast] for " "expressive delivery. Maximum 4,000 bytes.

language_codestring

Language for the speech output

Default: "en-US"
af-ZAam-ETar-001ar-EGaz-AZbe-BYbg-BGbn-BDca-ESceb-PHcmn-CNcmn-twcs-CZda-DKde-DEel-GRen-AUen-GBen-INen-USes-419es-ESes-MXet-EEeu-ESfa-IRfi-FIfil-PHfr-CAfr-FRgl-ESgu-INhe-ILhi-INhr-HRht-HThu-HUhy-AMid-IDis-ISit-ITja-JPjv-JVka-GEkn-INko-KRkok-INla-VAlb-LUlo-LAlt-LTlv-LVmai-INmg-MGmk-MKml-INmn-MNmr-INms-MYmy-MMnb-NOne-NPnl-NLnn-NOor-INpa-INpl-PLps-AFpt-BRpt-PTro-ROru-RUsd-INsi-LKsk-SKsl-SIsq-ALsr-RSsv-SEsw-KEta-INte-INth-THtr-TRuk-UAur-PKvi-VN
promptstring

Style instructions to control how the text is spoken. " "Use natural language to describe the desired tone, pace, accent, " 'and emotion. For example: "Say this in a calm, professional tone" ' 'or "Speak with excitement and energy". Maximum 4,000 bytes.

Default: "Say the following."
voicestring

Voice to use for speech generation

Default: "Kore"
AchernarAchirdAlgenibAlgiebaAlnilamAoedeAutonoeCallirrhoeCharonDespinaEnceladusErinomeFenrirGacruxIapetusKoreLaomedeiaLedaOrusPulcherrimaPuckRasalgethiSadachbiaSadaltagerSchedarSulafatUmbrielVindemiatrixZephyrZubenelgenubi
Version: e165f56d71b9Updated: 7/25/2026205.2K runs