← Back to all generators

lucataco/csm-1b

CSM (Conversational Speech Model) is a speech generation model from Sesame that generates RVQ audio codes from text and audio inputs

Capabilities

No capability data available

Cost

Community model (estimated from hardware time)

Input Parameters

max_audio_length_msinteger

Maximum audio length in milliseconds

Default: 10000min: 1000, max: 30000
speakerinteger

Speaker ID (0 or 1)

Default: 0
01
textstring

Text to convert to speech

Default: "Hello from Sesame."
Version: 3e59b10a9894Updated: 7/25/20261.2K runs