← Back to all generators

stability-ai/stable-diffusion

A latent text-to-image diffusion model capable of generating photo-realistic images given any text input

Capabilities

1:14:33:416:99:162:33:2Negative PromptSeed

Cost

Community model (estimated from hardware time)

Input Parameters

guidance_scalenumber

Scale for classifier-free guidance

Default: 7.5min: 1, max: 20
heightinteger

Height of generated image in pixels. Needs to be a multiple of 64

Default: 768
641281922563203844485125766407047688328969601024
negative_promptstring

Specify things to not see in the output

num_inference_stepsinteger

Number of denoising steps

Default: 50min: 1, max: 500
num_outputsinteger

Number of images to generate.

Default: 1min: 1, max: 4
promptstring

Input prompt

Default: "a vision of paradise. unreal engine"
schedulerstring

Choose a scheduler.

Default: "DPMSolverMultistep"
DDIMK_EULERDPMSolverMultistepK_EULER_ANCESTRALPNDMKLMS
seedinteger

Random seed. Leave blank to randomize the seed

widthinteger

Width of generated image in pixels. Needs to be a multiple of 64

Default: 768
641281922563203844485125766407047688328969601024
Version: ac732df83ceaUpdated: 7/25/2026110.9M runs