infery
← All models

Stable Diffusion V3

stable-diffusion-v3-medium

Image generationby Stability AI

Stable Diffusion 3 Medium (Text to Image) is a Multimodal Diffusion Transformer (MMDiT) model that improves image quality, typography, prompt understanding, and efficiency.

images

Details

Accepts
text

Pricing

Price
4.38 cr / image

Prices in credits (1 credit = $0.01).

Data schema

Input

FieldTypeDescription
seed The same seed and the same prompt given to the same version of Stable Diffusion will output the same image every time.
promptrequiredstringThe prompt to generate an image from.
sync_modebooleanIf `True`, the media will be returned as a data URI and the output data won't be available in the request history.
image_sizeThe size of the generated image.
num_imagesintegerThe number of images to generate.
guidance_scalenumber The CFG (Classifier Free Guidance) scale is a measure of how close you want the model to stick to your prompt when looking for a related image to show you.
negative_promptstringThe negative prompt to generate an image from.
prompt_expansionbooleanIf set to true, prompt will be upsampled with more details.
num_inference_stepsintegerThe number of inference steps to perform.
enable_safety_checkerbooleanIf set to true, the safety checker will be enabled. Disabling it requires account authorization; unauthorized requests are always checked, and images flagged as unsafe are returned as black images.

Output

FieldTypeDescription
seedinteger Seed of the generated Image. It will be the same value of the one passed in the input or the randomly generated that was used in case none was passed.
imagesarrayThe generated image files info.
promptstringThe prompt used for generating the image.
timingsobject
num_imagesintegerThe number of images generated.
has_nsfw_conceptsarrayWhether the generated images contain NSFW concepts.