Image Generation Optimization
How to get the best results from Stable Diffusion 3

How to get the best results from Stable Diffusion 3

6/18/2024

What this post added

This post details how to achieve optimal results with Stable Diffusion 3, focusing on prompt engineering, understanding different model versions (including text encoder configurations like fp8, fp16, and CLIP-only), and recommended settings for width, height, steps, and guidance scale. It highlights SD3's improved prompt adherence with longer prompts and advises against negative prompts, comparing its behavior to Midjourney v6 and DALL-E 3. Specific recommendations are provided for resolutions, steps (around 28), and CFG scale (3.5-4.5), along with observations on how steps affect image coherence and subject details. The post also touches on the potential for using different prompts for each of SD3's three text encoders.

Read the original post ↗