Skip to main content
📖 The AI Tool Bible

Stable Diffusion Web UI (AUTOMATIC1111)

The de facto local Stable Diffusion power-user UI.

Free· Free / open-source (AGPL-3.0). You provide your own compute (local GPU, rented cloud GPU, or a Colab instance).Image GenerationStable Diffusion 1.x / 2.x / SDXL and community fine-tunes (Safetensors checkpoints)
Visit website →
Best for

Hobbyists, illustrators, and researchers who want a free, local, fully controllable Stable Diffusion setup with access to every community model, LoRA, and extension.

Skip if

Non-technical users who want a one-click hosted product, teams needing SLAs and support, or anyone chasing the newest diffusion architectures on release day.

AUTOMATIC1111's Stable Diffusion Web UI is the most widely used local front-end for running Stable Diffusion image models on your own hardware. It wraps the diffusion pipeline in a Gradio browser UI that exposes essentially every knob the community has invented: txt2img, img2img, inpainting, outpainting, batch generation, X/Y/Z parameter grids, prompt attention weighting and prompt editing mid-generation, negative prompts, seed variation and subseed strength, face restoration (GFPGAN, CodeFormer), a stack of upscalers (RealESRGAN, ESRGAN, SwinIR, Latent), checkpoint merging, CLIP interrogation, and training workflows for textual inversion, hypernetworks, and LoRA. It loads SD 1.x, SD 2.x, and community-tuned checkpoints in .safetensors, and a huge third-party extension ecosystem adds ControlNet, regional prompting, animation, video, upscaling pipelines, and integrations. Typical workflows are experimenting with prompts and samplers at low step counts, then producing final images with an upscaler pass and optional inpainting cleanup; power users chain it into ControlNet for pose or depth control, or use LoRAs to steer style and characters. It is the reference tool for anyone learning how diffusion parameters actually behave, and remains a common daily driver for illustrators, concept artists, indie game devs, and researchers who want reproducible local generation without a subscription.

Editor's take

Still the reference implementation for learning how Stable Diffusion actually works — every parameter you read about in a paper or Reddit thread is a slider here. In 2026 I lean toward ComfyUI for cutting-edge model support and reproducible pipelines, but AUTOMATIC1111 remains the friendliest way to sit down and just make images with a local checkpoint.

— The AI Tool Bible editorial team

Pros

  • Runs entirely locally, so there are no per-image fees or content-policy filters imposed by a hosted API.
  • Supports nearly every Stable Diffusion checkpoint, LoRA, VAE, and embedding format the community produces.
  • Massive extension ecosystem (ControlNet, Regional Prompter, Dynamic Prompts, ADetailer, etc.) covers advanced workflows.
  • Fine-grained control over samplers, CFG, seeds, prompt weighting, and prompt scheduling that hosted tools rarely expose.
  • Built-in training paths for textual inversion, hypernetworks, and LoRA on your own datasets.
  • Works on modest hardware (reports of usable output at 4GB VRAM with low-precision modes) and supports Apple Silicon.

Cons

  • ⚠️ Setup requires Python 3.10, Git, and matching GPU drivers, which is a real barrier for non-technical users.
  • ⚠️ The Gradio UI is dense and inconsistent; discoverability of features is poor compared to newer node-based tools like ComfyUI.
  • ⚠️ Development cadence has slowed and it lags behind ComfyUI on newer model architectures (SDXL refinements, SD3, Flux support arrives late or via extensions).
  • ⚠️ No first-party hosted version — you are responsible for GPU cost, updates, and extension conflicts.
  • ⚠️ Extension quality varies wildly and a bad extension can break the whole install until you disable it.

Use cases

Local text-to-image generationInpainting and outpainting existing imagesLoRA and textual inversion trainingControlNet-guided composition (pose, depth, edges)Batch prompt exploration with X/Y/Z gridsUpscaling and face restoration passesConcept art and illustration referenceCheckpoint merging and model experimentation

Explore related

Compare with similar tools

All in Image Generation

Midjourney

Featured
Image Generation · Midjourney v7
9.4

The gold standard for aesthetic AI image generation.

Paid· Basic: $10 · Pro: $20 · Enterprise: Contact salesillustrationconcept art

Flux

Featured
Image Generation · Flux.1 [schnell / dev / pro]
9.0

Black Forest Labs' open-weights image model — rivals Midjourney quality.

Freemium· FLUX.2 [max]: $0.07 · FLUX.2 [pro]: $0.03 · FLUX.2 [klein] 9B: $0.015 · FLUX.2 [klein] 4B: $0.014 · FLUX.2 [flex]: $0.05open sourceself-hosted

Stable Diffusion

Image Generation · SD 3.5 / SDXL
8.8

Open-source image generation — run anywhere, fine-tune anything.

Free· Free open weights; optional Stability APIlocalfine-tuning

Nano Banana (Gemini Image)

Image Generation · Gemini 3 Pro Image (Nano Banana Pro), Gemini 3.1 Flash Image (Nano Banana 2), Gemini 3.1 Flash-Lite Image (Nano Banana 2 Lite)
8.7

Google DeepMind's Gemini-powered image generation and conversational editing model family

Paid· Consumer access via Gemini app (free tier + Google AI Pro/Ultra subscriptions). API usage-based: Nano Banana Pro (Gemini 3 Pro Image) ~$0.134/image at 1K-2K, ~$0.24/image at 4K; Nano Banana 2 (Gemini 3.1 Flash Image) ~$0.067/image at 1K, up to ~$0.151 at higher resolutions; Nano Banana 2 Lite priced lower for high-throughput use. Batch API roughly 50% off. Enterprise pricing via Gemini Enterprise Agent Platform and Vertex AI.Marketing hero imagesProduct mockups and packaging visualisations

DALL·E 3

Image Generation · DALL·E 3
8.6

OpenAI's image model — strong on prompt adherence and text-in-image.

Freemium· Basic: $10 · Pro: $20 · Enterprise: Contact salespostersinfographics

Canva Magic Studio

Image Generation · Multi-model: partners including OpenAI (Magic Write historically on GPT models), Google Imagen and Runway for image/video, plus Canva's in-house design and layout models
8.5

Canva's all-in-one AI creative suite for design, image, video, copy, and presentations

Freemium· Free: Free · Pro: €11.67 · Business: €14.17 · Enterprise: Contact salesSocial media post generationShort-form video ads