Skip to main content
📖 The AI Tool Bible
Stable Diffusion Web UI (AUTOMATIC1111) preview image
Stable Diffusion Web UI (AUTOMATIC1111) logo

Stable Diffusion Web UI (AUTOMATIC1111)

The de facto local Stable Diffusion power-user UI.

Free· Free / open-source (AGPL-3.0). You provide your own compute (local GPU, rented cloud GPU, or a Colab instance).Image GenerationStable Diffusion 1.x / 2.x / SDXL and community fine-tunes (Safetensors checkpoints)
Visit website →

In short

This tool provides a comprehensive local front-end for Stable Diffusion, exposing detailed parameters for text-to-image, inpainting, and model training. It is best for hobbyists and researchers seeking full control over local generation without subscription fees.

Best for

Hobbyists, illustrators, and researchers who want a free, local, fully controllable Stable Diffusion setup with access to every community model, LoRA, and extension.

Skip if

Non-technical users who want a one-click hosted product, teams needing SLAs and support, or anyone chasing the newest diffusion architectures on release day.

AUTOMATIC1111's Stable Diffusion Web UI is the most widely used local front-end for running Stable Diffusion image models on your own hardware. It wraps the diffusion pipeline in a Gradio browser UI that exposes essentially every knob the community has invented: txt2img, img2img, inpainting, outpainting, batch generation, X/Y/Z parameter grids, prompt attention weighting and prompt editing mid-generation, negative prompts, seed variation and subseed strength, face restoration (GFPGAN, CodeFormer), a stack of upscalers (RealESRGAN, ESRGAN, SwinIR, Latent), checkpoint merging, CLIP interrogation, and training workflows for textual inversion, hypernetworks, and LoRA. It loads SD 1.x, SD 2.x, and community-tuned checkpoints in .safetensors, and a huge third-party extension ecosystem adds ControlNet, regional prompting, animation, video, upscaling pipelines, and integrations. Typical workflows are experimenting with prompts and samplers at low step counts, then producing final images with an upscaler pass and optional inpainting cleanup; power users chain it into ControlNet for pose or depth control, or use LoRAs to steer style and characters. It is the reference tool for anyone learning how diffusion parameters actually behave, and remains a common daily driver for illustrators, concept artists, indie game devs, and researchers who want reproducible local generation without a subscription.

Editor's take

Still the reference implementation for learning how Stable Diffusion actually works — every parameter you read about in a paper or Reddit thread is a slider here. In 2026 I lean toward ComfyUI for cutting-edge model support and reproducible pipelines, but AUTOMATIC1111 remains the friendliest way to sit down and just make images with a local checkpoint.

— The AI Tool Bible editorial team

Pros

  • Runs entirely locally, so there are no per-image fees or content-policy filters imposed by a hosted API.
  • Supports nearly every Stable Diffusion checkpoint, LoRA, VAE, and embedding format the community produces.
  • Massive extension ecosystem (ControlNet, Regional Prompter, Dynamic Prompts, ADetailer, etc.) covers advanced workflows.
  • Fine-grained control over samplers, CFG, seeds, prompt weighting, and prompt scheduling that hosted tools rarely expose.
  • Built-in training paths for textual inversion, hypernetworks, and LoRA on your own datasets.
  • Works on modest hardware (reports of usable output at 4GB VRAM with low-precision modes) and supports Apple Silicon.

Cons

  • ⚠️ Setup requires Python 3.10, Git, and matching GPU drivers, which is a real barrier for non-technical users.
  • ⚠️ The Gradio UI is dense and inconsistent; discoverability of features is poor compared to newer node-based tools like ComfyUI.
  • ⚠️ Development cadence has slowed and it lags behind ComfyUI on newer model architectures (SDXL refinements, SD3, Flux support arrives late or via extensions).
  • ⚠️ No first-party hosted version — you are responsible for GPU cost, updates, and extension conflicts.
  • ⚠️ Extension quality varies wildly and a bad extension can break the whole install until you disable it.

Use cases

Local text-to-image generationInpainting and outpainting existing imagesLoRA and textual inversion trainingControlNet-guided composition (pose, depth, edges)Batch prompt exploration with X/Y/Z gridsUpscaling and face restoration passesConcept art and illustration referenceCheckpoint merging and model experimentation

Frequently asked

Is Stable Diffusion Web UI free to use?
Yes, it is free and open-source under the AGPL-3.0 license. Users must provide their own compute resources, such as a local GPU or rented cloud instance.
What types of models does it support?
It supports Stable Diffusion 1.x, 2.x, and SDXL models, as well as community fine-tunes in .safetensors format. It also handles LoRAs, VAEs, and embeddings.
Can I train custom models using this interface?
Yes, the tool includes built-in workflows for training textual inversion, hypernetworks, and LoRAs on your own datasets.
What are the system requirements for setup?
Setup requires Python 3.10, Git, and matching GPU drivers. It can run on modest hardware with as little as 4GB VRAM using low-precision modes and supports Apple Silicon.
Who is this tool not recommended for?
It is not ideal for non-technical users seeking one-click hosted products, teams requiring SLAs, or those needing immediate support for the newest diffusion architectures.

Explore related

Compare with similar tools

All in Image Generation
MI

Midjourney

Featured
Image Generation · Midjourney v7
9.4

The gold standard for aesthetic AI image generation.

Paid· Basic: $20 · Pro: $50 · Enterprise: Contact salesillustrationconcept art
FL

Flux

Featured
Image Generation · Flux.1 [schnell / dev / pro]
9.0

Black Forest Labs' open-weights image model — rivals Midjourney quality.

Freemium· FLUX.2 [max]: $0.07 · FLUX.2 [pro]: $0.03 · FLUX.2 [klein] 9B: $0.015 · FLUX.2 [klein] 4B: $0.014 · FLUX.2 [flex]: $0.05open sourceself-hosted
SD

Stable Diffusion

Image Generation · SD 3.5 / SDXL
8.8

Open-source image generation — run anywhere, fine-tune anything.

Free· Free open weights; optional Stability APIlocalfine-tuning
NB

Nano Banana (Gemini Image)

Image Generation · Gemini 3 Pro Image (Nano Banana Pro), Gemini 3.1 Flash Image (Nano Banana 2), Gemini 3.1 Flash-Lite Image (Nano Banana 2 Lite)
8.7

Google DeepMind's Gemini-powered image generation and conversational editing model family

Paid· Consumer access via Gemini app (free tier + Google AI Pro/Ultra subscriptions). API usage-based: Nano Banana Pro (Gemini 3 Pro Image) ~$0.134/image at 1K-2K, ~$0.24/image at 4K; Nano Banana 2 (Gemini 3.1 Flash Image) ~$0.067/image at 1K, up to ~$0.151 at higher resolutions; Nano Banana 2 Lite priced lower for high-throughput use. Batch API roughly 50% off. Enterprise pricing via Gemini Enterprise Agent Platform and Vertex AI.Marketing hero imagesProduct mockups and packaging visualisations
DE

DALL·E 3

Image Generation · DALL·E 3
8.6

OpenAI's image model — strong on prompt adherence and text-in-image.

Freemium· Basic: $10 · Pro: $20postersinfographics
CM

Canva Magic Studio

Image Generation · Multi-model: partners including OpenAI (Magic Write historically on GPT models), Google Imagen and Runway for image/video, plus Canva's in-house design and layout models
8.5

Canva's all-in-one AI creative suite for design, image, video, copy, and presentations

Freemium· Free: Free · Pro: €11.67 · Business: €14.17 · Enterprise: Contact salesSocial media post generationShort-form video ads