stable-diffusion-webui/README.md

# Stable Diffusion web UI
A browser interface based on Gradio library for Stable Diffusion.

![](txt2img_Screenshot.png)

Check the [custom scripts](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Custom-Scripts) wiki page for extra scripts developed by users.

## Features
[Detailed feature showcase with images](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Features):
- Original txt2img and img2img modes
- One click install and run script (but you still must install python and git)
- Outpainting
- Inpainting
- Prompt Matrix
- Stable Diffusion Upscale
- Attention, specify parts of text that the model should pay more attention to
    - a man in a ((tuxedo)) - will pay more attention to tuxedo
    - a man in a (tuxedo:1.21) - alternative syntax
    - select text and press ctrl+up or ctrl+down to automatically adjust attention to selected text (code contributed by anonymous user)
- Loopback, run img2img processing multiple times
- X/Y plot, a way to draw a 2 dimensional plot of images with different parameters
- Textual Inversion
    - have as many embeddings as you want and use any names you like for them
    - use multiple embeddings with different numbers of vectors per token
    - works with half precision floating point numbers
- Extras tab with:
    - GFPGAN, neural network that fixes faces
    - CodeFormer, face restoration tool as an alternative to GFPGAN
    - RealESRGAN, neural network upscaler
    - ESRGAN, neural network upscaler with a lot of third party models
    - SwinIR, neural network upscaler
    - LDSR, Latent diffusion super resolution upscaling
- Resizing aspect ratio options
- Sampling method selection
- Interrupt processing at any time
- 4GB video card support (also reports of 2GB working)
- Correct seeds for batches
- Prompt length validation
     - get length of prompt in tokens as you type
     - get a warning after generation if some text was truncated
- Generation parameters
     - parameters you used to generate images are saved with that image
     - in PNG chunks for PNG, in EXIF for JPEG
     - can drag the image to PNG info tab to restore generation parameters and automatically copy them into UI
     - can be disabled in settings
- Settings page
- Running arbitrary python code from UI (must run with --allow-code to enable)
- Mouseover hints for most UI elements
- Possible to change defaults/mix/max/step values for UI elements via text config
- Random artist button
- Tiling support, a checkbox to create images that can be tiled like textures
- Progress bar and live image generation preview
- Negative prompt, an extra text field that allows you to list what you don't want to see in generated image
- Styles, a way to save part of prompt and easily apply them via dropdown later
- Variations, a way to generate same image but with tiny differences
- Seed resizing, a way to generate same image but at slightly different resolution
- CLIP interrogator, a button that tries to guess prompt from an image
- Prompt Editing, a way to change prompt mid-generation, say to start making a watermelon and switch to anime girl midway
- Batch Processing, process a group of files using img2img
- Img2img Alternative
- Highres Fix, a convenience option to produce high resolution pictures in one click without usual distortions
- Reloading checkpoints on the fly
- Checkpoint Merger, a tab that allows you to merge two checkpoints into one
- [Custom scripts](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Custom-Scripts) with many extensions from community
- [Composable-Diffusion](https://energy-based-model.github.io/Compositional-Visual-Generation-with-Composable-Diffusion-Models/), a way to use multiple prompts at once
     - separate prompts using uppercase `AND`
     - also supports weights for prompts: `a cat :1.2 AND a dog AND a penguin :2.2`
- No token limit for prompts (original stable diffusion lets you use up to 75 tokens)
- DeepDanbooru integration, creates danbooru style tags for anime prompts (add --deepdanbooru to commandline args)

## Installation and Running
Make sure the required [dependencies](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Dependencies) are met and follow the instructions available for both [NVidia](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Install-and-Run-on-NVidia-GPUs) (recommended) and [AMD](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Install-and-Run-on-AMD-GPUs) GPUs.

Alternatively, use Google Colab:

- [Colab, maintained by Akaibu](https://colab.research.google.com/drive/1kw3egmSn-KgWsikYvOMjJkVDsPLjEMzl)
- [Colab, original by me, outdated](https://colab.research.google.com/drive/1Iy-xW9t1-OQWhb0hNxueGij8phCyluOh).

### Automatic Installation on Windows
1. Install [Python 3.10.6](https://www.python.org/downloads/windows/), checking "Add Python to PATH"
2. Install [git](https://git-scm.com/download/win).
3. Download the stable-diffusion-webui repository, for example by running `git clone https://github.com/AUTOMATIC1111/stable-diffusion-webui.git`.
4. Place `model.ckpt` in the `models` directory (see [dependencies](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Dependencies) for where to get it).
5. _*(Optional)*_ Place `GFPGANv1.4.pth` in the base directory, alongside `webui.py` (see [dependencies](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Dependencies) for where to get it).
6. Run `webui-user.bat` from Windows Explorer as normal, non-administrator, user.

### Automatic Installation on Linux
1. Install the dependencies:
```bash
# Debian-based:
sudo apt install wget git python3 python3-venv
# Red Hat-based:
sudo dnf install wget git python3
# Arch-based:
sudo pacman -S wget git python3
```
2. To install in `/home/$(whoami)/stable-diffusion-webui/`, run:
```bash
bash <(wget -qO- https://raw.githubusercontent.com/AUTOMATIC1111/stable-diffusion-webui/master/webui.sh)
```

### Installation on Apple Silicon

Find the instructions [here](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Installation-on-Apple-Silicon).

## Contributing
Here's how to add code to this repo: [Contributing](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Contributing)

## Documentation
The documentation was moved from this README over to the project's [wiki](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki).

## Credits
- Stable Diffusion - https://github.com/CompVis/stable-diffusion, https://github.com/CompVis/taming-transformers
- k-diffusion - https://github.com/crowsonkb/k-diffusion.git
- GFPGAN - https://github.com/TencentARC/GFPGAN.git
- CodeFormer - https://github.com/sczhou/CodeFormer
- ESRGAN - https://github.com/xinntao/ESRGAN
- SwinIR - https://github.com/JingyunLiang/SwinIR
- LDSR - https://github.com/Hafiidz/latent-diffusion
- Ideas for optimizations - https://github.com/basujindal/stable-diffusion
- Doggettx - Cross Attention layer optimization - https://github.com/Doggettx/stable-diffusion, original idea for prompt editing.
- Rinon Gal - Textual Inversion - https://github.com/rinongal/textual_inversion (we're not using his code, but we are using his ideas).
- Idea for SD upscale - https://github.com/jquesnelle/txt2imghd
- Noise generation for outpainting mk2 - https://github.com/parlance-zz/g-diffuser-bot
- CLIP interrogator idea and borrowing some code - https://github.com/pharmapsychotic/clip-interrogator
- Initial Gradio script - posted on 4chan by an Anonymous user. Thank you Anonymous user.
- DeepDanbooru - interrogator for anime diffusors https://github.com/KichangKim/DeepDanbooru
- (You)
first 2022-08-22 08:15:46 -06:00			`# Stable Diffusion web UI`
			`A browser interface based on Gradio library for Stable Diffusion.`

Update README.md 2022-09-23 18:22:02 -06:00			`![](txt2img_Screenshot.png)`
moved images to 2022-09-04 04:08:06 -06:00
add advertisement for custom scripts to readme 2022-09-29 23:18:05 -06:00			`Check the [custom scripts](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Custom-Scripts) wiki page for extra scripts developed by users.`

Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`## Features`
			`[Detailed feature showcase with images](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Features):`
moved images to 2022-09-04 04:08:06 -06:00			`- Original txt2img and img2img modes`
Removed mention of CUDA in the README The requirement to install CUDA was removed with https://github.com/AUTOMATIC1111/stable-diffusion-webui/commit/e92d4cf7476f1897fce376916dfb40755ea7920f#diff-b335630551682c19a781afebcf4d07bf978fb1f8ac04c6bf87428ed5106870f5L63 so that mention in README should be superfluous 2022-09-09 02:39:41 -06:00			`- One click install and run script (but you still must install python and git)`
moved images to 2022-09-04 04:08:06 -06:00			`- Outpainting`
			`- Inpainting`
Grammar Fix 2022-09-30 18:15:43 -06:00			`- Prompt Matrix`
			`- Stable Diffusion Upscale`
features updates unused code removed from outpainting mk2 2022-09-30 15:38:48 -06:00			`- Attention, specify parts of text that the model should pay more attention to`
Grammar Fix 2022-09-30 18:15:43 -06:00			`- a man in a ((tuxedo)) - will pay more attention to tuxedo`
			`- a man in a (tuxedo:1.21) - alternative syntax`
add info about cross attention javascript shortcut code 2022-10-08 01:15:29 -06:00			`- select text and press ctrl+up or ctrl+down to automatically adjust attention to selected text (code contributed by anonymous user)`
Grammar Fix 2022-09-30 18:15:43 -06:00			`- Loopback, run img2img processing multiple times`
features updates unused code removed from outpainting mk2 2022-09-30 15:38:48 -06:00			`- X/Y plot, a way to draw a 2 dimensional plot of images with different parameters`
moved images to 2022-09-04 04:08:06 -06:00			`- Textual Inversion`
features updates unused code removed from outpainting mk2 2022-09-30 15:38:48 -06:00			`- have as many embeddings as you want and use any names you like for them`
			`- use multiple embeddings with different numbers of vectors per token`
			`- works with half precision floating point numbers`
ESRGAN support 2022-09-04 09:54:12 -06:00			`- Extras tab with:`
Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`- GFPGAN, neural network that fixes faces`
			`- CodeFormer, face restoration tool as an alternative to GFPGAN`
			`- RealESRGAN, neural network upscaler`
features updates unused code removed from outpainting mk2 2022-09-30 15:38:48 -06:00			`- ESRGAN, neural network upscaler with a lot of third party models`
Update README.md Added SwinIR and new features to readme 2022-09-21 14:58:41 -06:00			`- SwinIR, neural network upscaler`
Update README.md 2022-09-21 16:53:35 -06:00			`- LDSR, Latent diffusion super resolution upscaling`
ESRGAN support 2022-09-04 09:54:12 -06:00			`- Resizing aspect ratio options`
moved images to 2022-09-04 04:08:06 -06:00			`- Sampling method selection`
			`- Interrupt processing at any time`
features updates unused code removed from outpainting mk2 2022-09-30 15:38:48 -06:00			`- 4GB video card support (also reports of 2GB working)`
chore: Fix typos 2022-10-08 13:12:24 -06:00			`- Correct seeds for batches`
moved images to 2022-09-04 04:08:06 -06:00			`- Prompt length validation`
Grammar Fix 2022-09-30 18:15:43 -06:00			`- get length of prompt in tokens as you type`
			`- get a warning after generation if some text was truncated`
features updates unused code removed from outpainting mk2 2022-09-30 15:38:48 -06:00			`- Generation parameters`
			`- parameters you used to generate images are saved with that image`
			`- in PNG chunks for PNG, in EXIF for JPEG`
			`- can drag the image to PNG info tab to restore generation parameters and automatically copy them into UI`
			`- can be disabled in settings`
moved images to 2022-09-04 04:08:06 -06:00			`- Settings page`
Grammar Fix 2022-09-30 18:15:43 -06:00			`- Running arbitrary python code from UI (must run with --allow-code to enable)`
Fix multiple grammar & spelling errors in ReadMe 2022-09-12 18:00:40 -06:00			`- Mouseover hints for most UI elements`
added UI config file: ui-config.json 2022-09-04 04:52:01 -06:00			`- Possible to change defaults/mix/max/step values for UI elements via text config`
fix for live progress breaking lowvram and medvram optimizations 2022-09-06 14:10:12 -06:00			`- Random artist button`
features updates unused code removed from outpainting mk2 2022-09-30 15:38:48 -06:00			`- Tiling support, a checkbox to create images that can be tiled like textures`
fix for live progress breaking lowvram and medvram optimizations 2022-09-06 14:10:12 -06:00			`- Progress bar and live image generation preview`
features updates unused code removed from outpainting mk2 2022-09-30 15:38:48 -06:00			`- Negative prompt, an extra text field that allows you to list what you don't want to see in generated image`
			`- Styles, a way to save part of prompt and easily apply them via dropdown later`
			`- Variations, a way to generate same image but with tiny differences`
			`- Seed resizing, a way to generate same image but at slightly different resolution`
			`- CLIP interrogator, a button that tries to guess prompt from an image`
			`- Prompt Editing, a way to change prompt mid-generation, say to start making a watermelon and switch to anime girl midway`
			`- Batch Processing, process a group of files using img2img`
Update README.md Added SwinIR and new features to readme 2022-09-21 14:58:41 -06:00			`- Img2img Alternative`
features updates unused code removed from outpainting mk2 2022-09-30 15:38:48 -06:00			`- Highres Fix, a convenience option to produce high resolution pictures in one click without usual distortions`
			`- Reloading checkpoints on the fly`
			`- Checkpoint Merger, a tab that allows you to merge two checkpoints into one`
			`- [Custom scripts](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Custom-Scripts) with many extensions from community`
added ctrl+up or ctrl+down hotkeys for attention 2022-10-06 14:44:54 -06:00			`- [Composable-Diffusion](https://energy-based-model.github.io/Compositional-Visual-Generation-with-Composable-Diffusion-Models/), a way to use multiple prompts at once`
			- separate prompts using uppercase `AND`
			- also supports weights for prompts: `a cat :1.2 AND a dog AND a penguin :2.2`
do not let user choose his own prompt token count limit 2022-10-08 05:25:47 -06:00			`- No token limit for prompts (original stable diffusion lets you use up to 75 tokens)`
made deepdanbooru optional, added to readme, automatic download of deepbooru model 2022-10-08 10:02:56 -06:00			`- DeepDanbooru integration, creates danbooru style tags for anime prompts (add --deepdanbooru to commandline args)`
moved images to 2022-09-04 04:08:06 -06:00
Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`## Installation and Running`
			`Make sure the required [dependencies](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Dependencies) are met and follow the instructions available for both [NVidia](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Install-and-Run-on-NVidia-GPUs) (recommended) and [AMD](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Install-and-Run-on-AMD-GPUs) GPUs.`
let me tell you about realesrgan 2022-09-11 03:13:26 -06:00
upgrade to gradio==3.4b3 t fixthe inpain bugs rework progressbar/preview to work with new gradio remove unnecessary create style button added link to alternative colab 2022-09-23 11:46:02 -06:00			`Alternatively, use Google Colab:`

removed some information that its owner decided he did not want to share 2022-09-23 11:55:54 -06:00			`- [Colab, maintained by Akaibu](https://colab.research.google.com/drive/1kw3egmSn-KgWsikYvOMjJkVDsPLjEMzl)`
upgrade to gradio==3.4b3 t fixthe inpain bugs rework progressbar/preview to work with new gradio remove unnecessary create style button added link to alternative colab 2022-09-23 11:46:02 -06:00			`- [Colab, original by me, outdated](https://colab.research.google.com/drive/1Iy-xW9t1-OQWhb0hNxueGij8phCyluOh).`
bat file for installing and launching 2022-09-02 00:49:35 -06:00
Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`### Automatic Installation on Windows`
			`1. Install [Python 3.10.6](https://www.python.org/downloads/windows/), checking "Add Python to PATH"`
			`2. Install [git](https://git-scm.com/download/win).`
			3. Download the stable-diffusion-webui repository, for example by running `git clone https://github.com/AUTOMATIC1111/stable-diffusion-webui.git`.
add info about where to get models to instructions. 2022-09-18 00:09:10 -06:00			4. Place `model.ckpt` in the `models` directory (see [dependencies](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Dependencies) for where to get it).
Update README.md 2022-09-19 10:24:58 -06:00			5. _(Optional)_ Place `GFPGANv1.4.pth` in the base directory, alongside `webui.py` (see [dependencies](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Dependencies) for where to get it).
add info about where to get models to instructions. 2022-09-18 00:09:10 -06:00			6. Run `webui-user.bat` from Windows Explorer as normal, non-administrator, user.
brought manual instructions up to date reworked launching with different parameters 2022-09-08 23:37:19 -06:00
Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`### Automatic Installation on Linux`
			`1. Install the dependencies:`
			```bash
			`# Debian-based:`
install/launch scripts for linux 2022-09-13 07:28:04 -06:00			`sudo apt install wget git python3 python3-venv`
Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`# Red Hat-based:`
install/launch scripts for linux 2022-09-13 07:28:04 -06:00			`sudo dnf install wget git python3`
Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`# Arch-based:`
			`sudo pacman -S wget git python3`
install/launch scripts for linux 2022-09-13 07:28:04 -06:00			```
Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			2. To install in `/home/$(whoami)/stable-diffusion-webui/`, run:
put manual instructions above WSL, clarify they work for windows and linux 2022-09-09 15:31:58 -06:00			```bash
Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`bash <(wget -qO- https://raw.githubusercontent.com/AUTOMATIC1111/stable-diffusion-webui/master/webui.sh)`
put manual instructions above WSL, clarify they work for windows and linux 2022-09-09 15:31:58 -06:00			```

Update README to link to wiki page for Apple Silicon installs 2022-09-22 03:35:12 -06:00			`### Installation on Apple Silicon`

			`Find the instructions [here](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Installation-on-Apple-Silicon).`

added contributing to readme 2022-09-30 02:57:19 -06:00			`## Contributing`
			`Here's how to add code to this repo: [Contributing](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki/Contributing)`

Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`## Documentation`
			`The documentation was moved from this README over to the project's [wiki](https://github.com/AUTOMATIC1111/stable-diffusion-webui/wiki).`
initial work on img2imgalt 2022-09-11 16:55:34 -06:00
added credits section 2022-09-04 10:09:00 -06:00			`## Credits`
			`- Stable Diffusion - https://github.com/CompVis/stable-diffusion, https://github.com/CompVis/taming-transformers`
			`- k-diffusion - https://github.com/crowsonkb/k-diffusion.git`
			`- GFPGAN - https://github.com/TencentARC/GFPGAN.git`
Added CodeFormer to Credits request from Shangchen Zhou 2022-09-21 01:38:06 -06:00			`- CodeFormer - https://github.com/sczhou/CodeFormer`
added credits section 2022-09-04 10:09:00 -06:00			`- ESRGAN - https://github.com/xinntao/ESRGAN`
Update README.md Added SwinIR and new features to readme 2022-09-21 14:58:41 -06:00			`- SwinIR - https://github.com/JingyunLiang/SwinIR`
Update README.md 2022-09-21 16:53:35 -06:00			`- LDSR - https://github.com/Hafiidz/latent-diffusion`
Update to cross attention from https://github.com/Doggettx/stable-diffusion #219 2022-09-10 03:06:19 -06:00			`- Ideas for optimizations - https://github.com/basujindal/stable-diffusion`
add prompt editing to readme 2022-09-15 06:54:11 -06:00			`- Doggettx - Cross Attention layer optimization - https://github.com/Doggettx/stable-diffusion, original idea for prompt editing.`
credit Rinon Gal 2022-10-02 14:22:48 -06:00			`- Rinon Gal - Textual Inversion - https://github.com/rinongal/textual_inversion (we're not using his code, but we are using his ideas).`
added credits section 2022-09-04 10:09:00 -06:00			`- Idea for SD upscale - https://github.com/jquesnelle/txt2imghd`
added Noise generation for outpainting mk2 to credits 2022-09-16 14:17:10 -06:00			`- Noise generation for outpainting mk2 - https://github.com/parlance-zz/g-diffuser-bot`
CLIP interrogator 2022-09-11 09:48:36 -06:00			`- CLIP interrogator idea and borrowing some code - https://github.com/pharmapsychotic/clip-interrogator`
added credits section 2022-09-04 10:09:00 -06:00			`- Initial Gradio script - posted on 4chan by an Anonymous user. Thank you Anonymous user.`
made deepdanbooru optional, added to readme, automatic download of deepbooru model 2022-10-08 10:02:56 -06:00			`- DeepDanbooru - interrogator for anime diffusors https://github.com/KichangKim/DeepDanbooru`
Rewrote a large portion of the README to point towards the wiki. 2022-09-15 04:50:55 -06:00			`- (You)`