What's on the canvas Text-to-image Hi-resolution fix Image-to-image Let an LLM write the prompt Read an image back into a prompt Detection to inpaint mask Prompt matrix Upscale and background removal Annotators PNG Info Image-to-video Running models on your own GPU Every output is an API Where this sits next to ComfyUI Build your own In our last post, we built five small gr.Workflow graphs and hinted at what it would take to build something as complex as AUTOMATIC1111's stable-diffusion-webui. In this post we walk you through Workflow1111, where we have rebuilt most of AUTOMATIC1111's feature set as a single workflow canvas.

Workflow1111 is a graph of eleven media pipelines built using seventy-three nodes. It brings together SOTA models for text-to-image, hi-resolution fix, image-to-image, prompt-matrix grids, VLM interrogate, detection-to-inpaint masks, ControlNet-style annotators, background removal, PNG Info storing, and image-to-video.