ComfyUI
GPL-3.0 node graph for image, video, audio and 3D, with Stable Diffusion, SDXL, Flux, Wan and LTX, LoRAs and ControlNets, running fully offline.
Know firstA node graph takes learning, and commits between stable tags can break custom nodes.
The open-source alternatives to Midjourney are ComfyUI, InvokeAI, SD.Next and SwarmUI. Each runs open image models such as Stable Diffusion, SDXL and Flux on your own hardware, and three of them also generate video. ComfyUI gives a node graph with the widest model support, InvokeAI a canvas for inpainting and outpainting, SD.Next a form-based interface that runs on modest or unusual GPUs, and SwarmUI a simple tab with ComfyUI's full graph behind it. The software is free; each model has its own licence.
4 of 4 carry an OSI licence. 4 of 4 run on your own machines. Figures read from GitHub on September 17, 2026.
One project per need, with the reason from its own README and the one thing to know before you choose it.
GPL-3.0 node graph for image, video, audio and 3D, with Stable Diffusion, SDXL, Flux, Wan and LTX, LoRAs and ControlNets, running fully offline.
Know firstA node graph takes learning, and commits between stable tags can break custom nodes.
Apache-2.0 with a unified canvas for generation, inpainting, outpainting and brushes, node workflows, and 28 model families.
Know firstHardware requirements are not stated, and three listed models run only through paid APIs.
Apache-2.0 web interface with LoRA, ControlNet and quantisation its README says cuts video memory up to four times, on AMD, Intel, Apple or DirectX GPUs and CPUs.
Know firstThe README lists no model families and the project has no GitHub releases, only dated tags.
MIT interface with a beginner Generate tab and the raw ComfyUI graph behind it, plus multi-GPU generation for one user.
Know firstIts README calls it Almost-Release status, and the last tagged release is a beta.
The facts that decide most choices. Each project's page carries the full panel.
| Project | Best for | Openness | Licence | Runs as | Own machines | Stars |
|---|---|---|---|---|---|---|
| ComfyUI | Control, and the newest models | OSI | GPL-3.0 | Desktop app, Portable build, pip (comfy-cli), From source | Yes | 133,702 |
| InvokeAI | A canvas for inpainting and editing | OSI | Apache-2.0 | Launcher (desktop), Local web server | Yes | 28,237 |
| SD.Next | Modest or unusual hardware | OSI | Apache-2.0 | Web UI (built-in installer), Docker | Yes | 7,342 |
| SwarmUI | A simple tab now, a graph later | OSI | MIT | Web UI (installer scripts), Docker, Google Colab | Yes | 4,571 |
Licence and star figures read from GitHub on September 17, 2026. Stars are shown as a dated fact and were not used to rank this page.
The GPL-3.0 node-graph engine for image, video, audio and 3D generation that supports the newest open models, from Stable Diffusion and Flux to Wan and LTX video, on NVIDIA, AMD, Intel and Apple hardware, offline.
GPL-3.0 node graph for image, video, audio and 3D, with Stable Diffusion, SDXL, Flux, Wan and LTX, LoRAs and ControlNets, running fully offline.
Know firstA node graph takes learning, and commits between stable tags can break custom nodes.
An Apache-2.0 creative image tool with a unified canvas for inpainting and outpainting, node workflows, and support for 28 model families from Stable Diffusion to Flux and Qwen Image, run as a local web app.
Apache-2.0 with a unified canvas for generation, inpainting, outpainting and brushes, node workflows, and 28 model families.
Know firstHardware requirements are not stated, and three listed models run only through paid APIs.
An Apache-2.0 all-in-one web interface for AI image and video generation, descended from AUTOMATIC1111's WebUI, with LoRA, ControlNet, quantisation that cuts video memory up to four times, and the widest hardware list here.
Apache-2.0 web interface with LoRA, ControlNet and quantisation its README says cuts video memory up to four times, on AMD, Intel, Apple or DirectX GPUs and CPUs.
Know firstThe README lists no model families and the project has no GitHub releases, only dated tags.
An MIT image and video generation web interface, formerly StableSwarmUI, with a simple Generate tab for beginners and a raw ComfyUI workflow tab for experts, that can spread work across several GPUs.
MIT interface with a beginner Generate tab and the raw ComfyUI graph behind it, plus multi-GPU generation for one user.
Know firstIts README calls it Almost-Release status, and the last tagged release is a beta.
The questions that settle it, in the order they usually come up.
ComfyUI's graph for every step, SwarmUI's simple tab to start, InvokeAI's canvas for editing, SD.Next's form for a familiar layout.
SD.Next covers AMD, Intel, Apple and DirectX GPUs and CPUs with memory-saving quantisation. ComfyUI covers NVIDIA, AMD, Intel, Apple and Ascend.
ComfyUI, SwarmUI and SD.Next generate video; InvokeAI lists one video model as API only.
Midjourney is not described here in its own words: its site refused every automated request on September 17, 2026, and this directory does not write a description it has not read. See midjourney.com for the vendor's own account. The cases below are where it still wins.
Stars, licence, last commit and last release are read from the GitHub API by a script every week and carry the date they were read. Nobody types them.
What each project does is written from its own README and LICENSE, read in full on the date shown, and phrased as the project's claim. The sources are listed at the end.
That a project is faster, better or cheaper than Midjourney. A price for anyone. A user count. A roadmap. If a fact is not on this page, it was not in the source.
ComfyUI (GPL-3.0) for the widest support of new open models and full control through a node graph, including video. InvokeAI (Apache-2.0) for a canvas with inpainting and outpainting. SD.Next (Apache-2.0) for AMD, Intel or low-memory hardware. SwarmUI (MIT) for a simple interface with ComfyUI behind it.
The software licences allow it, but the model does not always. Each open model carries its own licence and some are non-commercial, as SwarmUI's README notes. Check the model card of whichever model you generate with, not just the interface you run it in.
It depends on the model. SD.Next lists NVIDIA, AMD, Intel, Apple and DirectX GPUs and CPU-only use, with quantisation that its README says cuts video memory up to four times. ComfyUI covers NVIDIA, AMD, Intel, Apple Silicon and Ascend. None of the READMEs states a minimum amount of video memory.
Yes, three of them. ComfyUI lists Wan, LTX-Video, HunyuanVideo, CogVideoX and Mochi. SwarmUI lists MiniMax H3, Wan and LTX-2. SD.Next covers text to video and image to video. InvokeAI lists one video model, Wan, as API only.
Because its site refused every automated request when this page was written, and this directory does not write a description it has not read. Every other page here quotes the vendor's own words from a page that could be fetched.
Running one of these inside your own environment, with your own data and your own security rules, is the kind of work Reveneau does. Read how a forward deployed engagement works.