Open-source alternative to Midjourney: image generation on your own hardware

Midjourney is bought for visuals: ads, social posts, covers. Open image models do the same work on your own machine or server: you can generate as much as you like, prompts and reference images never leave your side, and a model can be adapted to your style or product with a separate fine-tune. In exchange you take on the GPU, the interface setup and the search for good settings — a usable picture rarely comes out of the first prompt, it usually takes several iterations. Composition, hands and text inside the image are still uneven in open models, so part of the output has to be discarded or retouched. The one thing to check before you start is the license of the exact version: within a single family it can be either permissive or strictly non-commercial.

Updated 22 Sep 2026Find a model in 4 questions

What to use instead

Image generationRU2022–2026

Kandinsky

Sber (Kandinsky Lab) · Russia

Sber's Russian family of image and video generation models. Understands Russian-language prompts and Russian cultural context well; released under MIT.

  • Images from Russian-language descriptions
  • Short promo videos from text or a photo
  • Instruction-based image editing
Sizes
2B – 19B
Hardware
from: 1 GPU
Commercial use allowedDetails
Image generationGGUF2024–2026

FLUX

Black Forest Labs · Germany

Image generation from the creators of Stable Diffusion. Renders text in images well and keeps the composition.

  • Images for product cards
  • Banners and covers
  • Photo editing by description (Kontext)
Sizes
4B – 32B
Hardware
from: 1 GPU
Commercial use with conditionsDetails
Image generationGGUF2022–2024

Stable Diffusion

Stability AI · UK

The model that started open image generation. A huge ecosystem of fine-tunes, styles and plugins; runs even on a home PC. The popular SDXL-Lightning and Hyper-SD accelerators were made by ByteDance.

  • Illustrations and banners for advertising
  • Backgrounds and scenes for product cards
  • Fine-tuning to a brand style
Sizes
0.9B – 8B
Hardware
from: Laptop
Commercial use with conditionsDetails
Image generationGGUF2025–2026

Qwen-Image

Alibaba · China

Image generation and editing, including text in images. Earlier versions allow commercial use; the latest 2.1 is non-commercial only.

  • Infographics for product cards
  • Photo editing by text command
  • Ad creatives
Sizes
7B – 20B
Hardware
from: 1 GPU
Commercial use with conditionsDetails
Image generationGGUF2025–2026

Z-Image

Alibaba (Tongyi-MAI) · China

A compact 6B model with photorealism on par with large models. The Turbo version produces an image in a few steps on a regular gaming GPU.

  • Photorealistic ad images
  • Images with English and Chinese text
  • Bulk visual generation
Sizes
6B
Hardware
from: 1 GPU
Commercial use allowedDetails

Kandinsky and Z-Image are Apache 2.0 and MIT, whereas FLUX [dev] versions are non-commercial, Stable Diffusion licensing depends on the generation, and the recent Qwen-Image-2.1 was released under a research license. For commercial campaigns, have the exact version and its license reviewed by a lawyer.

Other alternatives

Need a model for your task?

An open model can run on your own server: data stays in-house, there is no per-request fee, and the model can be fine-tuned on your documents.

  1. SelectThe model and size for your task and hardware budget
  2. DeployOn your server or in a closed network, with an API
  3. Fine-tuneOn your data, or connect a knowledge base
  4. IntegrateInto your CRM, ERP, bot, website or team chat
Discuss deployment