Open-source alternative to Sora and Runway: video on your own server

Sora and Runway are used for short clips: animating a product photo, building an intro, making a video for social media. Open video models do the same inside the company — scripts, frames and client footage never reach a third-party service, and the number of attempts is limited only by your hardware, which matters for video because a good take rarely comes first. Video is heavier than images though: smaller models run on a gaming GPU, larger ones expect server cards, and each clip takes noticeable time to render. Clip length, face stability and motion physics are weaker in open models than in mature cloud services, so a complex narrative piece is still out of reach. The realistic use is short clips and animated stills, with editing and sound added on top.

Updated 22 Sep 2026Find a model in 4 questions

What to use instead

Image generationRU2022–2026

Kandinsky

Sber (Kandinsky Lab) · Russia

Sber's Russian family of image and video generation models. Understands Russian-language prompts and Russian cultural context well; released under MIT.

  • Images from Russian-language descriptions
  • Short promo videos from text or a photo
  • Instruction-based image editing
Sizes
2B – 19B
Hardware
from: 1 GPU
Commercial use allowedDetails
Video2025–2026

Wan

Alibaba · China

Text-to-video and image-to-video; the small version runs on a gaming GPU. After 2.2 only applied models are open: editing (VACE), audio-driven talking characters (S2V), dancing to music (Dancer).

  • Short promo videos
  • Animating product photos
  • Videos for social media
Sizes
1,3B – 14B
Hardware
from: 1 GPU
Commercial use allowedDetails
Video2024–2026

LTX-Video / LTX-2

Lightricks · Israel

A fast video model; with LTX-2 it generates video with sound and speech in one go. Camera and pose control, lightweight versions available.

  • Ad videos with sound
  • Video from a product photo
  • Voiced scenes for social media
Sizes
2B – 22B
Hardware
from: 1 GPU
Commercial use with conditionsDetails
VideoGGUF2024–2025

HunyuanVideo

Tencent · China

Tencent's video model, one of the first open ones on par with closed services. Version 1.5 is lighter (8.3B) and runs on consumer GPUs.

  • Video from a text script
  • Animating images
  • Base for fine-tuning your own video models
Sizes
8.3B – 13B
Hardware
from: 1 GPU
Commercial use with conditionsDetails
Video2024

CogVideoX

Zhipu AI (Z.ai) and Tsinghua University · China

A 2–5B video model that runs on a single gaming GPU. A popular base for research and add-ons.

  • Short clips from text
  • Animating images
  • Video fine-tuning experiments
Sizes
2B – 5B
Hardware
from: 1 GPU
Commercial use with conditionsDetails

Wan and Kandinsky are Apache 2.0 and MIT, while LTX-2 and HunyuanVideo carry their own community licenses with restrictions (HunyuanVideo's does not apply in the EU, the UK and South Korea), and CogVideoX 5B allows commercial use after registration. Have the terms of your chosen version checked by a lawyer.

Other alternatives

Need a model for your task?

An open model can run on your own server: data stays in-house, there is no per-request fee, and the model can be fine-tuned on your documents.

  1. SelectThe model and size for your task and hardware budget
  2. DeployOn your server or in a closed network, with an API
  3. Fine-tuneOn your data, or connect a knowledge base
  4. IntegrateInto your CRM, ERP, bot, website or team chat
Discuss deployment