Open-weight AI models

Open-weight models released since 2022: text, code, images, video, speech, 3D. For each one — what it does, where it is used, what hardware it needs and whether commercial use is allowed. I can deploy any of them on your server and fine-tune it for your task.

564 families38 categoriesUpdated 22 Sep 2026Guide: how to choose a model (in Russian)
Text and code
Documents and search
Images and video
Speech and audio
Industries and science

Showing 564 of 564

Image generationGGUF2025–2026

Qwen-Image

Alibaba · China

Image generation and editing, including text in images. Earlier versions allow commercial use; the latest 2.1 is non-commercial only.

  • Infographics for product cards
  • Photo editing by text command
  • Ad creatives
Sizes
7B – 20B
Hardware
from: 1 GPU
Commercial use with conditionsDetails
MedicineGGUF2023–2026

HuatuoGPT

FreedomIntelligence (The Chinese University of Hong Kong, Shenzhen) · China

A large family of medical models: chat, an imaging version, the reasoning HuatuoGPT-o1 and the new HuatuoGPT-3 on Qwen3. Does not replace a doctor; decisions are made by a specialist.

  • Draft discharge summaries for a doctor to review
  • Searching medical literature
  • Hints for doctors when reviewing images (Vision)
Sizes
7B – 72B
Hardware
from: Laptop
Commercial use allowedDetails
Speech to textRUGGUF2025–2026

VibeVoice

Microsoft · USA

Microsoft speech models: long multi-voice dialogue synthesis, fast synthesis for live conversation, and recognition of long recordings split by speaker, including in Russian.

  • Transcribing long meetings with speaker labels
  • Voicing podcasts and dialogues
  • Real-time voice for assistants
Sizes
0.5B – 9B
Hardware
from: Laptop
Commercial use allowedDetails
Music and soundGGUF2025–2026

YuE

M-A-P and HKUST · China

Generates full songs with vocals and accompaniment from lyrics and a style description: English, Chinese, Japanese, Korean.

  • Songs and jingles from lyrics
  • Demo versions of tracks
  • Music for videos
Sizes
0.5B – 7B
Hardware
from: 1 GPU
Commercial use with conditionsDetails
Music and sound2026

HeartMuLa

HeartMuLa Team · not disclosed

An open model for generating songs with vocals in Chinese, English, Japanese, Korean and Spanish, plus a codec and a lyrics transcription model.

  • Songs and jingles from lyrics
  • Music for videos
  • Transcribing song lyrics
Sizes
3B
Hardware
from: 1 GPU
Commercial use allowedDetails
TextOllama2023–2026

DeepSeek

DeepSeek · China

DeepSeek's flagship line: from the first 7B/67B to V4-Pro with 1.6 trillion parameters. Closed-model quality under an open MIT license; V4-Flash-Vision-Exp and V4.1-Flash understand images, context up to 1M tokens.

  • Employee assistant on your own server
  • Analysis of long contracts and reports
  • Agents that work with tools and APIs
Sizes
7B – 1.6T-A49B
Hardware
from: Laptop
Commercial use allowedDetails
TextOllama2023–2026

InternLM / Intern-S

Shanghai AI Laboratory · China

Models from Shanghai AI Laboratory. The early InternLM line is general-purpose; the new Intern-S1/S2 is scientific: it understands formulas, molecules, charts and images.

  • Research assistant: papers, formulas, data
  • Analysis of scientific and technical documents
  • Corporate chat on small models
Sizes
1.8B – about 1T
Hardware
from: Laptop
Commercial use allowedDetails
TextGGUF2025–2026

MiMo

Xiaomi · China

Xiaomi models for reasoning and agents: from the compact MiMo-7B to MiMo-V2.6-Pro with 1.02 trillion parameters. The larger versions understand text, images, video and audio, with a 1M token context. Languages: English and Chinese.

  • Logic and calculation tasks
  • Agents with tools
  • Help for developers
Sizes
7B – 1,02T-A42B
Hardware
from: Laptop
Commercial use allowedDetails
TextRU2024–2026

GigaChat

Sber · Russia

Sber open models with strong Russian language support and local context, from 10B-A1.8B to 702B, all MIT. GigaChat3.1-Audio handles recordings up to two hours; GFusion is a fast diffusion text version.

  • Russian-language employee assistant on your own server
  • Customer replies and request handling in Russian
  • Working with contracts and internal policies
Sizes
10B-A1.8B – 702B-A36B
Hardware
from: Laptop
Commercial use allowedDetails
TextRU2022–2026

YandexGPT / AliceAI

Yandex · Russia

Yandex models trained from scratch with a focus on the Russian language and Russian context. The new AliceAI-Foundation 80B-A3B (Apache 2.0) is a base model only, with no instruct version: you fine-tune it for your own tasks. The efficient AliceAI-T5 35B-A0.6B is also available.

  • Russian-language assistant and chatbot
  • Answers based on the company knowledge base
  • Base for industry-specific fine-tuning
Sizes
8B – 100B
Hardware
from: Laptop
Commercial use with conditionsDetails
TextRUOllama2024–2026

Aya

Cohere Labs · Canada

Multilingual models from Cohere's research arm, covering 23 to 100+ languages. Tiny Aya (2026, 3.3B) runs on a regular PC, but for non-commercial use only.

  • Translation and correspondence in less common languages
  • Multilingual chat assistant
  • Analysis of images with text (Vision)
Sizes
3.3B – 35B
Hardware
from: Laptop
Commercial use with conditionsDetails
Tabular data2022–2026

TabPFN

Prior Labs (University of Freiburg) · Germany

A ready-made model for tables: it takes example rows and immediately predicts for new ones, without lengthy training or tuning. Only v2 is free for business; newer versions are non-commercial.

  • Predicting customer churn from a CRM export
  • Scoring applications and leads
  • Classifying customers from 1C data
Sizes
from a few to hundreds of millions of parameters
Hardware
from: Laptop
Commercial use with conditionsDetails
Tabular data2025–2026

Mitra

Amazon (AutoGluon team) · USA

Amazon's tabular model built into AutoGluon: classification and regression from examples with brief fine-tuning. Mitra-v2 handles more rows and columns.

  • Predicting churn and repeat purchases
  • Scoring applications
  • Predicting deal or order value
Sizes
about 76M
Hardware
from: Laptop
Commercial use allowedDetails
Tabular data2025–2026

TabDPT

Layer 6 AI (TD Bank) · Canada

A tabular model from a Canadian bank's AI lab, trained on real tables rather than only synthetic ones. Version 1.2 Turbo made computation orders of magnitude faster.

  • Scoring applications and customers
  • Predicting churn
  • Classifying transactions and customers
Sizes
about 60–80M
Hardware
from: Laptop
Commercial use allowedDetails
Tabular data2025–2026

LimiX

Stable AI (Beijing, with Tsinghua University) · China

A table model that alone can classify, predict numbers and fill in missing data. The lightweight LimiX-2M runs on an ordinary computer.

  • Filling gaps in 1C and CRM exports
  • Churn prediction and scoring
  • Classifying customers and products
Sizes
2M – 16M and LimiX-2
Hardware
from: Laptop
Commercial use with conditionsDetails
Faces2021–2026

InsightFace

InsightFace (deepinsight) · China

The most widely used open toolkit for face detection and recognition. Many identity-preserving image generators are built on it. The pretrained weights are non-commercial.

  • Detecting and comparing faces in photos
  • Face-based access in prototypes
  • Face processing as part of other AI systems
Sizes
packages from 16 MB to 407 MB
Hardware
from: Laptop
Non-commercial onlyDetails
Text2024–2026

MiniCPM

OpenBMB (ModelBest and Tsinghua University) · China

Compact text models that run directly on a device: laptop, phone or mini PC. The 1B and 2B MiniCPM5 models focus on tool calling and long context.

  • A local chat assistant without the cloud
  • Data extraction and text classification
  • Tool calling and simple agents on low-end hardware
Sizes
0.5B – 8B
Hardware
from: Laptop
Commercial use allowedDetails
Text2024–2026

K2 (K2-Think, K2-V2, K2-Horizon)

MBZUAI, Institute of Foundation Models (IFM, LLM360 project) · UAE

Fully open models from the UAE: data, training code and intermediate checkpoints are published along with the weights. K2-Horizon (2026) spans 0.9B to 375B with context up to 512K tokens.

  • Reasoning, maths and technical questions
  • Analysing long documents
  • Agents and writing code
Sizes
0.9B – 375B-A23B
Hardware
from: Laptop
Commercial use allowedDetails
TextGGUF2025–2026

LLaDA

Renmin University of China (GSAI) and Ant Group (inclusionAI) · China

Diffusion language models: text is written in blocks and then refined rather than word by word, which speeds up generation. LLaDA2.2 can edit what it has written and targets agents. LLaDA-Image is a separate product.

  • Fast generation of code and text
  • Agent scenarios with long context
  • Research into alternatives to standard LLMs
Sizes
8B – 100B (MoE)
Hardware
from: 1 GPU
Commercial use allowedDetails
3D2024–2026

DUSt3R / MASt3R

NAVER LABS Europe · France (NAVER, South Korea)

The family that started "single-pass" 3D reconstruction from a pair or set of photos without camera calibration. MASt3R added point matching and scale; MUSt3R and BLASt3R added video support.

  • 3D scene from several photos without calibration
  • Point matching between images
  • Mapping from video (SLAM)
Sizes
0.57B – 0.69B
Hardware
from: Laptop
Non-commercial onlyDetails
Music and soundGGUF2022–2026

MERT

m-a-p (Multimodal Art Projection) · UK / China

A music encoder: turns a track into a numeric representation used to detect genre, mood, key and rhythm. MERT-v2 handles full songs up to 6 minutes.

  • Automatic tagging of a music catalog
  • Finding similar tracks
  • Detecting genre, mood and tempo
Sizes
95M – 632M
Hardware
from: Laptop
Non-commercial onlyDetails
Autonomous driving2020–2026

comma.ai openpilot (supercombo)

comma.ai · USA

An open driver assistance system: a neural network keeps the lane and controls speed from a camera, plus a driver attention monitoring model. The models live right in the repository and are updated constantly.

  • A research testbed for driver assistance systems
  • Studying driver attention monitoring with an in-cabin camera
  • Comparison with your own lane-keeping algorithms
Sizes
compact, designed for an in-vehicle device
Hardware
from: Laptop
Commercial use allowedDetails
Visual document searchGGUF2026

EVIE

Tencent · China

Tencent models based on Qwen3.5 for searching scans and PDFs as images. According to the model card, among the top of the ViDoRe leaderboard at release.

  • Search across scans and PDFs without OCR
  • RAG over reports with tables and charts
  • Search across document archives
Sizes
4.5B – 8B
Hardware
from: 1 GPU
Commercial use allowedDetails
TextRU2026

Zarya (ai-forever)

SberDevices (ai-forever) · Russia

A Russian and English research prototype: the model writes text in blocks at once (diffusion) rather than word by word, which speeds up responses. The authors do not recommend it for production systems.

  • Experiments with faster generation
  • Fine-tuning small models for your own tasks
  • Research
Sizes
0.6B – 4B
Hardware
from: Laptop
Commercial use allowedDetails
Documents and OCR2026

jina-ocr-v1

Jina AI · Germany

Document parsing in a single model: a whole page becomes Markdown - text in correct reading order, tables and formulas in LaTeX. Built on DeepSeek-OCR, with only 0.6B of its 3.4B parameters active.

  • Converting scans and PDFs to Markdown
  • Recognizing tables and formulas
  • Parsing invoices, acts and reports
Sizes
3.4B-A0.6B
Hardware
from: 1 GPU
Non-commercial onlyDetails
Voice assistants2026

Samsone

Samsung · South Korea

Tiny audio-understanding models for smartphones: they listen to speech, music and ambient sounds and answer in text - describing a recording and answering questions about it. They run on the device itself; prompts and answers are in English - no other languages are present in the training data.

  • Describing an audio recording in words
  • Answering questions about a sound
  • Identifying the type of sound and the setting
Sizes
99M – 356M
Hardware
from: Laptop
Non-commercial onlyDetails
Deepfake detection2023–2026

TrustMark

Adobe Research and University of Surrey · USA

An image watermark for arbitrary resolutions built for the Content Authenticity Initiative: it can both apply a mark and remove one. The detector errs in both directions - a human reviews the output.

  • Marking images on the way out of your own pipeline
  • Checking the provenance of a submitted image
  • Linking with content provenance metadata
Sizes
model types Q and P with different mark capacity
Hardware
from: Laptop
Commercial use allowedDetails
TextRUOllama2023–2026

Qwen

Alibaba · China

A family of language models with strong Russian language support, from small versions for a laptop to a flagship on par with commercial APIs.

  • Chatbot and knowledge-base assistant
  • Replies to emails and customer requests
  • Document parsing and classification
Sizes
0,6B – 2,4T-A95B
Hardware
from: Laptop
Commercial use with conditionsDetails
VideoGGUF2025–2026

MAGI

Sand AI · China

Video generated chunk by chunk in sequence, so a clip can be extended indefinitely. MAGI-2 produces video with sound.

  • Long videos with continuation
  • Video with sound
  • Animating images
Sizes
4.5B – 114B-A6B
Hardware
from: 1 GPU
Commercial use allowedDetails
Video2025–2026

SANA-Video

NVIDIA · USA

NVIDIA's lightweight, fast video model. Produces 720p clips on a single GPU; a 4-step version enables quick generation.

  • Quick clips for social media
  • Bulk video generation
  • Video from an image
Sizes
2B – 5B
Hardware
from: 1 GPU
Commercial use allowedDetails
Computer vision2024–2026

MoGe

Microsoft Research · USA

Reconstructs the 3D geometry of a scene from one photo: depth in meters, a point cloud and surface normals.

  • Measuring rooms and objects from photos
  • 3D point cloud from a single shot
  • Preparing data for robots and AR
Sizes
ViT-S – ViT-G
Hardware
from: Laptop
Commercial use allowedDetails
Search and RAGRU2024–2026

FRIDA / Giga-Embeddings

Sber (SberDevices) · Russia

Sber embeddings built for Russian: according to the developers, among the best on Russian-language search benchmarks. FRIDA is compact, Giga-Embeddings is more powerful.

  • Search across Russian-language documents
  • RAG for chatbots in Russian
  • Classifying requests and reviews
Sizes
480M – 10B-A1.8B
Hardware
from: Laptop
Commercial use allowedDetails
Forecasting2024–2026

TimesFM

Google · USA

A ready-made Google forecasting model: forecasts any time series without training on your data.

  • Sales and demand forecasting
  • Purchase and inventory planning
  • Load and traffic forecasting
Sizes
200M – 500M
Hardware
from: Laptop
Commercial use with conditionsDetails
Robotics2025–2026

GigaBrain

GigaAI · China

A robot control model trained mostly on synthetic data from a world model. It reduces spending on collecting data from real robots.

  • Controlling a robot arm
  • Fine-tuning on a small amount of your own data
  • Sorting and assembly pilots
Sizes
3.5B
Hardware
from: 1 GPU
Commercial use allowedDetails
Speech to textGGUF2024–2026

Moonshine

Moonshine AI (Useful Sensors) · USA

Very small and fast speech recognition models for phones, tablets and embedded devices. Version 2 streams, producing text while the person is still speaking.

  • Voice control of devices
  • Offline recognition on a phone
  • Live subtitles
Sizes
27M – 245M
Hardware
from: Laptop
Commercial use allowedDetails
Speech to textGGUF2025–2026

Granite Speech

IBM · USA

IBM speech models for recognizing and translating speech in English, several European languages and Japanese. Designed for enterprise use.

  • Transcribing business meetings
  • Translating speech into text in another language
  • Voice assistants
Sizes
470M – 8B
Hardware
from: Laptop
Commercial use allowedDetails
Text to speechGGUF2025–2026

IndexTTS

bilibili · China

Speech synthesis with voice cloning and precise duration control, handy for video dubbing. Controls emotion separately from timbre.

  • Video dubbing matched to timing
  • Voice cloning
  • Emotional voiceover
Sizes
about 1B – 2B
Hardware
from: Laptop
Commercial use with conditionsDetails
TextOllama2023–2026

GLM (ChatGLM)

Zhipu AI (Z.ai) · China

One of the oldest Chinese open lines: from ChatGLM-6B to GLM-5.3. Strong at agentic tasks and programming; GLM-5.3-Flash understands images and is released under MIT.

  • Corporate chat assistant
  • Agents for routine office tasks
  • Help for developers
Sizes
1.5B – 744B-A40B
Hardware
from: Laptop
Commercial use with conditionsDetails
Text2024–2026

Hunyuan / Hy

Tencent · China

Tencent language models: from small 0.5B–7B to Hy4-preview with 770 billion parameters. Since 2026 the line has been renamed Hy, and new versions are released under Apache 2.0.

  • Corporate assistant
  • Translation and multilingual texts
  • Agents with tools
Sizes
0.5B – 770B-A49B
Hardware
from: Laptop
Commercial use with conditionsDetails
Text2025–2026

Ling / Ring

Ant Group (inclusionAI) · China

An Ant Group family: Ling for standard models, Ring for reasoning ones. There are trillion-parameter flagships and the efficient Ling-3.0-tiny, which needs only 1.3 billion active parameters.

  • Corporate assistant
  • Agents for office processes
  • Financial analytics (Fin version available)
Sizes
7.9B-A1.3B – 1T
Hardware
from: Laptop
Commercial use allowedDetails
TextRUOllama2024–2026

Cohere Command

Cohere · Canada

Business models: document search with source citations, tool calling, many languages. Command A+ (2026) was the first under Apache 2.0, followed by the North line: code, translation and compact vision.

  • Knowledge-base answers with source citations
  • Agents that work with internal systems
  • Translation and correspondence in different languages
Sizes
2.5B – 218B-A25B
Hardware
from: Laptop
Commercial use with conditionsDetails
TextOllama2024–2026

NVIDIA Nemotron

NVIDIA · USA

NVIDIA models for agents and reasoning, optimized to run fast on its GPUs. Nemotron 3 is a Mamba and MoE hybrid from 4B to 550B; Nano Omni handles video, audio and images (English only).

  • Agents with tool calling
  • Reasoning and calculation tasks
  • Answers based on long documents
Sizes
4B – 550B-A55B
Hardware
from: Laptop
Commercial use with conditionsDetails
TextOllama2024–2026

IBM Granite

IBM · USA

IBM enterprise models with transparent training data and ISO 42001 certification. Granite 4 is a memory-efficient Mamba and Transformer hybrid.

  • Answers based on internal documents (RAG)
  • Tool calling and agent work
  • Data extraction and classification
Sizes
350M – 34B
Hardware
from: Laptop
Commercial use allowedDetails
TextRUOllama2025–2026

Liquid LFM

Liquid AI · USA

Models with a new architecture for on-device use: fast on a regular CPU and on phones. Versions for data extraction, RAG and tools, plus LFM2.5-VL for images and voice LFM2.5-Audio.

  • Offline assistant on a laptop or phone
  • Data extraction from documents
  • Tool calling in apps
Sizes
230M – 24B-A2B
Hardware
from: Laptop
Commercial use with conditionsDetails
TextOllama2026

Muse Glimmer

Meta Superintelligence Labs · USA

An open Meta model for agents on affordable hardware: distilled from the closed Muse Spark, understands text and images, trained on 100+ languages.

  • Agents with tool calling
  • Analysis of screenshots, charts and documents
  • Multilingual assistant
Sizes
30B
Hardware
from: 1 GPU
Commercial use allowedDetails
Text analysis2024–2026

GLiNER

Urchade Zaratiana and Fastino AI · France / USA

Finds the entities you need in text without training: just list what to look for (name, amount, date). GLiNER2 also classifies text. Multilingual versions understand Russian.

  • Extracting names, amounts and dates from emails and contracts
  • Parsing requests into CRM fields
  • Classifying requests by topic
Sizes
about 50M to 500M
Hardware
from: Laptop
Commercial use allowedDetails
Documents and OCR2026

TeleOCR

TeleAI (China Telecom) · China

A new lightweight document parsing model that led the OmniDocBench v1.6 benchmark at release. Handles pages photographed on a phone and crumpled pages well. Languages on the card: Chinese, English, Japanese.

  • Recognising invoices and delivery notes photographed on a phone
  • Recognising tables and formulas
  • Converting documents to Markdown for RAG
Sizes
about 1.2B
Hardware
from: Laptop
Commercial use allowedDetails
Voice: speakers and sound2022–2026

UVR / MDX-Net / RoFormer (разделение звука)

Community: Ultimate Vocal Remover (Anjok07), ZFTurbo, MVSep · International community

A large open collection of models for separating vocals from music and noise: MDX-Net, BS-RoFormer, Mel-RoFormer, SCNet. The quality leaders for vocals among open solutions.

  • Clean vocals from a recording with music
  • Backing tracks and stems for karaoke
  • Removing background music and noise from videos
Sizes
from tens to hundreds of millions of parameters
Hardware
from: Laptop
Commercial use with conditionsDetails
Computer-use agents2025–2026

OpenCUA / Qwen-CUA

XLANG Lab (University of Hong Kong) · China

Fully open desktop agents: weights, data and training code. They work on Windows, macOS and Linux; the latest Qwen-CUA controls a computer with ordinary clicks and keystrokes.

  • Working in desktop software without an API
  • Moving data between systems
  • Running user scenarios for tests
Sizes
7B – about 400B (MoE)
Hardware
from: 1 GPU
Commercial use allowedDetails
Computer-use agentsGGUF2025–2026

UI-Venus

Ant Group (inclusionAI) · China

An Ant Group family for finding elements on screen and completing tasks in phone and computer interfaces. UI-Venus-2 was specifically trained to refuse dangerous actions.

  • Automating actions in mobile apps
  • Filling in forms in web interfaces
  • UI autotests
Sizes
2B – 72B
Hardware
from: Laptop
Commercial use with conditionsDetails
CodeOllama2026

Ornith

DeepReinforce · not disclosed

Models for agentic development: they build their own plan and scaffolding for a task and execute it in the terminal. Fine-tuned from Qwen 3.5 and Gemma 4; work with Claude Code, OpenHands and similar tools.

  • A developer agent in the terminal
  • Fixing bugs from a task description
  • Understanding and extending a large repository
Sizes
9B – 397B
Hardware
from: 1 GPU
Commercial use allowedDetails
Weather and climate2024–2026

Ai2 ACE2 (климатический эмулятор)

Allen Institute for AI (Ai2) · USA

A fast climate model emulator: simulates the atmosphere years and decades ahead on a single GPU. Coupled with an ocean model (SamudrACE) for long-term scenarios.

  • Decades-long climate scenarios to assess long-term asset risks
  • Large-scale what-if runs on temperature and precipitation
  • Preparing data for crop yield and energy demand models
Sizes
checkpoint of about 1.8 GB
Hardware
from: 1 GPU
Commercial use allowedDetails
Weather and climate2023–2026

Google DeepMind GraphCast / GenCast / WeatherNext 2

Google DeepMind · UK

Google DeepMind's family of global weather models: GraphCast (10-day forecast), GenCast (probabilistic ensemble) and WeatherNext 2 with cyclone forecasting. Since August 2026 the weights are cleared for commercial use.

  • Medium-range weather forecasts for planning shifts, voyages and deliveries
  • Probabilistic assessment of extreme weather for insurance portfolios
  • Tropical cyclone track forecasts for marine and port operations
Sizes
from lightweight 1° versions to full 0.25°
Hardware
from: 1 GPU
Commercial use allowedDetails
Biology and chemistry2022–2026

OpenFold / OpenFold3

AlQuraishi Lab (Columbia University) and the OpenFold consortium · USA

A fully open reproduction of AlphaFold 2 and then AlphaFold 3 under Apache 2.0, with training data. OpenFold3 predicts complexes of proteins, nucleic acids and ligands.

  • Predicting structures of proteins and ligand complexes
  • Fine-tuning on the company's own data (training code is open)
  • An in-house structural analysis service without sending data outside
Sizes
a single set of weights per version
Hardware
from: 1 GPU
Commercial use allowedDetails
Autonomous driving2025–2026

NVIDIA Alpamayo

NVIDIA · USA

Vision-language-action models for self-driving vehicles: they plan a trajectory from camera video and explain the decision in text. Used to develop and test autopilot systems, not as a ready-made autopilot.

  • Auto-labeling camera recordings to train your own driver assistance systems
  • Analyzing complex road scenes with text explanations
  • Testing autopilot systems in simulation on rare scenarios
Sizes
10B – 34B
Hardware
from: 1 GPU
Commercial use with conditionsDetails
Autonomous driving2026

Qwen-Drive

Alibaba (Qwen team) · China

An autonomous driving model based on Qwen3.5-4B: 3D detection of objects around the vehicle, answers to questions about the road scene and trajectory planning in one model.

  • A perception and planning prototype for autonomous vehicles on closed sites
  • Answering questions about camera recordings when reviewing incidents
  • Labeling road scenes to train your own models
Sizes
4B
Hardware
from: 1 GPU
Commercial use allowedDetails
Deepfake detection2025–2026

Community Forensics

University of Michigan · USA

A lightweight detector of generated images, trained on 2.7M samples from nearly 5000 different generators. It errs in both directions: the result is a reason for a human to check, not proof.

  • Checking submitted photos and illustrations
  • Filtering AI images in a content flow
  • Flagging suspicious images for manual review
Sizes
22M
Hardware
from: Laptop
Commercial use allowedDetails
Deepfake detection2023–2026

UniversalFakeDetect

University of Wisconsin-Madison · USA

An early and still used approach: a simple classifier trained on top of a frozen CLIP that transfers to unseen generators. It errs in both directions - the output needs a human check.

  • Checking images from new, unfamiliar generators
  • A baseline when comparing detectors
  • Fast rollout of a check without training a large model
Sizes
a linear classifier on top of CLIP ViT-L/14
Hardware
from: Laptop
Commercial use allowedDetails
TextRUOllama2023–2026

Mistral

Mistral AI · France

European models focused on speed. Mixtral was one of the first open mixture-of-experts models; there are versions for images (Pixtral, Medium 3.5), Lean proofs and moderation (Shieldstral).

  • Fast chat responses
  • Data extraction from text
  • Translation and multilingual work
Sizes
3B – 675B
Hardware
from: Laptop
Commercial use with conditionsDetails
Video2025–2026

Wan

Alibaba · China

Text-to-video and image-to-video; the small version runs on a gaming GPU. After 2.2 only applied models are open: editing (VACE), audio-driven talking characters (S2V), dancing to music (Dancer).

  • Short promo videos
  • Animating product photos
  • Videos for social media
Sizes
1,3B – 14B
Hardware
from: 1 GPU
Commercial use allowedDetails

Model comparison

Collections

Need a model for your task?

An open model can run on your own server: data stays in-house, there is no per-request fee, and the model can be fine-tuned on your documents.

  1. SelectThe model and size for your task and hardware budget
  2. DeployOn your server or in a closed network, with an API
  3. Fine-tuneOn your data, or connect a knowledge base
  4. IntegrateInto your CRM, ERP, bot, website or team chat
Discuss deployment