F5-TTS or Fish Speech: which voiceover model to pick

Both models clone a voice from a short sample, and both close off commercial use, so a product needs a separate agreement or a different model. A voice may only be cloned with the consent of the person it belongs to. F5-TTS is about 340M, runs on a CPU, officially covers English and Chinese, and gets Russian through community fine-tunes. Fish Speech is larger (0.5B to 4.5B) and needs a GPU, but the catalog lists 80+ languages including Russian plus emotion control. Fish Speech moves faster: the latest S2-Pro shipped in 2026-03 against an F5-TTS update in 2025-03, though the S2-Pro weights use the Fish Audio Research License and commercial use requires a separate agreement.

Comparison based on catalog data

ParameterF5-TTSFish Speech / OpenAudio
CategoryText to speechText to speech
DeveloperShanghai Jiao Tong University and partners, ChinaFish Audio, USA / China
ReleasesOct 2024 – Mar 2025Apr 2024 – Mar 2026
Sizesabout 340M0.5B – about 4.5B
HardwareLaptopLaptop, 1 GPU
Commercial useNon-commercial onlyNon-commercial only
LicenseCode MIT, official weights CC BY-NC 4.0, non-commercialEarly versions CC BY-NC-SA 4.0, S2-Pro under the Fish Audio Research License; commercial use only under a separate agreement
RussianNot supportedSupported
OllamaNoNo
Without GPUNoNo
Tasks
  • Voice cloning
  • Voicing audiobooks and videos
  • Research and prototypes
  • Voice cloning
  • Emotional voiceover
  • Multilingual voiceover

Choose F5-TTS if

  • No GPU available: F5-TTS runs on a regular CPU
  • You want a smaller model, about 340M instead of several billion
  • English and Chinese are enough, or community fine-tunes work for you
F5-TTS

Choose Fish Speech / OpenAudio if

  • You need Russian and broad language coverage: the catalog lists 80+
  • Emotional delivery matters for your voiceover
  • You want the newest release: S2-Pro shipped in 2026-03
Fish Speech / OpenAudio

Other comparisons

Need a model for your task?

An open model can run on your own server: data stays in-house, there is no per-request fee, and the model can be fine-tuned on your documents.

  1. SelectThe model and size for your task and hardware budget
  2. DeployOn your server or in a closed network, with an API
  3. Fine-tuneOn your data, or connect a knowledge base
  4. IntegrateInto your CRM, ERP, bot, website or team chat
Discuss deployment