Parakeet or Qwen3-ASR: which transcription model to pick

Both families recognize Russian speech and run without a GPU, though the larger NVIDIA models expect one. Parakeet offers a wider range (110M to 2.5B) and has streaming versions, Nemotron Speech Streaming and Nemotron 3.5 ASR Streaming, for real-time work, with the latest release in 2026-05. The catch with NVIDIA is licensing: it is mixed, with Parakeet and Canary mostly CC-BY-4.0, the first Canary-1B non-commercial, Parakeet Unified and Nemotron Speech Streaming under the NVIDIA Open Model License and Nemotron 3.5 ASR under OpenMDW, so each model has to be checked on its own. Qwen3-ASR is simpler: the whole set is Apache 2.0, covers 50+ languages, and the catalog notes it copes with noise, singing and accents.

Comparison based on catalog data

ParameterNVIDIA Parakeet / Canary / Nemotron SpeechQwen3-ASR
CategorySpeech to textSpeech to text
DeveloperNVIDIA, USAAlibaba (Qwen), China
ReleasesDec 2023 – May 2026Jan 2026 – Jun 2026
Sizes110M – 2.5B0.6B – 1.7B
HardwareLaptop, 1 GPULaptop
Commercial useCommercial use with conditionsCommercial use allowed
LicenseMixed: Parakeet and Canary mostly CC-BY-4.0 (the first Canary-1B is non-commercial), Parakeet Unified and Nemotron Speech Streaming under the NVIDIA Open Model License, Nemotron 3.5 ASR under OpenMDWApache 2.0
RussianSupportedSupported
OllamaNoNo
Without GPUYesYes
Tasks
  • Transcribing calls and meetings
  • Video subtitles
  • Real-time voice input
  • Tagging audio archives
  • Transcribing calls and meetings
  • Video subtitles
  • Multilingual recognition
  • Voice input in apps

Choose NVIDIA Parakeet / Canary / Nemotron Speech if

  • You need streaming: real-time voice input and live subtitles
  • You want a size range from 110M to 2.5B for different hardware
  • You are willing to check the license for each specific model
NVIDIA Parakeet / Canary / Nemotron Speech

Choose Qwen3-ASR if

  • You want one clear license: all of Qwen3-ASR is Apache 2.0
  • Transcription spans many languages: the catalog lists 50+
  • Your audio is noisy, accented or includes singing
Qwen3-ASR

Other comparisons

Need a model for your task?

An open model can run on your own server: data stays in-house, there is no per-request fee, and the model can be fine-tuned on your documents.

  1. SelectThe model and size for your task and hardware budget
  2. DeployOn your server or in a closed network, with an API
  3. Fine-tuneOn your data, or connect a knowledge base
  4. IntegrateInto your CRM, ERP, bot, website or team chat
Discuss deployment