Orca 2
Microsoft research models based on Llama 2, trained to choose a reasoning approach for each task. The orca-mini model in Ollama is a different project by independent developer Pankaj Mathur.
The last open version came out in Nov 2023. The family has not been updated for a long time: the model still works, but do not expect fixes or new sizes.
- Developer
- Microsoft Research, USA
- First release
- Nov 2023
- Latest release
- Nov 2023
- Sizes
- 7B – 13B
- License
- Non-commercial onlyMicrosoft Research License: research only, no commercial use
- Russian
- Not stated
- Ready-made builds
- GGUF
- Running
- Available in OllamaNeeds a GPU
- Industries
- Science and research, Education
What it does
- Research on reasoning methods
- Comparison with modern small models
- Training specialists
Where it is used
Hardware requirements
Versions
- Orca 2 (7B, 13B)
How to run it
I can set this up end to end: pick the model size, deploy it on your server and connect it to your systems. Quantization compresses a model so it takes less video memory and runs on more modest hardware. Answers change slightly, so quality is checked on your own examples.
Frequently asked questions
Can Orca 2 be used in a commercial project?
No. License: Microsoft Research License: research only, no commercial use. A commercial product needs a different model or a separate agreement with the rights holder.
What hardware does Orca 2 need?
At minimum: Laptop or regular PC, up to 8 GB of VRAM — smaller versions. Without a GPU the model is not practical. You can calculate the exact VRAM for your model size and context in the hardware calculator.
Does Orca 2 support Russian?
The model card does not list languages, so Russian support cannot be promised — it has to be tested on your own examples.
Where can I download Orca 2 and what does it cost?
The Orca 2 weights are open and free to download. You only pay for the hardware it runs on and for the setup. Source links are at the bottom of this page.
How I deploy it for clients
- SelectionI pick the model size for your task and hardware and test it on your examples.
- DeploymentI deploy it on your server or in a closed network and provide an API.
- Fine-tuningI fine-tune it on your data (LoRA) or connect a knowledge base — whichever is cheaper for the task.
- IntegrationI connect it to your CRM, ERP, bot, website or team chat and set up monitoring.
Similar models
Small Microsoft models trained on carefully selected data: strong at logic and math for their modest size. Versions with images and speech are available.
DetailsTextLlamaMeta · USACommercial use with conditionsThe models that started mass open source in AI. A huge ecosystem of fine-tuned versions and tools.
DetailsTextWizardLM, WizardCoder, WizardMathWizardLM (Microsoft and Peking University) · USA / ChinaCommercial use with conditionsFine-tunes of Llama, Mistral and StarCoder using Evol-Instruct, which automatically makes instructions more complex. WizardLM-2 was released in April 2024 and removed almost immediately, so only the 2023 versions are relevant.
DetailsSource: huggingface.co/microsoft/Orca-2-13b. Data checked against the model card on 22 Sep 2026. Have a lawyer review the license before commercial launch.


