Devstral
Mistral models for agentic development: they read the repository, edit files and run commands on their own. The 24B version fits on a single GPU.
- Developer
- Mistral AI (with All Hands AI), France
- First release
- May 2025
- Latest release
- Dec 2025
- Sizes
- 24B – 123B
- License
- Commercial use with conditionsSmall 24B versions: Apache 2.0; Devstral 2 123B: modified MIT (not for companies with revenue above 20 million dollars per month)
- Russian
- Supported
- Ready-made builds
- GGUF, AWQ, GPTQ, MLX (Apple)
- Running
- Available in OllamaNeeds a GPU
- Industries
- Software development
What it does
- A developer agent that fixes tickets from the tracker
- Extending internal systems from a description
- Automating routine code edits
Where it is used
Hardware requirements
Versions
- Devstral 2 123B и Devstral Small 2 24B
- Devstral Small 2507
- Devstral Small 2505
How to run it
I can set this up end to end: pick the model size, deploy it on your server and connect it to your systems. Quantization compresses a model so it takes less video memory and runs on more modest hardware. Answers change slightly, so quality is checked on your own examples.
Frequently asked questions
Can Devstral be used in a commercial project?
With conditions. License: Small 24B versions: Apache 2.0; Devstral 2 123B: modified MIT (not for companies with revenue above 20 million dollars per month). Restrictions vary — region, company revenue, attribution requirements. Have a lawyer check the terms before a commercial launch.
What hardware does Devstral need?
At minimum: One GPU with 16–80 GB — mid-size versions. Without a GPU the model is not practical. You can calculate the exact VRAM for your model size and context in the hardware calculator.
Does Devstral support Russian?
Yes, Russian is listed on the model card.
Where can I download Devstral and what does it cost?
The Devstral weights are open and free to download. You only pay for the hardware it runs on and for the setup. Source links are at the bottom of this page.
How I deploy it for clients
- SelectionI pick the model size for your task and hardware and test it on your examples.
- DeploymentI deploy it on your server or in a closed network and provide an API.
- Fine-tuningI fine-tune it on your data (LoRA) or connect a knowledge base — whichever is cheaper for the task.
- IntegrationI connect it to your CRM, ERP, bot, website or team chat and set up monitoring.
Comparisons
Similar models
The broadest open coding family: from 0.5B for autocompletion to 480B for agents. Qwen3-Coder-Next (80B, 3B active) works as a developer agent on a single GPU.
DetailsCodeKAT-Coder / KAT-DevKwaipilot (Kuaishou) · ChinaCommercial use allowedKuaishou models for agentic development, trained to solve real tasks in repositories. KAT-Coder-V2.5-Dev (35B, 3B active) is the open version of their closed flagship.
DetailsCodeSERA (Ai2 Open Coding Agents)Ai2 (Allen Institute for AI) · USACommercial use allowedFully open developer agents from Ai2: weights, data and training recipe are all public. Designed so a company can cheaply fine-tune the agent on its own repository.
DetailsSource: huggingface.co/mistralai/Devstral-2-123B-Instruct-2512. Data checked against the model card on 22 Sep 2026. Have a lawyer review the license before commercial launch.


