Skip to content

AI model by Mistral AI

Mistral Small 4 for business

Mistral's efficient open-weight model under Apache 2.0, one hybrid model for instructions, reasoning and coding, with tool calling and a 256K window.

  • Reviewed on September 24, 2026
  • Mistral AI

Capability tiersRelative, not benchmarks

Reasoning Medium
Coding Medium
Vision Very low
Speed High
Price level Lowest
Context size High
Tool calling High
Structured output High

Key facts about Mistral Small 4

Model ID at review
mistral-small-2603
Provider
Mistral AI
Main category
Open weights
Open weights
Yes, Apache 2.0
Last reviewed
September 24, 2026

Fit

Where it fits and where it does not

Good at

  • Efficient serving, with about 6.5 billion active parameters.

  • Instructions, reasoning and code in one model.

  • Tool calling and structured outputs.

  • Self-hosting under Apache 2.0.

  • Long inputs, with a 256K token window.

Not the right choice for

  • The hardest reasoning tasks.

  • Image input, which we could not confirm at review time.

  • Teams that do not want to run infrastructure, unless they use the API.

  • Use without testing on your own data.

Use cases

Business use cases we would use it for

Self-hosted assistants

Private assistants on modest hardware. See internal help desk.

Private processing

Classification and extraction of sensitive text in your own environment.

Code helpers

Internal coding helpers where code must not leave your network.

Our notes

When we would choose it

Mistral Small 4 was released in March 2026 under the Apache 2.0 license. Mistral describes it as one hybrid model for instruction following, reasoning and coding. It is a mixture-of-experts model with 119 billion parameters in total and about 6.5 billion active, which makes it efficient to serve for its capability. It supports tool calling, structured outputs and a 256K token context window.

For self-hosted projects, Small 4 is one of the most practical options at review time. The permissive license, efficient serving and support for tools and JSON output cover what most business assistants and automations need.

We could not confirm image input from the model card at review time, so the vision meter is set low. If you need to read images or scans, check the current documentation or look at Gemma 4 or Mistral Large 3.

The model is also available through the Mistral API, which is a simple way to test it before deciding whether to host it. We would run your real tasks through it and a hosted leader, then weigh the quality difference against the benefits of running it yourself. See open weights models.

Keep in mind that total size, not active size, decides how much memory you need. A mixture-of-experts model is fast per token but still has to be loaded in full. Size your servers on the total, and test throughput with realistic traffic before launch.

Before you commit

Things to check before you commit

  • Memory

    Total size is 119 billion parameters, so memory needs are higher than the active count suggests.

  • Vision

    Confirm image support in the current model card if you need it.

  • Quality

    Compare with Gemma 4 and hosted models on the same test set.

  • License

    Apache 2.0 is permissive. Confirm it fits your use.

Alternatives

Models to compare it with

Gemma 4

Google's open-weight model family under Apache 2.0, in sizes from phone-friendly to 31B, with image input and function calling.

Mistral Large 3

Mistral's open-weight general-purpose flagship under Apache 2.0, with image input, tool calling, structured outputs and a 256K token window.

Qwen3-Coder

Qwen's open-weight coding models under Apache 2.0, from the efficient Qwen3-Coder-Next to the large 480B model, with tool use and long context.

Keep exploring

FAQ

Questions people ask us

Have a question that is not here? Ask us directly.

Start a project

Not sure which model fits? Ask us to evaluate your use case.

Send a short brief. We reply with questions, a suggested plan and an estimate you can compare with other offers.

Your privacy choices

We use necessary storage to run this site. With your permission we also use Google Analytics to see which pages help people, and load maps from Google. You can change this at any time. Read the cookie policy.