Back to glossary

Term

Mistral 7B

Mistral 7B (September 2023) is Mistral AIs first open-weight model — a 7.3B-parameter model under Apache 2.0 that at the time outperformed the larger Llama 2 13B on benchmarks.

Mistral 7B — explained in more detail

Mistral 7B was the debut model of the French lab Mistral AI, released in September 2023. With around 7.3 billion parameters it belongs to the class of small, efficient models — and became well known because, at release, it outperformed the much larger Llama 2 13B on all tested benchmarks and beat Llama 1 34B in many areas. The model is under the permissive Apache 2.0 license and is therefore fully free to use, download and deploy commercially.

The efficiency comes from two architectural tricks: Grouped-Query Attention (GQA) speeds up inference, and Sliding Window Attention (SWA) allows longer sequences to be processed more cheaply. In benchmarks Mistral 7B reached about 60.1 on MMLU (knowledge, 5-shot) versus 55.6 for Llama 2 13B, and 30.5 on HumanEval (code) versus 18.3 — the gap was especially clear on code and reasoning. This combination of small size and strong performance made Mistral 7B one of the most influential open models of its time.

Example / Practical context

The practical appeal of Mistral 7B is that it is small enough to run on modest hardware — for example a single consumer GPU or, quantized, even capable laptops. A typical use: a developer downloads the model via Ollama or llama.cpp and runs a local chat or extraction service without sending data to a cloud API. Thanks to the Apache 2.0 license, Mistral 7B also served as a popular base for countless fine-tunes, including German-language variants and domain-specific offshoots.

Mistral 7B is a classic dense model, not a Mixture-of-Experts like the later Mixtral. Within Mistral AIs portfolio it sits at the small, efficient end and has over the years been superseded by larger and newer models (such as Mistral Small, Mistral Large) and newer small-model generations. It differs from proprietary API models (GPT, Claude, Gemini) through its open weights and free license. As a 7B model it is today far behind current frontier models in absolute capability, but it remains a reference point for the leap that efficient small models made in 2023.

See everything in one place:Mistral