Measured model

Measured on a Mac Studio (M5 Max, 36GB)Magistral-Small-2509

Mistral · Ollama name hf.co/mistralai/Magistral-Small-2509-GGUF:Q4_K_M

On a Mac Studio (M5 Max, 36GB), Magistral-Small-2509 answers a simple question in about 28.8 s, gets 28% of the Japanese knowledge questions right, 1% of the math problems, and uses 16.2GB of memory.

Results

MeasureResultAmong measured modelsNote
Reply timeabout 28.8 s31 of 33Until one simple common-sense question is answered (median of 50)
Knowledge (high school to university)28%33 of 33200 questions, ±7
Math (word problems)1%33 of 33100 questions, ±4
Memory used16.2GB19 of 33
EN→JA translation0 pts30 of 30out of 100 (chrF), about 29 s a sentence
JA→EN translation6 pts30 of 30out of 100 (chrF), about 28.9 s a sentence

"Among measured models" is the place among the 33 models measured on this site (lower is better for time and memory).

Try it on your Mac

  1. Install the Mac version of Ollama from ollama.com and start it.
  2. Type this in Terminal. The first time, the model downloads; then you can chat with it.ollama run hf.co/mistralai/Magistral-Small-2509-GGUF:Q4_K_M

It used 16.2GB here. To use it comfortably alongside other apps, a Mac with 36GB of memory or more is a good guide (the model then takes about half the memory).

About the model

Maker
Mistral
Size
23.6B parameters
Quantization
Q4_K_M
Thinks before answering
No
Longest input at once
40,960 tokens
Released (Hugging Face)
Sep 12, 2025
Generation speed
28.9 tok/s
Start-up (loading)
about 1.3 s
Measured on
Oct 3, 2026
Ollama
0.34.3
Hugging Face
mistralai/Magistral-Small-2509
Series
Mistral

What the parts of a name mean (size, quantization and so on): how to read a model name.

Models of similar memory

Compare the main models on the text page. The series page lists every model measured.