Measured model

Measured on a Mac Studio (M5 Max, 36GB)muse-glimmer:30b

Meta

On a Mac Studio (M5 Max, 36GB), muse-glimmer:30b answers a simple question in about 14.3 s, gets 85% of the Japanese knowledge questions right, 88% of the math problems, and uses 17.6GB of memory.

These numbers are from before the method changed on Oct 2, 2026 (see how it was measured).

Results

MeasureResultAmong measured modelsNote
Reply timeabout 14.3 s28 of 30Until one simple common-sense question is answered (median of 50)
Knowledge (high school to university)85%4 of 30200 questions, ±6
Math (word problems)88%3 of 30100 questions, ±8
Memory used17.6GB20 of 30
EN→JA translation36 pts1 of 28out of 100 (chrF), about 19 s a sentence
JA→EN translation60 pts1 of 28out of 100 (chrF), about 21.4 s a sentence

"Among measured models" is the place among the 30 models measured on this site (lower is better for time and memory).

Try it on your Mac

  1. Install the Mac version of Ollama from ollama.com and start it.
  2. Type this in Terminal. The first time, the model downloads; then you can chat with it.ollama run muse-glimmer:30b

It used 17.6GB here. To use it comfortably alongside other apps, a Mac with 36GB of memory or more is a good guide (the model then takes about half the memory).

About the model

Maker
Meta
Size
27.9B parameters
Quantization
Q4_K_M
Thinks before answering
Yes
Longest input at once
131,072 tokens
Released (Hugging Face)
Aug 9, 2026
Generation speed
25 tok/s
Start-up (loading)
about 6.1 s
Measured on
Sep 26, 2026
Ollama
0.34.3
Hugging Face
meta-models/Muse-Glimmer-30B
Series
Muse

Models of similar memory

Compare the main models on the text page. The series page lists every model measured.