Gemma 3 27B Instruct

Gemma 3 27B Instruct loads in 21 GB at Q4_K_M and is comfortable from 36 GB of memory, Macs included from the MacBook Pro M5 Pro 48GB. Below: the full memory math, the cheapest card that runs it well, and the fastest machine we track.

PARAMETERS
27B
FORMAT
Q4_K_M
MIN MEMORY
32 GB
BEST FOR
Quality, Coding

Memory math

Weights (Q4_K_M)
21 GB
+ KV cache (16k)
~4.0 GB
Total at 16k
~25.0 GB
Comfortable from
36 GB

KV = fp16 estimate (q8_0 cache roughly halves it). "Comfortable" = weights + KV within the engine's tiered budget (~70-85% of memory).

Hardware snapshot

Cheapest GPU
~32 tok/s est. — ~$700 used
Fastest
~52 tok/s est.
Runs on a Mac?
~15 tok/s est.

More Gemma models

Frequently asked questions

How much memory does Gemma 3 27B Instruct need?

21 GB for the Q4_K_M weights, plus ~4.0 GB of KV-cache at 16k context — about 25.0 GB total. Comfortable from 36 GB of VRAM or unified memory.

Does Gemma 3 27B Instruct run on a Mac?

Yes — from the MacBook Pro M5 Pro 48GB (~15 tok/s est.). Unified memory means the RAM budget is the only limit.

What is the cheapest GPU for Gemma 3 27B Instruct?

The AMD Radeon RX 7900 XTX is the cheapest tracked card that runs Gemma 3 27B Instruct comfortably — ~32 tok/s est. at ~$700 used.

Cite this page

ModelFit: Gemma 3 27B Instruct — specs, memory math and hardware verdicts.
https://modelfit.io/models/gemma3-27b/ (dataset updated 2026-09-03, CC BY 4.0).