Llama 4 Maverick
Llama 4 Maverick loads in 245 GB at Q4_K_M and is comfortable from 256 GB of memory. Below: the full memory math, the cheapest card that runs it well, and the fastest machine we track.
PARAMETERS
400B (17B active)
FORMAT
Q4_K_M
MIN MEMORY
320 GB
BEST FOR
Frontier quality, Long context
Memory math
Weights (Q4_K_M)
245 GB
+ KV cache (16k)
~6.0 GB
Total at 16k
~251.0 GB
Comfortable from
256 GB
KV = fp16 estimate (q8_0 cache roughly halves it). "Comfortable" = weights + KV within the engine's tiered budget (~70-85% of memory).
Hardware snapshot
Cheapest GPU
—
no consumer fit
Fastest
—
n/a
Runs on a Mac?
No
beyond current Macs
More Llama models
Frequently asked questions
How much memory does Llama 4 Maverick need?
245 GB for the Q4_K_M weights, plus ~6.0 GB of KV-cache at 16k context — about 251.0 GB total. Comfortable from 256 GB of VRAM or unified memory.
Does Llama 4 Maverick run on a Mac?
Not on the current Mac configs we track; Llama 4 Maverick needs more memory than they offer.
What is the cheapest GPU for Llama 4 Maverick?
No tracked consumer GPU runs Llama 4 Maverick comfortably.
Cite this page
ModelFit: Llama 4 Maverick — specs, memory math and hardware verdicts. https://modelfit.io/models/llama4-maverick/ (dataset updated 2026-09-03, CC BY 4.0).