Qwen3.5 35B-A3B Instruct
Qwen3.5 35B-A3B Instruct loads in 20 GB at Q4_K_M and is comfortable from 32 GB of memory, Macs included from the MacBook Pro M5 Pro 48GB. Below: the full memory math, the cheapest card that runs it well, and the fastest machine we track.
Memory math
KV = fp16 estimate (q8_0 cache roughly halves it). "Comfortable" = weights + KV within the engine's tiered budget (~70-85% of memory).
Hardware snapshot
Go deeper
More Qwen models
Frequently asked questions
How much memory does Qwen3.5 35B-A3B Instruct need?
20 GB for the Q4_K_M weights, plus ~0.3 GB of KV-cache at 16k context — about 20.3 GB total. Comfortable from 32 GB of VRAM or unified memory.
Does Qwen3.5 35B-A3B Instruct run on a Mac?
Yes — from the MacBook Pro M5 Pro 48GB (~39 tok/s est.). Unified memory means the RAM budget is the only limit.
What is the cheapest GPU for Qwen3.5 35B-A3B Instruct?
The AMD Radeon RX 7900 XTX is the cheapest tracked card that runs Qwen3.5 35B-A3B Instruct comfortably — ~72 tok/s est. at ~$700 used.
Cite this page
ModelFit: Qwen3.5 35B-A3B Instruct — specs, memory math and hardware verdicts. https://modelfit.io/models/qwen3.5-35b-a3b/ (dataset updated 2026-09-03, CC BY 4.0).