Computers capable of running GLM-4.6
What PC do you need to run GLM-4.6? 200 GB VRAM at Q4 (710 GB FP16) — see the exact AI workstations that host GLM-4.6 locally, from $9,499.
AI workstations that I can host GLM-4.6
Frequently asked
What PC do I need to run "GLM-4.6"?
At Q4 quantization GLM-4.6 needs about 200 GB of VRAM — the cheapest match in our catalog is the Mac Studio (M3 Ultra, 512 GB) (512 GB, $9,499). At Q8 plan for 360 GB, at FP16 710 GB.
How much VRAM does "GLM-4.6" need?
GLM-4.6 requires roughly 200 GB at Q4 (recommended for local use), 360 GB at Q8 and 710 GB at FP16, including KV-cache headroom.
Can I run "GLM-4.6" locally on my own computer?
Yes — with the right hardware. Any workstation with at least 200 GB VRAM runs GLM-4.6 locally, such as the Mac Studio (M3 Ultra, 512 GB). No cloud, no per-token bills.
Which GPU runs "GLM-4.6"?
The 80-core Apple GPU in the Mac Studio (M3 Ultra, 512 GB) (512 GB total) is the entry point; larger multi-GPU builds scale to 360–710 GB for Q8/FP16.
Cheap computer for GLM-4.6
The most affordable way to host GLM-4.6: a 200 GB-VRAM workstation — from $9,499 (Mac Studio (M3 Ultra, 512 GB)), which typically pays for itself vs. $1–4/hr cloud GPUs.
GLM-4.6 fine-tuning machine
Fine-tuning GLM-4.6 needs the FP16 size — about 710 GB of VRAM (fits: NVIDIA DGX Station Gen 2, Bizon ZX7000 — 8× H200).