GLM-5.2: hardware requirements
GLM-5.2 needs about 485 GB of memory at Q4 quantization — a 466 GB download plus room for context and overhead. That is beyond consumer hardware; use it through an API instead.
Parameters
753B
Tier
Server only
Licence
mit
Released
2026-06-16
Will it run on your machine?
8GB RAMNot enough memory16GB RAMNot enough memory32GB RAMNot enough memory64GB RAMNot enough memory128GB RAMNot enough memory8GB VRAMNot enough memory12GB VRAMNot enough memory16GB VRAMNot enough memory24GB VRAMNot enough memory32GB VRAMNot enough memory
Download sizes by quantization
Lower quantization means a smaller file and less memory, at some cost to quality. Q4_K_M is the usual starting point. What is this?
| Quant | Download | RAM @ 4K | RAM @ 32K | Source |
|---|---|---|---|---|
| Q4_K_Mrecommended | 466 GB | 485 GB | 595 GB | measured |
| Q5_K_M | 561 GB | 580 GB | 690 GB | measured |
| Q6_K | 626 GB | 645 GB | 755 GB | measured |
| Q8_0 | 801 GB | 820 GB | 930 GB | measured |
| BF16 | 1508 GB | 1527 GB | 1637 GB | measured |
“Measured” sizes are read from published GGUF files (unsloth/GLM-5.2-GGUF). “Estimated” sizes are derived from the parameter count and are typically within a few percent. The RAM columns add the context cache and runtime overhead to the download size.