granite-speech-5.0-470m-turboctc: hardware requirements
granite-speech-5.0-470m-turboctc needs about 857 MB of memory at Q4 quantization — a 286 MB download plus room for context and overhead. That fits a machine with 8GB RAM.
Parameters
473M
Tier
Runs on a laptop
Licence
See model card
Released
2026-08-25
Will it run on your machine?
8GB RAMRuns well16GB RAMRuns well24GB RAMRuns well32GB RAMRuns well64GB RAMRuns well128GB RAMRuns well8GB VRAMRuns well12GB VRAMRuns well16GB VRAMRuns well24GB VRAMRuns well32GB VRAMRuns well
Download sizes by quantization
Lower quantization means a smaller file and less memory, at some cost to quality. Q4_K_M is the usual starting point. What is this?
| Quant | Download | RAM @ 4K | RAM @ 32K | Source |
|---|---|---|---|---|
| Q4_K_Mrecommended | 286 MB | 857 MB | 857 MB | estimated |
| Q5_K_M | 335 MB | 912 MB | 912 MB | estimated |
| Q6_K | 388 MB | 971 MB | 971 MB | estimated |
| Q8_0 | 503 MB | 1.1 GB | 1.1 GB | estimated |
| BF16 | 946 MB | 1.6 GB | 1.6 GB | estimated |
“Measured” sizes are read from published GGUF files. “Estimated” sizes are derived from the parameter count and are typically within a few percent. The RAM columns add the context cache and runtime overhead to the download size.
Model card on HuggingFace →API specs & pricing →Verified 2026-08-26Tracked for 12 months after release