granite-embedding-97m-multilingual-r2: hardware requirements
granite-embedding-97m-multilingual-r2 needs about 655 MB of memory at Q4 quantization — a 105 MB download plus room for context and overhead. That fits a machine with 8GB RAM.
Parameters
97M
Tier
Runs on a laptop
Licence
See model card
Released
2026-05-14
Will it run on your machine?
8GB RAMRuns well16GB RAMRuns well32GB RAMRuns well64GB RAMRuns well128GB RAMRuns well8GB VRAMRuns well12GB VRAMRuns well16GB VRAMRuns well24GB VRAMRuns well32GB VRAMRuns well
Download sizes by quantization
Lower quantization means a smaller file and less memory, at some cost to quality. Q4_K_M is the usual starting point. What is this?
| Quant | Download | RAM @ 4K | RAM @ 32K | Source |
|---|---|---|---|---|
| Q4_K_Mrecommended | 105 MB | 655 MB | 655 MB | measured |
| Q5_K_M | 107 MB | 657 MB | 657 MB | measured |
| Q6_K | 113 MB | 664 MB | 664 MB | measured |
| Q8_0 | 115 MB | 666 MB | 666 MB | measured |
| BF16 | 206 MB | 768 MB | 768 MB | measured |
“Measured” sizes are read from published GGUF files (mykor/granite-embedding-97m-multilingual-r2-GGUF). “Estimated” sizes are derived from the parameter count and are typically within a few percent. The RAM columns add the context cache and runtime overhead to the download size.