Weights are not all of the memory.
The raw calculation is parameters × bits ÷ 8. Quantised formats also store scales and metadata, and some tensors stay at higher precision. The adjustable overhead is an assumption; if you know it, the real size of the model files is a better starting point.