VRAM explorer: which model fits your hardware
Pick an open model and see how much memory it needs and which machines it fits. No sign-up: the maths runs in your browser.
Work out the memory a model needs
Pick a model, a quantisation and a context. The bars show which machines it fits.
Estimated memory needed
Weights = parameters × bits per weight ÷ 8 (Q4_0 4.5 · Q4_K_M 4.85 · Q8_0 8.5 · FP16 16). Cache = MB per token × context tokens. Runtime = 1.5 GB.
Measured with ollama ps on our GB10, real use comes in 1.4 to 1.6 GB below this figure: we would rather you have memory to spare than run short. From each machine we also take off the share the operating system keeps.
Estimated speed: –
Which machines it fits
Tip:
Shall we set it up on your machine?
Get this analysis by email
Analysis saved. We will email it to you.
Next step: Hardware selector ROI calculator Grants
Edge AI Hardware Selector
Answer 3 questions and we recommend the ideal hardware for your case.
1. What is your primary use case?
Write the same sum yourself in the Dev Playground (12 min) →
Tell us what you want to run
Tell us what you want to run and on what budget. We will tell you which hardware you need, which model fits, and what to expect from it, before you spend anything.
69 free guides · 17 compliance templates