VRAMFit: LLM Calculator
VRAMFit: LLM Calculator has 1 ratings, averaging 5.0 stars, #1,215 of 1,705 in Developer Tools.
Description
Before you download gigabytes of model weights, know whether they'll actually fit.
VRAMFit tells you the largest local large language model you can run on your hardware — whether that's an NVIDIA GPU's VRAM or the unified memory on an Apple Silicon Mac. Set your memory, pick a quantization, and get a clear, instant answer.
Built for people who run models locally: llama.cpp, Ollama, LM Studio, and anyone deciding what GPU to buy next.
MEASURED, NOT GUESSED
Most calculators multiply a parameter count by a rule of thumb. VRAMFit ships the real size of the published file for every model at every quantization it is actually released in, and names the exact build each figure came from — so you can check it against the download before you start it.
That matters most where rules of thumb break. A ternary model has no 4-bit build to estimate. A natively 4-bit model barely changes size as you move up the scale. VRAMFit says what those models actually weigh, or tells you plainly that the build you picked does not exist.
WHAT IT DOES
• Enter your VRAM (or Apple Silicon RAM) and instantly see what fits
• Sorts the catalog into what runs, what runs one setting down, and what won't
• Accounts for quantization — from full precision down to 4-bit and FP4
• Factors in context length and KV cache, the memory costs people usually forget
• Switches between discrete GPU and Apple Silicon unified-memory math
FREE
• The full calculator, with quantization, context, and FP4
• A starter set of models spanning what fits and what nearly fits
• GPU and Apple Silicon modes
VRAMFit Pro (one-time upgrade)
• The complete catalog of 73 current open models, measured
• Per-quantization memory breakdown for every model
• Advanced tools: KV-cache compression and multi-token prediction estimates
PRIVATE BY DESIGN
VRAMFit runs entirely on your device. No account, no sign-in, no tracking, and nothing about you leaves your iPhone. The numbers you enter stay with you.
Stop guessing whether a model will fit. Check first, download once.
Growth change in ratings
| Per | Growth | Growth % | Actual change | Measured over |
|---|---|---|---|---|
| day | Not enough history yet | |||
| week | Not enough history yet | |||
| month | Not enough history yet | |||
| year | Not enough history yet | |||
How these figures are measured
Change in ratings, scaled to each window from the real gap between crawls. A window needs a full period of history before it reports anything.
Tracking since 09/18/2026. History accumulates as changes are observed — check back soon.
Alternatives to VRAMFit: LLM Calculator
| Name | Ratings | Rating |
|---|---|---|
|
|
32 | 4.6 |
|
|
10 | 3.3 |
|
|
6 | 5.0 |
|
|
5 | 4.0 |
|
|
4 | 5.0 |
|
|
4 | 4.3 |
|
|
3 | 3.7 |
|
|
3 | 5.0 |
|
|
1 | 4.0 |
|
|
1 | 5.0 |