VRAMFit: LLM Calculator

VRAMFit: LLM Calculator

Will that LLM fit your GPU?

1 ratings
Free

Rating summary
1
0
0
0
0

Details

  • Released
  • Updated
  • July 21, 2026
  • September 3, 2026

Features

VRAMFit: LLM Calculator screenshot #1 for iPhone
VRAMFit: LLM Calculator screenshot #2 for iPhone
VRAMFit: LLM Calculator screenshot #3 for iPhone
VRAMFit: LLM Calculator screenshot #4 for iPhone
VRAMFit: LLM Calculator screenshot #5 for iPhone
VRAMFit: LLM Calculator screenshot #6 for iPhone
VRAMFit: LLM Calculator screenshot #7 for iPhone
VRAMFit: LLM Calculator screenshot #8 for iPhone
🖼️Get Icon
Icons↘︎

About

Determine the largest local large language model that fits your hardware, whether it's an NVIDIA GPU or Apple Silicon Mac. Input your memory, select quantization, and get an instant, accurate answer based on measured model sizes.

Calculate LLM fit for GPU and Apple Silicon
Accounts for quantization (4-bit, FP4)
Includes context length and KV cache costs
Provides measured, not guessed, model sizes
Sorts models by what fits and what nearly fits
Show original description

What's New in VRAMFit

1.1.0

September 3, 2026

Model sizes are now measured, not calculated. VRAMFit used to size a model by multiplying its parameters by a rule of thumb for each quantization. It now ships the real size of the published file for every model at every setting it is actually released in, and names the exact build each number came from. • A much larger catalog of current open models, every one measured • Models that only exist at one quantization say so, instead of showing a size for a build that was never published • Natively 4-bit models no longer appear to grow as you move up the scale • The list is grouped into what fits, what fits one setting down, and what won't — so a model that just misses tells you which setting would run it • Rewritten help explaining where the numbers come from and what they still cannot know Estimates are higher than before at every setting. That is the correction: the old rule of thumb ran low.

More

Developer apps

FAQ

What is VRAMFit?

VRAMFit is a developer tool that helps you determine the largest local large language model (LLM) that will fit into your computer's VRAM or unified memory. It provides measured sizes for models at various quantizations, factoring in context length and KV cache.

How does VRAMFit calculate model sizes?

Unlike other calculators that use rules of thumb, VRAMFit uses the actual file size of published models at specific quantizations. This ensures accuracy, especially for models with unique characteristics like ternary or natively 4-bit builds.

What hardware does VRAMFit support?

VRAMFit supports both NVIDIA GPUs for VRAM calculations and Apple Silicon Macs for unified memory calculations. You can switch between these modes within the app.

Does VRAMFit have ads?

No, VRAMFit is ad-free. The free version offers the full calculator functionality, and an optional one-time upgrade to VRAMFit Pro unlocks the complete model catalog and advanced tools.

Is VRAMFit private?

Yes, VRAMFit is designed to be private. It runs entirely on your device, requiring no account, sign-in, or tracking. Your entered data stays with you.

What is the difference between VRAMFit and VRAMFit Pro?

The free version of VRAMFit includes the full calculator with quantization, context, and FP4 support, along with a starter set of models. VRAMFit Pro, a one-time upgrade, provides the complete catalog of 73 current open models, per-quantization breakdowns, and advanced tools like KV-cache compression estimates.

How often is VRAMFit updated?

The latest version of VRAMFit is 1.1.0, last updated on September 3, 2026. Updates typically include new model data and feature enhancements.

What is the age rating for VRAMFit?

VRAMFit has an age rating of 4+, making it suitable for a wide range of users interested in LLM hardware compatibility.