Before you download gigabytes of model weights, know whether they'll actually fit.
VRAMFit tells you the largest local large language model you can run on your hardware — whether that's an NVIDIA GPU's VRAM or the unified memory on an Apple Silicon Mac. Set your memory, pick a quantization, and get a clear, instant answer.
Built for people who run models locally: llama.cpp, Ollama, LM Studio, and anyone deciding what GPU to buy next.
WHAT IT DOES
• Enter your VRAM (or Apple Silicon RAM) and instantly see what fits
• Accounts for quantization — from full precision down to 4-bit and FP4
• Factors in context length and KV cache, the memory costs people usually forget
• Switches between discrete GPU and Apple Silicon unified-memory math
• Ships with a 2026 catalog of popular open models, sized for you
FREE
• The full calculator, with quantization, context, and FP4
• A starter set of 5 popular models
• GPU and Apple Silicon modes
VRAMFit Pro (one-time upgrade)
• The complete 20-model 2026 catalog
• Per-quantization memory breakdown for every model
• Advanced tools: KV-cache compression and multi-token prediction estimates
PRIVATE BY DESIGN
VRAMFit runs entirely on your device. No account, no sign-in, no tracking, and nothing about you leaves your iPhone. The numbers you enter stay with you.
Stop guessing whether a model will fit. Check first, download once.
用户评价
立即分享产品体验
你的真实体验,为其他用户提供宝贵参考
💎 分享获得宝石
【分享体验 · 获得宝石 · 增加抽奖机会】
将你的产品体验分享给更多人,获得更多宝石奖励!
💎 宝石奖励
每当有用户点击你分享的体验链接并点赞"对我有用",你将获得:
🔗 如何分享
复制下方专属链接,分享到社交媒体、群聊或好友:
💡 小贴士
分享时可以添加你的个人推荐语,让更多人了解这款产品的优点!
示例分享文案:
"推荐一款我最近体验过的应用,界面设计很精美,功能也很实用。有兴趣的朋友可以看看我的详细体验评价~"
关注 Mergeek 公众号
领奖遇到问题?联系小门助手