ModelFitCan I run this LLM? Real file sizes, live from Hugging Face.
What LLM can I run? Pick your VRAM / RAM
Type your GPU VRAM (or usable system RAM for CPU inference). Every model that fits is listed with its
best quant — computed from measured GGUF file sizes, 8K context, ~8% headroom. Data refreshes daily from Hugging Face.