Can I run this LLM? Real file sizes, live from Hugging Face.
All models › deepseek-v4 vs gemma-4-12b-it

deepseek-v4 vs gemma-4-12b-it — hardware requirements compared (measured)

Both sides use real GGUF file sizes from Hugging Face — not params×bits formulas. Updated daily.

deepseek-v4gemma-4-12b-it
Parameters403B12B
Downloads (30d)1956k898k
Quantizations published621
Smallest build needs118.6 GB (Q2KDown-AProjQ8-SExpQ8-OutQ8-chat-v2-imatrix-fixed-0731)6.6 GB (UD-IQ2_M)
Recommended quant needs118.6 GB (Q2KDown-AProjQ8-SExpQ8-OutQ8)6.6 GB (UD-IQ2_M)
Lossless build needs968.5 GB30.1 GB

Verdict: deepseek-v4 needs more memory at the recommended quant. Capability is a different question (see benchmark leaderboards) — this page answers only “which one fits my machine, and at what quality cost?”

Will it run on your GPU?

At the recommended quant, 8K context:

HardwareVRAM/RAMdeepseek-v4gemma-4-12b-it
RTX 3060 12GB12 GBnofits
RTX 4070 12GB12 GBnofits
RTX 4090 24GB24 GBnofits
RTX 5090 32GB32 GBnofits
Mac 24GB18 GBnofits
Mac 64GB48 GBnofits

Full measured tables

deepseek-v4every quant, exact GB gemma-4-12b-itevery quant, exact GB
Share this page: 𝕏 Post Reddit Hacker News Telegram WhatsApp More…