No — Gemma 3 12B won't fit 2x RTX 3090 (48GB) at Q4
Gemma 3 12B needs 8.1GB+ and exceeds 2x RTX 3090 (48GB). Try a smaller quant or model.
Multimodal 12B. Q4 ~8 GB, Q5 ~10 GB. Excellent on 12 GB cards.
Can Gemma 3 12B run on 2x RTX 3090 (48GB)?
Gemma 3 12B needs 8.1GB+ and exceeds 2x RTX 3090 (48GB). Try a smaller quant or model. Model context 128k. Compare with all models or re-check with your exact specs on the bench.
Unload with ollama stop, cap layers --num-gpu 28, cut context to 2k. See the OOM fixer on the homepage.
Start Q4. Bigger model at Q4 beats smaller at Q8. Only go Q8 if 2GB+ headroom remains.
Slower than GPU — this combo needs offload. Pick a smaller model for interactive chat.
Gemma — check terms before commercial use. Filter commercial-safe on the bench.