ModelRigSYS: READY
Home / Models / Gemma 3 4B
LLM / MODEL FILE

Gemma 3 4B

Google model for small multimodal assistant. Memory figures are estimates, not measured benchmarks.

Parameters4.3B
Q4 estimate4.5 GB
WorkloadLLM
Best proofTest exact runtime
QUANTIZATION

Memory scenarios

CLOSEST FITS

Hardware near the Q4 line

LIMITFitting in memory does not guarantee acceptable speed.

Runtime, drivers, memory bandwidth, context, batch size, offloading, and model architecture all matter.