ModelRigSYS: READY
Home / Models / Qwen2.5 32B
LLM / MODEL FILE

Qwen2.5 32B

Qwen model for higher-quality chat and coding. Memory figures are estimates, not measured benchmarks.

Parameters32.5B
Q4 estimate24.2 GB
WorkloadLLM
Best proofTest exact runtime
QUANTIZATION

Memory scenarios

CLOSEST FITS

Hardware near the Q4 line

LIMITFitting in memory does not guarantee acceptable speed.

Runtime, drivers, memory bandwidth, context, batch size, offloading, and model architecture all matter.