community pool / model answer

GLM-5.2-Int4-Int8Mix

chat · QuantTrio/GLM-5.2-Int4-Int8Mix

35 tok/s · model median · 0 runs in the model record.

Reference rigQuantizationMedian resultBasisTTFT (ms)Peak VRAMMax contextEngines

Context tested

The table above is the tested context for each exact rig and quantization cell. Empty fields remain empty rather than being filled with an estimate.