community pool / model answer

Qwen2.5 Coder 7B Instruct GGUF

code · unsloth/Qwen2.5-Coder-7B-Instruct-GGUF

No data yet · 3 runs in the model record.

Reference rigQuantizationMedian resultBasisTTFT (ms)Peak VRAMMax contextEngines
L4 24GB (modal)4-bit47.2 tok/smeasured5724.8 GB4,096llama.cpp

Context tested

The table above is the tested context for each exact rig and quantization cell. Empty fields remain empty rather than being filled with an estimate.