community pool / model answer

Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive

chat · HauhauCS/Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive

73.9 tok/s · model median · 6 runs in the model record.

Reference rigQuantizationMedian resultBasisTTFT (ms)Peak VRAMMax contextEngines
M5 Max 128GB6-bit99.7 tok/sreported20- GB8,192llama.cpp
Radeon AI Pro R9700 32GB4-bit126.6 tok/sreported-- GB2,048llama.cpp
Radeon AI Pro R9700 32GB ×38-bit76.3 tok/sreported35449 GB787llama.cpp
Ryzen AI Max 395 128GB4-bit64.9 tok/sreported20124.2 GB32,768llama.cpp
T4 16GB2-bit59.3 tok/sreported-- GB2,048llama.cpp

Context tested

The table above is the tested context for each exact rig and quantization cell. Empty fields remain empty rather than being filled with an estimate.