community-verified · measured pool data

Can the Radeon AI Pro R9700 32GB ×3 run Qwen3-Coder-Next?

Yes — community-measured at 64.5 tok/s (4 runs at 5-bit).

Verified single-stream runs on the Radeon AI Pro R9700 32GB ×3 from the localmaxxing public pool. Every number on this page is measured — nothing here is estimated.

Quantization Median tok/s TTFT (ms) Peak VRAM Max context tested Engines
4-bit 68 59.1 717 llama.cpp
5-bit 64.5 431.9 70 GB 786 llama.cpp
Go deeper
I have a goal — find models that fit my use case → I have hardware — browse everything my rig can run → All machine × model answer pages →