community-verified · measured pool data

Can the Ryzen AI Max 395 128GB run gemma-4-26B-A4B-it?

Yes — community-measured at 60.7 tok/s (4 runs at 4-bit).

Verified single-stream runs on the Ryzen AI Max 395 128GB from the localmaxxing public pool. Every number on this page is measured — nothing here is estimated.

Quantization Median tok/s TTFT (ms) Peak VRAM Max context tested Engines
4-bit 60.7 21.5 GB 32,768 llama.cpp
8-bit 46.1 745.6 27.8 GB 16,384 llama.cpp
Go deeper
I have a goal — find models that fit my use case → I have hardware — browse everything my rig can run → All machine × model answer pages →