Qwen3.8-Flash-Next Coder on a 12 GB RTX 3060 and 32 GB of DDR4. Strata prefills 6-7x faster than stock llama.cpp. One batch setting closes most of that.
A weekly email with every new guide and measured benchmark.