Would this be super fast for Qwen 27b & next? 32GB isn't much overhead for Q8 + 256k context

Wait 5 sec.

I have a 128gb m5 max now. It runs Qwen 3.8 27b well but as context grows the PP and output is so slow. Qwen Next is faster but still too slow to do lots of agentic work (coding) with. I am not a video gamer and I feel really bad if I contribute to hurting them if I buy this computer just for LLMs, but I would love to have a very speedy local Qwen 3.8 27b :p https://www.corsair.com/us/en/p/gaming-computers/cs-9060022-na/vengeance-a8200-gaming-pc-amd-ryzen-9-9950x3d-geforce-rtx-5090-64gb-ddr5-6tb-m-2-ssd-win11-pro-cs-9060022-na From what I've read with those two qwen models, I think my PP would be over 3,000 and my tks out would be 80 to 150? That sounds amazing to me. The only issue is the 5090 only has 32GB ram though so I'm scared I would have to use a very low quant like Q4 and that I couldn't fit a full 262k context. I was wondering if anyone knows if this build above would be worth it? Thank you!   submitted by   /u/lots_of_puppies [link]   [comments]