Best small model + setup together with qwen3.8-flash-next

Wait 5 sec.

hello, I currently use swift qwen3.8 gsq rco in iq3 xxs in strata. The model is only used for hermes agent. I have 64GB of ram + 32GB R9700 where the model runs on. But I also have an rx7800xt with 16GB. For hermes agent i want to run a small but intelligent enough model as an executioner model, where I need a few parallel sessions. Ive had a few ideas: 1. R9700 + RAM: Qwen Flash next, rx7800: minicpm5-2b 2. rx7800xt + ram: qwen flash next, r9700: idk, maybe qwen3.8 27b q4 3. r9700 + ram: qwen flash next, rx7800xt: idk, maybe qwen3.8 27b q3 what would yall do?   submitted by   /u/thatscoolbutno123 [link]   [comments]