You must log in or register to comment.
Thats kinda how the whole Chinese competition started - optimization to make due with limited hw.
Also I just changed the inference engine I use for local models and Qwen 3.8 27B went from 50tps to 150tps. The new engine is optimized for my hw.



