• Avid Amoeba@lemmy.ca
    link
    fedilink
    English
    arrow-up
    12
    ·
    12 hours ago

    Thats kinda how the whole Chinese competition started - optimization to make due with limited hw.

    Also I just changed the inference engine I use for local models and Qwen 3.8 27B went from 50tps to 150tps. The new engine is optimized for my hw.