In the release page: “The other model(s) in Qwen3.8-series would be released later”, we know that Qwen3.8-27B is dropping next, but that statement implies there might be other models beyond the 27B!

  • Domi@lemmy.secnd.me
    link
    fedilink
    English
    arrow-up
    1
    ·
    2 days ago

    128GB is plenty to try different models and see what works best. Have fun!

    In case you didn’t see it yet, your best friend for Strix Halo machines is https://strixhalo.wiki/

    Lots of info on how to setup llama.cpp (with pre-made toolboxes), configure the BIOS settings for maximum available VRAM and Linux for the best performance.

    I would recommend setting um llama-swap together with the Strix Halo Toolboxes with llama.cpp. So you can freely swap between models.