• SirDimples@programming.dev
    link
    fedilink
    English
    arrow-up
    0
    ·
    21 days ago

    I’ve been testing this model since yesterday through the Qwen API in opencode and deepseek harness. I actually prefer it over GLM-5.3-Flash! it’s super good at agentic coding and comfortably fast, now my goal is to be able to one day run it locally 😳 I think a Strix Halo with 128GB and a fast nvme should be able to run it well at near loseless quant