The company also said that it chose Qwen3.6 over the newer Qwen3.8 (released earlier this month) because the latter runs slower on "today's Macs" since it needs reasoning enabled.
They’re also targeting 64GB M5 Pro and up as “today’s Macs” which perform fine with reasoning enabled.
It’s maybe a bit interesting that Jetbrains are making moves to integrate with local models more closely, but I think claiming this is “easier” is incorrect for most users.
Sad a Qwen3.8-35B-A3B model wasn't released.
Also, I think this is just an ad.
brew install llama.cpp
llama serve -hf unsloth/Qwen3.6-35B-A3B-GGUF:UD-Q4_K_M
open http://127.0.0.1:8080