It uses an OpenAI-compatible API and can power private apps or agentic sessions; I’m currently using local models for creative writing and business projects. Qwen3.8 still needs some per-machine tuning and can use a lot of thinking tokens, but in one particularly extreme test it one-shot an absolutely gorgeous app after about 19 hours of thinking.