37 pointsby fibo2 hours ago3 comments
  • simoiacos8 minutes ago
    Nothing comparable but inspired from DwarfStar I wrote a little inference engine for Intel Xe-LP (no XMX) 32GB laptops. The only model supported right now is a quantized Gemma-4, but I don't exclude in the future to support other MoE of similar size. Too bad we have no Qwen 3.8 35B-A3B yet.

    I'm also looking into expanding the protocol and the engine to support various steering techniques.

    https://github.com/simoneiacomino/xenolith

  • doctorpanglossan hour ago
    the problem is the dsv4 checkpoint so quantized isn't very good
  • 123-1129221 minutes ago
    Cannot read the vibe coded website because it hangs and lags.

    So recently opposition to the local LLM narrative pops up here and there and we get reassurance immediately. Is this automated by AI now? Make a sentiment analysis and produce slop articles that suppress last week's opposition?