109 pointsby mfiguiere5 hours ago5 comments
  • mmastrac42 minutes ago
    I've been working with an automatic incremental context compactor enabled and it's been surprisingly helpful. It was particularly effective with DS41f - I think I was running at an effective session length of 5M, with the model running around 300k-400k and it was holding on both speed and intelligence.

    TBH I also ran the 400tok/s preview and that was just nuts. I just let the thing compact over and over over the course of a day attacking a couple of tough problems

  • arikrahman4 hours ago
    I am very impressed with the KV Cache Compression work as well as the prefix cacheing making queries converge on practically free.
  • vivzkestrel3 hours ago
    404 on the blog page? https://zartbot.github.io/blog/
  • N_Lens2 hours ago
    [dead]
  • smy200113 hours ago
    Removed
    • girvo3 hours ago
      The website design definitely is, but I don’t know if the content is? This reads pretty human to me and is quite interesting to boot!
    • yunfei3 hours ago
      you don't know him?
    • sebmellen3 hours ago
      The writing feels human to me… and I call out AI slop as much as possible.