> As for Chollet, the whole purpose of the ARC-AGI benchmark is to determine true progress toward AGI. While LLMs are performing better on this especially tough benchmark, none have come close to hitting the mark.
I'm curious what "the mark" refers to, since it seems like every new iteration of ARC-AGI is getting saturated faster than the last one.